Thirteen agent harnesses × 27 capabilities = 351 scored cells, each scored 0–2 against the vendor’s own documentation. Tap any cell for its evidence and source.
The number is documented operational breadth — how much surface a project documents and ships. Not coding quality, not output quality, and not a ranking: a deliberately lean core scores lower on purpose. The Method tab inside the map has the full rubric.
Open the map full size in a new tab · static exports: checklist (Markdown) · checklist (CSV, one row per cell) · peer-review audit trail
