The state of repos
What the corpus says: how repos score on agent readiness, and where they make AI agents stumble.
- Repos judged
- 62
- Average score
- 74
- Median score
- 75
Score distribution
Latest score per repo, bucketed into the bands the judge actually grades by.
- Genuinely excellent (90–100)1
- Solid, with gaps (70–89)52
- Rough (50–69)6
- A real problem (0–49)3
Where repos stumble the most
Category averages across the corpus, weakest first — plus how many prioritized issues landed in each.
| Category | Average score | Issues found |
|---|---|---|
| Reproducible Setup | 55 | 27 |
| Verification Surface | 59 | 26 |
| Context Economy | 64 | 26 |
| Code Navigability | 67 | 29 |
| Machine Orientation | 69 | 25 |
| Agent Safety | 69 | 26 |