Skip to content

The state of repos

What the corpus says: how repos score on agent readiness, and where they make AI agents stumble.

Repos judged
62
Average score
74
Median score
75

Score distribution

Latest score per repo, bucketed into the bands the judge actually grades by.

Where repos stumble the most

Category averages across the corpus, weakest first — plus how many prioritized issues landed in each.

CategoryAverage scoreIssues found
Reproducible Setup5527
Verification Surface5926
Context Economy6426
Code Navigability6729
Machine Orientation6925
Agent Safety6926

Browse by language