Skip to content

The state of repos

What the corpus says: how repos score on agent readiness, and where they make AI agents stumble.

Repos judged
63
Average score
74
Median score
75

Score distribution

Latest score per repo, bucketed into the bands the judge actually grades by.

Where repos stumble the most

Category averages across the corpus, weakest first — plus how many prioritized issues landed in each.

CategoryAverage scoreIssues found
Reproducible Setup6433
Verification Surface6828
Context Economy6931
Code Navigability7133
Agent Safety7230
Machine Orientation7427

Browse by language