BenchAtlas

Sources / arc-prize

ARC Prize Leaderboard

healthydefault: independent

arcprize.org/leaderboard · trust rank 80 · adapter arc-prize

Observations
1,226
append-only score facts
Raw snapshots
29
immutable fetched payloads
Last success
2026-09-12 05:30Z
end of last succeeded/partial run
Freshness
2.8 h
hours since last success

License & attribution

Task datasets are Apache-2.0 and the harness/raw results MIT; the leaderboard JSONs are unlicensed with no access prohibition. Cited per ARC Prize policy (arcprize.org/policy).

What we ingest

Static leaderboard JSONs (evaluations, models, datasets): per-configuration accuracy and cost-per-task across ARC-AGI-1/2/3. Semi-private and public eval splits never mix — the semi-private split is the canonical variant, and cost never enters the composite.

Ingestion runs

last 10 of the run history
StartedDurationStatus
2026-09-12 05:30Z8.9 spartial
2026-09-11 05:30Z13.0 spartial
2026-09-10 05:30Z17.9 spartial
2026-09-09 05:30Z8.1 spartial
2026-09-08 05:30Z10.4 spartial
2026-09-07 05:30Z9.8 spartial
2026-09-06 05:30Z7.6 spartial
2026-09-05 05:30Z11.6 spartial
2026-09-04 05:30Z11.0 spartial
2026-09-03 05:30Z23.7 spartial

partial = run finished but some rows are blocked on identity review — nothing is published until every referenced model/agent/benchmark is resolved.

Raw snapshots

latest 5
FetchedURLSHA-256
2026-09-04 05:30Zhttps://arcprize.org/media/data/models.json2a3ac802eeda
2026-09-04 05:30Zhttps://arcprize.org/media/data/evaluations.json17b35f029522
2026-09-02 05:30Zhttps://arcprize.org/media/data/models.json95b03912ddee
2026-09-02 05:30Zhttps://arcprize.org/media/data/evaluations.jsond20f203e9811
2026-08-22 05:30Zhttps://arcprize.org/media/data/models.json17aeed69df45

Snapshots are immutable raw payloads — every published number traces back to one. Re-fetching an identical payload bumps the re-fetch time instead of storing a duplicate. Scores from this source appear on benchmark profiles and model pages with this source linked.