model profile
Rank · overall#6
Score87.8
Benchmark scores
6 benchmarksTerminal-Bench 2.178.4%Accuracy · Jul 11, 2026 · ± 1.3
Chatbot Arena (Coding)1562Coding Elo
Artificial Analysis55Intelligence Index
AA Coding Index76.7Coding Index
ARC-AGI83.9%Score
FrontierMath (Tier 4)70.7%Accuracy · ± 7.2
Across harnesses
1 agentsPer-benchmark score · composite per harness. Tap a row to open that board.
Scraped and aggregated from public leaderboards · 2026-07-22