model profile
Rank · overall#412
Score28.2
Benchmark scores
4 benchmarksSWE-bench Verified34.8%% Resolved · 2025-08-07
Chatbot Arena (Coding)1363Coding Elo
ARC-AGI2.6%Score
FrontierMath (Tier 4)2.4%Accuracy · ± 2.4
Across harnesses
0 agentsThis model appears only in model-level benchmarks — no harness ran it directly.
Scraped and aggregated from public leaderboards · 2026-07-23