model profile
Rank · overall#27
Score74.2
Benchmark scores
4 benchmarksCodingAgentBench92%Pass rate
Chatbot Arena (Coding)1433Coding Elo
MathArena Apex13.5%Accuracy
BrowseComp69%Accuracy
Across harnesses
10 agentsQwen CodeCLI
92%———
99OpenCodeTUI
88%———
96Copilot CLICLI
84%———
96OpenHandsAgent
84%———
96piCLI
76%———
89AiderCLI
72%———
86Codex CLICLI
68%———
84CrushTUI
60%———
77GooseCLI
24%———
61PlandexCLI
8%———
44Per-benchmark score · composite per harness. Tap a row to open that board.
Scraped and aggregated from public leaderboards · 2026-08-16