Data sourceLLM Stats
Ranking methodtrueskill
Ranked atSep 29, 2026, 1:36 PM
Models included50
How to read the overall board

The overall board aggregates several coding benchmarks. Use it to compare broad relative position, then inspect individual tasks, price, and deployment constraints.

Showing 50 / 50
RankModelOrganizationScoreConservativeLicenseMin input priceEvaluated
1anthropic0.751.6Closed$412
2anthropic0.748.9Closed$212
3openai0.748.7Closed$108
4openai0.645.1Closed$517
5anthropic0.744.6Closed$109
6anthropic0.843.9Closed—6
7deepseek0.543.1Open source$0.229
8moonshotai0.742.0Open weight$2.859
9zai-org0.541.5Open weight$1.212
10openai0.641.4Closed$217
11meta0.741.2Closed$0.13
12xiaomi0.641.1Open source$0.4359
13anthropic0.640.9Closed$54
14anthropic0.740.1Closed$511
15deepseek0.739.3Open source$1.35