Benchmarks

Top models across a combined benchmark plus Artificial Analysis, LMArena, LiveBench, FrontierCode, Epoch AI, ARC Prize, EQ-Bench, Design Arena, and WiseOne AI benchmark categories.

Combined

Equal-weight blend of Artificial Analysis Intelligence Index, LMArena Overall, LiveBench Overall, NanoGPT Usage Share. Each source is min-max normalized to 0-100 across its current leaderboard and weighted at 25%. Missing or unavailable source entries contribute 0.

Top 20 price vs performance

X-axis: $/M blended tokens

1.

Claude Opus 5.5
Anthropic logo

by Anthropic

63.1%

2.

GPT 6 Astra
OpenAI logo

by OpenAI

37.6%

34.1%

4.

GPT 6.1 Sol
OpenAI logo

by OpenAI

33.1%

5.

Claude Fable 5.1
Anthropic logo

by Anthropic

31.4%

6.

Claude Fable 5 1 Effort

25.0%

7.

Gemini 4 Argon

Google logo

by Google

25.0%

8.

Claude Fable 5
Anthropic logo

by Anthropic

24.4%

9.

Claude Fable 5 Effort

23.0%

10.

Claude Sonnet 5.5
Anthropic logo

by Anthropic

22.1%

Weighted blend of latest source snapshots

NanoGPT Composite