Skip to main content
Radar

Five or fewer confirmed AI changes, with original sources, on mornings when something changed.A free source-linked morning brief.

Start free brief

LLM Price vs Performance Chart

Compare the active public score against current API pricing. The shortlist favors the cheapest model that keeps at least 90% of the leader’s score; the efficiency frontier shows every undominated tradeoff.

30 confirmed releases in the last 30 daysStart free brief
90% shortlist

Grok 4.5

Score 75.4 · $6.00/1M · 88% below the leader’s price

Highest Score

Claude Mythos 5

Score: 83.2 · $50.00/1M out

Availability restrictions apply

Price floor

Ministral 3 3B

Score: 17.8 · $0.10/1M out

Score Axis
Cost Basis
Source Type
Price Range
Model Families

80 priced models match these filters. The blended view assumes three input tokens for every output token.

Efficiency Frontier

Lowest-cost score leaders (Overall)

Score/$ is shown as a narrow comparison aid, not a universal value verdict.

RankModelScore / cost / value
1
Ministral 3 3BMistral · frontier tradeoff
17.8 score$0.10 / 1M out178.0 score/$
2
Ministral 3 8BMistral · frontier tradeoff
20.3 score$0.15 / 1M out135.3 score/$
3
Ministral 3 14BMistral · frontier tradeoff
33.9 score$0.20 / 1M out169.5 score/$
4
Step 3.5 FlashStepFun · frontier tradeoff
54.3 score$0.30 / 1M out181.0 score/$
5
DeepSeek V3.2DeepSeek · frontier tradeoff
54.7 score$0.42 / 1M out130.2 score/$
7
MiniMax M3MiniMax · frontier tradeoff
68.7 score$1.20 / 1M out57.3 score/$
8
Grok 4.5xAI · within 90% of leader
75.4 score$6.00 / 1M out12.6 score/$
9
Gemini 3.6 FlashGoogle · within 90% of leader
75.5 score$7.50 / 1M out10.1 score/$
10
Kimi K3Moonshot AI · within 90% of leader
80.5 score$15.00 / 1M out5.4 score/$

Frequently Asked Questions

What is the LLM price-performance chart?

This chart plots each AI model by its benchmark score (vertical axis) against its API output price per million tokens (horizontal axis). Models in the upper-left quadrant offer the best value — high performance at low cost. The efficiency frontier line connects the best-value models at each price point.

What is the efficiency frontier?

The efficiency frontier (Pareto frontier) connects models where no other model offers both a higher score and a lower price. Models on this line represent the optimal price-performance tradeoff. If a model is below and to the right of the frontier, there exists a cheaper model with a better score.

Which LLM has the best price-to-performance ratio?

Grok 4.5 is the lowest-priced model that retains at least 90% of the current overall leader's score under these filters. We use this threshold instead of declaring the largest Score/$ ratio the universal winner, because normalized benchmark points are an index rather than units of completed work.

How are scores calculated?

The chart reads the active public ranking lane for each category and pairs that score with current catalog pricing. Overall, coding, and agentic views use the current BenchAlign methodology when enabled; other categories use the provisional composite. The 90% shortlist is a decision aid, not a claim that benchmark points convert directly into dollars.

Keep the value shortlist current

One weekly email when price changes or new benchmark evidence alter the models worth shortlisting.

Read a sample issue

Join 2,000+ readers.

One email each week. Unsubscribe anytime.