Skip to main content
Radar

Five or fewer confirmed AI changes, with original sources, on mornings when something changed.A free source-linked morning brief.

Start free brief

Citable dataset

LLM Pricing Statistics (2026)

Updated August 15, 2026 · Auto-generated from BenchLM's live dataset on every data refresh

Frontier LLM token prices are 88% below their March 2023 level as of August 2026, per the BenchLM Token Price Index (12 on a base of 100).

Frontier price drop since March 2023

Frontier LLM token prices are 88% below their March 2023 level as of August 2026, per the BenchLM Token Price Index (12 on a base of 100).

−88% (index: 12)

Price direction, last 12 months (frontier vs mid-tier)

Over the 12 months to August 2026, the BenchLM Token Price Index shows frontier LLM token prices up 36.4% year-over-year, while mid-tier prices went down 35.8%.

frontier +36.4%, mid-tier -35.8%

Models with tracked API pricing

As of August 15, 2026, BenchLM tracks live API pricing for 145 AI models.

145

Median price per 1M tokens (input / output)

As of August 15, 2026, the median LLM API price across 145 models tracked by BenchLM is $1.00 per 1M input tokens and $3.75 per 1M output tokens.

$1.00 / $3.75

Spread between cheapest and most expensive model

As of August 15, 2026, the most expensive LLM API tracked by BenchLM (o1-pro) costs roughly 4773x more per blended 1M tokens than the cheapest (Qwen3.7 Flash).

4773x

Cheapest frontier-tier model (top 10 overall)

As of August 15, 2026, the cheapest model in BenchLM's overall top 10 is Gemini 3.6 Flash at $1.50 per 1M input tokens and $7.50 per 1M output tokens.

Gemini 3.6 Flash ($1.50 in / $7.50 out)

Open-weight median discount vs proprietary

As of August 15, 2026, open-weight models on BenchLM have a median blended API price 78% lower than proprietary models ($0.53 vs $2.41 per 1M tokens at a 3:1 input:output ratio).

78% cheaper

Methodology & sources

Prices are USD per 1M tokens from BenchLM's pricing dataset (145 models, updated August 15, 2026). "Blended" price assumes a 3:1 input:output token ratio. "Frontier" means the top 10 models on BenchLM's overall ranking. Time-series figures come from the BenchLM Token Price Index (median blended price of flagship models, base March 2023 = 100).

Cite these statistics

Every number on this page is generated from BenchLM's live dataset and refreshed with each data update. Link any statistic directly using its anchor, or cite the page as:

BenchLM.ai, "LLM Statistics" (August 15, 2026), https://benchlm.ai/stats/llm-pricing

Frequently Asked Questions

How much does an LLM API cost per million tokens?

As of August 15, 2026, the median price across 145 models tracked by BenchLM is $1.00 per 1M input tokens and $3.75 per 1M output tokens, with roughly a 4773x spread between the cheapest and most expensive models.

Are open-weight models cheaper than proprietary models?

Yes. As of August 15, 2026, open-weight models tracked by BenchLM have a median blended API price 78% lower than proprietary models ($0.53 vs $2.41 per 1M tokens).

What is the cheapest frontier-quality model?

As of August 15, 2026, the cheapest model in BenchLM's overall top 10 is Gemini 3.6 Flash, at $1.50 per 1M input tokens and $7.50 per 1M output tokens.

How much have LLM token prices dropped since GPT-4?

Frontier LLM token prices are 88% below their March 2023 (GPT-4 launch) level as of August 2026, per the BenchLM Token Price Index — though the last 12 months moved up, not down, at the frontier.

Know when it’s worth switching models

The model to choose, the cheaper alternative, and the release we would wait on.

Read a sample issue

Join 2,000+ readers.

One email each week. Unsubscribe anytime.