# The frontier is outrunning the scorecards

Sent July 23, 2026

GPT-5.6 became generally available on July 9. Kimi K3 arrived seven days later. In between, another frontier lab released a 975-billion-parameter open-weight model. The release cycle is compressing, and the next model can now arrive before the last one has enough independent evidence to compare cleanly.

## What changed in seven days

- [GPT-5.6 Sol](https://benchlm.ai/models/gpt-5-6-sol): Available in the API at $5 input and $30 output per million tokens. The family adds lower-cost Terra and Luna tiers, plus a multi-agent ultra mode.
- [Kimi K3](https://benchlm.ai/models/kimi-3): Available in Kimi products and the API with a 1M-token window. API pricing is $3 input and $15 output per million cache-miss tokens.

## The limit

Kimi K3 was still unranked when this issue was sent. Its launch table mixed model and harness effects, and the promised weights were due by July 27. Official results were useful launch evidence, not a clean substitute for independent runs in a comparable setup.

## What to expect next

- Families, not single flagships: Labs can refresh capability tiers on separate schedules, which makes a model name less durable than its exact version and effort setting.
- More agent-dependent scores: The host model, harness, reasoning budget, and parallelism increasingly determine the result together.
- Shorter buying windows: Keep a small production fixture and compare cost per completed task. A general leaderboard can narrow the field; it cannot reproduce your stack.

## Analysis worth opening

- [Kimi K3’s Official Benchmark Table](https://benchlm.ai/blog/posts/kimi-3-release-data-coming-soon): What the launch results established, what they did not, and why the model remained unranked at send time.
- [Thinking Machines Chose Open Weights First](https://benchlm.ai/blog/posts/thinking-machines-chose-open-weights-first): Why a new frontier lab made its 975B-parameter first release portable, and where the hardware bill set a practical boundary.
- [AI Coding Agents Need Receipts](https://benchlm.ai/blog/posts/ai-coding-agents): The test a coding-agent ranking should pass before it tells anyone what to buy.

Archive copy reflects the rankings, prices, and availability stated when this issue was sent. Current pages may show newer evidence.

Canonical page: https://benchlm.ai/newsletter/issues/2026-07-23
