Best Non-Reasoning LLMs in 2026
As of September 22, 2026, the top model in best non-reasoning llms on the BenchLM leaderboard is Claude Opus 4.7 with a score of 66.4.
Bottom line: Non-reasoning models are faster and cheaper than chain-of-thought alternatives. Gemini 3.1 Pro leads this tier — proving that strong reasoning scores are possible without dedicated thinking tokens.
This ranking moves when models ship. Get the releases, price changes and retirements that affect your shortlist. Follow model changes
Full Rankings (72 models)
What changed
Gemini 3.1 Pro leads non-reasoning models — best reasoning (97), knowledge (96), and multilingual (100).
Claude Opus 4.6 most consistent non-reasoning model across all 8 categories.
Claude Sonnet 4.6 strong mid-tier with best multimodal (95) in this tier.
How to choose
Key Takeaways
The top model is Claude Opus 4.7 by Anthropic with a BenchAlign v5.6 score of 66.4 and Estimated evidence.
The best open-weight model is Inkling-Small at position #5.
72 models are included in this ranking.
Score in Context
What these scores mean
Non-reasoning models are standard completion/chat models without dedicated chain-of-thought. They are ranked by the same overall BenchLM score and are typically faster and cheaper per token.
Known limitations
The "non-reasoning" label excludes models with explicit chain-of-thought (like o3, DeepSeek R1). Some non-reasoning models still reason internally — the distinction is about architecture and pricing, not capability.
About this ranking
Last verified: September 22, 2026
Top standard AI models (no chain-of-thought reasoning) ranked by benchmark performance. Faster and cheaper than reasoning models.
Unless noted otherwise, ranking surfaces on this page use BenchLM’s provisional leaderboard lane rather than the stricter sourced-only verified leaderboard.
Claude Opus 4.7 leads this ranking with a score of 66.4, followed by Claude Opus 4.6 (64.3) and Gemini 3 Pro (61.3). There is meaningful separation between the top models, suggesting genuine performance differences.
The best open-weight option is Inkling-Small (ranked #5 with a score of 55.9). While proprietary models lead, open-weight options are within striking distance for teams willing to trade a few points of performance for full model control.
This ranking uses provisional overall weighted scores from the active scoring formula. For detailed model profiles, click any model name above. To compare two specific models head-to-head, use the "vs #" links.
Explore More
Know when it’s worth switching models
The model to choose, the cheaper alternative, and the release we would wait on.
Read a sample issueJoin 2,000+ readers.
One email each week. Unsubscribe anytime.