Model profile
Qwen3.8 Max Preview
Evidence coverage
0 of 321 tracked benchmarks are published. 0 are verified and 0 provisional. 0 of 8 categories are measured.
- Published / tracked
- 0 / 321
- Verified
- 0
- Provisional
- 0
- Categories with evidence
- 0 / 8
Evidence by category
- Agentic0 benchmarksNot measured
- Coding0 benchmarksNot measured
- Reasoning0 benchmarksNot measured
- Knowledge0 benchmarksNot measured
- Math0 benchmarksNot measured
- Multilingual0 benchmarksNot measured
- Multimodal0 benchmarksNot measured
- Inst. Following0 benchmarksNot measured
BenchLM is tracking Qwen3.8 Max Preview, but sourced benchmark results are not published on the site yet. This page currently shows the model metadata we can verify now, and score-level benchmark coverage will appear once public evaluations land.
Qwen3.8 Max Preview is a proprietary model. It uses explicit chain-of-thought reasoning, which typically improves performance on math and complex reasoning tasks at the cost of higher latency and token usage.
Preview access is available through Alibaba's Token Plan, Qoder, and QoderWork. Qwen has announced future open weights for the 2.4T-parameter Qwen3.8 family, but no weight files or release date are public yet.
BenchLM links it directly to Qwen3.7 Max as the earlier related model in that lineage. This profile does not yet have sourced benchmarks on BenchLM, so the benchmark sections below are intentionally marked as coming soon.
Peer position
Exact provisional scores and ranks for the closest listed peers. A score can appear before a model clears the evidence threshold for a rank, so equal scores can have different rank states.
Range 77.44–83.93
- Claude Mythos 5AnthropicCompare#183.93Claude Mythos 5 is #1 with a score of 83.93.
- Claude Fable 5AnthropicCompare#283.68Claude Fable 5 is #2 with a score of 83.68.
- GPT-5.6 SolOpenAICompare#381.96GPT-5.6 Sol is #3 with a score of 81.96.
- Kimi K3Moonshot AICompare#480.96Kimi K3 is #4 with a score of 80.96.
- Claude Opus 4.8AnthropicCompare#578.34Claude Opus 4.8 is #5 with a score of 78.34.
- Muse Spark 1.1MetaCompare#677.44Muse Spark 1.1 is #6 with a score of 77.44.
- Qwen3.8 Max PreviewCurrent modelAlibabaUnrankedNot measuredQwen3.8 Max Preview is Unranked with a score of Not measured.
Benchmark data coming soon
BenchLM has created this model profile, but no sourced benchmark rows are published here yet. Category charts, rankings, and score tables will appear after public evaluations are available and attached to source records.
Category evidence
Scores and ranks appear only where this model has published benchmark evidence. Categories without displayable source records remain not measured.
| Category | Score | Rank | Percentile | Weight | Benchmarks | Evidence |
|---|---|---|---|---|---|---|
| AgenticWeight 22%0 benchmarksNot measured | Not measured | Not ranked | Not available | 22% | 0 benchmarks | Not measured |
| CodingWeight 20%0 benchmarksNot measured | Not measured | Not ranked | Not available | 20% | 0 benchmarks | Not measured |
| ReasoningWeight 17%0 benchmarksNot measured | Not measured | Not ranked | Not available | 17% | 0 benchmarks | Not measured |
| KnowledgeWeight 12%0 benchmarksNot measured | Not measured | Not ranked | Not available | 12% | 0 benchmarks | Not measured |
| MathWeight 5%0 benchmarksNot measured | Not measured | Not ranked | Not available | 5% | 0 benchmarks | Not measured |
| MultilingualWeight 7%0 benchmarksNot measured | Not measured | Not ranked | Not available | 7% | 0 benchmarks | Not measured |
| MultimodalWeight 12%0 benchmarksNot measured | Not measured | Not ranked | Not available | 12% | 0 benchmarks | Not measured |
| Inst. FollowingWeight 5%0 benchmarksNot measured | Not measured | Not ranked | Not available | 5% | 0 benchmarks | Not measured |
Frequently Asked Questions
How does Qwen3.8 Max Preview perform overall in AI benchmarks?
We are tracking Qwen3.8 Max Preview, but sourced benchmark coverage is still coming soon. The profile lists only the model metadata that has a verified public source.
Does Qwen3.8 Max Preview have full benchmark coverage on BenchLM?
Not yet. Qwen3.8 Max Preview currently has 0 published benchmark scores out of the 321 benchmarks BenchLM tracks. BenchLM only exposes non-generated public benchmark rows, so missing categories stay blank until a sourced evaluation is available.
What is the context window size of Qwen3.8 Max Preview?
Qwen3.8 Max Preview's context window has not been published in a source we can verify yet. The profile leaves this field unavailable instead of borrowing a limit from an earlier model in the family.
Related Resources
Don't miss the next GPT moment
Which models moved up, what is new, and what it costs. One email each week.
Free. One email per week.