Mar 10, 2026 · you are here
Grok 4.20 Multi-agentNot publicly ranked · Price not listed
Model profile · xAI
Data as of July 28, 2026 · How the score is built
Each documented value carries its source. Missing fields stay visible as not sourced or not published, rather than disappearing from the page.
The sequence follows explicit supersedes links. Scores and prices remain blank when the corresponding public row or first-party rate is unavailable.
Mar 10, 2026 · you are here
Grok 4.20 Multi-agentNot publicly ranked · Price not listed
Multi-agent
The visual layer above carries the decisions. These notes preserve the model, ranking, coverage, and family context behind the numbers.
We track Grok 4.20 Multi-agent, but no sourced benchmark result is published on the site yet. This page shows the metadata we can verify now; score-level coverage will appear when public evaluations land.
Grok 4.20 Multi-agent is a proprietary model with a 2M context window. It uses an explicit reasoning mode, which can improve complex problem solving while adding latency and token use.
Tracked as unresolved. xAI's current docs use the Grok 4.20 Multi-agent name, but there is still no clean exact-value benchmark table and the capability docs still describe multi-agent as beta.
Grok 4.20 Multi-agent sits in the Grok 4.20 family with Grok 4.20. The profile has no source-displayable benchmark row yet.
Grok 4.20 Multi-agent does not have any source-displayable benchmark rows yet, so this profile does not assign a public score or rank. Documented specifications remain visible, while score-led charts and claims stay unavailable until a published evaluation can be attached to the exact model.
Grok 4.20 Multi-agent belongs to the Grok 4.20 family. Related tracked variants include Grok 4.20. A sibling link indicates shared lineage or a documented configuration relationship; it does not mean the variants have identical pricing, context limits, benchmark evidence, or deployment behavior. Compare before switching.
No. Grok 4.20 Multi-agent currently has 0 source-displayable rows across 369 tracked benchmark slots. The profile exposes published, non-generated evidence and leaves missing categories blank until an exact evaluation is available. Coverage describes how much was measured; it is not a penalty added to an individual benchmark result.
Grok 4.20 Multi-agent has a reported context window of 2M in the exact-model catalog record. The value stays visible, but the profile marks its source link as unavailable instead of presenting it as directly documented. Maximum output length remains separate because providers often publish a different limit.
Related resources
Last updated July 28, 2026. Runtime fields remain blank until a sourced snapshot exists.
The model to choose, the cheaper alternative, and the release we would wait on.