Skip to main content

Model profile · xAI

Grok 4.20 Multi-agent

CurrentReleased Mar 10, 2026ProprietaryReasoning2M context
Grok 4.20 Multi-agent is tracked, but not publicly ranked yet. The profile exposes 0 sourced benchmark rows and leaves unsupported fields blank until a published record exists.

Data as of July 28, 2026 · How the score is built

Spec sheet

Each documented value carries its source. Missing fields stay visible as not sourced or not published, rather than disappearing from the page.

API model ID
Not published
Context window
2M
Maximum output
Not sourced yet
Knowledge cutoff
Not sourced yet
Input modalities
Not sourced yet
Output modalities
Not sourced yet
Parameters
Not disclosed by the provider
Availability
Not sourced yet
Cloud regions
Not tracked yet
Lifecycle
Current
API capabilities
Tool calling, structured outputs, and batch support are not tracked yet
Prompt caching
Not documented in the pricing record
Self-host
Weights are not published
Rate limits
Not tracked yet

Lineage

The sequence follows explicit supersedes links. Scores and prices remain blank when the corresponding public row or first-party rate is unavailable.

Multi-agent

How to read this profile

The visual layer above carries the decisions. These notes preserve the model, ranking, coverage, and family context behind the numbers.

We track Grok 4.20 Multi-agent, but no sourced benchmark result is published on the site yet. This page shows the metadata we can verify now; score-level coverage will appear when public evaluations land.

Grok 4.20 Multi-agent is a proprietary model with a 2M context window. It uses an explicit reasoning mode, which can improve complex problem solving while adding latency and token use.

Tracked as unresolved. xAI's current docs use the Grok 4.20 Multi-agent name, but there is still no clean exact-value benchmark table and the capability docs still describe multi-agent as beta.

Grok 4.20 Multi-agent sits in the Grok 4.20 family with Grok 4.20. The profile has no source-displayable benchmark row yet.

Frequently asked questions

How does Grok 4.20 Multi-agent perform overall in AI benchmarks?

Grok 4.20 Multi-agent does not have any source-displayable benchmark rows yet, so this profile does not assign a public score or rank. Documented specifications remain visible, while score-led charts and claims stay unavailable until a published evaluation can be attached to the exact model.

Which sibling models are related to Grok 4.20 Multi-agent?

Grok 4.20 Multi-agent belongs to the Grok 4.20 family. Related tracked variants include Grok 4.20. A sibling link indicates shared lineage or a documented configuration relationship; it does not mean the variants have identical pricing, context limits, benchmark evidence, or deployment behavior. Compare before switching.

Does Grok 4.20 Multi-agent have full benchmark coverage on BenchLM?

No. Grok 4.20 Multi-agent currently has 0 source-displayable rows across 369 tracked benchmark slots. The profile exposes published, non-generated evidence and leaves missing categories blank until an exact evaluation is available. Coverage describes how much was measured; it is not a penalty added to an individual benchmark result.

What is the context window size of Grok 4.20 Multi-agent?

Grok 4.20 Multi-agent has a reported context window of 2M in the exact-model catalog record. The value stays visible, but the profile marks its source link as unavailable instead of presenting it as directly documented. Maximum output length remains separate because providers often publish a different limit.

Last updated July 28, 2026. Runtime fields remain blank until a sourced snapshot exists.

Make the right model choice this week

The model to choose, the cheaper alternative, and the release we would wait on.