Grok 4.20 Multi-Agent is tuned for collaborative agentic workflows while keeping the same 2M-token context window and multimodal support. Content policy rejections can still be charged: SpaceXAI may pass through a $0.05 moderation-failure fee, or a $0.055 usage-guidelines violation fee, depending on which rejection upstream returns.
Added Mar 31, 2026
Context Window
2.0M
Max Output
131.1K
Avg output tokens (7d)
7.4K tokens
Input Price (Auto)
$1.25/1M
Output Price (Auto)
$2.50/1M
Cache Read (Auto)
$0.20/1M
Capabilities
Benchmarks
Benchmarks
Performance metrics and benchmarks
Sourced from LMArena.
Arena Score
1470.3
Overall Rank
#43 / 409
Votes
64,674
Confidence Interval
1466.6 - 1474.0
Category Scores
Coding
#57 / 404
18,015 votes
1508.5
Math
#62 / 393
3,382 votes
1455.2
Longer Query
#75 / 387
27,886 votes
1457.1
Creative Writing
#44 / 407
11,073 votes
1446.6
Instruction Following
#73 / 409
21,971 votes
1443.5
Hard Prompts
#55 / 409
42,162 votes
1483.2
Additional Categories22
Polish
#29 / 229
1,314 votes
1481.0
French
#34 / 287
2,278 votes
1492.1
German
#34 / 306
1,055 votes
1474.4
Korean
#34 / 276
1,045 votes
1434.0
Russian
#36 / 373
7,166 votes
1477.4
Non English
#42 / 409
35,388 votes
1459.4
Spanish
#43 / 292
2,036 votes
1463.9
Exclude Ties
#44 / 409
48,862 votes
1477.2
Industry Entertainment And Sports And Media
#49 / 407
13,968 votes
1439.2
Multi Turn
#49 / 407
10,630 votes
1473.4
English
#50 / 409
29,285 votes
1473.2
Industry Software And It Services
#50 / 409
25,579 votes
1502.0
Industry Legal And Government
#52 / 382
5,185 votes
1470.1
Industry Medicine And Healthcare
#52 / 379
4,846 votes
1481.1
Industry Life And Physical And Social Science
#57 / 407
10,692 votes
1479.3
Industry Writing And Literature And Language
#64 / 408
15,946 votes
1442.9
Japanese
#66 / 271
619 votes
1419.2
Industry Business And Management And Financial Operations
#67 / 402
12,921 votes
1454.5
Industry Mathematical
#67 / 390
3,549 votes
1456.7
Expert
#69 / 359
6,544 votes
1479.9
Hard Prompts English
#70 / 407
19,915 votes
1480.5
Chinese
#75 / 385
4,055 votes
1492.6
Published 2026-09-25 · Matched as grok-4.20-multi-agent-beta-0309
LMArena DatasetProviders
Choose explicit providers for this model. Auto routing remains available as the default option.
Loading provider options…
Related text models
Compare Grok 4.20 Multi-Agent with similar models from the same provider or model family.
Grok 4.7
x-ai/grok-4.7Grok 4.7 is SpaceXAI's flagship reasoning model for coding, long-running agentic tasks, and knowledge work. It supports text, image, and file inputs with tool calling and structured outputs. Reasoning is always active, and requests above 200k input tokens use higher long-context rates. Content policy rejections can still be charged: SpaceXAI may pass through a $0.05 moderation-failure fee, or a $0.055 usage-guidelines violation fee, depending on which rejection upstream returns.
Grok 4.6
x-ai/grok-4.6Grok 4.6 is SpaceXAI's frontier reasoning model for coding, knowledge work, and STEM. It supports text, image, and file inputs with tool calling and structured outputs. Reasoning is always active, and requests above 200k input tokens use higher long-context rates. Content policy rejections can still be charged: SpaceXAI may pass through a $0.05 moderation-failure fee, or a $0.055 usage-guidelines violation fee, depending on which rejection upstream returns.
Grok 4.5
x-ai/grok-4.5Grok 4.5 is SpaceXAI's frontier reasoning model for coding, knowledge work, and STEM. It supports text, image, and file inputs with tool calling and structured outputs. Reasoning is always active. Content policy rejections can still be charged: SpaceXAI may pass through a $0.05 moderation-failure fee, or a $0.055 usage-guidelines violation fee, depending on which rejection upstream returns.
Grok Build 0.1
x-ai/grok-build-0.1Grok Build 0.1 is SpaceXAI's fast coding model trained specifically for agentic software engineering workflows. It supports text and image inputs with text output, and is optimized for interactive coding agents, tool use, and multi-step development tasks. Currently in early access. Content policy rejections can still be charged: SpaceXAI may pass through a $0.05 moderation-failure fee, or a $0.055 usage-guidelines violation fee, depending on which rejection upstream returns.
Grok Latest
x-ai/grok-latestCompatibility alias that routes to the newest Grok model. Currently routes to Grok 4.6. Content policy rejections can still be charged: SpaceXAI may pass through a $0.05 moderation-failure fee, or a $0.055 usage-guidelines violation fee, depending on which rejection upstream returns.
Grok 4.3
x-ai/grok-4.3Grok 4.3 is SpaceXAI's reasoning model for text and image inputs, built for agentic workflows, instruction following, factual accuracy, long-document analysis, and deep research. Reasoning is always active and requests above 200k total tokens are charged at the higher long-context rate. Content policy rejections can still be charged: SpaceXAI may pass through a $0.05 moderation-failure fee, or a $0.055 usage-guidelines violation fee, depending on which rejection upstream returns.