Grok 4.20 Multi-Agent is tuned for collaborative agentic workflows while keeping the same 2M-token context window and multimodal support. Content policy rejections can still be charged: SpaceXAI may pass through a $0.05 moderation-failure fee, or a $0.055 usage-guidelines violation fee, depending on which rejection upstream returns.
Added Mar 31, 2026
Context Window
2.0M
Max Output
131.1K
Avg output tokens (7d)
7.6K tokens
Input Price (Auto)
$1.25/1M
Output Price (Auto)
$2.50/1M
Cache Read (Auto)
$0.20/1M
Capabilities
Benchmarks
Benchmarks
Performance metrics and benchmarks
Sourced from LMArena.
Arena Score
1470.2
Overall Rank
#39 / 401
Votes
60,775
Confidence Interval
1466.4 - 1474.0
Category Scores
Coding
#50 / 396
16,883 votes
1508.1
Math
#63 / 384
3,221 votes
1452.8
Longer Query
#68 / 379
25,586 votes
1456.7
Creative Writing
#39 / 399
10,213 votes
1447.9
Instruction Following
#66 / 401
20,349 votes
1443.3
Hard Prompts
#50 / 401
39,416 votes
1483.5
Additional Categories22
Polish
#25 / 222
1,284 votes
1480.7
German
#27 / 299
1,024 votes
1476.5
French
#29 / 281
2,178 votes
1491.2
Korean
#29 / 269
1,012 votes
1436.4
Russian
#30 / 365
6,466 votes
1479.0
Non English
#35 / 401
32,959 votes
1459.8
Spanish
#38 / 283
1,965 votes
1463.5
Exclude Ties
#40 / 401
45,918 votes
1477.1
Industry Software And It Services
#42 / 401
24,008 votes
1502.2
Industry Entertainment And Sports And Media
#43 / 399
12,941 votes
1441.0
English
#45 / 401
27,815 votes
1473.0
Industry Medicine And Healthcare
#47 / 369
4,479 votes
1480.5
Multi Turn
#48 / 399
10,025 votes
1473.3
Industry Legal And Government
#49 / 373
4,866 votes
1471.1
Industry Life And Physical And Social Science
#49 / 399
9,904 votes
1480.0
Industry Writing And Literature And Language
#55 / 400
14,761 votes
1444.8
Chinese
#57 / 372
3,395 votes
1495.5
Industry Business And Management And Financial Operations
#59 / 394
12,174 votes
1455.8
Japanese
#59 / 265
587 votes
1420.6
Industry Mathematical
#61 / 378
3,257 votes
1457.5
Hard Prompts English
#62 / 399
18,839 votes
1480.7
Expert
#64 / 350
6,024 votes
1479.1
Published 2026-09-11 · Matched as grok-4.20-multi-agent-beta-0309
LMArena DatasetProviders
Choose explicit providers for this model. Auto routing remains available as the default option.
Loading provider options…
Related text models
Compare Grok 4.20 Multi-Agent with similar models from the same provider or model family.
Grok 4.6
x-ai/grok-4.6Grok 4.6 is SpaceXAI's frontier reasoning model for coding, knowledge work, and STEM. It supports text, image, and file inputs with tool calling and structured outputs. Reasoning is always active, and requests above 200k input tokens use higher long-context rates. Content policy rejections can still be charged: SpaceXAI may pass through a $0.05 moderation-failure fee, or a $0.055 usage-guidelines violation fee, depending on which rejection upstream returns.
Grok 4.5
x-ai/grok-4.5Grok 4.5 is SpaceXAI's frontier reasoning model for coding, knowledge work, and STEM. It supports text, image, and file inputs with tool calling and structured outputs. Reasoning is always active. Content policy rejections can still be charged: SpaceXAI may pass through a $0.05 moderation-failure fee, or a $0.055 usage-guidelines violation fee, depending on which rejection upstream returns.
Grok Build 0.1
x-ai/grok-build-0.1Grok Build 0.1 is SpaceXAI's fast coding model trained specifically for agentic software engineering workflows. It supports text and image inputs with text output, and is optimized for interactive coding agents, tool use, and multi-step development tasks. Currently in early access. Content policy rejections can still be charged: SpaceXAI may pass through a $0.05 moderation-failure fee, or a $0.055 usage-guidelines violation fee, depending on which rejection upstream returns.
Grok Latest
x-ai/grok-latestCompatibility alias that routes to the newest Grok model. Currently routes to Grok 4.6. Content policy rejections can still be charged: SpaceXAI may pass through a $0.05 moderation-failure fee, or a $0.055 usage-guidelines violation fee, depending on which rejection upstream returns.
Grok 4.3
x-ai/grok-4.3Grok 4.3 is SpaceXAI's reasoning model for text and image inputs, built for agentic workflows, instruction following, factual accuracy, long-document analysis, and deep research. Reasoning is always active and requests above 200k total tokens are charged at the higher long-context rate. Content policy rejections can still be charged: SpaceXAI may pass through a $0.05 moderation-failure fee, or a $0.055 usage-guidelines violation fee, depending on which rejection upstream returns.
Grok 4.20
x-ai/grok-4.20SpaceXAI's Grok 4.20 flagship release with tool calling, multimodal input support, and a 2M-token context window. Content policy rejections can still be charged: SpaceXAI may pass through a $0.05 moderation-failure fee, or a $0.055 usage-guidelines violation fee, depending on which rejection upstream returns.