Grok 4.20 Multi-Agent is tuned for collaborative agentic workflows while keeping the same 2M-token context window and multimodal support. Content policy rejections can still be charged: SpaceXAI may pass through a $0.05 moderation-failure fee, or a $0.055 usage-guidelines violation fee, depending on which rejection upstream returns.
Added Mar 31, 2026
Context Window
2.0M
Max Output
131.1K
Avg output tokens (7d)
8.1K tokens
Input Price (Auto)
$1.25/1M
Output Price (Auto)
$2.50/1M
Cache Read (Auto)
$0.20/1M
Capabilities
Benchmarks
Benchmarks
Performance metrics and benchmarks
Sourced from LMArena.
Arena Score
1470.5
Overall Rank
#33 / 389
Votes
60,964
Confidence Interval
1466.7 - 1474.3
Category Scores
Coding
#45 / 384
16,935 votes
1508.3
Math
#55 / 374
3,224 votes
1453.0
Longer Query
#61 / 367
25,683 votes
1457.4
Creative Writing
#32 / 387
10,243 votes
1448.5
Instruction Following
#57 / 389
20,392 votes
1444.0
Hard Prompts
#45 / 389
39,544 votes
1483.8
Additional Categories22
German
#24 / 293
1,021 votes
1475.1
French
#25 / 271
2,182 votes
1492.9
Polish
#26 / 215
1,293 votes
1481.5
Russian
#27 / 352
6,323 votes
1478.8
Korean
#29 / 264
1,002 votes
1434.1
Spanish
#30 / 272
1,898 votes
1465.0
Non English
#31 / 389
32,512 votes
1460.4
Exclude Ties
#34 / 389
46,052 votes
1477.4
Industry Software And It Services
#36 / 389
24,082 votes
1502.4
English
#37 / 389
28,451 votes
1473.3
Industry Entertainment And Sports And Media
#38 / 387
12,987 votes
1441.3
Multi Turn
#40 / 387
10,033 votes
1473.6
Industry Medicine And Healthcare
#42 / 357
4,495 votes
1480.4
Industry Legal And Government
#43 / 361
4,888 votes
1471.3
Industry Life And Physical And Social Science
#44 / 387
9,942 votes
1480.8
Chinese
#49 / 360
3,249 votes
1498.1
Industry Writing And Literature And Language
#51 / 388
14,801 votes
1445.2
Industry Business And Management And Financial Operations
#52 / 382
12,190 votes
1456.1
Hard Prompts English
#55 / 388
19,322 votes
1481.2
Industry Mathematical
#55 / 368
3,251 votes
1457.3
Expert
#56 / 339
6,036 votes
1479.4
Japanese
#60 / 256
575 votes
1416.1
Published 2026-08-12 · Matched as grok-4.20-multi-agent-beta-0309
LMArena DatasetProviders
Choose explicit providers for this model. Auto routing remains available as the default option.
Loading provider options…
Related text models
Compare Grok 4.20 Multi-Agent with similar models from the same provider or model family.
Grok 4.6
x-ai/grok-4.6Grok 4.6 is SpaceXAI's frontier reasoning model for coding, knowledge work, and STEM. It supports text, image, and file inputs with tool calling and structured outputs. Reasoning is always active, and requests above 200k input tokens use higher long-context rates. Content policy rejections can still be charged: SpaceXAI may pass through a $0.05 moderation-failure fee, or a $0.055 usage-guidelines violation fee, depending on which rejection upstream returns.
Grok 4.5
x-ai/grok-4.5Grok 4.5 is SpaceXAI's frontier reasoning model for coding, knowledge work, and STEM. It supports text, image, and file inputs with tool calling and structured outputs. Reasoning is always active. Content policy rejections can still be charged: SpaceXAI may pass through a $0.05 moderation-failure fee, or a $0.055 usage-guidelines violation fee, depending on which rejection upstream returns.
Grok Build 0.1
x-ai/grok-build-0.1Grok Build 0.1 is SpaceXAI's fast coding model trained specifically for agentic software engineering workflows. It supports text and image inputs with text output, and is optimized for interactive coding agents, tool use, and multi-step development tasks. Currently in early access. Content policy rejections can still be charged: SpaceXAI may pass through a $0.05 moderation-failure fee, or a $0.055 usage-guidelines violation fee, depending on which rejection upstream returns.
Grok Latest
x-ai/grok-latestCompatibility alias that routes to the newest Grok model. Currently routes to Grok 4.6. Content policy rejections can still be charged: SpaceXAI may pass through a $0.05 moderation-failure fee, or a $0.055 usage-guidelines violation fee, depending on which rejection upstream returns.
Grok 4.3
x-ai/grok-4.3Grok 4.3 is SpaceXAI's reasoning model for text and image inputs, built for agentic workflows, instruction following, factual accuracy, long-document analysis, and deep research. Reasoning is always active and requests above 200k total tokens are charged at the higher long-context rate. Content policy rejections can still be charged: SpaceXAI may pass through a $0.05 moderation-failure fee, or a $0.055 usage-guidelines violation fee, depending on which rejection upstream returns.
Grok 4.20
x-ai/grok-4.20SpaceXAI's Grok 4.20 flagship release with tool calling, multimodal input support, and a 2M-token context window. Content policy rejections can still be charged: SpaceXAI may pass through a $0.05 moderation-failure fee, or a $0.055 usage-guidelines violation fee, depending on which rejection upstream returns.