Grok 4.20 Multi-Agent

x-ai/grok-4.20-multi-agent

Grok 4.20 Multi-Agent

x-ai/grok-4.20-multi-agent

Grok 4.20 Multi-Agent is tuned for collaborative agentic workflows while keeping the same 2M-token context window and multimodal support. Content policy rejections can still be charged: SpaceXAI may pass through a $0.05 moderation-failure fee, or a $0.055 usage-guidelines violation fee, depending on which rejection upstream returns.

Added Mar 31, 2026

Context Window

2.0M

Max Output

131.1K

Avg output tokens (7d)

8.1K tokens

98%

Input Price (Auto)

$1.25/1M

Output Price (Auto)

$2.50/1M

Cache Read (Auto)

$0.20/1M

Capabilities

Benchmarks

Performance metrics and benchmarks

Sourced from LMArena.

Arena Score

1470.5

Overall Rank

#33 / 389

Votes

60,964

Confidence Interval

1466.7 - 1474.3

Category Scores

Coding

#45 / 384

16,935 votes

1508.3

Math

#55 / 374

3,224 votes

1453.0

Longer Query

#61 / 367

25,683 votes

1457.4

Creative Writing

#32 / 387

10,243 votes

1448.5

Instruction Following

#57 / 389

20,392 votes

1444.0

Hard Prompts

#45 / 389

39,544 votes

1483.8

Additional Categories
22

German

#24 / 293

1,021 votes

1475.1

French

#25 / 271

2,182 votes

1492.9

Polish

#26 / 215

1,293 votes

1481.5

Russian

#27 / 352

6,323 votes

1478.8

Korean

#29 / 264

1,002 votes

1434.1

Spanish

#30 / 272

1,898 votes

1465.0

Non English

#31 / 389

32,512 votes

1460.4

Exclude Ties

#34 / 389

46,052 votes

1477.4

Industry Software And It Services

#36 / 389

24,082 votes

1502.4

English

#37 / 389

28,451 votes

1473.3

Industry Entertainment And Sports And Media

#38 / 387

12,987 votes

1441.3

Multi Turn

#40 / 387

10,033 votes

1473.6

Industry Medicine And Healthcare

#42 / 357

4,495 votes

1480.4

Industry Legal And Government

#43 / 361

4,888 votes

1471.3

Industry Life And Physical And Social Science

#44 / 387

9,942 votes

1480.8

Chinese

#49 / 360

3,249 votes

1498.1

Industry Writing And Literature And Language

#51 / 388

14,801 votes

1445.2

Industry Business And Management And Financial Operations

#52 / 382

12,190 votes

1456.1

Hard Prompts English

#55 / 388

19,322 votes

1481.2

Industry Mathematical

#55 / 368

3,251 votes

1457.3

Expert

#56 / 339

6,036 votes

1479.4

Japanese

#60 / 256

575 votes

1416.1

Published 2026-08-12 · Matched as grok-4.20-multi-agent-beta-0309

LMArena Dataset

Providers

Choose explicit providers for this model. Auto routing remains available as the default option.

Loading provider options…

Compare Grok 4.20 Multi-Agent with similar models from the same provider or model family.

Grok 4.6

x-ai/grok-4.6

Grok 4.6 is SpaceXAI's frontier reasoning model for coding, knowledge work, and STEM. It supports text, image, and file inputs with tool calling and structured outputs. Reasoning is always active, and requests above 200k input tokens use higher long-context rates. Content policy rejections can still be charged: SpaceXAI may pass through a $0.05 moderation-failure fee, or a $0.055 usage-guidelines violation fee, depending on which rejection upstream returns.

Grok 4.5

x-ai/grok-4.5

Grok 4.5 is SpaceXAI's frontier reasoning model for coding, knowledge work, and STEM. It supports text, image, and file inputs with tool calling and structured outputs. Reasoning is always active. Content policy rejections can still be charged: SpaceXAI may pass through a $0.05 moderation-failure fee, or a $0.055 usage-guidelines violation fee, depending on which rejection upstream returns.

Grok Build 0.1

x-ai/grok-build-0.1

Grok Build 0.1 is SpaceXAI's fast coding model trained specifically for agentic software engineering workflows. It supports text and image inputs with text output, and is optimized for interactive coding agents, tool use, and multi-step development tasks. Currently in early access. Content policy rejections can still be charged: SpaceXAI may pass through a $0.05 moderation-failure fee, or a $0.055 usage-guidelines violation fee, depending on which rejection upstream returns.

Grok Latest

x-ai/grok-latest

Compatibility alias that routes to the newest Grok model. Currently routes to Grok 4.6. Content policy rejections can still be charged: SpaceXAI may pass through a $0.05 moderation-failure fee, or a $0.055 usage-guidelines violation fee, depending on which rejection upstream returns.

Grok 4.3

x-ai/grok-4.3

Grok 4.3 is SpaceXAI's reasoning model for text and image inputs, built for agentic workflows, instruction following, factual accuracy, long-document analysis, and deep research. Reasoning is always active and requests above 200k total tokens are charged at the higher long-context rate. Content policy rejections can still be charged: SpaceXAI may pass through a $0.05 moderation-failure fee, or a $0.055 usage-guidelines violation fee, depending on which rejection upstream returns.

Grok 4.20

x-ai/grok-4.20

SpaceXAI's Grok 4.20 flagship release with tool calling, multimodal input support, and a 2M-token context window. Content policy rejections can still be charged: SpaceXAI may pass through a $0.05 moderation-failure fee, or a $0.055 usage-guidelines violation fee, depending on which rejection upstream returns.