Grok 4.20 Multi-Agent is tuned for collaborative agentic workflows while keeping the same 2M-token context window and multimodal support. Content policy rejections can still be charged: SpaceXAI may pass through a $0.05 moderation-failure fee, or a $0.055 usage-guidelines violation fee, depending on which rejection upstream returns.
Added Mar 31, 2026
Context Window
2.0M
Max Output
131.1K
Avg output tokens (7d)
8.1K tokens
Input Price (Auto)
$1.25/1M
Output Price (Auto)
$2.50/1M
Cache Read (Auto)
$0.20/1M
Capabilities
Benchmarks
Benchmarks
Performance metrics and benchmarks
Sourced from LMArena.
Arena Score
1470.2
Overall Rank
#36 / 394
Votes
60,864
Confidence Interval
1466.4 - 1474.0
Category Scores
Coding
#48 / 389
16,914 votes
1508.0
Math
#55 / 377
3,223 votes
1452.8
Longer Query
#66 / 372
25,623 votes
1456.7
Creative Writing
#36 / 392
10,225 votes
1447.9
Instruction Following
#61 / 394
20,377 votes
1443.4
Hard Prompts
#47 / 394
39,477 votes
1483.5
Additional Categories22
Polish
#24 / 217
1,275 votes
1482.1
French
#26 / 273
2,167 votes
1492.8
Russian
#27 / 357
6,306 votes
1479.2
German
#28 / 294
1,023 votes
1474.6
Korean
#29 / 265
1,000 votes
1434.5
Non English
#33 / 394
32,448 votes
1460.3
Spanish
#35 / 274
1,921 votes
1464.1
Exclude Ties
#37 / 394
45,979 votes
1477.0
Industry Entertainment And Sports And Media
#39 / 392
12,959 votes
1440.6
Industry Software And It Services
#40 / 394
24,055 votes
1502.0
English
#42 / 394
28,415 votes
1472.9
Multi Turn
#43 / 392
10,019 votes
1473.5
Industry Medicine And Healthcare
#46 / 363
4,482 votes
1480.6
Industry Life And Physical And Social Science
#47 / 392
9,913 votes
1480.4
Industry Legal And Government
#48 / 367
4,873 votes
1471.4
Chinese
#52 / 364
3,239 votes
1498.4
Industry Writing And Literature And Language
#52 / 393
14,774 votes
1444.8
Hard Prompts English
#57 / 393
19,286 votes
1480.8
Industry Business And Management And Financial Operations
#57 / 387
12,190 votes
1455.9
Industry Mathematical
#57 / 371
3,253 votes
1457.3
Expert
#59 / 344
6,033 votes
1479.4
Japanese
#60 / 257
576 votes
1416.8
Published 2026-08-21 · Matched as grok-4.20-multi-agent-beta-0309
LMArena DatasetProviders
Choose explicit providers for this model. Auto routing remains available as the default option.
Loading provider options…
Related text models
Compare Grok 4.20 Multi-Agent with similar models from the same provider or model family.
Grok 4.6
x-ai/grok-4.6Grok 4.6 is SpaceXAI's frontier reasoning model for coding, knowledge work, and STEM. It supports text, image, and file inputs with tool calling and structured outputs. Reasoning is always active, and requests above 200k input tokens use higher long-context rates. Content policy rejections can still be charged: SpaceXAI may pass through a $0.05 moderation-failure fee, or a $0.055 usage-guidelines violation fee, depending on which rejection upstream returns.
Grok 4.5
x-ai/grok-4.5Grok 4.5 is SpaceXAI's frontier reasoning model for coding, knowledge work, and STEM. It supports text, image, and file inputs with tool calling and structured outputs. Reasoning is always active. Content policy rejections can still be charged: SpaceXAI may pass through a $0.05 moderation-failure fee, or a $0.055 usage-guidelines violation fee, depending on which rejection upstream returns.
Grok Build 0.1
x-ai/grok-build-0.1Grok Build 0.1 is SpaceXAI's fast coding model trained specifically for agentic software engineering workflows. It supports text and image inputs with text output, and is optimized for interactive coding agents, tool use, and multi-step development tasks. Currently in early access. Content policy rejections can still be charged: SpaceXAI may pass through a $0.05 moderation-failure fee, or a $0.055 usage-guidelines violation fee, depending on which rejection upstream returns.
Grok Latest
x-ai/grok-latestCompatibility alias that routes to the newest Grok model. Currently routes to Grok 4.6. Content policy rejections can still be charged: SpaceXAI may pass through a $0.05 moderation-failure fee, or a $0.055 usage-guidelines violation fee, depending on which rejection upstream returns.
Grok 4.3
x-ai/grok-4.3Grok 4.3 is SpaceXAI's reasoning model for text and image inputs, built for agentic workflows, instruction following, factual accuracy, long-document analysis, and deep research. Reasoning is always active and requests above 200k total tokens are charged at the higher long-context rate. Content policy rejections can still be charged: SpaceXAI may pass through a $0.05 moderation-failure fee, or a $0.055 usage-guidelines violation fee, depending on which rejection upstream returns.
Grok 4.20
x-ai/grok-4.20SpaceXAI's Grok 4.20 flagship release with tool calling, multimodal input support, and a 2M-token context window. Content policy rejections can still be charged: SpaceXAI may pass through a $0.05 moderation-failure fee, or a $0.055 usage-guidelines violation fee, depending on which rejection upstream returns.