Grok 4.20 Multi-Agent is tuned for collaborative agentic workflows while keeping the same 2M-token context window and multimodal support. Content policy rejections can still be charged: SpaceXAI may pass through a $0.05 moderation-failure fee, or a $0.055 usage-guidelines violation fee, depending on which rejection upstream returns.
Added Mar 31, 2026
Context Window
2.0M
Max Output
131.1K
Avg output tokens (7d)
8.1K tokens
Input Price (Auto)
$1.25/1M
Output Price (Auto)
$2.50/1M
Cache Read (Auto)
$0.20/1M
Capabilities
Benchmarks
Benchmarks
Performance metrics and benchmarks
Sourced from LMArena.
Arena Score
1470.5
Overall Rank
#35 / 395
Votes
60,845
Confidence Interval
1466.7 - 1474.3
Category Scores
Coding
#50 / 390
16,902 votes
1508.1
Math
#57 / 379
3,230 votes
1452.8
Longer Query
#66 / 373
25,618 votes
1457.0
Creative Writing
#35 / 393
10,223 votes
1448.0
Instruction Following
#63 / 395
20,379 votes
1443.4
Hard Prompts
#48 / 395
39,474 votes
1483.6
Additional Categories22
German
#23 / 296
1,029 votes
1476.9
Polish
#25 / 219
1,286 votes
1480.4
Korean
#26 / 265
1,012 votes
1435.3
Russian
#27 / 359
6,468 votes
1479.7
French
#28 / 276
2,190 votes
1491.6
Non English
#33 / 395
32,987 votes
1460.1
Spanish
#35 / 277
1,968 votes
1463.9
Exclude Ties
#37 / 395
45,954 votes
1477.4
Industry Entertainment And Sports And Media
#40 / 393
12,962 votes
1441.0
Industry Software And It Services
#41 / 395
24,040 votes
1502.1
English
#42 / 395
27,857 votes
1473.1
Industry Medicine And Healthcare
#44 / 363
4,472 votes
1480.6
Multi Turn
#44 / 393
10,019 votes
1473.1
Industry Life And Physical And Social Science
#46 / 393
9,913 votes
1480.3
Industry Legal And Government
#48 / 368
4,869 votes
1471.4
Industry Writing And Literature And Language
#51 / 394
14,775 votes
1445.0
Chinese
#55 / 367
3,409 votes
1496.0
Industry Business And Management And Financial Operations
#56 / 388
12,191 votes
1456.0
Japanese
#57 / 261
587 votes
1421.9
Hard Prompts English
#59 / 393
18,866 votes
1480.7
Industry Mathematical
#60 / 374
3,277 votes
1457.1
Expert
#61 / 345
6,034 votes
1479.0
Published 2026-08-27 · Matched as grok-4.20-multi-agent-beta-0309
LMArena DatasetProviders
Choose explicit providers for this model. Auto routing remains available as the default option.
Loading provider options…
Related text models
Compare Grok 4.20 Multi-Agent with similar models from the same provider or model family.
Grok 4.6
x-ai/grok-4.6Grok 4.6 is SpaceXAI's frontier reasoning model for coding, knowledge work, and STEM. It supports text, image, and file inputs with tool calling and structured outputs. Reasoning is always active, and requests above 200k input tokens use higher long-context rates. Content policy rejections can still be charged: SpaceXAI may pass through a $0.05 moderation-failure fee, or a $0.055 usage-guidelines violation fee, depending on which rejection upstream returns.
Grok 4.5
x-ai/grok-4.5Grok 4.5 is SpaceXAI's frontier reasoning model for coding, knowledge work, and STEM. It supports text, image, and file inputs with tool calling and structured outputs. Reasoning is always active. Content policy rejections can still be charged: SpaceXAI may pass through a $0.05 moderation-failure fee, or a $0.055 usage-guidelines violation fee, depending on which rejection upstream returns.
Grok Build 0.1
x-ai/grok-build-0.1Grok Build 0.1 is SpaceXAI's fast coding model trained specifically for agentic software engineering workflows. It supports text and image inputs with text output, and is optimized for interactive coding agents, tool use, and multi-step development tasks. Currently in early access. Content policy rejections can still be charged: SpaceXAI may pass through a $0.05 moderation-failure fee, or a $0.055 usage-guidelines violation fee, depending on which rejection upstream returns.
Grok Latest
x-ai/grok-latestCompatibility alias that routes to the newest Grok model. Currently routes to Grok 4.6. Content policy rejections can still be charged: SpaceXAI may pass through a $0.05 moderation-failure fee, or a $0.055 usage-guidelines violation fee, depending on which rejection upstream returns.
Grok 4.3
x-ai/grok-4.3Grok 4.3 is SpaceXAI's reasoning model for text and image inputs, built for agentic workflows, instruction following, factual accuracy, long-document analysis, and deep research. Reasoning is always active and requests above 200k total tokens are charged at the higher long-context rate. Content policy rejections can still be charged: SpaceXAI may pass through a $0.05 moderation-failure fee, or a $0.055 usage-guidelines violation fee, depending on which rejection upstream returns.
Grok 4.20
x-ai/grok-4.20SpaceXAI's Grok 4.20 flagship release with tool calling, multimodal input support, and a 2M-token context window. Content policy rejections can still be charged: SpaceXAI may pass through a $0.05 moderation-failure fee, or a $0.055 usage-guidelines violation fee, depending on which rejection upstream returns.
