Grok 4.20 Multi-Agent

x-ai/grok-4.20-multi-agent

Grok 4.20 Multi-Agent

x-ai/grok-4.20-multi-agent

Grok 4.20 Multi-Agent is tuned for collaborative agentic workflows while keeping the same 2M-token context window and multimodal support. Content policy rejections can still be charged: SpaceXAI may pass through a $0.05 moderation-failure fee, or a $0.055 usage-guidelines violation fee, depending on which rejection upstream returns.

Added Mar 31, 2026

Context Window

2.0M

Max Output

131.1K

Avg output tokens (7d)

8.1K tokens

98%

Input Price (Auto)

$1.25/1M

Output Price (Auto)

$2.50/1M

Cache Read (Auto)

$0.20/1M

Capabilities

Benchmarks

Performance metrics and benchmarks

Sourced from LMArena.

Arena Score

1470.5

Overall Rank

#35 / 395

Votes

60,845

Confidence Interval

1466.7 - 1474.3

Category Scores

Coding

#50 / 390

16,902 votes

1508.1

Math

#57 / 379

3,230 votes

1452.8

Longer Query

#66 / 373

25,618 votes

1457.0

Creative Writing

#35 / 393

10,223 votes

1448.0

Instruction Following

#63 / 395

20,379 votes

1443.4

Hard Prompts

#48 / 395

39,474 votes

1483.6

Additional Categories
22

German

#23 / 296

1,029 votes

1476.9

Polish

#25 / 219

1,286 votes

1480.4

Korean

#26 / 265

1,012 votes

1435.3

Russian

#27 / 359

6,468 votes

1479.7

French

#28 / 276

2,190 votes

1491.6

Non English

#33 / 395

32,987 votes

1460.1

Spanish

#35 / 277

1,968 votes

1463.9

Exclude Ties

#37 / 395

45,954 votes

1477.4

Industry Entertainment And Sports And Media

#40 / 393

12,962 votes

1441.0

Industry Software And It Services

#41 / 395

24,040 votes

1502.1

English

#42 / 395

27,857 votes

1473.1

Industry Medicine And Healthcare

#44 / 363

4,472 votes

1480.6

Multi Turn

#44 / 393

10,019 votes

1473.1

Industry Life And Physical And Social Science

#46 / 393

9,913 votes

1480.3

Industry Legal And Government

#48 / 368

4,869 votes

1471.4

Industry Writing And Literature And Language

#51 / 394

14,775 votes

1445.0

Chinese

#55 / 367

3,409 votes

1496.0

Industry Business And Management And Financial Operations

#56 / 388

12,191 votes

1456.0

Japanese

#57 / 261

587 votes

1421.9

Hard Prompts English

#59 / 393

18,866 votes

1480.7

Industry Mathematical

#60 / 374

3,277 votes

1457.1

Expert

#61 / 345

6,034 votes

1479.0

Published 2026-08-27 · Matched as grok-4.20-multi-agent-beta-0309

LMArena Dataset

Providers

Choose explicit providers for this model. Auto routing remains available as the default option.

Loading provider options…

Compare Grok 4.20 Multi-Agent with similar models from the same provider or model family.

Grok 4.6

x-ai/grok-4.6

Grok 4.6 is SpaceXAI's frontier reasoning model for coding, knowledge work, and STEM. It supports text, image, and file inputs with tool calling and structured outputs. Reasoning is always active, and requests above 200k input tokens use higher long-context rates. Content policy rejections can still be charged: SpaceXAI may pass through a $0.05 moderation-failure fee, or a $0.055 usage-guidelines violation fee, depending on which rejection upstream returns.

Grok 4.5

x-ai/grok-4.5

Grok 4.5 is SpaceXAI's frontier reasoning model for coding, knowledge work, and STEM. It supports text, image, and file inputs with tool calling and structured outputs. Reasoning is always active. Content policy rejections can still be charged: SpaceXAI may pass through a $0.05 moderation-failure fee, or a $0.055 usage-guidelines violation fee, depending on which rejection upstream returns.

Grok Build 0.1

x-ai/grok-build-0.1

Grok Build 0.1 is SpaceXAI's fast coding model trained specifically for agentic software engineering workflows. It supports text and image inputs with text output, and is optimized for interactive coding agents, tool use, and multi-step development tasks. Currently in early access. Content policy rejections can still be charged: SpaceXAI may pass through a $0.05 moderation-failure fee, or a $0.055 usage-guidelines violation fee, depending on which rejection upstream returns.

Grok Latest

x-ai/grok-latest

Compatibility alias that routes to the newest Grok model. Currently routes to Grok 4.6. Content policy rejections can still be charged: SpaceXAI may pass through a $0.05 moderation-failure fee, or a $0.055 usage-guidelines violation fee, depending on which rejection upstream returns.

Grok 4.3

x-ai/grok-4.3

Grok 4.3 is SpaceXAI's reasoning model for text and image inputs, built for agentic workflows, instruction following, factual accuracy, long-document analysis, and deep research. Reasoning is always active and requests above 200k total tokens are charged at the higher long-context rate. Content policy rejections can still be charged: SpaceXAI may pass through a $0.05 moderation-failure fee, or a $0.055 usage-guidelines violation fee, depending on which rejection upstream returns.

Grok 4.20

x-ai/grok-4.20

SpaceXAI's Grok 4.20 flagship release with tool calling, multimodal input support, and a 2M-token context window. Content policy rejections can still be charged: SpaceXAI may pass through a $0.05 moderation-failure fee, or a $0.055 usage-guidelines violation fee, depending on which rejection upstream returns.