llmboard.ai
Benchmarks
CompareRankings
llmboard.ai
Benchmarks
CompareRankings
HomeModelsGrok 4.20 Multi Agent Beta

xAI model product

Grok 4.20 Multi Agent Beta

Grok 4.20 Multi-Agent Beta is xAI's multi-agent variant of the Grok 4.20 model family, designed for orchestrating and coordinating multiple AI agents in complex workflows. Released as a beta on March 9, 2026, it features a 2 million token context window and supports advanced multi-agent collaboration patterns.

Updated Aug 10, 2026. Default version: Grok-4.20 Multi-Agent Beta

Compare
LLMBoard scoreN/ANot scored
CoverageN/ANo published score
Context window2MTokens
Official input priceN/AOfficial price unavailable

On this page

  • Specification
  • Capability
  • Benchmarks
  • Arena
  • Pricing
  • Versions
  • About
  • Similar models
  • FAQ

Model specification

Structured fields from the published default version.

Version
Grok-4.20 Multi-Agent Beta
Released
Mar 9, 2026
Knowledge cutoff
Unknown
Parameters
N/A
Context window
2M
Max output
30K
Inputs
image, text
Outputs
text
Open weights
No
License
Proprietary

Capability profile

This profile uses the latest version under this unique model that has a calculated LLMBoard score. Arena and price are excluded.

No capability profile

No published version under this unique model currently has enough benchmark coverage to calculate a score.

Benchmark results

Published benchmark records for the scored version currently unavailable.

No benchmark rows

No published benchmark result is linked to the scored version.

Arena results

Preference and agent-evaluation signals from published Arena datasets.

100 rows
Columns

Show columns

search style controloverall61206.887,118N/AJul 21, 2026
searchoverall101205.587,118N/AJul 21, 2026
search factualityoverall121199.087,067N/AJul 21, 2026
text factualitygerman131430.0484N/AAug 6, 2026
visionentity recognition131267.4144N/AAug 6, 2026
vision style controlentity recognition141230.9144N/AAug 6, 2026
visioncreative writing vision161289.61,350N/AAug 6, 2026
vision style controlcreative writing vision171273.81,350N/AAug 6, 2026
text factualityjapanese181423.2390N/AAug 6, 2026
text factualitypolish201440.7813N/AAug 6, 2026
text style controlpolish211482.21,278N/AAug 6, 2026
text style controlgerman221475.81,010N/AAug 6, 2026
text style controlfrench231492.82,146N/AAug 6, 2026
text style controlkorean241435.9999N/AAug 6, 2026
visionhumor241275.5790N/AAug 6, 2026
text factualityfrench251484.01,627N/AAug 6, 2026
text factualityspanish251449.21,372N/AAug 6, 2026
vision style controlhumor251258.5790N/AAug 6, 2026
text style controlrussian261479.76,266N/AAug 6, 2026
text factualitycreative writing281446.08,446N/AAug 6, 2026
text style controlcreative writing281448.810,123N/AAug 6, 2026
text style controlspanish281465.61,862N/AAug 6, 2026
textpolish291456.71,278N/AAug 6, 2026
text style controlnon english301461.032,203N/AAug 6, 2026
textgerman311457.11,010N/AAug 6, 2026
textrussian321458.56,266N/AAug 6, 2026
text style controloverall321471.060,322N/AAug 6, 2026
text factualityrussian331463.64,973N/AAug 6, 2026
text style controlexclude ties331477.945,597N/AAug 6, 2026
textkorean351417.7999N/AAug 6, 2026
text factualitynon english351443.332,121N/AAug 6, 2026
text style controlindustry software and it services351503.023,774N/AAug 6, 2026
vision style controlchinese351296.31,248N/AAug 6, 2026
textcreative writing361437.410,123N/AAug 6, 2026
text factualityindustry medicine and healthcare361485.13,519N/AAug 6, 2026
text style controlindustry entertainment and sports and media361441.812,810N/AAug 6, 2026
text style controlenglish371473.628,118N/AAug 6, 2026
visionchinese371304.61,248N/AAug 6, 2026
textoverall381450.760,322N/AAug 6, 2026
text factualityoverall381456.960,212N/AAug 6, 2026
vision style controloverall381252.422,353N/AAug 6, 2026
textexclude ties391448.045,597N/AAug 6, 2026
textfrench391467.82,146N/AAug 6, 2026
textnon english391441.432,203N/AAug 6, 2026
text factualityexclude ties391458.345,512N/AAug 6, 2026
text style controlindustry legal and government391473.44,789N/AAug 6, 2026
text style controlindustry medicine and healthcare391482.24,414N/AAug 6, 2026
visionoverall401261.622,353N/AAug 6, 2026
vision style controlocr401260.915,804N/AAug 6, 2026
text factualityenglish411464.827,969N/AAug 6, 2026
text style controlmulti turn411473.99,895N/AAug 6, 2026
visionenglish411259.79,291N/AAug 6, 2026
vision style controlhomework411272.83,103N/AAug 6, 2026
text factualityindustry entertainment and sports and media421431.811,443N/AAug 6, 2026
text style controlindustry life and physical and social science421481.19,770N/AAug 6, 2026
visionhomework421274.73,103N/AAug 6, 2026
visionocr421262.715,804N/AAug 6, 2026
text factualitychinese431489.52,534N/AAug 6, 2026
text style controlcoding431509.216,757N/AAug 6, 2026
vision style controlenglish431247.29,291N/AAug 6, 2026
text factualityindustry legal and government441466.73,814N/AAug 6, 2026
text factualityindustry mathematical441444.02,512N/AAug 6, 2026
text factualityindustry writing and literature and language441442.513,008N/AAug 6, 2026
text style controlhard prompts441484.438,976N/AAug 6, 2026
vision style controldiagram441268.25,942N/AAug 6, 2026
textindustry entertainment and sports and media451423.912,810N/AAug 6, 2026
text factualitymath451441.52,454N/AAug 6, 2026
text factualitymulti turn461464.18,292N/AAug 6, 2026
textindustry writing and literature and language481432.314,578N/AAug 6, 2026
text factualityindustry life and physical and social science481473.58,194N/AAug 6, 2026
text style controlchinese481498.53,183N/AAug 6, 2026
visiondiagram481262.15,942N/AAug 6, 2026
textenglish491452.528,118N/AAug 6, 2026
text factualityhard prompts491472.838,873N/AAug 6, 2026
text factualityhard prompts english491481.317,705N/AAug 6, 2026
text factualityindustry software and it services491487.022,952N/AAug 6, 2026
text style controlindustry writing and literature and language491445.514,578N/AAug 6, 2026
textindustry legal and government501451.54,789N/AAug 6, 2026
textmulti turn501451.59,895N/AAug 6, 2026
textspanish501447.71,862N/AAug 6, 2026
textindustry life and physical and social science521456.59,770N/AAug 6, 2026
textindustry medicine and healthcare521453.84,414N/AAug 6, 2026
text factualityexpert521474.24,702N/AAug 6, 2026
text style controlindustry business and management and financial operations521456.811,994N/AAug 6, 2026
text factualitycoding531498.415,300N/AAug 6, 2026
text style controlexpert531481.45,772N/AAug 6, 2026
text style controlindustry mathematical531457.73,188N/AAug 6, 2026
text style controlhard prompts english551481.719,026N/AAug 6, 2026
text style controlinstruction following551444.419,938N/AAug 6, 2026
text style controlmath551453.23,207N/AAug 6, 2026
textindustry software and it services561461.623,774N/AAug 6, 2026
textchinese571480.73,183N/AAug 6, 2026
textjapanese571402.4551N/AAug 6, 2026
text style controljapanese571416.8551N/AAug 6, 2026
texthard prompts581448.938,976N/AAug 6, 2026
text factualityinstruction following601440.018,501N/AAug 6, 2026
text style controllonger query601458.125,028N/AAug 6, 2026
text factualityindustry business and management and financial operations631447.310,199N/AAug 6, 2026
textmath641439.33,207N/AAug 6, 2026
text factualitylonger query641454.723,594N/AAug 6, 2026

Pricing

Official vendor API PAYG pricing is summarized first. The table then lists individual provider offerings without treating their minimum as the official price.

Official API
N/A
Official provider
N/A
Lowest third-party
From $1.25 input, $2.5 output per 1M via Vercel AI Gateway
Tracked offerings
2
2 rows
Columns

Show columns

Vercel AI Gatewayxai/grok-4.20-multi-agent-betaglobal$1.25$2.52MAug 7, 2026
302.AIgrok-4.20-multi-agent-beta-0309global$2$62MAug 7, 2026

Official prices use only the vendor's configured official Provider and positive standard USD PAYG rates. Third-party offers remain explicitly labeled.

Grok 4.20 Multi Agent Beta versions

All published versions linked to this unique model. The score columns identify the version used by the current overall ranking.

1 rows
Columns

Show columns

Grok-4.20 Multi-Agent BetaMar 9, 2026N/AN/A2M30KNoProprietary

What is Grok 4.20 Multi Agent Beta?

A concise description based on the published model registry.

Grok 4.20 Multi-Agent Beta is xAI's multi-agent variant of the Grok 4.20 model family, designed for orchestrating and coordinating multiple AI agents in complex workflows. Released as a beta on March 9, 2026, it features a 2 million token context window and supports advanced multi-agent collaboration patterns.

Use the benchmark, Arena and pricing sections above as separate evidence. A missing field means the current data snapshot does not support that claim.

Data snapshot: 2026-08-07. Editorial model content is not available in the backend.

Models similar to Grok 4.20 Multi Agent Beta

Recommendations prioritize the same model type and family, then the closest published LLMBoard score.

#156
XA

Grok 1.5

xAI

19.3 LLMBoard

DetailsCompare
#37
XA

Grok 4 Fast

xAI

66.2 LLMBoard

DetailsCompare
#170
AC

Qwen3.5 0.8B

Alibaba Cloud / Qwen Team

0.3 LLMBoard

DetailsCompare
#169
GO

Gemma 3 1B

Google

2.7 LLMBoard

DetailsCompare
#168
GO

Gemma 3n E2B

Google

6.9 LLMBoard

DetailsCompare
#167
AC

Qwen3 VL 4B

Alibaba Cloud / Qwen Team

7.7 LLMBoard

DetailsCompare

FAQ

Common questions about Grok 4.20 Multi Agent Beta.

When was Grok 4.20 Multi Agent Beta released?

Grok 4.20 Multi Agent Beta's default version was released on Mar 9, 2026.

How much does Grok 4.20 Multi Agent Beta cost?

No official standard PAYG price is currently available for Grok 4.20 Multi Agent Beta. The lowest tracked third-party offer starts at $1.25 input and $2.5 output via Vercel AI Gateway.

Who created Grok 4.20 Multi Agent Beta?

Grok 4.20 Multi Agent Beta is published under xAI in the model registry.

What is the context window for Grok 4.20 Multi Agent Beta?

The default version has a 2M token context window.

Is Grok 4.20 Multi Agent Beta open weight?

No. The default version is not marked as having publicly available weights.

How many API providers offer Grok 4.20 Multi Agent Beta?

2 published provider offerings are linked to the default version.

What models should I compare Grok 4.20 Multi Agent Beta with?

No nearby ranked alternatives are currently available.

Rankings

OverallCodingText ArenaPricing

Modalities

Image GenerationVideo GenerationSpeech-to-TextEmbeddings

Benchmarks

All BenchmarksReasoningMathCoding

Vendors

All VendorsOpenAIAnthropicGoogle
llmboard.aiCopyright 2026 llmboard.ai