llmboard.ai
Benchmarks
CompareRankings
llmboard.ai
Benchmarks
CompareRankings
HomeModelsGrok 4 Heavy

xAI model product

Grok 4 Heavy

Grok 4 Heavy is the multi-agent version of Grok 4, released alongside the standard model in summer 2025. This system spawns multiple Grok 4 agents in parallel that work independently on problems and then collaborate by comparing their solutions, similar to a study group. The agents share insights and tricks they discover, with the system intelligently combining their work rather than simply using majority voting. Grok 4 Heavy uses approximately 10x more test-time compute than regular Grok 4, enabling it to solve significantly more complex problems. On the Humanities Last Exam, it achieves over 50% accuracy on text-only problems, and it scored a perfect result on the AIME 2025 mathematics competition. The system represents a major advancement in multi-agent AI collaboration and reasoning capabilities.

Updated Aug 10, 2026. Default version: Grok-4 Heavy

Compare
LLMBoard scoreN/ANot scored
CoverageN/ANo published score
Context windowN/ATokens
Official input priceN/AOfficial price unavailable

On this page

  • Specification
  • Capability
  • Benchmarks
  • Arena
  • Pricing
  • Versions
  • About
  • Similar models
  • FAQ

Model specification

Structured fields from the published default version.

Version
Grok-4 Heavy
Released
Unknown
Knowledge cutoff
Dec 31, 2024
Parameters
N/A
Context window
N/A
Max output
N/A
Inputs
image, text
Outputs
text
Open weights
No
License
Proprietary

Capability profile

This profile uses the latest version under this unique model that has a calculated LLMBoard score. Arena and price are excluded.

No capability profile

No published version under this unique model currently has enough benchmark coverage to calculate a score.

Benchmark results

Published benchmark records for the scored version currently unavailable.

No benchmark rows

No published benchmark result is linked to the scored version.

Arena results

Preference and agent-evaluation signals from published Arena datasets.

No published Arena match

The default version has no published Arena rows, or its source alias has not been resolved.

Pricing

Official vendor API PAYG pricing is summarized first. The table then lists individual provider offerings without treating their minimum as the official price.

Official API
N/A
Official provider
N/A
Lowest third-party
N/A
Tracked offerings
0
No published price snapshot

The default version has no provider offering with current input or output token prices.

Official prices use only the vendor's configured official Provider and positive standard USD PAYG rates. Third-party offers remain explicitly labeled.

Grok 4 Heavy versions

All published versions linked to this unique model. The score columns identify the version used by the current overall ranking.

1 rows
Columns

Show columns

Grok-4 HeavyN/AN/AN/AN/AN/ANoProprietary

What is Grok 4 Heavy?

A concise description based on the published model registry.

Grok 4 Heavy is the multi-agent version of Grok 4, released alongside the standard model in summer 2025. This system spawns multiple Grok 4 agents in parallel that work independently on problems and then collaborate by comparing their solutions, similar to a study group. The agents share insights and tricks they discover, with the system intelligently combining their work rather than simply using majority voting. Grok 4 Heavy uses approximately 10x more test-time compute than regular Grok 4, enabling it to solve significantly more complex problems. On the Humanities Last Exam, it achieves over 50% accuracy on text-only problems, and it scored a perfect result on the AIME 2025 mathematics competition. The system represents a major advancement in multi-agent AI collaboration and reasoning capabilities.

Use the benchmark, Arena and pricing sections above as separate evidence. A missing field means the current data snapshot does not support that claim.

Data snapshot: 2026-08-07. Editorial model content is not available in the backend.

Models similar to Grok 4 Heavy

Recommendations prioritize the same model type and family, then the closest published LLMBoard score.

#156
XA

Grok 1.5

xAI

19.3 LLMBoard

DetailsCompare
#37
XA

Grok 4 Fast

xAI

66.2 LLMBoard

DetailsCompare
#170
AC

Qwen3.5 0.8B

Alibaba Cloud / Qwen Team

0.3 LLMBoard

DetailsCompare
#169
GO

Gemma 3 1B

Google

2.7 LLMBoard

DetailsCompare
#168
GO

Gemma 3n E2B

Google

6.9 LLMBoard

DetailsCompare
#167
AC

Qwen3 VL 4B

Alibaba Cloud / Qwen Team

7.7 LLMBoard

DetailsCompare

FAQ

Common questions about Grok 4 Heavy.

When was Grok 4 Heavy released?

The current registry does not publish a release date for the default version.

How much does Grok 4 Heavy cost?

No official standard PAYG price is currently available for Grok 4 Heavy.

Who created Grok 4 Heavy?

Grok 4 Heavy is published under xAI in the model registry.

What is the context window for Grok 4 Heavy?

The current registry does not publish a context window for the default version.

Is Grok 4 Heavy open weight?

No. The default version is not marked as having publicly available weights.

How many API providers offer Grok 4 Heavy?

No published provider offering is currently linked to the default version.

What models should I compare Grok 4 Heavy with?

No nearby ranked alternatives are currently available.

Rankings

OverallCodingText ArenaPricing

Modalities

Image GenerationVideo GenerationSpeech-to-TextEmbeddings

Benchmarks

All BenchmarksReasoningMathCoding

Vendors

All VendorsOpenAIAnthropicGoogle
llmboard.aiCopyright 2026 llmboard.ai