llmboard.ai
Benchmarks
CompareRankings
llmboard.ai
Benchmarks
CompareRankings
HomeModelsGrok 4.1

xAI model product

Grok 4.1

Grok 4.1 (Non-Thinking) brings significant improvements to the real-world usability of Grok. The model is exceptionally capable in creative, emotional, and collaborative interactions. It is more perceptive to nuanced intent, compelling to speak with, and coherent in personality, while fully retaining the razor-sharp intelligence and reliability of its predecessors. Grok 4.1 uses no thinking tokens for immediate responses, making it faster while maintaining high quality. The model features reduced hallucinations compared to previous versions, with significant improvements in factual accuracy for information-seeking prompts. To achieve this, xAI used large scale reinforcement learning infrastructure to optimize style, personality, helpfulness, and alignment, developing new methods that use frontier agentic reasoning models as reward models to autonomously evaluate and iterate on responses at scale. Grok 4.1 includes comprehensive safety mitigations including refusal training, input filters for restricted knowledge, and adversarial robustness measures. The model demonstrates strong refusal rates on harmful queries (5% answer rate on chat refusals, 4% on agentic refusals) and improved honesty training to reduce deception.

Updated Aug 10, 2026. Default version: Grok-4.1

Compare
LLMBoard scoreN/ANot scored
CoverageN/ANo published score
Context window256KTokens
Official input priceN/AOfficial price unavailable

On this page

  • Specification
  • Capability
  • Benchmarks
  • Arena
  • Pricing
  • Versions
  • About
  • Similar models
  • FAQ

Model specification

Structured fields from the published default version.

Version
Grok-4.1
Released
Nov 17, 2025
Knowledge cutoff
Unknown
Parameters
N/A
Context window
256K
Max output
8K
Inputs
image, text
Outputs
text
Open weights
No
License
Proprietary

Capability profile

This profile uses the latest version under this unique model that has a calculated LLMBoard score. Arena and price are excluded.

No capability profile

No published version under this unique model currently has enough benchmark coverage to calculate a score.

Benchmark results

Published benchmark records for the scored version currently unavailable.

No benchmark rows

No published benchmark result is linked to the scored version.

Arena results

Preference and agent-evaluation signals from published Arena datasets.

86 rows
Columns

Show columns

text factualitykorean151406.7471N/AAug 6, 2026
text factualitygerman161423.8664N/AAug 6, 2026
text style controlkorean251435.31,093N/AAug 6, 2026
text style controlpolish291473.31,929N/AAug 6, 2026
text factualitypolish341416.41,341N/AAug 6, 2026
text style controlindustry medicine and healthcare341485.34,347N/AAug 6, 2026
textpolish371448.51,929N/AAug 6, 2026
text style controlgerman381464.71,196N/AAug 6, 2026
text style controlchinese401503.03,629N/AAug 6, 2026
text style controlindustry business and management and financial operations401462.512,736N/AAug 6, 2026
text style controlindustry legal and government401472.64,870N/AAug 6, 2026
text style controljapanese401430.5612N/AAug 6, 2026
textkorean421408.71,093N/AAug 6, 2026
text style controlenglish441469.431,456N/AAug 6, 2026
text style controlexclude ties441465.648,720N/AAug 6, 2026
text style controloverall461459.567,246N/AAug 6, 2026
textgerman481444.51,196N/AAug 6, 2026
text style controlnon english491446.535,777N/AAug 6, 2026
text style controlfrench511476.21,654N/AAug 6, 2026
text style controlindustry life and physical and social science511478.210,672N/AAug 6, 2026
text style controlcreative writing521432.09,702N/AAug 6, 2026
text style controlindustry entertainment and sports and media541429.412,499N/AAug 6, 2026
text factualityfrench551448.01,114N/AAug 6, 2026
text style controlhard prompts571474.538,104N/AAug 6, 2026
textjapanese581401.2612N/AAug 6, 2026
text style controlindustry software and it services581488.724,526N/AAug 6, 2026
text style controlmulti turn591462.712,257N/AAug 6, 2026
textcreative writing601412.49,702N/AAug 6, 2026
text factualityspanish601414.71,152N/AAug 6, 2026
text style controlrussian601448.46,468N/AAug 6, 2026
textindustry legal and government611444.04,870N/AAug 6, 2026
text style controlhard prompts english611477.418,663N/AAug 6, 2026
text factualitycreative writing631417.76,395N/AAug 6, 2026
textenglish641446.331,456N/AAug 6, 2026
textindustry entertainment and sports and media641406.212,499N/AAug 6, 2026
text style controlspanish641444.41,725N/AAug 6, 2026
text factualitychinese651473.32,427N/AAug 6, 2026
textexclude ties661431.648,720N/AAug 6, 2026
text style controlindustry writing and literature and language661426.814,740N/AAug 6, 2026
textoverall671436.667,246N/AAug 6, 2026
textnon english681423.635,777N/AAug 6, 2026
text factualityindustry medicine and healthcare681463.72,775N/AAug 6, 2026
text style controlcoding681492.415,554N/AAug 6, 2026
textchinese691473.63,629N/AAug 6, 2026
text factualityindustry legal and government691449.63,138N/AAug 6, 2026
text style controllonger query691450.519,804N/AAug 6, 2026
textrussian701425.86,468N/AAug 6, 2026
textmulti turn721439.912,257N/AAug 6, 2026
textfrench741445.71,654N/AAug 6, 2026
text style controlinstruction following741431.218,548N/AAug 6, 2026
textindustry life and physical and social science751444.110,672N/AAug 6, 2026
text factualityindustry entertainment and sports and media751409.18,579N/AAug 6, 2026
textindustry business and management and financial operations761424.412,736N/AAug 6, 2026
textindustry medicine and healthcare771444.54,347N/AAug 6, 2026
text factualityindustry writing and literature and language781417.810,246N/AAug 6, 2026
text factualityrussian781429.34,028N/AAug 6, 2026
textindustry software and it services801450.824,526N/AAug 6, 2026
texthard prompts811436.438,104N/AAug 6, 2026
textspanish821423.81,725N/AAug 6, 2026
text factualitymath821407.72,624N/AAug 6, 2026
text factualityindustry mathematical831401.32,280N/AAug 6, 2026
textindustry writing and literature and language841407.414,740N/AAug 6, 2026
text style controlmath861427.84,195N/AAug 6, 2026
texthard prompts english871439.818,663N/AAug 6, 2026
text factualityexclude ties871430.347,902N/AAug 6, 2026
text factualitymulti turn871436.18,525N/AAug 6, 2026
text style controlexpert871454.54,591N/AAug 6, 2026
text factualityindustry life and physical and social science881451.07,044N/AAug 6, 2026
textcoding891444.415,554N/AAug 6, 2026
text factualitynon english891418.335,055N/AAug 6, 2026
text factualityoverall941430.666,313N/AAug 6, 2026
textlonger query961416.519,804N/AAug 6, 2026
text factualityexpert961430.72,982N/AAug 6, 2026
text factualitycoding971469.311,129N/AAug 6, 2026
text factualityenglish981440.530,584N/AAug 6, 2026
text factualityindustry software and it services1011459.020,764N/AAug 6, 2026
text factualityhard prompts1021446.037,335N/AAug 6, 2026
text factualityindustry business and management and financial operations1021424.98,718N/AAug 6, 2026
text factualitylonger query1021430.115,201N/AAug 6, 2026
text factualityhard prompts english1031448.813,967N/AAug 6, 2026
text factualityinstruction following1031409.813,965N/AAug 6, 2026
textinstruction following1061400.418,548N/AAug 6, 2026
textmath1081417.44,195N/AAug 6, 2026
text style controlindustry mathematical1081423.43,339N/AAug 6, 2026
textexpert1151413.84,591N/AAug 6, 2026
textindustry mathematical1251407.33,339N/AAug 6, 2026

Pricing

Official vendor API PAYG pricing is summarized first. The table then lists individual provider offerings without treating their minimum as the official price.

Official API
N/A
Official provider
N/A
Lowest third-party
From $2 input, $10 output per 1M via 302.AI
Tracked offerings
1
1 rows
Columns

Show columns

302.AIgrok-4.1global$2$10200KAug 7, 2026

Official prices use only the vendor's configured official Provider and positive standard USD PAYG rates. Third-party offers remain explicitly labeled.

Grok 4.1 versions

All published versions linked to this unique model. The score columns identify the version used by the current overall ranking.

1 rows
Columns

Show columns

Grok-4.1Nov 17, 2025N/AN/A256K8KNoProprietary

What is Grok 4.1?

A concise description based on the published model registry.

Grok 4.1 (Non-Thinking) brings significant improvements to the real-world usability of Grok. The model is exceptionally capable in creative, emotional, and collaborative interactions. It is more perceptive to nuanced intent, compelling to speak with, and coherent in personality, while fully retaining the razor-sharp intelligence and reliability of its predecessors. Grok 4.1 uses no thinking tokens for immediate responses, making it faster while maintaining high quality. The model features reduced hallucinations compared to previous versions, with significant improvements in factual accuracy for information-seeking prompts. To achieve this, xAI used large scale reinforcement learning infrastructure to optimize style, personality, helpfulness, and alignment, developing new methods that use frontier agentic reasoning models as reward models to autonomously evaluate and iterate on responses at scale. Grok 4.1 includes comprehensive safety mitigations including refusal training, input filters for restricted knowledge, and adversarial robustness measures. The model demonstrates strong refusal rates on harmful queries (5% answer rate on chat refusals, 4% on agentic refusals) and improved honesty training to reduce deception.

Use the benchmark, Arena and pricing sections above as separate evidence. A missing field means the current data snapshot does not support that claim.

Data snapshot: 2026-08-07. Editorial model content is not available in the backend.

Models similar to Grok 4.1

Recommendations prioritize the same model type and family, then the closest published LLMBoard score.

#156
XA

Grok 1.5

xAI

19.3 LLMBoard

DetailsCompare
#37
XA

Grok 4 Fast

xAI

66.2 LLMBoard

DetailsCompare
#170
AC

Qwen3.5 0.8B

Alibaba Cloud / Qwen Team

0.3 LLMBoard

DetailsCompare
#169
GO

Gemma 3 1B

Google

2.7 LLMBoard

DetailsCompare
#168
GO

Gemma 3n E2B

Google

6.9 LLMBoard

DetailsCompare
#167
AC

Qwen3 VL 4B

Alibaba Cloud / Qwen Team

7.7 LLMBoard

DetailsCompare

FAQ

Common questions about Grok 4.1.

When was Grok 4.1 released?

Grok 4.1's default version was released on Nov 17, 2025.

How much does Grok 4.1 cost?

No official standard PAYG price is currently available for Grok 4.1. The lowest tracked third-party offer starts at $2 input and $10 output via 302.AI.

Who created Grok 4.1?

Grok 4.1 is published under xAI in the model registry.

What is the context window for Grok 4.1?

The default version has a 256K token context window.

Is Grok 4.1 open weight?

No. The default version is not marked as having publicly available weights.

How many API providers offer Grok 4.1?

1 published provider offerings are linked to the default version.

What models should I compare Grok 4.1 with?

No nearby ranked alternatives are currently available.

Rankings

OverallCodingText ArenaPricing

Modalities

Image GenerationVideo GenerationSpeech-to-TextEmbeddings

Benchmarks

All BenchmarksReasoningMathCoding

Vendors

All VendorsOpenAIAnthropicGoogle
llmboard.aiCopyright 2026 llmboard.ai