llmboard.ai
Benchmarks
CompareRankings
llmboard.ai
Benchmarks
CompareRankings
HomeModelsQwen3 235B A22B Thinking

Alibaba Cloud / Qwen Team model product

Qwen3 235B A22B Thinking

Qwen3-235B-A22B-Thinking-2507 is a state-of-the-art thinking-enabled Mixture-of-Experts (MoE) model with 235B total parameters (22B activated). It features 94 layers, 128 experts (8 activated), and supports 262K native context length. This version delivers significantly improved reasoning performance, achieving state-of-the-art results among open-source thinking models on logical reasoning, mathematics, science, coding, and academic benchmarks. Key enhancements include markedly better general capabilities (instruction following, tool usage, text generation), enhanced 256K long-context understanding, and increased thinking depth. The model supports only thinking mode with automatic <think> tag inclusion.

Updated Aug 10, 2026. Default version: Qwen3-235B-A22B-Thinking-2507

Compare
LLMBoard score65.7Qwen3-235B-A22B-Thinking-2507
Coverage80%11 benchmark families
Context window262.1KTokens
Official input priceN/AOfficial price unavailable

On this page

  • Specification
  • Capability
  • Benchmarks
  • Arena
  • Pricing
  • Versions
  • About
  • Compare
  • Similar models
  • FAQ

Model specification

Structured fields from the published default version.

Version
Qwen3-235B-A22B-Thinking-2507
Released
Jul 25, 2025
Knowledge cutoff
Unknown
Parameters
235B
Context window
262.1K
Max output
131.1K
Inputs
text
Outputs
text
Open weights
No
License
Apache 2.0

Capability profile

This profile uses the latest version under this unique model that has a calculated LLMBoard score. Arena and price are excluded.

Qwen3-235B-A22B-Thinking-2507 category scores

Benchmark results

Published benchmark records for the scored version Qwen3-235B-A22B-Thinking-2507.

25 rows
Columns

Show columns

CFEval2134.012100.0%CAug 7, 2026
Multi-IF0.8120100.0%CAug 7, 2026
WritingBench0.9115100.0%CAug 7, 2026
LiveBench 202411250.821492.3%CAug 7, 2026
Arena-Hard v20.831686.7%CAug 7, 2026
Creative Writing v30.931383.3%CAug 7, 2026
BFCL-v30.761972.2%CAug 7, 2026
OJBench0.36937.5%CAug 7, 2026
MMLU-Redux0.974887.2%CAug 7, 2026
Include0.883176.7%CAug 7, 2026
MMLU-ProX0.883277.4%CAug 7, 2026
PolyMATH0.682368.2%CAug 7, 2026
HMMT250.8112558.3%CAug 7, 2026
SuperGPQA0.6113469.7%CAug 7, 2026
Tau2 Airline0.6152336.4%CAug 7, 2026
Tau2 Retail0.7162640.0%CAug 7, 2026
TAU-bench Airline0.5172327.3%CAug 7, 2026
TAU-bench Retail0.7172533.3%CAug 7, 2026
LiveCodeBench v60.7255353.9%CAug 7, 2026
MMLU-Pro0.82512981.3%CAug 7, 2026
IFEval0.9286557.8%CAug 7, 2026
Tau2 Telecom0.5313511.8%CAug 7, 2026
AIME 20250.93711468.1%CAug 7, 2026
Humanity's Last Exam0.2609235.2%CAug 7, 2026
GPQA0.87923366.4%CAug 7, 2026

Arena results

Preference and agent-evaluation signals from published Arena datasets.

56 rows
Columns

Show columns

textexpert571456.0410N/AAug 6, 2026
textindustry medicine and healthcare641449.8519N/AAug 6, 2026
textchinese701473.1410N/AAug 6, 2026
textjapanese741383.7403N/AAug 6, 2026
textindustry mathematical761435.2491N/AAug 6, 2026
text style controlexpert791462.1410N/AAug 6, 2026
text style controljapanese901375.0403N/AAug 6, 2026
textindustry business and management and financial operations981411.21,544N/AAug 6, 2026
textgerman991405.9204N/AAug 6, 2026
textindustry life and physical and social science1001431.81,513N/AAug 6, 2026
textkorean1021360.5200N/AAug 6, 2026
textenglish1041427.04,364N/AAug 6, 2026
texthard prompts english1051429.42,004N/AAug 6, 2026
text style controlindustry mathematical1061424.0491N/AAug 6, 2026
textcreative writing1071386.71,075N/AAug 6, 2026
textexclude ties1121399.86,338N/AAug 6, 2026
text style controlgerman1121398.2204N/AAug 6, 2026
textoverall1131413.78,975N/AAug 6, 2026
textmath1131412.3486N/AAug 6, 2026
textindustry software and it services1141430.42,992N/AAug 6, 2026
text style controlchinese1141448.6410N/AAug 6, 2026
textindustry writing and literature and language1151387.61,866N/AAug 6, 2026
textnon english1171397.24,603N/AAug 6, 2026
textrussian1171398.1436N/AAug 6, 2026
text style controlkorean1171347.8200N/AAug 6, 2026
texthard prompts1181415.43,841N/AAug 6, 2026
textpolish1181385.9765N/AAug 6, 2026
textcoding1191423.81,612N/AAug 6, 2026
textmulti turn1201405.71,403N/AAug 6, 2026
text style controlindustry medicine and healthcare1201433.7519N/AAug 6, 2026
textindustry entertainment and sports and media1211369.61,536N/AAug 6, 2026
text style controlrussian1211402.1436N/AAug 6, 2026
textspanish1231390.4154N/AAug 6, 2026
textinstruction following1251385.12,099N/AAug 6, 2026
textindustry legal and government1261407.9539N/AAug 6, 2026
textlonger query1261397.71,616N/AAug 6, 2026
text style controlcreative writing1311373.61,075N/AAug 6, 2026
text style controlnon english1311389.84,603N/AAug 6, 2026
text style controlindustry business and management and financial operations1331402.51,544N/AAug 6, 2026
text style controlpolish1351374.5765N/AAug 6, 2026
text style controlhard prompts english1371426.92,004N/AAug 6, 2026
text style controlindustry software and it services1371432.72,992N/AAug 6, 2026
text style controloverall1381398.98,975N/AAug 6, 2026
text style controlexclude ties1391380.26,338N/AAug 6, 2026
text style controlspanish1391382.1154N/AAug 6, 2026
text style controlcoding1401441.51,612N/AAug 6, 2026
text style controlindustry writing and literature and language1401377.61,866N/AAug 6, 2026
text style controlinstruction following1401385.32,099N/AAug 6, 2026
text style controlhard prompts1411417.43,841N/AAug 6, 2026
text style controlmath1421397.2486N/AAug 6, 2026
text style controlindustry life and physical and social science1431410.81,513N/AAug 6, 2026
text style controllonger query1471400.61,616N/AAug 6, 2026
text style controlenglish1481404.54,364N/AAug 6, 2026
text style controlmulti turn1481393.01,403N/AAug 6, 2026
text style controlindustry entertainment and sports and media1521350.51,536N/AAug 6, 2026
text style controlindustry legal and government1611390.5539N/AAug 6, 2026

Pricing

Official vendor API PAYG pricing is summarized first. The table then lists individual provider offerings without treating their minimum as the official price.

Official API
N/A
Official provider
N/A
Lowest third-party
From $0.20 input, $0.60 output per 1M via submodel
Tracked offerings
11
10 rows
Columns

Show columns

ModelScopeQwen/Qwen3-235B-A22B-Thinking-2507globalN/AN/A262.1KAug 7, 2026
iFlowqwen3-235b-a22b-thinking-2507globalN/AN/A256KAug 7, 2026
submodelQwen/Qwen3-235B-A22B-Thinking-2507global$0.20$0.60262.1KAug 7, 2026
OpenRouterqwen/qwen3-235b-a22b-thinking-2507global$0.23$2.3262.1KAug 7, 2026
LLM Gatewayqwen3-235b-a22b-thinking-2507global$0.30$3262KAug 7, 2026
NovitaAIqwen/qwen3-235b-a22b-thinking-2507global$0.30$3131.1KAug 7, 2026
Hugging FaceQwen/Qwen3-235B-A22B-Thinking-2507global$0.30$3262.1KAug 7, 2026
Jiekou.AIqwen/qwen3-235b-a22b-thinking-2507global$0.30$3131.1KAug 7, 2026
Vercel AI Gatewayalibaba/qwen3-235b-a22b-thinkingglobal$0.40$4131.1KAug 7, 2026
Venice AIqwen3-235b-a22b-thinking-2507global$0.45$3.5128KAug 7, 2026

Official prices use only the vendor's configured official Provider and positive standard USD PAYG rates. Third-party offers remain explicitly labeled.

Qwen3 235B A22B Thinking versions

All published versions linked to this unique model. The score columns identify the version used by the current overall ranking.

1 rows
Columns

Show columns

Qwen3-235B-A22B-Thinking-2507Jul 25, 202565.7235B262.1K131.1KNoApache 2.0

What is Qwen3 235B A22B Thinking?

A concise description based on the published model registry.

Qwen3-235B-A22B-Thinking-2507 is a state-of-the-art thinking-enabled Mixture-of-Experts (MoE) model with 235B total parameters (22B activated). It features 94 layers, 128 experts (8 activated), and supports 262K native context length. This version delivers significantly improved reasoning performance, achieving state-of-the-art results among open-source thinking models on logical reasoning, mathematics, science, coding, and academic benchmarks. Key enhancements include markedly better general capabilities (instruction following, tool usage, text generation), enhanced 256K long-context understanding, and increased thinking depth. The model supports only thinking mode with automatic <think> tag inclusion.

Use the benchmark, Arena and pricing sections above as separate evidence. A missing field means the current data snapshot does not support that claim.

Data snapshot: 2026-08-07. Editorial model content is not available in the backend.

Qwen3 235B A22B Thinking vs nearby models

Open a comparison with the three ranked models immediately above and below this model.

Qwen3 235B A22B ThinkingvsGemma 4 26B A4BQwen3 235B A22B ThinkingvsGrok 4 FastQwen3 235B A22B ThinkingvsGLM 4.7Qwen3 235B A22B ThinkingvsGLM 5.1Qwen3 235B A22B ThinkingvsClaude Sonnet 3.7Qwen3 235B A22B ThinkingvsGemini 2.5 Flash

Models similar to Qwen3 235B A22B Thinking

Recommendations prioritize the same model type and family, then the closest published LLMBoard score.

#50-4.3
AC

Qwen3.5 122B A10B

Alibaba Cloud / Qwen Team

61.4 LLMBoard

DetailsCompare
#26+6.5
AC

Qwen3.6 27B

Alibaba Cloud / Qwen Team

72.2 LLMBoard

DetailsCompare
#25+6.7
AC

Qwen3.6 Plus

Alibaba Cloud / Qwen Team

72.4 LLMBoard

DetailsCompare
#57-7.3
AC

Qwen3 VL 235B A22B Thinking

Alibaba Cloud / Qwen Team

58.5 LLMBoard

DetailsCompare
#23+7.4
AC

Qwen3.7 Plus

Alibaba Cloud / Qwen Team

73.1 LLMBoard

DetailsCompare
#58-8.6
AC

Qwen3.5 27B

Alibaba Cloud / Qwen Team

57.1 LLMBoard

DetailsCompare

FAQ

Common questions about Qwen3 235B A22B Thinking.

When was Qwen3 235B A22B Thinking released?

Qwen3 235B A22B Thinking's default version was released on Jul 25, 2025.

How much does Qwen3 235B A22B Thinking cost?

No official standard PAYG price is currently available for Qwen3 235B A22B Thinking. The lowest tracked third-party offer starts at $0.20 input and $0.60 output via submodel.

Who created Qwen3 235B A22B Thinking?

Qwen3 235B A22B Thinking is published under Alibaba Cloud / Qwen Team in the model registry.

What is the context window for Qwen3 235B A22B Thinking?

The default version has a 262.1K token context window.

Is Qwen3 235B A22B Thinking open weight?

No. The default version is not marked as having publicly available weights.

How many API providers offer Qwen3 235B A22B Thinking?

11 published provider offerings are linked to the default version.

What models should I compare Qwen3 235B A22B Thinking with?

Nearby ranked alternatives include Gemma 4 26B A4B, Grok 4 Fast, GLM 4.7.

Rankings

OverallCodingText ArenaPricing

Modalities

Image GenerationVideo GenerationSpeech-to-TextEmbeddings

Benchmarks

All BenchmarksReasoningMathCoding

Vendors

All VendorsOpenAIAnthropicGoogle
llmboard.aiCopyright 2026 llmboard.ai