llmboard.ai
Benchmarks
CompareRankings
llmboard.ai
Benchmarks
CompareRankings
HomeModelsQwen3 Next 80B A3B Thinking

Alibaba Cloud / Qwen Team model product

Qwen3 Next 80B A3B Thinking

Qwen3-Next-80B-A3B-Thinking is the thinking variant of the Qwen3-Next series, featuring the same groundbreaking architecture as the instruct model. Leveraging GSPO, it addresses stability and efficiency challenges of hybrid attention + high-sparsity MoE in RL training. It uses Hybrid Attention combining Gated DeltaNet and Gated Attention for efficient ultra-long context modeling, High-Sparsity MoE with 512 experts (10 activated + 1 shared), and Multi-Token Prediction. With 80B total parameters and only 3B activated, it demonstrates outstanding performance on complex reasoning tasks — outperforming Qwen3-30B-A3B-Thinking-2507, Qwen3-32B-Thinking, and even the proprietary Gemini-2.5-Flash-Thinking across multiple benchmarks. Architecture: 48 layers, 15T training tokens, hybrid layout of 12*(3*(Gated DeltaNet->MoE)->(Gated Attention->MoE)). Supports only thinking mode with automatic <think> tag inclusion, may generate longer thinking content.

Updated Aug 10, 2026. Default version: Qwen3-Next-80B-A3B-Thinking

Compare
LLMBoard score54.1Qwen3-Next-80B-A3B-Thinking
Coverage80%10 benchmark families
Context window131.1KTokens
Official input price$0.144Alibaba (China) API

On this page

  • Specification
  • Capability
  • Benchmarks
  • Arena
  • Pricing
  • Versions
  • About
  • Compare
  • Similar models
  • FAQ

Model specification

Structured fields from the published default version.

Version
Qwen3-Next-80B-A3B-Thinking
Released
Sep 10, 2025
Knowledge cutoff
Apr 1, 2025
Parameters
80B
Context window
131.1K
Max output
32.8K
Inputs
text
Outputs
text
Open weights
Yes
License
Apache 2.0

Capability profile

This profile uses the latest version under this unique model that has a calculated LLMBoard score. Arena and price are excluded.

Qwen3-Next-80B-A3B-Thinking category scores

Benchmark results

Published benchmark records for the scored version Qwen3-Next-80B-A3B-Thinking.

23 rows
Columns

Show columns

CFEval2071.0220.0%CAug 7, 2026
LiveBench 202411250.831484.6%CAug 7, 2026
BFCL-v30.751977.8%CAug 7, 2026
Multi-IF0.852079.0%CAug 7, 2026
OJBench0.37925.0%CAug 7, 2026
WritingBench0.891542.9%CAug 7, 2026
Arena-Hard v20.6101640.0%CAug 7, 2026
PolyMATH0.6102359.1%CAug 7, 2026
TAU-bench Retail0.7112558.3%CAug 7, 2026
Tau2 Airline0.6122350.0%CAug 7, 2026
MMLU-ProX0.8133261.3%CAug 7, 2026
Include0.8143156.7%CAug 7, 2026
TAU-bench Airline0.5152336.4%CAug 7, 2026
HMMT250.7162537.5%CAug 7, 2026
SuperGPQA0.6163454.5%CAug 7, 2026
MMLU-Redux0.9184863.8%CAug 7, 2026
Tau2 Retail0.7222616.0%CAug 7, 2026
IFEval0.9236565.6%CAug 7, 2026
MMLU-Pro0.83212975.8%CAug 7, 2026
Tau2 Telecom0.432358.8%CAug 7, 2026
LiveCodeBench v60.7335338.5%CAug 7, 2026
AIME 20250.95311454.0%CAug 7, 2026
GPQA0.89823358.2%CAug 7, 2026

Arena results

Preference and agent-evaluation signals from published Arena datasets.

58 rows
Columns

Show columns

textgerman1401359.2285N/AAug 6, 2026
textmath1401395.0826N/AAug 6, 2026
textindustry medicine and healthcare1421390.8720N/AAug 6, 2026
textchinese1431417.4713N/AAug 6, 2026
textkorean1441306.6331N/AAug 6, 2026
text style controlgerman1441365.8285N/AAug 6, 2026
textenglish1451392.86,615N/AAug 6, 2026
textjapanese1451282.1180N/AAug 6, 2026
textindustry mathematical1471388.0666N/AAug 6, 2026
textpolish1471350.8712N/AAug 6, 2026
text style controljapanese1471296.6180N/AAug 6, 2026
text style controlpolish1491364.6712N/AAug 6, 2026
text style controlchinese1501412.6713N/AAug 6, 2026
text style controlkorean1511316.8331N/AAug 6, 2026
text style controlmath1511389.2826N/AAug 6, 2026
texthard prompts english1521385.93,486N/AAug 6, 2026
textexpert1531366.7616N/AAug 6, 2026
textindustry business and management and financial operations1531361.82,495N/AAug 6, 2026
textindustry software and it services1531393.14,825N/AAug 6, 2026
textcoding1541391.42,672N/AAug 6, 2026
textindustry life and physical and social science1551381.72,119N/AAug 6, 2026
text style controlindustry mathematical1571385.9666N/AAug 6, 2026
textfrench1581356.7196N/AAug 6, 2026
textoverall1591367.613,661N/AAug 6, 2026
texthard prompts1591370.06,674N/AAug 6, 2026
textspanish1591348.6398N/AAug 6, 2026
textexclude ties1611333.29,737N/AAug 6, 2026
textinstruction following1651342.83,507N/AAug 6, 2026
text style controlindustry software and it services1651409.44,825N/AAug 6, 2026
text style controlcoding1661420.72,672N/AAug 6, 2026
text style controlenglish1661386.56,615N/AAug 6, 2026
textindustry legal and government1671363.9866N/AAug 6, 2026
text style controlexpert1671382.7616N/AAug 6, 2026
text style controlfrench1671366.7196N/AAug 6, 2026
text style controlindustry medicine and healthcare1671393.7720N/AAug 6, 2026
textnon english1691339.17,046N/AAug 6, 2026
textlonger query1701351.42,817N/AAug 6, 2026
textrussian1711337.8747N/AAug 6, 2026
text style controloverall1711369.113,661N/AAug 6, 2026
text style controlinstruction following1711358.23,507N/AAug 6, 2026
text style controlspanish1711350.4398N/AAug 6, 2026
text style controlexclude ties1721337.09,737N/AAug 6, 2026
text style controlhard prompts1731383.96,674N/AAug 6, 2026
text style controlindustry business and management and financial operations1731367.32,495N/AAug 6, 2026
text style controlnon english1751348.37,046N/AAug 6, 2026
text style controlrussian1751352.5747N/AAug 6, 2026
textindustry writing and literature and language1761327.33,064N/AAug 6, 2026
text style controlhard prompts english1761395.83,486N/AAug 6, 2026
textindustry entertainment and sports and media1771312.02,441N/AAug 6, 2026
text style controlindustry life and physical and social science1771379.72,119N/AAug 6, 2026
text style controllonger query1771368.42,817N/AAug 6, 2026
textmulti turn1781345.62,295N/AAug 6, 2026
text style controlmulti turn1821349.92,295N/AAug 6, 2026
text style controlindustry writing and literature and language1841338.13,064N/AAug 6, 2026
textcreative writing1851310.81,782N/AAug 6, 2026
text style controlcreative writing1871322.71,782N/AAug 6, 2026
text style controlindustry entertainment and sports and media1891315.12,441N/AAug 6, 2026
text style controlindustry legal and government1921361.2866N/AAug 6, 2026

Pricing

Official vendor API PAYG pricing is summarized first. The table then lists individual provider offerings without treating their minimum as the official price.

Official API
$0.144 input, $1.43 output per 1M
Official provider
Alibaba (China)
Lowest third-party
From $0.149 input, $1.2 output per 1M via Cortecs
Tracked offerings
4
3 rows
Columns

Show columns

Alibaba (China)qwen3-next-80b-a3b-thinkingglobal$0.144$1.43131.1KAug 7, 2026
Cortecsqwen3-next-80b-a3b-thinkingglobal$0.149$1.2128KAug 7, 2026
LLM Gatewayqwen3-next-80b-a3b-thinkingglobal$0.15$1.2131.1KAug 7, 2026

Official prices use only the vendor's configured official Provider and positive standard USD PAYG rates. Third-party offers remain explicitly labeled.

Qwen3 Next 80B A3B Thinking versions

All published versions linked to this unique model. The score columns identify the version used by the current overall ranking.

1 rows
Columns

Show columns

Qwen3-Next-80B-A3B-ThinkingSep 10, 202554.180B131.1K32.8KYesApache 2.0

What is Qwen3 Next 80B A3B Thinking?

A concise description based on the published model registry.

Qwen3-Next-80B-A3B-Thinking is the thinking variant of the Qwen3-Next series, featuring the same groundbreaking architecture as the instruct model. Leveraging GSPO, it addresses stability and efficiency challenges of hybrid attention + high-sparsity MoE in RL training. It uses Hybrid Attention combining Gated DeltaNet and Gated Attention for efficient ultra-long context modeling, High-Sparsity MoE with 512 experts (10 activated + 1 shared), and Multi-Token Prediction. With 80B total parameters and only 3B activated, it demonstrates outstanding performance on complex reasoning tasks — outperforming Qwen3-30B-A3B-Thinking-2507, Qwen3-32B-Thinking, and even the proprietary Gemini-2.5-Flash-Thinking across multiple benchmarks. Architecture: 48 layers, 15T training tokens, hybrid layout of 12*(3*(Gated DeltaNet->MoE)->(Gated Attention->MoE)). Supports only thinking mode with automatic <think> tag inclusion, may generate longer thinking content.

Use the benchmark, Arena and pricing sections above as separate evidence. A missing field means the current data snapshot does not support that claim.

Data snapshot: 2026-08-07. Editorial model content is not available in the backend.

Qwen3 Next 80B A3B Thinking vs nearby models

Open a comparison with the three ranked models immediately above and below this model.

Qwen3 Next 80B A3B Thinkingvso4 miniQwen3 Next 80B A3B ThinkingvsMiniMax M2.1Qwen3 Next 80B A3B ThinkingvsLlama 3.3 70BQwen3 Next 80B A3B ThinkingvsQwen3.5 35B A3BQwen3 Next 80B A3B ThinkingvsQwen2.5 72BQwen3 Next 80B A3B ThinkingvsGemini 2.0 Flash

Models similar to Qwen3 Next 80B A3B Thinking

Recommendations prioritize the same model type and family, then the closest published LLMBoard score.

#71-0.9
AC

Qwen3.5 35B A3B

Alibaba Cloud / Qwen Team

53.3 LLMBoard

DetailsCompare
#72-1.1
AC

Qwen2.5 72B

Alibaba Cloud / Qwen Team

53.1 LLMBoard

DetailsCompare
#78-2.4
AC

Qwen3 VL 32B Thinking

Alibaba Cloud / Qwen Team

51.7 LLMBoard

DetailsCompare
#79-2.6
AC

Qwen2.5 VL 32B

Alibaba Cloud / Qwen Team

51.5 LLMBoard

DetailsCompare
#59+2.8
AC

Qwen2.5 32B

Alibaba Cloud / Qwen Team

56.9 LLMBoard

DetailsCompare
#58+2.9
AC

Qwen3.5 27B

Alibaba Cloud / Qwen Team

57.1 LLMBoard

DetailsCompare

FAQ

Common questions about Qwen3 Next 80B A3B Thinking.

When was Qwen3 Next 80B A3B Thinking released?

Qwen3 Next 80B A3B Thinking's default version was released on Sep 10, 2025.

How much does Qwen3 Next 80B A3B Thinking cost?

Qwen3 Next 80B A3B Thinking's official API price is $0.144 per million input tokens and $1.43 per million output tokens via Alibaba (China). The lowest tracked third-party offer starts at $0.149 input and $1.2 output via Cortecs.

Who created Qwen3 Next 80B A3B Thinking?

Qwen3 Next 80B A3B Thinking is published under Alibaba Cloud / Qwen Team in the model registry.

What is the context window for Qwen3 Next 80B A3B Thinking?

The default version has a 131.1K token context window.

Is Qwen3 Next 80B A3B Thinking open weight?

Yes. The default version is marked as open weight under Apache 2.0.

How many API providers offer Qwen3 Next 80B A3B Thinking?

4 published provider offerings are linked to the default version.

What models should I compare Qwen3 Next 80B A3B Thinking with?

Nearby ranked alternatives include o4 mini, MiniMax M2.1, Llama 3.3 70B.

Rankings

OverallCodingText ArenaPricing

Modalities

Image GenerationVideo GenerationSpeech-to-TextEmbeddings

Benchmarks

All BenchmarksReasoningMathCoding

Vendors

All VendorsOpenAIAnthropicGoogle
llmboard.aiCopyright 2026 llmboard.ai