llmboard.ai
Benchmarks
CompareRankings
llmboard.ai
Benchmarks
CompareRankings
HomeModelsGLM 5.1

Zhipu AI model product

GLM 5.1

GLM-5.1 is Z.AI's next-generation flagship foundation model designed for long-horizon agentic engineering tasks. Built on a 754B MoE architecture (40B active parameters), it can work continuously and autonomously on a single task for up to 8 hours, completing the full loop from planning and execution to iterative optimization and delivery. GLM-5.1 achieves state-of-the-art on SWE-Bench Pro (58.4) and demonstrates strong performance across coding, reasoning, and agentic benchmarks. It supports 200K context length, 128K max output tokens, thinking mode, function calling, structured output, context caching, and MCP integration. Overall performance is aligned with Claude Opus 4.6 with particular strengths in sustained execution and complex engineering optimization.

Updated Aug 10, 2026. Default version: GLM-5.1

Compare
LLMBoard score64.7GLM-5.1
Coverage100%7 benchmark families
Context window200KTokens
Official input priceN/AOfficial price unavailable

On this page

  • Specification
  • Capability
  • Benchmarks
  • Arena
  • Pricing
  • Versions
  • About
  • Compare
  • Similar models
  • FAQ

Model specification

Structured fields from the published default version.

Version
GLM-5.1
Released
Apr 7, 2026
Knowledge cutoff
Unknown
Parameters
754B
Context window
200K
Max output
131.1K
Inputs
text
Outputs
text
Open weights
Yes
License
MIT

Capability profile

This profile uses the latest version under this unique model that has a calculated LLMBoard score. Arena and price are excluded.

GLM-5.1 category scores

Benchmark results

Published benchmark records for the scored version GLM-5.1.

18 rows
Columns

Show columns

Vending-Bench 25634.42466.7%CAug 7, 2026
AIME 20261.031787.5%CAug 7, 2026
TAU3-Bench0.73550.0%CAug 7, 2026
NL2Repo0.481446.1%CAug 7, 2026
CyberGym0.791120.0%CAug 7, 2026
HMMT 20250.9103371.9%CAug 7, 2026
IMO-AnswerBench0.8101950.0%CAug 7, 2026
FrontierSWE0.3111528.6%BAug 7, 2026
HMMT Feb 260.811110.0%CAug 7, 2026
Terminal-Bench 2.00.7114979.2%CAug 7, 2026
Finance Agent v20.4122656.0%BAug 7, 2026
Humanity's Last Exam0.5159284.6%CAug 7, 2026
MCP Atlas0.7173044.8%CAug 7, 2026
SWE-Bench Pro0.6184460.5%CAug 7, 2026
BrowseComp0.8215864.9%CAug 7, 2026
Toolathlon0.4243123.3%CAug 7, 2026
LiveBench0.7293824.3%BAug 7, 2026
GPQA0.94623380.6%CAug 7, 2026

Arena results

Preference and agent-evaluation signals from published Arena datasets.

88 rows
Columns

Show columns

texthard prompts english141483.211,412N/AAug 6, 2026
textenglish171477.616,813N/AAug 6, 2026
text style controlenglish171483.616,813N/AAug 6, 2026
text factualitychinese181517.31,765N/AAug 6, 2026
textindustry medicine and healthcare191473.62,682N/AAug 6, 2026
textmath191479.31,875N/AAug 6, 2026
text factualitymath191469.81,438N/AAug 6, 2026
text style controlhard prompts english201500.711,412N/AAug 6, 2026
text style controlmath201482.71,875N/AAug 6, 2026
textcoding211490.210,330N/AAug 6, 2026
webdevwebdev-html211532.6888N/AAug 6, 2026
textchinese221520.72,206N/AAug 6, 2026
textcreative writing221453.56,516N/AAug 6, 2026
textindustry software and it services221486.414,744N/AAug 6, 2026
text factualitycreative writing221452.15,806N/AAug 6, 2026
text factualityfrench221486.0967N/AAug 6, 2026
textindustry entertainment and sports and media231443.28,071N/AAug 6, 2026
textkorean241426.6713N/AAug 6, 2026
textmulti turn241474.16,070N/AAug 6, 2026
text factualityjapanese241416.9255N/AAug 6, 2026
textgerman251462.6565N/AAug 6, 2026
textindustry writing and literature and language251455.29,234N/AAug 6, 2026
text factualityhard prompts english251499.810,904N/AAug 6, 2026
textoverall261464.336,716N/AAug 6, 2026
textexclude ties261467.427,654N/AAug 6, 2026
textindustry life and physical and social science261479.16,077N/AAug 6, 2026
textpolish261462.1726N/AAug 6, 2026
webdevoverall261515.17,853N/AAug 6, 2026
webdevwebdev261515.17,853N/AAug 6, 2026
webdevwebdev-react261508.86,224N/AAug 6, 2026
text factualityenglish271474.916,718N/AAug 6, 2026
text factualityindustry life and physical and social science271487.95,419N/AAug 6, 2026
text factualityindustry writing and literature and language271455.28,636N/AAug 6, 2026
text style controlindustry life and physical and social science271488.96,077N/AAug 6, 2026
text style controllonger query271483.216,213N/AAug 6, 2026
texthard prompts281474.424,242N/AAug 6, 2026
textnon english281449.819,902N/AAug 6, 2026
text factualityindustry medicine and healthcare281489.02,180N/AAug 6, 2026
text factualitymulti turn281476.25,324N/AAug 6, 2026
text style controlchinese281520.82,206N/AAug 6, 2026
textindustry legal and government291470.42,826N/AAug 6, 2026
textindustry mathematical291477.51,968N/AAug 6, 2026
textrussian291460.93,876N/AAug 6, 2026
textspanish291459.21,235N/AAug 6, 2026
text factualityindustry entertainment and sports and media291439.77,547N/AAug 6, 2026
text factualityindustry legal and government291479.12,353N/AAug 6, 2026
text factualityindustry mathematical291459.91,508N/AAug 6, 2026
text factualityrussian291466.23,274N/AAug 6, 2026
text style controlcreative writing291448.76,516N/AAug 6, 2026
text style controlexpert291498.83,664N/AAug 6, 2026
textjapanese301426.9365N/AAug 6, 2026
textlonger query301469.916,213N/AAug 6, 2026
text style controlcoding301516.710,330N/AAug 6, 2026
text style controlmulti turn301480.36,070N/AAug 6, 2026
textexpert311481.43,664N/AAug 6, 2026
textindustry business and management and financial operations311451.17,452N/AAug 6, 2026
text factualityspanish311445.4839N/AAug 6, 2026
text style controlindustry mathematical311484.01,968N/AAug 6, 2026
text style controlindustry software and it services311505.014,744N/AAug 6, 2026
textfrench321474.41,311N/AAug 6, 2026
text factualitycoding331515.39,790N/AAug 6, 2026
text factualityexclude ties331465.027,539N/AAug 6, 2026
text factualityexpert331489.83,143N/AAug 6, 2026
text factualityhard prompts331482.924,161N/AAug 6, 2026
text factualitylonger query331478.215,725N/AAug 6, 2026
text style controlfrench331486.31,311N/AAug 6, 2026
text style controlindustry writing and literature and language331455.29,234N/AAug 6, 2026
text factualityoverall341458.536,569N/AAug 6, 2026
text style controlindustry entertainment and sports and media341442.88,071N/AAug 6, 2026
textinstruction following351452.712,475N/AAug 6, 2026
text factualityindustry business and management and financial operations351463.96,735N/AAug 6, 2026
text factualityindustry software and it services351497.314,363N/AAug 6, 2026
text factualityinstruction following351458.511,957N/AAug 6, 2026
text style controlhard prompts351490.824,242N/AAug 6, 2026
text style controlindustry legal and government351475.72,826N/AAug 6, 2026
text style controlindustry medicine and healthcare351485.22,682N/AAug 6, 2026
text style controlinstruction following351462.812,475N/AAug 6, 2026
text factualitynon english361443.219,824N/AAug 6, 2026
text style controloverall361467.836,716N/AAug 6, 2026
text style controlexclude ties361475.027,654N/AAug 6, 2026
text style controlindustry business and management and financial operations381463.97,452N/AAug 6, 2026
text style controlgerman391463.6565N/AAug 6, 2026
text style controljapanese411429.7365N/AAug 6, 2026
text style controlspanish411456.51,235N/AAug 6, 2026
text style controlnon english421451.119,902N/AAug 6, 2026
text style controlkorean431419.7713N/AAug 6, 2026
text style controlrussian431464.83,876N/AAug 6, 2026
text style controlpolish471458.0726N/AAug 6, 2026

Pricing

Official vendor API PAYG pricing is summarized first. The table then lists individual provider offerings without treating their minimum as the official price.

Official API
N/A
Official provider
N/A
Lowest third-party
From $0.45 input, $2.15 output per 1M via CrofAI
Tracked offerings
19
18 rows
Columns

Show columns

Zhipu AI Coding Planglm-5.1globalN/AN/A200KAug 7, 2026
Alibaba Token Planglm-5.1globalN/AN/A202.8KAug 7, 2026
Alibaba Token Plan (China)glm-5.1globalN/AN/A202.8KAug 7, 2026
CrofAIglm-5.1global$0.45$2.15202.8KAug 7, 2026
EBCloudGLM-5.1global$0.8571$3.43200KAug 7, 2026
302.AIglm-5.1global$0.86$3.5200KAug 7, 2026
Alibaba (China)glm-5.1global$0.87$3.48202.8KAug 7, 2026
LLM Gatewayglm-5.1global$0.931$2.93204.8KAug 7, 2026
DigitalOceanglm-5.1global$0.975$4.3163.8KAug 7, 2026
WaferGLM-5.1global$1$3.2202.8KAug 7, 2026
DInferenceglm-5.1global$1.25$3.89200KAug 7, 2026
Cortecsglm-5.1global$1.38$4.35202.8KAug 7, 2026
OpenCode Goglm-5.1global$1.4$4.4202.8KAug 7, 2026
OpenCode Zenglm-5.1global$1.4$4.4204.8KAug 7, 2026
Aurikoglm-5.1global$1.4$4.4200KAug 7, 2026
Z.AIglm-5.1global$1.4$4.4200KAug 7, 2026
Charm Hyperglm-5.1global$1.52$4.79202.8KAug 7, 2026
GreenPTglm-5.1global$1.76$5.52200KAug 7, 2026

Official prices use only the vendor's configured official Provider and positive standard USD PAYG rates. Third-party offers remain explicitly labeled.

GLM 5.1 versions

All published versions linked to this unique model. The score columns identify the version used by the current overall ranking.

1 rows
Columns

Show columns

GLM-5.1Apr 7, 202664.7754B200K131.1KYesMIT

What is GLM 5.1?

A concise description based on the published model registry.

GLM-5.1 is Z.AI's next-generation flagship foundation model designed for long-horizon agentic engineering tasks. Built on a 754B MoE architecture (40B active parameters), it can work continuously and autonomously on a single task for up to 8 hours, completing the full loop from planning and execution to iterative optimization and delivery. GLM-5.1 achieves state-of-the-art on SWE-Bench Pro (58.4) and demonstrates strong performance across coding, reasoning, and agentic benchmarks. It supports 200K context length, 128K max output tokens, thinking mode, function calling, structured output, context caching, and MCP integration. Overall performance is aligned with Claude Opus 4.6 with particular strengths in sustained execution and complex engineering optimization.

Use the benchmark, Arena and pricing sections above as separate evidence. A missing field means the current data snapshot does not support that claim.

Data snapshot: 2026-08-07. Editorial model content is not available in the backend.

GLM 5.1 vs nearby models

Open a comparison with the three ranked models immediately above and below this model.

GLM 5.1vsGrok 4 FastGLM 5.1vsGLM 4.7GLM 5.1vsQwen3 235B A22B ThinkingGLM 5.1vsClaude Sonnet 3.7GLM 5.1vsGemini 2.5 FlashGLM 5.1vsNova Pro

Models similar to GLM 5.1

Recommendations prioritize the same model type and family, then the closest published LLMBoard score.

#38+1.5
ZA

GLM 4.7

Zhipu AI

66.2 LLMBoard

DetailsCompare
#48-2.7
ZA

GLM 4.5

Zhipu AI

62.0 LLMBoard

DetailsCompare
#65-9.6
ZA

GLM 4.5 Air

Zhipu AI

55.1 LLMBoard

DetailsCompare
#10+15.6
ZA

GLM 5.2

Zhipu AI

80.3 LLMBoard

DetailsCompare
#39+1.0
AC

Qwen3 235B A22B Thinking

Alibaba Cloud / Qwen Team

65.7 LLMBoard

DetailsCompare
#41-1.1
AN

Claude Sonnet 3.7

Anthropic

63.6 LLMBoard

DetailsCompare

FAQ

Common questions about GLM 5.1.

When was GLM 5.1 released?

GLM 5.1's default version was released on Apr 7, 2026.

How much does GLM 5.1 cost?

No official standard PAYG price is currently available for GLM 5.1. The lowest tracked third-party offer starts at $0.45 input and $2.15 output via CrofAI.

Who created GLM 5.1?

GLM 5.1 is published under Zhipu AI in the model registry.

What is the context window for GLM 5.1?

The default version has a 200K token context window.

Is GLM 5.1 open weight?

Yes. The default version is marked as open weight under MIT.

How many API providers offer GLM 5.1?

19 published provider offerings are linked to the default version.

What models should I compare GLM 5.1 with?

Nearby ranked alternatives include Grok 4 Fast, GLM 4.7, Qwen3 235B A22B Thinking.

Rankings

OverallCodingText ArenaPricing

Modalities

Image GenerationVideo GenerationSpeech-to-TextEmbeddings

Benchmarks

All BenchmarksReasoningMathCoding

Vendors

All VendorsOpenAIAnthropicGoogle
llmboard.aiCopyright 2026 llmboard.ai