llmboard.ai
Benchmarks
CompareRankings
llmboard.ai
Benchmarks
CompareRankings
HomeModelsLongCat Flash Thinking

Meituan model product

LongCat Flash Thinking

LongCat-Flash-Thinking-2601 is an upgraded version of LongCat-Flash-Thinking with 560B total parameters (MoE, ~27B activated). It achieves open-source SOTA performance on core evaluation benchmarks including Agentic Search, Agentic Tool Use, and Tool-Integrated Reasoning (TIR). Features Heavy Thinking mode that contributes +4-6 points on demanding agentic reasoning benchmarks. Mid-training with structured agentic trajectories improves pass@k by up to +12 points, and context management yields +17.5 improvement.

Updated Aug 10, 2026. Default version: LongCat-Flash-Thinking-2601

Compare
LLMBoard score61.2LongCat-Flash-Thinking-2601
Coverage100%6 benchmark families
Context window128KTokens
Official input priceN/AOfficial price unavailable

On this page

  • Specification
  • Capability
  • Benchmarks
  • Arena
  • Pricing
  • Versions
  • About
  • Compare
  • Similar models
  • FAQ

Model specification

Structured fields from the published default version.

Version
LongCat-Flash-Thinking-2601
Released
Jan 14, 2026
Knowledge cutoff
Unknown
Parameters
560B
Context window
128K
Max output
128K
Inputs
text
Outputs
text
Open weights
No
License
MIT

Capability profile

This profile uses the latest version under this unique model that has a calculated LLMBoard score. Arena and price are excluded.

LongCat-Flash-Thinking-2601 category scores

Benchmark results

Published benchmark records for the scored version LongCat-Flash-Thinking-2601.

11 rows
Columns

Show columns

Tau2 Airline0.8123100.0%CAug 7, 2026
Tau2 Telecom1.023597.1%CAug 7, 2026
BrowseComp-zh0.741375.0%CAug 7, 2026
Tau2 Retail0.942688.0%CAug 7, 2026
LiveCodeBench0.867393.1%CAug 7, 2026
AIME 20251.0911492.9%CAug 7, 2026
IMO-AnswerBench0.818195.6%CAug 7, 2026
BrowseComp0.6385835.1%CAug 7, 2026
Humanity's Last Exam0.3489248.4%CAug 7, 2026
SWE-Bench Verified0.75810444.7%CAug 7, 2026
GPQA0.88523363.8%CAug 7, 2026

Arena results

Preference and agent-evaluation signals from published Arena datasets.

No published Arena match

The default version has no published Arena rows, or its source alias has not been resolved.

Pricing

Official vendor API PAYG pricing is summarized first. The table then lists individual provider offerings without treating their minimum as the official price.

Official API
N/A
Official provider
N/A
Lowest third-party
N/A
Tracked offerings
0
No published price snapshot

The default version has no provider offering with current input or output token prices.

Official prices use only the vendor's configured official Provider and positive standard USD PAYG rates. Third-party offers remain explicitly labeled.

LongCat Flash Thinking versions

All published versions linked to this unique model. The score columns identify the version used by the current overall ranking.

2 rows
Columns

Show columns

LongCat-Flash-Thinking-2601Jan 14, 202661.2560B128K128KNoMIT
LongCat-Flash-ThinkingSep 22, 2025N/A560B128K128KNoMIT

What is LongCat Flash Thinking?

A concise description based on the published model registry.

LongCat-Flash-Thinking-2601 is an upgraded version of LongCat-Flash-Thinking with 560B total parameters (MoE, ~27B activated). It achieves open-source SOTA performance on core evaluation benchmarks including Agentic Search, Agentic Tool Use, and Tool-Integrated Reasoning (TIR). Features Heavy Thinking mode that contributes +4-6 points on demanding agentic reasoning benchmarks. Mid-training with structured agentic trajectories improves pass@k by up to +12 points, and context management yields +17.5 improvement.

Use the benchmark, Arena and pricing sections above as separate evidence. A missing field means the current data snapshot does not support that claim.

Data snapshot: 2026-08-07. Editorial model content is not available in the backend.

LongCat Flash Thinking vs nearby models

Open a comparison with the three ranked models immediately above and below this model.

LongCat Flash ThinkingvsGLM 4.5LongCat Flash Thinkingvso3 miniLongCat Flash ThinkingvsQwen3.5 122B A10BLongCat Flash ThinkingvsNemotron Nano 9BLongCat Flash ThinkingvsKimi K2 ThinkingLongCat Flash ThinkingvsNemotron 3 Super

Models similar to LongCat Flash Thinking

Recommendations prioritize the same model type and family, then the closest published LLMBoard score.

#75-8.5
ME

LongCat Flash Chat

Meituan

52.8 LLMBoard

DetailsCompare
#52-0.0
NV

Nemotron Nano 9B

NVIDIA

61.2 LLMBoard

DetailsCompare
#50+0.1
AC

Qwen3.5 122B A10B

Alibaba Cloud / Qwen Team

61.4 LLMBoard

DetailsCompare
#49+0.3
OP

o3 mini

OpenAI

61.6 LLMBoard

DetailsCompare
#48+0.8
ZA

GLM 4.5

Zhipu AI

62.0 LLMBoard

DetailsCompare
#47+1.3
NV

Llama 3.1 Nemotron Ultra 253B

NVIDIA

62.6 LLMBoard

DetailsCompare

FAQ

Common questions about LongCat Flash Thinking.

When was LongCat Flash Thinking released?

LongCat Flash Thinking's default version was released on Jan 14, 2026.

How much does LongCat Flash Thinking cost?

No official standard PAYG price is currently available for LongCat Flash Thinking.

Who created LongCat Flash Thinking?

LongCat Flash Thinking is published under Meituan in the model registry.

What is the context window for LongCat Flash Thinking?

The default version has a 128K token context window.

Is LongCat Flash Thinking open weight?

No. The default version is not marked as having publicly available weights.

How many API providers offer LongCat Flash Thinking?

No published provider offering is currently linked to the default version.

What models should I compare LongCat Flash Thinking with?

Nearby ranked alternatives include GLM 4.5, o3 mini, Qwen3.5 122B A10B.

Rankings

OverallCodingText ArenaPricing

Modalities

Image GenerationVideo GenerationSpeech-to-TextEmbeddings

Benchmarks

All BenchmarksReasoningMathCoding

Vendors

All VendorsOpenAIAnthropicGoogle
llmboard.aiCopyright 2026 llmboard.ai