llmboard.ai
Benchmarks
CompareRankings
llmboard.ai
Benchmarks
CompareRankings
HomeModelsLongCat Flash Lite

Meituan model product

LongCat Flash Lite

LongCat-Flash-Lite is a lightweight MoE model from Meituan with 68.5B total parameters and only 2.9B-4.5B activated per token. It explores N-gram embedding expansion as a new scaling direction, supporting 256K context length via YaRN. Optimized for agent tooling and programming tasks, achieving 500-700 tokens per second inference speed while maintaining strong performance on coding, math, and agentic benchmarks.

Updated Aug 10, 2026. Default version: LongCat-Flash-Lite

Compare
LLMBoard scoreN/ANot scored
CoverageN/ANo published score
Context window256KTokens
Official input priceN/AOfficial price unavailable

On this page

  • Specification
  • Capability
  • Benchmarks
  • Arena
  • Pricing
  • Versions
  • About
  • Similar models
  • FAQ

Model specification

Structured fields from the published default version.

Version
LongCat-Flash-Lite
Released
Feb 5, 2026
Knowledge cutoff
Unknown
Parameters
68.5B
Context window
256K
Max output
128K
Inputs
text
Outputs
text
Open weights
No
License
MIT

Capability profile

This profile uses the latest version under this unique model that has a calculated LLMBoard score. Arena and price are excluded.

No capability profile

No published version under this unique model currently has enough benchmark coverage to calculate a score.

Benchmark results

Published benchmark records for the scored version currently unavailable.

No benchmark rows

No published benchmark result is linked to the scored version.

Arena results

Preference and agent-evaluation signals from published Arena datasets.

No published Arena match

The default version has no published Arena rows, or its source alias has not been resolved.

Pricing

Official vendor API PAYG pricing is summarized first. The table then lists individual provider offerings without treating their minimum as the official price.

Official API
N/A
Official provider
N/A
Lowest third-party
N/A
Tracked offerings
0
No published price snapshot

The default version has no provider offering with current input or output token prices.

Official prices use only the vendor's configured official Provider and positive standard USD PAYG rates. Third-party offers remain explicitly labeled.

LongCat Flash Lite versions

All published versions linked to this unique model. The score columns identify the version used by the current overall ranking.

1 rows
Columns

Show columns

LongCat-Flash-LiteFeb 5, 2026N/A68.5B256K128KNoMIT

What is LongCat Flash Lite?

A concise description based on the published model registry.

LongCat-Flash-Lite is a lightweight MoE model from Meituan with 68.5B total parameters and only 2.9B-4.5B activated per token. It explores N-gram embedding expansion as a new scaling direction, supporting 256K context length via YaRN. Optimized for agent tooling and programming tasks, achieving 500-700 tokens per second inference speed while maintaining strong performance on coding, math, and agentic benchmarks.

Use the benchmark, Arena and pricing sections above as separate evidence. A missing field means the current data snapshot does not support that claim.

Data snapshot: 2026-08-07. Editorial model content is not available in the backend.

Models similar to LongCat Flash Lite

Recommendations prioritize the same model type and family, then the closest published LLMBoard score.

#75
ME

LongCat Flash Chat

Meituan

52.8 LLMBoard

DetailsCompare
#51
ME

LongCat Flash Thinking

Meituan

61.2 LLMBoard

DetailsCompare
#170
AC

Qwen3.5 0.8B

Alibaba Cloud / Qwen Team

0.3 LLMBoard

DetailsCompare
#169
GO

Gemma 3 1B

Google

2.7 LLMBoard

DetailsCompare
#168
GO

Gemma 3n E2B

Google

6.9 LLMBoard

DetailsCompare
#167
AC

Qwen3 VL 4B

Alibaba Cloud / Qwen Team

7.7 LLMBoard

DetailsCompare

FAQ

Common questions about LongCat Flash Lite.

When was LongCat Flash Lite released?

LongCat Flash Lite's default version was released on Feb 5, 2026.

How much does LongCat Flash Lite cost?

No official standard PAYG price is currently available for LongCat Flash Lite.

Who created LongCat Flash Lite?

LongCat Flash Lite is published under Meituan in the model registry.

What is the context window for LongCat Flash Lite?

The default version has a 256K token context window.

Is LongCat Flash Lite open weight?

No. The default version is not marked as having publicly available weights.

How many API providers offer LongCat Flash Lite?

No published provider offering is currently linked to the default version.

What models should I compare LongCat Flash Lite with?

No nearby ranked alternatives are currently available.

Rankings

OverallCodingText ArenaPricing

Modalities

Image GenerationVideo GenerationSpeech-to-TextEmbeddings

Benchmarks

All BenchmarksReasoningMathCoding

Vendors

All VendorsOpenAIAnthropicGoogle
llmboard.aiCopyright 2026 llmboard.ai