llmboard.aiLeaderboard Center
Overall
Overall RankingOpen Models
Tools
Model DirectoryCompare Models
Capabilities
CodingReasoningMathKnowledgeInstruction Following
Price & Efficiency
Price & ValueCapability vs. PriceRuntime Performance
Modalities
Image GenerationVideo GenerationSpeech ModelsEmbeddings
Core Benchmarks
GPQAMMLU-ProAIME 2025SWE-Bench VerifiedMMLUHumanity's Last ExamLiveCodeBenchMATHHumanEvalMMMU-ProView all benchmarks
Methods
Scoring & Data
393 models668 benchmarks

Leaderboard Center

Overall RankingCodingCore BenchmarksPrice & ValueRuntime Performance

Modalities

All ModelsImage GenerationImage EditingVideo GenerationImage-to-VideoVideo EditingText-to-SpeechSpeech-to-TextEmbeddings

Data & Methods

Scoring MethodAll BenchmarksReasoningMath

Vendors

All VendorsOpenAIAnthropicGoogle
llmboard.aiCopyright 2026 llmboard.ai

Meituan model product

LongCat Flash Chat

3B parameters (~27B average) based on contextual demands.

Updated Aug 12, 2026. Default version: LongCat-Flash-Chat

Compare
LLMBoard score36.9LongCat-Flash-Chat
Coverage100%16 benchmark families
Context window128KTokens
Official input priceN/AOfficial price unavailable

On this page

  • Specification
  • Capability
  • Benchmarks
  • Arena
  • Runtime
  • Pricing
  • Versions
  • About
  • Compare
  • Similar models
  • FAQ

LongCat Flash Chat Specifications

Technical details for the model's default version.

Version
LongCat-Flash-Chat
Released
Aug 29, 2025
Knowledge cutoff
Unknown
Parameters
560B
Context window
128K
Max output
128K
Inputs
text
Outputs
text
Open weights
No
License
MIT

LongCat Flash Chat Capability Profile

This profile uses the model's current scored version. Arena ratings and prices are shown separately.

LongCat-Flash-Chat category scores

LongCat Flash Chat Benchmark Results

Benchmark scores for LongCat-Flash-Chat.

16 rows
Columns

Show columns

CMMLU84.3%03660.0%CAug 11, 2026
ZebraLogic89.3%04857.1%CAug 11, 2026
Terminal-Bench39.5%092566.7%CAug 11, 2026
MMLU89.7%1210088.9%CAug 11, 2026
MATH-50096.4%133261.3%CAug 11, 2026
Tau2 Airline58.0%132345.5%CAug 11, 2026
DROP79.1%163048.3%CAug 11, 2026
HumanEval88.4%196672.3%CAug 11, 2026
Tau2 Retail71.3%192628.0%CAug 11, 2026
IFEval89.6%206570.3%CAug 11, 2026
Tau2 Telecom73.7%243532.4%CAug 11, 2026
MMLU-Pro82.7%3312975.0%CAug 11, 2026
LiveCodeBench48.0%487334.7%CAug 11, 2026
SWE-Bench Verified60.4%7910525.0%CAug 11, 2026
AIME 202561.3%9711415.0%CAug 11, 2026
GPQA73.2%11723450.2%CAug 11, 2026

LongCat Flash Chat Arena Results

Preference and agent-evaluation results for the default version.

5 rows
Columns

Show columns

text style controloverall811435.728,034N/AAug 10, 2026
textoverall821425.928,034N/AAug 10, 2026
textoverall891422.311,433N/AAug 10, 2026
text factualityoverall1001428.027,955N/AAug 10, 2026
text style controloverall1361401.011,433N/AAug 10, 2026

LongCat Flash Chat Runtime Performance

Provider-specific output speed and catalog latency for LongCat-Flash-Chat. Runtime does not affect the capability score.

1 rows
Columns

Show columns

Meituan100 tok/s3 s128K128KAug 11, 2026

Output Speed is generated output tokens received per second. Catalog latency is reported separately from observed provider TTFT.

LongCat Flash Chat Pricing

Official vendor API pricing appears first, followed by individual provider offers.

Official API
N/A
Official provider
N/A
Lowest third-party
N/A
Tracked offerings
0
No provider prices

The default version has no current input or output token prices.

Official prices use only the vendor's configured official Provider and positive standard USD PAYG rates. Third-party offers remain explicitly labeled.

LongCat Flash Chat Versions

Available versions of this model. The score column identifies the version used in the overall ranking.

1 rows
Columns

Show columns

LongCat-Flash-ChatAug 29, 202536.9560B128K128KNoMIT

What is LongCat Flash Chat?

Key information about LongCat Flash Chat and its available data.

3B parameters (~27B average) based on contextual demands. It features Zero-Computation Experts for efficient routing and supports 128K context.

Optimized for conversational and agentic tasks, it shows competitive performance across reasoning, coding, instruction following, and domain benchmarks with particular strengths in tool use and complex multi-step interactions. Achieves over 100 tokens per second on H800 GPUs.

Data as of 2026-08-11.

LongCat Flash Chat vs nearby models

Open a comparison with the three ranked models immediately above and below this model.

LongCat Flash ChatvsNorth Mini Code 1.0LongCat Flash ChatvsKimi K2LongCat Flash ChatvsMiniMax M1 80KLongCat Flash ChatvsQwen3 Next 80B A3BLongCat Flash Chatvso1LongCat Flash Chatvso3 mini

Models similar to LongCat Flash Chat

Recommendations prioritize the same model type and family, then the closest LLMBoard score.

#152-3.8
ME

LongCat Flash Lite

Meituan

33.0 LLMBoard

DetailsCompare
#55+24.0
ME

LongCat Flash Thinking

Meituan

60.8 LLMBoard

DetailsCompare
#137+0.1
MI

MiniMax M1 80K

MiniMax

37.0 LLMBoard

DetailsCompare
#139-0.3
AC

Qwen3 Next 80B A3B

Alibaba Cloud / Qwen Team

36.6 LLMBoard

DetailsCompare
#136+0.3
MA

Kimi K2

Moonshot AI

37.2 LLMBoard

DetailsCompare
#135+0.5
CO

North Mini Code 1.0

Cohere

37.4 LLMBoard

DetailsCompare

FAQ

Common questions about LongCat Flash Chat.

When was LongCat Flash Chat released?

LongCat Flash Chat's default version was released on Aug 29, 2025.

How much does LongCat Flash Chat cost?

No official standard PAYG price is currently available for LongCat Flash Chat.

Who created LongCat Flash Chat?

LongCat Flash Chat was created by Meituan.

What is the context window for LongCat Flash Chat?

The default version has a 128K token context window.

Is LongCat Flash Chat open weight?

No. The default version is not marked as having publicly available weights.

How many API providers offer LongCat Flash Chat?

No provider offering is currently linked to the default version.

What models should I compare LongCat Flash Chat with?

Nearby ranked alternatives include North Mini Code 1.0, Kimi K2, MiniMax M1 80K.