llmboard.aiLeaderboard Center
Overall
Overall RankingOpen Models
Tools
Model DirectoryCompare Models
Capabilities
CodingReasoningMathKnowledgeInstruction Following
Price & Efficiency
Price & ValueCapability vs. PriceRuntime Performance
Modalities
Image GenerationVideo GenerationSpeech ModelsEmbeddings
Core Benchmarks
GPQAMMLU-ProAIME 2025SWE-Bench VerifiedMMLUHumanity's Last ExamLiveCodeBenchMATHHumanEvalMMMU-ProView all benchmarks
Methods
Scoring & Data
393 models668 benchmarks

Leaderboard Center

Overall RankingCodingCore BenchmarksPrice & ValueRuntime Performance

Modalities

All ModelsImage GenerationImage EditingVideo GenerationImage-to-VideoVideo EditingText-to-SpeechSpeech-to-TextEmbeddings

Data & Methods

Scoring MethodAll BenchmarksReasoningMath

Vendors

All VendorsOpenAIAnthropicGoogle
llmboard.aiCopyright 2026 llmboard.ai

Meituan model product

LongCat Flash Thinking

LongCat-Flash-Thinking is Meituan's reasoning model built on the LongCat-Flash foundation with 560B total parameters (MoE, ~27B activated).

Updated Aug 12, 2026. Default version: LongCat-Flash-Thinking-2601

Compare
LLMBoard score60.8LongCat-Flash-Thinking-2601
Coverage80%11 benchmark families
Context window128KTokens
Official input priceN/AOfficial price unavailable

On this page

  • Specification
  • Capability
  • Benchmarks
  • Arena
  • Runtime
  • Pricing
  • Versions
  • About
  • Compare
  • Similar models
  • FAQ

LongCat Flash Thinking Specifications

Technical details for the model's default version.

Version
LongCat-Flash-Thinking-2601
Released
Jan 14, 2026
Knowledge cutoff
Unknown
Parameters
560B
Context window
128K
Max output
128K
Inputs
text
Outputs
text
Open weights
No
License
MIT

LongCat Flash Thinking Capability Profile

This profile uses the model's current scored version. Arena ratings and prices are shown separately.

LongCat-Flash-Thinking-2601 category scores

LongCat Flash Thinking Benchmark Results

Benchmark scores for LongCat-Flash-Thinking-2601.

11 rows
Columns

Show columns

Tau2 Airline76.5%0123100.0%CAug 11, 2026
Tau2 Telecom99.3%023597.1%CAug 11, 2026
BrowseComp-zh69.0%041375.0%CAug 11, 2026
Tau2 Retail88.6%042688.0%CAug 11, 2026
LiveCodeBench82.8%067393.1%CAug 11, 2026
AIME 202599.6%0911492.9%CAug 11, 2026
IMO-AnswerBench78.6%18195.6%CAug 11, 2026
BrowseComp56.6%385835.1%CAug 11, 2026
Humanity's Last Exam25.2%489348.9%CAug 11, 2026
SWE-Bench Verified70.0%5910544.2%CAug 11, 2026
GPQA80.5%8623463.5%CAug 11, 2026

LongCat Flash Thinking Arena Results

Preference and agent-evaluation results for the default version.

No Arena results

The default version does not have a matching Arena result yet.

LongCat Flash Thinking Runtime Performance

Provider-specific output speed and catalog latency for LongCat-Flash-Thinking-2601. Runtime does not affect the capability score.

1 rows
Columns

Show columns

Meituan100 tok/s3 s128K128KAug 11, 2026

Output Speed is generated output tokens received per second. Catalog latency is reported separately from observed provider TTFT.

LongCat Flash Thinking Pricing

Official vendor API pricing appears first, followed by individual provider offers.

Official API
N/A
Official provider
N/A
Lowest third-party
N/A
Tracked offerings
0
No provider prices

The default version has no current input or output token prices.

Official prices use only the vendor's configured official Provider and positive standard USD PAYG rates. Third-party offers remain explicitly labeled.

LongCat Flash Thinking Versions

Available versions of this model. The score column identifies the version used in the overall ranking.

2 rows
Columns

Show columns

LongCat-Flash-Thinking-2601Jan 14, 202660.8560B128K128KNoMIT
LongCat-Flash-ThinkingSep 22, 2025N/A560B128K128KNoMIT

What is LongCat Flash Thinking?

Key information about LongCat Flash Thinking and its available data.

LongCat-Flash-Thinking is Meituan's reasoning model built on the LongCat-Flash foundation with 560B total parameters (MoE, ~27B activated). It introduces a training pipeline specifically tuned for advanced reasoning, featuring Re-thinking Mode that delivers parallel reasoning paths for sophisticated decision-making.

Achieves strong performance on mathematical reasoning, agentic tool use, and formal theorem proving benchmarks.

Data as of 2026-08-11.

LongCat Flash Thinking vs nearby models

Open a comparison with the three ranked models immediately above and below this model.

LongCat Flash ThinkingvsMiMoLongCat Flash ThinkingvsGPT-5.1-InstantLongCat Flash ThinkingvsKimi K2 ThinkingLongCat Flash ThinkingvsQwen3.6 27BLongCat Flash ThinkingvsGPT-5.4-miniLongCat Flash ThinkingvsStep 3.5 Flash

Models similar to LongCat Flash Thinking

Recommendations prioritize the same model type and family, then the closest LLMBoard score.

#138-24.0
ME

LongCat Flash Chat

Meituan

36.9 LLMBoard

DetailsCompare
#152-27.8
ME

LongCat Flash Lite

Meituan

33.0 LLMBoard

DetailsCompare
#54+0.1
MA

Kimi K2 Thinking

Moonshot AI

61.0 LLMBoard

DetailsCompare
#56-0.3
AC

Qwen3.6 27B

Alibaba Cloud / Qwen Team

60.5 LLMBoard

DetailsCompare
#57-0.4
OP

GPT-5.4-mini

OpenAI

60.5 LLMBoard

DetailsCompare
#53+0.5
OP

GPT-5.1-Instant

OpenAI

61.3 LLMBoard

DetailsCompare

FAQ

Common questions about LongCat Flash Thinking.

When was LongCat Flash Thinking released?

LongCat Flash Thinking's default version was released on Jan 14, 2026.

How much does LongCat Flash Thinking cost?

No official standard PAYG price is currently available for LongCat Flash Thinking.

Who created LongCat Flash Thinking?

LongCat Flash Thinking was created by Meituan.

What is the context window for LongCat Flash Thinking?

The default version has a 128K token context window.

Is LongCat Flash Thinking open weight?

No. The default version is not marked as having publicly available weights.

How many API providers offer LongCat Flash Thinking?

No provider offering is currently linked to the default version.

What models should I compare LongCat Flash Thinking with?

Nearby ranked alternatives include MiMo, GPT-5.1-Instant, Kimi K2 Thinking.