llmboard.aiAI model intelligence
Home

Model Rankings

OverallOpen ModelsAgentCodingReasoningMathKnowledgeInstruction FollowingTextVision
Image GenerationImage Editing
Video GenerationImage to VideoVideo Editing
Text to SpeechSpeech to Text
Embeddings

Efficiency

Chat Token PricingImage PricingVideo PricingAudio Pricing
Chat Speed & LatencyProvider Reliability

Benchmarks

GPQAMMLU-ProAIME 2025SWE-Bench VerifiedMMLUHumanity's Last ExamLiveCodeBenchMATHHumanEvalMMMU-Pro
All Benchmarks

Tools

Model DirectoryCompare Models

Scoring & Data

Scoring & Data
393 models668 benchmarks

Leaderboard Center

Overall RankingCodingCore BenchmarksPrice & ValueRuntime Performance

Modalities

All ModelsImage GenerationImage EditingVideo GenerationImage-to-VideoVideo EditingText-to-SpeechSpeech-to-TextEmbeddings

Data & Methods

Scoring MethodAll BenchmarksReasoningMath

Vendors

All VendorsOpenAIAnthropicGoogle
llmboard.aiCopyright 2026 llmboard.ai

Meituan model product

LongCat Flash Thinking

LongCat-Flash-Thinking is Meituan's reasoning model built on the LongCat-Flash foundation with 560B total parameters (MoE, ~27B activated).

Updated Aug 17, 2026. Default version: LongCat-Flash-Thinking-2601

Compare
LLMBoard Score60.3LongCat-Flash-Thinking-2601
Coverage80%11 benchmark families
Context window128KTokens
Official input priceN/AOfficial price unavailable

On this page

  • Capability
  • Benchmarks
  • Arena
  • Pricing
  • Runtime
  • Specification
  • Versions
  • Compare
  • Similar models
  • About
  • FAQ

LongCat Flash Thinking Capability Profile

This profile uses the model's current scored version. Arena ratings and prices are shown separately.

LongCat-Flash-Thinking-2601 LLMBoard score breakdown

LongCat Flash Thinking Benchmark Results

Benchmark scores for LongCat-Flash-Thinking-2601.

11 rows
Columns

Show columns

Sort by
Benchmark
Score
Rank
Participants
Percentile
Evidence
Evaluated
BenchmarkTau2 AirlineScore76.5%Rank01Participants23Percentile100.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkTau2 TelecomScore99.3%Rank02Participants35Percentile97.1%EvidenceCEvaluatedAug 17, 2026
BenchmarkBrowseComp-zhScore69.0%Rank04Participants13Percentile75.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkTau2 RetailScore88.6%Rank04Participants26Percentile88.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkLiveCodeBenchScore82.8%Rank08Participants75Percentile90.5%EvidenceCEvaluatedAug 17, 2026
BenchmarkAIME 2025Score99.6%Rank09Participants115Percentile93.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkIMO-AnswerBenchScore78.6%Rank19Participants20Percentile5.3%EvidenceCEvaluatedAug 17, 2026
BenchmarkBrowseCompScore56.6%Rank39Participants62Percentile37.7%EvidenceCEvaluatedAug 17, 2026
BenchmarkHumanity's Last ExamScore25.2%Rank53Participants99Percentile46.9%EvidenceCEvaluatedAug 17, 2026
BenchmarkSWE-Bench VerifiedScore70.0%Rank64Participants111Percentile42.7%EvidenceCEvaluatedAug 17, 2026
BenchmarkGPQAScore80.5%Rank90Participants239Percentile62.6%EvidenceCEvaluatedAug 17, 2026

LongCat Flash Thinking Arena Results

Preference and agent-evaluation results for the default version.

No Arena results

The default version does not have a matching Arena result yet.

LongCat Flash Thinking Pricing

Official vendor API pricing appears first, followed by individual provider offers.

Official API
N/A
Official provider
N/A
Lowest third-party
N/A
Tracked offerings
0
No provider prices

The default version has no current input or output token prices.

Official prices use only the vendor's configured official Provider and positive standard USD PAYG rates. Third-party offers remain explicitly labeled.

LongCat Flash Thinking Runtime Performance

Provider-specific output speed and catalog latency for LongCat-Flash-Thinking-2601. Runtime does not affect the capability score.

1 rows
Columns

Show columns

Sort by
Provider
Output Speed
Catalog Latency
Max Input
Max Output
Updated
ProviderMeituanOutput Speed100 tok/sCatalog Latency3 sMax Input128KMax Output128KUpdatedAug 17, 2026

Output Speed is generated output tokens received per second. Catalog latency is reported separately from observed provider TTFT.

LongCat Flash Thinking Specifications

Technical details for the model's default version.

Version
LongCat-Flash-Thinking-2601
Released
Jan 14, 2026
Knowledge cutoff
Unknown
Parameters
560B
Context window
128K
Max output
128K
Inputs
text
Outputs
text
Open weights
No
License
MIT

LongCat Flash Thinking Versions

Available versions of this model. The score column identifies the version used in the overall ranking.

2 rows
Columns

Show columns

Sort by
Version
Released
LLMBoard
Parameters
Context
Max output
Open weights
License
VersionLongCat-Flash-Thinking-2601ReleasedJan 14, 2026LLMBoard60.3Parameters560BContext128KMax output128KOpen weightsNoLicenseMIT
VersionLongCat-Flash-ThinkingReleasedSep 22, 2025LLMBoardN/AParameters560BContext128KMax output128KOpen weightsNoLicenseMIT

LongCat Flash Thinking vs nearby models

Open a comparison with the three ranked models immediately above and below this model.

LongCat Flash ThinkingvsGPT-5.1-InstantLongCat Flash ThinkingvsKimi K2 ThinkingLongCat Flash ThinkingvsStep 3.5 FlashLongCat Flash ThinkingvsGPT-5.4-miniLongCat Flash ThinkingvsQwen3.6 27BLongCat Flash ThinkingvsGemma 4 31B

Models similar to LongCat Flash Thinking

Recommendations prioritize the same model type and family, then the closest LLMBoard score.

#149-24.1
ME

LongCat Flash Chat

Meituan

36.1 LLMBoard

DetailsCompare
#164-28.1
ME

LongCat Flash Lite

Meituan

32.2 LLMBoard

DetailsCompare
#62+0.0
ST

Step 3.5 Flash

StepFun

60.3 LLMBoard

DetailsCompare
#64-0.1
OP

GPT-5.4-mini

OpenAI

60.2 LLMBoard

DetailsCompare
#61+0.3
MA

Kimi K2 Thinking

Moonshot AI

60.6 LLMBoard

DetailsCompare
#65-0.3
AC

Qwen3.6 27B

Alibaba Cloud / Qwen Team

59.9 LLMBoard

DetailsCompare

What is LongCat Flash Thinking?

Key information about LongCat Flash Thinking and its available data.

LongCat-Flash-Thinking is Meituan's reasoning model built on the LongCat-Flash foundation with 560B total parameters (MoE, ~27B activated). It introduces a training pipeline specifically tuned for advanced reasoning, featuring Re-thinking Mode that delivers parallel reasoning paths for sophisticated decision-making.

Achieves strong performance on mathematical reasoning, agentic tool use, and formal theorem proving benchmarks.

Data as of 2026-08-17.

FAQ

Common questions about LongCat Flash Thinking.

When was LongCat Flash Thinking released?

LongCat Flash Thinking's default version was released on Jan 14, 2026.

How much does LongCat Flash Thinking cost?

No official standard PAYG price is currently available for LongCat Flash Thinking.

Who created LongCat Flash Thinking?

LongCat Flash Thinking was created by Meituan.

What is the context window for LongCat Flash Thinking?

The default version has a 128K token context window.

Is LongCat Flash Thinking open weight?

No. The default version is not marked as having publicly available weights.

How many API providers offer LongCat Flash Thinking?

No provider offering is currently linked to the default version.

What models should I compare LongCat Flash Thinking with?

Nearby ranked alternatives include GPT-5.1-Instant, Kimi K2 Thinking, Step 3.5 Flash.