llmboard.aiAI model intelligence
Home

Model Rankings

OverallOpen ModelsAgentCodingReasoningMathKnowledgeInstruction FollowingTextVision
Image GenerationImage Editing
Video GenerationImage to VideoVideo Editing
Text to SpeechSpeech to Text
Embeddings

Efficiency

Chat Token PricingImage PricingVideo PricingAudio Pricing
Chat Speed & LatencyProvider Reliability

Benchmarks

GPQAMMLU-ProAIME 2025SWE-Bench VerifiedMMLUHumanity's Last ExamLiveCodeBenchMATHHumanEvalMMMU-Pro
All Benchmarks

Tools

Model Directory

Scoring & Data

Scoring & Data
1224 models729 benchmarks

Leaderboard Center

Overall RankingCodingCore BenchmarksPrice & ValueRuntime Performance

Modalities

All ModelsImage GenerationImage EditingVideo GenerationImage-to-VideoVideo EditingText-to-SpeechSpeech-to-TextEmbeddings

Data & Methods

Scoring MethodAll Benchmarks
llmboard.aiCopyright 2026 llmboard.ai

Meituan model product

LongCat Flash Thinking

LongCat Flash Thinking is an upgraded Mixture-of-Experts model with 560B total parameters and approximately 27B activated parameters.

Updated Sep 8, 2026. Default version: LongCat-Flash-Thinking-2601

LLMBoard Score59.5LongCat-Flash-Thinking-2601
Coverage20%11 benchmark families
Context window128KTokens
Official input priceN/AOfficial price unavailable

On this page

  • Capability
  • Benchmarks
  • Arena
  • Pricing
  • Runtime
  • Specification
  • Versions
  • Similar models
  • About
  • FAQ

LongCat Flash Thinking Capability Profile

This profile uses the model's current scored version. Arena ratings and prices are shown separately.

LongCat-Flash-Thinking-2601 LLMBoard score breakdown

LongCat Flash Thinking Benchmark Results

Benchmark scores for LongCat-Flash-Thinking-2601.

11 rows
Columns

Show columns

Sort by
Benchmark
Score
Rank
Participants
Percentile
Evidence
Evaluated
BenchmarkTau2 AirlineScore76.50%Rank01Participants24Percentile100.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkTau2 TelecomScore99.30%Rank02Participants36Percentile97.14%EvidenceCEvaluatedSep 8, 2026
BenchmarkBrowseComp-zhScore69.00%Rank04Participants13Percentile75.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkTau2 RetailScore88.60%Rank04Participants27Percentile88.46%EvidenceCEvaluatedSep 8, 2026
BenchmarkLiveCodeBenchScore82.80%Rank08Participants75Percentile90.54%EvidenceCEvaluatedSep 8, 2026
BenchmarkAIME 2025Score99.60%Rank09Participants119Percentile93.22%EvidenceCEvaluatedSep 8, 2026
BenchmarkIMO-AnswerBenchScore78.60%Rank19Participants20Percentile5.26%EvidenceCEvaluatedSep 8, 2026
BenchmarkBrowseCompScore56.60%Rank40Participants63Percentile37.10%EvidenceCEvaluatedSep 8, 2026
BenchmarkHumanity's Last ExamScore25.20%Rank57Participants103Percentile45.10%EvidenceCEvaluatedSep 8, 2026
BenchmarkSWE-Bench VerifiedScore70.00%Rank64Participants113Percentile43.75%EvidenceCEvaluatedSep 8, 2026
BenchmarkGPQAScore80.50%Rank95Participants247Percentile61.79%EvidenceCEvaluatedSep 8, 2026

LongCat Flash Thinking Arena Results

Preference and agent-evaluation results for the default version.

No Arena results

The default version does not have a matching Arena result yet.

LongCat Flash Thinking Pricing

Official vendor API pricing appears first, followed by individual provider offers.

Official API
N/A
Official provider
N/A
Lowest third-party
N/A
Tracked offerings
0
No provider prices

The default version has no current input or output token prices.

Official prices use only the vendor's configured official Provider and positive standard USD PAYG rates. Third-party offers remain explicitly labeled.

LongCat Flash Thinking Runtime Performance

Provider-specific output speed and catalog latency for LongCat-Flash-Thinking-2601. Runtime does not affect the capability score.

No runtime data

No provider-specific speed or latency record is linked to the default version yet.

Output Speed is generated output tokens received per second. Catalog latency is reported separately from observed provider TTFT.

LongCat Flash Thinking Specifications

Technical details for the model's default version.

Version
LongCat-Flash-Thinking-2601
Released
Jan 14, 2026
Knowledge cutoff
Unknown
Parameters
560B
Context window
128K
Max output
128K
Inputs
text
Outputs
text
Open weights
Yes
License
MIT

LongCat Flash Thinking Versions

Available versions of this model. The score column identifies the version used in the overall ranking.

2 rows
Columns

Show columns

Sort by
Version
Released
LLMBoard
Parameters
Context
Max output
Open weights
License
VersionLongCat-Flash-Thinking-2601ReleasedJan 14, 2026LLMBoard59.49Parameters560BContext128KMax output128KOpen weightsYesLicenseMIT
VersionLongCat-Flash-ThinkingReleasedSep 22, 2025LLMBoardN/AParameters560BContext128KMax output128KOpen weightsYesLicenseMIT

Models similar to LongCat Flash Thinking

Recommendations prioritize the same model type and family, then the closest LLMBoard score.

#159-23.64
ME

LongCat Flash Chat

Meituan

35.85 LLMBoard

Details
#175-27.62
ME

LongCat Flash Lite

Meituan

31.87 LLMBoard

Details
#73-0.07
ST

Step 3.5 Flash

StepFun

59.42 LLMBoard

Details
#71+0.25
BA

ERNIE 5.0

Baidu

59.74 LLMBoard

Details
#70+0.50
MA

Kimi K2 Thinking

Moonshot AI

59.99 LLMBoard

Details
#69+0.51
XI

MiMo

Xiaomi

60.00 LLMBoard

Details

What is LongCat Flash Thinking?

Key information about LongCat Flash Thinking and its available data.

LongCat-Flash-Thinking-2601 is an upgraded version of LongCat-Flash-Thinking from Meituan, featuring Heavy Thinking mode and training with structured agentic trajectories and context management. It is evaluated on Agentic Search, Agentic Tool Use, and Tool-Integrated Reasoning benchmarks.

Data as of 2026-09-08.

FAQ

Common questions about LongCat Flash Thinking.

When was LongCat Flash Thinking released?

LongCat Flash Thinking's default version was released on Jan 14, 2026.

How much does LongCat Flash Thinking cost?

No official standard PAYG price is currently available for LongCat Flash Thinking.

Who created LongCat Flash Thinking?

LongCat Flash Thinking was created by Meituan.

What is the context window for LongCat Flash Thinking?

The default version has a 128K token context window.

Is LongCat Flash Thinking open weight?

Yes. The default version is marked as open weight under MIT.

How many API providers offer LongCat Flash Thinking?

No provider offering is currently linked to the default version.

Browse runtime rankings