llmboard.aiAI model intelligence
Home

Model Rankings

OverallOpen ModelsAgentCodingReasoningMathKnowledgeInstruction FollowingTextVision
Image GenerationImage Editing
Video GenerationImage to VideoVideo Editing
Text to SpeechSpeech to Text
Embeddings

Efficiency

Chat Token PricingImage PricingVideo PricingAudio Pricing
Chat Speed & LatencyProvider Reliability

Benchmarks

GPQAMMLU-ProAIME 2025SWE-Bench VerifiedMMLUHumanity's Last ExamLiveCodeBenchMATHHumanEvalMMMU-Pro
All Benchmarks

Tools

Model Directory

Scoring & Data

Scoring & Data
1224 models729 benchmarks

Leaderboard Center

Overall RankingCodingCore BenchmarksPrice & ValueRuntime Performance

Modalities

All ModelsImage GenerationImage EditingVideo GenerationImage-to-VideoVideo EditingText-to-SpeechSpeech-to-TextEmbeddings

Data & Methods

Scoring MethodAll Benchmarks
llmboard.aiCopyright 2026 llmboard.ai

Alibaba Cloud / Qwen Team model product

Qwen3 Next 80B A3B Thinking

Qwen3 Next 80B A3B Thinking is a thinking-only large language model from Alibaba Cloud / Qwen Team with 80B total parameters and 3B activated parameters.

Updated Sep 8, 2026. Default version: Qwen3-Next-80B-A3B-Thinking

LLMBoard Score39.6Qwen3-Next-80B-A3B-Thinking
Coverage80%22 benchmark families
Context window65.5KTokens
Official input price$0.50Alibaba API

On this page

  • Capability
  • Benchmarks
  • Arena
  • Pricing
  • Runtime
  • Specification
  • Versions
  • Similar models
  • About
  • FAQ

Qwen3 Next 80B A3B Thinking Capability Profile

This profile uses the model's current scored version. Arena ratings and prices are shown separately.

Qwen3-Next-80B-A3B-Thinking LLMBoard score breakdown

Qwen3 Next 80B A3B Thinking Benchmark Results

Benchmark scores for Qwen3-Next-80B-A3B-Thinking.

25 rows
Columns

Show columns

Sort by
Benchmark
Score
Rank
Participants
Percentile
Evidence
Evaluated
BenchmarkCFEvalScore2,071.00 pointsRank02Participants2Percentile0.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkLiveBench 20241125Score76.60%Rank03Participants14Percentile84.62%EvidenceCEvaluatedSep 8, 2026
BenchmarkBFCL-v3Score72.00%Rank05Participants19Percentile77.78%EvidenceCEvaluatedSep 8, 2026
BenchmarkMulti-IFScore77.80%Rank06Participants23Percentile77.27%EvidenceCEvaluatedSep 8, 2026
BenchmarkOJBenchScore29.70%Rank07Participants9Percentile25.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkWritingBenchScore84.60%Rank09Participants15Percentile42.86%EvidenceCEvaluatedSep 8, 2026
BenchmarkPolyMATHScore56.30%Rank10Participants23Percentile59.09%EvidenceCEvaluatedSep 8, 2026
BenchmarkTAU-bench RetailScore69.60%Rank11Participants25Percentile58.33%EvidenceCEvaluatedSep 8, 2026
BenchmarkArena-Hard v2Score62.30%Rank12Participants19Percentile38.89%EvidenceCEvaluatedSep 8, 2026
BenchmarkTau2 AirlineScore60.50%Rank12Participants24Percentile52.17%EvidenceCEvaluatedSep 8, 2026
BenchmarkMMLU-ProXScore78.70%Rank13Participants32Percentile61.29%EvidenceCEvaluatedSep 8, 2026
BenchmarkIncludeScore78.90%Rank14Participants31Percentile56.67%EvidenceCEvaluatedSep 8, 2026
BenchmarkTAU-bench AirlineScore49.00%Rank15Participants23Percentile36.36%EvidenceCEvaluatedSep 8, 2026
BenchmarkSuperGPQAScore60.80%Rank16Participants34Percentile54.55%EvidenceCEvaluatedSep 8, 2026
BenchmarkHMMT25Score73.90%Rank18Participants28Percentile37.04%EvidenceCEvaluatedSep 8, 2026
BenchmarkMMLU-ReduxScore92.50%Rank18Participants48Percentile63.83%EvidenceCEvaluatedSep 8, 2026
BenchmarkTau2 RetailScore67.80%Rank23Participants27Percentile15.38%EvidenceCEvaluatedSep 8, 2026
BenchmarkIFEvalScore88.90%Rank24Participants68Percentile65.67%EvidenceCEvaluatedSep 8, 2026
BenchmarkTau2 TelecomScore43.90%Rank33Participants36Percentile8.57%EvidenceCEvaluatedSep 8, 2026
BenchmarkMMLU-ProScore82.70%Rank36Participants138Percentile74.45%EvidenceCEvaluatedSep 8, 2026
BenchmarkLiveCodeBench v6Score68.70%Rank41Participants62Percentile34.43%EvidenceCEvaluatedSep 8, 2026
BenchmarkAIME 2025Score87.80%Rank55Participants119Percentile54.24%EvidenceCEvaluatedSep 8, 2026
BenchmarkGPQAScore77.20%Rank108Participants247Percentile56.50%EvidenceCEvaluatedSep 8, 2026
BenchmarkLM Arena TextScore1,367.75 ratingRank124Participants210Percentile41.15%EvidenceAEvaluatedSep 2, 2026
BenchmarkLM Arena Text Style ControlScore1,369.04 ratingRank131Participants210Percentile37.80%EvidenceAEvaluatedSep 2, 2026

Qwen3 Next 80B A3B Thinking Arena Results

Preference and agent-evaluation results for the default version.

30 of 58 rows
Columns

Show columns

Sort by
Arena
Category
Rank
Rating / score
Votes
Observations
Result date
ArenatextCategorymathRank102Rating / score1,398.86Votes816ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryindustry medicine and healthcareRank108Rating / score1,392.91Votes710ObservationsN/AResult dateSep 2, 2026
ArenatextCategorygermanRank111Rating / score1,358.49Votes294ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryenglishRank112Rating / score1,392.37Votes5,862ObservationsN/AResult dateSep 2, 2026
ArenatextCategorykoreanRank112Rating / score1,308.05Votes347ObservationsN/AResult dateSep 2, 2026
ArenatextCategorypolishRank114Rating / score1,346.20Votes711ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategorygermanRank114Rating / score1,364.34Votes294ObservationsN/AResult dateSep 2, 2026
ArenatextCategorychineseRank115Rating / score1,406.50Votes1,042ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategorymathRank115Rating / score1,393.01Votes816ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryindustry life and physical and social scienceRank116Rating / score1,384.14Votes2,097ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryindustry mathematicalRank117Rating / score1,389.22Votes668ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryjapaneseRank117Rating / score1,282.39Votes201ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategorykoreanRank117Rating / score1,317.51Votes347ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryexpertRank118Rating / score1,375.59Votes620ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryindustry business and management and financial operationsRank118Rating / score1,360.64Votes2,479ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategorypolishRank118Rating / score1,359.54Votes711ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryhard prompts englishRank119Rating / score1,384.74Votes3,000ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryindustry software and it servicesRank119Rating / score1,393.77Votes4,687ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategoryjapaneseRank119Rating / score1,295.17Votes201ObservationsN/AResult dateSep 2, 2026
ArenatextCategorycodingRank120Rating / score1,392.12Votes2,645ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryspanishRank122Rating / score1,355.25Votes455ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryexclude tiesRank124Rating / score1,333.58Votes9,532ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryoverallRank124Rating / score1,367.75Votes13,391ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryhard promptsRank125Rating / score1,371.11Votes6,628ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategoryindustry mathematicalRank125Rating / score1,386.78Votes668ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryfrenchRank126Rating / score1,352.07Votes215ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategorychineseRank126Rating / score1,401.87Votes1,042ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategoryindustry medicine and healthcareRank126Rating / score1,396.40Votes710ObservationsN/AResult dateSep 2, 2026
ArenatextCategorynon englishRank127Rating / score1,341.82Votes7,529ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryindustry legal and governmentRank128Rating / score1,363.82Votes839ObservationsN/AResult dateSep 2, 2026

Qwen3 Next 80B A3B Thinking Pricing

Official vendor API pricing appears first, followed by individual provider offers.

Official API
$0.50 input, $6 output per 1M
Official provider
Alibaba
Lowest third-party
From $0.15 input, $0.65 output per 1M via NanoGPT
Tracked offerings
14
13 rows
Columns

Show columns

Sort by
Provider
Provider model ID
Region
Input / 1M
Output / 1M
Context
Updated
ProviderAlibaba (China)Provider model IDqwen3-next-80b-a3b-thinkingRegionglobalInput / 1M$0.144Output / 1M$1.43Context131.1KUpdatedSep 8, 2026
ProviderNanoGPTProvider model IDqwen/qwen3-next-80b-a3b-thinkingRegionglobalInput / 1M$0.15Output / 1M$0.65Context256KUpdatedSep 8, 2026
ProviderJalapeno CloudProvider model IDQwen3-Next-80B-A3B-ThinkingRegionglobalInput / 1M$0.15Output / 1M$1.5Context131.1KUpdatedSep 8, 2026
ProviderJiekou.AIProvider model IDqwen/qwen3-next-80b-a3b-thinkingRegionglobalInput / 1M$0.15Output / 1M$1.5Context65.5KUpdatedSep 8, 2026
ProviderOpenRouterProvider model IDqwen/qwen3-next-80b-a3b-thinkingRegionglobalInput / 1M$0.15Output / 1M$1.2Context262.1KUpdatedSep 8, 2026
ProviderMerge GatewayProvider model IDqwen/qwen3-next-80b-a3b-thinkingRegionglobalInput / 1M$0.15Output / 1M$1.2Context131.1KUpdatedSep 8, 2026
ProviderVercel AI GatewayProvider model IDalibaba/qwen3-next-80b-a3b-thinkingRegionglobalInput / 1M$0.15Output / 1M$1.2Context131.1KUpdatedSep 8, 2026
ProviderEden AIProvider model IDqwen/qwen3-next-80b-a3b-thinkingRegionglobalInput / 1M$0.15Output / 1M$1.2Context131.1KUpdatedSep 8, 2026
ProviderNovitaAIProvider model IDqwen/qwen3-next-80b-a3b-thinkingRegionglobalInput / 1M$0.15Output / 1M$1.5Context131.1KUpdatedSep 8, 2026
ProviderDevPass (LLM Gateway)Provider model IDqwen3-next-80b-a3b-thinkingRegionglobalInput / 1M$0.15Output / 1M$1.2Context131.1KUpdatedSep 8, 2026
ProviderKilo GatewayProvider model IDqwen/qwen3-next-80b-a3b-thinkingRegionglobalInput / 1M$0.15Output / 1M$1.2Context262.1KUpdatedSep 8, 2026
ProviderHugging FaceProvider model IDQwen/Qwen3-Next-80B-A3B-ThinkingRegionglobalInput / 1M$0.30Output / 1M$2Context262.1KUpdatedSep 8, 2026
ProviderAlibabaProvider model IDqwen3-next-80b-a3b-thinkingRegionglobalInput / 1M$0.50Output / 1M$6Context131.1KUpdatedSep 8, 2026

Official prices use only the vendor's configured official Provider and positive standard USD PAYG rates. Third-party offers remain explicitly labeled.

Qwen3 Next 80B A3B Thinking Runtime Performance

Provider-specific output speed and catalog latency for Qwen3-Next-80B-A3B-Thinking. Runtime does not affect the capability score.

No runtime data

No provider-specific speed or latency record is linked to the default version yet.

Output Speed is generated output tokens received per second. Catalog latency is reported separately from observed provider TTFT.

Qwen3 Next 80B A3B Thinking Specifications

Technical details for the model's default version.

Version
Qwen3-Next-80B-A3B-Thinking
Released
Sep 10, 2025
Knowledge cutoff
Unknown
Parameters
80B
Context window
65.5K
Max output
65.5K
Inputs
text
Outputs
text
Open weights
Yes
License
Apache 2.0

Qwen3 Next 80B A3B Thinking Versions

Available versions of this model. The score column identifies the version used in the overall ranking.

1 row
Columns

Show columns

Sort by
Version
Released
LLMBoard
Parameters
Context
Max output
Open weights
License
VersionQwen3-Next-80B-A3B-ThinkingReleasedSep 10, 2025LLMBoard39.60Parameters80BContext65.5KMax output65.5KOpen weightsYesLicenseApache 2.0

Models similar to Qwen3 Next 80B A3B Thinking

Recommendations prioritize the same model type and family, then the closest LLMBoard score.

#146-0.19
AC

Qwen3 VL 32B Thinking

Alibaba Cloud / Qwen Team

39.41 LLMBoard

Details
#144+0.57
AC

Qwen3.5 9B

Alibaba Cloud / Qwen Team

40.17 LLMBoard

Details
#139+1.59
AC

Qwen3 235B A22B

Alibaba Cloud / Qwen Team

41.19 LLMBoard

Details
#137+2.35
AC

Qwen3 VL 235B A22B

Alibaba Cloud / Qwen Team

41.95 LLMBoard

Details
#136+2.67
AC

Qwen3 Max

Alibaba Cloud / Qwen Team

42.27 LLMBoard

Details
#163-4.45
AC

Qwen3 Next 80B A3B

Alibaba Cloud / Qwen Team

35.15 LLMBoard

Details

What is Qwen3 Next 80B A3B Thinking?

Key information about Qwen3 Next 80B A3B Thinking and its available data.

Qwen3 Next 80B A3B Thinking is the thinking variant of the Qwen3-Next series, using hybrid attention, a high-sparsity mixture-of-experts architecture with 512 experts, and Multi-Token Prediction. It supports only thinking mode with automatic <think> tag inclusion.

Data as of 2026-09-08.

FAQ

Common questions about Qwen3 Next 80B A3B Thinking.

When was Qwen3 Next 80B A3B Thinking released?

Qwen3 Next 80B A3B Thinking's default version was released on Sep 10, 2025.

How much does Qwen3 Next 80B A3B Thinking cost?

Qwen3 Next 80B A3B Thinking's official API price is $0.50 per million input tokens and $6 per million output tokens via Alibaba. The lowest tracked third-party offer starts at $0.15 input and $0.65 output via NanoGPT.

Who created Qwen3 Next 80B A3B Thinking?

Qwen3 Next 80B A3B Thinking was created by Alibaba Cloud / Qwen Team.

What is the context window for Qwen3 Next 80B A3B Thinking?

The default version has a 65.5K token context window.

Is Qwen3 Next 80B A3B Thinking open weight?

Yes. The default version is marked as open weight under Apache 2.0.

How many API providers offer Qwen3 Next 80B A3B Thinking?

14 provider offerings are linked to the default version.

Browse runtime rankings