llmboard.aiAI model intelligence
Home

Model Rankings

OverallOpen ModelsAgentCodingReasoningMathKnowledgeInstruction FollowingTextVision
Image GenerationImage Editing
Video GenerationImage to VideoVideo Editing
Text to SpeechSpeech to Text
Embeddings

Efficiency

Chat Token PricingImage PricingVideo PricingAudio Pricing
Chat Speed & LatencyProvider Reliability

Benchmarks

GPQAMMLU-ProAIME 2025SWE-Bench VerifiedMMLUHumanity's Last ExamLiveCodeBenchMATHHumanEvalMMMU-Pro
All Benchmarks

Tools

Model Directory

Scoring & Data

Scoring & Data
1224 models729 benchmarks

Leaderboard Center

Overall RankingCodingCore BenchmarksPrice & ValueRuntime Performance

Modalities

All ModelsImage GenerationImage EditingVideo GenerationImage-to-VideoVideo EditingText-to-SpeechSpeech-to-TextEmbeddings

Data & Methods

Scoring MethodAll Benchmarks
llmboard.aiCopyright 2026 llmboard.ai

Alibaba Cloud / Qwen Team model product

Qwen3.8 Max

4 trillion total parameters, 95 billion active parameters, and a 262,144-token native context window.

Updated Sep 8, 2026. Default version: Qwen3.8 Max

LLMBoard Score86.9Qwen3.8 Max
Coverage40%33 benchmark families
Context window1MTokens
Official input price$2Alibaba API

On this page

  • Capability
  • Benchmarks
  • Arena
  • Pricing
  • Runtime
  • Specification
  • Versions
  • Similar models
  • About
  • FAQ

Qwen3.8 Max Capability Profile

This profile uses the model's current scored version. Arena ratings and prices are shown separately.

Qwen3.8 Max LLMBoard score breakdown

Qwen3.8 Max Benchmark Results

Benchmark scores for Qwen3.8 Max.

30 of 54 rows
Columns

Show columns

Sort by
Benchmark
Score
Rank
Participants
Percentile
Evidence
Evaluated
BenchmarkAndroidBenchScore75.10%Rank01Participants1Percentile100.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkAndroidWorldScore85.30%Rank01Participants7Percentile100.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkCoWorkBenchScore74.80%Rank01Participants6Percentile100.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkERQAScore77.80%Rank01Participants26Percentile100.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkHealthBenchScore60.20%Rank01Participants9Percentile100.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkIFBenchScore82.80%Rank01Participants39Percentile100.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkLongBench v2Score66.30%Rank01Participants17Percentile100.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkMobileWorldScore77.80%Rank01Participants3Percentile100.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkOSWorld-VerifiedScore86.10%Rank01Participants24Percentile100.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkPaperBenchScore93.00%Rank01Participants3Percentile100.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkPerceptionBenchScore63.50%Rank01Participants2Percentile100.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkPLawBenchScore73.20%Rank01Participants1Percentile100.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkPRBench-FinanceScore58.30%Rank01Participants1Percentile100.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkPRBench-LegalScore57.60%Rank01Participants1Percentile100.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkQwenQoderBenchScore58.40%Rank01Participants1Percentile100.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkQwenReactBenchScore1,724.00 pointsRank01Participants1Percentile100.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkQwenSVGScore1,713.00 pointsRank01Participants2Percentile100.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkQwenSWEBenchScore80.70%Rank01Participants2Percentile100.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkSkillsBenchScore70.20%Rank01Participants9Percentile100.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkVideoMME w sub.Score90.40%Rank01Participants10Percentile100.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkVision2WebScore69.00%Rank01Participants5Percentile100.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkWorkspace BenchScore67.70%Rank01Participants4Percentile100.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkMLS-Bench LiteScore41.00%Rank02Participants3Percentile50.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkWideSearchScore81.90%Rank02Participants11Percentile90.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkAgents' Last ExamScore52.40%Rank03Participants16Percentile86.67%EvidenceCEvaluatedSep 8, 2026
BenchmarkHumanity's Last Exam (with tools, text-only)Score56.20%Rank03Participants5Percentile50.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkLVBenchScore81.80%Rank03Participants28Percentile92.59%EvidenceCEvaluatedSep 8, 2026
BenchmarkMRCR v2 (8-needle)Score92.90%Rank03Participants24Percentile91.30%EvidenceCEvaluatedSep 8, 2026
BenchmarkRealWorldQAScore88.00%Rank03Participants31Percentile93.33%EvidenceCEvaluatedSep 8, 2026
BenchmarkLM Arena Vision Style ControlScore1,299.71 ratingRank03Participants107Percentile98.11%EvidenceAEvaluatedAug 27, 2026

Qwen3.8 Max Arena Results

Preference and agent-evaluation results for the default version.

30 of 100 rows
Columns

Show columns

Sort by
Arena
Category
Rank
Rating / score
Votes
Observations
Result date
Arenavision style controlCategoryocrRank02Rating / score1,314.73Votes5,092ObservationsN/AResult dateAug 27, 2026
ArenatextCategorykoreanRank03Rating / score1,471.41Votes309ObservationsN/AResult dateSep 2, 2026
Arenatext factualityCategoryindustry life and physical and social scienceRank03Rating / score1,516.49Votes1,872ObservationsN/AResult dateSep 2, 2026
Arenavision style controlCategoryoverallRank03Rating / score1,299.71Votes7,244ObservationsN/AResult dateAug 27, 2026
ArenawebdevCategoryimage to webdevRank03Rating / score1,618.44Votes1,869ObservationsN/AResult dateAug 25, 2026
ArenatextCategoryindustry life and physical and social scienceRank04Rating / score1,518.80Votes2,045ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryindustry medicine and healthcareRank04Rating / score1,509.56Votes873ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryspanishRank04Rating / score1,511.67Votes371ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategoryindustry medicine and healthcareRank04Rating / score1,516.57Votes873ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategoryspanishRank04Rating / score1,496.28Votes371ObservationsN/AResult dateSep 2, 2026
Arenavision style controlCategoryenglishRank04Rating / score1,295.05Votes2,403ObservationsN/AResult dateAug 27, 2026
Arenavision style controlCategoryhomeworkRank04Rating / score1,334.05Votes630ObservationsN/AResult dateAug 27, 2026
ArenawebdevCategoryoverallRank04Rating / score1,686.07Votes1,868ObservationsN/AResult dateSep 5, 2026
ArenawebdevCategorywebdevRank04Rating / score1,686.07Votes1,868ObservationsN/AResult dateSep 5, 2026
ArenawebdevCategorywebdev-reactRank04Rating / score1,702.44Votes1,326ObservationsN/AResult dateSep 5, 2026
ArenavisionCategorydiagramRank05Rating / score1,326.63Votes1,837ObservationsN/AResult dateAug 27, 2026
ArenavisionCategoryenglishRank05Rating / score1,311.56Votes2,403ObservationsN/AResult dateAug 27, 2026
ArenavisionCategoryocrRank05Rating / score1,321.94Votes5,092ObservationsN/AResult dateAug 27, 2026
ArenavisionCategoryoverallRank05Rating / score1,312.86Votes7,244ObservationsN/AResult dateAug 27, 2026
Arenavision style controlCategorydiagramRank05Rating / score1,326.40Votes1,837ObservationsN/AResult dateAug 27, 2026
ArenawebdevCategorywebdev-htmlRank05Rating / score1,657.16Votes184ObservationsN/AResult dateSep 5, 2026
ArenatextCategoryindustry business and management and financial operationsRank06Rating / score1,485.88Votes2,602ObservationsN/AResult dateSep 2, 2026
Arenatext factualityCategoryhard prompts englishRank06Rating / score1,515.76Votes3,547ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategoryindustry life and physical and social scienceRank06Rating / score1,520.73Votes2,045ObservationsN/AResult dateSep 2, 2026
ArenavisionCategoryhomeworkRank06Rating / score1,336.40Votes630ObservationsN/AResult dateAug 27, 2026
ArenatextCategorychineseRank07Rating / score1,536.69Votes863ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryfrenchRank07Rating / score1,500.94Votes488ObservationsN/AResult dateSep 2, 2026
Arenatext factualityCategoryindustry business and management and financial operationsRank07Rating / score1,488.41Votes2,412ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategorykoreanRank07Rating / score1,465.07Votes309ObservationsN/AResult dateSep 2, 2026
ArenavisionCategorycreative writing visionRank07Rating / score1,318.54Votes459ObservationsN/AResult dateAug 27, 2026

Qwen3.8 Max Pricing

Official vendor API pricing appears first, followed by individual provider offers.

Official API
$2 input, $6 output per 1M
Official provider
Alibaba
Lowest third-party
From $1.6 input, $4.8 output per 1M via Vancine
Tracked offerings
31
30 of 31 rows
Columns

Show columns

Sort by
Provider
Provider model ID
Region
Input / 1M
Output / 1M
Context
Updated
ProviderAlibaba Token PlanProvider model IDqwen3.8-maxRegionglobalInput / 1MN/AOutput / 1MN/AContext1MUpdatedSep 8, 2026
ProviderKenariProvider model IDqwen3-8-maxRegionglobalInput / 1MN/AOutput / 1MN/AContext1MUpdatedSep 8, 2026
ProviderSCNet Token PlanProvider model IDQwen3.8-MaxRegionglobalInput / 1MN/AOutput / 1MN/AContext1MUpdatedSep 8, 2026
ProviderAlibaba Token Plan (China)Provider model IDqwen3.8-maxRegionglobalInput / 1MN/AOutput / 1MN/AContext1MUpdatedSep 8, 2026
ProviderVancineProvider model IDqwen3.8-maxRegionglobalInput / 1M$1.6Output / 1M$4.8Context1MUpdatedSep 8, 2026
ProviderDeep InfraProvider model IDQwen/Qwen3.8-MaxRegionglobalInput / 1M$1.65Output / 1M$4.95Context256KUpdatedSep 8, 2026
ProviderAIHubMixProvider model IDqwen3.8-maxRegionglobalInput / 1M$1.69Output / 1M$5.07Context991KUpdatedSep 8, 2026
ProviderAlibaba (China)Provider model IDqwen3.8-maxRegionglobalInput / 1M$1.78Output / 1M$5.33Context1MUpdatedSep 8, 2026
ProviderSCX.aiProvider model IDQwen3.8-MaxRegionglobalInput / 1M$1.82Output / 1M$5.45Context1MUpdatedSep 8, 2026
ProviderDevPass (LLM Gateway)Provider model IDqwen3.8-maxRegionglobalInput / 1M$1.82Output / 1M$5.45Context1MUpdatedSep 8, 2026
ProviderCrossModelProvider model IDqwen/qwen3.8-maxRegionglobalInput / 1M$1.88Output / 1M$5.63Context1MUpdatedSep 8, 2026
ProviderOrcaRouterProvider model IDqwen/qwen3.8-maxRegionglobalInput / 1M$2Output / 1M$6Context1MUpdatedSep 8, 2026
ProviderNanoGPTProvider model IDqwen3.8-maxRegionglobalInput / 1M$2Output / 1M$6Context991KUpdatedSep 8, 2026
ProviderOpenCode GoProvider model IDqwen3.8-maxRegionglobalInput / 1M$2Output / 1M$6Context1MUpdatedSep 8, 2026
ProviderLLM GatewayProvider model IDalibaba/qwen3.8-maxRegionglobalInput / 1M$2Output / 1M$6Context1MUpdatedSep 8, 2026
ProviderMerge GatewayProvider model IDqwen/qwen3.8-maxRegionglobalInput / 1M$2Output / 1M$6Context1MUpdatedSep 8, 2026
ProviderCloudflare AI GatewayProvider model IDalibaba/qwen3.8-maxRegionglobalInput / 1M$2Output / 1M$6Context1MUpdatedSep 8, 2026
ProviderDigitalOceanProvider model IDqwen3.8-maxRegionglobalInput / 1M$2Output / 1M$6Context1MUpdatedSep 8, 2026
ProviderAlibabaProvider model IDqwen3.8-maxRegionglobalInput / 1M$2Output / 1M$6Context1MUpdatedSep 8, 2026
ProviderOfoxProvider model IDbailian/qwen3.8-maxRegionglobalInput / 1M$2Output / 1M$6Context1MUpdatedSep 8, 2026
ProviderRequestyProvider model IDqwen3.8-maxRegionglobalInput / 1M$2Output / 1M$6Context1MUpdatedSep 8, 2026
ProviderCharm HyperProvider model IDqwen3.8-maxRegionglobalInput / 1M$2Output / 1M$6Context1MUpdatedSep 8, 2026
ProviderModalProvider model IDQwen/Qwen3.8-2.4T-A95BRegionglobalInput / 1M$2Output / 1M$6Context1MUpdatedSep 8, 2026
ProviderClinePassProvider model IDcline-pass/qwen3.8-maxRegionglobalInput / 1M$2Output / 1M$6Context1MUpdatedSep 8, 2026
ProviderEmpirioLabs AIProvider model IDqwen3-8-maxRegionglobalInput / 1M$2Output / 1M$6Context1MUpdatedSep 8, 2026
ProviderVercel AI GatewayProvider model IDalibaba/qwen3.8-maxRegionglobalInput / 1M$2Output / 1M$6Context1MUpdatedSep 8, 2026
ProviderEden AIProvider model IDqwen/qwen3.8-maxRegionglobalInput / 1M$2Output / 1M$6Context1MUpdatedSep 8, 2026
ProviderFireworks AIProvider model IDaccounts/fireworks/models/qwen3p8-maxRegionglobalInput / 1M$2Output / 1M$6Context262.1KUpdatedSep 8, 2026
ProviderAbacusProvider model IDqwen3.8-maxRegionglobalInput / 1M$2Output / 1M$6Context1MUpdatedSep 8, 2026
Providerabove.devProvider model IDqwen3.8-maxRegionglobalInput / 1M$2.2Output / 1M$6.6Context1MUpdatedSep 8, 2026

Qwen3.8 Max Runtime Performance

Provider-specific output speed and catalog latency for Qwen3.8 Max. Runtime does not affect the capability score.

1 row
Columns

Show columns

Sort by
Provider
Output Speed
Catalog Latency
Max Input
Max Output
Updated
ProviderDeepInfraOutput Speed60.87 tok/sCatalog Latency8.02 sMax Input256KMax Output131.1KUpdatedSep 8, 2026

Output Speed is generated output tokens received per second. Catalog latency is reported separately from observed provider TTFT.

Qwen3.8 Max Specifications

Technical details for the model's default version.

Version
Qwen3.8 Max
Released
Aug 2, 2026
Knowledge cutoff
Unknown
Parameters
2.4T
Context window
1M
Max output
131.1K
Inputs
image, text, video
Outputs
text
Open weights
Yes
License
Qwen3.8-Max License

Qwen3.8 Max Versions

Available versions of this model. The score column identifies the version used in the overall ranking.

1 row
Columns

Show columns

Sort by
Version
Released
LLMBoard
Parameters
Context
Max output
Open weights
License
VersionQwen3.8 MaxReleasedAug 2, 2026LLMBoard86.93Parameters2.4TContext1MMax output131.1KOpen weightsYesLicenseQwen3.8-Max License

Models similar to Qwen3.8 Max

Recommendations prioritize the same model type and family, then the closest LLMBoard score.

#21-5.49
AC

Qwen3.8 Flash Next

Alibaba Cloud / Qwen Team

81.44 LLMBoard

Details
#22-5.58
AC

Qwen3.8 Flash

Alibaba Cloud / Qwen Team

81.35 LLMBoard

Details
#31-10.30
AC

Qwen3.7 Max

Alibaba Cloud / Qwen Team

76.63 LLMBoard

Details
#36-11.99
AC

Qwen3.8 27B

Alibaba Cloud / Qwen Team

74.94 LLMBoard

Details
#42-15.43
AC

Qwen3.7 Plus

Alibaba Cloud / Qwen Team

71.50 LLMBoard

Details
#57-21.20
AC

Qwen3.6 Plus

Alibaba Cloud / Qwen Team

65.73 LLMBoard

Details

What is Qwen3.8 Max?

Key information about Qwen3.8 Max and its available data.

4T-A95B open-weight checkpoint, is a mixture-of-experts model from Alibaba Cloud / Qwen Team with always-on thinking and a context window extensible to about 1 million tokens. 8 Max API family with text and image input.

Data as of 2026-09-08.

FAQ

Common questions about Qwen3.8 Max.

When was Qwen3.8 Max released?

Qwen3.8 Max's default version was released on Aug 2, 2026.

How much does Qwen3.8 Max cost?

Qwen3.8 Max's official API price is $2 per million input tokens and $6 per million output tokens via Alibaba. The lowest tracked third-party offer starts at $1.6 input and $4.8 output via Vancine.

Who created Qwen3.8 Max?

Qwen3.8 Max was created by Alibaba Cloud / Qwen Team.

What is the context window for Qwen3.8 Max?

The default version has a 1M token context window.

Is Qwen3.8 Max open weight?

Yes. The default version is marked as open weight under Qwen3.8-Max License.

How many API providers offer Qwen3.8 Max?

31 provider offerings are linked to the default version.

Official prices use only the vendor's configured official Provider and positive standard USD PAYG rates. Third-party offers remain explicitly labeled.