llmboard.aiAI model intelligence
Home

Model Rankings

OverallOpen ModelsAgentCodingReasoningMathKnowledgeInstruction FollowingTextVision
Image GenerationImage Editing
Video GenerationImage to VideoVideo Editing
Text to SpeechSpeech to Text
Embeddings

Efficiency

Chat Token PricingImage PricingVideo PricingAudio Pricing
Chat Speed & LatencyProvider Reliability

Benchmarks

GPQAMMLU-ProAIME 2025SWE-Bench VerifiedMMLUHumanity's Last ExamLiveCodeBenchMATHHumanEvalMMMU-Pro
All Benchmarks

Tools

Model Directory

Scoring & Data

Scoring & Data
1224 models729 benchmarks

Leaderboard Center

Overall RankingCodingCore BenchmarksPrice & ValueRuntime Performance

Modalities

All ModelsImage GenerationImage EditingVideo GenerationImage-to-VideoVideo EditingText-to-SpeechSpeech-to-TextEmbeddings

Data & Methods

Scoring MethodAll Benchmarks
llmboard.aiCopyright 2026 llmboard.ai

Alibaba Cloud / Qwen Team model product

Qwen3.8 Flash

8 Flash is a managed QwenCloud / OpenRouter API model with multimodal understanding, tool use, structured output, and a 1M-token context window.

Updated Sep 8, 2026. Default version: Qwen3.8 Flash

LLMBoard Score81.3Qwen3.8 Flash
Coverage40%22 benchmark families
Context window1MTokens
Official input price$0.15Alibaba API

On this page

  • Capability
  • Benchmarks
  • Arena
  • Pricing
  • Runtime
  • Specification
  • Versions
  • Similar models
  • About
  • FAQ

Qwen3.8 Flash Capability Profile

This profile uses the model's current scored version. Arena ratings and prices are shown separately.

Qwen3.8 Flash LLMBoard score breakdown

Qwen3.8 Flash Benchmark Results

Benchmark scores for Qwen3.8 Flash.

22 rows
Columns

Show columns

Sort by
Benchmark
Score
Rank
Participants
Percentile
Evidence
Evaluated
BenchmarkClawEval-MMScore60.40%Rank02Participants6Percentile80.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkLiveCodeBench v6Score91.90%Rank02Participants62Percentile98.36%EvidenceCEvaluatedSep 8, 2026
BenchmarkRealWorldQAScore88.50%Rank02Participants31Percentile96.67%EvidenceCEvaluatedSep 8, 2026
BenchmarkRecreationBenchScore49.90%Rank02Participants3Percentile50.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkAndroidWorldScore84.50%Rank03Participants7Percentile66.67%EvidenceCEvaluatedSep 8, 2026
BenchmarkCoWorkBenchScore73.90%Rank03Participants6Percentile60.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkERQAScore72.30%Rank03Participants26Percentile92.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkMathVisionScore95.70%Rank03Participants35Percentile94.12%EvidenceCEvaluatedSep 8, 2026
BenchmarkVision2WebScore64.00%Rank03Participants5Percentile50.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkJob BenchScore55.70%Rank04Participants8Percentile57.14%EvidenceCEvaluatedSep 8, 2026
BenchmarkAgents' Last ExamScore51.20%Rank05Participants16Percentile73.33%EvidenceCEvaluatedSep 8, 2026
BenchmarkCharXiv-RScore90.60%Rank05Participants55Percentile92.59%EvidenceCEvaluatedSep 8, 2026
BenchmarkIFBenchScore81.30%Rank05Participants39Percentile89.47%EvidenceCEvaluatedSep 8, 2026
BenchmarkSWE-bench MultilingualScore81.00%Rank05Participants43Percentile90.48%EvidenceCEvaluatedSep 8, 2026
BenchmarkToolathlonScore73.50%Rank06Participants41Percentile87.50%EvidenceCEvaluatedSep 8, 2026
BenchmarkLVBenchScore76.60%Rank07Participants28Percentile77.78%EvidenceCEvaluatedSep 8, 2026
BenchmarkNL2RepoScore48.10%Rank10Participants22Percentile57.14%EvidenceCEvaluatedSep 8, 2026
BenchmarkOSWorld 2.0Score19.40%Rank11Participants11Percentile0.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkSWE-Bench ProScore62.50%Rank13Participants55Percentile77.78%EvidenceCEvaluatedSep 8, 2026
BenchmarkDeepSWE 1.1Score58.70%Rank18Participants31Percentile43.33%EvidenceCEvaluatedSep 8, 2026
BenchmarkGPQAScore91.70%Rank20Participants247Percentile92.28%EvidenceCEvaluatedSep 8, 2026
BenchmarkHumanity's Last ExamScore35.90%Rank46Participants103Percentile55.88%EvidenceCEvaluatedSep 8, 2026

Qwen3.8 Flash Arena Results

Preference and agent-evaluation results for the default version.

No Arena results

The default version does not have a matching Arena result yet.

Qwen3.8 Flash Pricing

Official vendor API pricing appears first, followed by individual provider offers.

Official API
$0.15 input, $0.47 output per 1M
Official provider
Alibaba
Lowest third-party
From $0.12 input, $0.38 output per 1M via Vancine
Tracked offerings
19
19 rows
Columns

Show columns

Sort by
Provider
Provider model ID
Region
Input / 1M
Output / 1M
Context
Updated
ProviderAlibaba Token PlanProvider model IDqwen3.8-flashRegionglobalInput / 1MN/AOutput / 1MN/AContext1MUpdatedSep 8, 2026
ProviderAlibaba Token Plan (China)Provider model IDqwen3.8-flashRegionglobalInput / 1MN/AOutput / 1MN/AContext1MUpdatedSep 8, 2026
ProviderNaNProvider model IDqwen3.8-flashRegionglobalInput / 1MN/AOutput / 1MN/AContext262.1KUpdatedSep 8, 2026
ProviderSCNet Token PlanProvider model IDQwen3.8-FlashRegionglobalInput / 1MN/AOutput / 1MN/AContext1MUpdatedSep 8, 2026
ProviderAlibaba (China)Provider model IDqwen3.8-flashRegionglobalInput / 1M$0.1188Output / 1M$0.4007Context1MUpdatedSep 8, 2026
ProviderVancineProvider model IDqwen3.8-flashRegionglobalInput / 1M$0.12Output / 1M$0.38Context1MUpdatedSep 8, 2026
ProviderCrossModelProvider model IDqwen/qwen3.8-flashRegionglobalInput / 1M$0.13Output / 1M$0.43Context1MUpdatedSep 8, 2026
ProviderOpenRouterProvider model IDqwen/qwen3.8-flashRegionglobalInput / 1M$0.15Output / 1M$0.47Context1MUpdatedSep 8, 2026
ProviderOpenCode GoProvider model IDqwen3.8-flashRegionglobalInput / 1M$0.15Output / 1M$0.47Context1MUpdatedSep 8, 2026
ProviderLLM GatewayProvider model IDalibaba/qwen3.8-flashRegionglobalInput / 1M$0.15Output / 1M$0.47Context1MUpdatedSep 8, 2026
ProviderAlibabaProvider model IDqwen3.8-flashRegionglobalInput / 1M$0.15Output / 1M$0.47Context1MUpdatedSep 8, 2026
ProviderCharm HyperProvider model IDqwen3.8-flashRegionglobalInput / 1M$0.15Output / 1M$0.47Context1MUpdatedSep 8, 2026
ProviderEden AIProvider model IDqwen/qwen3.8-flashRegionglobalInput / 1M$0.15Output / 1M$0.47Context1MUpdatedSep 8, 2026
ProviderDevPass (LLM Gateway)Provider model IDqwen3.8-flashRegionglobalInput / 1M$0.15Output / 1M$0.47Context1MUpdatedSep 8, 2026
ProviderKilo GatewayProvider model IDqwen/qwen3.8-flashRegionglobalInput / 1M$0.15Output / 1M$0.47Context1MUpdatedSep 8, 2026
ProviderNanoGPTProvider model IDalibaba/qwen3.8-flashRegionglobalInput / 1M$0.16Output / 1M$0.47Context991.8KUpdatedSep 8, 2026
ProviderRequestyProvider model IDqwen3.8-flashRegionglobalInput / 1M$0.16Output / 1M$0.47Context1MUpdatedSep 8, 2026
ProviderEmpirioLabs AIProvider model IDqwen3-8-flashRegionglobalInput / 1M$0.16Output / 1M$0.47Context1MUpdatedSep 8, 2026
ProviderVercel AI GatewayProvider model IDalibaba/qwen3.8-flashRegionglobalInput / 1M$0.16Output / 1M$0.47Context991KUpdatedSep 8, 2026

Qwen3.8 Flash Runtime Performance

Provider-specific output speed and catalog latency for Qwen3.8 Flash. Runtime does not affect the capability score.

No runtime data

No provider-specific speed or latency record is linked to the default version yet.

Output Speed is generated output tokens received per second. Catalog latency is reported separately from observed provider TTFT.

Qwen3.8 Flash Specifications

Technical details for the model's default version.

Version
Qwen3.8 Flash
Released
Aug 26, 2026
Knowledge cutoff
Unknown
Parameters
125B
Context window
1M
Max output
131.1K
Inputs
image, text, video
Outputs
text
Open weights
Yes
License
Proprietary

Qwen3.8 Flash Versions

Available versions of this model. The score column identifies the version used in the overall ranking.

1 row
Columns

Show columns

Sort by
Version
Released
LLMBoard
Parameters
Context
Max output
Open weights
License
VersionQwen3.8 FlashReleasedAug 26, 2026LLMBoard81.35Parameters125BContext1MMax output131.1KOpen weightsYesLicenseProprietary

Models similar to Qwen3.8 Flash

Recommendations prioritize the same model type and family, then the closest LLMBoard score.

#21+0.09
AC

Qwen3.8 Flash Next

Alibaba Cloud / Qwen Team

81.44 LLMBoard

Details
#31-4.72
AC

Qwen3.7 Max

Alibaba Cloud / Qwen Team

76.63 LLMBoard

Details
#10+5.58
AC

Qwen3.8 Max

Alibaba Cloud / Qwen Team

86.93 LLMBoard

Details
#36-6.41
AC

Qwen3.8 27B

Alibaba Cloud / Qwen Team

74.94 LLMBoard

Details
#42-9.85
AC

Qwen3.7 Plus

Alibaba Cloud / Qwen Team

71.50 LLMBoard

Details
#57-15.62
AC

Qwen3.6 Plus

Alibaba Cloud / Qwen Team

65.73 LLMBoard

Details

What is Qwen3.8 Flash?

Key information about Qwen3.8 Flash and its available data.

8 Flash is an Alibaba Cloud / Qwen Team large language model available through the QwenCloud and OpenRouter APIs, based on Flash-Next. It supports text, image, and video understanding, thinking, function calling, built-in tools, and structured output.

Data as of 2026-09-08.

FAQ

Common questions about Qwen3.8 Flash.

When was Qwen3.8 Flash released?

Qwen3.8 Flash's default version was released on Aug 26, 2026.

How much does Qwen3.8 Flash cost?

Qwen3.8 Flash's official API price is $0.15 per million input tokens and $0.47 per million output tokens via Alibaba. The lowest tracked third-party offer starts at $0.12 input and $0.38 output via Vancine.

Who created Qwen3.8 Flash?

Qwen3.8 Flash was created by Alibaba Cloud / Qwen Team.

What is the context window for Qwen3.8 Flash?

The default version has a 1M token context window.

Is Qwen3.8 Flash open weight?

Yes. The default version is marked as open weight under Proprietary.

How many API providers offer Qwen3.8 Flash?

Official prices use only the vendor's configured official Provider and positive standard USD PAYG rates. Third-party offers remain explicitly labeled.

19 provider offerings are linked to the default version.

Browse runtime rankings