llmboard.aiAI model intelligence
Home

Model Rankings

OverallOpen ModelsAgentCodingReasoningMathKnowledgeInstruction FollowingTextVision
Image GenerationImage Editing
Video GenerationImage to VideoVideo Editing
Text to SpeechSpeech to Text
Embeddings

Efficiency

Chat Token PricingImage PricingVideo PricingAudio Pricing
Chat Speed & LatencyProvider Reliability

Benchmarks

GPQAMMLU-ProAIME 2025SWE-Bench VerifiedMMLUHumanity's Last ExamLiveCodeBenchMATHHumanEvalMMMU-Pro
All Benchmarks

Tools

Model Directory

Scoring & Data

Scoring & Data
1224 models729 benchmarks

Leaderboard Center

Overall RankingCodingCore BenchmarksPrice & ValueRuntime Performance

Modalities

All ModelsImage GenerationImage EditingVideo GenerationImage-to-VideoVideo EditingText-to-SpeechSpeech-to-TextEmbeddings

Data & Methods

Scoring MethodAll Benchmarks
llmboard.aiCopyright 2026 llmboard.ai

Alibaba Cloud / Qwen Team model product

Qwen3.7 Max

7 Max is a proprietary large language model from Alibaba Cloud / Qwen Team for agent-driven workflows, coding, office automation, MCP, and multi-agent orchestration.

Updated Sep 8, 2026. Default version: Qwen3.7 Max

LLMBoard Score76.6Qwen3.7 Max
Coverage100%37 benchmark families
Context window1MTokens
Official input price$2.5Alibaba API

On this page

  • Capability
  • Benchmarks
  • Arena
  • Pricing
  • Runtime
  • Specification
  • Versions
  • Similar models
  • About
  • FAQ

Qwen3.7 Max Capability Profile

This profile uses the model's current scored version. Arena ratings and prices are shown separately.

Qwen3.7 Max LLMBoard score breakdown

Qwen3.7 Max Benchmark Results

Benchmark scores for Qwen3.7 Max.

30 of 52 rows
Columns

Show columns

Sort by
Benchmark
Score
Rank
Participants
Percentile
Evidence
Evaluated
BenchmarkBFCL-V4Score75.00%Rank01Participants18Percentile100.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkHMMT Feb 26Score97.10%Rank01Participants12Percentile100.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkKernel Bench L3Score96.00%Rank01Participants1Percentile100.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkMAXIFEScore89.20%Rank01Participants11Percentile100.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkMMLU-ProXScore87.00%Rank01Participants32Percentile100.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkMMLU-ReduxScore95.00%Rank01Participants48Percentile100.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkMRCR 128K (8-needle)Score90.40%Rank01Participants2Percentile100.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkPolyMATHScore86.50%Rank01Participants23Percentile100.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkQwenWebBenchScore1,568.00 pointsRank01Participants2Percentile100.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkSuperGPQAScore73.60%Rank01Participants34Percentile100.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkZClawBenchScore64.30%Rank01Participants4Percentile100.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkIncludeScore86.20%Rank02Participants31Percentile96.67%EvidenceCEvaluatedSep 8, 2026
BenchmarkMCP-MarkScore60.80%Rank02Participants8Percentile85.71%EvidenceCEvaluatedSep 8, 2026
BenchmarkMMLU-ProScore89.60%Rank02Participants138Percentile99.27%EvidenceCEvaluatedSep 8, 2026
BenchmarkNOVA-63Score59.00%Rank02Participants11Percentile90.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkQwenSVGScore1,608.00 pointsRank02Participants2Percentile0.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkQwenWorldBenchScore57.30%Rank02Participants2Percentile0.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkSpreadSheetBench-v1Score87.00%Rank02Participants3Percentile50.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkVITA-BenchScore47.90%Rank02Participants10Percentile88.89%EvidenceCEvaluatedSep 8, 2026
BenchmarkCritPTScore11.40%Rank03Participants6Percentile60.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkGlobal PIQAScore91.40%Rank03Participants13Percentile83.33%EvidenceCEvaluatedSep 8, 2026
BenchmarkIFEvalScore94.30%Rank03Participants68Percentile97.01%EvidenceCEvaluatedSep 8, 2026
BenchmarkLiveCodeBench v6Score91.60%Rank03Participants62Percentile96.72%EvidenceCEvaluatedSep 8, 2026
BenchmarkSkillsBenchScore59.20%Rank03Participants9Percentile75.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkWMT24++Score85.80%Rank03Participants23Percentile90.91%EvidenceCEvaluatedSep 8, 2026
BenchmarkIMO-AnswerBenchScore90.00%Rank04Participants20Percentile84.21%EvidenceCEvaluatedSep 8, 2026
BenchmarkSciCodeScore53.50%Rank04Participants24Percentile86.96%EvidenceCEvaluatedSep 8, 2026
BenchmarkClaw-EvalScore65.20%Rank05Participants14Percentile69.23%EvidenceCEvaluatedSep 8, 2026
BenchmarkCoWorkBenchScore67.20%Rank05Participants6Percentile20.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkMathArena ApexScore44.50%Rank05Participants9Percentile50.00%EvidenceCEvaluatedSep 8, 2026

Qwen3.7 Max Arena Results

Preference and agent-evaluation results for the default version.

30 of 66 rows
Columns

Show columns

Sort by
Arena
Category
Rank
Rating / score
Votes
Observations
Result date
Arenatext factualityCategoryindustry software and it servicesRank09Rating / score1,517.62Votes1,550ObservationsN/AResult dateSep 2, 2026
Arenatext factualityCategorylonger queryRank09Rating / score1,497.28Votes1,609ObservationsN/AResult dateSep 2, 2026
Arenatext factualityCategorynon englishRank09Rating / score1,474.12Votes1,899ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryindustry software and it servicesRank12Rating / score1,497.44Votes1,550ObservationsN/AResult dateSep 2, 2026
ArenatextCategorylonger queryRank12Rating / score1,484.98Votes1,609ObservationsN/AResult dateSep 2, 2026
ArenatextCategorymathRank12Rating / score1,493.48Votes218ObservationsN/AResult dateSep 2, 2026
ArenatextCategorynon englishRank12Rating / score1,472.15Votes1,899ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryrussianRank12Rating / score1,483.79Votes375ObservationsN/AResult dateSep 2, 2026
Arenatext factualityCategoryexclude tiesRank12Rating / score1,486.82Votes2,848ObservationsN/AResult dateSep 2, 2026
Arenatext factualityCategoryoverallRank12Rating / score1,478.94Votes3,710ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategorylonger queryRank13Rating / score1,494.12Votes1,609ObservationsN/AResult dateSep 2, 2026
ArenatextCategorychineseRank15Rating / score1,524.24Votes264ObservationsN/AResult dateSep 2, 2026
ArenatextCategorycodingRank15Rating / score1,497.59Votes1,118ObservationsN/AResult dateSep 2, 2026
ArenatextCategorymulti turnRank16Rating / score1,480.03Votes658ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryoverallRank16Rating / score1,474.24Votes3,710ObservationsN/AResult dateSep 2, 2026
Arenatext factualityCategoryhard promptsRank16Rating / score1,498.49Votes2,527ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategorycodingRank16Rating / score1,524.79Votes1,118ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategorymathRank16Rating / score1,489.74Votes218ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryexclude tiesRank17Rating / score1,478.62Votes2,848ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategoryindustry software and it servicesRank17Rating / score1,513.21Votes1,550ObservationsN/AResult dateSep 2, 2026
Arenatext factualityCategoryinstruction followingRank18Rating / score1,473.43Votes1,283ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryindustry entertainment and sports and mediaRank19Rating / score1,445.79Votes681ObservationsN/AResult dateSep 2, 2026
Arenatext factualityCategoryenglishRank19Rating / score1,477.49Votes1,811ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategoryexpertRank19Rating / score1,506.97Votes348ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategoryrussianRank20Rating / score1,484.63Votes375ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryindustry life and physical and social scienceRank21Rating / score1,478.98Votes657ObservationsN/AResult dateSep 2, 2026
ArenatextCategorycreative writingRank22Rating / score1,451.35Votes482ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryenglishRank22Rating / score1,470.84Votes1,811ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryindustry medicine and healthcareRank22Rating / score1,469.84Votes296ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryindustry writing and literature and languageRank22Rating / score1,458.09Votes851ObservationsN/AResult dateSep 2, 2026

Qwen3.7 Max Pricing

Official vendor API pricing appears first, followed by individual provider offers.

Official API
$2.5 input, $7.5 output per 1M
Official provider
Alibaba
Lowest third-party
From $0.825 input, $2.48 output per 1M via Merge Gateway
Tracked offerings
33
30 of 33 rows
Columns

Show columns

Sort by
Provider
Provider model ID
Region
Input / 1M
Output / 1M
Context
Updated
ProviderAlibaba Token PlanProvider model IDqwen3.7-maxRegionglobalInput / 1MN/AOutput / 1MN/AContext1MUpdatedSep 8, 2026
ProviderAlibaba Token Plan (China)Provider model IDqwen3.7-maxRegionglobalInput / 1MN/AOutput / 1MN/AContext1MUpdatedSep 8, 2026
ProviderMerge GatewayProvider model IDqwen/qwen3.7-maxRegionglobalInput / 1M$0.825Output / 1M$2.48Context1MUpdatedSep 8, 2026
ProviderOrcaRouterProvider model IDqwen/qwen3.7-maxRegionglobalInput / 1M$1.25Output / 1M$3.75Context1MUpdatedSep 8, 2026
ProviderCloudflare AI GatewayProvider model IDalibaba/qwen3.7-maxRegionglobalInput / 1M$1.25Output / 1M$3.75Context1MUpdatedSep 8, 2026
ProviderTogether AIProvider model IDQwen/Qwen3.7-MaxRegionglobalInput / 1M$1.25Output / 1M$3.75Context1MUpdatedSep 8, 2026
ProviderNovitaAIProvider model IDqwen/qwen3.7-maxRegionglobalInput / 1M$1.25Output / 1M$3.75Context1MUpdatedSep 8, 2026
ProviderDevPass (LLM Gateway)Provider model IDqwen3.7-maxRegionglobalInput / 1M$1.25Output / 1M$3.75Context1MUpdatedSep 8, 2026
ProviderKilo GatewayProvider model IDqwen/qwen3.7-maxRegionglobalInput / 1M$1.25Output / 1M$3.75Context1MUpdatedSep 8, 2026
ProviderPioneerProvider model IDqwen3.7-maxRegionglobalInput / 1M$1.25Output / 1M$3.75Context991KUpdatedSep 8, 2026
ProviderOpenRouterProvider model IDqwen/qwen3.7-maxRegionglobalInput / 1M$1.48Output / 1M$4.43Context1MUpdatedSep 8, 2026
ProviderAIHubMixProvider model IDqwen3.7-maxRegionglobalInput / 1M$1.69Output / 1M$5.07Context991KUpdatedSep 8, 2026
ProviderCrossModelProvider model IDqwen/qwen3.7-maxRegionglobalInput / 1M$1.88Output / 1M$5.63Context1MUpdatedSep 8, 2026
ProviderNanoGPTProvider model IDqwen3.7-maxRegionglobalInput / 1M$2.5Output / 1M$7.5Context1MUpdatedSep 8, 2026
ProviderDeep InfraProvider model IDQwen/Qwen3.7-MaxRegionglobalInput / 1M$2.5Output / 1M$7.5Context256KUpdatedSep 8, 2026
ProviderImpossiblProvider model IDqwen/qwen3.7-maxRegionglobalInput / 1M$2.5Output / 1M$7.5Context1MUpdatedSep 8, 2026
ProviderAlibaba (China)Provider model IDqwen3.7-maxRegionglobalInput / 1M$2.5Output / 1M$7.5Context1MUpdatedSep 8, 2026
ProviderOpenCode GoProvider model IDqwen3.7-maxRegionglobalInput / 1M$2.5Output / 1M$7.5Context1MUpdatedSep 8, 2026
ProviderLLM GatewayProvider model IDalibaba/qwen3.7-maxRegionglobalInput / 1M$2.5Output / 1M$7.5Context1MUpdatedSep 8, 2026
ProviderAlibaba Coding Plan (China)Provider model IDqwen3.7-maxRegionglobalInput / 1M$2.5Output / 1M$7.5Context1MUpdatedSep 8, 2026
ProviderAlibabaProvider model IDqwen3.7-maxRegionglobalInput / 1M$2.5Output / 1M$7.5Context1MUpdatedSep 8, 2026
ProviderZenMuxProvider model IDqwen/qwen3.7-maxRegionglobalInput / 1M$2.5Output / 1M$7.5Context1MUpdatedSep 8, 2026
ProviderAlibaba Coding PlanProvider model IDqwen3.7-maxRegionglobalInput / 1M$2.5Output / 1M$7.5Context1MUpdatedSep 8, 2026
ProviderGMI CloudProvider model IDQwen/Qwen3.7-MaxRegionglobalInput / 1M$2.5Output / 1M$7.5Context1MUpdatedSep 8, 2026
ProviderOfoxProvider model IDbailian/qwen3.7-maxRegionglobalInput / 1M$2.5Output / 1M$7.5Context1MUpdatedSep 8, 2026
ProviderRequestyProvider model IDqwen3.7-maxRegionglobalInput / 1M$2.5Output / 1M$7.5Context1MUpdatedSep 8, 2026
ProviderCharm HyperProvider model IDqwen3.7-maxRegionglobalInput / 1M$2.5Output / 1M$7.5Context1MUpdatedSep 8, 2026
ProviderClinePassProvider model IDcline-pass/qwen3.7-maxRegionglobalInput / 1M$2.5Output / 1M$7.5Context1MUpdatedSep 8, 2026
ProviderEmpirioLabs AIProvider model IDqwen3-7-maxRegionglobalInput / 1M$2.5Output / 1M$7.5Context1MUpdatedSep 8, 2026
ProviderVercel AI GatewayProvider model IDalibaba/qwen3.7-maxRegionglobalInput / 1M$2.5Output / 1M$7.5Context991KUpdatedSep 8, 2026

Qwen3.7 Max Runtime Performance

Provider-specific output speed and catalog latency for Qwen3.7 Max. Runtime does not affect the capability score.

No runtime data

No provider-specific speed or latency record is linked to the default version yet.

Output Speed is generated output tokens received per second. Catalog latency is reported separately from observed provider TTFT.

Qwen3.7 Max Specifications

Technical details for the model's default version.

Version
Qwen3.7 Max
Released
May 19, 2026
Knowledge cutoff
Unknown
Parameters
N/A
Context window
1M
Max output
65.5K
Inputs
text
Outputs
text
Open weights
No
License
Proprietary

Qwen3.7 Max Versions

Available versions of this model. The score column identifies the version used in the overall ranking.

1 row
Columns

Show columns

Sort by
Version
Released
LLMBoard
Parameters
Context
Max output
Open weights
License
VersionQwen3.7 MaxReleasedMay 19, 2026LLMBoard76.63ParametersN/AContext1MMax output65.5KOpen weightsNoLicenseProprietary

Models similar to Qwen3.7 Max

Recommendations prioritize the same model type and family, then the closest LLMBoard score.

#36-1.69
AC

Qwen3.8 27B

Alibaba Cloud / Qwen Team

74.94 LLMBoard

Details
#22+4.72
AC

Qwen3.8 Flash

Alibaba Cloud / Qwen Team

81.35 LLMBoard

Details
#21+4.81
AC

Qwen3.8 Flash Next

Alibaba Cloud / Qwen Team

81.44 LLMBoard

Details
#42-5.13
AC

Qwen3.7 Plus

Alibaba Cloud / Qwen Team

71.50 LLMBoard

Details
#10+10.30
AC

Qwen3.8 Max

Alibaba Cloud / Qwen Team

86.93 LLMBoard

Details
#57-10.90
AC

Qwen3.6 Plus

Alibaba Cloud / Qwen Team

65.73 LLMBoard

Details

What is Qwen3.7 Max?

Key information about Qwen3.7 Max and its available data.

7 Max is Alibaba Cloud / Qwen Team's proprietary large language model, with a 1 million token context window and up to 65,536 output tokens. It is designed for coding agents, office automation, MCP, multi-agent orchestration, and long-horizon autonomous execution.

Data as of 2026-09-08.

FAQ

Common questions about Qwen3.7 Max.

When was Qwen3.7 Max released?

Qwen3.7 Max's default version was released on May 19, 2026.

How much does Qwen3.7 Max cost?

Qwen3.7 Max's official API price is $2.5 per million input tokens and $7.5 per million output tokens via Alibaba. The lowest tracked third-party offer starts at $0.825 input and $2.48 output via Merge Gateway.

Who created Qwen3.7 Max?

Qwen3.7 Max was created by Alibaba Cloud / Qwen Team.

What is the context window for Qwen3.7 Max?

The default version has a 1M token context window.

Is Qwen3.7 Max open weight?

No. The default version is not marked as having publicly available weights.

How many API providers offer Qwen3.7 Max?

33 provider offerings are linked to the default version.

Official prices use only the vendor's configured official Provider and positive standard USD PAYG rates. Third-party offers remain explicitly labeled.

Browse runtime rankings