llmboard.aiAI model intelligence
Home

Model Rankings

OverallOpen ModelsAgentCodingReasoningMathKnowledgeInstruction FollowingTextVision
Image GenerationImage Editing
Video GenerationImage to VideoVideo Editing
Text to SpeechSpeech to Text
Embeddings

Efficiency

Chat Token PricingImage PricingVideo PricingAudio Pricing
Chat Speed & LatencyProvider Reliability

Benchmarks

GPQAMMLU-ProAIME 2025SWE-Bench VerifiedMMLUHumanity's Last ExamLiveCodeBenchMATHHumanEvalMMMU-Pro
All Benchmarks

Tools

Model Directory

Scoring & Data

Scoring & Data
1224 models729 benchmarks

Leaderboard Center

Overall RankingCodingCore BenchmarksPrice & ValueRuntime Performance

Modalities

All ModelsImage GenerationImage EditingVideo GenerationImage-to-VideoVideo EditingText-to-SpeechSpeech-to-TextEmbeddings

Data & Methods

Scoring MethodAll Benchmarks
llmboard.aiCopyright 2026 llmboard.ai

Zhipu AI model product

GLM 5.1

1 is a Zhipu AI large language model for coding, reasoning, and agentic engineering tasks.

Updated Sep 8, 2026. Default version: GLM-5.1

LLMBoard Score67.6GLM-5.1
Coverage40%18 benchmark families
Context window200KTokens
Official input price$1.4Zhipu AI API

On this page

  • Capability
  • Benchmarks
  • Arena
  • Pricing
  • Runtime
  • Specification
  • Versions
  • Similar models
  • About
  • FAQ

GLM 5.1 Capability Profile

This profile uses the model's current scored version. Arena ratings and prices are shown separately.

GLM-5.1 LLMBoard score breakdown

GLM 5.1 Benchmark Results

Benchmark scores for GLM-5.1.

28 rows
Columns

Show columns

Sort by
Benchmark
Score
Rank
Participants
Percentile
Evidence
Evaluated
BenchmarkVending-Bench 2Score5,634.41 usdRank02Participants4Percentile66.67%EvidenceCEvaluatedSep 8, 2026
BenchmarkTAU3-BenchScore70.60%Rank03Participants8Percentile71.43%EvidenceCEvaluatedSep 8, 2026
BenchmarkAIME 2026Score95.30%Rank05Participants22Percentile80.95%EvidenceCEvaluatedSep 8, 2026
BenchmarkHMMT 2025Score94.00%Rank10Participants33Percentile71.88%EvidenceCEvaluatedSep 8, 2026
BenchmarkIMO-AnswerBenchScore83.80%Rank11Participants20Percentile47.37%EvidenceCEvaluatedSep 8, 2026
BenchmarkTerminal-Bench 2.0Score69.00%Rank11Participants51Percentile80.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkFinance Agent v2Score44.79%Rank12Participants26Percentile56.00%EvidenceBEvaluatedSep 8, 2026
BenchmarkFrontierSWEScore31.00%Rank12Participants16Percentile26.67%EvidenceBEvaluatedSep 8, 2026
BenchmarkHMMT Feb 26Score82.60%Rank12Participants12Percentile0.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkCyberGymScore68.70%Rank13Participants15Percentile14.29%EvidenceCEvaluatedSep 8, 2026
BenchmarkNL2RepoScore42.70%Rank15Participants22Percentile33.33%EvidenceCEvaluatedSep 8, 2026
BenchmarkHumanity's Last ExamScore52.30%Rank19Participants103Percentile82.35%EvidenceCEvaluatedSep 8, 2026
BenchmarkMCP AtlasScore71.80%Rank20Participants34Percentile42.42%EvidenceCEvaluatedSep 8, 2026
BenchmarkBrowseCompScore79.30%Rank22Participants63Percentile66.13%EvidenceCEvaluatedSep 8, 2026
BenchmarkSWE-Bench ProScore58.40%Rank23Participants55Percentile59.26%EvidenceCEvaluatedSep 8, 2026
BenchmarkLM Arena TextScore1,462.74 ratingRank25Participants210Percentile88.52%EvidenceAEvaluatedSep 2, 2026
BenchmarkLiveBenchScore70.18%Rank29Participants38Percentile24.32%EvidenceBEvaluatedSep 8, 2026
BenchmarkLM Arena Agent Task Outcome ExplicitScore0.17%Rank29Participants49Percentile41.67%EvidenceAEvaluatedSep 5, 2026
BenchmarkLM Arena Agent SteerabilityScore-1.27%Rank30Participants49Percentile39.58%EvidenceAEvaluatedSep 5, 2026
BenchmarkLM Arena Agent Praise ComplaintScore-2.28%Rank32Participants49Percentile35.42%EvidenceAEvaluatedSep 5, 2026
BenchmarkLM Arena Agent LeaderboardScore-2.22%Rank33Participants49Percentile33.33%EvidenceAEvaluatedSep 5, 2026
BenchmarkToolathlonScore40.70%Rank34Participants41Percentile17.50%EvidenceCEvaluatedSep 8, 2026
BenchmarkLM Arena Text FactualityScore1,457.10 ratingRank34Participants121Percentile72.50%EvidenceAEvaluatedSep 2, 2026
BenchmarkLM Arena Text Style ControlScore1,466.02 ratingRank35Participants210Percentile83.73%EvidenceAEvaluatedSep 2, 2026
BenchmarkLM Arena WebdevScore1,508.11 ratingRank36Participants97Percentile63.54%EvidenceAEvaluatedSep 5, 2026
BenchmarkLM Arena Agent Bash Recovery StepsScore-6.23%Rank39Participants49Percentile20.83%EvidenceAEvaluatedSep 5, 2026
BenchmarkLM Arena Agent Tool HallucinationScore-1.50%Rank46Participants49Percentile6.25%EvidenceAEvaluatedSep 5, 2026
BenchmarkGPQAScore86.20%Rank54Participants247Percentile78.46%EvidenceCEvaluatedSep 8, 2026

GLM 5.1 Arena Results

Preference and agent-evaluation results for the default version.

30 of 96 rows
Columns

Show columns

Sort by
Arena
Category
Rank
Rating / score
Votes
Observations
Result date
Arenatext factualityCategorychineseRank14Rating / score1,517.38Votes2,250ObservationsN/AResult dateSep 2, 2026
Arenatext factualityCategoryjapaneseRank14Rating / score1,428.22Votes337ObservationsN/AResult dateSep 2, 2026
Arenatext factualityCategorypolishRank16Rating / score1,440.97Votes501ObservationsN/AResult dateSep 2, 2026
Arenatext factualityCategoryindustry medicine and healthcareRank18Rating / score1,491.29Votes2,713ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryenglishRank19Rating / score1,473.77Votes20,337ObservationsN/AResult dateSep 2, 2026
ArenatextCategorygermanRank19Rating / score1,467.22Votes731ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryindustry medicine and healthcareRank19Rating / score1,472.07Votes3,314ObservationsN/AResult dateSep 2, 2026
ArenatextCategorymathRank20Rating / score1,474.38Votes2,244ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryspanishRank20Rating / score1,465.99Votes1,545ObservationsN/AResult dateSep 2, 2026
Arenatext factualityCategorycreative writingRank20Rating / score1,449.47Votes7,651ObservationsN/AResult dateSep 2, 2026
ArenatextCategorycreative writingRank21Rating / score1,452.01Votes8,453ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryindustry entertainment and sports and mediaRank21Rating / score1,442.38Votes10,409ObservationsN/AResult dateSep 2, 2026
ArenatextCategorychineseRank22Rating / score1,517.05Votes2,745ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryhard prompts englishRank22Rating / score1,479.41Votes13,702ObservationsN/AResult dateSep 2, 2026
ArenatextCategorykoreanRank22Rating / score1,424.79Votes907ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryindustry business and management and financial operationsRank23Rating / score1,451.41Votes9,086ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryindustry life and physical and social scienceRank23Rating / score1,477.62Votes7,487ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryindustry writing and literature and languageRank23Rating / score1,454.49Votes11,710ObservationsN/AResult dateSep 2, 2026
Arenatext factualityCategoryindustry life and physical and social scienceRank23Rating / score1,487.49Votes6,710ObservationsN/AResult dateSep 2, 2026
Arenatext factualityCategorykoreanRank23Rating / score1,405.55Votes498ObservationsN/AResult dateSep 2, 2026
Arenatext factualityCategoryspanishRank23Rating / score1,450.08Votes1,091ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategorymathRank23Rating / score1,477.38Votes2,244ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryjapaneseRank24Rating / score1,430.81Votes571ObservationsN/AResult dateSep 2, 2026
ArenatextCategorymulti turnRank24Rating / score1,471.51Votes7,434ObservationsN/AResult dateSep 2, 2026
Arenatext factualityCategoryindustry legal and governmentRank24Rating / score1,479.71Votes3,017ObservationsN/AResult dateSep 2, 2026
Arenatext factualityCategoryindustry writing and literature and languageRank24Rating / score1,454.46Votes11,092ObservationsN/AResult dateSep 2, 2026
Arenatext factualityCategorymulti turnRank24Rating / score1,475.78Votes6,533ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategorychineseRank24Rating / score1,517.21Votes2,745ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategoryenglishRank24Rating / score1,479.85Votes20,337ObservationsN/AResult dateSep 2, 2026
ArenatextCategorycodingRank25Rating / score1,486.69Votes12,715ObservationsN/AResult dateSep 2, 2026

GLM 5.1 Pricing

Official vendor API pricing appears first, followed by individual provider offers.

Official API
$1.4 input, $4.4 output per 1M
Official provider
Zhipu AI
Lowest third-party
From $0.45 input, $2.15 output per 1M via CrofAI
Tracked offerings
52
30 of 51 rows
Columns

Show columns

Sort by
Provider
Provider model ID
Region
Input / 1M
Output / 1M
Context
Updated
ProviderAlibaba Token PlanProvider model IDglm-5.1RegionglobalInput / 1MN/AOutput / 1MN/AContext202.8KUpdatedSep 8, 2026
ProviderKenariProvider model IDglm-5-1RegionglobalInput / 1MN/AOutput / 1MN/AContext200KUpdatedSep 8, 2026
ProviderZhipu AI Coding PlanProvider model IDglm-5.1RegionglobalInput / 1MN/AOutput / 1MN/AContext200KUpdatedSep 8, 2026
ProviderSCNet Token PlanProvider model IDGLM-5.1RegionglobalInput / 1MN/AOutput / 1MN/AContext200KUpdatedSep 8, 2026
ProviderAlibaba Token Plan (China)Provider model IDglm-5.1RegionglobalInput / 1MN/AOutput / 1MN/AContext202.8KUpdatedSep 8, 2026
ProviderCrofAIProvider model IDglm-5.1RegionglobalInput / 1M$0.45Output / 1M$2.15Context202.8KUpdatedSep 8, 2026
ProviderHPC-AIProvider model IDzai-org/glm-5.1RegionglobalInput / 1M$0.615Output / 1M$2.46Context202KUpdatedSep 8, 2026
ProviderNanoGPTProvider model IDz-ai/glm-5.1RegionglobalInput / 1M$0.75Output / 1M$2.6Context200KUpdatedSep 8, 2026
ProviderAlibaba (China)Provider model IDglm-5.1RegionglobalInput / 1M$0.825Output / 1M$3.3Context202.8KUpdatedSep 8, 2026
ProviderEmpirioLabs AIProvider model IDglm-5-1RegionglobalInput / 1M$0.825Output / 1M$3.3Context202KUpdatedSep 8, 2026
ProviderEBCloudProvider model IDGLM-5.1RegionglobalInput / 1M$0.8571Output / 1M$3.43Context200KUpdatedSep 8, 2026
Provider302.AIProvider model IDglm-5.1RegionglobalInput / 1M$0.86Output / 1M$3.5Context200KUpdatedSep 8, 2026
ProviderZenMuxProvider model IDz-ai/glm-5.1RegionglobalInput / 1M$0.8781Output / 1M$3.51Context200KUpdatedSep 8, 2026
ProviderDevPass (LLM Gateway)Provider model IDglm-5.1RegionglobalInput / 1M$0.931Output / 1M$2.93Context204.8KUpdatedSep 8, 2026
ProviderOpenRouterProvider model IDz-ai/glm-5.1RegionglobalInput / 1M$0.966Output / 1M$3.04Context204.8KUpdatedSep 8, 2026
ProviderGMI CloudProvider model IDzai-org/GLM-5.1-FP8RegionglobalInput / 1M$0.98Output / 1M$3.08Context202.8KUpdatedSep 8, 2026
ProviderPioneerProvider model IDzai-org/GLM-5.1RegionglobalInput / 1M$0.98Output / 1M$3.08Context202KUpdatedSep 8, 2026
ProviderHugging FaceProvider model IDzai-org/GLM-5.1RegionglobalInput / 1M$1Output / 1M$3.2Context202.8KUpdatedSep 8, 2026
ProviderCrossModelProvider model IDz-ai/glm-5.1RegionglobalInput / 1M$1Output / 1M$3.8Context200KUpdatedSep 8, 2026
ProviderWaferProvider model IDGLM-5.1RegionglobalInput / 1M$1Output / 1M$3.2Context202.8KUpdatedSep 8, 2026
ProviderDeep InfraProvider model IDzai-org/GLM-5.1RegionglobalInput / 1M$1.05Output / 1M$3.5Context202.8KUpdatedSep 8, 2026
ProviderFastRouterProvider model IDz-ai/glm-5.1RegionglobalInput / 1M$1.05Output / 1M$3.5Context200KUpdatedSep 8, 2026
ProviderCrusoeProvider model IDzai/GLM-5.1RegionglobalInput / 1M$1.2Output / 1M$4.4Context200KUpdatedSep 8, 2026
ProviderDInferenceProvider model IDglm-5.1RegionglobalInput / 1M$1.25Output / 1M$3.89Context200KUpdatedSep 8, 2026
ProviderDigitalOceanProvider model IDglm-5.1RegionglobalInput / 1M$1.3Output / 1M$4.3Context163.8KUpdatedSep 8, 2026
ProviderBasetenProvider model IDzai-org/GLM-5.1RegionglobalInput / 1M$1.3Output / 1M$4.3Context202.8KUpdatedSep 8, 2026
ProviderCharm HyperProvider model IDglm-5.1RegionglobalInput / 1M$1.33Output / 1M$4.31Context202.8KUpdatedSep 8, 2026
ProviderJalapeno CloudProvider model IDGLM-5.1RegionglobalInput / 1M$1.38Output / 1M$4.4Context202.8KUpdatedSep 8, 2026
ProviderNovitaAIProvider model IDzai-org/glm-5.1RegionglobalInput / 1M$1.38Output / 1M$4.4Context204.8KUpdatedSep 8, 2026
ProviderCortecsProvider model IDglm-5.1RegionglobalInput / 1M$1.38Output / 1M$4.35Context202.8KUpdatedSep 8, 2026

GLM 5.1 Runtime Performance

Provider-specific output speed and catalog latency for GLM-5.1. Runtime does not affect the capability score.

No runtime data

No provider-specific speed or latency record is linked to the default version yet.

Output Speed is generated output tokens received per second. Catalog latency is reported separately from observed provider TTFT.

GLM 5.1 Specifications

Technical details for the model's default version.

Version
GLM-5.1
Released
Apr 7, 2026
Knowledge cutoff
Unknown
Parameters
754B
Context window
200K
Max output
128K
Inputs
text
Outputs
text
Open weights
Yes
License
MIT

GLM 5.1 Versions

Available versions of this model. The score column identifies the version used in the overall ranking.

1 row
Columns

Show columns

Sort by
Version
Released
LLMBoard
Parameters
Context
Max output
Open weights
License
VersionGLM-5.1ReleasedApr 7, 2026LLMBoard67.64Parameters754BContext200KMax output128KOpen weightsYesLicenseMIT

Models similar to GLM 5.1

Recommendations prioritize the same model type and family, then the closest LLMBoard score.

#56-1.79
ZA

GLM 5

Zhipu AI

65.85 LLMBoard

Details
#76-9.09
ZA

GLM 4.7

Zhipu AI

58.55 LLMBoard

Details
#26+9.38
ZA

GLM 5.2

Zhipu AI

77.02 LLMBoard

Details
#88-13.38
ZA

GLM 5V Turbo

Zhipu AI

54.26 LLMBoard

Details
#94-14.85
ZA

GLM 4.6

Zhipu AI

52.79 LLMBoard

Details
#15+16.64
ZA

GLM 5.3 Flash

Zhipu AI

84.28 LLMBoard

Details

What is GLM 5.1?

Key information about GLM 5.1 and its available data.

1 is a 754B-parameter mixture-of-experts model with 40B active parameters, a 200K context length, and up to 128K output tokens. It supports thinking mode, function calling, structured output, context caching, and MCP integration.

Data as of 2026-09-08.

FAQ

Common questions about GLM 5.1.

When was GLM 5.1 released?

GLM 5.1's default version was released on Apr 7, 2026.

How much does GLM 5.1 cost?

GLM 5.1's official API price is $1.4 per million input tokens and $4.4 per million output tokens via Zhipu AI. The lowest tracked third-party offer starts at $0.45 input and $2.15 output via CrofAI.

Who created GLM 5.1?

GLM 5.1 was created by Zhipu AI.

What is the context window for GLM 5.1?

The default version has a 200K token context window.

Is GLM 5.1 open weight?

Yes. The default version is marked as open weight under MIT.

How many API providers offer GLM 5.1?

52 provider offerings are linked to the default version.

Official prices use only the vendor's configured official Provider and positive standard USD PAYG rates. Third-party offers remain explicitly labeled.

Browse runtime rankings