llmboard.aiAI model intelligence
Home

Model Rankings

OverallOpen ModelsAgentCodingReasoningMathKnowledgeInstruction FollowingTextVision
Image GenerationImage Editing
Video GenerationImage to VideoVideo Editing
Text to SpeechSpeech to Text
Embeddings

Efficiency

Chat Token PricingImage PricingVideo PricingAudio Pricing
Chat Speed & LatencyProvider Reliability

Benchmarks

GPQAMMLU-ProAIME 2025SWE-Bench VerifiedMMLUHumanity's Last ExamLiveCodeBenchMATHHumanEvalMMMU-Pro
All Benchmarks

Tools

Model Directory

Scoring & Data

Scoring & Data
1224 models729 benchmarks

Leaderboard Center

Overall RankingCodingCore BenchmarksPrice & ValueRuntime Performance

Modalities

All ModelsImage GenerationImage EditingVideo GenerationImage-to-VideoVideo EditingText-to-SpeechSpeech-to-TextEmbeddings

Data & Methods

Scoring MethodAll Benchmarks
llmboard.aiCopyright 2026 llmboard.ai

Anthropic model product

Claude Opus 5

Claude Opus 5 is Anthropic's general-access large language model for agentic coding, knowledge work, scientific research, computer use, and problem-solving.

Updated Sep 8, 2026. Default version: Claude Opus 5

LLMBoard Score93.2Claude Opus 5
Coverage40%12 benchmark families
Context window1MTokens
Official input price$5Anthropic API

On this page

  • Capability
  • Benchmarks
  • Arena
  • Pricing
  • Runtime
  • Specification
  • Versions
  • Similar models
  • About
  • FAQ

Claude Opus 5 Capability Profile

This profile uses the model's current scored version. Arena ratings and prices are shown separately.

Claude Opus 5 LLMBoard score breakdown

Claude Opus 5 Benchmark Results

Benchmark scores for Claude Opus 5.

27 rows
Columns

Show columns

Sort by
Benchmark
Score
Rank
Participants
Percentile
Evidence
Evaluated
BenchmarkBioMysteryBenchScore90.10%Rank01Participants5Percentile100.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkFrontier-Bench v0.1Score43.30%Rank01Participants1Percentile100.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkLM Arena Agent Bash Recovery StepsScore15.69%Rank01Participants49Percentile100.00%EvidenceAEvaluatedSep 5, 2026
BenchmarkLM Arena Agent SteerabilityScore16.09%Rank01Participants49Percentile100.00%EvidenceAEvaluatedSep 5, 2026
BenchmarkLM Arena DocumentScore1,520.11 ratingRank01Participants33Percentile100.00%EvidenceAEvaluatedJul 30, 2026
BenchmarkARC-AGI-3Score30.20%Rank02Participants5Percentile75.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkFrontierCodeScore53.40%Rank02Participants4Percentile66.67%EvidenceCEvaluatedSep 8, 2026
BenchmarkFrontierCode 1.1Score53.40%Rank02Participants17Percentile93.75%EvidenceBEvaluatedSep 8, 2026
BenchmarkHumanity's Last ExamScore64.70%Rank02Participants103Percentile99.02%EvidenceCEvaluatedSep 8, 2026
BenchmarkLegal Agent BenchmarkScore11.70%Rank02Participants13Percentile91.67%EvidenceCEvaluatedSep 8, 2026
BenchmarkLM Arena Agent LeaderboardScore12.68%Rank02Participants49Percentile97.92%EvidenceAEvaluatedSep 5, 2026
BenchmarkLM Arena TextScore1,505.02 ratingRank02Participants210Percentile99.52%EvidenceAEvaluatedSep 2, 2026
BenchmarkLM Arena VisionScore1,321.96 ratingRank02Participants107Percentile99.06%EvidenceAEvaluatedAug 27, 2026
BenchmarkBrowseCompScore90.80%Rank03Participants63Percentile96.77%EvidenceCEvaluatedSep 8, 2026
BenchmarkOSWorld 2.0Score70.60%Rank03Participants11Percentile80.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkTerminal-Bench 4.0Score51.80%Rank03Participants13Percentile83.33%EvidenceBEvaluatedSep 8, 2026
BenchmarkLM Arena Agent Task Outcome ExplicitScore14.89%Rank03Participants49Percentile95.83%EvidenceAEvaluatedSep 5, 2026
BenchmarkLM Arena WebdevScore1,687.61 ratingRank03Participants97Percentile97.92%EvidenceAEvaluatedSep 5, 2026
BenchmarkHealthBench ProfessionalScore59.80%Rank04Participants10Percentile66.67%EvidenceCEvaluatedSep 8, 2026
BenchmarkLM Arena Agent Praise ComplaintScore18.27%Rank04Participants49Percentile93.75%EvidenceAEvaluatedSep 5, 2026
BenchmarkLM Arena Document Style ControlScore1,492.66 ratingRank04Participants33Percentile90.63%EvidenceAEvaluatedJul 30, 2026
BenchmarkLM Arena Text FactualityScore1,488.67 ratingRank04Participants121Percentile97.50%EvidenceAEvaluatedSep 2, 2026
BenchmarkLM Arena Text Style ControlScore1,493.33 ratingRank07Participants210Percentile97.13%EvidenceAEvaluatedSep 2, 2026
BenchmarkLM Arena Vision Style ControlScore1,290.35 ratingRank07Participants107Percentile94.34%EvidenceAEvaluatedAug 27, 2026
BenchmarkDeepSWE 1.1Score68.80%Rank08Participants31Percentile76.67%EvidenceCEvaluatedSep 8, 2026
BenchmarkAutomationBenchScore26.00%Rank11Participants18Percentile41.18%EvidenceCEvaluatedSep 8, 2026
BenchmarkLM Arena Agent Tool HallucinationScore0.76%Rank20Participants49Percentile60.42%EvidenceAEvaluatedSep 5, 2026

Claude Opus 5 Arena Results

Preference and agent-evaluation results for the default version.

30 of 100 rows
Columns

Show columns

Sort by
Arena
Category
Rank
Rating / score
Votes
Observations
Result date
Arenaagent bash recovery stepsCategoryoverallRank01Rating / score0.16VotesN/AObservations46.1KResult dateSep 5, 2026
Arenaagent steerabilityCategoryoverallRank01Rating / score0.16VotesN/AObservations21.3KResult dateSep 5, 2026
ArenadocumentCategoryoverallRank01Rating / score1,520.11Votes1,663ObservationsN/AResult dateJul 30, 2026
ArenatextCategorychineseRank01Rating / score1,594.00Votes1,228ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryexpertRank01Rating / score1,555.05Votes3,872ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryfrenchRank01Rating / score1,522.00Votes1,163ObservationsN/AResult dateSep 2, 2026
ArenatextCategorygermanRank01Rating / score1,517.42Votes295ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryindustry business and management and financial operationsRank01Rating / score1,504.37Votes3,227ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryindustry legal and governmentRank01Rating / score1,534.32Votes1,471ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryindustry mathematicalRank01Rating / score1,546.47Votes1,845ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryjapaneseRank01Rating / score1,513.02Votes292ObservationsN/AResult dateSep 2, 2026
ArenatextCategorykoreanRank01Rating / score1,524.84Votes363ObservationsN/AResult dateSep 2, 2026
ArenatextCategorymathRank01Rating / score1,542.99Votes740ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryspanishRank01Rating / score1,524.93Votes440ObservationsN/AResult dateSep 2, 2026
Arenatext factualityCategorychineseRank01Rating / score1,567.05Votes968ObservationsN/AResult dateSep 2, 2026
Arenatext factualityCategoryindustry business and management and financial operationsRank01Rating / score1,505.24Votes2,982ObservationsN/AResult dateSep 2, 2026
Arenatext factualityCategoryindustry legal and governmentRank01Rating / score1,538.11Votes1,177ObservationsN/AResult dateSep 2, 2026
Arenatext factualityCategoryindustry life and physical and social scienceRank01Rating / score1,529.01Votes2,615ObservationsN/AResult dateSep 2, 2026
Arenatext factualityCategoryindustry mathematicalRank01Rating / score1,523.24Votes1,456ObservationsN/AResult dateSep 2, 2026
Arenatext factualityCategoryindustry medicine and healthcareRank01Rating / score1,528.33Votes1,058ObservationsN/AResult dateSep 2, 2026
Arenatext factualityCategorykoreanRank01Rating / score1,500.08Votes282ObservationsN/AResult dateSep 2, 2026
Arenatext factualityCategorymathRank01Rating / score1,535.65Votes557ObservationsN/AResult dateSep 2, 2026
Arenatext factualityCategoryrussianRank01Rating / score1,514.62Votes1,534ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategorychineseRank01Rating / score1,569.38Votes1,228ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategoryindustry mathematicalRank01Rating / score1,537.17Votes1,845ObservationsN/AResult dateSep 2, 2026
ArenavisionCategorychineseRank01Rating / score1,405.90Votes409ObservationsN/AResult dateAug 27, 2026
Arenavision style controlCategorychineseRank01Rating / score1,370.88Votes409ObservationsN/AResult dateAug 27, 2026
ArenawebdevCategoryimage to webdevRank01Rating / score1,664.39Votes1,765ObservationsN/AResult dateAug 25, 2026
ArenaagentCategoryoverallRank02Rating / score0.13VotesN/AObservations2.7MResult dateSep 5, 2026
ArenatextCategorycodingRank02Rating / score1,531.77Votes9,354ObservationsN/AResult dateSep 2, 2026

Claude Opus 5 Pricing

Official vendor API pricing appears first, followed by individual provider offers.

Official API
$5 input, $25 output per 1M
Official provider
Anthropic
Lowest third-party
From $5 input, $25 output per 1M via OrcaRouter
Tracked offerings
31
29 rows
Columns

Show columns

Sort by
Provider
Provider model ID
Region
Input / 1M
Output / 1M
Context
Updated
ProviderKenariProvider model IDclaude-opus-5RegionglobalInput / 1MN/AOutput / 1MN/AContext1MUpdatedSep 8, 2026
ProviderOrcaRouterProvider model IDanthropic/claude-opus-5RegionglobalInput / 1M$5Output / 1M$25Context1MUpdatedSep 8, 2026
ProviderNanoGPTProvider model IDanthropic/claude-opus-5RegionglobalInput / 1M$5Output / 1M$25Context1MUpdatedSep 8, 2026
ProviderVertexProvider model IDclaude-opus-5@defaultRegionglobalInput / 1M$5Output / 1M$25Context1MUpdatedSep 8, 2026
ProviderOpenRouterProvider model IDanthropic/claude-opus-5RegionglobalInput / 1M$5Output / 1M$25Context1MUpdatedSep 8, 2026
ProviderVertex (Anthropic)Provider model IDclaude-opus-5@defaultRegionglobalInput / 1M$5Output / 1M$25Context1MUpdatedSep 8, 2026
ProviderCrossModelProvider model IDanthropic/claude-opus-5RegionglobalInput / 1M$5Output / 1M$25Context1MUpdatedSep 8, 2026
ProviderAnthropicProvider model IDclaude-opus-5RegionglobalInput / 1M$5Output / 1M$25Context1MUpdatedSep 8, 2026
ProviderAmazon BedrockProvider model IDanthropic.claude-opus-5RegionglobalInput / 1M$5Output / 1M$25Context1MUpdatedSep 8, 2026
ProviderOpperProvider model IDanthropic/claude-opus-5RegionglobalInput / 1M$5Output / 1M$25Context1MUpdatedSep 8, 2026
ProviderLLM GatewayProvider model IDanthropic/claude-opus-5RegionglobalInput / 1M$5Output / 1M$25Context1MUpdatedSep 8, 2026
ProviderMerge GatewayProvider model IDanthropic/claude-opus-5RegionglobalInput / 1M$5Output / 1M$25Context1MUpdatedSep 8, 2026
ProviderCloudflare AI GatewayProvider model IDanthropic/claude-opus-5RegionglobalInput / 1M$5Output / 1M$25Context1MUpdatedSep 8, 2026
ProviderOfoxProvider model IDanthropic/claude-opus-5RegionglobalInput / 1M$5Output / 1M$25Context1MUpdatedSep 8, 2026
ProviderNeonProvider model IDclaude-opus-5RegionglobalInput / 1M$5Output / 1M$25Context1MUpdatedSep 8, 2026
ProviderAzure Cognitive ServicesProvider model IDclaude-opus-5RegionglobalInput / 1M$5Output / 1M$25Context1MUpdatedSep 8, 2026
ProviderOpenCode ZenProvider model IDclaude-opus-5RegionglobalInput / 1M$5Output / 1M$25Context1MUpdatedSep 8, 2026
ProviderRequestyProvider model IDclaude-opus-5RegionglobalInput / 1M$5Output / 1M$25Context1MUpdatedSep 8, 2026
ProviderAzureProvider model IDclaude-opus-5RegionglobalInput / 1M$5Output / 1M$25Context1MUpdatedSep 8, 2026
ProviderAIHubMixProvider model IDclaude-opus-5RegionglobalInput / 1M$5Output / 1M$25Context1MUpdatedSep 8, 2026
ProviderVercel AI GatewayProvider model IDanthropic/claude-opus-5RegionglobalInput / 1M$5Output / 1M$25Context1MUpdatedSep 8, 2026
ProviderEden AIProvider model IDanthropic/claude-opus-5RegionglobalInput / 1M$5Output / 1M$25Context1MUpdatedSep 8, 2026
ProviderDevPass (LLM Gateway)Provider model IDclaude-opus-5RegionglobalInput / 1M$5Output / 1M$25Context1MUpdatedSep 8, 2026
ProviderGitHub CopilotProvider model IDclaude-opus-5RegionglobalInput / 1M$5Output / 1M$25Context1MUpdatedSep 8, 2026
ProviderKilo GatewayProvider model IDanthropic/claude-opus-5RegionglobalInput / 1M$5Output / 1M$25Context1MUpdatedSep 8, 2026
ProviderPioneerProvider model IDclaude-opus-5RegionglobalInput / 1M$5Output / 1M$25Context1MUpdatedSep 8, 2026
ProviderAbacusProvider model IDclaude-opus-5RegionglobalInput / 1M$5Output / 1M$25Context1MUpdatedSep 8, 2026
ProviderCortecsProvider model IDclaude-opus-5RegionglobalInput / 1M$5.5Output / 1M$27.5Context1MUpdatedSep 8, 2026
ProviderVenice AIProvider model IDclaude-opus-5RegionglobalInput / 1M$6Output / 1M$30Context1MUpdatedSep 8, 2026

Claude Opus 5 Runtime Performance

Provider-specific output speed and catalog latency for Claude Opus 5. Runtime does not affect the capability score.

1 row
Columns

Show columns

Sort by
Provider
Output Speed
Catalog Latency
Max Input
Max Output
Updated
ProviderAnthropicOutput Speed62.88 tok/sCatalog Latency10.36 sMax Input1MMax Output128KUpdatedSep 8, 2026

Output Speed is generated output tokens received per second. Catalog latency is reported separately from observed provider TTFT.

Claude Opus 5 Specifications

Technical details for the model's default version.

Version
Claude Opus 5
Released
Jul 24, 2026
Knowledge cutoff
Unknown
Parameters
N/A
Context window
1M
Max output
128K
Inputs
image, text
Outputs
text
Open weights
No
License
Proprietary

Claude Opus 5 Versions

Available versions of this model. The score column identifies the version used in the overall ranking.

1 row
Columns

Show columns

Sort by
Version
Released
LLMBoard
Parameters
Context
Max output
Open weights
License
VersionClaude Opus 5ReleasedJul 24, 2026LLMBoard93.24ParametersN/AContext1MMax output128KOpen weightsNoLicenseProprietary

Models similar to Claude Opus 5

Recommendations prioritize the same model type and family, then the closest LLMBoard score.

#4-1.61
AN

Claude Fable 5

Anthropic

91.63 LLMBoard

Details
#5-1.94
AN

Claude Mythos

Anthropic

91.30 LLMBoard

Details
#2+4.71
AN

Claude Fable 5.1

Anthropic

97.95 LLMBoard

Details
#13-7.85
AN

Claude Opus 4.8

Anthropic

85.39 LLMBoard

Details
#19-10.98
AN

Claude Opus 4.7

Anthropic

82.26 LLMBoard

Details
#23-13.08
AN

Claude Sonnet 5

Anthropic

80.16 LLMBoard

Details

What is Claude Opus 5?

Key information about Claude Opus 5 and its available data.

1, GDPval-AA, and life-sciences evaluations. 8.

Data as of 2026-09-08.

FAQ

Common questions about Claude Opus 5.

When was Claude Opus 5 released?

Claude Opus 5's default version was released on Jul 24, 2026.

How much does Claude Opus 5 cost?

Claude Opus 5's official API price is $5 per million input tokens and $25 per million output tokens via Anthropic. The lowest tracked third-party offer starts at $5 input and $25 output via OrcaRouter.

Who created Claude Opus 5?

Claude Opus 5 was created by Anthropic.

What is the context window for Claude Opus 5?

The default version has a 1M token context window.

Is Claude Opus 5 open weight?

No. The default version is not marked as having publicly available weights.

How many API providers offer Claude Opus 5?

31 provider offerings are linked to the default version.

Official prices use only the vendor's configured official Provider and positive standard USD PAYG rates. Third-party offers remain explicitly labeled.