llmboard.aiAI model intelligence
Home

Model Rankings

OverallOpen ModelsAgentCodingReasoningMathKnowledgeInstruction FollowingTextVision
Image GenerationImage Editing
Video GenerationImage to VideoVideo Editing
Text to SpeechSpeech to Text
Embeddings

Efficiency

Chat Token PricingImage PricingVideo PricingAudio Pricing
Chat Speed & LatencyProvider Reliability

Benchmarks

GPQAMMLU-ProAIME 2025SWE-Bench VerifiedMMLUHumanity's Last ExamLiveCodeBenchMATHHumanEvalMMMU-Pro
All Benchmarks

Tools

Model Directory

Scoring & Data

Scoring & Data
1224 models729 benchmarks

Leaderboard Center

Overall RankingCodingCore BenchmarksPrice & ValueRuntime Performance

Modalities

All ModelsImage GenerationImage EditingVideo GenerationImage-to-VideoVideo EditingText-to-SpeechSpeech-to-TextEmbeddings

Data & Methods

Scoring MethodAll Benchmarks
llmboard.aiCopyright 2026 llmboard.ai

Anthropic model product

Claude Opus 4.6

6 is an Anthropic large language model with a 1M token context window in beta and 128K output tokens.

Updated Sep 8, 2026. Default version: Claude Opus 4.6

LLMBoard Score78.2Claude Opus 4.6
Coverage60%26 benchmark families
Context window1MTokens
Official input price$5Anthropic API

On this page

  • Capability
  • Benchmarks
  • Arena
  • Pricing
  • Runtime
  • Specification
  • Versions
  • Similar models
  • About
  • FAQ

Claude Opus 4.6 Capability Profile

This profile uses the model's current scored version. Arena ratings and prices are shown separately.

Claude Opus 4.6 LLMBoard score breakdown

Claude Opus 4.6 Benchmark Results

Benchmark scores for Claude Opus 4.6.

30 of 44 rows
Columns

Show columns

Sort by
Benchmark
Score
Rank
Participants
Percentile
Evidence
Evaluated
BenchmarkGraphwalks parents >128kScore95.40%Rank01Participants7Percentile100.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkOpenRCAScore34.90%Rank01Participants1Percentile100.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkTau2 RetailScore91.90%Rank01Participants27Percentile100.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkTau2 TelecomScore99.30%Rank01Participants36Percentile100.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkVending-Bench 2Score8,017.59 usdRank01Participants4Percentile100.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkFigQAScore78.30%Rank02Participants3Percentile50.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkLM Arena DocumentScore1,509.94 ratingRank02Participants33Percentile96.88%EvidenceAEvaluatedJul 30, 2026
BenchmarkLM Arena SearchScore1,253.42 ratingRank02Participants28Percentile96.30%EvidenceAEvaluatedAug 24, 2026
BenchmarkLM Arena Text FactualityScore1,493.07 ratingRank02Participants121Percentile99.17%EvidenceAEvaluatedSep 2, 2026
BenchmarkLM Arena Text Style ControlScore1,504.84 ratingRank02Participants210Percentile99.52%EvidenceAEvaluatedSep 2, 2026
BenchmarkDeepSearchQAScore91.30%Rank03Participants10Percentile77.78%EvidenceCEvaluatedSep 8, 2026
BenchmarkFinance AgentScore60.70%Rank03Participants8Percentile71.43%EvidenceCEvaluatedSep 8, 2026
BenchmarkOSWorldScore72.70%Rank03Participants20Percentile89.47%EvidenceCEvaluatedSep 8, 2026
BenchmarkLM Arena Document Style ControlScore1,494.69 ratingRank03Participants33Percentile93.75%EvidenceAEvaluatedJul 30, 2026
BenchmarkLM Arena Search FactualityScore1,227.85 ratingRank03Participants28Percentile92.59%EvidenceAEvaluatedAug 24, 2026
BenchmarkLM Arena TextScore1,503.14 ratingRank03Participants210Percentile99.04%EvidenceAEvaluatedSep 2, 2026
BenchmarkLM Arena Search Style ControlScore1,223.31 ratingRank04Participants28Percentile88.89%EvidenceAEvaluatedAug 24, 2026
BenchmarkLM Arena VisionScore1,314.76 ratingRank04Participants107Percentile97.17%EvidenceAEvaluatedAug 27, 2026
BenchmarkLM Arena Vision Style ControlScore1,298.76 ratingRank04Participants107Percentile97.17%EvidenceAEvaluatedAug 27, 2026
BenchmarkAIME 2025Score99.79%Rank06Participants119Percentile95.76%EvidenceCEvaluatedSep 8, 2026
BenchmarkARC-AGI v2Score68.80%Rank06Participants18Percentile70.59%EvidenceCEvaluatedSep 8, 2026
BenchmarkGraphwalks BFS >128kScore61.50%Rank06Participants11Percentile50.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkLegal Agent BenchmarkScore4.20%Rank06Participants13Percentile58.33%EvidenceBEvaluatedSep 8, 2026
BenchmarkMMMLUScore91.10%Rank06Participants49Percentile89.58%EvidenceCEvaluatedSep 8, 2026
BenchmarkMRCR v2 (8-needle)Score76.00%Rank06Participants24Percentile78.26%EvidenceCEvaluatedSep 8, 2026
BenchmarkSWE-Bench VerifiedScore80.80%Rank07Participants113Percentile94.64%EvidenceCEvaluatedSep 8, 2026
BenchmarkFrontierSWEScore56.00%Rank09Participants16Percentile46.67%EvidenceBEvaluatedSep 8, 2026
BenchmarkSWE-bench MultilingualScore77.83%Rank09Participants43Percentile80.95%EvidenceCEvaluatedSep 8, 2026
BenchmarkCyberGymScore73.80%Rank10Participants15Percentile35.71%EvidenceCEvaluatedSep 8, 2026
BenchmarkLiveBenchScore76.33%Rank11Participants38Percentile72.97%EvidenceBEvaluatedSep 8, 2026

Claude Opus 4.6 Arena Results

Preference and agent-evaluation results for the default version.

30 of 100 rows
Columns

Show columns

Sort by
Arena
Category
Rank
Rating / score
Votes
Observations
Result date
ArenatextCategorycodingRank01Rating / score1,535.49Votes18,800ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryhard prompts englishRank01Rating / score1,531.28Votes20,869ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryinstruction followingRank01Rating / score1,523.12Votes22,974ObservationsN/AResult dateSep 2, 2026
Arenatext factualityCategorycodingRank01Rating / score1,553.89Votes16,764ObservationsN/AResult dateSep 2, 2026
Arenatext factualityCategorycreative writingRank01Rating / score1,492.28Votes10,367ObservationsN/AResult dateSep 2, 2026
Arenatext factualityCategoryexpertRank01Rating / score1,544.35Votes5,152ObservationsN/AResult dateSep 2, 2026
Arenatext factualityCategoryhard prompts englishRank01Rating / score1,534.98Votes18,856ObservationsN/AResult dateSep 2, 2026
Arenatext factualityCategoryindustry entertainment and sports and mediaRank01Rating / score1,482.55Votes13,795ObservationsN/AResult dateSep 2, 2026
Arenatext factualityCategoryinstruction followingRank01Rating / score1,508.58Votes20,978ObservationsN/AResult dateSep 2, 2026
Arenatext factualityCategorymulti turnRank01Rating / score1,507.67Votes10,006ObservationsN/AResult dateSep 2, 2026
Arenatext factualityCategoryspanishRank01Rating / score1,505.20Votes1,938ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategoryenglishRank01Rating / score1,513.92Votes32,084ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategoryhard promptsRank01Rating / score1,533.15Votes45,743ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategoryhard prompts englishRank01Rating / score1,536.25Votes20,869ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategoryindustry life and physical and social scienceRank01Rating / score1,528.01Votes11,862ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategoryinstruction followingRank01Rating / score1,513.51Votes22,974ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategorylonger queryRank01Rating / score1,524.26Votes29,792ObservationsN/AResult dateSep 2, 2026
ArenadocumentCategoryoverallRank02Rating / score1,509.94Votes37,271ObservationsN/AResult dateJul 30, 2026
ArenasearchCategoryoverallRank02Rating / score1,253.42Votes134,699ObservationsN/AResult dateAug 24, 2026
ArenatextCategorycreative writingRank02Rating / score1,504.70Votes12,678ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryenglishRank02Rating / score1,511.82Votes32,084ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryexpertRank02Rating / score1,545.17Votes6,409ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryfrenchRank02Rating / score1,509.43Votes2,387ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryhard promptsRank02Rating / score1,527.17Votes45,743ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryindustry business and management and financial operationsRank02Rating / score1,499.03Votes14,421ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryindustry entertainment and sports and mediaRank02Rating / score1,493.50Votes15,836ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryindustry mathematicalRank02Rating / score1,527.06Votes3,631ObservationsN/AResult dateSep 2, 2026
ArenatextCategorylonger queryRank02Rating / score1,520.29Votes29,792ObservationsN/AResult dateSep 2, 2026
ArenatextCategorymulti turnRank02Rating / score1,512.19Votes12,303ObservationsN/AResult dateSep 2, 2026
Arenatext factualityCategorychineseRank02Rating / score1,545.37Votes3,418ObservationsN/AResult dateSep 2, 2026

Claude Opus 4.6 Pricing

Official vendor API pricing appears first, followed by individual provider offers.

Official API
$5 input, $25 output per 1M
Official provider
Anthropic
Lowest third-party
From $4.3 input, $21 output per 1M via Poe
Tracked offerings
39
30 of 37 rows
Columns

Show columns

Sort by
Provider
Provider model ID
Region
Input / 1M
Output / 1M
Context
Updated
ProviderPoeProvider model IDanthropic/claude-opus-4.6RegionglobalInput / 1M$4.3Output / 1M$21Context983KUpdatedSep 8, 2026
ProviderOrcaRouterProvider model IDanthropic/claude-opus-4.6RegionglobalInput / 1M$5Output / 1M$25Context1MUpdatedSep 8, 2026
ProviderVertexProvider model IDclaude-opus-4-6@defaultRegionglobalInput / 1M$5Output / 1M$25Context1MUpdatedSep 8, 2026
ProviderImpossiblProvider model IDanthropic/claude-opus-4-6RegionglobalInput / 1M$5Output / 1M$25Context1MUpdatedSep 8, 2026
ProviderJiekou.AIProvider model IDclaude-opus-4-6RegionglobalInput / 1M$5Output / 1M$25Context1MUpdatedSep 8, 2026
ProviderOpenRouterProvider model IDanthropic/claude-opus-4.6RegionglobalInput / 1M$5Output / 1M$25Context1MUpdatedSep 8, 2026
ProviderVertex (Anthropic)Provider model IDclaude-opus-4-6@defaultRegionglobalInput / 1M$5Output / 1M$25Context1MUpdatedSep 8, 2026
ProviderAurikoProvider model IDclaude-opus-4-6RegionglobalInput / 1M$5Output / 1M$25Context1MUpdatedSep 8, 2026
ProviderAnthropicProvider model IDclaude-opus-4-6RegionglobalInput / 1M$5Output / 1M$25Context1MUpdatedSep 8, 2026
ProviderDatabricksProvider model IDdatabricks-claude-opus-4-6RegionglobalInput / 1M$5Output / 1M$25Context1MUpdatedSep 8, 2026
ProviderAmazon BedrockProvider model IDanthropic.claude-opus-4-6-v1RegionglobalInput / 1M$5Output / 1M$25Context1MUpdatedSep 8, 2026
ProviderOpperProvider model IDanthropic/claude-opus-4-6RegionglobalInput / 1M$5Output / 1M$25Context1MUpdatedSep 8, 2026
ProviderLLM GatewayProvider model IDanthropic/claude-opus-4-6RegionglobalInput / 1M$5Output / 1M$25Context1MUpdatedSep 8, 2026
ProviderNEAR AI CloudProvider model IDanthropic/claude-opus-4-6RegionglobalInput / 1M$5Output / 1M$25Context200KUpdatedSep 8, 2026
ProviderMerge GatewayProvider model IDanthropic/claude-opus-4-6RegionglobalInput / 1M$5Output / 1M$25Context1MUpdatedSep 8, 2026
ProviderCloudflare AI GatewayProvider model IDanthropic/claude-opus-4.6RegionglobalInput / 1M$5Output / 1M$25Context1MUpdatedSep 8, 2026
ProviderZenMuxProvider model IDanthropic/claude-opus-4.6RegionglobalInput / 1M$5Output / 1M$25Context1MUpdatedSep 8, 2026
ProviderPerplexity AgentProvider model IDanthropic/claude-opus-4-6RegionglobalInput / 1M$5Output / 1M$25Context200KUpdatedSep 8, 2026
ProviderGMI CloudProvider model IDanthropic/claude-opus-4.6RegionglobalInput / 1M$5Output / 1M$25Context409.6KUpdatedSep 8, 2026
ProviderFreeModelProvider model IDclaude-opus-4-6RegionglobalInput / 1M$5Output / 1M$25Context1MUpdatedSep 8, 2026
ProviderOfoxProvider model IDanthropic/claude-opus-4.6RegionglobalInput / 1M$5Output / 1M$25Context1MUpdatedSep 8, 2026
ProviderNeonProvider model IDclaude-opus-4-6RegionglobalInput / 1M$5Output / 1M$25Context1MUpdatedSep 8, 2026
ProviderAzure Cognitive ServicesProvider model IDclaude-opus-4-6RegionglobalInput / 1M$5Output / 1M$25Context1MUpdatedSep 8, 2026
Provider302.AIProvider model IDclaude-opus-4-6RegionglobalInput / 1M$5Output / 1M$25Context1MUpdatedSep 8, 2026
ProviderOpenCode ZenProvider model IDclaude-opus-4-6RegionglobalInput / 1M$5Output / 1M$25Context1MUpdatedSep 8, 2026
ProviderRequestyProvider model IDclaude-opus-4-6RegionglobalInput / 1M$5Output / 1M$25Context1MUpdatedSep 8, 2026
ProviderAzureProvider model IDclaude-opus-4-6RegionglobalInput / 1M$5Output / 1M$25Context1MUpdatedSep 8, 2026
ProviderFrogBotProvider model IDclaude-opus-4-6RegionglobalInput / 1M$5Output / 1M$25Context200KUpdatedSep 8, 2026
ProviderAIHubMixProvider model IDclaude-opus-4-6RegionglobalInput / 1M$5Output / 1M$25Context1MUpdatedSep 8, 2026
ProviderVercel AI GatewayProvider model IDanthropic/claude-opus-4.6RegionglobalInput / 1M$5Output / 1M$25Context1MUpdatedSep 8, 2026

Claude Opus 4.6 Runtime Performance

Provider-specific output speed and catalog latency for Claude Opus 4.6. Runtime does not affect the capability score.

1 row
Columns

Show columns

Sort by
Provider
Output Speed
Catalog Latency
Max Input
Max Output
Updated
ProviderAnthropicOutput Speed123.15 tok/sCatalog Latency2.15 sMax Input1MMax Output128KUpdatedSep 8, 2026

Output Speed is generated output tokens received per second. Catalog latency is reported separately from observed provider TTFT.

Claude Opus 4.6 Specifications

Technical details for the model's default version.

Version
Claude Opus 4.6
Released
Feb 5, 2026
Knowledge cutoff
Unknown
Parameters
N/A
Context window
1M
Max output
128K
Inputs
image, text
Outputs
text
Open weights
No
License
Proprietary

Claude Opus 4.6 Versions

Available versions of this model. The score column identifies the version used in the overall ranking.

1 row
Columns

Show columns

Sort by
Version
Released
LLMBoard
Parameters
Context
Max output
Open weights
License
VersionClaude Opus 4.6ReleasedFeb 5, 2026LLMBoard78.18ParametersN/AContext1MMax output128KOpen weightsNoLicenseProprietary

Models similar to Claude Opus 4.6

Recommendations prioritize the same model type and family, then the closest LLMBoard score.

#23+1.98
AN

Claude Sonnet 5

Anthropic

80.16 LLMBoard

Details
#19+4.08
AN

Claude Opus 4.7

Anthropic

82.26 LLMBoard

Details
#43-6.81
AN

Claude Sonnet 4.6

Anthropic

71.37 LLMBoard

Details
#13+7.21
AN

Claude Opus 4.8

Anthropic

85.39 LLMBoard

Details
#50-9.96
AN

Claude Opus 4.5

Anthropic

68.22 LLMBoard

Details
#5+13.12
AN

Claude Mythos

Anthropic

91.30 LLMBoard

Details

What is Claude Opus 4.6?

Key information about Claude Opus 4.6 and its available data.

6 is an LLM from Anthropic with adaptive thinking, configurable low, medium, high, and max effort controls, and context compaction for long-running tasks. It is priced at $5 per million input tokens and $25 per million output tokens.

Data as of 2026-09-08.

FAQ

Common questions about Claude Opus 4.6.

When was Claude Opus 4.6 released?

Claude Opus 4.6's default version was released on Feb 5, 2026.

How much does Claude Opus 4.6 cost?

Claude Opus 4.6's official API price is $5 per million input tokens and $25 per million output tokens via Anthropic. The lowest tracked third-party offer starts at $4.3 input and $21 output via Poe.

Who created Claude Opus 4.6?

Claude Opus 4.6 was created by Anthropic.

What is the context window for Claude Opus 4.6?

The default version has a 1M token context window.

Is Claude Opus 4.6 open weight?

No. The default version is not marked as having publicly available weights.

How many API providers offer Claude Opus 4.6?

Official prices use only the vendor's configured official Provider and positive standard USD PAYG rates. Third-party offers remain explicitly labeled.

39 provider offerings are linked to the default version.