llmboard.aiAI model intelligence
Home

Model Rankings

OverallOpen ModelsAgentCodingReasoningMathKnowledgeInstruction FollowingTextVision
Image GenerationImage Editing
Video GenerationImage to VideoVideo Editing
Text to SpeechSpeech to Text
Embeddings

Efficiency

Chat Token PricingImage PricingVideo PricingAudio Pricing
Chat Speed & LatencyProvider Reliability

Benchmarks

GPQAMMLU-ProAIME 2025SWE-Bench VerifiedMMLUHumanity's Last ExamLiveCodeBenchMATHHumanEvalMMMU-Pro
All Benchmarks

Tools

Model Directory

Scoring & Data

Scoring & Data
1224 models729 benchmarks

Leaderboard Center

Overall RankingCodingCore BenchmarksPrice & ValueRuntime Performance

Modalities

All ModelsImage GenerationImage EditingVideo GenerationImage-to-VideoVideo EditingText-to-SpeechSpeech-to-TextEmbeddings

Data & Methods

Scoring MethodAll Benchmarks
llmboard.aiCopyright 2026 llmboard.ai

Anthropic model product

Claude Haiku 3

Claude Haiku 3 is an Anthropic LLM designed for fast responses to simple queries and requests.

Updated Sep 8, 2026. Default version: Claude 3 Haiku

LLMBoard Score0.0Claude 3 Haiku
Coverage40%10 benchmark families
Context window200KTokens
Official input priceN/AOfficial price unavailable

On this page

  • Capability
  • Benchmarks
  • Arena
  • Pricing
  • Runtime
  • Specification
  • Versions
  • Similar models
  • About
  • FAQ

Claude Haiku 3 Capability Profile

This profile uses the model's current scored version. Arena ratings and prices are shown separately.

Claude 3 Haiku LLMBoard score breakdown

Claude Haiku 3 Benchmark Results

Benchmark scores for Claude 3 Haiku.

14 rows
Columns

Show columns

Sort by
Benchmark
Score
Rank
Participants
Percentile
Evidence
Evaluated
BenchmarkBIG-Bench HardScore73.70%Rank10Participants21Percentile55.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkHellaSwagScore85.90%Rank11Participants27Percentile61.54%EvidenceCEvaluatedSep 8, 2026
BenchmarkARC-CScore89.20%Rank12Participants34Percentile66.67%EvidenceCEvaluatedSep 8, 2026
BenchmarkDROPScore78.40%Rank18Participants30Percentile41.38%EvidenceCEvaluatedSep 8, 2026
BenchmarkMGSMScore75.10%Rank20Participants31Percentile36.67%EvidenceCEvaluatedSep 8, 2026
BenchmarkGSM8kScore88.90%Rank29Participants48Percentile40.43%EvidenceCEvaluatedSep 8, 2026
BenchmarkHumanEvalScore75.90%Rank48Participants66Percentile27.69%EvidenceCEvaluatedSep 8, 2026
BenchmarkMATHScore38.90%Rank67Participants71Percentile5.71%EvidenceCEvaluatedSep 8, 2026
BenchmarkMMLUScore75.20%Rank75Participants101Percentile26.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkLM Arena Vision Style ControlScore1,000.63 ratingRank104Participants107Percentile2.83%EvidenceAEvaluatedAug 27, 2026
BenchmarkLM Arena VisionScore950.15 ratingRank106Participants107Percentile0.94%EvidenceAEvaluatedAug 27, 2026
BenchmarkLM Arena Text Style ControlScore1,261.08 ratingRank196Participants210Percentile6.70%EvidenceAEvaluatedSep 2, 2026
BenchmarkLM Arena TextScore1,194.46 ratingRank200Participants210Percentile4.78%EvidenceAEvaluatedSep 2, 2026
BenchmarkGPQAScore33.30%Rank230Participants247Percentile6.91%EvidenceCEvaluatedSep 8, 2026

Claude Haiku 3 Arena Results

Preference and agent-evaluation results for the default version.

30 of 64 rows
Columns

Show columns

Sort by
Arena
Category
Rank
Rating / score
Votes
Observations
Result date
Arenavision style controlCategorychineseRank91Rating / score974.52Votes1,019ObservationsN/AResult dateAug 27, 2026
ArenavisionCategorychineseRank92Rating / score914.78Votes1,019ObservationsN/AResult dateAug 27, 2026
Arenavision style controlCategoryoverallRank104Rating / score1,000.63Votes13,380ObservationsN/AResult dateAug 27, 2026
ArenavisionCategoryenglishRank106Rating / score948.59Votes8,449ObservationsN/AResult dateAug 27, 2026
ArenavisionCategoryoverallRank106Rating / score950.15Votes13,380ObservationsN/AResult dateAug 27, 2026
Arenavision style controlCategoryenglishRank106Rating / score990.84Votes8,449ObservationsN/AResult dateAug 27, 2026
Arenatext style controlCategorypolishRank143Rating / score1,298.36Votes264ObservationsN/AResult dateSep 2, 2026
ArenatextCategorypolishRank151Rating / score1,233.35Votes264ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategorykoreanRank162Rating / score1,208.91Votes2,364ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategoryjapaneseRank164Rating / score1,171.57Votes2,063ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategoryfrenchRank165Rating / score1,276.73Votes1,772ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryjapaneseRank167Rating / score1,102.41Votes2,063ObservationsN/AResult dateSep 2, 2026
ArenatextCategorykoreanRank169Rating / score1,108.21Votes2,364ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategorygermanRank170Rating / score1,246.18Votes3,508ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryfrenchRank171Rating / score1,194.39Votes1,772ObservationsN/AResult dateSep 2, 2026
ArenatextCategorygermanRank175Rating / score1,173.40Votes3,508ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategoryspanishRank175Rating / score1,238.59Votes1,704ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryspanishRank176Rating / score1,164.75Votes1,704ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategorymathRank188Rating / score1,231.44Votes14,983ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategoryindustry mathematicalRank189Rating / score1,238.69Votes13,491ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategoryrussianRank189Rating / score1,266.68Votes11,956ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategoryindustry medicine and healthcareRank191Rating / score1,279.36Votes6,385ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryindustry mathematicalRank192Rating / score1,187.33Votes13,491ObservationsN/AResult dateSep 2, 2026
ArenatextCategorymathRank192Rating / score1,188.05Votes14,983ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategorycodingRank193Rating / score1,300.90Votes20,898ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategoryindustry writing and literature and languageRank193Rating / score1,245.90Votes29,876ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategorymulti turnRank193Rating / score1,244.57Votes19,344ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategorynon englishRank193Rating / score1,245.85Votes53,164ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategorychineseRank194Rating / score1,248.79Votes16,707ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategoryexpertRank194Rating / score1,252.77Votes6,336ObservationsN/AResult dateSep 2, 2026

Claude Haiku 3 Pricing

Official vendor API pricing appears first, followed by individual provider offers.

Official API
N/A
Official provider
N/A
Lowest third-party
From $0.25 input, $1.25 output per 1M via OpenRouter
Tracked offerings
2
2 rows
Columns

Show columns

Sort by
Provider
Provider model ID
Region
Input / 1M
Output / 1M
Context
Updated
ProviderOpenRouterProvider model IDanthropic/claude-3-haikuRegionglobalInput / 1M$0.25Output / 1M$1.25Context200KUpdatedSep 8, 2026
ProviderHeliconeProvider model IDclaude-3-haiku-20240307RegionglobalInput / 1M$0.25Output / 1M$1.25Context200KUpdatedSep 8, 2026

Official prices use only the vendor's configured official Provider and positive standard USD PAYG rates. Third-party offers remain explicitly labeled.

Claude Haiku 3 Runtime Performance

Provider-specific output speed and catalog latency for Claude 3 Haiku. Runtime does not affect the capability score.

2 rows
Columns

Show columns

Sort by
Provider
Output Speed
Catalog Latency
Max Input
Max Output
Updated
ProviderAnthropicOutput Speed100.00 tok/sCatalog Latency0.50 sMax Input200KMax Output200KUpdatedSep 8, 2026
ProviderGoogleOutput Speed42.00 tok/sCatalog Latency0.40 sMax Input200KMax Output200KUpdatedSep 8, 2026

Output Speed is generated output tokens received per second. Catalog latency is reported separately from observed provider TTFT.

Claude Haiku 3 Specifications

Technical details for the model's default version.

Version
Claude 3 Haiku
Released
Mar 13, 2024
Knowledge cutoff
Unknown
Parameters
N/A
Context window
200K
Max output
200K
Inputs
image, text
Outputs
text
Open weights
No
License
Proprietary

Claude Haiku 3 Versions

Available versions of this model. The score column identifies the version used in the overall ranking.

1 row
Columns

Show columns

Sort by
Version
Released
LLMBoard
Parameters
Context
Max output
Open weights
License
VersionClaude 3 HaikuReleasedMar 13, 2024LLMBoard0.00ParametersN/AContext200KMax output200KOpen weightsNoLicenseProprietary

Models similar to Claude Haiku 3

Recommendations prioritize the same model type and family, then the closest LLMBoard score.

#265+1.96
AN

Claude Sonnet 3

Anthropic

1.96 LLMBoard

Details
#250+7.16
AN

Claude Haiku 3.5

Anthropic

7.16 LLMBoard

Details
#226+14.88
AN

Claude Opus 3

Anthropic

14.88 LLMBoard

Details
#180+30.76
AN

Claude Sonnet 3.5

Anthropic

30.76 LLMBoard

Details
#153+37.50
AN

Claude Sonnet 3.7

Anthropic

37.50 LLMBoard

Details
#150+38.71
AN

Claude Sonnet 4

Anthropic

38.71 LLMBoard

Details

What is Claude Haiku 3?

Key information about Claude Haiku 3 and its available data.

Claude Haiku 3 is the fastest and most compact model in Anthropic’s Claude 3 family. It is designed for near-instant responsiveness and answering simple queries and requests.

Data as of 2026-09-08.

FAQ

Common questions about Claude Haiku 3.

When was Claude Haiku 3 released?

Claude Haiku 3's default version was released on Mar 13, 2024.

How much does Claude Haiku 3 cost?

No official standard PAYG price is currently available for Claude Haiku 3. The lowest tracked third-party offer starts at $0.25 input and $1.25 output via OpenRouter.

Who created Claude Haiku 3?

Claude Haiku 3 was created by Anthropic.

What is the context window for Claude Haiku 3?

The default version has a 200K token context window.

Is Claude Haiku 3 open weight?

No. The default version is not marked as having publicly available weights.

How many API providers offer Claude Haiku 3?

2 provider offerings are linked to the default version.