llmboard.aiAI model intelligence
Home

Model Rankings

OverallOpen ModelsAgentCodingReasoningMathKnowledgeInstruction FollowingTextVision
Image GenerationImage Editing
Video GenerationImage to VideoVideo Editing
Text to SpeechSpeech to Text
Embeddings

Efficiency

Chat Token PricingImage PricingVideo PricingAudio Pricing
Chat Speed & LatencyProvider Reliability

Benchmarks

GPQAMMLU-ProAIME 2025SWE-Bench VerifiedMMLUHumanity's Last ExamLiveCodeBenchMATHHumanEvalMMMU-Pro
All Benchmarks

Tools

Model Directory

Scoring & Data

Scoring & Data
1224 models729 benchmarks

Leaderboard Center

Overall RankingCodingCore BenchmarksPrice & ValueRuntime Performance

Modalities

All ModelsImage GenerationImage EditingVideo GenerationImage-to-VideoVideo EditingText-to-SpeechSpeech-to-TextEmbeddings

Data & Methods

Scoring MethodAll Benchmarks
llmboard.aiCopyright 2026 llmboard.ai

OpenAI model product

GPT-5.4-mini

4-mini is an OpenAI mini model for coding, computer use, and subagents.

Updated Sep 8, 2026. Default version: GPT-5.4 mini

LLMBoard Score58.9GPT-5.4 mini
Coverage0%15 benchmark families
Context window400KTokens
Official input price$0.75OpenAI API

On this page

  • Capability
  • Benchmarks
  • Arena
  • Pricing
  • Runtime
  • Specification
  • Versions
  • Similar models
  • About
  • FAQ

GPT-5.4-mini Capability Profile

This profile uses the model's current scored version. Arena ratings and prices are shown separately.

GPT-5.4 mini LLMBoard score breakdown

GPT-5.4-mini Benchmark Results

Benchmark scores for GPT-5.4 mini.

21 rows
Columns

Show columns

Sort by
Benchmark
Score
Rank
Participants
Percentile
Evidence
Evaluated
BenchmarkGraphwalks BFS <128kScore76.30%Rank04Participants11Percentile70.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkGraphwalks parents <128kScore71.50%Rank05Participants11Percentile60.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkFinance Agent v2Score45.36%Rank10Participants26Percentile64.00%EvidenceBEvaluatedSep 8, 2026
BenchmarkOmniDocBench 1.5Score87.37%Rank12Participants18Percentile35.29%EvidenceCEvaluatedSep 8, 2026
BenchmarkLegal Agent BenchmarkScore0.00%Rank13Participants13Percentile0.00%EvidenceBEvaluatedSep 8, 2026
BenchmarkTau2 TelecomScore93.40%Rank13Participants36Percentile65.71%EvidenceCEvaluatedSep 8, 2026
BenchmarkMRCR v2 (8-needle)Score33.60%Rank14Participants24Percentile43.48%EvidenceCEvaluatedSep 8, 2026
BenchmarkOSWorld-VerifiedScore72.10%Rank16Participants24Percentile34.78%EvidenceCEvaluatedSep 8, 2026
BenchmarkTerminal-Bench 2.0Score60.00%Rank20Participants51Percentile62.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkMMMU-ProScore76.60%Rank27Participants69Percentile61.76%EvidenceCEvaluatedSep 8, 2026
BenchmarkMCP AtlasScore57.70%Rank31Participants34Percentile9.09%EvidenceCEvaluatedSep 8, 2026
BenchmarkLM Arena Vision Style ControlScore1,252.07 ratingRank32Participants107Percentile70.75%EvidenceAEvaluatedAug 27, 2026
BenchmarkToolathlonScore42.90%Rank33Participants41Percentile20.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkSWE-Bench ProScore54.40%Rank39Participants55Percentile29.63%EvidenceCEvaluatedSep 8, 2026
BenchmarkGPQAScore88.00%Rank43Participants247Percentile82.93%EvidenceCEvaluatedSep 8, 2026
BenchmarkLM Arena VisionScore1,245.48 ratingRank45Participants107Percentile58.49%EvidenceAEvaluatedAug 27, 2026
BenchmarkLM Arena Text FactualityScore1,448.26 ratingRank46Participants121Percentile62.50%EvidenceAEvaluatedSep 2, 2026
BenchmarkHumanity's Last ExamScore28.20%Rank55Participants103Percentile47.06%EvidenceCEvaluatedSep 8, 2026
BenchmarkLM Arena Text Style ControlScore1,448.20 ratingRank56Participants210Percentile73.68%EvidenceAEvaluatedSep 2, 2026
BenchmarkLM Arena WebdevScore1,397.19 ratingRank61Participants97Percentile37.50%EvidenceAEvaluatedSep 5, 2026
BenchmarkLM Arena TextScore1,412.18 ratingRank91Participants210Percentile56.94%EvidenceAEvaluatedSep 2, 2026

GPT-5.4-mini Arena Results

Preference and agent-evaluation results for the default version.

30 of 100 rows
Columns

Show columns

Sort by
Arena
Category
Rank
Rating / score
Votes
Observations
Result date
Arenavision style controlCategoryhomeworkRank22Rating / score1,292.58Votes3,354ObservationsN/AResult dateAug 27, 2026
Arenatext factualityCategorypolishRank24Rating / score1,428.15Votes716ObservationsN/AResult dateSep 2, 2026
Arenatext factualityCategoryfrenchRank25Rating / score1,481.60Votes1,677ObservationsN/AResult dateSep 2, 2026
ArenavisionCategoryhomeworkRank25Rating / score1,286.90Votes3,354ObservationsN/AResult dateAug 27, 2026
Arenatext factualityCategoryjapaneseRank27Rating / score1,393.18Votes406ObservationsN/AResult dateSep 2, 2026
Arenavision style controlCategoryentity recognitionRank27Rating / score1,185.96Votes153ObservationsN/AResult dateAug 27, 2026
Arenatext factualityCategorykoreanRank28Rating / score1,394.70Votes618ObservationsN/AResult dateSep 2, 2026
ArenavisionCategoryentity recognitionRank29Rating / score1,204.55Votes153ObservationsN/AResult dateAug 27, 2026
Arenavision style controlCategorydiagramRank29Rating / score1,282.54Votes6,468ObservationsN/AResult dateAug 27, 2026
Arenavision style controlCategoryoverallRank32Rating / score1,252.07Votes24,076ObservationsN/AResult dateAug 27, 2026
Arenatext factualityCategoryindustry business and management and financial operationsRank33Rating / score1,459.76Votes10,157ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategorypolishRank33Rating / score1,460.90Votes1,164ObservationsN/AResult dateSep 2, 2026
Arenavision style controlCategorychineseRank33Rating / score1,289.16Votes1,355ObservationsN/AResult dateAug 27, 2026
Arenavision style controlCategoryhumorRank33Rating / score1,236.33Votes803ObservationsN/AResult dateAug 27, 2026
Arenavision style controlCategoryocrRank33Rating / score1,263.34Votes17,069ObservationsN/AResult dateAug 27, 2026
Arenatext factualityCategoryrussianRank34Rating / score1,461.51Votes4,925ObservationsN/AResult dateSep 2, 2026
Arenatext factualityCategorymathRank36Rating / score1,446.20Votes2,569ObservationsN/AResult dateSep 2, 2026
Arenavision style controlCategoryenglishRank36Rating / score1,248.92Votes10,099ObservationsN/AResult dateAug 27, 2026
Arenatext factualityCategoryindustry legal and governmentRank37Rating / score1,463.16Votes3,764ObservationsN/AResult dateSep 2, 2026
ArenavisionCategorydiagramRank39Rating / score1,264.34Votes6,468ObservationsN/AResult dateAug 27, 2026
ArenavisionCategorychineseRank40Rating / score1,289.00Votes1,355ObservationsN/AResult dateAug 27, 2026
ArenavisionCategoryhumorRank40Rating / score1,234.36Votes803ObservationsN/AResult dateAug 27, 2026
Arenatext factualityCategorynon englishRank41Rating / score1,438.67Votes31,857ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategoryindustry business and management and financial operationsRank41Rating / score1,459.59Votes12,110ObservationsN/AResult dateSep 2, 2026
Arenatext factualityCategoryhard promptsRank42Rating / score1,472.54Votes38,810ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategorymulti turnRank42Rating / score1,468.29Votes11,175ObservationsN/AResult dateSep 2, 2026
Arenavision style controlCategorycreative writing visionRank42Rating / score1,225.34Votes1,500ObservationsN/AResult dateAug 27, 2026
ArenavisionCategoryocrRank43Rating / score1,252.79Votes17,069ObservationsN/AResult dateAug 27, 2026
Arenatext factualityCategorymulti turnRank44Rating / score1,460.33Votes9,225ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategoryrussianRank45Rating / score1,454.19Votes6,274ObservationsN/AResult dateSep 2, 2026

GPT-5.4-mini Pricing

Official vendor API pricing appears first, followed by individual provider offers.

Official API
$0.75 input, $4.5 output per 1M
Official provider
OpenAI
Lowest third-party
From $0.375 input, $4 output per 1M via Xpersona
Tracked offerings
37
30 of 36 rows
Columns

Show columns

Sort by
Provider
Provider model ID
Region
Input / 1M
Output / 1M
Context
Updated
ProviderKenariProvider model IDgpt-5-4-miniRegionglobalInput / 1MN/AOutput / 1MN/AContext400KUpdatedSep 8, 2026
ProviderXpersonaProvider model IDgpt-5.4-miniRegionglobalInput / 1M$0.375Output / 1M$4Context272KUpdatedSep 8, 2026
ProviderPoeProvider model IDopenai/gpt-5.4-miniRegionglobalInput / 1M$0.68Output / 1M$4Context400KUpdatedSep 8, 2026
ProviderOrcaRouterProvider model IDopenai/gpt-5.4-miniRegionglobalInput / 1M$0.75Output / 1M$4.5Context400KUpdatedSep 8, 2026
ProviderNanoGPTProvider model IDopenai/gpt-5.4-miniRegionglobalInput / 1M$0.75Output / 1M$4.5Context400KUpdatedSep 8, 2026
ProviderImpossiblProvider model IDopenai/gpt-5.4-miniRegionglobalInput / 1M$0.75Output / 1M$4.5Context400KUpdatedSep 8, 2026
ProviderOpenRouterProvider model IDopenai/gpt-5.4-miniRegionglobalInput / 1M$0.75Output / 1M$4.5Context400KUpdatedSep 8, 2026
ProviderCrossModelProvider model IDopenai/gpt-5.4-miniRegionglobalInput / 1M$0.75Output / 1M$4.5Context400KUpdatedSep 8, 2026
ProviderDatabricksProvider model IDdatabricks-gpt-5-4-miniRegionglobalInput / 1M$0.75Output / 1M$4.5Context400KUpdatedSep 8, 2026
ProviderOpenAIProvider model IDgpt-5.4-miniRegionglobalInput / 1M$0.75Output / 1M$4.5Context400KUpdatedSep 8, 2026
ProviderOpperProvider model IDopenai/gpt-5.4-miniRegionglobalInput / 1M$0.75Output / 1M$4.5Context400KUpdatedSep 8, 2026
ProviderLLM GatewayProvider model IDopenai/gpt-5.4-miniRegionglobalInput / 1M$0.75Output / 1M$4.5Context400KUpdatedSep 8, 2026
ProviderVivgridProvider model IDgpt-5.4-miniRegionglobalInput / 1M$0.75Output / 1M$4.5Context400KUpdatedSep 8, 2026
ProviderNEAR AI CloudProvider model IDopenai/gpt-5.4-miniRegionglobalInput / 1M$0.75Output / 1M$4.5Context400KUpdatedSep 8, 2026
ProviderMerge GatewayProvider model IDopenai/gpt-5.4-miniRegionglobalInput / 1M$0.75Output / 1M$4.5Context400KUpdatedSep 8, 2026
ProviderCloudflare AI GatewayProvider model IDopenai/gpt-5.4-miniRegionglobalInput / 1M$0.75Output / 1M$4.5Context128KUpdatedSep 8, 2026
ProviderFastRouterProvider model IDopenai/gpt-5.4-miniRegionglobalInput / 1M$0.75Output / 1M$4.5Context400KUpdatedSep 8, 2026
ProviderZenMuxProvider model IDopenai/gpt-5.4-miniRegionglobalInput / 1M$0.75Output / 1M$4.5Context400KUpdatedSep 8, 2026
ProviderFreeModelProvider model IDgpt-5.4-miniRegionglobalInput / 1M$0.75Output / 1M$4.5Context400KUpdatedSep 8, 2026
ProviderOfoxProvider model IDopenai/gpt-5.4-miniRegionglobalInput / 1M$0.75Output / 1M$4.5Context400KUpdatedSep 8, 2026
ProviderNeonProvider model IDgpt-5-4-miniRegionglobalInput / 1M$0.75Output / 1M$4.5Context400KUpdatedSep 8, 2026
ProviderAzure Cognitive ServicesProvider model IDgpt-5.4-miniRegionglobalInput / 1M$0.75Output / 1M$4.5Context400KUpdatedSep 8, 2026
Provider302.AIProvider model IDgpt-5.4-miniRegionglobalInput / 1M$0.75Output / 1M$4.5Context400KUpdatedSep 8, 2026
ProviderOpenCode ZenProvider model IDgpt-5.4-miniRegionglobalInput / 1M$0.75Output / 1M$4.5Context400KUpdatedSep 8, 2026
ProviderRequestyProvider model IDgpt-5.4-miniRegionglobalInput / 1M$0.75Output / 1M$4.5Context400KUpdatedSep 8, 2026
ProviderAzureProvider model IDgpt-5.4-miniRegionglobalInput / 1M$0.75Output / 1M$4.5Context400KUpdatedSep 8, 2026
ProviderFrogBotProvider model IDgpt-5-4-miniRegionglobalInput / 1M$0.75Output / 1M$4.5Context400KUpdatedSep 8, 2026
ProviderAIHubMixProvider model IDgpt-5.4-miniRegionglobalInput / 1M$0.75Output / 1M$4.5Context400KUpdatedSep 8, 2026
ProviderVercel AI GatewayProvider model IDopenai/gpt-5.4-miniRegionglobalInput / 1M$0.75Output / 1M$4.5Context400KUpdatedSep 8, 2026
ProviderEden AIProvider model IDopenai/gpt-5.4-miniRegionglobalInput / 1M$0.75Output / 1M$4.5Context400KUpdatedSep 8, 2026

GPT-5.4-mini Runtime Performance

Provider-specific output speed and catalog latency for GPT-5.4 mini. Runtime does not affect the capability score.

No runtime data

No provider-specific speed or latency record is linked to the default version yet.

Output Speed is generated output tokens received per second. Catalog latency is reported separately from observed provider TTFT.

GPT-5.4-mini Specifications

Technical details for the model's default version.

Version
GPT-5.4 mini
Released
Mar 17, 2026
Knowledge cutoff
Aug 31, 2025
Parameters
N/A
Context window
400K
Max output
128K
Inputs
image, text
Outputs
text
Open weights
No
License
Proprietary

GPT-5.4-mini Versions

Available versions of this model. The score column identifies the version used in the overall ranking.

1 row
Columns

Show columns

Sort by
Version
Released
LLMBoard
Parameters
Context
Max output
Open weights
License
VersionGPT-5.4 miniReleasedMar 17, 2026LLMBoard58.88ParametersN/AContext400KMax output128KOpen weightsNoLicenseProprietary

Models similar to GPT-5.4-mini

Recommendations prioritize the same model type and family, then the closest LLMBoard score.

#81-1.57
OP

GPT-5.2-Codex

OpenAI

57.31 LLMBoard

Details
#68+1.84
OP

GPT-5.1-Instant

OpenAI

60.72 LLMBoard

Details
#67+1.91
OP

GPT-5.1

OpenAI

60.79 LLMBoard

Details
#64+2.64
OP

GPT-5.3-Codex

OpenAI

61.52 LLMBoard

Details
#98-6.78
OP

GPT-5.5-Instant

OpenAI

52.10 LLMBoard

Details
#99-7.19
OP

o3

OpenAI

51.69 LLMBoard

Details

What is GPT-5.4-mini?

Key information about GPT-5.4-mini and its available data.

4-mini is an OpenAI large language model designed for high-volume workloads. It supports coding, reasoning, multimodal understanding, computer use, subagents, and tool use.

Data as of 2026-09-08.

FAQ

Common questions about GPT-5.4-mini.

When was GPT-5.4-mini released?

GPT-5.4-mini's default version was released on Mar 17, 2026.

How much does GPT-5.4-mini cost?

GPT-5.4-mini's official API price is $0.75 per million input tokens and $4.5 per million output tokens via OpenAI. The lowest tracked third-party offer starts at $0.375 input and $4 output via Xpersona.

Who created GPT-5.4-mini?

GPT-5.4-mini was created by OpenAI.

What is the context window for GPT-5.4-mini?

The default version has a 400K token context window.

Is GPT-5.4-mini open weight?

No. The default version is not marked as having publicly available weights.

How many API providers offer GPT-5.4-mini?

37 provider offerings are linked to the default version.

Official prices use only the vendor's configured official Provider and positive standard USD PAYG rates. Third-party offers remain explicitly labeled.

Browse runtime rankings