llmboard.aiAI model intelligence
Home

Model Rankings

OverallOpen ModelsAgentCodingReasoningMathKnowledgeInstruction FollowingTextVision
Image GenerationImage Editing
Video GenerationImage to VideoVideo Editing
Text to SpeechSpeech to Text
Embeddings

Efficiency

Chat Token PricingImage PricingVideo PricingAudio Pricing
Chat Speed & LatencyProvider Reliability

Benchmarks

GPQAMMLU-ProAIME 2025SWE-Bench VerifiedMMLUHumanity's Last ExamLiveCodeBenchMATHHumanEvalMMMU-Pro
All Benchmarks

Tools

Model Directory

Scoring & Data

Scoring & Data
1224 models729 benchmarks

Leaderboard Center

Overall RankingCodingCore BenchmarksPrice & ValueRuntime Performance

Modalities

All ModelsImage GenerationImage EditingVideo GenerationImage-to-VideoVideo EditingText-to-SpeechSpeech-to-TextEmbeddings

Data & Methods

Scoring MethodAll Benchmarks
llmboard.aiCopyright 2026 llmboard.ai

Model catalog

Text-to-Speech Model Leaderboard

This text-to-speech leaderboard compares the available benchmark ranking, score status, per-character pricing and media capabilities without inventing scores for unranked models.

Data as of 2026-09-08

On this page

  • Catalog
  • Overview
  • Pricing
  • Modalities
  • FAQ

Text-to-Speech Leaderboard Models

Browse the text-to-speech leaderboard ranking alongside the full catalog. Models without enough comparable benchmark evidence remain unscored.

30 of 111 rows
Columns

Show columns

Sort by
Model
LLMBoard score
Input -> output
Native price
Provider
Context
Released
ModelCASonic 3.6CartesiaLLMBoard score66.63Input -> outputtext -> audioNative priceN/AProviderN/AContextN/AReleasedN/A
ModelINRealtime TTS-2InworldLLMBoard score66.28Input -> outputtext -> audioNative priceN/AProviderN/AContextN/AReleasedN/A
ModelALQwen-Audio-3.0-TTS-PlusAlibabaLLMBoard score65.93Input -> outputtext -> audioNative priceN/AProviderN/AContextN/AReleasedN/A
ModelSPSimba 3.2SpeechifyAILLMBoard score65.58Input -> outputtext -> audioNative priceN/AProviderN/AContextN/AReleasedN/A
ModelVLLuna TTSVUI LabsLLMBoard score65.23Input -> outputtext -> audioNative priceN/AProviderN/AContextN/AReleasedN/A
ModelINRealtime TTS-2 Flash - Research PreviewInworldLLMBoard score64.88Input -> outputtext -> audioNative priceN/AProviderN/AContextN/AReleasedN/A
ModelBRBreeze TTS 2BreezeBlueLLMBoard score64.53Input -> outputtext -> audioNative priceN/AProviderN/AContextN/AReleasedN/A
ModelELv3 ConversationalElevenLabsLLMBoard score64.18Input -> outputtext -> audioNative priceN/AProviderN/AContextN/AReleasedN/A
ModelGOGemini 3.1 Flash TTSGoogleLLMBoard score63.83Input -> outputtext -> audioNative priceN/AProviderN/AContextN/AReleasedN/A
ModelINInworld TTS 1 MaxInworldLLMBoard score63.52Input -> outputtext -> audioNative priceN/AProviderN/AContextN/AReleasedN/A
ModelSTStepAudio 2.5 TTS (Aug 2026)StepFunLLMBoard score63.48Input -> outputtext -> audioNative priceN/AProviderN/AContextN/AReleasedN/A
ModelCASonic 3.5CartesiaLLMBoard score63.13Input -> outputtext -> audioNative priceN/AProviderN/AContextN/AReleasedN/A
ModelSOSoniox TTS Real-Time v2SonioxLLMBoard score62.43Input -> outputtext -> audioNative priceN/AProviderN/AContextN/AReleasedN/A
ModelINInworld TTS 1InworldLLMBoard score61.83Input -> outputtext -> audioNative priceN/AProviderN/AContextN/AReleasedN/A
ModelSMLightning V3.1 Pro (Jul 2026)Smallest.aiLLMBoard score61.46Input -> outputtext -> audioNative priceN/AProviderN/AContextN/AReleasedN/A
ModelMAFalcon 2Murf AILLMBoard score61.38Input -> outputtext -> audioNative priceN/AProviderN/AContextN/AReleasedN/A
ModelGRGradium TTS (Aug 2026)GradiumLLMBoard score60.68Input -> outputtext -> audioNative priceN/AProviderN/AContextN/AReleasedN/A
ModelASAsync Flash v1.5asyncLLMBoard score60.16Input -> outputtext -> audioNative priceN/AProviderN/AContextN/AReleasedN/A
ModelFAFish Audio S2.1 ProFish AudioLLMBoard score60.16Input -> outputtext -> audioNative priceN/AProviderN/AContextN/AReleasedN/A
ModelMISpeech 2.8 HDMiniMaxLLMBoard score60.09Input -> outputtext -> audioNative priceN/AProviderN/AContextN/AReleasedN/A
ModelSTStep TTS 2 (Mar 2026)StepFunLLMBoard score59.63Input -> outputtext -> audioNative priceN/AProviderN/AContextN/AReleasedN/A
ModelSMLightning V3.1 Pro TTS (Jun 2026)Smallest.aiLLMBoard score58.93Input -> outputtext -> audioNative priceN/AProviderN/AContextN/AReleasedN/A
ModelMIAzure HD 2.5MicrosoftLLMBoard score58.58Input -> outputtext -> audioNative priceN/AProviderN/AContextN/AReleasedN/A
ModelFAFish Audio S2 ProFish AudioLLMBoard score58.23Input -> outputtext -> audioNative priceN/AProviderN/AContextN/AReleasedN/A
ModelXASpaceXAI TTSxAILLMBoard score57.53Input -> outputtext -> audioNative priceN/AProviderN/AContextN/AReleasedN/A
ModelMISpeech 2.8 TurboMiniMaxLLMBoard score57.21Input -> outputtext -> audioNative priceN/AProviderN/AContextN/AReleasedN/A
ModelSPSimba 3.0SpeechifyAILLMBoard score57.18Input -> outputtext -> audioNative priceN/AProviderN/AContextN/AReleasedN/A
ModelASAsync Pro v1.0asyncLLMBoard score56.83Input -> outputtext -> audioNative priceN/AProviderN/AContextN/AReleasedN/A
ModelMISpeech 2.6 TurboMiniMaxLLMBoard score56.47Input -> outputtext -> audioNative priceN/AProviderN/AContextN/AReleasedN/A
ModelOPTTS 1 HDOpenAILLMBoard score56.13Input -> outputtext -> audioNative price$30 / 1M charactersProviderOpenAIContextN/AReleasedN/A

This leaderboard uses evaluation and benchmark evidence only for scores. Price, context and model specifications remain separate from the ranking.

Leaderboard overview

Models111
Vendors44
Scored models98
Ranking status98 provisional
Models11144 vendors
Multimodal inputs0Products accepting more than one input modality

Vendor catalog mix

111 models
Google10
MiniMax10
ElevenLabs6
Cartesia6
Microsoft5
Fish Audio4
Inworld4
Amazon4

Text-to-Speech Leaderboard Pricing

Use price as context for the leaderboard, not as part of its benchmark ranking. Each listing keeps its original unit.

per million characters

Up to $10 / 1M characters
2
$10 / 1M characters-$30 / 1M characters
4
Above $30 / 1M characters
1

Text-to-Speech Input and Output Modalities

Compare input and output support alongside the leaderboard; modality support does not change the benchmark ranking.

ModelIn: textIn: imageIn: audioIn: videoOut: textOut: imageOut: audioOut: video
ArcanaRimeYes-----Yes-
AAsync Flash v1.0asyncYes-----Yes-
AAsync Flash v1.5asyncYes-----Yes-
AAsync Pro v1.0asyncYes-----Yes-
Aura AsteriaDeepgramYes-----Yes-
Aura LunaDeepgramYes-----Yes-
Aura StellaDeepgramYes-----Yes-
Azure HD 2.5MicrosoftYes-----Yes-
Azure NeuralMicrosoftYes-----Yes-
BABland Speech v3Bland AIYes-----Yes-
BBreeze TTS 2BreezeBlueYes-----Yes-
RAChatterboxResemble AIYes-----Yes-

About Text-to-Speech Model Leaderboard

How does this text-to-speech leaderboard work?

98 models currently have a domain-specific LLMBoard score. The leaderboard ranking uses matched benchmark evidence, while coverage and provisional status identify incomplete evidence.

Which text-to-speech model has the lowest listed price?

Speech 02 Turbo has the lowest current native price at $7.5 / 1M characters via MiniMax (minimax.io).

Which text-to-speech model has the highest LLMBoard score?

Sonic 3.6 leads the available LLMBoard scores at 66.63.

How many text-to-speech models support multiple input modalities?

0 of 111 catalog models have more than one listed input modality.

How many text-to-speech models have no native price?

104 catalog models currently have no positive native-unit price listing.

How are prices displayed?

Price availability

With native price
7
Without native price
104

Prices retain the unit returned by the catalog, such as per image, per second, per minute or per million characters. They are never shown as token prices.

Why can a model have no price?

Some models do not have a current price listing in their native billing unit.

Lowest listed native priceSpeech 02 Turbo$7.5 / 1M characters via MiniMax (minimax.io)