llmboard.aiAI model intelligence
Home

Model Rankings

OverallOpen ModelsAgentCodingReasoningMathKnowledgeInstruction FollowingTextVision
Image GenerationImage Editing
Video GenerationImage to VideoVideo Editing
Text to SpeechSpeech to Text
Embeddings

Efficiency

Chat Token PricingImage PricingVideo PricingAudio Pricing
Chat Speed & LatencyProvider Reliability

Benchmarks

GPQAMMLU-ProAIME 2025SWE-Bench VerifiedMMLUHumanity's Last ExamLiveCodeBenchMATHHumanEvalMMMU-Pro
All Benchmarks

Tools

Model Directory

Scoring & Data

Scoring & Data
1281 models767 benchmarks

Leaderboard Center

Overall RankingCodingCore BenchmarksPrice & ValueRuntime Performance

Modalities

All ModelsImage GenerationImage EditingVideo GenerationImage-to-VideoVideo EditingText-to-SpeechSpeech-to-TextEmbeddings

Data & Methods

Scoring MethodAll Benchmarks
llmboard.aiCopyright 2026 llmboard.ai

Thinking Machines Lab model product

Inkling

Inkling is a general-purpose multimodal model that accepts text, image, and audio inputs and generates text outputs.

Updated Sep 25, 2026. Default version: Inkling

LLMBoard Score63.6Inkling
Coverage100%16 benchmark families
Context window524.3KTokens
Official input priceN/AOfficial price unavailable

On this page

  • Capability
  • Benchmarks
  • Arena
  • Pricing
  • Runtime
  • Specification
  • Versions
  • Similar models
  • About
  • FAQ

Inkling Capability Profile

This profile uses the model's current scored version. Arena ratings and prices are shown separately.

Inkling LLMBoard score breakdown

Inkling Benchmark Results

Benchmark scores for Inkling.

30 of 35 rows
Columns

Show columns

Sort by
Benchmark
Score
Rank
Participants
Percentile
Evidence
Evaluated
BenchmarkMMAUScore77.20%Rank01Participants4Percentile100.00%EvidenceCEvaluatedSep 24, 2026
BenchmarkSimpleQA VerifiedScore43.90%Rank01Participants4Percentile100.00%EvidenceCEvaluatedSep 24, 2026
BenchmarkVoiceBench AvgScore91.40%Rank01Participants3Percentile100.00%EvidenceCEvaluatedSep 24, 2026
BenchmarkAA-Omniscience IndexScore2.10 pointsRank02Participants3Percentile50.00%EvidenceCEvaluatedSep 24, 2026
BenchmarkAIME 2026Score97.10%Rank02Participants26Percentile96.00%EvidenceCEvaluatedSep 24, 2026
BenchmarkGlobal-MMLU-LiteScore88.70%Rank02Participants16Percentile93.33%EvidenceCEvaluatedSep 24, 2026
BenchmarkTau3 BankingScore23.70%Rank05Participants11Percentile60.00%EvidenceCEvaluatedSep 24, 2026
BenchmarkHumanity's Last Exam (no tools, text-only)Score29.70%Rank07Participants8Percentile14.29%EvidenceCEvaluatedSep 24, 2026
BenchmarkIFBenchScore79.80%Rank08Participants42Percentile82.93%EvidenceCEvaluatedSep 24, 2026
BenchmarkHumanity's Last Exam (with tools, text-only)Score46.00%Rank09Participants10Percentile11.11%EvidenceCEvaluatedSep 24, 2026
BenchmarkGDPval-AAScore1,238.00 pointsRank11Participants12Percentile9.09%EvidenceCEvaluatedSep 24, 2026
BenchmarkMCP AtlasScore76.00%Rank14Participants36Percentile62.86%EvidenceCEvaluatedSep 24, 2026
BenchmarkLiveBench math (2026-06-25)Score88.36 scoreRank21Participants41Percentile50.00%EvidenceBEvaluatedN/A
BenchmarkLM Arena Agent Bash Recovery StepsScore1.96%Rank22Participants42Percentile48.78%EvidenceAEvaluatedSep 15, 2026
BenchmarkLiveBench instruction (2026-06-25)Score70.10 scoreRank22Participants41Percentile47.50%EvidenceBEvaluatedN/A
BenchmarkSWE-Bench VerifiedScore77.60%Rank27Participants116Percentile77.39%EvidenceCEvaluatedSep 24, 2026
BenchmarkLM Arena Agent Tool HallucinationScore-0.03%Rank30Participants42Percentile29.27%EvidenceAEvaluatedSep 15, 2026
BenchmarkCharXiv-RScore78.10%Rank31Participants58Percentile47.37%EvidenceCEvaluatedSep 24, 2026
BenchmarkAA Omniscience AccuracyScore41.55%Rank31Participants201Percentile85.00%EvidenceBEvaluatedN/A
BenchmarkTerminal-Bench 2.1Score63.80%Rank32Participants42Percentile24.39%EvidenceCEvaluatedSep 24, 2026
BenchmarkLiveBench reasoning (2026-06-25)Score78.35 scoreRank32Participants41Percentile22.50%EvidenceBEvaluatedN/A
BenchmarkMMMU-ProScore73.50%Rank38Participants72Percentile47.89%EvidenceCEvaluatedSep 24, 2026
BenchmarkLM Arena Agent LeaderboardScore-10.02%Rank38Participants42Percentile9.76%EvidenceAEvaluatedSep 15, 2026
BenchmarkLM Arena Agent SteerabilityScore-10.51%Rank40Participants42Percentile4.88%EvidenceAEvaluatedSep 15, 2026
BenchmarkLM Arena Agent Task Outcome ExplicitScore-19.32%Rank40Participants42Percentile4.88%EvidenceAEvaluatedSep 15, 2026
BenchmarkAA SciCode SubtasksScore46.99%Rank40Participants89Percentile55.68%EvidenceBEvaluatedN/A
BenchmarkLM Arena Agent Praise ComplaintScore-22.20%Rank42Participants42Percentile0.00%EvidenceAEvaluatedSep 15, 2026
BenchmarkAA HLE Text No ToolsScore31.88%Rank46Participants200Percentile77.39%EvidenceBEvaluatedN/A
BenchmarkLM Arena Text FactualityScore1,449.67 ratingRank47Participants125Percentile62.90%EvidenceAEvaluatedSep 13, 2026
BenchmarkAA CritPtScore5.43%Rank47Participants200Percentile76.88%EvidenceBEvaluatedN/A

Inkling Arena Results

Preference and agent-evaluation results for the default version.

30 of 92 rows
Columns

Show columns

Sort by
Arena
Category
Rank
Rating / score
Votes
Observations
Result date
Arenatext factualityCategorymathRank09Rating / score1,487.33Votes881ObservationsN/AResult dateSep 13, 2026
Arenatext factualityCategoryspanishRank18Rating / score1,459.48Votes478ObservationsN/AResult dateSep 13, 2026
ArenatextCategorymathRank19Rating / score1,477.41Votes1,120ObservationsN/AResult dateSep 13, 2026
ArenatextCategorypolishRank21Rating / score1,464.38Votes403ObservationsN/AResult dateSep 13, 2026
Arenaagent bash recovery stepsCategoryoverallRank22Rating / score0.02VotesN/AObservations45.1KResult dateSep 15, 2026
Arenatext factualityCategoryindustry mathematicalRank23Rating / score1,466.97Votes1,088ObservationsN/AResult dateSep 13, 2026
Arenatext style controlCategorymathRank24Rating / score1,476.88Votes1,120ObservationsN/AResult dateSep 13, 2026
Arenaagent tool hallucinationCategoryoverallRank30Rating / score0.00VotesN/AObservations1.5MResult dateSep 15, 2026
Arenatext factualityCategoryindustry legal and governmentRank34Rating / score1,469.16Votes1,748ObservationsN/AResult dateSep 13, 2026
Arenatext style controlCategorypolishRank35Rating / score1,460.40Votes403ObservationsN/AResult dateSep 13, 2026
Arenatext factualityCategorychineseRank37Rating / score1,495.02Votes1,483ObservationsN/AResult dateSep 13, 2026
Arenatext factualityCategoryexpertRank37Rating / score1,484.23Votes2,595ObservationsN/AResult dateSep 13, 2026
ArenaagentCategoryoverallRank38Rating / score-0.10VotesN/AObservations1.6MResult dateSep 15, 2026
Arenatext factualityCategoryindustry business and management and financial operationsRank39Rating / score1,457.34Votes4,718ObservationsN/AResult dateSep 13, 2026
Arenaagent steerabilityCategoryoverallRank40Rating / score-0.11VotesN/AObservations30.9KResult dateSep 15, 2026
Arenaagent task outcome explicitCategoryoverallRank40Rating / score-0.19VotesN/AObservations24.1KResult dateSep 15, 2026
ArenatextCategoryjapaneseRank40Rating / score1,412.38Votes408ObservationsN/AResult dateSep 13, 2026
Arenatext factualityCategoryindustry software and it servicesRank40Rating / score1,490.46Votes10,436ObservationsN/AResult dateSep 13, 2026
ArenatextCategorykoreanRank41Rating / score1,401.90Votes500ObservationsN/AResult dateSep 13, 2026
Arenaagent praise complaintCategoryoverallRank42Rating / score-0.22VotesN/AObservations9.4KResult dateSep 15, 2026
ArenatextCategoryindustry mathematicalRank42Rating / score1,459.85Votes1,369ObservationsN/AResult dateSep 13, 2026
Arenatext factualityCategorynon englishRank42Rating / score1,439.32Votes15,006ObservationsN/AResult dateSep 13, 2026
ArenatextCategoryindustry business and management and financial operationsRank43Rating / score1,438.57Votes5,051ObservationsN/AResult dateSep 13, 2026
ArenatextCategoryexpertRank44Rating / score1,464.45Votes2,971ObservationsN/AResult dateSep 13, 2026
Arenatext factualityCategoryindustry life and physical and social scienceRank45Rating / score1,472.45Votes3,847ObservationsN/AResult dateSep 13, 2026
Arenatext factualityCategoryindustry medicine and healthcareRank45Rating / score1,472.79Votes1,450ObservationsN/AResult dateSep 13, 2026
ArenatextCategorychineseRank46Rating / score1,490.07Votes1,820ObservationsN/AResult dateSep 13, 2026
ArenatextCategorynon englishRank46Rating / score1,430.27Votes15,116ObservationsN/AResult dateSep 13, 2026
ArenatextCategoryindustry software and it servicesRank47Rating / score1,464.96Votes10,523ObservationsN/AResult dateSep 13, 2026
Arenatext factualityCategorycodingRank47Rating / score1,499.77Votes7,297ObservationsN/AResult dateSep 13, 2026

Inkling Pricing

Official vendor API pricing appears first, followed by individual provider offers.

Official API
N/A
Official provider
N/A
Lowest third-party
From $0.95 input, $4.05 output per 1M via Deep Infra
Tracked offerings
21
21 rows
Columns

Show columns

Sort by
Provider
Provider model ID
Region
Input / 1M
Output / 1M
Context
Updated
ProviderNvidiaProvider model IDthinkingmachines/inklingRegionglobalInput / 1MN/AOutput / 1MN/AContext1MUpdatedSep 25, 2026
ProviderDeep InfraProvider model IDthinkingmachines/InklingRegionglobalInput / 1M$0.95Output / 1M$4.05Context524.3KUpdatedSep 25, 2026
ProviderKilo GatewayProvider model IDthinkingmachines/inklingRegionglobalInput / 1M$0.95Output / 1M$4.05Context524.3KUpdatedSep 25, 2026
ProviderDevPass (LLM Gateway)Provider model IDinklingRegionglobalInput / 1M$0.95Output / 1M$4.05Context524.3KUpdatedSep 25, 2026
ProviderNanoGPTProvider model IDthinkingmachines/inklingRegionglobalInput / 1M$1Output / 1M$4.05Context1MUpdatedSep 25, 2026
ProviderHugging FaceProvider model IDthinkingmachines/InklingRegionglobalInput / 1M$1Output / 1M$4.05Context1MUpdatedSep 25, 2026
ProviderOpenRouterProvider model IDthinkingmachines/inklingRegionglobalInput / 1M$1Output / 1M$4.05Context1MUpdatedSep 25, 2026
ProviderMerge GatewayProvider model IDthinkingmachines/inklingRegionglobalInput / 1M$1Output / 1M$4.05Context1MUpdatedSep 25, 2026
ProviderTogether AIProvider model IDthinkingmachines/InklingRegionglobalInput / 1M$1Output / 1M$4.05Context524.3KUpdatedSep 25, 2026
ProviderNeonProvider model IDinklingRegionglobalInput / 1M$1Output / 1M$4.05Context1MUpdatedSep 25, 2026
ProviderBasetenProvider model IDthinkingmachines/inklingRegionglobalInput / 1M$1Output / 1M$4.05Context1MUpdatedSep 25, 2026
ProviderVercel AI GatewayProvider model IDthinkingmachines/inklingRegionglobalInput / 1M$1Output / 1M$4.05Context256KUpdatedSep 25, 2026
ProviderFireworks AIProvider model IDaccounts/fireworks/models/inklingRegionglobalInput / 1M$1Output / 1M$4.05Context1MUpdatedSep 25, 2026
ProviderCharm HyperProvider model IDinklingRegionglobalInput / 1M$1.09Output / 1M$4.41Context1MUpdatedSep 25, 2026
ProviderModalProvider model IDthinkingmachines/Inkling-NVFP4RegionglobalInput / 1M$1.2Output / 1M$5Context1MUpdatedSep 25, 2026
ProviderVenice AIProvider model IDinklingRegionglobalInput / 1M$1.25Output / 1M$5.06Context524.3KUpdatedSep 25, 2026
ProviderImpossiblProvider model IDthinkingmachines/inklingRegionglobalInput / 1M$1.87Output / 1M$4.68Context65.5KUpdatedSep 25, 2026
ProviderLLMTRProvider model IDthinkingmachines/inklingRegionglobalInput / 1M$1.87Output / 1M$4.68Context262.1KUpdatedSep 25, 2026
ProviderRequestyProvider model IDinklingRegionglobalInput / 1M$1.87Output / 1M$4.68Context65.5KUpdatedSep 25, 2026
ProviderThinking MachinesProvider model IDthinkingmachines/InklingRegionglobalInput / 1M$1.87Output / 1M$4.68Context65.5KUpdatedSep 25, 2026
ProviderAbacusProvider model IDthinkingmachines/InklingRegionglobalInput / 1M$3.74Output / 1M$9.36Context262.1KUpdatedSep 25, 2026

Inkling Runtime Performance

Provider-specific output speed and catalog latency for Inkling. Runtime does not affect the capability score.

No runtime data

No provider-specific speed or latency record is linked to the default version yet.

Output Speed is generated output tokens received per second. Catalog latency is reported separately from observed provider TTFT.

Inkling Specifications

Technical details for the model's default version.

Version
Inkling
Released
Jul 21, 2026
Knowledge cutoff
Unknown
Parameters
975B
Context window
524.3K
Max output
524.3K
Inputs
audio, image, text
Outputs
text
Open weights
Yes
License
Apache 2.0

Inkling Versions

Available versions of this model. The score column identifies the version used in the overall ranking.

1 row
Columns

Show columns

Sort by
Version
Released
LLMBoard
Parameters
Context
Max output
Open weights
License
VersionInklingReleasedJul 21, 2026LLMBoard63.57Parameters975BContext524.3KMax output524.3KOpen weightsYesLicenseApache 2.0

Models similar to Inkling

Recommendations prioritize the same model type and family, then the closest LLMBoard score.

#72-0.39
TM

Inkling Small

Thinking Machines Lab

63.18 LLMBoard

Details
#73-0.69
GO

Gemini 3.8 Flash Cyber

Google

62.88 LLMBoard

Details
#70+0.77
AC

Qwen3.5 397B A17B

Alibaba Cloud / Qwen Team

64.34 LLMBoard

Details
#69+0.79
GO

Gemini 3 Flash

Google

64.36 LLMBoard

Details
#68+1.49
MA

Kimi K2.5

Moonshot AI

65.06 LLMBoard

Details
#67+1.85
MA

Kimi K2.7 Code

Moonshot AI

65.42 LLMBoard

Details

What is Inkling?

Key information about Inkling and its available data.

Inkling is a general-purpose multimodal model from Thinking Machines Lab. It accepts text, image, and audio inputs and generates text outputs.

Data as of 2026-09-24.

FAQ

Common questions about Inkling.

When was Inkling released?

Inkling's default version was released on Jul 21, 2026.

How much does Inkling cost?

No official standard PAYG price is currently available for Inkling. The lowest tracked third-party offer starts at $0.95 input and $4.05 output via Deep Infra.

Who created Inkling?

Inkling was created by Thinking Machines Lab.

What is the context window for Inkling?

The default version has a 524.3K token context window.

Is Inkling open weight?

Yes. The default version is marked as open weight under Apache 2.0.

How many API providers offer Inkling?

21 provider offerings are linked to the default version.

Official prices use only the vendor's configured official Provider and positive standard USD PAYG rates. Third-party offers remain explicitly labeled.

Browse runtime rankings