LLMBoard ranking
Compare the current agent model ranking from the LLMBoard Agent Score, eligible agent benchmarks, official token pricing and supporting task evidence.
Data as of 2026-09-08
This leaderboard ranking orders models by their domain-specific LLMBoard score aggregated from eligible benchmark evidence. Pricing, coverage and the source signal remain separate.
Rank | Model | LLMBoard Agent Score | Official input / 1M | Official output / 1M | Score updated |
|---|
| Rank01 | ModelAN | LLMBoard Agent Score86.35 | Official input / 1M$5 | Official output / 1M$25 | Score updated |
| Rank02 | ModelAN | LLMBoard Agent Score85.92 | Official input / 1M$10 | Official output / 1M$50 | Score updated |
| Rank03 | ModelOP | LLMBoard Agent Score82.84 | Official input / 1M$4 | Official output / 1M$20 | Score updated |
| Rank04 | ModelAN | LLMBoard Agent Score82.04 | Official input / 1M$10 | Official output / 1M$50 | Score updated |
| Rank05 | ModelMA | LLMBoard Agent Score78.26 | Official input / 1M$3 | Official output / 1M$15 | Score updated |
| Rank06 | ModelOP | LLMBoard Agent Score77.93 | Official input / 1M$5 | Official output / 1M$30 | Score updated |
| Rank07 | ModelDE | LLMBoard Agent Score75.09 | Official input / 1MN/A | Official output / 1MN/A | Score updated |
| Rank08 | ModelZA | LLMBoard Agent Score73.87 | Official input / 1M$1.4 | Official output / 1M$4.4 | Score updated |
| Rank09 | ModelXA | LLMBoard Agent Score72.64 | Official input / 1M$2 | Official output / 1M$6 | Score updated |
| Rank10 | ModelXA | LLMBoard Agent Score72.21 | Official input / 1M$2 | Official output / 1M$6 | Score updated |
| Rank11 | ModelAN | LLMBoard Agent Score70.70 | Official input / 1M$2 | Official output / 1M$10 | Score updated |
| Rank12 | ModelAN | LLMBoard Agent Score70.42 | Official input / 1M$5 | Official output / 1M$25 | Score updated |
| Rank13 | ModelAN | LLMBoard Agent Score70.09 | Official input / 1M$5 | Official output / 1M$25 | Score updated |
| Rank14 | ModelTE | LLMBoard Agent Score69.85 | Official input / 1MN/A | Official output / 1MN/A | Score updated |
| Rank15 | ModelZA | LLMBoard Agent Score66.97 | Official input / 1M$1.4 | Official output / 1M$4.4 | Score updated |
| Rank16 | ModelAN | LLMBoard Agent Score66.59 | Official input / 1M$5 | Official output / 1M$25 | Score updated |
| Rank17 | ModelZA | LLMBoard Agent Score65.64 | Official input / 1M$0.075 | Official output / 1M$0.25 | Score updated |
| Rank18 | ModelOP | LLMBoard Agent Score64.79 | Official input / 1M$2 | Official output / 1M$12 | Score updated |
| Rank19 | ModelGO | LLMBoard Agent Score62.62 | Official input / 1M$0.75 | Official output / 1M$3.75 | Score updated |
| Rank20 | ModelAC | LLMBoard Agent Score60.73 | Official input / 1M$2 | Official output / 1M$6 | Score updated |
| Rank21 | ModelOP | LLMBoard Agent Score59.78 | Official input / 1M$2.5 | Official output / 1M$15 | Score updated |
| Rank22 | ModelDE | LLMBoard Agent Score59.12 | Official input / 1M$0.14 | Official output / 1M$0.28 | Score updated |
| Rank23 | ModelOP | LLMBoard Agent Score58.32 | Official input / 1M$0.20 | Official output / 1M$1.2 | Score updated |
| Rank24 | ModelMA | LLMBoard Agent Score56.33 | Official input / 1M$0.95 | Official output / 1M$4 | Score updated |
| Rank25 | ModelMA | LLMBoard Agent Score52.55 | Official input / 1M$0.95 | Official output / 1M$4 | Score updated |
| Rank26 | ModelME | LLMBoard Agent Score47.73 | Official input / 1M$1.25 | Official output / 1M$4.25 | Score updated |
| Rank27 | ModelAC | LLMBoard Agent Score47.68 | Official input / 1MN/A | Official output / 1MN/A | Score updated |
| Rank28 | ModelAN | LLMBoard Agent Score47.07 | Official input / 1M$3 | Official output / 1M$15 | Score updated |
| Rank29 | ModelAC | LLMBoard Agent Score45.70 | Official input / 1MN/A | Official output / 1MN/A | Score updated |
| Rank30 | ModelGO | LLMBoard Agent Score45.56 | Official input / 1M$0.75 | Official output / 1M$3.75 | Score updated |
This leaderboard overview connects current LLMBoard leaders with benchmark coverage and supporting source activity.
Use confidence and observed task volume to judge how firmly each leaderboard position is supported; these source proportions remain separate from benchmark coverage.
Vendor representation among the top-ranked models in this leaderboard.
A practical summary of the first five entries in this agent AI model leaderboard, with benchmark evidence and pricing kept in context.
Common questions about the Agent Model Leaderboard.
Claude Opus 5 is currently ranked first with 86.35 LLMBoard Agent Score.
The current leaders are Claude Fable 5.1 (rank #1), Claude Opus 5 (rank #2), and Claude Fable 5 (rank #3).
GLM 5.3 Flash has the lowest matched official input price at $0.075 per 1M tokens.
The fastest matched records are GLM 5.3 (472.69 tok/s via FriendliAI), Gemini 3.5 Flash (210.94 tok/s via Google), and GPT-5.5 (134.94 tok/s via OpenAI).
No. Source Arena results describe a particular task or preference signal. Capability benchmarks, prices and runtime can produce different rankings.
The LLMBoard score normalizes eligible evidence by task dimension, applies domain weights and reports coverage separately. The ranking follows that score.
This page currently compares 49 models.
A model appears after LLMBoard calculates a score for this domain. A source Arena result is optional supporting evidence.
Ranking basisThis agent AI model leaderboard uses the domain-specific LLMBoard score, evidence coverage and status. The leaderboard ranking keeps matched price and speed data separate from benchmark evidence.
Selection summary
Claude Opus 5 currently leads the agent ranking at 86.35. Compare evidence coverage, source activity and price separately before choosing a model for production.
Use this leaderboard with the supporting benchmark results and coverage details above. A leaderboard position summarizes the selected ranking signal; it does not replace workload-specific testing.