llmboard.aiAI model intelligence
Home

Model Rankings

OverallOpen ModelsAgentCodingReasoningMathKnowledgeInstruction FollowingTextVision
Image GenerationImage Editing
Video GenerationImage to VideoVideo Editing
Text to SpeechSpeech to Text
Embeddings

Efficiency

Chat Token PricingImage PricingVideo PricingAudio Pricing
Chat Speed & LatencyProvider Reliability

Benchmarks

GPQAMMLU-ProAIME 2025SWE-Bench VerifiedMMLUHumanity's Last ExamLiveCodeBenchMATHHumanEvalMMMU-Pro
All Benchmarks

Tools

Model Directory

Scoring & Data

Scoring & Data
1224 models729 benchmarks

Leaderboard Center

Overall RankingCodingCore BenchmarksPrice & ValueRuntime Performance

Modalities

All ModelsImage GenerationImage EditingVideo GenerationImage-to-VideoVideo EditingText-to-SpeechSpeech-to-TextEmbeddings

Data & Methods

Scoring MethodAll Benchmarks
llmboard.aiCopyright 2026 llmboard.ai

retrieval benchmark

ViDoRe v3 Overall Leaderboard

ViDoRe v3 Overall is the current overall leaderboard for ViDoRe(v3).

Updated Sep 8, 2026

Models40
Model coverage40
MetricMean task score
EvidenceA

On this page

  • Ranking
  • Highlights
  • Distribution
  • Top models
  • About
  • FAQ

ViDoRe v3 Overall Ranking

Higher mean task score ranks better on this benchmark.

30 of 40 rows
Columns

Show columns

Sort by
Rank
Model
Mean task score
Percentile
Participants
Evidence
Evaluated
Rank01ModelWOwebAI-ColVec1.1-8bWebai OfficialMean task score64.95%Percentile100.00%Participants40EvidenceAEvaluatedN/A
Rank02ModelVUVultronRetrieverPrime-Qwen3.5-8BVultrMean task score64.26%Percentile97.44%Participants40EvidenceAEvaluatedN/A
Rank03ModelWOwebAI-ColVec1.1-4bWebai OfficialMean task score63.90%Percentile94.87%Participants40EvidenceAEvaluatedN/A
Rank04ModelVUVultronRetrieverCore-Qwen3.5-4.5BVultrMean task score63.57%Percentile92.31%Participants40EvidenceAEvaluatedN/A
Rank05ModelNVnemotron-colembed-vl-8b-v2NVIDIAMean task score63.42%Percentile89.74%Participants40EvidenceAEvaluatedN/A
Rank06ModelWOwebAI-ColVec1-9bWebai OfficialMean task score63.00%Percentile87.18%Participants40EvidenceAEvaluatedN/A
Rank07ModelWOwebAI-ColVec1-4bWebai OfficialMean task score62.22%Percentile84.62%Participants40EvidenceAEvaluatedN/A
Rank08ModelTOtomoro-colqwen3-embed-8bTomoroaiMean task score61.59%Percentile82.05%Participants40EvidenceAEvaluatedN/A
Rank09ModelNVnemotron-colembed-vl-4b-v2NVIDIAMean task score61.54%Percentile79.49%Participants40EvidenceAEvaluatedN/A
Rank10ModelAScolqwen3.5-4.5B-v3Athrael SojuMean task score61.46%Percentile76.92%Participants40EvidenceAEvaluatedN/A
Rank11ModelOAOps-Colqwen3-4BOpensearch AiMean task score61.17%Percentile74.36%Participants40EvidenceAEvaluatedN/A
Rank12ModelTOtomoro-colqwen3-embed-4bTomoroaiMean task score60.20%Percentile71.79%Participants40EvidenceAEvaluatedN/A
Rank13ModelNVllama-nemotron-colembed-vl-3b-v2NVIDIAMean task score59.79%Percentile69.23%Participants40EvidenceAEvaluatedN/A
Rank14ModelACQwen3-VL-Embedding-8BAlibaba Cloud / Qwen TeamMean task score57.78%Percentile66.67%Participants40EvidenceAEvaluatedN/A
Rank15ModelNAcolnomic-embed-multimodal-7bNomic AiMean task score57.33%Percentile64.10%Participants40EvidenceAEvaluatedN/A
Rank16ModelNVllama-nemoretriever-colembed-3b-v1NVIDIAMean task score57.26%Percentile61.54%Participants40EvidenceAEvaluatedN/A
Rank17ModelVUVultronRetrieverFlash-Qwen3.5-0.8BVultrMean task score56.16%Percentile58.97%Participants40EvidenceAEvaluatedN/A
Rank18ModelNAcolnomic-embed-multimodal-3bNomic AiMean task score55.78%Percentile56.41%Participants40EvidenceAEvaluatedN/A
Rank19ModelNVllama-nemoretriever-colembed-1b-v1NVIDIAMean task score55.59%Percentile53.85%Participants40EvidenceAEvaluatedN/A
Rank20ModelACQwen3-VL-Embedding-2BAlibaba Cloud / Qwen TeamMean task score52.64%Percentile51.28%Participants40EvidenceAEvaluatedN/A
Rank21ModelVIcolqwen2.5-v0.2VidoreMean task score51.90%Percentile48.72%Participants40EvidenceAEvaluatedN/A
Rank22ModelJIjina-embeddings-v4JinaaiMean task score48.52%Percentile46.15%Participants40EvidenceAEvaluatedN/A
Rank23ModelJIjina-embeddings-v5-omni-smallJinaaiMean task score45.61%Percentile43.59%Participants40EvidenceAEvaluatedN/A
Rank24ModelJIjina-embeddings-v5-omni-nanoJinaaiMean task score45.56%Percentile41.03%Participants40EvidenceAEvaluatedN/A
Rank25ModelVIcolqwen2-v1.0VidoreMean task score44.65%Percentile38.46%Participants40EvidenceAEvaluatedN/A
Rank26ModelVIcolpali-v1.3VidoreMean task score43.05%Percentile35.90%Participants40EvidenceAEvaluatedN/A
Rank27ModelMOcolmodernvbertModernvbertMean task score26.07%Percentile33.33%Participants40EvidenceAEvaluatedN/A
Rank28ModelVIcolSmol-256MVidoreMean task score21.40%Percentile30.77%Participants40EvidenceAEvaluatedN/A
Rank29ModelGOsiglip-large-patch16-384GoogleMean task score16.44%Percentile28.21%Participants40EvidenceAEvaluatedN/A
Rank30ModelGOsiglip-so400m-patch14-384GoogleMean task score15.75%Percentile25.64%Participants40EvidenceAEvaluatedN/A

ViDoRe v3 Overall Highlights

The leading models and scores on this benchmark.

ViDoRe v3 Overall Score Distribution

A closer view of the leading scores on this benchmark.

ViDoRe v3 Overall

The Top AI Models for ViDoRe v3 Overall

The first five results on this benchmark, with official price and output speed added where the model identity can be matched.

What is ViDoRe v3 Overall?

What ViDoRe v3 Overall measures and how its scores work.

ViDoRe v3 Overall is a retrieval benchmark leaderboard.

It reports Mean task score as a ratio.

Scores are shown in ratio. This benchmark is verified and has an evidence level of A.

Family
ViDoRe v3
Modality
multimodal
Primary category
retrieval
Score direction
higher
LLMBoard eligible
No
Evaluation key
overall

MTEB. Benchmark scores retain their original unit. Overall score eligibility is shown separately.

FAQ

Common questions about ViDoRe v3 Overall.

Which model scores highest on ViDoRe v3 Overall?

webAI-ColVec1.1-8b is currently ranked first with 64.95%.

What are the top three models on ViDoRe v3 Overall?

The current leaders are webAI-ColVec1.1-8b (64.95%), VultronRetrieverPrime-Qwen3.5-8B (64.26%), and webAI-ColVec1.1-4b (63.90%).

Which ViDoRe v3 Overall model has the lowest official input price?

No matched official input price is currently available.

Which models are fastest among ViDoRe v3 Overall results?

No matched runtime record is currently available.

Does the highest mean task score result prove overall model quality?

No. This benchmark measures one defined capability or task. The overall LLMBoard score uses a separate aggregation across eligible benchmark evidence.

What does ViDoRe v3 Overall measure?

It reports Mean task score as a ratio.

Is a higher mean task score better?

Yes. Higher values rank better for this benchmark.

How many models are compared?

40 model results are currently shown.

Does this benchmark affect the overall score?

No. This benchmark is shown for reference but does not contribute to the overall score.

Rank #1webAI-ColVec1.1-8b64.95%
Rank #2VultronRetrieverPrime-Qwen3.5-8B64.26%
Rank #3webAI-ColVec1.1-4b63.90%
Rank #4VultronRetrieverCore-Qwen3.5-4.5B63.57%

Ranking basisThis vidore v3 overall AI model leaderboard uses descending mean task score in the benchmark original unit. The leaderboard ranking keeps matched price and speed data separate from benchmark evidence.

  1. 01
    WO
    webAI-ColVec1.1-8bWebai Official
    Mean task score
    64.95%

    Strengths

    • Ranks #1 of 40 compared models
    • 100th percentile on this benchmark
    • A evidence result

    Considerations

    • This result measures ViDoRe v3 Overall, not total model capability
  2. 02
    VU
    VultronRetrieverPrime-Qwen3.5-8BVultr
    Mean task score
    64.26%

    Strengths

    • Ranks #2 of 40 compared models
    • 97th percentile on this benchmark
    • A evidence result

    Considerations

    • This result measures ViDoRe v3 Overall, not total model capability
  3. 03
    WO
    Webai Official
    Mean task score
    63.90%

    Strengths

    • Ranks #3 of 40 compared models
    • 95th percentile on this benchmark
    • A evidence result

    Considerations

    • This result measures ViDoRe v3 Overall, not total model capability
  4. 04
    VU
    Vultr
    Mean task score
    63.57%

    Strengths

    • Ranks #4 of 40 compared models
    • 92th percentile on this benchmark
    • A evidence result

    Considerations

    • This result measures ViDoRe v3 Overall, not total model capability
  5. 05
    NV
    NVIDIA
    Mean task score
    63.42%

    Strengths

    • Ranks #5 of 40 compared models
    • 90th percentile on this benchmark
    • A evidence result

    Considerations

    • This result measures ViDoRe v3 Overall, not total model capability

Selection summary

Best AI Models for ViDoRe v3 Overall

webAI-ColVec1.1-8b currently leads ViDoRe v3 Overall with 64.95%. It is the top model on this specific benchmark, while the best LLM for the broader task should also be checked against other benchmarks, price, and runtime.

Use this leaderboard with the supporting benchmark results and coverage details above. A leaderboard position summarizes the selected ranking signal; it does not replace workload-specific testing.

webAI-ColVec1.1-4b
VultronRetrieverCore-Qwen3.5-4.5B
nemotron-colembed-vl-8b-v2
Benchmark rank #1webAI-ColVec1.1-8b64.95%
Benchmark rank #2VultronRetrieverPrime-Qwen3.5-8B64.26%
Benchmark rank #3webAI-ColVec1.1-4b63.90%