llmboard.ai
Benchmarks
CompareRankings
llmboard.ai
Benchmarks
CompareRankings
HomeBenchmarksImage To Text

Benchmark category

Image To Text Benchmarks

Published image to text evaluations and source-native model rankings. Each benchmark keeps its original scale and methodology.

Data as of 2026-08-07

Benchmarks12
Highest coverage26
Score eligible0

Image To Text benchmark registry

Select a benchmark to inspect model-level results, evidence fields and scoring direction.

6 rows
Columns

Show columns

DocVQADocVQAmultimodalScore26featuredBNo
OCRBenchOCRBenchmultimodalScore22featuredBNo
TextVQATextVQAmultimodalScore15featuredBNo
SimpleVQASimpleVQAmultimodalScore13featuredBNo
OCRBench-V2 (en)OCRBench-V2 (en)multimodalScore12featuredBNo
OCRBench-V2 (zh)OCRBench-V2 (zh)multimodalScore11featuredBNo

Top image to text result sets

High-coverage benchmarks with at least two published model results.

DocVQA

View benchmark

OCRBench

View benchmark

TextVQA

View benchmark

SimpleVQA

View benchmark

What are image to text benchmarks?

How this category is assembled on llmboard.ai.

This page groups benchmarks whose primary or display category matches image to text. It does not average incompatible metrics into a new category score.

Open an individual benchmark to inspect score direction, evidence level, participant count and source-native results.

Category membership is derived from the current benchmark registry response.

Rankings

OverallCodingText ArenaPricing

Modalities

Image GenerationVideo GenerationSpeech-to-TextEmbeddings

Benchmarks

All BenchmarksReasoningMathCoding

Vendors

All VendorsOpenAIAnthropicGoogle
llmboard.aiCopyright 2026 llmboard.ai