llmboard.ai
Benchmarks
CompareRankings
llmboard.ai
Benchmarks
CompareRankings
HomeBenchmarksMultimodal

Benchmark category

Multimodal Benchmarks

Published multimodal evaluations and source-native model rankings. Each benchmark keeps its original scale and methodology.

Data as of 2026-08-07

Benchmarks166
Highest coverage65
Score eligible0

Multimodal benchmark registry

Select a benchmark to inspect model-level results, evidence fields and scoring direction.

36 rows
Columns

Show columns

MMMU-ProMMMU-PromultimodalScore65featuredBNo
MMMUMMMUmultimodalScore63featuredBNo
CharXiv-RCharXiv-RmultimodalScore47featuredBNo
MathVistaMathVistamultimodalScore39featuredBNo
AI2DAI2DmultimodalScore32featuredBNo
MathVisionMathVisionmultimodalScore32featuredBNo
DocVQADocVQAmultimodalScore26featuredBNo
VideoMMMUVideoMMMUmultimodalScore26featuredBNo
ChartQAChartQAmultimodalScore24featuredBNo
LVBenchLVBenchmultimodalScore24featuredBNo
ScreenSpot ProScreenSpot PromultimodalScore24featuredCNo
MathVista-MiniMathVista-MinimultimodalScore23featuredBNo
PolyMATHPolyMATHmultimodalScore23featuredBNo
MMStarMMStarmultimodalScore22featuredBNo
OSWorld-VerifiedOSWorld-VerifiedmultimodalScore22featuredBNo
OSWorldOSWorldmultimodalScore20featuredBNo
CC-OCRCC-OCRmultimodalScore18featuredBNo
MMBench-V1.1MMBench-V1.1multimodalScore18featuredBNo
MM-MT-BenchMM-MT-BenchmultimodalScore17featuredBNo
MVBenchMVBenchmultimodalScore17featuredBNo
Video-MMEVideo-MMEmultimodalScore17featuredBNo
CharXiv-DCharXiv-DmultimodalScore16featuredBNo
OmniDocBench 1.5OmniDocBench 1.5multimodalScore16featuredBNo
ScreenSpotScreenSpotmultimodalScore16featuredBNo
TextVQATextVQAmultimodalScore15featuredBNo
BLINKBLINKmultimodalScore13featuredBNo
SimpleVQASimpleVQAmultimodalScore13featuredBNo
CharadesSTACharadesSTAmultimodalScore12featuredBNo
InfoVQAtestInfoVQAtestmultimodalScore12featuredBNo
MedXpertQAMedXpertQAmultimodalScore12featuredBNo
DocVQAtestDocVQAtestmultimodalScore11featuredBNo
MMMU (val)MMMU (val)multimodalScore11featuredBNo
MuirBenchMuirBenchmultimodalScore11featuredBNo
MLVUMLVUmultimodalScore10featuredBNo
VideoMME w sub.VideoMME w sub.multimodalScore10featuredBNo
VideoMME w/o sub.VideoMME w/o sub.multimodalScore10featuredBNo

Top multimodal result sets

High-coverage benchmarks with at least two published model results.

MMMU-Pro

View benchmark

MMMU

View benchmark

CharXiv-R

View benchmark

MathVista

View benchmark

What are multimodal benchmarks?

How this category is assembled on llmboard.ai.

This page groups benchmarks whose primary or display category matches multimodal. It does not average incompatible metrics into a new category score.

Open an individual benchmark to inspect score direction, evidence level, participant count and source-native results.

Category membership is derived from the current benchmark registry response.

Rankings

OverallCodingText ArenaPricing

Modalities

Image GenerationVideo GenerationSpeech-to-TextEmbeddings

Benchmarks

All BenchmarksReasoningMathCoding

Vendors

All VendorsOpenAIAnthropicGoogle
llmboard.aiCopyright 2026 llmboard.ai