Share your thoughts, 1 month free Claude Pro on us
See more
Feedback
Search any
task
Search any
task
SOTA Vision-Language Understanding benchmarks and papers with code | Wizwand
Share your thoughts, 1 month free Claude Pro on us
See more
Home
/
Tasks
Vision-Language Understanding
Benchmarks
Dataset Name
SOTA Method
Dataset Name
SOTA Method
Metric
Trend
Results
Last Updated
MMBench
Qwen3-VL-4B-Instruct
Accuracy
88.7
88
18d ago
MM-Vet
REVIS
Total Score
72.16
43
29d ago
Vision-Language Benchmark Suite Aggregate
Vanilla
Aggregate Performance (%)
100
34
3mo ago
MVista
ZPPO
Accuracy
80.5
28
1mo ago
MMBench cn
Vanilla
Accuracy
60.6
27
1mo ago
MME
Vanilla
Average Score
100
18
1mo ago
Naturalbench
SemDeDup
General Score
13.2
13
2mo ago
Winoground (test)
LLaVA-OV
Text Score
76.75
12
1mo ago
VStarBench
Qwen3-VL-8B
Accuracy
83.77
12
2mo ago
MVerse VI
Qwen3-VL-8B
Accuracy
43.78
12
2mo ago
MVerse VO
Qwen3-VL-8B
Accuracy
39.97
12
2mo ago
NegBench VOC2007
InternVL-3.0-8B
Accuracy
95.37
11
2mo ago
NegBench COCO
InternVL-3.0-8B
Accuracy
93.19
11
2mo ago
SEED-Bench Image
Vanilla
Average Accuracy
100
11
3mo ago
Vision-Language Evaluation Suite MMB, MMStar, MMMU, Hallusion, AI2D, OCR, SEED, SQA (test val)
Qwen2-VL-7B (Teacher)
MMB Score
80.7
10
4mo ago
MMStar (test)
ETTC
Accuracy
63.73
7
1mo ago
MMBench CN v1.1
Qwen3-Omni
Accuracy
87.7
5
2mo ago
MMBench EN v1.1
MiniCPM-o 4.5
Accuracy
89
5
2mo ago
Winoground
CROCScore
Text Accuracy
61.5
5
3mo ago
MMBench (test)
Baseline
Overall Accuracy
78.8
4
1mo ago
Quantiphy
Phi-4-Multimodal
Average MRA
27.4
4
2mo ago
MMBench (dev)
Prompt Highlighter
Accuracy
69.7
4
4mo ago
MME Perception
Prompt Highlighter
MME Score
1,552.5
4
4mo ago
MMStar
SmoothSMoE annealed (k=2)
Accuracy
42.64
3
1mo ago
Vision-Language Evaluation Suite (ChartQA, DocVQA, AI2D, VQA, AndroidControl, CountBenchQA)
Our Method
ChartQA Accuracy
68.1
2
3mo ago
Showing 25 of 27 rows
25 / page
50 / page
100 / page
1
2
Search any
task
Search any
task
Privacy Policy
Terms of Service
FAQs
Swarm Docs