Share your thoughts, 1 month free Claude Pro on usSee more
WorkDL logo mark

MMLongBench-Doc

Benchmarks

Task NameDataset NameSOTA ResultTrend
Multimodal Long-document UnderstandingMMLongBench-Doc 1.0 (test)
Reports Score46.7
12
Multimodal VQAMMLongBench-Doc 2K Packed
Accuracy (Residential)55.66
11
Question AnsweringMMLongBench-Doc Multi-page (MP)
EM24.18
10
Document RetrievalMMLongBench-Doc Multi-page (MP)
Recall61.38
10
Question AnsweringMMLongBench-Doc Single-page
EM29.34
10
Document RetrievalMMLongBench-Doc Single-page (SP)
Top-1 Accuracy75.77
10
Long-context document understandingMMLongBench-Doc 32K context slice
MMLongBench-Doc (32K) Score45
4
Long-context document understandingMMLongBench-Doc 16K context slice
Score48.57
4
Showing 8 of 8 rows