Share your thoughts, 1 month free Claude Pro on usSee more
WorkDL logo mark

OpenEQA

Benchmarks

Task NameDataset NameSOTA ResultTrend
Embodied Question AnsweringOpenEQA v1.0 (test)
LLM-Match67.7
32
Episodic-memory Question AnsweringOpenEQA v1 (ScanNet)
LLM-Match87.7
29
Embodied Question AnsweringOpenEQA
Accuracy69.3
26
Embodied Question AnsweringOpenEQA EM-EQA
Accuracy86.8
24
Open-ended Video ReasoningOpenEQA (test)
Accuracy (< 30s)58.6
18
3D Question AnsweringOpenEQA
LLM-Match56.2
12
Active Question AnsweringOpenEQA v1 (HM3D)
LLM-Match85.1
12
Episodic-memory Question AnsweringOpenEQA v1 (HM3D)
LLM Match Score85.1
12
Embodied Question AnsweringOpenEQA EM-EQA Episodes up to 32 frames
LLM-Match Score86.8
10
Zero-shot understanding of embodied scenesOpenEQA
Score51.2
10
Embodied Question AnsweringOpenEQA (test)
GPT-4 Score86.8
10
Embodied Question AnsweringOpenEQA EM-EQA
LLM-Match86.8
8
Episodic Memory Embodied Question AnsweringOpenEQA 1/10 subset Unseen (val)
LLM Match Score61.5
7
Embodied Question AnsweringOpenEQA v1 (test)
Score85.1
5
Active Embodied Question AnsweringOpenEQA 184
LLM-SR52.6
4
Showing 15 of 15 rows