Share your thoughts, 1 month free Claude Pro on us
See more
Home
/
Benchmarks
Grounded Human-Object Interaction Video Generation on Real Scene 50 cases (test)
Loading...
28.49
VLM-QA
InteractAvatar
24.8292
25.7796
26.73
27.6804
Feb 2, 2026
VLM-QA
HQ
OQ
PI
DINOsub
CLIPtext
Syncconf
Updated 5mo ago
Evaluation Results
Method
Method
Links
VLM-QA
HQ
OQ
PI
DINOsub
CLIPtext
Syncconf
InteractAvatar
2026.02
28.49
91
13.5
79.4
66.5
28.9
5.51
HuMo
2026.02
26.23
84
11.6
78.1
63.5
28.8
5.33
HY-Avatar
2026.02
25.29
81.7
10.8
72.4
64.4
28.5
5.45
Wan-S2V
2026.02
24.97
85.1
11.5
74
64.6
28.5
5.36
Feedback
Search any
task
Search any
task