Share your thoughts, 1 month free Claude Pro on us
See more
Home
/
Benchmarks
Long-context Question Answering on LongBench Code Repository Understanding v2
Loading...
52
Accuracy
RLM
27.04
33.52
40
46.48
Jul 1, 2026
Accuracy
Latency (s)
Updated 24d ago
Evaluation Results
Method
Method
Links
Accuracy
Latency (s)
RLM
Backbone=Qwen3.6-35B-A3B
2026.07
52
187.8
MPLM
Backbone=Qwen3-30B-A3B
2026.07
35
65.5
RLM
Backbone=Qwen3-30B-A3B
2026.07
32.5
93.3
MPLM
Backbone=Qwen3.6-35B-A3B
2026.07
28
56.7
Feedback
Search any
task
Search any
task