Share your thoughts, 1 month free Claude Pro on us
See more
Home
/
Benchmarks
Offline-to-Online Reinforcement Learning on D4RL Aggregate
Loading...
77.6
Average Normalized Score
Loss Smoothing
19.464
34.557
49.65
64.743
May 14, 2026
May 22, 2026
May 30, 2026
Jun 7, 2026
Jun 15, 2026
Jun 23, 2026
Jul 1, 2026
Average Normalized Score
Updated 23d ago
Evaluation Results
Method
Method
Links
Average Normalized Score
Loss Smoothing
2026.07
77.6
ROAD
Mixing Ratio Strategy=...
2026.05
71.95
OPT
2026.07
65.9
0.1
Mixing Ratio Strategy=...
2026.05
62.15
BR
Mixing Ratio Strategy=...
2026.05
61.8
Decreasing
Mixing Ratio Strategy=...
2026.05
59
0.3
Mixing Ratio Strategy=...
2026.05
58.66
Uniform
Mixing Ratio Strategy=...
2026.05
58.45
Hard-Switch
2026.07
58.2
0.2
Mixing Ratio Strategy=...
2026.05
57.28
0.4
Mixing Ratio Strategy=...
2026.05
56.49
0.5
Mixing Ratio Strategy=...
2026.05
56.26
0.0
Mixing Ratio Strategy=...
2026.05
54.07
Cal-QL
2026.07
48.8
Online
2026.07
39.4
Offline
2026.07
27.3
AWAC
2026.07
21.7
Feedback
Search any
task
Search any
task