Share your thoughts, 1 month free Claude Pro on us
See more
Home
/
Benchmarks
Optimal Policy Estimation on Continuous Simulation Setting epsilon = 0.9
Loading...
0.06
Mean Regret
Super
0.0464
0.1382
0.23
0.3218
Sep 29, 2022
Mean Regret
Standard Deviation
Updated 1mo ago
Evaluation Results
Method
Method
Links
Mean Regret
Standard Deviation
Super
Sample size (n)=1000,...
2022.09
0.06
0.0063
SZonly
Sample size (n)=1000,...
2022.09
0.12
0.0529
Sonly
Sample size (n)=1000,...
2022.09
0.4
0.002
Feedback
Search any
task
Search any
task