Share your thoughts, 1 month free Claude Pro on us
See more
Home
/
Benchmarks
Optimal Policy Estimation on Continuous Simulation Setting epsilon = 0.7
Loading...
0.1
Mean Regret
Super
0.088
0.169
0.25
0.331
Sep 29, 2022
Mean Regret
Standard Deviation
Updated 1mo ago
Evaluation Results
Method
Method
Links
Mean Regret
Standard Deviation
Super
Sample size (n)=1000,...
2022.09
0.1
0.0027
SZonly
Sample size (n)=1000,...
2022.09
0.12
0.0021
Sonly
Sample size (n)=1000,...
2022.09
0.4
0.0024
Feedback
Search any
task
Search any
task