Share your thoughts, 1 month free Claude Pro on us
See more
Home
/
Benchmarks
Optimal Policy Estimation on Continuous Simulation Setting (epsilon = 0.5)
Loading...
0.11
Mean Regret
SZonly
0.0984
0.1767
0.255
0.3333
Sep 29, 2022
Mean Regret
Standard Deviation
Updated 1mo ago
Evaluation Results
Method
Method
Links
Mean Regret
Standard Deviation
SZonly
Sample size (n)=1000,...
2022.09
0.11
0.0018
Super
Sample size (n)=1000,...
2022.09
0.11
0.0018
Sonly
Sample size (n)=1000,...
2022.09
0.4
0.0023
Feedback
Search any
task
Search any
task