Share your thoughts, 1 month free Claude Pro on us
See more
Feedback
Search any
task
Search any
task
SOTA Regret Minimization benchmarks and papers with code | Wizwand
Share your thoughts, 1 month free Claude Pro on us
See more
Home
/
Tasks
Regret Minimization
Benchmarks
Dataset Name
SOTA Method
Dataset Name
SOTA Method
Metric
Trend
Results
Last Updated
Synthetic Subspace
Oracle GP-TS
Average Total Regret
86
10
4mo ago
Synthetic Lengthscale
Oracle GP-TS
Average Total Regret
28.1
10
4mo ago
Synthetic Kernel
Oracle GP-TS
Average total regret
35
10
4mo ago
Discrete feature values setting n=5000 (simulation)
SUPER
Mean Regret
0.17
9
1mo ago
Two dim reward function synthetic (test)
RMEL
Oracle Regret
2,589.32
9
5mo ago
Sine reward function synthetic (test)
RMEL
Oracle Regret
289.86
9
5mo ago
Triangle reward function synthetic (test)
Zooming
Oracle Regret
366.58
9
5mo ago
Two dim reward function weak adversaries Appendix A.7 (test)
RMEL
Oracle Regret
2,589.32
9
5mo ago
Sine reward function weak adversaries Appendix A.7 (test)
RMEL
Oracle Regret
280.62
9
5mo ago
Triangle reward function weak adversaries Appendix A.7 (test)
Zooming
Oracle Regret/Score
366.58
9
5mo ago
Episodic MNL mixture MDP
Hwang and Oh
Regret Bound
1
7
1mo ago
Episodic Tabular MDPs Adversarial Regime
Dann et al. (2023, Theorem 4.3)
Regret Upper Bound
2
6
25d ago
Combinatorial Bandits Theoretical
[10]
Regret
0.5
6
1mo ago
Synthetic Patient Cohort Ablation Grid
UCB-BOLD
Cumulative Normalized Regret CVaR (1-alpha=0.5)
0.13
6
2mo ago
PNW Precip 1980-1994 (test)
HP-GP-TS
Average Total Regret
167.7
6
4mo ago
PeMS
PE-GP-TS
Average Total Regret
1,214.2
6
4mo ago
Intel
EEI
Average Total Regret
51.6
6
4mo ago
Episodic Tabular MDPs Stochastic Regime with Adversarial Corruption
Dann et al. (2023, Theorem 4.3)
Regret Upper Bound
2
5
25d ago
Multi-armed Bandit anytime setting
Naive implementation
Regret Coefficient
2
5
5mo ago
Uniform Auction Non-Stationary Adversary, K=5, T=5000
A3M
Final Regret
0.0029
4
26d ago
Regret minimization with static historical data
KLUCB-H
Regret Upper Bound
0
4
2mo ago
K-armed bandits Gaussian rewards
UCB1
Regret
114.4
4
19d ago
F_LR Stochastic Low-Rank Reward
Noisy power method (NPM)
Regret
2
4
4mo ago
F_EV Stochastic Eigenvalue Reward settings
Lower Bound
Regret
2
4
4mo ago
Matching Bandits Theoretical Bound
Liu et al.
Regret
2
3
26d ago
Showing 25 of 109 rows
25 / page
50 / page
100 / page
1
2
3
4
5
Search any
task
Search any
task
Privacy Policy
Terms of Service
FAQs
Swarm Docs