Share your thoughts, 1 month free Claude Pro on usSee more
WorkDL logo mark

MDP

Benchmarks

Task NameDataset NameSOTA ResultTrend
Entropy EstimationMDP M1
Estimated Entropy1.023
6
Entropy EstimationMDP M2
Estimated Entropy1.324
3
Gittins index estimationMDP
Maximum Estimation Error0.135
3
Minimal Entropy Strategy SynthesisMDP M4
EntAns0.693
1
Minimal Entropy Strategy SynthesisMDP M3
Entropy Answer1.099
1
Entropy EstimationMDP M5
Exp(Estimated Entropy)2.954
1
Entropy EstimationMDP M4
Exp(Estimated Entropy)2
1
Entropy EstimationMDP M3
Exp(Estimated Entropy)3
1
Near-optimal policy identificationMDP
Metric-
0
Transfer Reinforcement LearningMDP with Tucker rank (d, d, d)
Metric-
0
Transfer Reinforcement LearningMDP with Tucker rank (S, S, d)
Metric-
0
Compute epsilon-optimal policyMDP sample setting
Metric-
0
Showing 12 of 12 rows