Share your thoughts, 1 month free Claude Pro on usSee more
WorkDL logo mark

TOY

Benchmarks

Task NameDataset NameSOTA ResultTrend
MSI CompletionToy MSI database
PSNR48.878
28
Sequential RecommendationToy
NDCG@54.38
18
Sequential RecommendationToy
NDCG@104.85
18
Explicit AttackToy
Avg Queries (E)500
17
Toy packingToy Unseen States (test)
Success Rate75.5
16
RationalizationToy (test)
HI-F176.02
12
Single-object generationToy4K
PSNR23.98
11
Sampling Time Series Super-Resolution (SSR)Toy #2
MSE0.016
8
2d multi-goalToy
Recovery Time (%)3.2
8
ClassificationToy Synthetic Skew (test)
F1 Score99.93
7
Unsupervised point predictionToy
RMSE1.754
6
ClassificationToy (test)
F1 Score99.92
5
Generative RecommendationToy
NDCG@50.0341
4
Averaging Time Series Super-Resolution (ASR)Toy #3
MSE0.024
4
Averaging Time Series Super-Resolution (ASR)Toy #1
MSE0.013
4
Sampling Time Series Super-Resolution (SSR)Toy #3
MSE0.015
4
Sampling Time Series Super-Resolution (SSR)Toy #1
MSE0.008
4
State EstimationToy
RMSE78.03
4
Cell detectionTOY
AP @ IoU=0.5099.98
4
Set-level AttributionToy
Shape Accuracy100
3
Real2Sim Reconstruction and Interaction PredictionToy4K real-world experiment
Stability73.3
2
High-dimensional predictionToy-512
Average Regret0.29
2
High-dimensional predictionToy-256
Average Regret1.29
2
High-dimensional predictionToy-128
Average Regret4.18
2
High-dimensional predictionToy-64
Average Regret5.61
2
Showing 25 of 31 rows