Share your thoughts, 1 month free Claude Pro on usSee more
WorkDL logo mark

Med

Benchmarks

Task NameDataset NameSOTA ResultTrend
Finding target T_uniformMeD
Average Similarity1
25
Data Leakage AttackMed
AP (alpha=0.5)92.5
24
Speculative DecodingMed
Throughput (tokens/s)128.51
22
Question Answeringmed02
RL Score0.19
14
Continual Semantic SegmentationMed Semi-Supervised-JASCL
Session 0 Dice Score73.6
9
Author DisambiguationMed
MRR87.1
9
Paper-Venue PredictionMed
NDCG0.446
9
Paper-Field (L2) ClassificationMed
NDCG38.4
9
Paper-Field (L1) ClassificationMed
NDCG70.9
9
Retrieval EfficiencyMed.
Retrieved Tokens23,657
8
ClusteringMED
Purity0.59
6
Finding target quiz distributionMeD
Average Iterations17
5
Machine TranslationMed. En-De out-of-domain WMT14 (test)
BLEU30.8
5
Natural Language InferenceMED (test)
Upward Acc91.4
5
Traveling Salesman ProblemMed Small (test)
Solution Gap (%)1.113
2
Showing 15 of 15 rows