Share your thoughts, 1 month free Claude Pro on us
See more
Home
/
Benchmarks
Point-level consensus correctness prediction on LT
Loading...
0.465
AUPRC
CAKE(PR)
0.39948
0.41649
0.4335
0.45051
Feb 20, 2026
AUPRC
AUROC
Updated 1mo ago
Evaluation Results
Method
Method
Links
AUPRC
AUROC
CAKE(PR)
Clustering algorithm=k...
2026.02
0.465
0.661
CAKE(HM)
Clustering algorithm=k...
2026.02
0.465
0.664
Bootstrap stability
Clustering algorithm=k...
2026.02
0.409
0.602
Entropy agreement
Clustering algorithm=k...
2026.02
0.402
0.617
Feedback
Search any
task
Search any
task