Our new X account is live! Follow @wizwand_team for updates
Home
/
Benchmarks
Downstream Performance Prediction on CMMLU
Loading...
0.0033
MSE
CSV
0.001985
0.010997
0.02001
0.029023
Jun 16, 2025
MSE
Updated 4d ago
Evaluation Results
Method
Method
Links
MSE
CSV
loss type=CSV (Capabil...
2025.06
0.0033
All token loss
loss type=All token
2025.06
0.0269
Label token loss
loss type=Label token
2025.06
0.0367
Feedback
Search any
task
Search any
task