Our new X account is live! Follow @wizwand_team for updates
Home
/
Benchmarks
LLM Routing on MMLU In-Domain
Loading...
0.6788
AUROC
ProbeDirichlet
0.428576
0.493538
0.5585
0.623462
Feb 12, 2026
AUROC
Updated 4d ago
Evaluation Results
Method
Method
Links
AUROC
ProbeDirichlet
signal_modality=Hidden...
2026.02
0.6788
EmbeddingMLP
signal_modality=Embedd...
2026.02
0.5489
SemanticEntropy
signal_modality=Verbos...
2026.02
0.5393
SelfAsk
signal_modality=Verbos...
2026.02
0.5375
Entropy
signal_modality=Logit-...
2026.02
0.4926
ConfidenceMargin
signal_modality=Logit-...
2026.02
0.4656
MaxLogits
signal_modality=Logit-...
2026.02
0.4382
Feedback
Search any
task
Search any
task