Share your thoughts, 1 month free Claude Pro on us
See more
Home
/
Benchmarks
Multi-label Biomedical Classification on Ohsumed (val)
Loading...
80.86
Macro F1
EvoPool
4.2744
24.1572
44.04
63.9228
Jun 1, 2026
Macro F1
Updated 1d ago
Evaluation Results
Method
Method
Links
Macro F1
EvoPool
Backbone=gpt-4o-mini
2026.06
80.86
LLM annotation
Backbone=gpt-4o-mini
2026.06
42.64
Alchemist
Backbone=gpt-4o-mini
2026.06
18.72
DataSculpt
Backbone=gpt-4o-mini
2026.06
7.22
Feedback
Search any
task
Search any
task