Share your thoughts, 1 month free Claude Pro on usSee more
WorkDL logo mark

DCLM

Benchmarks

Task NameDataset NameSOTA ResultTrend
Language ModelingDCLM (val)
Loss2.589
78
Zero-shot EvaluationDCLM CORE V2
CORE_V2 Score48
17
Language Model EvaluationDCLM Core
DCLM Core Score49.3
12
Language ModelingDCLM
Loss1.863
11
Zero-shot EvaluationDCLM CORE
CORE Score0.268
9
Natural Language UnderstandingDCLM
DCLM Core Score48.9
9
Language ModelingDCLM benchmark
Macro Avg Value14.58
9
General Language CapabilityDCLM CORE v2 (test)
Commonsense Score61.9
7
Language UnderstandingDCLM evaluation suite (test)
HellaSwag Accuracy62.9
7
Aggregate EvaluationDCLM CORE (val)
CORE27.2
5
Scaling Law FittingDCLM strongly regularized baseline 20 points (Full fit)
RMSE0.008
4
Zero-shot Language EvaluationDCLM Pro
WinoGrande57.93
2
Showing 12 of 12 rows