Share your thoughts, 1 month free Claude Pro on usSee more
WorkDL logo mark

WG-M

Benchmarks

Task NameDataset NameSOTA ResultTrend
Commonsense ReasoningWG-M
Accuracy76.45
18
Common-sense reasoningWG-M (In-Distribution)
Accuracy83.61
12
Commonsense ReasoningWG-M (test)
Accuracy83.23
10
Showing 3 of 3 rows