Share your thoughts, 1 month free Claude Pro on usSee more
WorkDL logo mark

Multi-Agent Reinforcement Learning on MAMuJoCo Walker2d 6x1 (test)

1,116.75Average Episodic Return

Sim2O

-37.8476261.9037561.655861.4063Feb 13, 2026Mar 6, 2026Mar 27, 2026Apr 17, 2026May 8, 2026May 29, 2026Jun 19, 2026
Updated 1mo ago

Evaluation Results

MethodLinks
2026.06
1,116.75
2026.06
1,062.81
2026.06
1,017.93
2026.06
884.7
2026.06
872.62
2026.02
28.56
2026.02
23.33
2026.02
18.72
2026.02
18.61
2026.02
12.16
2026.02
10.86
2026.02
7.4
2026.02
6.56