Share your thoughts, 1 month free Claude Pro on us
See more
Home
/
Benchmarks
In-context adaptation on Hangman
Loading...
6.48
Cumulative Reward
Cross-task RL
3.3496
4.1623
4.975
5.7877
Jun 13, 2026
Cumulative Reward
Gain
Final Reward
Updated 1mo ago
Evaluation Results
Method
Method
Links
Cumulative Reward
Gain
Final Reward
Cross-task RL
Training Strategy=Cros...
2026.06
6.48
5
0.65
Single-task RL
Training Strategy=Sing...
2026.06
5.53
5
0.55
Base (Qwen3-8B)
Training Strategy=None...
2026.06
3.47
3
0.33
Feedback
Search any
task
Search any
task