Share your thoughts, 1 month free Claude Pro on us
See more
Home
/
Benchmarks
In-context adaptation on Wordle
Loading...
6.23
Cumulative Reward
Cross-task RL
2.3924
3.3887
4.385
5.3813
Jun 13, 2026
Cumulative Reward
Gain (%)
Final Reward
Updated 1mo ago
Evaluation Results
Method
Method
Links
Cumulative Reward
Gain (%)
Final Reward
Cross-task RL
Training Strategy=Cros...
2026.06
6.23
11
0.61
Single-task RL
Training Strategy=Sing...
2026.06
3.26
61
0.24
Base (Qwen3-8B)
Training Strategy=None...
2026.06
2.54
8
0.26
Feedback
Search any
task
Search any
task