Our new X account is live! Follow @wizwand_team for updates
Home
/
Benchmarks
Few-shot example selection on Task #8
Loading...
0.42
Score
TRIPLE-CSAR
0.3472
0.3661
0.385
0.4039
Feb 15, 2024
Score
Updated 4d ago
Evaluation Results
Method
Method
Links
Score
TRIPLE-CSAR
LLM=GPT-3.5
2024.02
0.42
Uniform
LLM=GPT-3.5
2024.02
0.4
TRIPLE-SAR
LLM=GPT-3.5
2024.02
0.38
Random
LLM=GPT-3.5
2024.02
0.35
Feedback
Search any
task
Search any
task