Share your thoughts, 1 month free Claude Pro on us
See more
Home
/
Benchmarks
Persuasion prediction on IBM Argument Quality (test)
Loading...
59
F1
PI-sub
21.04
30.895
40.75
50.605
Jun 12, 2026
F1
Updated 1mo ago
Evaluation Results
Method
Method
Links
F1
PI-sub
Feature granularity=55...
2026.06
59
GPT-4o
Evaluation protocol=ze...
2026.06
58.6
PI-mean
Feature granularity=15...
2026.06
57.9
RoBERTa
Evaluation protocol=fi...
2026.06
22.5
Feedback
Search any
task
Search any
task