Share your thoughts, 1 month free Claude Pro on us
See more
Home
/
Benchmarks
Safety Content Detection on Standard non-adversarial dataset Phishing
Loading...
98
TPR
GPT-4
90.72
92.61
94.5
96.39
Jan 27, 2026
TPR
FPR
Updated 5mo ago
Evaluation Results
Method
Method
Links
TPR
FPR
GPT-4
Rule Definitions=Included
2026.01
98
0
GAVEL
2026.01
95
0
GPT-4
Rule Definitions=None
2026.01
91
0
Feedback
Search any
task
Search any
task