Share your thoughts, 1 month free Claude Pro on us
See more
Home
/
Benchmarks
Mitigation Recommendation on GiAnt Corpus (test)
Loading...
78.2
BERTScore
DeepSeek-V3
76.848
77.199
77.55
77.901
Jun 5, 2026
BERTScore
LLM Judge Rating
Updated 1mo ago
Evaluation Results
Method
Method
Links
BERTScore
LLM Judge Rating
DeepSeek-V3
2026.06
78.2
4.09
Qwen2.5-Coder-Instruct
2026.06
78
4.01
GPT-4o
2026.06
77.8
4.03
GPT-3.5-turbo
2026.06
76.9
3.64
Feedback
Search any
task
Search any
task