Share your thoughts, 1 month free Claude Pro on us
See more
Home
/
Benchmarks
Code Summarization on GiAnt Corpus Risk Analysis
Loading...
0.737
BERTScore
GPT-4o
0.73284
0.73392
0.735
0.73608
Jun 5, 2026
BERTScore
LLM Judge Rating
Updated 1mo ago
Evaluation Results
Method
Method
Links
BERTScore
LLM Judge Rating
GPT-4o
2026.06
0.737
2.97
DeepSeek-V3
2026.06
0.735
2.89
GPT-3.5-turbo
2026.06
0.734
2.64
Qwen3-Coder-Instruct
2026.06
0.733
2.86
Feedback
Search any
task
Search any
task