Share your thoughts, 1 month free Claude Pro on us
See more
Home
/
Benchmarks
Code Summarization on GiAnt Corpus
Loading...
0.758
BERTScore
GPT-4o
0.73928
0.74414
0.749
0.75386
Jun 5, 2026
BERTScore
LLM-Judge Rating
Updated 1mo ago
Evaluation Results
Method
Method
Links
BERTScore
LLM-Judge Rating
GPT-4o
2026.06
0.758
2.51
Qwen3-Coder-Instruct
2026.06
0.757
2.48
DeepSeek-V3
2026.06
0.747
2.4
GPT-3.5-turbo
2026.06
0.74
1.97
Feedback
Search any
task
Search any
task