Share your thoughts, 1 month free Claude Pro on us
See more
Home
/
Benchmarks
Travelled distance estimation on Vessel Trajectory Descriptions
Loading...
0.89
Mean Error (%)
openai/gpt-oss-120b
0.374
3.857
7.34
10.823
Mar 8, 2026
Mean Error (%)
Stdev Error (%)
Max Error (%)
Updated 4mo ago
Evaluation Results
Method
Method
Links
Mean Error (%)
Stdev Error (%)
Max Error (%)
openai/gpt-oss-120b
LLM Model=openai/gpt-o...
2026.03
0.89
4.12
35.32
openai/gpt-oss-20b
LLM Model=openai/gpt-o...
2026.03
1.06
4.29
35.32
llama-3.3-70b-versatile
LLM Model=llama-3.3-70...
2026.03
2.44
4.52
34.55
qwen/qwen3-32b
LLM Model=qwen/qwen3-32b
2026.03
4.92
43.48
848.22
llama-3.1-8b-instant
LLM Model=llama-3.1-8b...
2026.03
13.79
26.36
248.28
Feedback
Search any
task
Search any
task