Share your thoughts, 1 month free Claude Pro on us
See more
Home
/
Benchmarks
Root Cause Analysis on AIOps Hard 2025
Loading...
27.4
Accuracy
Qwen3-Next-80b-a3b-Thinking
-1.096
6.302
13.7
21.098
Apr 6, 2026
Accuracy
Updated 28d ago
Evaluation Results
Method
Method
Links
Accuracy
Qwen3-Next-80b-a3b-Thinking
LLM Category=Open-sour...
2026.04
27.4
Qwen-Plus-2025-09-11
LLM Category=Closed-so...
2026.04
24.7
Qwen3-Max-2025-09-23
LLM Category=Closed-so...
2026.04
22.1
GPT-5.2
LLM Category=Closed-so...
2026.04
21.6
OpsLLM-32B
LLM Category=OpsLLM
2026.04
21.1
Deepseek-v3.2-exp
LLM Category=Open-sour...
2026.04
19.2
OpsLLM-14B
LLM Category=OpsLLM
2026.04
13.6
Qwen-Turbo-2025-07-15
LLM Category=Closed-so...
2026.04
13.5
Zhiyu-32B
LLM Category=Open-sour...
2026.04
9.1
OpsLLM-7B
LLM Category=OpsLLM
2026.04
8.8
Moonshot-Kimi-K2-Instruct
LLM Category=Open-sour...
2026.04
7.7
R1-Distill-SRE-Qwen-32B-INT8
LLM Category=Open-sour...
2026.04
6.3
Qwen2.5-32B-Instruct
LLM Category=Base LLM
2026.04
6
Qwen2.5-14B-Instruct
LLM Category=Base LLM
2026.04
4.6
Qwen2.5-7B-Instruct
LLM Category=Base LLM
2026.04
4.3
aiops-qwen-4b
LLM Category=Open-sour...
2026.04
1.6
R1-Distill-SRE-Qwen-7B
LLM Category=Open-sour...
2026.04
0
Feedback
Search any
task
Search any
task