Counterfactual Graph for Multi-Agent LLM Calibration

About

Multi-agent LLM systems often treat agreement as evidence: when many agents in a panel give the same answer, that answer is assumed to be more reliable. We show that this assumption can fail after agents communicate. Communication can induce correlated failures and false consensus, so the same vote share may reflect reliable agreement in one topology but over-confidence in another. We propose CAGE-CAL, a counterfactual agent-graph calibration framework for multi-agent LLMs. For each query, CAGE-CAL compares an observed post-communication agent graph with a matched counterfactual no-communication graph, capturing both pairwise failure correlations and group-level dependencies. Rather than simply counting how many agents agree, CAGE-CAL estimates the counterfactual shift between observed and no-communication dependence, and calibrates confidence accordingly. Across five benchmarks, CAGE-CAL improves reliability discrimination with competitive ECE, and its calibrated confidence further improves topology selection over the best fixed-topology strategy.

Jiatan Huang, Mingchen Li, Ziming Li, Sunjae Kwon, Hong Yu, Chuxu Zhang• 2026

Related benchmarks

Task	Dataset	Result
Uncertainty Estimation	TriviaQA (test)	AUROC86.12	110
Question Answering	TriviaQA	BS (%)9.55	65
Calibration	TriviaQA	--	39
Calibration	TruthfulQA	--	32
Uncertainty Quantification	MMLU Pro (test)	AUROC77.74	24
Calibration	GSM8K	ECE1.64	11
Calibration	BBH	ECE6.12	11
Calibration	Mean macro-average across benchmarks	Expected Calibration Error (ECE)5.56	11
Language Understanding	MMLU-Pro	Brier Score19.26	11
Math Reasoning	GSM8K	Brier Score4.33	11

Showing 10 of 21 rows

Other info

Follow for update

@wizwand_team Discord