MedAgents: Large Language Models as Collaborators for Zero-shot Medical Reasoning

About

Large language models (LLMs), despite their remarkable progress across various general domains, encounter significant barriers in medicine and healthcare. This field faces unique challenges such as domain-specific terminologies and reasoning over specialized knowledge. To address these issues, we propose MedAgents, a novel multi-disciplinary collaboration framework for the medical domain. MedAgents leverages LLM-based agents in a role-playing setting that participate in a collaborative multi-round discussion, thereby enhancing LLM proficiency and reasoning capabilities. This training-free framework encompasses five critical steps: gathering domain experts, proposing individual analyses, summarising these analyses into a report, iterating over discussions until a consensus is reached, and ultimately making a decision. Our work focuses on the zero-shot setting, which is applicable in real-world scenarios. Experimental results on nine datasets (MedQA, MedMCQA, PubMedQA, and six subtasks from MMLU) establish that our proposed MedAgents framework excels at mining and harnessing the medical expertise within LLMs, as well as extending its reasoning abilities. Our code can be found at https://github.com/gersteinlab/MedAgents.

Xiangru Tang, Anni Zou, Zhuosheng Zhang, Ziming Li, Yilun Zhao, Xingyao Zhang, Arman Cohan, Mark Gerstein• 2023

Related benchmarks

Task	Dataset	Result
Medical Question Answering	MedMCQA	Accuracy74.8	521
Question Answering	PubMedQA	Accuracy76.8	145
Medical Question Answering	MedMCQA (test)	Accuracy74.8	134
Medical Question Answering	MedQA	Accuracy47.3	124
Medical Question Answering	PubMedQA	Accuracy56.1	117
Medical Visual Question Answering	PMC-VQA	Accuracy56.5	103
Question Answering	MedQA	Accuracy83.7	96
Medical Visual Question Answering	PathVQA	Accuracy66.7	80
Question Answering	MedQA (test)	Accuracy83.7	67
Medical Question Answering	PubMedQA	Accuracy76.8	65

Showing 10 of 78 rows

...

Other info

Code

Follow for update

@wizwand_team Discord