GreaseLM: Graph REASoning Enhanced Language Models for Question Answering

About

Answering complex questions about textual narratives requires reasoning over both stated context and the world knowledge that underlies it. However, pretrained language models (LM), the foundation of most modern QA systems, do not robustly represent latent relationships between concepts, which is necessary for reasoning. While knowledge graphs (KG) are often used to augment LMs with structured representations of world knowledge, it remains an open question how to effectively fuse and reason over the KG representations and the language context, which provides situational constraints and nuances. In this work, we propose GreaseLM, a new model that fuses encoded representations from pretrained LMs and graph neural networks over multiple layers of modality interaction operations. Information from both modalities propagates to the other, allowing language context representations to be grounded by structured world knowledge, and allowing linguistic nuances (e.g., negation, hedging) in the context to inform the graph representations of knowledge. Our results on three benchmarks in the commonsense reasoning (i.e., CommonsenseQA, OpenbookQA) and medical question answering (i.e., MedQA-USMLE) domains demonstrate that GreaseLM can more reliably answer questions that require reasoning over both situational constraints and structured knowledge, even outperforming models 8x larger.

Xikun Zhang, Antoine Bosselut, Michihiro Yasunaga, Hongyu Ren, Percy Liang, Christopher D. Manning, Jure Leskovec• 2022

Related benchmarks

Task	Dataset	Result
Commonsense Reasoning	HellaSwag	Accuracy82.8	1896
Commonsense Reasoning	PIQA	Accuracy79.6	757
Commonsense Reasoning	CSQA	Accuracy74.2	366
Commonsense Reasoning	ARC Challenge	Accuracy44.7	243
Commonsense Reasoning	OBQA	Accuracy66.9	187
Question Answering	PubMedQA (test)	Accuracy72.4	170
Question Answering	OpenBookQA (OBQA) (test)	OBQA Accuracy84.8	130
Question Answering	MedQA-USMLE (test)	Accuracy45.1	101
Commonsense Question Answering	CosmosQA	Accuracy80.6	68
Question Answering	MedQA (test)	Accuracy38.5	67

Showing 10 of 20 rows

Other info

Follow for update

@wizwand_team Discord