Share your thoughts, 1 month free Claude Pro on usSee more
WorkDL logo mark

Logic-LM: Empowering Large Language Models with Symbolic Solvers for Faithful Logical Reasoning

About

Large Language Models (LLMs) have shown human-like reasoning abilities but still struggle with complex logical problems. This paper introduces a novel framework, Logic-LM, which integrates LLMs with symbolic solvers to improve logical problem-solving. Our method first utilizes LLMs to translate a natural language problem into a symbolic formulation. Afterward, a deterministic symbolic solver performs inference on the formulated problem. We also introduce a self-refinement module, which utilizes the symbolic solver's error messages to revise symbolic formalizations. We demonstrate Logic-LM's effectiveness on five logical reasoning datasets: ProofWriter, PrOntoQA, FOLIO, LogicalDeduction, and AR-LSAT. On average, Logic-LM achieves a significant performance boost of 39.2% over using LLM alone with standard prompting and 18.4% over LLM with chain-of-thought prompting. Our findings suggest that Logic-LM, by combining LLMs with symbolic logic, offers a promising avenue for faithful logical reasoning. Code and data are publicly available at https://github.com/teacherpeterpan/Logic-LLM.

Liangming Pan, Alon Albalak, Xinyi Wang, William Yang Wang• 2023

Related benchmarks

TaskDatasetResultRank
Logical reasoningFOLIO
Accuracy71.6
126
Logical reasoningFOLIO (test)
Accuracy73.83
60
Logical reasoningAR-LSAT
Accuracy43.04
60
Logical reasoningProofWriter (test)
Accuracy84.35
57
Logical reasoningProntoQA (test)
Accuracy90.5
57
Logical reasoningProofWriter
Accuracy64.7
44
Policy SelectionMeasles
Precision@372.2
24
Policy SelectionSupply Chain
Precision@387
24
Simulator selectionCOVID-19
Top-1 Regret1.78
24
Simulator selectionSupply Chain
Top-1 Regret0.54
24
Showing 10 of 34 rows

Other info

Follow for update