Share your thoughts, 1 month free Claude Pro on usSee more
WorkDL logo mark

SoftCoT: Soft Chain-of-Thought for Efficient Reasoning with LLMs

About

Chain-of-Thought (CoT) reasoning enables Large Language Models (LLMs) to solve complex reasoning tasks by generating intermediate reasoning steps. However, most existing approaches focus on hard token decoding, which constrains reasoning within the discrete vocabulary space and may not always be optimal. While recent efforts explore continuous-space reasoning, they often require full-model fine-tuning and suffer from catastrophic forgetting, limiting their applicability to state-of-the-art LLMs that already perform well in zero-shot settings with a proper instruction. To address this challenge, we propose a novel approach for continuous-space reasoning that does not require modifying the LLM. Specifically, we employ a lightweight fixed assistant model to speculatively generate instance-specific soft thought tokens as the initial chain of thoughts, which are then mapped into the LLM's representation space via a trainable projection module. Experimental results on five reasoning benchmarks demonstrate that our method enhances LLM reasoning performance through supervised, parameter-efficient fine-tuning. Source code is available at https://github.com/xuyige/SoftCoT.

Yige Xu, Xu Guo, Zhiwei Zeng, Chunyan Miao• 2025

Related benchmarks

TaskDatasetResultRank
Code GenerationHumanEval (test)
Pass@171.83
701
Mathematical ReasoningMATH 500
Accuracy (Acc)82.8
600
Multimodal UnderstandingMMStar--
511
Diagram Question AnsweringAI2D
AI2D Accuracy80.4
509
Code GenerationMBPP (test)
Pass@156.04
411
Mathematical ReasoningAIME 24
Accuracy16.67
358
Science Question AnsweringScienceQA (SQA)
Accuracy89.5
338
Mathematical ReasoningSVAMP (test)
Accuracy40
298
Commonsense ReasoningStrategyQA
Accuracy71.18
208
Mathematical ReasoningAQUA
Accuracy80.63
167
Showing 10 of 51 rows

Other info

Code

Follow for update