ReFT: Representation Finetuning for Language Models

About

Parameter-efficient finetuning (PEFT) methods seek to adapt large neural models via updates to a small number of weights. However, much prior interpretability work has shown that representations encode rich semantic information, suggesting that editing representations might be a more powerful alternative. We pursue this hypothesis by developing a family of Representation Finetuning (ReFT) methods. ReFT methods operate on a frozen base model and learn task-specific interventions on hidden representations. We define a strong instance of the ReFT family, Low-rank Linear Subspace ReFT (LoReFT), and we identify an ablation of this method that trades some performance for increased efficiency. Both are drop-in replacements for existing PEFTs and learn interventions that are 15x--65x more parameter-efficient than LoRA. We showcase LoReFT on eight commonsense reasoning tasks, four arithmetic reasoning tasks, instruction-tuning, and GLUE. In all these evaluations, our ReFTs deliver the best balance of efficiency and performance, and almost always outperform state-of-the-art PEFTs. We release a generic ReFT training library publicly at https://github.com/stanfordnlp/pyreft.

Zhengxuan Wu, Aryaman Arora, Zheng Wang, Atticus Geiger, Dan Jurafsky, Christopher D. Manning, Christopher Potts• 2024

Related benchmarks

Task	Dataset	Result
Commonsense Reasoning	HellaSwag	Accuracy96.31	1896
Commonsense Reasoning	WinoGrande	--	1442
Mathematical Reasoning	GSM8K (test)	Accuracy64.7	816
Instruction Following	AlpacaEval 2.0	Win Rate61.68	722
Code Generation	HumanEval (test)	Pass@168.66	612
Natural Language Understanding	GLUE (dev)	SST-2 (Acc)96.2	529
Mathematical Reasoning	GSM8K	Accuracy40.1	499
Code Generation	MBPP (test)	Pass@154.4	405
Commonsense Reasoning	Common Sense Reasoning Tasks	Avg Score83.3	321
Reading Comprehension	RACE high	Accuracy85.33	295

Showing 10 of 59 rows

Other info

Follow for update

@wizwand_team Discord