Share your thoughts, 1 month free Claude Pro on usSee more
WorkDL logo mark

Fixed-Point Reasoners: Stable and Adaptive Deep Looped Transformers

About

Looped architectures provide an inductive bias toward learning step-by-step procedures for tasks that require compositional reasoning. The number of effective layers reached by looping determines the quality of the solution these models find. Like deep architectures, looped architectures are prone to a signal propagation problem induced by depth as the halting decision is postponed. In this paper, we address this signal propagation issue using pre-norm layers and residual scaling. Building on these architectural modifications, we propose FPRM, a Transformer-based Fixed-Point Reasoning Model that uses fixed-point convergence as an end-to-end halting mechanism in a looped architecture. We show that fixed-point halting allows FPRM to adapt its compute to task difficulty. FPRM is effective on common reasoning benchmarks, namely Sudoku, Maze, state-tracking, and ARC-AGI.

Sajad Movahedi, Vera Milovanovi\'c, Shlomo Libo Feigin, Alexander Theus, Thomas Hofmann, Valentina Boeva, T. Konstantin Rusch, Antonio Orvieto• 2026

Related benchmarks

TaskDatasetResultRank
Puzzle SolvingSudoku-Extreme (test)
Pass@1 Success Rate94.2
9
Puzzle SolvingMaze-hard (test)
Pass@187
6
Abstraction and ReasoningARC-1 (test)
Pass@247.5
4
Abstraction and ReasoningARC 2 (test)
Pass@26.2
4
Showing 4 of 4 rows

Other info

Follow for update