A Mechanistic Analysis of Looped Reasoning Language Models

· 2026 · cs.LG · arXiv 2604.11791

7 Pith papers cite this work. Polarity classification is still indexing.

7 Pith papers citing it

open full Pith review browse 7 citing papers arXiv PDF

abstract

Reasoning has become a central capability in large language models. Recent research has shown that reasoning performance can be improved by looping an LLM's layers in the latent dimension, resulting in looped reasoning language models. Despite promising results, few works have investigated how their internal dynamics differ from those of standard feedforward models. In this paper, we conduct a mechanistic analysis of the latent states in looped language models, focusing in particular on how the stages of inference observed in feedforward models compare to those observed in looped ones. To this end, we analyze cyclic recurrence and show that for many of the studied models each layer in the cycle converges to a distinct fixed point; consequently, the recurrent block follows a consistent cyclic trajectory in the latent space. We provide evidence that as these fixed points are reached, attention-head behavior stabilizes, leading to constant behavior across recurrences. Empirically, we discover that recurrent blocks learn stages of inference that closely mirror those of feedforward models, repeating these stages in depth with each iteration. We study how recurrent block size, input injection, and normalization influence the emergence and stability of these cyclic fixed points. We believe these findings help translate mechanistic insights into practical guidance for architectural design.

citation-role summary

background 1 baseline 1

citation-polarity summary

background 1 baseline 1

representative citing papers

Looped SSMs: Depth-Recurrence and Input Reshaping for Time Series Classification

cs.LG · 2026-05-15 · unverdicted · novelty 7.0

Looped SSMs with shared parameters across depth match or exceed standard SSMs with more parameters on time series classification, with additional gains from input reshaping techniques.

Fixed-Point Reasoners: Stable and Adaptive Deep Looped Transformers

cs.AI · 2026-06-16 · unverdicted · novelty 6.0

FPRM is a Transformer-based model using fixed-point convergence for adaptive halting in looped architectures, claimed effective on Sudoku, Maze, state-tracking, and ARC-AGI benchmarks.

Memory-Efficient Looped Transformer: Decoupling Compute from Memory in Looped Language Models

cs.CL · 2026-05-08 · unverdicted · novelty 6.0 · 2 refs

MELT decouples reasoning depth from memory in looped language models by sharing a single gated KV cache per layer and training it via chunk-wise distillation from Ouro starting models.

Chain-of-Thought and Compressed Looped Transformers: A Memory-Budget Separation

cs.LG · 2026-05-29 · unverdicted · novelty 5.0

Compressed looped Transformers with fixed small recurrent state cannot decide P-complete problems under logspace reductions, while polynomial-length chain-of-thought can.

Hyperloop Transformers

cs.LG · 2026-04-23 · unverdicted · novelty 5.0

Hyperloop Transformers outperform standard and mHC Transformers with roughly 50% fewer parameters by looping a middle block of layers and applying hyper-connections only after each loop.

Simply Stabilizing the Loop via Fully Looped Transformer

cs.LG · 2026-05-11

SMolLM: Small Language Models Learn Small Molecular Grammar

cs.LG · 2026-05-07

citing papers explorer

Showing 1 of 1 citing paper after filters.

Fixed-Point Reasoners: Stable and Adaptive Deep Looped Transformers cs.AI · 2026-06-16 · unverdicted · none · ref 75 · internal anchor
FPRM is a Transformer-based model using fixed-point convergence for adaptive halting in looped architectures, claimed effective on Sudoku, Maze, state-tracking, and ARC-AGI benchmarks.

A Mechanistic Analysis of Looped Reasoning Language Models

citation-role summary

citation-polarity summary

fields

years

verdicts

roles

polarities

representative citing papers

citing papers explorer