Loop-aligned supervision lets a looped Transformer generate CoT chains beyond training length, and those chains improve an auto-regressive CoT model's length generalization.
This task identifies the length of longest strictly increasing subsequence in a numerical sequence
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CL 1years
2025 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Enhancing Auto-regressive Chain-of-Thought through Loop-Aligned Reasoning
Loop-aligned supervision lets a looped Transformer generate CoT chains beyond training length, and those chains improve an auto-regressive CoT model's length generalization.