CONSONANCE.for your information
Monday, 5 October 2026frenvi

Worth reading closely

01 — inference 42 upvotes

Decoding Looped Transformers Better for (Almost) Free

QUESTION — How can intermediate states from recurrent loops in Transformers be leveraged to improve decoding quality without additional training?

The paper introduces LoopCD, a training-free contrastive decoding framework that leverages intermediate states from recurrent loops in Looped Transformers. By contrasting the final prediction with an earlier recurrent pass—either in logit space or hidden-state space—LoopCD guides token selection effectively. This approach delivers substantial decoding performance gains while enabling a reduction of recurrent loops by half, thereby significantly cutting forward FLOPs during inference.

LoopCD-Logits raises Ouro-2.6B-Thinking's AIME 2024 pass@1 from 61.88% to 73.33%.

LoopCD-Hidden lifts Huginn's HumanEval pass@1 from 22.56% to 31.71%.

The approach enables halving the number of recurrent loops, reducing forward FLOPs by 22.5% to 48.2%.

neosknight · 01 Oct 2026 read the original ↗
↑