Recurrent Looped Transformer
Recurrent Looped Transformer (RLT)passes the decoder's final hidden state to the next token, together with that token's causal encoder representation. The decoder reads encoder-derived global KV memory and maintains a sliding-window attention (SWA) cache at every layer. The same update runs over prompt and response tokens.
More effective reasoning depth!
https://github.com/yifanzhang-pro/recurrent-looped-tranformer
scorecomments26 sightings
first seen 2026-10-03 04:50 UTClast seen 2026-10-09 01:28 UTCscore then 57score now 54gained -3sightings 26
open on reddit ↗
💬 26