From Chain-of-Thought to Loops: Non-Autoregressive Latent Reasoning via Looped Transformers
Chain-of-thought (CoT) reasoning often improves language-model performance by giving models additional computation before answering. However, explicit CoT expresses this computation as a sequence of autoregressively generated tokens. Latent reasoning replaces these tokens with compact continuous states, but most autore...