A Model with No Head and Many Thoughts
This work introduces Soft Latent Thinking, a method that replaces the LM head during reasoning with a lightweight projector, enabling autoregressive rollout in embedding space where reasoning steps remain continuous rather than tokenized.
N. Koriagin, Yaroslav Aksenov, George Bredis et al.
· 0 citations