Skip to content

Author

Eric Elmoznino

1 paper indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Amortizing intractable inference in large language models

This work interprets chain-of-thought reasoning as a latent variable modeling problem and demonstrates that this distribution-matching paradigm of LLM fine-tuning can serve as an effective alternative to maximum-likelihood training and reward-maximizing policy optimization.

Edward J. Hu, Moksh Jain, Eric Elmoznino et al. · 110 citations · ⚡19

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.