Preprint
Aug 2026
Simple Actors and Deep Critics for Scalable Reinforcement Learning
This work revisits where capacity should be invested in an offline actor--critic method and proposes LAC (Light Actor, deep Critic), a lightweight deterministic actor that matches the strongest diffusion- and flow-matching baselines while achieving up to 4x lower inference latency.
Gu-Heon Kang, Jaehwi Lee, Minhae Kwon
· 0 citations