Learning from Mixed-Quality Deployment Experience for Robot Manipulation
Predictive Action Chunk Learning first learns a predictive chunk-level critic that evaluates temporally extended action sequences and augments temporal difference learning with future latent prediction, providing richer supervision for long-horizon value estimation.