Preprint
Sep 2026
Towards High-DoF Dexterous Manipulation through VLA Post-Training
A unified four-step post-training pipeline comprising a learned temporal hand-action codec, supervised fine-tuning, DAgger, and real-world residual reinforcement learning provides a practical path for adapting VLA foundation models to reliable real-world dexterous manipulation.
Jun-Lei Zhu, Shen-Zhe Yao, Chao-Gui Huang et al.
· 0 citations