ST-Policy: Encoding Spatiotemporal Continuity for Diffusion Policy
Diffusion-based visuomotor policies are a powerful paradigm for robotic imitation learning, yet their sequence models apply a uniform inductive bias at every position of the action sequence, ignoring the spatiotemporal continuity inherent in robotic motion. We identify a fundamental limitation of causal state space mod...