ActFovea: Runtime Safeguarding for VLA Policies via Spatiotemporal Visual-Action Consistency
ActFovea is introduced, a plug-and-play safeguarding framework that detects and mitigates runtime failures of vision-language-action policies without retraining or modifying the underlying VLA policy.