Skip to content

Author

Xue-Yang Guo

2 papers indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Conference Aug 2026

Stage-Aware Vision-Language-Action Model for Long-Horizon Robotic Manipulation

Vision-Language-Action (VLA) models learn generalist robot manipulation policies by mapping language instructions and visual observations to continuous actions through imitation learning. However, their performance degrades on long-horizon tasks, particularly when sub-tasks admit multiple valid execution orders. Since...

Hao-Ran Shi, Yi-Han Zhou, Ming-Cong Li et al. · 0 citations
Preprint Sep 2026

Gaze Prompts: Temporally Dense Human Attention for Vision-Language-Action Fine-Tuning

Vision-Language-Action (VLA) fine-tuning pairs images with actions at every step, yet typically provides only a task-level language instruction, leaving moment-to-moment visual relevance implicit. We introduce \emph{eye-tracker-supervised gaze prompting}, which uses gaze recorded during VR teleoperation to provide fram...

Yi-Han Zhou, Rui Yan, Ming-Cong Li et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.