Jul 2026
SeededGrasp: Language-Guided Grasping in Complex Scenes with Multiple Embodiments
This work proposes SeededGrasp, a novel data-efficient framework that enables a VLM to predict a seed point to be used as conditioning for a subsequent lightweight grasp-generation model, enabling multi-embodiment support while bypassing the need for expensive end-to-end training.
Yang Xu, Gurpreet Singh Mukker, Raymond Wang et al.
· arXiv.org · 0 citations