GrabVG: Graph-Attentive Binding for Visual Grounding in UAV Imagery
This work proposes GrabVG, a novel visual grounding framework inspired by human visual search that generates a compact set of reliable object hypotheses through distillation-guided proposal induction and text-aware hypothesis filtering, substantially reducing background distractions and semantic mismatches.
Chaowei Wang, Yan Di, Jingjun Sun et al.
· 0 citations