GeoSeg-OV: Bridging Geospatial Gaps with Structural Guidance for Open-Vocabulary Remote Sensing Segmentation
GeoSeg-OV is proposed, which decouples auxiliary VFM features from visual-text matching and repurposes them as structural guidance for cost aggregation and decoding, and introduces Cost-Aware Decoding (CAD) to adaptively refine and fuse multi-scale semantic and structural guidance based on the current decoder context.