Long-horizon Earth observation reasoning requires models to organize multi-stage geographic evolution, localize spatial changes, detect temporal anomalies, and infer future from extended image sequences. However, existing remote sensing vision-language models mainly focus on isolated images, image pairs, or short seque...
Yupan Ding, Jing Xiao, Zhenyuan Zhang et al.· 0 citations
A radiation, rotation, and scale invariant (RRSI) feature descriptor that enables feature encoding, interaction, and fusion across intra-modal, dual-head sampled, and inter-modal regions, and introduces a bidirectional cross-modal generative reconstruction constraint during training.
Yuan-Xin Ye, Teng-Feng Tang, Tao Peng et al.· 0 citations
EO-VGGT is presented, a framework that adapts a frozen perspective-driven model to orbital observations via explicit physical geometry embedding via explicit physical geometry embedding for robust feed-forward satellite 3D reconstruction.
Qiyan Luo, Y. Pi, Lekang Wen et al.· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.