Skip to content

Author

Ruijiao Li

We have 2 of 17 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Open access 2026

A High-Efficiency Diffusion Model-Inspired Network for Pixel-Level Self-Supervised Hyperspectral Anomaly Change Detection

Hyperspectral image change detection (CD) has garnered significant attention in the field of remote sensing. A critical task within CD is anomaly CD. Current generative anomaly CD methods, such as diffusion models, typically rely on computationally expensive iterative sampling to extract features, severely limiting their real-time application capabilities. Furthermore, in the absence of labeled data, existing unsupervised algorithms struggle to effectively distinguish subtle target variations from background artifacts caused by shadows or registration errors. To address these challenges, we propose a novel hyperspectral anomaly CD method named efficient latent denoising-inspired network (ELDI-Net). It employs a one-step manifold projection paradigm, achieving high computational efficiency while preserving the noise-robustness advantages of generative models. Specifically, we introduce a one-step latent manifold projection framework that transforms traditional iterative denoising into a deterministic latent mapping via an encoder–projector–decoder architecture, achieving a substantial improvement in inference speed. In addition, a spectral adaptive calibration projection module is constructed, employing channel-adaptive calibration to suppress spectral redundancy while effectively preserving critical features of subtle targets. A bidirectional focus alignment mechanism is designed for implicit semantic denoising under self-supervised conditions, suppressing pseudovariation artifacts through twin cross-prediction. Finally, a GCI strategy is introduced to eliminate directional sensor noise. Experimental results on three datasets demonstrate that the proposed ELDI-Net method achieves superior or highly competitive performance compared to multiple state-of-the-art approaches.

Xing Hu, Xiang-Cheng Liu, Chen-Xi Guo et al. · 0 citations
Sep 2026

MOP: A Multimodal Object-Aware Policy for Robotic Manipulation via Geometry-Guided Fusion and Trajectory Prediction

3D imitation learning has demonstrated capability in diverse visuomotor tasks but often struggles with small objects or high-precision manipulation due to geometric sparsity. To address this, we propose the Multimodal Object-aware Policy (MOP), a novel framework that incorporates a geometry-guided fusion module to adaptively integrate 2D semantic features with 3D geometry for precise control. Additionally, we introduce a lightweight Object Position Prediction (OPP) module that serves as an auxiliary supervision signal. The training labels for this module are generated by using a cost-effective vision-based software tracker, effectively replacing expensive hardware-based motion capture systems. We evaluate our policy across 56 tasks on 3 simulation benchmarks. Experimental results demonstrate that MOP significantly outperforms baselines, achieving higher success rates with lower variance and efficient inference. Real-world experiments on 4 tasks further validate the robustness and transferability of our approach.

Keyu Zhang, Yicheng Yang, Lifeng Wang et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.