Skip to content

OASIS-Map: Object-Level Change Detection in Multi-Session Mapping using Semantic Correspondence Matching

Jul 2026 · arXiv.org · Vol abs/2607.14899 · 0 citations · 39 references
Computer Science

TL;DR

This work proposes OASIS-Map, a multi-session mapping system that maintains a spatio-temporally consistent object-level map by establishing dense patch-level semantic correspondences between temporal observations that detect where the scene has changed and incrementally associate objects across revisits as the robot re-observes the environment.

Abstract

Map representations which are consistent across repeated visits to a real-world semi-static environment are very useful for long-term robotic inspection. In such settings, the scene may evolve while the robot is absent, with objects appearing, disappearing, moving, or being replaced, quickly making a static map outdated. Existing change-detection methods reason through geometry, category-level semantics, or object persistence. However, achieving reliable object association across revisits remains a key challenge, especially under partial views, occlusion, and imperfect segmentation. In this work, we propose OASIS-Map, a multi-session mapping system that maintains a spatio-temporally consistent object-level map by establishing dense patch-level semantic correspondences between temporal observations. These correspondences detect where the scene has changed and incrementally associate objects across revisits as the robot re-observes the environment. We demonstrate OASIS-Map on three challenging real-world scenarios: object rearrangements in 3RScan, visually similar car replacements in a car park, and large-scale scene changes in an outdoor market. We achieve 0.783 F1 on change detection in a car replacement scenario in a car park and 0.667 F1 on moved object association in 3RScan. https://dynamic.robots.ox.ac.uk/projects/oasis-map/

View source

Similar papers

Open access Jul 2026

SuperMap: A Spatio-Temporal SLAM System for Visual-Language Navigation

This work presents SuperMap, a 4D spatio-temporal mapping framework for language-guided navigation that integrates high-frequency geometric SLAM with asynchronous open-vocabulary perception and releases the full system as open-source to provide the community with a deployable baseline for open-vocabulary spatio-tempora...

Shibo Zhao, Guofei Chen, Honghao Zhu et al. · 2 citations
Open access Jul 2026

SAM3R: Object-Centric 3D Mapping via Foundation-Model-Guided Data Association in Changing Scenes

For embodied agents to navigate and reason indoor spaces, they need object-level 3D representations that stay consistent over time as new frames arrive from a monocular camera. Current online 3D instance segmentation methods either depend on posed RGB-D input with ground-truth depth or couple tightly to the internal re...

M. Mohrat, E. Derevyanka, I. Obrubov et al. · 0 citations
2026

TS-MapLoc: Large-Scale Indoor Object-Level Localization With Topological-Semantic Maps

Large-scale indoor mapping and positioning with vision sensors is fundamental to a wide range of applications, such as robotic navigation and augmented reality. However, the rapidly increasing number of detectable objects and the expanded spatial coverage jointly introduce matching ambiguity and high computational cost...

Cui-Yun Fang, Fan Wang, Ye-Dong Jiang et al. · 0 citations
Jul 2026

VLA-ReID: Video-Level Association for Re-Identification in Multi-Object Tracking with Highly Similar Objects

Video-Level Association re-ID (VLA-ReID), which reformulates re-ID as video-level association modeling, and uses aggregated historical trajectory features as queries and all current-frame detections as candidates as candidates, enabling direct optimization of their global association at each frame.

Yanrong Qin, X. Cao, Yao Yao · 0 citations
Open access Jul 2026

SAR-SLAM: Semantic-Aware Recognition for Dynamic SLAM in Robotic Applications

This paper introduces SAR-SLAM (Semantic-Aware Recognition SLAM), an RGB-D SLAM framework that robustly handles dynamic scenes containing moving people and objects using dual semantic geometric processing, and remains competitive with state-of-the-art dynamic SLAM methods across a range of dynamic scenarios.

Basheer Al-Tawil, Magnus Jung, Thorsten Hempel et al. · 0 citations
Conference Jul 2026

GORI: Image-Guided Selective 3D Object Re-Association for 3D Scene Graphs and Task Planning

3D scene graphs provide structured environmental representations that enable robots to perform language-grounded tasks such as navigation and manipulation. A key challenge in constructing 3D scene graphs is preserving consistent object identity. During robot navigation, newly observed 3D segments need to be associated...

Jei Kong, Seungjae Lee, Jeewon Kim et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.