Skip to content

PolarBEVPlace: Rotation-Robust LiDAR Place Recognition via Radial Distortion-Aware Polar Transform

Sep 2026 · IEEE Robotics and Automation Letters · Vol 11, pp. 10585-10592 · 0 citations · 23 references

Abstract

LiDAR-based place recognition is fundamental to loop closure and global localization in SLAM. Bird’s Eye View (BEV) projection has emerged as an effective paradigm for 3D point cloud processing, yet rotation robustness remains expensive: prior approaches either discretize SO(2) at <inline-formula><tex-math notation="LaTeX">$\mathcal {O}(N_{R})$</tex-math></inline-formula> inference cost or require high-parameter dual-branch cross-attention. Polar BEV representations convert heading-induced 2D rotations into 1D cyclic translations, but introduce a systematic <italic>radial distortion</italic> unaddressed by existing architectures: expanding annuli with decreasing LiDAR point density cause feature activations to collapse at large radii, while global attention mixes geometrically incompatible cross-radius representations. We propose two lightweight mechanisms to address these failure modes: <bold>Radius-Aware CoordConv</bold> concatenates a deterministic radial coordinate channel to recover feature energy in sparse outer rings; <bold>Ringwise Circular Attention (RCA)</bold> confines self-attention to 1D azimuthal rings at each fixed radius, preventing cross-radius feature mixing by construction and reducing attention complexity from <inline-formula><tex-math notation="LaTeX">$\mathcal {O}((H_\theta W_{r})^{2})$</tex-math></inline-formula> to <inline-formula><tex-math notation="LaTeX">$\mathcal {O}(W_{r} H_\theta ^{2})$</tex-math></inline-formula>. The resulting single-stream network, PolarBEVPlace, achieves a mean Recall@1 of 94.0% across 14 sequences on KITTI, NCLT, and UrbanNav-HK at 2.30 ms per scan with 1.06 M parameters—a 5.8× latency reduction and 10.7× MAC reduction over BEVPlace++, outperforming or matching it on 13 of 14 sequences.

View source

Similar papers

Preprint Sep 2026

DXPR: Depth-Based Vision-LiDAR Cross-Modal Place Recognition Using Vision Foundation Models

We present DXPR, a depth-based cross-modal place recognition (CMPR) framework that uses vision foundation models (VFMs) to match monocular camera queries against a LiDAR map without modality-specific encoders. This enables robots and autonomous vehicles to robustly localize using only cameras within pre-built LiDAR map...

Yu-Hang Han, Youngseok Jang, Seungwon Roh et al. · 0 citations
Preprint Aug 2026

CVSD-Reg: Cross-Modal Visual Semantic Prior Distillation for Robust LiDAR Registration

CVSD-Reg is proposed, a robust global LiDAR registration framework that distills visual semantic priors from a vision foundation model into LiDAR representations and generalizes to both single-sensor and zero-shot cross-sensor scenarios without sensor-specific adaptation and remains entirely camera-free at inference.

Eunsoo Im, Junghun Suh, Gyeonggwan Lee et al. · 0 citations
Open access 2026

Sparse Voxel Meets View: Bridging Local Geometry and Global Semantics via Linear Attention in Place Recognition

Achieving accurate LiDAR-based place recognition is a crucial step towards reliable autonomous navigation, as it enables robust loop closure and global localization without GPS. However, existing 3D point cloud descriptors often give up fine local geometry for broad semantic context. We propose SBC-Net, a dual-branch a...

Minseong Park, DoHyeong Kwon, Jae-Jin Jeon et al. · 0 citations
Preprint Sep 2026

UpDown-SC: Gravity-Canonicalized Dual-Envelope Scan Context for Indoor LiDAR Place Recognition

UpDown-SC is presented, a training-free polar descriptor that first canonicalizes gravity and then represents two complementary surfaces: the upper envelope of lower/middle structures and the lower envelope of overhead structures that gives the best or second-best F1max and AUPR under threshold-based acceptance while r...

Jie Xu, Yong-Xin Yang, Zi-Yi Jin et al. · 0 citations
Open access 2026

Decoupling Mask Quality From Completion Design: A Diagnostic Framework for Occlusion-Aware LiDAR 3-D Object Detection

OccBEV-Oracle, a lightweight completion neck inserted between the BEV backbone and dense head of a CenterPoint-style detector, and a direct measurement of where the neck edits the BEV map, show that the aligned GT-ROM mask itself supplies a strong localization prior.

Jun Wang, Quanxin Zheng, Jian-Ping Yu · 0 citations
Review Open access Sep 2026

GFE-Net: Geometry-Enhanced Feature Extraction Network for Semantic Segmentation of Large-Scale LiDAR Point Clouds

Accurate semantic segmentation of large-scale outdoor LiDAR point clouds remains a challenging endeavor, primarily due to ambiguous class transitions at object interfaces, non-uniform sampling density across the surveyed area, and shared geometric signatures among distinct object categories. This paper proposes GFE-Net...

Hui Liu, Guang-Ming Zhang, Chuang Chen et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.