Skip to content
Preprint

GeoPose: Patient-agnostic CTA-to-DSA registration through projection-space calibration

Aug 2026 · 0 citations · 34 references
Computer Science

TL;DR

GeoPose is proposed, a population-trained framework that estimates the C-arm pose in a learned canonical frame and transfers it to the native frame of an unseen CTA through projection-space calibration and transform composition, and provides rapid native-frame registration with fixed population-level weights and the geometric correspondence required for downstream biplanar 3D vascular reconstruction.

Abstract

Aligning intraoperative biplanar digital subtraction angiography (DSA) to pre-procedural computed tomography angiography (CTA) requires rapid and accurate 3D-to-2D registration. Optimization-based methods are sensitive to initialization and may require hundreds of iterations, whereas learning-based approaches commonly rely on patient-specific training. We propose GeoPose, a population-trained framework that estimates the C-arm pose in a learned canonical frame and transfers it to the native frame of an unseen CTA through projection-space calibration and transform composition. A population-trained residual network refines the pose, followed optionally by low-budget image-driven optimization. GeoPose requires neither patient-specific adaptation nor explicit inter-volume preregistration. On 80 DSA observations from 20 held-out patients, optimization-free GeoPose achieved a carotid mean projected centerline distance (mPCD) of 5.8 mm and a clDice of 0.45, compared with 14.5 mm and 0.28 for the best-performing baseline, while requiring only 0.15 s. After 25 optimization iterations, GeoPose reached an mPCD of 4.6 mm and a clDice of 0.58 in approximately two seconds. Under the same budget, native-initialized optimization achieved 14.6 mm and 0.15, respectively. GeoPose thus provides rapid native-frame registration with fixed population-level weights and the geometric correspondence required for downstream biplanar 3D vascular reconstruction.

View source

Similar papers

Preprint Sep 2026

XPos3R: Cross-Modal Transformer for Intraoperative 2D/3D Registration

This work proposes XPos3R, a generalizable pose regression method that eliminates preoperative preparation, and introduces an asymmetric encoder-decoder architecture that improves cross-modal feature alignment while maintaining computational efficiency.

Shi-Yan Su, Ruyi Zha, Hong-Dong Li et al. · 0 citations
Open access Sep 2026

GeoCM-Pose: geometry-aware monocular dental 2D/3D registration benchmarked against reference-assisted methods

Objective. To develop GeoCM-Pose, a geometry-aware monocular dental 2D/3D registration method that predicts metric model-to-camera 6DoF pose from one image under weak texture, repetitive anatomy and partial visibility. Approach. GeoCM-Pose comprises cross-modal adaptation (CMA) and geometry-aware pose regression (GAPR)...

Zhi-Xian Qiu, Jin-Gang Jiang, Jie Pan et al. · 0 citations
Preprint Sep 2026

Colon3R: Cross-Domain 3D Reconstruction from Monocular Colonoscopic Video

Monocular colonoscopic 3D reconstruction is important for surgical robotic colonoscopy, but remains challenging due to weak texture, specular reflections, limited view overlap, and non-rigid tissue motion. Conventional multi-view 3D reconstruction methods rely on stable correspondences and approximate rigidity, which a...

Zhi-Hao Xing, Ying-Yu Wang, Liang Zhao et al. · 0 citations
Open access Oct 2026

RGB-D/CT Registration for Preemptive Patient Pose Alignment in Longitudinal CT Scans

Abstract Registration of longitudinal Computed Tomography (CT) scans is critical for monitoring treatment responses, yet its precision is often compromised by inconsistencies in patient positioning between sessions. Historically, the medical imaging community has relied purely on post-hoc ("postscan") image registratio...

Elisabeth Schiele, Melina Wördehoff, J. Decker et al. · 0 citations
Conference Aug 2026

AtlasCT: Report-Conditioned 3D CT Synthesis with a Learnable Population Atlas Prior

Medical image synthesis can reduce data scarcity, but volumetric generation must preserve anatomy across planes. In chest computed tomography (CT), report conditioning specifies pathology but gives little spatial guidance, while mask-guided methods require a case-specific segmentation at inference. AtlasCT removes that...

Jia-He Hou, John Moraros, Shui-Hua Wang · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.