Skip to content
Open access

Cross-Domain TransNet for sparse-view CT reconstruction

Jul 2026 · Frontiers in Nuclear Medicine · Vol 6 · 0 citations · 42 references
Medicine

TL;DR

The results show that Cross-Domain TransNet consistently improves reconstruction quality, effectively suppresses noise, and reduces artifacts, outperforming both conventional reconstruction algorithms and state-of-the-art deep learning approaches.

Abstract

Introduction Sparse-view computed tomography (CT) reconstruction is crucial for clinical diagnostics, as reducing radiation exposure is essential to minimize risks to patients. Existing dual-domain reconstruction methods leverage both image and projection domains but often process them sequentially, overlooking their implicit correlations. Methods To address this limitation, we propose Cross-Domain TransNet, a Transformer-based dual-domain framework for sparse-view CT reconstruction. The proposed model captures long-range dependencies within each domain and integrates image and sinogram representations through a hybrid self-attention mechanism. In addition, a Convolution Fusion Layer (CFL) is introduced to enhance feature interactions and facilitate more effective utilization of dual-domain information. Results Extensive experiments on the NIH-AAPM dataset demonstrate the superior performance and generalization capability of the proposed method under various sparse-view settings. The results show that Cross-Domain TransNet consistently improves reconstruction quality, effectively suppresses noise, and reduces artifacts, outperforming both conventional reconstruction algorithms and state-of-the-art deep learning approaches. Conclusion Cross-Domain TransNet provides an effective and robust solution for sparse-view CT reconstruction. By fully exploiting complementary information from both image and projection domains, the proposed framework enhances diagnostic image quality while supporting radiation dose reduction.

Read PDF

Similar papers

Jul 2026

Dual-Domain Cross-Prompt Learning for Efficient Sparse-View CT.

This work designs an implicit pixel-wise learnable step size to adapt to the spatial gradient heterogeneity of CT images and develops a cross-prompt guiding mechanism to enable inter-domain prompt interaction, which facilitates efficient prompt generation and enhances the convergence stability of the model.

Wenchao Du, Qiao Mu, Huanhuan Cui et al. · 0 citations
#edge computing Open access Sep 2026

CAFDIM: a group convolution and self-attention fusion-based dual-domain iterative method for sparse-view CT reconstruction

Objective. Sparse-view computed tomography (CT) reduces radiation dose and acquisition time by decreasing the number of projection views, but it also makes image reconstruction severely ill-posed, leading to structural distortion and severe artifacts. This study aims to develop an effective reconstruction framework for improving both projection-data fidelity and reconstructed image quality in sparse-view CT. Approach. We propose a group convolution- and self-attention fusion-based dual-domain iterative method (CAFDIM) for sparse-view CT reconstruction. CAFDIM follows a model-informed dual-domain iterative design. The framework consists of the initialization enhancement network, gradient update block, projection-domain repair network, image-domain repair network, and momentum update block. The projection-domain branch employs a deep sparse block to enhance sparse projection features before full-view projection restoration, while the image-domain branch uses edge-guided residual refinement to improve anatomical structure preservation. To enhance local-global feature representation, a Convolution-Attention Fusion Block is embedded into both repair branches by combining group convolution with Pixel Shift Self-Attention. Results. Experiments on simulated and real clinical projection datasets demonstrate that CAFDIM effectively suppresses sparse-view artifacts, preserves anatomical structures, and achieves superior reconstruction accuracy, visual quality, and generalization ability compared with state-of-the-art methods. Significance. CAFDIM provides an effective and efficient dual-domain reconstruction framework for sparse-view CT, showing strong potential for clinical applications in sparse-view and low-dose CT imaging.

Ji-Zhong Duan, Cheng-Hong Sun, Hai-Bo Tao et al. · 0 citations
Open access Jul 2026

Trans2-CBCT: A Dual-Transformer Framework for Sparse-View CBCT Reconstruction

Cone-beam computed tomography (CBCT) with sparse projection views offers reduced radiation dose and faster scans but introduces severe streak artifacts and spatial coverage gaps. We address these challenges within a unified framework. First, we replace conventional UNet/ResNet encoders with TransUNet, a hybrid CNN–Transformer architecture that jointly models local details and long-range spatial context. It is adapted to CBCT reconstruction by concatenating multi-scale feature maps and introducing a lightweight attenuation-prediction head. Trans-CBCT outperforms the best baseline by 1.17 dB in PSNR and by 0.0163 in SSIM on LUNA16 with only six projection views. Second, we incorporate a neighbor-aware Point Transformer with explicit 3D positional encodings and a neighbor-aware attention module aggregating information from each point’s k-nearest spatial neighbors to enforce volumetric coherence. The resulting Trans2-CBCT achieves an additional 0.63 dB increase in PSNR and 0.0117 increase in SSIM over Trans-CBCT. In experiments with 6-10 views, Trans-CBCT and Trans2-CBCT consistently outperform all prior methods in both PSNR and SSIM on LUNA16. On the ToothFairy dataset, Trans2-CBCT leads in five of the six measurements, outperforming all baselines in PSNR. These results highlight the effectiveness of combining hybrid CNN–Transformer features with geometry-aware point-based reasoning for sparse-view CBCT reconstruction.

Minmin Yang, Yunhui Zhu, Huantao Ren et al. · 0 citations
Jul 2026

K-NeAS: Scalable Multi-Material CT Reconstruction Using Neural SDFs

This paper proposes K-NeAS, a unified and scalable architecture for automated, multi-material surface reconstruction that replaces independent material networks with a shared latent backbone and introduces a fully differentiable $K$-material sequential soft selector to model an arbitrary number of overlapping tissues.

Daksh K. Shah, Emmanouil Nikolakakis, Razvan V. Marinescu · 0 citations
Jul 2026

Physics in Medicine & Biology

D 3 R-Net establishes a robust and interpretable dual-domain reconstruction framework for ULDCT imaging and consistently out-performs competing methods in terms of quantitative metrics and visual image quality across all evaluated scenarios.

Jia-Bing Xiang, Yu-Hang Yang, Yan-Xin Wang et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.