Skip to content

Unsupervised pelvic CBCT-to-CT translation via cross-slice attention for 3D context modeling

Sep 2026 · Biomedical engineering and physics express · Vol 12, pp. 055026 · 0 citations · 24 references
Medicine Physics

TL;DR

CAMIT provides an effective trade-off between 2D efficiency and 3D contextual modeling, supporting treatment-day anatomical assessment and adaptive workflow guidance, and is evaluated on an unpaired pelvic CBCT/CT dataset.

Abstract

Accurate CBCT-to-CT translation is important for image-guided and adaptive radiotherapy in prostate cancer, where treatment-day anatomy can differ substantially from planning CT. Most unsupervised methods are based on 2D slice-wise translation and therefore underuse volumetric context, while fully 3D models are often computationally prohibitive for routine training and deployment. To address this gap, we propose cross-slice attention based medical image translation (CAMIT), an unsupervised framework that incorporates 3D contextual information without full-volume 3D convolution. CAMIT adopts a two-stage strategy: modality-specific autoencoder pretraining to obtain compact latent representations, followed by latent-space domain translation with a cross-slice attention module that models long-range inter-slice dependencies from randomly sampled slices. We evaluated CAMIT on an unpaired pelvic CBCT/CT dataset (40 training, 5 validation, and 10 testing cases) using both quantitative and qualitative analyses. CAMIT achieved a peak signal-to-noise ratio of 27.45 and an structural similarity index of 0.67, outperforming representative state-of-the-art unsupervised baselines; statistical testing further confirmed significant performance gains. These results indicate that CAMIT provides an effective trade-off between 2D efficiency and 3D contextual modeling, supporting treatment-day anatomical assessment and adaptive workflow guidance.

View source

Similar papers

Preprint Aug 2026

Unsupervised Adaptation of 3D CT Foundation Models for 3D CBCT Segmentation

This work proposes a novel unsupervised domain adaptation (UDA) framework based on redundancy-reducing feature alignment, enabling 3D CBCT segmentation with no target-domain annotations or inference-time adaptation, and demonstrates that this approach consistently outperforms existing pretrained foundation model and UD...

Gauthier Miralles, Loic Le Folgoc, Vincent Jugnon et al. · 0 citations
Conference Open access Sep 2026

Slice-aware MedSAM Adaptation with cross-slice consistency for 2.5D multi-organ CT segmentation

Abdominal CT multi-organ segmentation is important for computer-aided diagnosis and quantitative analysis, but it remains challenging due to large variations in organ size, shape, and boundary clarity. Existing 2D methods lack inter-slice contextual modeling, while 3D methods require high computational and memory costs...

Qing-Yang Guan, Ke-Wen Qin, Wei-Hai Pang et al. · 0 citations
Open access Aug 2026

Planning CT-Guided Dual-Branch Attention GAN for Longitudinal CBCT Outpainting in Adaptive Radiotherapy.

Limited longitudinal field of view in cone-beam computed tomography (CBCT) remains a significant challenge for image-guided adaptive radiotherapy. We developed a deep learning-based CBCT outpainting framework that synthesizes anatomically consistent extensions beyond the scanned volume. The model employs a dual-branch...

Soyoung Chung, J. Kim, Min-Gyo Chung et al. · 0 citations
#computer vision Review Sep 2026

NV-Reason-CT: 3D Visual Language Model for CT Analysis

The NV-Reason-CT model, a generative vision--language model for chest and abdominal CT combining native 3D visual encoding with radiologist-guided reasoning, and the model and training code are released to support reproducible research on explainable AI for volumetric medical imaging.

Andriy Myronenko, Dong Yang, Yu-Cheng Tang et al. · 0 citations
Open access Sep 2026

Prior-Enhanced Axial TransUNet: integrating MedSAM priors and sparse region-aware attention for kidney-tumor segmentation

Accurate kidney and renal-tumor segmentation is challenging because lesion size, location, morphology, and boundary contrast vary substantially across abdominal CT scans. Most existing methods rely on a single form of local evidence and struggle to recover the boundaries of small lesions while maintaining global anat...

Si-Yuan Liang, Cheng-Chuan Xu, Chao Lu et al. · 0 citations
Conference Aug 2026

AtlasCT: Report-Conditioned 3D CT Synthesis with a Learnable Population Atlas Prior

Medical image synthesis can reduce data scarcity, but volumetric generation must preserve anatomy across planes. In chest computed tomography (CT), report conditioning specifies pathology but gives little spatial guidance, while mask-guided methods require a case-specific segmentation at inference. AtlasCT removes that...

Jia-He Hou, John Moraros, Shui-Hua Wang · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.