arXiv Science⌕ Search

arXiv · 2609.36449

Quantifying the Impact of Upright Patient Positioning on Cardiac Substructures Using Deep Learning

Abstract

Upright patient positioners with diagnostic-quality vertical CT at treatment isocenter may improve image-guided radiation therapy (RT). However, cardiac substructure (CS) geometry in upright patients remains insufficiently characterized. This work evaluated if a supine-trained deep-learning (DL) CS segmentation model generalizes to upright CT images and quantified posture-dependent CS changes in thoracic patients, to assess potential CS-sparing benefits. 8 thoracic proton therapy patients underwent paired supine/upright 4DCT imaging. 20 CS and lungs were manually labeled on both datasets for positional comparisons. A previously developed supine-trained, DL pipeline generated the same CS, and performance was evaluated using Dice similarity coefficient (DSC) and 95% Hausdorff distance (HD95). Upright CTs were rigidly registered to corresponding supine CTs by aligning the thoracic vertebrae, and CS centroid shifts were measured in the registered coordinate frame and relative to the carina. Paired differences were assessed using Wilcoxon signed-rank tests (p<0.05). The DL model successfully predicted all 20 CS on both upright (DSC, 0.65(0.24); HD95, 7.6(5.6)mm) and supine (DSC, 0.72(0.19); HD95, 5.8(2.6)mm) images yet with lower (p<0.05) performance upright. Upright positioning significantly increased median lung volume by 20.6% (range, -8.9%-42.8%). After vertebral alignment, most CS centroids shifted significantly inferior (median heart shift, 23mm; range, 18-37mm) and anterior (median heart shift, 5.0mm; range, 1.0-13.0mm) when upright. Relative to the carina, most CS shifted significantly inferior and closer anterior-posterior. A supine-trained DL model generalized to upright CT images for CS segmentation. Upright positioning produced increased lung volumes and significant inferior CS displacement, suggesting favorable geometry changes that may support cardiac-sparing workflows.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Nicholas Summerfield, Yuhao Yan, Chase Ruff, Mark Pankuch, Shae Gans, Niek Schreuder, Carri K Glide-Hurst. 2026-09-29. Quantifying the Impact of Upright Patient Positioning on Cardiac Substructures Using Deep Learning. https://arxiv.org/abs/2609.36449

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

Joint-decoupled iterative CBCT reconstruction with hybrid scatter estimation and voxel-adaptive beam hardening correction

Cone-beam computed tomography (CBCT) is fundamentally challenged by scatter and beam hardening artifacts, which originate from X-ray scattering and the polychromatic nature of the X-ray spectrum, respectively. These two types of artifacts are intricately coupled in reconstructed images and manifest with similar streaking and cupping features, severely compromising high-precision CBCT imaging. This paper proposes a physics-driven iterative framework rooted in the polychromatic Polyquant attenuation model, which decouples these artifacts by establishing an optimization loop between scatter estimation and relative electron density (RED) reconstruction. We develop a hybrid strategy for scatter estimation, in which the first-order scattering component is analytically derived based on a polychromatic physical model to preserve high-frequency structural information, whereas the smoother multiple scattering component is efficiently estimated via an object-adaptive convolution module. Subsequently, for beam-hardening correction, we introduce a voxel-adaptive update mechanism that solves linearized, scatter-corrected polychromatic equations to derive optimal weights, enabling direct RED refinement without manual parameter tuning. The proposed method was validated through comprehensive studies on biomedical phantoms, utilizing both Monte Carlo simulations and physical experiments. Representative results demonstrate that the proposed method outperforms state-of-the-art techniques, with the mean relative error decreased from 11.96\% to 1.27\% for the anthropomorphic head phantom and from 12.55\% to 5.46\% for the physical Yin-Yang phantom.

physics.med-ph↗

MBFormer: Microbubble Transformer for 3D Time-Series Data Processing to Improve Bound Bubble Detection in Nonde-structive Ultrasound Molecular Imaging

Development of nondestructive ultrasound molecular imaging (UMI) is essential for early cancer detection through real-time screening using clinical ultrasound systems. Current techniques face challenges in accurately detecting targeted microbubbles (MBs) bound to specific biomarkers, primarily due to false-positive detections of unbound free-floating MBs. We propose a transformer model for time-series video processing to improve the differentiation of bound MBs. We propose a hierarchical transformer, termed MBFormer (microbubble transformer), featuring a positional-embedding-free encoder and a lightweight decoder. Leveraging attention within 3D spatio-temporal data to effectively capture stationary signals from bound MBs while suppressing nonstationary signals from unbound MBs. Since MBs appears as relatively small textures compared with conventional segmentation targets in medical imaging, such as organs and tumors, we optimized the model with two hierarchical layers, each with an attention block, to process ultrasound video data. The network outputs the molecular signal amplitude to visualize fine MB textures. Performance was evaluated using an in vivo breast cancer model, compared against a prior CNN-based UMI method and SegFormer3D baseline, a representative 3D transformer. MBFormer (AUC = 0.943) outperformed both CNN (AUC = 0.897) and SegFormer3D (AUC = 0.766) in detecting bound MBs. The CNN showed residual molecular signal from free MBs in the cardiac chambers, whereas SegFormer3D failed to detect fine MB textures. Overall, MBFormer demonstrated enhanced detection of bound MBs while suppressing free MBs and achieved a frame rate of 16.7 to 18.1 FPS, demonstrating its potential for real-time application. We anticipate that this transformer-based UMI model can facilitate real-time, free-hand nondestructive UMI in clinical systems.

physics.med-ph↗

Physics-informed self-supervised generation of digital brain MRI phantoms from weighted images using differentiable MRI simulation

Purpose: To develop a physics-informed, self-supervised framework for generating digital brain MRI phantoms directly from conventional weighted MR images without requiring ground-truth parametric maps or anatomical segmentation. Methods: The framework predicts T1, T2, and proton density (PD) maps from T1-, T2-, and PD-weighted images and reconstructs the input images through an MRI signal model. Three generative architectures (variational autoencoder (VAE), generative adversarial network (GAN), and flow-based model) were compared. Models were pretrained on 3,739 synthetic brain slices generated using digital phantoms and an analytical MRI signal model, followed by fine-tuning on 90 real brain slices from three healthy volunteers. The best-performing architecture was subsequently fine-tuned using the differentiable MR-Zero numerical MRI simulator and evaluated on 30 held-out real slices. Results: The flow-based model demonstrated the highest overall performance and preserved fine anatomical details better than the VAE and GAN. After analytical-model fine-tuning, it achieved MS-SSIM values of 0.955-0.985 and PSNR values of 26.68-32.63 dB across T1-, T2-, and PD-weighted images. Fine-tuning with MR-Zero increased T1-weighted reconstruction quality from 0.955 to 0.977 (MS-SSIM) and from 26.68 to 30.60 dB (PSNR), and provided high robustness of metrics across different MR image weightings. The resulting digital phantoms also enabled simulation of images using previously unseen acquisition protocols. Conclusion: The proposed framework enables physics-informed generation of reusable digital brain MRI phantoms from weighted images using limited real-world data. Combining synthetic pretraining with differentiable numerical MRI simulation provides a practical approach for physically grounded MRI data augmentation without requiring reference parametric maps.

physics.med-ph↗