arXiv Science⌕ Search

arXiv · 2609.30618

MBFormer: Microbubble Transformer for 3D Time-Series Data Processing to Improve Bound Bubble Detection in Nonde-structive Ultrasound Molecular Imaging

Abstract

Development of nondestructive ultrasound molecular imaging (UMI) is essential for early cancer detection through real-time screening using clinical ultrasound systems. Current techniques face challenges in accurately detecting targeted microbubbles (MBs) bound to specific biomarkers, primarily due to false-positive detections of unbound free-floating MBs. We propose a transformer model for time-series video processing to improve the differentiation of bound MBs. We propose a hierarchical transformer, termed MBFormer (microbubble transformer), featuring a positional-embedding-free encoder and a lightweight decoder. Leveraging attention within 3D spatio-temporal data to effectively capture stationary signals from bound MBs while suppressing nonstationary signals from unbound MBs. Since MBs appears as relatively small textures compared with conventional segmentation targets in medical imaging, such as organs and tumors, we optimized the model with two hierarchical layers, each with an attention block, to process ultrasound video data. The network outputs the molecular signal amplitude to visualize fine MB textures. Performance was evaluated using an in vivo breast cancer model, compared against a prior CNN-based UMI method and SegFormer3D baseline, a representative 3D transformer. MBFormer (AUC = 0.943) outperformed both CNN (AUC = 0.897) and SegFormer3D (AUC = 0.766) in detecting bound MBs. The CNN showed residual molecular signal from free MBs in the cardiac chambers, whereas SegFormer3D failed to detect fine MB textures. Overall, MBFormer demonstrated enhanced detection of bound MBs while suppressing free MBs and achieved a frame rate of 16.7 to 18.1 FPS, demonstrating its potential for real-time application. We anticipate that this transformer-based UMI model can facilitate real-time, free-hand nondestructive UMI in clinical systems.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Jihye Baek, Jeong Hoon Lee, Hoda Hashemi, Arutselvan Natarajan, Farbod Tabesh, Ramasamy Paulmurugan, Jeremy J. Dahl. 2026-09-28. MBFormer: Microbubble Transformer for 3D Time-Series Data Processing to Improve Bound Bubble Detection in Nonde-structive Ultrasound Molecular Imaging. https://arxiv.org/abs/2609.30618

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

An Uncertainty-Guided Digital Twin Framework for Online Adaptive Proton Therapy in Head and Neck Cancer: A Feasibility Study

Objective: Head and neck (HN) proton therapy spans six to seven weeks of anatomical change, while offline replanning takes about a week. We present an uncertainty-guided digital twin (UGDT) framework that forecasts treatment-day anatomy before treatment and evaluate whether it generates online adaptive proton therapy (APT) plans of clinical quality. Approach: A library of 302 longitudinal deformations from 88 previously treated HN patients was transported onto each new patient's treatment planning CT (TPCT) using two-step multi-atlas deformable image registration (DIR) built on a pretrained CT foundation model, generating about 284 predicted CTs (pdCTs) with contours per patient. Dispersion of propagated clinical target volume (CTV) contours defined a patient-specific robust margin. In ten patients, the quality assurance CT (QACT) triggering a replan represented treatment-day anatomy, and the physician-approved replan was the baseline. The pdCT most similar to the QACT (pdCT-H) and one from the lowest quartile (pdCT-L) were planned to within about 5% of baseline plan quality, forward-calculated on the QACT, and reoptimized to generate online APT plans. Main results: pdCT plans scored within -0.7% (pdCT-H) and -1.0% (pdCT-L) of baseline. Forward calculation on QACT reduced high-dose CTV D98% to 88.3% and 85.5%. After online reoptimization, D98% recovered to 98.3 +/- 0.3% and 98.2 +/- 0.3%, versus 98.5 +/- 0.4% at baseline. Spinal cord and brainstem doses remained below tolerance, and plan quality scores were within -1.1% (p = 0.19) and -1.7% (p = 0.01) of baseline. Significance: UGDT generated online APT plans comparable in quality to physician-approved offline replans using anatomy forecast before treatment, enabling a transition from reactive offline replanning toward anticipatory online adaptation.

physics.med-ph↗

BART Online Open-Source Sequence Toolbox for Computational MRI

Purpose In advanced computational MRI techniques, acquisition and reconstruction techniques are jointly designed. For reproducibility, it is therefore important to provide an open implementation of both. At the same time, any use in a clinical environment usually requires a close integration with the MRI scanner. Ensuring long-time reproducibility and maintenance then poses additional challenges. In this work, we aim to provide a fully integrated open-source framework that can meet these demands. Methods A software framework to develop pulse sequences is added to BART, an open-source toolbox for computational MRI. In addition, a vendor-specific driver sequence is developed that can be used to run the sequence on a clinical MRI scanner enabling online adjustment of all relevant sequence parameters. Using the Pulseq format, the exact same sequence can also be reproduced offline using a widely used vendor-neutral open-source standard. Using the Pulseq format, the exact same sequence can also be reproduced offline. As proof-of-concept, quantitative MRI methods for T1 and joint water/fat R2*, B0 mapping using radial FLASH and model-based reconstruction are implemented in the proposed framework. Consistency between online and offline acquisition is validated in phantom and in vivo experiments. Results Quantitative MRI methods highlighting specific challenges of acquisition and reconstruction were successfully implemented in BART. Acquisition parameters and FOV can be adapted online on a clinical MRI system. Quantitative parameter maps from model-based reconstruction agree for online and offline regenerated Pulseq acquisitions. Conclusion This work enables reproducibility of advanced computational MRI methods within a comprehensive end-to-end open-source framework.

physics.med-ph↗

Quantifying the Impact of Upright Patient Positioning on Cardiac Substructures Using Deep Learning

Upright patient positioners with diagnostic-quality vertical CT at treatment isocenter may improve image-guided radiation therapy (RT). However, cardiac substructure (CS) geometry in upright patients remains insufficiently characterized. This work evaluated if a supine-trained deep-learning (DL) CS segmentation model generalizes to upright CT images and quantified posture-dependent CS changes in thoracic patients, to assess potential CS-sparing benefits. 8 thoracic proton therapy patients underwent paired supine/upright 4DCT imaging. 20 CS and lungs were manually labeled on both datasets for positional comparisons. A previously developed supine-trained, DL pipeline generated the same CS, and performance was evaluated using Dice similarity coefficient (DSC) and 95% Hausdorff distance (HD95). Upright CTs were rigidly registered to corresponding supine CTs by aligning the thoracic vertebrae, and CS centroid shifts were measured in the registered coordinate frame and relative to the carina. Paired differences were assessed using Wilcoxon signed-rank tests (p<0.05). The DL model successfully predicted all 20 CS on both upright (DSC, 0.65(0.24); HD95, 7.6(5.6)mm) and supine (DSC, 0.72(0.19); HD95, 5.8(2.6)mm) images yet with lower (p<0.05) performance upright. Upright positioning significantly increased median lung volume by 20.6% (range, -8.9%-42.8%). After vertebral alignment, most CS centroids shifted significantly inferior (median heart shift, 23mm; range, 18-37mm) and anterior (median heart shift, 5.0mm; range, 1.0-13.0mm) when upright. Relative to the carina, most CS shifted significantly inferior and closer anterior-posterior. A supine-trained DL model generalized to upright CT images for CS segmentation. Upright positioning produced increased lung volumes and significant inferior CS displacement, suggesting favorable geometry changes that may support cardiac-sparing workflows.

physics.med-ph↗