arXiv ScienceSearch

arXiv subjects

Alexandre Bousse

Publications and source records attributed to Alexandre Bousse.

At least 19 recordsLinked to original sources

Continuous 3-D Latent Diffusion for Medical Image Generation and Reconstruction

High-resolution three-dimensional (3-D) medical diffusion models remain constrained by the cost of processing full volumes, even when denoising is performed in a compact latent space. We introduce a continuous 3-D latent diffusion model (LDM) framework for computed tomography (CT) and magnetic resonance imaging (MRI) generation and measurement-guided reconstruction. Its central component is a compact autoencoder (AE) with a coordinate-conditioned local implicit image function (LIIF) decoder that represents a volume as a continuous function of spatial coordinates. By evaluating the convolutional decoder once on the latent grid and restricting repeated computation to a lightweight implicit head, the proposed design avoids overlapping sub-volume decoding while remaining differentiable for inverse-problem optimization. We evaluate the framework on CT volumes of 512^3 voxels and MRI volumes of 256^3 voxels. On high-resolution CT, the proposed AE is approximately x12-32 faster than the evaluated reference autoencoders, achieves the lowest peak graphics processing unit (GPU) memory use, and retains comparable structural fidelity despite a moderate reduction in voxel-level accuracy. The resulting frozen 3-D latent prior generates coherent full volumes without visible patch seams and can be applied, without task-specific retraining, to sparse-view CT and accelerated MRI reconstruction through hard data consistency. Although direct pixel-domain reconstruction remains more accurate, the results demonstrate that a single volumetric latent prior can support both unconditional generation and measurement-conditioned reconstruction on one GPU. Overall, the framework provides a practical trade-off between continuous volumetric decoding, computational efficiency, and fine-detail preservation. Our code will be made available at https://github.com/mellak/.

physics.med-ph

Multilevel Stochastic Plug-and-Play for Sparse-View CT Reconstruction

Sparse-view computed tomography (SVCT) reduces radiation exposure and acquisition time, but the limited number of projection views makes the reconstruction problem severely ill-posed and leads to streak artifacts when analytical methods are used. Plug-and-Play (PnP) methods provide an effective way to combine data fidelity with learned image priors, while stochastic PnP methods further improve robustness by matching the denoiser input distribution through re-noising. However, these methods often require many iterations to converge, which limits their practical efficiency. In this work, we propose a multilevel (ML) stochastic PnP method for SVCT that accelerates stochastic PnP reconstruction. We highlight that, in the stochastic setting, directly enforcing prior coherence across levels would require accurately estimating fine-level prior gradients through multiple denoiser function evaluations, which substantially increases the computational cost. Motivated by this observation, we perform the multilevel steps in multiresolution analysis (MRA) approximation spaces. This choice is supported by the structure of the wavelet decomposition, which causes the prior-coherence correction to vanish in expectation, thereby avoiding costly estimation of fine-level stochastic prior gradients for the coarse-level corrections. Experiments on SVCT reconstruction show that our method, called Multilevel Stochastic Plug-and-Play (ML-SPnP), achieves reconstruction quality comparable to state-of-the-art methods while substantially reducing runtime.

cs.CV

Direct Dual-Energy CT Material Decomposition using Model-based Denoising Diffusion Model

Dual-energy X-ray Computed Tomography (DECT) constitutes an advanced technology which enables automatic decomposition of materials in clinical images without manual segmentation using the dependency of the X-ray linear attenuation with energy. However, most methods perform material decomposition in the image domain as a post-processing step after reconstruction but this procedure does not account for the beam-hardening effect and it results in sub-optimal results. In this work, we propose a deep learning procedure called Dual-Energy Decomposition Model-based Diffusion (DEcomp-MoD) for quantitative material decomposition which directly converts the DECT projection data into material images. The algorithm is based on incorporating the knowledge of the spectral DECT model into the deep learning training loss and combining a score-based denoising diffusion learned prior in the material image domain. Importantly the inference optimization loss takes as inputs directly the sinogram and converts to material images through a model-based conditional diffusion model which guarantees consistency of the results. We evaluate the performance with both quantitative and qualitative estimation of the proposed DEcomp-MoD method on synthetic DECT sinograms from the low-dose AAPM dataset. Finally, we show that DEcomp-MoD outperform state-of-the-art unsupervised score-based model and supervised deep learning networks, with the potential to be deployed for clinical diagnosis.

eess.IV

Joint Reconstruction of Activity and Attenuation in PET by Diffusion Posterior Sampling in Wavelet Coefficient Space

Attenuation correction (AC) is necessary for accurate activity quantification in positron emission tomography (PET). Conventional reconstruction methods typically rely on attenuation maps derived from a co-registered computed tomography (CT) or magnetic resonance (MR) scan. However, this additional scan may complicate the imaging workflow, introduce misalignment artifacts and increase radiation exposure. In this paper, we propose a joint reconstruction of activity and attenuation (JRAA) approach that eliminates the need for auxiliary anatomical imaging by relying solely on emission data. This framework combines wavelet diffusion model (WDM) and diffusion posterior sampling (DPS) to reconstruct fully three-dimensional (3-D) data. Experimental results on simulated data show our method outperforms maximum likelihood activity and attenuation (MLAA) and MLAA-UNet with U-Net-based postprocessing, and yields high-quality noise-free reconstructions across various count settings with time-of-flight (TOF). It is also able to reconstruct non-TOF data, although the reconstruction quality significantly degrades in low-count (LC) conditions, limiting its practical effectiveness in such settings. Nonetheless, a non-TOF Biograph mMR real data reconstruction with joint scatter estimation highlights the potential of the method for clinical applications. This approach represents a step towards stand-alone PET imaging by reducing the dependence on anatomical modalities while maintaining quantification accuracy, even in LC scenarios when TOF information is available. Our code is available on GitHub at https://github.com/clemphg/jraa-dps.

physics.med-ph

Adaptive Diffusion Models for Sparse-View Motion-Corrected Head Cone-beam CT

Cone-beam computed tomography (CBCT) is an imaging modality widely used in head and neck diagnostics due to its accessibility and lower radiation dose. However, its relatively long acquisition times make it susceptible to patient motion, especially under sparse-view settings used to reduce dose, which can result in severe image artifacts. In this work, we propose a novel framework combining joint reconstruction and motion estimation (JRM) with an adaptive diffusion model (ADM) that simultaneously addresses motion compensation and sparse-view reconstruction in head CBCT. Leveraging recent advances in diffusion-based generative models, our method integrates a wavelet-domain diffusion prior into an iterative reconstruction pipeline to guide the solution toward anatomically plausible volumes while estimating rigid motion parameters in a blind fashion. We evaluate our method on simulated motion-affected CBCT data derived from real clinical computed tomography (CT) volumes. Experimental results demonstrate that JRM- ADM achieves consistent quantitative improvements over both traditional and learning-based baselines. In highly undersampled cases, JRM-ADM improves peak signal-to-noise ratio (PSNR) by more than 4 dB and structural similarity index measure (SSIM) by 0.10 compared to the baseline motion-corrected (MC) reconstruction method. These results highlight the potential of our approach to enable motion-robust, low-dose CBCT imaging, paving the way for improved clinical viability. The project page is available at https://antoinedepaepe.github.io/jrm-adm-io/.

physics.med-ph

Material Decomposition in Photon-Counting Computed Tomography with Diffusion Models: Comparative Study and Hybridization with Variational Regularizers

Photon-counting computed tomography (PCCT) has emerged as a promising imaging technique, enabling spectral imaging and material decomposition (MD). However, images typically suffer from a low signal-to-noise ratio (SNR) due to constraints such as low photon counts and sparse-view settings which provoke artifacts. To prevent this, variational methods minimize a data-fit function coupled with handcrafted regularizers that mimic a prior by enforcing image properties such as gradient sparsity. In the last few years, diffusion models (DMs) have become predominant in the field of generative models and have been used as a learned prior for image reconstruction. This work investigates the use of DMs as regularizers for MD tasks in PCCT, specifically using diffusion posterior sampling (DPS) guidance. Three DPS-based approaches -- image-domain two-step DPS (im-TDPS), projection-domain two-step DPS (proj-TDPS), and one-step DPS (ODPS) -- are evaluated. The first two methods achieve MD in two steps by performing reconstruction and MD separately. The last method, ODPS, samples the material images directly from the measurement data. The results indicate that ODPS achieves superior performance compared to im-TDPS and proj-TDPS, providing sharper, noise-free and crosstalk-free images. Furthermore, we introduce a novel hybrid method for scenarios involving materials absent from the training dataset which combines DM priors with standard variational handcrafted regularizers for the materials unknown to the DM. This hybrid method demonstrates improved MD quality compared to a standard variational method and does not require additional training of the DM neural network (NN).

physics.med-ph

Semi-Supervised Learning for Dose Prediction in Targeted Radionuclide: A Synthetic Data Study

Targeted Radionuclide Therapy (TRT) is a modern strategy in radiation oncology that aims to administer a potent radiation dose specifically to cancer cells using cancer-targeting radiopharmaceuticals. Accurate radiation dose estimation tailored to individual patients is crucial. Deep learning, particularly with pre-therapy imaging, holds promise for personalizing TRT doses. However, current methods require large time series of SPECT imaging, which is hardly achievable in routine clinical practice, and thus raises issues of data availability. Our objective is to develop a semi-supervised learning (SSL) solution to personalize dosimetry using pre-therapy images. The aim is to develop an approach that achieves accurate results when PET/CT images are available, but are associated with only a few post-therapy dosimetry data provided by SPECT images. In this work, we introduce an SSL method using a pseudo-label generation approach for regression tasks inspired by the FixMatch framework. The feasibility of the proposed solution was preliminarily evaluated through an in-silico study using synthetic data and Monte Carlo simulation. Experimental results for organ dose prediction yielded promising outcomes, showing that the use of pseudo-labeled data provides better accuracy compared to using only labeled data.

physics.med-ph

Dual-Input Dynamic Convolution for Positron Range Correction in PET Image Reconstruction

Positron range (PR) blurring degrades positron emission tomography (PET) image resolution, particularly for high-energy emitters like gallium-68 (68 Ga). We introduce Dual-Input Dynamic Convolution (DDConv), a novel computationally efficient approach trained with voxel-specific PR point spread functions (PSFs) from Monte Carlo (MC) simulations and designed to be utilized within an iterative reconstruction algorithm to perform PR correction (PRC). By dynamically inferring local blurring kernels through a trained convolutional neural network (CNN), DDConv captures complex tissue interfaces more accurately than prior methods. Additionally, it also computes the transpose operator, ensuring consistency within iterative PET reconstruction. Comparisons with a state-of-the-art, tissue-dependent correction confirm the advantages of DDConv in recovering higher-resolution details in heterogeneous regions, including bone-soft tissue and lung-soft tissue boundaries. Experiments across digital phantoms and MC-simulated data show that DDConv offers near-MC accuracy and outperforms the state-of-the-art technique, namely spatially-variant and tissue-dependent (SVTD), especially in areas with complex material interfaces. Results from real phantom experiments further confirm DD-Conv's robustness and practical applicability: while both DD-Conv and SVTD performed similarly in homogeneous soft-tissue regions, DDConv provided more accurate activity recovery and sharper delineation at heterogeneous lung-soft tissue interfaces. Our code available at https://github.com/mellak/ddconv-prc.

physics.med-ph

Solving Blind Inverse Problems: Adaptive Diffusion Models for Motion-corrected Sparse-view 4DCT

Four-dimensional computed tomography (4DCT) is essential for medical imaging applications like radiotherapy, which demand precise respiratory motion representation. Traditional methods for reconstructing 4DCT data suffer from artifacts and noise, especially in sparse-view, low-dose contexts. Motion-corrected (MC) reconstruction is a blind inverse problem that we propose to solve with a novel diffusion model (DM) framework that calibrates an adaptive unknown forward model for motion correction. Furthermore, we used a wavelet diffusion model (WDM) to address computational cost and memory usage. By leveraging the prior probability distribution function (PDF) from the DMs, we enhance the joint reconstruction and motion estimation (JRM) process, improving image quality and preserving resolution. Experiments on extended cardiac-torso (XCAT) phantom data demonstrate that our method outperforms existing techniques, yielding artifact-free, high-resolution reconstructions even under irregular breathing conditions. These results showcase the potential of combining DMs with motion correction to advance sparse-view 4DCT imaging.

physics.med-ph

Evaluation of Deep Learning-based Scatter Correction on a Long-axial Field-of-view PET scanner

Objective: Long-axial field-of-view (LAFOV) positron emission tomography (PET) systems allow higher sensitivity, with an increased number of detected lines of response induced by a larger angle of acceptance. However, this extended angle increases the number of multiple scatters and the scatter contribution within oblique planes. As scattering affects both quality and quantification of the reconstructed image, it is crucial to correct this effect with more accurate methods than the state-of-the-art single scatter simulation (SSS) that can reach its limits with such an extended field-of-view (FOV). In this work, which is an extension of our previous assessment of deep learning-based scatter estimation (DLSE) carried out on a conventional PET system, we aim to evaluate the DLSE method performance on LAFOV total-body PET. Approach: The proposed DLSE method based on a convolutional neural network (CNN) U-Net architecture uses emission and attenuation sinograms to estimate scatter sinogram. The network was trained from Monte-Carlo (MC) simulations of XCAT phantoms [18F]-FDG PET acquisitions using a Siemens Biograph Vision Quadra scanner model, with multiple morphologies and dose distributions. We firstly evaluated the method performance on simulated data in both sinogram and image domain by comparing it to the MC ground truth and SSS scatter sinograms. We then tested the method on seven [18F]-FDG and seven [18F]-PSMA clinical datasets, and compare it to SSS estimations. Results: DLSE showed superior accuracy on phantom data, greater robustness to patient size and dose variations compared to SSS, and better lesion contrast recovery. It also yielded promising clinical results, improving lesion contrasts in [18F]-FDG datasets and performing consistently with [18F]-PSMA datasets despite no training with [18F]-PSMA.

physics.med-ph

Joint Reconstruction of the Activity and the Attenuation in PET by Diffusion Posterior Sampling: a Feasibility Study

This study introduces a novel framework for joint reconstruction of the activity and the attenuation (JRAA) in positron emission tomography (PET) using diffusion posterior sampling (DPS). By leveraging diffusion models (DMs), this approach directly addresses activity-attenuation dependencies, mitigating crosstalk issues prevalent in non-time-of-flight (TOF) settings. Experimental evaluations, conducted using 2-dimensional (2-D) XCAT phantom data, demonstrate that DPS significantly outperforms traditional maximum likelihood activity and attenuation (MLAA) methods, producing consistent and high-quality reconstructions even in the absence of TOF information. Ongoing work aims to extend our method to real 3-dimensional (3-D) data with encouraging preliminary findings.

physics.med-ph

Synergistic PET/CT Reconstruction Using a Joint Generative Model

We propose in this work a framework for synergistic positron emission tomography (PET)/computed tomography (CT) reconstruction using a joint generative model as a penalty. We use a synergistic penalty function that promotes PET/CT pairs that are likely to occur together. The synergistic penalty function is based on a generative model, namely $\beta$-variational autoencoder ($\beta$-VAE). The model generates a PET/CT image pair from the same latent variable which contains the information that is shared between the two modalities. This sharing of inter-modal information can help reduce noise during reconstruction. Our result shows that our method was able to utilize the information between two modalities. The proposed method was able to outperform individually reconstructed images of PET (i.e., by maximum likelihood expectation maximization (MLEM)) and CT (i.e., by weighted least squares (WLS)) in terms of peak signal-to-noise ratio (PSNR). Future work will focus on optimizing the parameters of the $\beta$-VAE network and further exploration of other generative network models.

physics.med-ph

Direct3{\gamma}: A Pipeline for Direct Three-gamma PET Image Reconstruction

This paper presents a novel image reconstruction pipeline for three-gamma (3-{\gamma}) positron emission tomography (PET) aimed at improving spatial resolution and reducing noise in nuclear medicine; the proposed Direct3{\gamma} pipeline addresses the inherent challenges in 3-{\gamma} PET systems, such as detector imperfections and uncertainty in photon interaction points, with a key feature being its ability to determine the order of interactions through a model trained on Monte Carlo (MC) simulations using the Geant4 Application for Tomography Emission (GATE) toolkit, thus providing the necessary information to construct Compton cones which intersect with the line of response (LOR) to estimate the emission point; the pipeline processes 3-{\gamma} PET raw data, reconstructs histoimages by propagating energy and spatial uncertainties along the LOR, and applies a 3-D convolutional neural network (CNN) to refine these intermediate images into high-quality reconstructions, further enhancing image quality through supervised learning and adversarial losses that preserve fine structural details; experimental results show that Direct3{\gamma} consistently outperforms conventional 200-ps time-of-flight (TOF) PET in terms of structural similarity index measure (SSIM) and peak signal-to-noise ratio (PSNR).

physics.med-ph

CConnect: Synergistic Convolutional Regularization for Cartesian T2* Mapping

Magnetic resonance imaging (MRI) is fundamental for the assessment of many diseases, due to its excellent tissue contrast characterization. This is based on quantitative techniques, such as T1 , T2 , and T2* mapping. Quantitative MRI requires the acquisition of several contrast-weighed images followed by a fitting to an exponential model or dictionary matching, which results in undesirably long acquisition times. Undersampling reconstruction techniques are commonly employed to speed up the scan, with the drawback of introducing aliasing artifacts. However, most undersampling reconstruction techniques require long computational times or do not exploit redundancies across the different contrast-weighted images. This study introduces a new regularization technique to overcome aliasing artifacts, namely CConnect, which uses an innovative regularization term that leverages several trained convolutional neural networks (CNNs) to connect and exploit information across image contrasts in a latent space. We validate our method using in-vivo T2* mapping of the brain, with retrospective undersampling factors of 4, 5 and 6, demonstrating its effectiveness in improving reconstruction in comparison to state-of-the-art techniques. Comparisons against joint total variation, nuclear low rank and a deep learning (DL) de-aliasing post-processing method, with respect to structural similarity index measure (SSIM) and peak signal-to-noise ratio (PSNR) metrics are presented.

physics.med-ph

Multibranch Generative Models for Multichannel Imaging with an Application to PET/CT Synergistic Reconstruction

This paper presents a novel approach for learned synergistic reconstruction of medical images using multibranch generative models. Leveraging variational autoencoders (VAEs), our model learns from pairs of images simultaneously, enabling effective denoising and reconstruction. Synergistic image reconstruction is achieved by incorporating the trained models in a regularizer that evaluates the distance between the images and the model. We demonstrate the efficacy of our approach on both Modified National Institute of Standards and Technology (MNIST) and positron emission tomography (PET)/computed tomography (CT) datasets, showcasing improved image quality for low-dose imaging. Despite challenges such as patch decomposition and model limitations, our results underscore the potential of generative models for enhancing medical imaging reconstruction.

eess.IV

Spectral CT Two-step and One-step Material Decomposition using Diffusion Posterior Sampling

This paper proposes a novel approach to spectral computed tomography (CT) material decomposition that uses the recent advances in generative diffusion models (DMs) for inverse problems. Spectral CT and more particularly photon-counting CT (PCCT) can perform transmission measurements at different energy levels which can be used for material decomposition. It is an ill-posed inverse problem and therefore requires regularization. DMs are a class of generative model that can be used to solve inverse problems via diffusion posterior sampling (DPS). In this paper we adapt DPS for material decomposition in a PCCT setting. We propose two approaches, namely Two-step Diffusion Posterior Sampling (TDPS) and One-step Diffusion Posterior Sampling (ODPS). Early results from an experiment with simulated low-dose PCCT suggest that DPSs have the potential to outperform state-of-the-art model-based iterative reconstruction (MBIR). Moreover, our results indicate that TDPS produces material images with better peak signal-to-noise ratio (PSNR) than images produced with ODPS with similar structural similarity (SSIM).

physics.med-ph

Fast-Track of F-18 Positron paths simulations using GANs

In recent years, the use of Monte Carlo (MC) simulations in the domain of Medical Physics has become a state-of-the-art technology that consumes lots of computational resources for the accurate prediction of particle interactions. The use of generative adversarial network (GAN) has been recently proposed as an alternative to improve the efficiency and extending the applications of computational tools in both medical imaging and therapeutic applications. This study introduces a new approach to simulate positron paths originating from Fluorine 18 (18 F) isotopes through the utilization of GANs. The proposed methodology developed a pure conditional transformer least squares (LS)-GAN model, designed to generate positron paths, and to track their interaction within the surrounding material. Conditioning factors include the pre-determined number of interactions, and the initial momentum of the emitted positrons, as derived from the emission spectrum of 18 F. By leveraging these conditions, the model aims to quickly and accurately simulate electromagnetic interactions of positron paths. Results were compared to the outcome produced with Geant4 Application for Tomography Emission (GATE) MC simulations toolkit. Less than 10 % of difference was observed in the calculation of the mean and maximum length of the path and the 1-D point spread function (PSF) for three different materials (Water, Bone, Lung).

physics.med-ph

Diffusion Posterior Sampling for Synergistic Reconstruction in Spectral Computed Tomography

Using recent advances in generative artificial intelligence (AI) brought by diffusion models, this paper introduces a new synergistic method for spectral computed tomography (CT) reconstruction. Diffusion models define a neural network to approximate the gradient of the log-density of the training data, which is then used to generate new images similar to the training ones. Following the inverse problem paradigm, we propose to adapt this generative process to synergistically reconstruct multiple images at different energy bins from multiple measurements. The experiments suggest that using multiple energy bins simultaneously improves the reconstruction by inverse diffusion and outperforms state-of-the-art synergistic reconstruction techniques.

physics.med-ph