arXiv ScienceSearch

arXiv · 2507.05144

Clinical test cases for commissioning, QA, and benchmarking of model-based dose calculation algorithms in 192Ir HDR gynecologic tandem and ring brachytherapy

Abstract

Purpose: To develop clinically relevant test cases for commissioning Model-Based Dose Calculation Algorithms (MBDCAs) for 192Ir High Dose Rate (HDR) gynecologic brachytherapy following the workflow proposed by the TG-186 report and the WGDCAB report 372. Acquisition and Validation Methods: Two cervical cancer intracavitary HDR brachytherapy patient models were created, using either uniformly structured regions or realistic segmentation. The computed tomography (CT) images of the models were converted to DICOM CT images via MATLAB and imported into two Treatment Planning Systems (TPSs) with MBDCA capability. The clinical segmentation was expanded to include additional organs at risk. The actual clinical treatment plan was generally maintained, with the source replaced by a generic 192Ir HDR source. Dose to medium in medium calculations were performed using the MBDCA option of each TPS, and three different Monte Carlo (MC) simulation codes. MC results agreed within statistical uncertainty, while comparisons between MBDCA and MC dose distributions highlighted both strengths and limitations of the studied MBDCAs, suggesting potential approaches to overcome the challenges. Data Format and Usage Notes: The datasets for the developed cases are available online at http://doi.org/ 10.5281/zenodo.15720996. The DICOM files include the treatment plan for each case, TPS, and the corresponding reference MC dose data. The package also contains a TPS- and case-specific user guide for commissioning the MBDCAs, and files needed to replicate the MC simulations. Potential Applications: The provided datasets and proposed methodology offer a commissioning framework for TPSs using MBDCAs, and serve as a benchmark for brachytherapy researchers using MC methods. They also facilitate intercomparisons of MBDCA performance and provide a quality assurance resource for evaluating future TPS software updates.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

V. Peppa, M. Robitaille, F. Akbari, S. A. Enger, R. M. Thomson, F. Mourtada, G. P. Fonseca. 2026-01-26. Clinical test cases for commissioning, QA, and benchmarking of model-based dose calculation algorithms in 192Ir HDR gynecologic tandem and ring brachytherapy. https://arxiv.org/abs/2507.05144

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

Fail-closed conformal multiresolution error control for conservative three-dimensional dose remapping in synthetic phantoms

Conservative dose remapping can be formulated by transporting mass and deposited energy across nonmatching grids, but fixed high-sampling calculations use the same work for easy and difficult cases. We developed conformal multiresolution error control (CoMERC), a fail-closed multilevel quasi-Monte Carlo (QMC) controller for reference-relative target-cell allocation error in a fixed piecewise-constant source model with known deformation. Two randomized nested replicates were evaluated at a probe level of 4 and candidate levels of 8 and 16. Four endpoints were controlled jointly: reference-mass-weighted global, high-gradient, and density-gradient-threshold root-mean-square error, plus eligible-cell maximum absolute error. A split-conformal multiplier was calibrated from 240 synthetic cases, frozen, and evaluated on 400 independent cases from the same generator and 120 cases from a prespecified shifted generator. In the primary set, joint coverage was 394/400 (98.50%; one-sided 95% lower limit, 97.06%), 388/400 cases (97.00%; lower limit, 95.18%) received an output, and 0/388 released outputs exceeded any tolerance (one-sided 95% upper limit, 0.769%). Final actions were level 8 for 190 cases, level 16 for 198, and ABSTAIN for 12. The frozen policy implied mean relative production-sample work of 0.5844 compared with always running both randomized replicates through level 16; the prespecified 95th-percentile bootstrap upper limit was 0.6194. All seven primary gates passed. In the shifted-generator set, 113/120 cases were released and one released output exceeded a tolerance. CoMERC provided marginal finite-sample error control and lower modeled production-sample work in the locked synthetic population, but not patient-level, registration-level, conditional-on-release, or clinical safety validation.

physics.med-ph

Dosimetric equivalence of deep learning prostate contours after LDR brachytherapy: pre-declared margins, patient-level acceptance thresholds and the incremental predictive value of DVH indices

Background and purpose: Dose Volume Histogram (DVH) indices remain the main dose-effect metrics for toxicity prediction, but they depend on how contours were made. Inter-observer variability (IOV) is unavoidable and clinically accepted, so the question for automatic segmentation addresses equivalence: do automatic contours produce DVH errors comparable to human IOV, and can these indices predict patient-reported toxicity? Materials and methods: In 429 patients treated with iodine-125 LDR-BT monotherapy, indices from expert manual delineation were compared with a deterministic and a Bayesian nnUNet on a fixed dose distribution, by two-one-sided tests against margins set from CT contouring IOV. Logistic regression gave patient-level thresholds at 90\,\% probability of equivalence. In 380 patients, eleven DHV indices were added to a clinical baseline predicting change in International Prostate Symptom Score (IPSS) at six horizons over 5 years, with nested cross-validation, bootstrap intervals and corrected $t$-tests. Results: All cohort-level comparisons of DVH indices were declared equivalent, with no interval consuming more than 44\,\% of its equivalence margin. Individual agreement was weaker with equivalence rates from 55.9\,\% to 89.7\,\%, depending on the DVH index and automatic segmentation model. Thresholds ranged from 0.864 to 0.958 Dice. No DVH block improved IPSS prediction at any horizon. The largest improvement declared by bootstrap intervals was 0.12 IPSS points, well below the minimal clinically important difference, across all learners. Conclusion: Automatic contours matched expert dosimetry within human IOV on average, but individual equivalence requires a volume dependent quality metric threshold definition. Our DVH indices panel provides no significant predictive power for IPSS, irrespective of segmentation source and horizon.

physics.med-ph

The Electrodynamic Basis of Dichroism-Mediated Polarization Perception

Humans see the polarization of light through entoptic percepts arising from the macula's Henle fiber layer, where xanthophyll pigments absorb preferentially across the radiating fibers. From the layer's complex dielectric tensor alone, Maxwell's equations in Berreman $4\times4$ form yield its Mueller matrix, set by four scalars $\{A,B,C,D\}$. Its intensity channel depends only on the dichroic pair $A$ and $B$, which fix the percepts at a maximum contrast $|B|/A\approx0.05$. One relation generates the whole dichroism-mediated family: Haidinger's brushes under a uniform field, their dark arms perpendicular to the $\mathbf{E}$-vector; fractured brushes under spatially varying fields; and $N$-fold brushes under vector-vortex illumination.

physics.med-ph