arXiv ScienceSearch

arXiv subjects

Adam Wright

Publications and source records attributed to Adam Wright.

6 recordsLinked to original sources

Teaching agentic AI to generalize expert diagnostic reasoning in rare diseases

Rare disease diagnosis depends on expert reasoning that is scarce and difficult to transfer. Large language models rank the correct disease first in only 35.4% of benchmark cases and often rely on learned phenotype-disease associations rather than reusable diagnostic reasoning strategies. We developed liteOdyssey through Policy Iteration with Human Feedback, a process in which model failures and expert corrections are iteratively consolidated into a clinician-gated, natural-language policy executed by a language model. Across 1,243 public benchmark cases spanning 722 rare diseases, liteOdyssey ranked the correct disease first in 59.3% of cases versus 26.5% without the policy, with comparable gains in cases involving diseases excluded from policy development. The same policy transferred across model families and sizes without retraining. Adaptation of the policy to the Undiagnosed Diseases Network (UDN) improved diagnostic accuracy among 515 UDN patients, with gains confirmed by blinded physician adjudication. These results show that expert reasoning can be externalized into an inspectable and revisable natural-language policy that generalizes across rare diseases, transfers across model backbones, and adapts to a real-world patient cohort.

cs.AI

An artificial intelligence framework for end-to-end rare disease phenotyping from clinical notes using large language models

Phenotyping is fundamental to rare disease diagnosis, but manual curation of structured phenotypes from clinical notes is labor-intensive and difficult to scale. Existing artificial intelligence approaches typically optimize individual components of phenotyping but do not operationalize the full clinical workflow of extracting features from clinical text, standardizing them to Human Phenotype Ontology (HPO) terms, and prioritizing diagnostically informative HPO terms. We developed RARE-PHENIX, an end-to-end AI framework for rare disease phenotyping that integrates large language model-based phenotype extraction, ontology-grounded standardization to HPO terms, and supervised ranking of diagnostically informative phenotypes. We trained RARE-PHENIX using data from 2,671 patients across 11 Undiagnosed Diseases Network clinical sites, and externally validated it on 16,357 real-world clinical notes from Vanderbilt University Medical Center. Using clinician-curated HPO terms as the gold standard, RARE-PHENIX consistently outperformed a state-of-the-art deep learning baseline (PhenoBERT) across ontology-based similarity and precision-recall-F1 metrics in end-to-end evaluation (i.e., ontology-based similarity of 0.70 vs. 0.58). Ablation analyses demonstrated performance improvements with the addition of each module in RARE-PHENIX (extraction, standardization, and prioritization), supporting the value of modeling the full clinical phenotyping workflow. By modeling phenotyping as a clinically aligned workflow rather than a single extraction task, RARE-PHENIX provides structured, ranked phenotypes that are more concordant with clinician curation and has the potential to support human-in-the-loop rare disease diagnosis in real-world settings.

cs.AI

Cosmological Constraints from Gas Mass Fractions of Massive, Relaxed Galaxy Clusters

We present updated cosmological constraints from measurements of the gas mass fractions ($f_{gas}$) of massive, dynamically relaxed galaxy clusters. Our new data set has greater leverage on models of dark energy, thanks to the addition of the Perseus Cluster at low redshifts, two new clusters at redshifts $z>0.97$, and significantly longer observations of four clusters at $0.6<z<0.9$. Our low-redshift ($z<0.16$) $f_{gas}$ data, combined with the cosmic baryon fraction measured from the cosmic microwave background (CMB), imply a Hubble constant of $h = 0.722 \pm 0.067$. Combining the full $f_{gas}$ data set with priors on the cosmic baryon density and the Hubble constant, we constrain the dark energy density to be $\Omega_\Lambda = 0.865 \pm 0.119$ in non-flat $\Lambda$CDM (cosmological constant) models, and its equation of state to be $w = -1.13_{-0.20}^{+0.17}$ in flat, constant-w models, respectively 41 and 29 per cent tighter than our previous work, and comparable to the best constraints available from other probes. Combining $f_{gas}$, CMB, supernova, and baryon acoustic oscillation data, we also constrain models with global curvature and evolving dark energy. For the massive, relaxed clusters employed here, we find the scaling of $f_{gas}$ with mass to be consistent with a constant, with an intrinsic scatter that corresponds to just 3 per cent in distance.

astro-ph.CO

Development of deep learning algorithms to categorize free-text notes pertaining to diabetes: convolution neural networks achieve higher accuracy than support vector machines

Health professionals can use natural language processing (NLP) technologies when reviewing electronic health records (EHR). Machine learning free-text classifiers can help them identify problems and make critical decisions. We aim to develop deep learning neural network algorithms that identify EHR progress notes pertaining to diabetes and validate the algorithms at two institutions. The data used are 2,000 EHR progress notes retrieved from patients with diabetes and all notes were annotated manually as diabetic or non-diabetic. Several deep learning classifiers were developed, and their performances were evaluated with the area under the ROC curve (AUC). The convolutional neural network (CNN) model with a separable convolution layer accurately identified diabetes-related notes in the Brigham and Womens Hospital testing set with the highest AUC of 0.975. Deep learning classifiers can be used to identify EHR progress notes pertaining to diabetes. In particular, the CNN-based classifier can achieve a higher AUC than an SVM-based classifier.

cs.CL

Weighing the Giants IV: Cosmology and Neutrino Mass

We employ robust weak gravitational lensing measurements to improve cosmological constraints from measurements of the galaxy cluster mass function and its evolution, using X-ray selected clusters detected in the ROSAT All-Sky Survey. Our lensing analysis constrains the absolute mass scale of such clusters at the 8 per cent level, including both statistical and systematic uncertainties. Combining it with the survey data and X-ray follow-up observations, we find a tight constraint on a combination of the mean matter density and late-time normalization of the matter power spectrum, $\sigma_8(\Omega_m/0.3)^{0.17}=0.81\pm0.03$, with marginalized, one-dimensional constraints of $\Omega_m=0.26\pm0.03$ and $\sigma_8=0.83\pm0.04$. For these two parameters, this represents a factor of two improvement in precision with respect to previous work, primarily due to the reduced systematic uncertainty in the absolute mass calibration provided by the lensing analysis. Our new results are in good agreement with constraints from cosmic microwave background (CMB) data, both WMAP and Planck (plus WMAP polarization), under the assumption of a flat $\Lambda$CDM cosmology with minimal neutrino mass. Consequently, we find no evidence for non-minimal neutrino mass from the combination of cluster data with CMB, supernova and baryon acoustic oscillation measurements, regardless of which all-sky CMB data set is used (and independent of the recent claimed detection of B-modes on degree scales). We also present improved constraints on models of dark energy (both constant and evolving), modifications of gravity, and primordial non-Gaussianity. Assuming flatness, the constraints for a constant dark energy equation of state from the cluster data alone are at the 15 per cent level, improving to $\sim 6$ per cent when the cluster data are combined with other leading probes.

astro-ph.CO

Robust Weak-lensing Mass Calibration of Planck Galaxy Clusters

In light of the tension in cosmological constraints reported by the Planck team between their SZ-selected cluster counts and Cosmic Microwave Background (CMB) temperature anisotropies, we compare the Planck cluster mass estimates with robust, weak-lensing mass measurements from the Weighing the Giants (WtG) project. For the 22 clusters in common between the Planck cosmology sample and WtG, we find an overall mass ratio of $\left< M_{Planck}/M_{\rm WtG} \right> = 0.688 \pm 0.072$. Extending the sample to clusters not used in the Planck cosmology analysis yields a consistent value of $\left< M_{Planck}/M_{\rm WtG} \right> = 0.698 \pm 0.062$ from 38 clusters in common. Identifying the weak-lensing masses as proxies for the true cluster mass (on average), these ratios are $\sim 1.6\sigma$ lower than the default mass bias of 0.8 assumed in the Planck cluster analysis. Adopting the WtG weak-lensing-based mass calibration would substantially reduce the tension found between the Planck cluster count cosmology results and those from CMB temperature anisotropies, thereby dispensing of the need for "new physics" such as uncomfortably large neutrino masses (in the context of the measured Planck temperature anisotropies and other data). We also find modest evidence (at 95 per cent confidence) for a mass dependence of the calibration ratio and discuss its potential origin in light of systematic uncertainties in the temperature calibration of the X-ray measurements used to calibrate the Planck cluster masses. Our results exemplify the critical role that robust absolute mass calibration plays in cluster cosmology, and the invaluable role of accurate weak-lensing mass measurements in this regard.

astro-ph.CO