arXiv ScienceSearch

arXiv subjects

Matthew Young

Publications and source records attributed to Matthew Young.

9 recordsLinked to original sources

Compiled AI: Deterministic Code Generation for LLM-Based Workflow Automation

We study compiled AI, a paradigm in which large language models generate executable code artifacts during a compilation phase, after which workflows execute deterministically without further model invocation. This paradigm has antecedents in prior work on declarative pipeline optimization (DSPy) and hybrid neural-symbolic planning (LLM+P); our contribution is a systems-oriented study of its application to high-stakes enterprise workflows, with particular emphasis on healthcare settings where reliability and auditability are critical. By constraining generation to narrow business-logic functions embedded in validated templates, compiled AI trades runtime flexibility for predictability, auditability, cost efficiency, and reduced security exposure. We introduce (i) a system architecture for constrained LLM-based code generation, (ii) a four-stage generation-and-validation pipeline that converts probabilistic model output into production-ready code artifacts, and (iii) an evaluation framework measuring operational metrics including token amortization, determinism, reliability, security, and cost. We evaluate on two task types: function-calling (BFCL, n=400) and document intelligence (DocILE, n=5,680 invoices). On function-calling, compiled AI achieves 96% task completion with zero execution tokens, breaking even with runtime inference at approximately 17 transactions and reducing token consumption by 57x at 1,000 transactions. On document intelligence, our Code Factory variant matches Direct LLM on key field extraction (KILE: 80.0%) while achieving the highest line item recognition accuracy (LIR: 80.4%). Security evaluation across 135 test cases demonstrates 96.7% accuracy on prompt injection detection and 87.5% on static code safety analysis with zero false positives.

cs.SE

From Statistical Fidelity to Clinical Consistency: Scalable Generation and Auditing of Synthetic Patient Trajectories

Access to electronic health records (EHRs) for digital health research is often limited by privacy regulations and institutional barriers. Synthetic EHRs have been proposed as a way to enable safe and sovereign data sharing; however, existing methods may produce records that capture overall statistical properties of real data but present inconsistencies across clinical processes and observations. We developed an integrated pipeline to make synthetic patient trajectories clinically consistent through two synergistic steps: high-fidelity generation and scalable auditing. Using the MIMIC-IV database, we trained a knowledge-grounded generative model that represents nearly 32,000 distinct clinical events, including demographics, laboratory measurements, medications, procedures, and diagnoses, while enforcing structural integrity. To support clinical consistency at scale, we incorporated an automated auditing module leveraging large language models to filter out clinical inconsistencies (e.g., contraindicated medications) that escape probabilistic generation. We generated 18,071 synthetic patient records derived from a source cohort of 180,712 real patients. While synthetic clinical event probabilities demonstrated robust agreement (mean bias effectively 0.00) and high correlation (R2=0.99) with the real counterparts, review of a random sample of synthetic records (N=20) by three clinicians identified inconsistencies in 45-60% of them. Automated auditing reduced the difference between real and synthetic data (Cohen's effect size d between 0.59 and 1.60 before auditing, and between 0.18 and 0.67 after auditing). Downstream models trained on audited data matched or even exceeded real-data performance. We found no evidence of privacy risks, with membership inference performance indistinguishable from random guessing (F1-score=0.51).

cs.LG

Characterization of MKIDs for CMB observation at 220 GHz with the South Pole Telescope

We present an updated design of the 220 GHz microwave kinetic inductance detector (MKID) pixel for SPT-3G+, the next-generation camera for the South Pole Telescope. We show results of the dark testing of a 63-pixel array with mean inductor quality factor $Q_i = 4.8 \times 10^5$, aluminum inductor transition temperature $T_c = 1.19$ K, and kinetic inductance fraction $\alpha_k = 0.32$. We optically characterize both the microstrip-coupled and CPW-coupled resonators, and find both have a spectral response close to prediction with an optical efficiency of $\eta \sim 70\%$. However, we find slightly lower optical response on the lower edge of the band than predicted, with neighboring dark detectors showing more response in this region, though at level consistent with less than 5\% frequency shift relative to the optical detectors. The detectors show polarized response consistent with expectations, with a cross-polar response of $\sim 10\%$ for both detector orientations.

astro-ph.IM

DARTS: DenseUnet-based Automatic Rapid Tool for brain Segmentation

Quantitative, volumetric analysis of Magnetic Resonance Imaging (MRI) is a fundamental way researchers study the brain in a host of neurological conditions including normal maturation and aging. Despite the availability of open-source brain segmentation software, widespread clinical adoption of volumetric analysis has been hindered due to processing times and reliance on manual corrections. Here, we extend the use of deep learning models from proof-of-concept, as previously reported, to present a comprehensive segmentation of cortical and deep gray matter brain structures matching the standard regions of aseg+aparc included in the commonly used open-source tool, Freesurfer. The work presented here provides a real-life, rapid deep learning-based brain segmentation tool to enable clinical translation as well as research application of quantitative brain segmentation. The advantages of the presented tool include short (~1 minute) processing time and improved segmentation quality. This is the first study to perform quick and accurate segmentation of 102 brain regions based on the surface-based protocol (DMK protocol), widely used by experts in the field. This is also the first work to include an expert reader study to assess the quality of the segmentation obtained using a deep-learning-based model. We show the superior performance of our deep-learning-based models over the traditional segmentation tool, Freesurfer. We refer to the proposed deep learning-based tool as DARTS (DenseUnet-based Automatic Rapid Tool for brain Segmentation). Our tool and trained models are available at https://github.com/NYUMedML/DARTS

q-bio.QM

Electric Sheep Team Description Paper Humanoid League Kid-Size 2019

In this paper we introduce the newly formed New Zealand based RoboCup Humanoid Kid-Size team, Electric Sheep. We describe our developed humanoid robot platform, particularly our unique take on the chassis, electronics and use of several motor types to create a low-cost entry platform. To support this hardware, we discuss our software framework, vision processing, walking and game-play strategy methodology. Lastly we give an overview of future research interests within the team and intentions of future contributions for the league and the goal of RoboCup.

cs.RO

The Long Night: Modeling the Climate of Westeros

Many previous authors have attempted to find explanations for Westeros's climate, characterized by a generally moderate, Earth-like climate punctuated by extremely long and cold winters, separated by thousands of years. One explanation that has been proposed is that the planet orbits in a Sitnikov configuration, where two equal-mass stars (or a star and a black hole) orbit each other on slightly eccentric orbits, and the planet moves along a line through the barycenter perpendicular to the primaries' orbital plane (Freistetter & Gr\"utzbauch 2018). We modify an intermediate-complexity GCM to include the effects of such an orbit and integrate it for thousands of years to determine whether such an orbit can a) be habitable and b) explain the climatic variations observed by the inhabitants of Westeros, in both double-star and star-black hole configurations. While configurations with low primary eccentricity and initial conditions that permit only small excursions from the ecliptic plane are habitable, these orbits are too stable to explain Westerosi climate. We find that while orbits with more bounded chaos are able to produce rare anomalously long and cold winters similar to Westeros's Long Night, huge variations in incident stellar flux on normal orbital timescales should render these planets uninhabitable. We note that the presence of an orbital megastructure, either around the planet or the barycenter, could block some of the sunlight during crossings of the primaries' orbital plane and preserve Westeros's habitability. While we find that bounded chaotic Sitnikov orbits are a viable explanation for Westeros's Long Night, we propose that chaotic variations of the planet's axial tilt or semimajor axis, potentially due to torques from nearby planets or stars, may be a more realistic explanation than Sitnikov orbits.

physics.pop-ph

Distribution of mass of holomorphic cusp forms

We prove an upper bound for the L^4-norm and for the L^2-norm restricted to the vertical geodesic of a holomorphic Hecke cusp form of large weight. The method is based on Watson's formula and estimating a mean value of certain L-functions of degree 6. Further applications to restriction problems of Siegel modular forms and subconvexity bounds of degree 8 L-functions are given.

math.NT

Lower-Order Terms of the 1-Level Density of Families of Elliptic Curves

The Katz-Sarnak philosophy predicts that statistics of zeros of families of L-functions are strikingly universal. However, subtle arithmetical differences between families of the same symmetry type can be detected by calculating lower-order terms of the statistics of interest. In this paper we calculate lower-order terms of the 1-level density of some families of elliptic curves. We show that there are essentially two different effects on the distribution of low-lying zeros. First, low-lying zeros are more numerous in families of elliptic curves E with relatively large numbers of points (mod p). Second, and somewhat surprisingly, a family with a relatively large number of primes of bad reduction has relatively fewer low-lying zeros. We also show that the lower order term can grow arbitrarily large by taking a biased family with a relatively large number of points (mod p) for all small primes p.

math.NT