arXiv ScienceSearch

arXiv subjects

Eliuvish Han Cui

Publications and source records attributed to Eliuvish Han Cui.

4 recordsLinked to original sources

Spending Scarce Confirmatory PET Measurements: Target-Aligned Validation in A4/LEARN

Anti-amyloid therapies and blood-based biomarkers are changing Alzheimer disease workups into a two-stage measurement workflow: screen broadly with cheaper information, then spend scarce confirmatory amyloid measurements where they support the decision that will be reported. Amyloid positron-emission tomography (PET) remains one such protocol measurement for amyloid burden, but PET slots, trial budgets, and payer-facing evidence packages are finite. This paper asks a deliberately operational question: when is simple transparent PET validation enough, and when is a fitted residual-uncertainty score worth the added complexity? For a weighted protocol target, the first-order value of validating subject i is the product of target influence and residual protocol uncertainty. Generic uncertainty sampling uses only the second factor and can spend PET measurements on subjects that are hard to predict but weak for the scientific, clinical, or commercial claim. We apply this rule to the A4/LEARN PET archive, treating observed PET as a design laboratory for scarce-confirmation studies. For the primary APOE4 carrier versus non-carrier contrast in Centiloid 24-or-higher PET positivity, simple APOE4-balanced validation recovers nearly all of the target-specific gain: at PET budget 200, the confidence-interval width ratio relative to random validation is 0.923 for APOE4 balancing and 0.914 for target-specific scoring, while generic uncertainty sampling is 0.980. Other targets behave differently: target-specific scoring gives larger gains for an age-slope analysis and for cutoff-indexed PET positivity. The practical message is simple: spend scarce protocol measurements according to the claim being validated, not only according to prediction uncertainty.

stat.AP

The Statistical Compass

This monograph develops probability and stochastic-process ideas as a translation language for statistics: from designed observations and data objects to targets, stability statements, inference, and use. The chapters move from motivating examples and randomization through probability measures, kernels, likelihoods, data objects, weak convergence, empirical fields, functional data, M- and Z-estimation, testing, local approximations, event-time processes, and prediction. Historical and biomedical examples are used to keep abstract objects tied to records, mechanisms, and decisions. The aim is to give readers a common grammar for classical probability, modern data structures, and statistical practice.

stat.AP

Brownian Motion with a Pulse: A Biostatistician's Guide to Diffusions, Bridges, Functional PCA, and First-Passage Models

Brownian motion is a compact mathematical language for continuous-time uncertainty in biostatistics. This tutorial develops the process from construction and path properties to tools that recur in applied biomedical work: the Markov and strong Markov properties, the Karhunen-Loeve expansion, functional principal component analysis (Functional PCA), reflection principles, local time, stochastic differential equations (SDEs), Brownian bridges, and empirical-process limits. The applications emphasize longitudinal biomarkers, degradation modelling, first-passage endpoints, dynamic frailty, group-sequential monitoring, calibration diagnostics, recurrent-event processes, electronic health records, and wearable streams. A short cross-domain section uses literary and historical archives to make Brownian-bridge thinking concrete without shifting the paper away from biostatistics, and includes a reproducible chapter-level experiment on Frankenstein. The Black-Merton-Scholes model is included as a solved SDE template, not as a finance application in its own right. The aim is to connect rigorous probability with modelling decisions faced by biostatisticians when biological processes evolve between noisy observation times.

stat.AP

From Risk Sets to Martingales: A Counting-Process Framework for Event-History Learning

Counting-process notation separates predictable risk-set information from observed event jumps through decompositions of the form dN(t)=Y(t)alpha(t)dt+dM(t). This article develops a unified event-history learning framework for censored, truncated, recurrent, multistate, and covariate-dependent data. Rather than cataloguing survival methods, the treatment translates each partially observed learning target into five recurring objects: risk process, jump process, compensator, estimating equation, and limiting argument. The framework connects right-censored survival curves, product-integral estimators, bivariate and interval-censored survival estimators, log-rank tests, Cox-Andersen-Gill regression, additive hazards, accelerated failure-time models, panel-count data, landmark prediction, semi-Markov models, Bayesian nonparametric transition models, and instrumental-variable methods. The original contribution is threefold. Computationally, the article turns risk-set sweeps, product-integral updates, interval-likelihood calculations, semi-Markov elapsed-time bookkeeping, Bayesian transition-hazard updating, and cross-fitted validation into reusable algorithms and simulation diagnostics. Theoretically, it gives proof templates for the recurring martingale, likelihood, product-integral, and empirical-process arguments, and proves a new out-of-fold compensator validation identity for cross-fitted censored learners. For applications, it maps biomedical, reliability, operational, economic, financial, literary, historical, and causal survival examples onto the same risk-set and compensator language. The resulting account provides a common mathematical language for deriving, checking, and comparing classical and machine-learning methods for censored, recurrent, and multistate event-history data.

stat.AP