arXiv Science⌕ Search

arXiv · 2610.09658

Sharp Partial Identification for Survival Model Comparison Without Target Outcomes

Abstract

We compare two locked survival prediction models in a target population before target survival outcomes are available. At a prespecified horizon, the estimand is the target Brier-risk contrast. Under a bounded conditional log-odds shift model on a prespecified deployment summary, we derive a sharp identified set preserving the shared unidentified target outcome law. Direct identification is never wider than separately identifying the risks and subtracting their bounds, with strict tightening under a Brier-specific same-side-1/2 condition. For right-censored source data, conditional Cox censoring estimation, inverse-probability-of-censoring-weighted logistic outcome modeling, and a joint pairs bootstrap yield simultaneous confidence envelopes over a finite sensitivity grid. In simulations, separate-to-direct width ratios ranged from 1.00 to 5.73 across controlled prediction geometries. Targeted simulations showed finite-sample undercoverage of the outer envelope at the small, heavily censored non-small-cell lung cancer (NSCLC) information scale (0.847-0.861 versus 0.95 nominal), compared with 0.946 at the Rotterdam-GBSG scale. In the cross-institutional NSCLC application, all 40 prespecified evaluations resulted in DEFER despite reduced identification uncertainty. In a supporting Rotterdam-to-GBSG analysis, candidate superiority was certified under small sensitivity allowances; one locked configuration yielded ADOPT CANDIDATE under direct identification but DEFER under separate-risk subtraction. Direct identification can materially reduce identification uncertainty and change the operational conclusion when signal and sampling precision are sufficient, while retaining DEFER when directional certification is unsupported.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Won Gi Choi, Sun-Ho Kim, Min Soo Kim. 2026-10-07. Sharp Partial Identification for Survival Model Comparison Without Target Outcomes. https://arxiv.org/abs/2610.09658

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

Robust Bayesian methods using amortized simulation-based inference

Bayesian simulation-based inference (SBI) methods are used in statistical models where simulation is feasible but the likelihood is intractable. Standard SBI methods can perform poorly in cases of model misspecification, and there has been much recent work on modified SBI approaches which are robust to misspecified likelihoods. However, less attention has been given to the issue of inappropriate prior specification, which is the focus of this work. In conventional Bayesian modelling, there will often be a wide range of prior distributions consistent with limited prior knowledge expressed by an expert. Choosing a single prior can lead to an inappropriate choice, possibly conflicting with the likelihood information. Robust Bayesian methods, where a class of priors is considered instead of a single prior, can address this issue. For each density in the prior class, a posterior can be computed, and the range of the resulting inferences is informative about posterior sensitivity to the prior imprecision. We consider density ratio classes for the prior and implement robust Bayesian SBI using amortized neural methods developed recently in the literature. We also discuss methods for checking for conflict between a density ratio class of priors and the likelihood, and sequential updating methods for examining conflict between different groups of summary statistics. The methods are illustrated for several simulated and real examples.

stat.ME↗

On the permutation equivariance principle for causal estimands

In many causal inference problems, multiple action variables share a common causal role yet lack a natural ordering. \revblue{We consider $K\geq2$ action variables, each evaluated under treatment or control conditions,} and formalize permutation equivariance, the principle that permuting the variables permutes the corresponding estimands in a trackable manner, hence preserving their scientific meaning. We characterize this principle algebraically and present a complete class of weighted permutation equivariant estimands capturing main effects and interactions of all orders. We discuss the interpretation and choice of weights and characterize residual-free estimands, whose inclusion--exclusion sum recovers the endpoint contrast between the all-treated and all-control configurations. \revblue{Applying our general framework to network interference yields a new hierarchy of direct, indirect, and overall effects, whose first-order aggregates recover the average effects of \citet{hu2022average}. We also identify the overall effects as mixed derivatives of expected welfare under independent Bernoulli assignment.} We illustrate the framework through the contexts of factorial studies, causal mediation, and network interference.

stat.ME↗

Continuous mixtures of Gaussian processes as models for spatial extremes

Spatial modelling of extreme values allows studying the risk of joint occurrence of extreme events at different locations and is of significant interest in climatic and other environmental sciences. A popular class of dependence models for spatial extremes is that of random location-scale mixtures, in which a spatial "baseline" process is multiplied or shifted by a random variable, potentially altering its extremal dependence behaviour. Gaussian location-scale mixtures retain benefits of their Gaussian baseline processes while overcoming some of their limitations, such as symmetry, light tails and weak tail dependence. We review properties of Gaussian location-scale mixtures and develop novel constructions with interesting features, together with a general algorithm for conditional simulation from these models. We leverage their flexibility to propose extended extreme-value models, that allow for appropriately modelling not only the tails but also the bulk of the data. This is important in many applications and avoids the need to explicitly select the events considered as extreme. We propose new solutions for likelihood inference in parametric models of Gaussian location-scale mixtures, in order to avoid the numerical bottleneck given by the latent location and scale variables that can lead to high computational cost of standard likelihood evaluations. The effectiveness of the models and of the inference methods is confirmed with simulated data examples, and we present an application to wildfire-related weather variables in Portugal. Although not detailed here, the approaches would also be straightforward to use for modelling multivariate (non spatial) data.

stat.ME↗