arXiv Science⌕ Search

arXiv · 2610.08194

The Noise Is the Signal: Correlated Sampling Error Is Rank-Informative for Proxy Metric Selection

Abstract

North-star metrics such as customer lifetime value are often too slow and noisy to decide a short A/B test. Teams therefore rely on a proxy metric, commonly chosen by how closely its effects tracked the north star's across past experiments. Validating that choice, or any method for making it, is hard: the only benchmark is the noisy north star, and the number of available past experiments is limited. In addition, proxy and north-star effects are estimated on the same customers, so their sampling errors are correlated. Recent work at major experimentation platforms removes this shared error as contamination, improving estimates of the true-effect covariance. Choosing a proxy, however, is a ranking problem, and a better estimate need not give a better ranking. We measure agreement free of shared error by estimating the two effects on disjoint random halves of each experiment's customers. In an archive of 262 experiments and 69 candidate proxies, the shared error ranks the candidates in a similar order to this agreement (Spearman correlation 0.65): it carries information about proxy quality. The more of it a correction removes, the worse the ranking because removal discards part of the signal but leaves the main sources of ranking noise, the noisy north star and the limited number of experiments, untouched. Archive-calibrated simulations, in which the correct ranking is known, confirm this even when every correction receives the true sampling covariance. Held-out real experiments, evaluated on disjoint customer halves so that shared error cannot bias the comparison, closely reproduce the predicted ordering (Spearman correlation 0.93). Correction can still pay off with more experiments, but the number needed rises steeply with the north star's noise. We map this crossover and give platform teams three inexpensive checks for deciding from their own archive whether and how strongly to correct.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Sandro Provenzano. 2026-10-06. The Noise Is the Signal: Correlated Sampling Error Is Rank-Informative for Proxy Metric Selection. https://arxiv.org/abs/2610.08194

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

Identifying treatment effects on categorical outcomes in IV models

This paper studies population treatment effects when the outcome is unordered categorical, the treatment is binary, and the instrument is binary. I introduce an assumption called association similarity. For each instrument value, association similarity requires the odds-ratio association between potential treatment and potential outcome to be the same whether the outcome is evaluated under treatment or under no treatment. Under association similarity and a testable categorywise dominance condition, the full potential-outcome distributions are point identified. The identification is constructive and allows estimation using sample frequencies. I also consider weaker restrictions on departures from association similarity that lead to sharp partial identification. I illustrate the method with an application about the effect of insurance on health outcomes.

econ.EM↗

Robust Inference for Dyadic Data with Spatially Dependent Nodes

We develop inference for complete dyadic samples with spatial dependence governed by geographic distances between nodes, covering settings beyond the scope of existing dyadic inference theory. We propose spatial variance and corrected block jackknife estimators consistent in nondegenerate and degenerate Gaussian cases. Under degeneracy, spatial subsampling consistently estimates possibly non-Gaussian limits. Combining the jackknife with subsampling yields the max jackknifes-ubsampling (MJS) interval, which provides pointwise asymptotically exact coverage in both Gaussian cases and conservative coverage in the specified non-Gaussian case. Fixed-effect extensions show that estimating node effects can change the fast limiting distribution. Simulations and an empirical illustration are provided.

econ.EM↗

Vine Copula VAR:From Recursive Margins to Joint Forecast Inference

Joint-event forecasts often combine a dependence estimate based on past forecast errors with newly estimated marginal distributions. When each historical error retains the marginal fit available at its issue date, inference must account for an overlapping sequence of estimation errors. We derive their joint influence with the terminal forecast estimates in a stable Vine Copula VAR with normal innovation margins and a fixed, correctly specified Gaussian or positive Clayton vine. An intercept identity and the stable VAR filter reduce the historical correction to harmonically weighted innovation moments, while terminal slope uncertainty remains. The resulting covariance estimator gives asymptotically valid repeated-sample intervals for fixed one-sided event probabilities at the realized forecast state. In the Gaussian submodel, retaining issued transforms adds a positive semidefinite covariance term relative to refitting margins on the same observations. Monte Carlo simulations show that terminal-margin uncertainty is quantitatively more important than this additional term and that logit intervals improve lower-tail coverage in the designs studied. A real-time forecasting application to U.S. macroeconomic releases shows how marginal estimation contributes to uncertainty in predicted probabilities of joint contractions and identifies limitations of the stationary marginal model.

econ.EM↗