arXiv ScienceSearch

arXiv subjects

Ruodu Wang

Publications and source records attributed to Ruodu Wang.

At least 19 recordsLinked to original sources

Admissibility and Complete Classes for False Discovery Rate Control with E-values

The false discovery rate (FDR) is the most widely used error metric in modern multiple testing. We provide a comprehensive analysis of the admissibility of e-value-based procedures with FDR control. We consider both simultaneous and point procedures and introduce strong and weak notions of dominance. Every simultaneous procedure is strongly, and hence weakly, dominated by an admissible weighted-mean closed e-Benjamini--Hochberg ($\overline{\mathrm{eBH}}$) procedure, and thus weighted-mean $\overline{\mathrm{eBH}}$ procedures form a complete class. Every constant-free weighted-mean $\overline{\mathrm{eBH}}$ procedure is admissible at every level, and we propose two ways for choosing constant-free weights based on side information, such as pre-screening data. Within the symmetric class, the usual mean $\overline{\mathrm{eBH}}$ procedure is the largest element if and only if the FDR level is small enough; otherwise this class has no largest element. We also obtain results on the admissibility of symmetric weighted-mean $\overline{\mathrm{eBH}}$ procedures with non-zero constant terms. Point e-testing procedures have a parallel theory of admissibility. These results highlight the central role of weighted-mean $\overline{\mathrm{eBH}}$ procedures in multiple testing.

stat.ME

Disappointment Aversion and Expectiles

This paper recasts Gul's (1991) theory of disappointment aversion in a Savage framework, with general outcomes, new explicit axioms of disappointment aversion, and novel explicit representations. These permit broader applications of the theory and a better understanding of its decision-theoretic foundations. Our results exploit an unexpected connection between Gul's model and the econometric framework of Newey and Powell (1987) of asymmetric least squares estimation. Our main axiomatization result shows that a preference relation over Savage acts is probabilistically sophisticated, invariant biseparable, and disappointment hedging if and only if it admits a representation \emph{à la} Gul, and hence all explicit equivalent representations that we present in the paper. We also derive a neurocomputational foundation of the theory based on recent neuroscience findings and a novel reinforcement learning result.

econ.TH

Equilibrium in closed constant-function market maker economies

We study equilibria in a closed, fee-free constant-function market maker (CFMM) economy with two assets and two traders. An interior state is a unilateral no-trade equilibrium exactly when the CFMM marginal price equals both traders' marginal rates of substitution. For an interior initial state, individually rational unilateral equilibria are Pareto optimal relative to the fixed CFMM invariant. A weak representative agent is obtained at each fixed equilibrium by weighted sup-convolution, whereas a state-independent strong representative agent exists exactly when traders share a common homothetic preference. Every interior feasible state is reachable through finitely many valid trades, and alternating utility-maximizing trades converge to a Pareto optimal unilateral equilibrium. We also derive conditions under which trading order produces a first-mover advantage or disadvantage in the first round.

q-fin.TR

Eliciting reference measures of law-invariant functionals

Law-invariant functionals are central to risk management and assign identical values to random prospects sharing the same distribution under an atomless reference probability measure. This measure is typically assumed fixed. Here, we adopt the reverse perspective: given only observed functional values, we aim to either recover the reference measure or identify a candidate measure to test for law invariance when that property is not {\em a priori} satisfied. Our approach is based on a key observation about law-invariant functionals defined on law-invariant domains. These functionals define lower (upper) supporting sets in dual spaces of signed measures, and the suprema (infima) of these supporting sets---if they exist---are scalar multiples of the reference measure. In specific cases, this observation can be formulated as a sandwich theorem. We illustrate the methodology through a detailed analysis of prominent examples: the entropic risk measure, Expected Shortfall, and Value-at-Risk. For the latter, our elicitation procedure initially fails due to the triviality of supporting set extrema. We therefore develop a suitable modification.

q-fin.RM

Combining e-values using demi-supermartingales

We present a new method for combining e-variables through demi-supermartingales, which settles an old conjecture in the literature on nonparametric mean testing. It also provides an explicit concentration bound for a certain Kullback--Leibler-type statistic arising in the stochastic multi-armed bandit literature. All of these combination results hold for independent e-variables as well as for the class of co-valid e-variables, whose dependence structure lies somewhere between independence and sequential validity. The results are further generalized to compound e-variables. The proofs proceed by analyzing elementary symmetric polynomials and their behavior as nonnegative demi-supermartingales.

stat.ME

The exact dimensional threshold for Spearman rank-correlation compatibility

We show that, for a given dimension, the set of Spearman's rank correlation matrices and that of linear correlation matrices coincide if and only if the dimension is no larger than nine. For this, we construct an extreme rank-four counterexample in dimension ten and prove its incompatibility using moment identities and Cauchy-Schwarz. Appending unit directions produces counterexamples in every higher dimension. This, together with existing results, completes the dimensional classification and settles a long-standing open question in quantitative risk management.

math.ST

Gaffke's confidence interval for the mean of bounded data is inadmissible but asymptotically efficient

Given observations $\mathbf x=(x_1,\dots,x_n)$, Gaffke (2005) defined \[ K_n(\mathbf x)=\mathbb{P}_{\mathbf D}\!\left\{\sum_{i=1}^n x_iD_i\le 1\right\}, \qquad (D_0,D_1,\ldots,D_n)\sim\mathrm{Dirichlet}(1,\ldots,1), \] and conjectured that it is a $p$-value whenever the inputs are independent e-values. Recently, Vlassis and Thomas (2026) proved this conjecture. Inverting the tests for observations in $[0,1]$ gives the confidence interval studied by Learned-Miller and Thomas (2020), which reduces to Clopper--Pearson for Bernoulli data. We give a finite- and large-sample account of Gaffke's test and interval. First, for every $\mathbf x\in[0,\infty)^n$ and every elementary symmetric polynomial $e_k$, \( K_n(\mathbf x)e_k(\mathbf x)\le {n\choose k}, \) so the Gaffke $p$-value never larger than the SymPol $p$-value of Ming et al. (2026). However, Gaffke's p-value is inadmissible. For $n=2$, we construct a valid rule that is strictly smaller on mixed configurations and is the unique admissible rule that dominates $K_2$. A neutral-face extension proves inadmissibility of $K_n$ for every $n\ge2$. If one independent uniform random variable is allowed, there is an even simpler full-dimensional improvement: on the upper orthant, where $K_n(\mathbf x)=1/\prod_i x_i$, replace it by $U/\prod_i x_i$. The equal-tail Gaffke confidence interval $I_n$ is nevertheless first-order asymptotically efficient: for iid observations on $[0,1]$ with unknown variance $σ^2>0$, \[ \sqrt n\,\operatorname{Width}(I_n)\longrightarrow 2σz_{1-α/2}\qquad\text{almost surely}. \] Our simulations also find that, among a variety of bounded-mean intervals considered, the Gaffke interval is the shortest, including comparisons with a recent empirical Berry--Esseen procedure having the same first-order Gaussian target.

math.ST

Optimal risk sharing, equilibria, and welfare with empirically realistic risk attitudes

This paper examines optimal risk sharing. It brings in empirical realism, reckoning with the risk seeking found empirically. We provide results on Pareto optimality, competitive equilibria, utility frontiers, and the first and second theorems of welfare. Empirical studies have found prevailing risk seeking in several subdomains. Thus, as a first step to increase empirical realism, we allow for some risk-seeking agents, still assuming expected utility. Yet more empirical realism is obtained by generalizing expected utility and allowing agents' attitudes to combine risk aversion in some domains with risk seeking in others. We provide results and show directions for future research.

econ.TH

Probability of worthwhile effect of monotone-response treatments

Experiments may, by design, prevent one from observing on a single subject both the response to a treatment and to its absence. Because of this, marginal distributions for both cases may be observable but not their joint distribution, thus obscuring the distribution of the treatment effect. We examine the case where we impose that the treatment effect is nonnegative, also called monotone treatment response, a common assumption relevant to many practical applications. We solve the problems of best- and worst-case probabilities that the treatment effect exceeds a given value, using an explicit construction for the dependence scheme in each case. Such problems can equivalently be described, in different contexts, as risk aggregation under dependence uncertainty and an order constraint, and as optimal transport with a particular cost function.

econ.EM

Confidence intervals for causal effects in sequential decision making

We derive confidence intervals and confidence sequences for causal effects in situations where the back-door criterion is applicable. Our tightest confidence intervals hold in the standard setting where the training data consists of IID observations over a system described by a given causal diagram. When interventions are allowed to depend on the past data, our confidence intervals become wider and involve a term coming from the law of the iterated logarithm, even where the number of observations is known in advance. In the sequential setting where the number of observations is not given, our confidence intervals, arranged into a confidence sequence for causal effects, involve more iterated logarithm terms and become even wider.

math.ST

Universal Value-at-Risk superadditivity

Value-at-Risk (VaR) is a standard regulatory risk measure, and its failure of subadditivity is well known. Much less appreciated is that for sufficiently heavy-tailed losses, VaR can be superadditive uniformly across all probability levels, a phenomenon strictly stronger than the asymptotic superadditivity studied in extreme value theory. We call this property universal VaR superadditivity (UVS). We study UVS and its stronger weighted version (WUVS) as properties of random vectors rather than of marginal distributions. This perspective unifies and extends a recent line of work on iid infinite-mean models. UVS, except for trivial cases, imposes an infinite-mean structure. We establish preservation properties of UVS and WUVS under increasing and convex transformations, weak convergence, and certain distributional mixtures, and use these tools to prove UVS and WUVS for non-identically distributed risks in several large families including completely subscalable, super-Cauchy, and inverted subadditive risks, extending results previously available only in the iid case. In many results, we also establish strict versions of UVS and WUVS, which lead to stronger decision-theoretic implications. As a consequence, for any portfolio satisfying WUVS, every distortion risk measure is superadditive, so an optimal allocation concentrates on a single asset, and diversification is never beneficial.

q-fin.RM

Diamond transports in quadratic-form and distorted optimal transport

The diamond transport is generated by the uniform law on a diamond-shaped copula support. Since a classical optimal transport (OT) objective is affine in the coupling, this transport cannot be the unique minimizer in the classical setting. We study a broader family of transports, called diamond-type transports, in non-classical settings such as quadratic-form optimal transport (QOT) and distorted optimal transport (DOT), which are generally nonconvex. Our main results are within the QOT framework: for symmetric one-dimensional marginals, the diamond transport is an optimizer for a large class of QOT problems whose costs depend on within-coordinate distances. Examples include product costs under positive-definiteness and convexity conditions and, in particular, mixed rectangular costs. For rectangular costs, we show that the diamond transport is the unique minimizer except for boundary cases. In the DOT framework, diamond-type transports are minimizers for a natural class of cost, and the diamond transport is the unique minimizer in specialized examples. We also identify the intersection between DOT and QOT, which corresponds precisely to quadratic distortion functions.

math.OC

Validity and Power of Heavy-Tailed Combination Tests under Asymptotic Dependence

Heavy-tailed combination tests, such as the Cauchy combination test and harmonic mean p-value method, are widely used for testing global null hypotheses by aggregating dependent p-values. Existing theoretical guarantees, however, are largely restricted to the case of asymptotically independent p-values, leaving the behavior of these tests under broader dependence structures poorly understood. We develop a unified framework based on multivariate regularly varying copulas, a flexible class defined by a mild regularity condition on the joint behavior of p-values near zero, that accommodates a wide range of dependence structures. Within this framework, heavy-tailed combination tests are asymptotically valid when the transformation distribution has tail index $γ\leq 1$, with $γ= 1$ maximizing power while preserving validity. We further show that combination tests with $γ= 1$ achieve strictly greater asymptotic power than Bonferroni's method if and only if the p-values are not asymptotically independent and signals are not extremely sparse, with the power advantage growing as dependence strengthens. Bonferroni emerges as the $γ\to 0$ limit and becomes overly conservative under asymptotic dependence. These results provide theoretical support for using truncated Cauchy or Pareto combination tests, offering a principled approach to enhance power while controlling false positives under complex dependence.

math.ST

Quadratic-form Optimal Transport

We introduce the framework of quadratic-form optimal transport (QOT), whose transport cost has the form $\iint c\,\mathrm{d}π\otimes\mathrm{d}π$ for some coupling $π$ between two marginals. Interesting examples of quadratic-form transport cost and their optimization include inequality measurement, the variance of a bivariate function, covariance, Kendall's tau, the Gromov--Wasserstein distance, quadratic assignment problems, and quadratic regularization of classic optimal transport. QOT leads to substantially different mathematical structures compared to classic transport problems and many technical challenges. We illustrate the fundamental properties of QOT and provide several cases where explicit solutions are obtained. For a wide class of cost functions, including the rectangular cost functions, the QOT problem is solved by a new coupling called the diamond transport, whose copula is supported on a diamond in the unit square.

math.PR

Adaptive Window Selection for Financial Risk Forecasting

Risk forecasts in financial regulation and internal management are calculated through historical data. The unknown structural changes of financial data pose a substantial challenge in selecting an appropriate look-back window for risk modeling and forecasting. We develop a data-driven online learning method, called the bootstrap-based adaptive window selection (BAWS), that adaptively determines the window size in a sequential manner. A central component of BAWS is to compare the realized scores against a data-dependent threshold based on the bootstrap method. We provide an asymptotic justification for the bootstrap threshold, covering non-smooth scores such as the VaR check loss and the joint VaR--ES score, with an extension to stationary weakly dependent data via the moving block bootstrap. A single-break analysis further shows that BAWS rejects overlong windows crossing sufficiently large breaks. The proposed method is applicable to the forecasting of risk measures that are elicitable individually or jointly, such as the Value-at-Risk (VaR) and the pair of VaR and the corresponding Expected Shortfall. Through simulation studies and an empirical analysis, we demonstrate that BAWS often improves upon the standard rolling window approach and the recently developed method of stability-based adaptive window selection, especially when there are structural changes in the data-generating process.

q-fin.RM

Online monotone density estimation and log-optimal calibration

We study the problem of online monotone density estimation, where density estimators must be constructed in a predictable manner from sequentially observed data. We propose two online estimators: an online analogue of the classical Grenander estimator, and an expert aggregation estimator inspired by exponential weighting methods from the online learning literature. In the well-specified stochastic setting, where the underlying density is monotone, we show that the expected cumulative log-likelihood gap between the online estimators and the true density admits an $O(n^{1/3})$ bound. We further establish a $\sqrt{n\log{n}}$ pathwise regret bound for the expert aggregation estimator relative to the best offline monotone estimator chosen in hindsight, under minimal regularity assumptions on the observed sequence. As an application of independent interest, we show that the problem of constructing log-optimal p-to-e calibrators for sequential hypothesis testing can be formulated as an online monotone density estimation problem. We adapt the proposed estimators to build empirically adaptive p-to-e calibrators and establish their optimality. Numerical experiments illustrate the theoretical results.

stat.ML

Newsvendor under Ambiguity and Misspecification

Problem definition: We consider a newsvendor problem with unknown demand distribution, where we distinguish ambiguity under which the newsvendor does not differentiate demand distributions of common characteristics and misspecification under which such characteristics might be misspecified. Methodology/results: The newsvendor hedges against ambiguity and misspecification by maximizing the worst-case expected profit regularized by a distribution's distance to an ambiguity set. Focusing on the popular mean-variance ambiguity set and optimal-transport cost for the misspecification, we show that the decision criterion of misspecification aversion possesses insightful interpretations as distributional transforms. We derive the closed-form optimal order quantity that generalizes the solution of the Scarf model under only ambiguity aversion. We establish the finite-sample performance guarantee, which consists of two parts: in-sample optimal value and out-of-sample effect of misspecification that can be further decoupled into estimation error and distribution shift. We also extend the framework to multiple products, distributional characteristics specified via optimal transport, and misspecification measured by total variation distance. Managerial implications: The closed-form solution highlights the impact of misspecification aversion: the optimal order quantity under misspecification aversion can decrease as the price or variance increases, reversing the monotonicity of that under only ambiguity aversion. Hence, ambiguity and misspecification, as different layers of distributional uncertainty, can result in distinct operational consequences. The finite-sample performance guarantee theoretically justifies the necessity of incorporating misspecification aversion in a non-stationary environment, which is also well demonstrated in our experiments with real-world data.

math.OC

E-backtesting

In the recent Basel Accords, the Expected Shortfall (ES) replaces the Value-at-Risk (VaR) as the standard risk measure for market risk in the banking sector, making it the most important risk measure in financial regulation. One of the most challenging tasks in risk modeling practice is to backtest ES forecasts provided by financial institutions. To design a model-free backtesting procedure for ES, we make use of the recently developed techniques of e-values and e-processes. Backtest e-statistics are introduced to formulate e-processes for risk measure forecasts, and unique forms of backtest e-statistics for VaR and ES are characterized using recent results on identification functions. For a given backtest e-statistic, a few criteria for optimally constructing the e-processes are studied. The proposed method can be naturally applied to many other risk measures and statistical quantities. We conduct extensive simulation studies and data analysis to illustrate the advantages of the model-free backtesting method, and compare it with the ones in the literature.

q-fin.RM