arXiv ScienceSearch

arXiv subjects

Larry Goldstein

Publications and source records attributed to Larry Goldstein.

At least 19 recordsLinked to original sources

Bias and Division in the Free World

Sampling bias is a foundational concept in statistics; associated bias transforms, such as size bias, have come to play important roles in probability theory of late. The first author and G. Reinert introduced zero bias, a transform whose unique fixed point is the normal distribution; it has become a standard tool in Stein's method and Gaussian approximation. Very recently, connections between zero bias and the class of infinitely divisible distributions have been found. In this paper, we develop a free probabilistic analog of the zero bias transform, proving its existence and regularity. The free zero bias has the semicircle law (free probability's central limit distribution) as its unique fixed point. We offer a construction of the free zero bias that mirrors a classical one incorporating square bias with a mollifier, and in the process develop a surprisingly new class of distributional operations through their Cauchy transforms. We then explore connections between the free zero bias, and size bias, with the class of freely infinitely divisible distributions. We develop a new self-contained treatment of the subject, together with a new characterization of free infinite divisibility using bias transforms. We also develop a parallel treatment of positively freely infinitely divisible distributions, which can also be characterized by a new kind of Levy--Khintchine formula that has no known classical analogue, and we use this to both give several new descriptions of such distributions and furnish new examples using these bias methods.

math.PR

Gaussian random field approximation via Stein's method with applications to wide random neural networks

We derive upper bounds on the Wasserstein distance ($W_1$), with respect to $\sup$-norm, between any continuous $\mathbb{R}^d$ valued random field indexed by the $n$-sphere and the Gaussian, based on Stein's method. We develop a novel Gaussian smoothing technique that allows us to transfer a bound in a smoother metric to the $W_1$ distance. The smoothing is based on covariance functions constructed using powers of Laplacian operators, designed so that the associated Gaussian process has a tractable Cameron-Martin or Reproducing Kernel Hilbert Space. This feature enables us to move beyond one dimensional interval-based index sets that were previously considered in the literature. Specializing our general result, we obtain the first bounds on the Gaussian random field approximation of wide random neural networks of any depth and Lipschitz activation functions at the random field level. Our bounds are explicitly expressed in terms of the widths of the network and moments of the random weights. We also obtain tighter bounds when the activation function has three bounded derivatives.

math.PR

Zero Bias Enhanced Stein Couplings

The Stein couplings of Chen and Roellin (2010) vastly expanded the range of applications for which coupling constructions in Stein's method for normal approximation could be applied, and subsumed both Stein's classical exchangeable pair, as well as the size bias coupling. A further simple generalization includes zero bias couplings, and also allows for situations where the coupling is not exact. The zero bias versions result in bounds for which often tedious computations of a variance of a conditional expectation is not required. An example to the Lightbulb process shows that even though the method may be simple to apply, it may yield improvements over previous results that had achieved bounds with optimal rates and small, explicit constants.

math.PR

The Game of Poker Chips, Dominoes and Survival

The Game of Poker Chips, Dominoes and Survival fosters team building and high level cooperation in large groups, and is a tool applied in management training exercises. Each player, initially given two colored poker chips, is allowed to make exchanges with the game coordinator according to two rules, and must secure a domino before time is called in order to `survive'. Though the rules are simple, it is not evident by their form that the survival of the entire group requires that they cooperate at a high level. From the point of view of the game coordinator, the difficulty of the game for the group can be controlled not only by the time limit, but also by the initial distribution of chips, in a way we make precise by a time complexity type argument. That analysis also provides insight into good strategies for group survival, those taking the least amount of time. In addition, coordinators may also want to be aware of when the game is `solvable', that is, when their initial distribution of chips permits the survival of all group members if given sufficient time to make exchanges. It turns out that the game is solvable if and only if the initial distribution contains seven chips that have one of two particular color distributions. In addition to being a lively game to play in management training or classroom settings, the analysis of the game after play can make for an engaging exercise in any basic discrete mathematics course to give a basic introduction to elements of game theory, logical reasoning, number theory and the computation of algorithmic complexities.

math.CO

Stein's method via induction

Applying an inductive technique for Stein and zero bias couplings yields Berry-Esseen theorems for normal approximation for two new examples. The conditions of the main results do not require that the couplings be bounded. Our two applications, one to the Erdős-Rényi, random graph with a fixed number of edges, and one to Jack measure on tableaux, demonstrate that the method can handle non-bounded variables with non-trivial global dependence, and can produce bounds in the Kolmogorov metric with the optimal rate.

math.PR

Relaxing the Gaussian assumption in Shrinkage and SURE in high dimension

Shrinkage estimation is a fundamental tool of modern statistics, pioneered by Charles Stein upon his discovery of the famous paradox involving the multivariate Gaussian. A large portion of the subsequent literature only considers the efficiency of shrinkage, and that of an associated procedure known as Stein's Unbiased Risk Estimate, or SURE, in the Gaussian setting of that original work. We investigate what extensions to the domain of validity of shrinkage and SURE can be made away from the Gaussian through the use of tools developed in the probabilistic area now known as Stein's method. We show that shrinkage is efficient away from the Gaussian under very mild conditions on the distribution of the noise. SURE is also proved to be adaptive under similar assumptions, and in particular in a way that retains the classical asymptotics of Pinsker's theorem. Notably, shrinkage and SURE are shown to be efficient under mild distributional assumptions, and particularly for general isotropic log-concave measures.

math.ST

$M$-estimation and deconvolution in a diffusion model with application to biosensor transdermal blood alcohol monitoring

We develop $M$-estimation and deconvolution methodology with the goal of making well-founded statistical inference on an individual's blood alcohol level based on noisy measurements of their skin alcohol content. We first apply our results to a nonlinear least squares estimator of the key parameter that specifies the blood/skin alcohol relation in a diffusion model, and establish its existence, consistency, and asymptotic normality. To make inference on the unknown underlying blood alchohol curve, we develop a basis space deconvolution approach with regulazation, and determine the asymptotic distribution of the error process, thus allowing us to compute uniform confidence bands on the curve. Simulation studies show agreement between the performance of our curve estimators and their asymptotic distributions at low noise levels, and we apply our methods to a real skin alcohol data set collected via a transdermal biosensor.

stat.AP

A Berry-Esseen bound for the uniform multinomial occupancy model

The inductive size bias coupling technique and Stein's method yield a Berry-Esseen theorem for the number of urns having occupancy $d \ge 2$ when $n$ balls are uniformly distributed over $m$ urns. In particular, there exists a constant $C$ depending only on $d$ such that $$ \sup_{z \in \mathbb{R}}|P(W_{n,m} \le z) -P(Z \le z)| \le C \left( \frac{1+(\frac{n}{m})^3}{σ_{n,m}} \right) \quad \mbox{for all $n \ge d$ and $m \ge 2$,} $$ where $W_{n,m}$ and $σ_{n,m}^2$ are the standardized count and variance, respectively, of the number of urns with $d$ balls, and $Z$ is a standard normal random variable. Asymptotically, the bound is optimal up to constants if $n$ and $m$ tend to infinity together in a way such that $n/m$ stays bounded.

math.PR

Dickman approximation in simulation, summations and perpetuities

The generalized Dickman distribution ${\cal D}_θ$ with parameter $θ>0$ is the unique solution to the distributional equality $W=_d W^*$, where \begin{eqnarray} W^*=_d U^{1/θ}(W+1) \qquad (1) \end{eqnarray} with $W$ non-negative with probability one, $U \sim {\cal U}[0,1]$ independent of $W$, and $=_d$ denoting equality in distribution. Members of this family appear in number theory, stochastic geometry, perpetuities and the study of algorithms. We obtain bounds in Wasserstein type distances between ${\cal D}_θ$ and \begin{eqnarray} W_n= \frac{1}{n} \sum_{i=1}^n Y_k B_k \qquad (2) \end{eqnarray} where $B_1,\ldots,B_n, Y_1, \ldots, Y_n$ are independent with $B_k \sim {\rm Ber}(1/k), E[Y_k]=k, {\rm Var}(Y_k)=σ_k^2$ and provide an application to the minimal directed spanning tree in $\mathbb{R}^2$, and also obtain such bounds when the Bernoulli variables in $(2)$ are replaced by Poissons. We also give simple proofs and provide bounds with optimal rates for the Dickman convergence of the weighted sums, arising in probabilistic number theory, of the form \begin{eqnarray} S_n=\frac{1}{\log(p_n)} \sum_{k=1}^n X_k \log(p_k) \end{eqnarray} where $(p_k)_{k \ge 1}$ is an enumeration of the prime numbers in increasing order and $X_k$ is Geometric with parameter $(1-1/p_k)$, Bernoulli with success probability $1/(1+p_k)$ or Poisson with mean $λ_k$. In addition, we broaden the class of generalized Dickman distributions by studying the fixed points of the transformation \begin{eqnarray*} s(W^*)=_d U^{1/θ}s(W+1) \end{eqnarray*} generalizing $(1)$, that allows the use of non-identity utility functions $s(\cdot)$ in Vervaat perpetuities. We obtain distributional bounds for recursive methods that can be used to simulate from this family.

math.PR

Size bias for one and all

Size bias occurs famously in waiting-time paradoxes, undesirably in sampling schemes, and unexpectedly in connection with Stein's method, tightness, analysis of the lognormal distribution, Skorohod embedding, infinite divisibility, and number theory. In this paper we review the basics and survey some of these unexpected connections.

math.PR

A BKR operation for events occurring for disjoint reasons with high probability

Given events $A$ and $B$ on a product space $S=\prod_{i=1}^n S_i$, the set $A \Box B$ consists of all vectors ${\bf x}=(x_1,\ldots,x_n) \in S$ for which there exist disjoint coordinate subsets $K$ and $L$ of $\{1,\ldots,n\}$ such that given the coordinates $x_i, i \in K$ one has that ${\bf x} \in A$ regardless of the values of ${\bf x}$ on the remaining coordinates, and likewise that ${\bf x} \in B$ given the coordinates {$x_j, j \in L$}. For a finite product of discrete spaces endowed with a product measure, the BKR inequality $$ P(A \Box B) \le P(A)P(B) \quad (1) $$ was conjectured by van den Berg and Kesten [3] and proved by Reimer [13]. In [7] inequality (1) was extended to general product probability spaces, replacing $A \Box B$ by the set $A \Box_{11} B$ consisting of those outcomes ${\bf x}$ which only assure with probability one that ${\bf x} \in A$ and ${\bf x} \in B$ based only on the revealed coordinates in $K$ and $L$ as above. A strengthening of the original BKR inequality (1) results, due to the fact that $A \Box B \subseteq A \Box_{11} B$. In particular, it may be the case that $A \Box B$ is empty, while $A \Box_{11} B$ is not. We propose the further extension $A \Box_{st} B$ depending on probability thresholds $s$ and $t$, where $A \Box_{11} B$ is the special case where both $s$ and $t$ take the value one. The outcomes ${\bf x}$ in $A \Box_{st} B$ are those for which disjoint sets of coordinates $K$ and $L$ exist such that given the values of $\bf x$ on the revealed set of coordinates $K$, the probability that $A$ occurs is at least $s$, and given the coordinates of $\bf x$ in $L$, the probability of $B$ is at least $t$. We provide simple examples that illustrate the utility of these extensions.

math.PR

Non-Gaussian Observations in Nonlinear Compressed Sensing via Stein Discrepancies

Performance guarantees for compression in nonlinear models under non-Gaussian observations can be achieved through the use of distributional characteristics that are sensitive to the distance to normality, and which in particular return the value of zero under Gaussian or linear sensing. The use of these characteristics, or discrepancies, improves some previous results in this area by relaxing conditions and tightening performance bounds. In addition, these characteristics are tractable to compute when Gaussian sensing is corrupted by either additive errors or mixing.

math.ST

Stein's method for positively associated random variables with applications to the Ising and voter models, bond percolation, and contact process

We provide non-asymptotic $L^1$ bounds to the normal for four well-known models in statistical physics and particle systems in $\mathbb{Z}^d$; the ferromagnetic nearest-neighbor Ising model, the supercritical bond percolation model, the voter model and the contact process. In the Ising model, we obtain an $L^1$ distance bound between the total magnetization and the normal distribution at any temperature when the magnetic moment parameter is nonzero, and when the inverse temperature is below critical and the magnetic moment parameter is zero. In the percolation model we obtain such a bound for the total number of points in a finite region belonging to an infinite cluster in dimensions $d \ge 2$, in the voter model for the occupation time of the origin in dimensions $d \ge 7$, and for finite time integrals of non-constant increasing cylindrical functions evaluated on the one dimensional supercritical contact process started in its unique invariant distribution. The tool developed for these purposes is a version of Stein's method adapted to positively associated random variables. In one dimension, letting $\boldsymbolξ=(ξ_1,\ldots,ξ_m)$ be a positively associated mean zero random vector with components that obey the bound $|ξ_i| \le B, i=1,\ldots,m$, and whose sum $W = \sum_{i=1}^m ξ_i$ has variance 1, it holds that $$ d_1 \left(\mathcal{L}(W),\mathcal{L}(Z) \right) \leq 5B + \sqrt{\frac{8}π}\sum_{i \neq j} \mathbb{E}[ξ_i ξ_j] $$ where $Z$ has the standard normal distribution and $d_1(\cdot,\cdot)$ is the $L^1$ metric. Our methods apply in the multidimensional case with the $L^1$ metric replaced by a smooth function metric.

math.PR

Bounded size biased couplings, log concave distributions and concentration of measure for occupancy models

Threshold-type counts based on multivariate occupancy models with log concave marginals admit bounded size biased couplings under weak conditions, leading to new concentration of measure results for random graphs, germ-grain models in stochastic geometry, multinomial allocation models and multivariate hypergeometric sampling. The work generalizes and improves upon previous results in a number of directions.

math.PR

Bounds to the normal for proximity region graphs

In a proximity region graph ${\cal G}$ in $\mathbb{R}^d$, two distinct points $x,y$ of a point process $μ$ are connected when the 'forbidden region' $S(x,y)$ these points determine has empty intersection with $μ$. The Gabriel graph, where $S(x,y)$ is the open disc with diameter the line segment connecting $x$ and $y$, is one canonical example. When $μ$ is a Poisson or binomial process, under broad conditions on the regions $S(x,y)$, bounds on the Kolmogorov and Wasserstein distances to the normal are produced for functionals of ${\cal G}$, including the total number of edges and the total length. Variance lower bounds, not requiring strong stabilization, are also proven to hold for a class of such functionals.

math.PR

Size biased couplings and the spectral gap for random regular graphs

Let $λ$ be the second largest eigenvalue in absolute value of a uniform random $d$-regular graph on $n$ vertices. It was famously conjectured by Alon and proved by Friedman that if $d$ is fixed independent of $n$, then $λ=2\sqrt{d-1} +o(1)$ with high probability. In the present work we show that $λ=O(\sqrt{d})$ continues to hold with high probability as long as $d=O(n^{2/3})$, making progress towards a conjecture of Vu that the bound holds for all $1\le d\le n/2$. Prior to this work the best result was obtained by Broder, Frieze, Suen and Upfal (1999) using the configuration model, which hits a barrier at $d=o(n^{1/2})$. We are able to go beyond this barrier by proving concentration of measure results directly for the uniform distribution on $d$-regular graphs. These come as consequences of advances we make in the theory of concentration by size biased couplings. Specifically, we obtain Bennett-type tail estimates for random variables admitting certain unbounded size biased couplings.

math.PR

Non asymptotic distributional bounds for the Dickman Approximation of the running time of the Quickselect algorithm

Given a non-negative random variable $W$ and $\theta>0$, let the generalized Dickman transformation map the distribution of $W$ to that of $$ W^*=_d U^{1/\theta}(W+1), $$ where $U \sim {\cal U}[0,1]$, a uniformly distributed variable on the unit interval, independent of $W$, and where $=_d$ denotes equality in distribution. It is well known that $W^*$ and $W$ are equal in distribution if and only if $W$ has the generalized Dickman distribution ${\cal D}_\theta$. We demonstrate that the Wasserstein distance $d_1$ between $W$, a non-negative random variable with finite mean, and $D_\theta$ having distribution ${\cal D}_\theta$ obeys the inequality $$ d_1(W,D_\theta) \le (1+\theta)d_1(W,W^*). $$ The specialization of this bound to the case $\theta=1$ and coupling constructions yield $$ d_1(W_{n,1},D) \le \frac{8\log (n/2)+10}{n} \quad \mbox{for all $n \ge 1$, where} \quad W_{n,1}=\frac{1}{n}C_{n,1}-1, $$ and $C_{n,m}$ is the number of comparisons made by the Quickselect algorithm to find the $m^{th}$ smallest element of a list of $n$ distinct numbers. A similar bound holds for $m \ge 2$, and together recover the results of [12] that show distributional convergence of $W_n$ to the standard Dickman distribution in the asymptotic regime $m=o(n)$. By developing an exact expression for the expected running time $E[C_{n,m}]$, lower bounds are provided that show the rate is not improvable for all $m \not = 2$.

math.PR

On Strong Embeddings by Stein's Method

Strong embeddings, that is, couplings between a partial sum process of a sequence of random variables and a Brownian motion, have found numerous applications in probability and statistics. We extend Chatterjee's novel use of Stein's method for $\{-1,+1\}$ valued variables to a general class of discrete distributions, and provide $\log n$ rates for the coupling of partial sums of independent variables to a Brownian motion, and results for coupling sums of suitably standardized exchangeable variables to a Brownian bridge.

math.PR