arXiv ScienceSearch

arXiv subjects

Zexin Pan

Publications and source records attributed to Zexin Pan.

At least 19 recordsLinked to original sources

Robust high-dimensional integration using medians of coarsely scrambled Sobol' sequences

We study the numerical approximation of high-dimensional integrals using randomized quasi-Monte Carlo (RQMC) methods, with a focus on scrambled Sobol' sequences. While asymptotically faster than Monte Carlo, classical RQMC suffers from error bounds that grow exponentially in the dimension $s$, and its root mean squared error (RMSE) convergence rate is generally no better than $O(N^{-3/2})$. To overcome these limitations, we combine two recent developments: Suzuki's coarse scrambling and the median trick. Coarse scrambling randomizes Sobol' sequences according to their generating base polynomials and significantly reduces the maximal gain coefficient. By appropriately choosing the base polynomials, we show that the coefficient can be made uniformly bounded in $s$, yielding an $O(N^{-1/2})$ RMSE for $L^2$ integrands with a dimension-independent constant. The median trick then enables near-optimal convergence for function classes beyond $L^2$: for $L^p$ integrands with $p\in(1,2)$, the median of independent coarsely scrambled estimates achieves an error of $O(N^{-1+1/p})$ with high probability; for integrands in the Haar wavelet space $\mathcal H_{\mathrm{wav},α,s,p,q}$ with $α_p:=α-(1/p-1/2)_+>0$, we prove a high-probability error bound of $O(N^{-α_p-1/2+\varepsilon})$ for any $\varepsilon>0$; the bound improves to $O(N^{-r-α_p-1/2+\varepsilon})$ when the integrand has dominating mixed derivatives of order $r\ge1$ that belong to $\mathcal H_{\mathrm{wav},α,s,p,q}$. The latter two bounds are uniform in $s$ under suitable conditions on the ANOVA components of the integrands. Numerical experiments confirm the predicted convergence rates and demonstrate the robustness of the proposed approach.

math.NA

Universal $L^2$-approximation using median digital-net algorithms

We propose a median digital-net algorithm for $L^2$-approximation of non-periodic functions over $[0,1]^s$, inspired by the recently developed median lattice algorithms for the periodic setting. The algorithm requires no smoothness or weight parameters but only a sufficiently large candidate Walsh index set $K$. It proceeds in three stages: generating multiple estimates of the Walsh coefficients in $K$ using independent randomized digital-net samples; taking the respective median of both the estimates and their absolute values; then, based on these median values, identifying the dominant coefficients and constructing a truncated Walsh series as the final approximation. We prove that if the target function has dominating mixed partial derivatives up to order $α$, all having finite Vitali variation of fractional order $λ$, then the algorithm achieves an $L^2$-error of $\mathcal{O}(M^{-α-λ+η})$ with high probability, where $M$ is the total number of function evaluations and $η>0$ is arbitrarily small. Furthermore, the implied constant grows at most polynomially in the dimension $s$ under suitable decay conditions on the ANOVA components of the target function. On the implementation side, we provide both parameter-dependent and -independent constructions of the index set $K$, and employ the fast Walsh--Hadamard transform and Gray code ordering to accelerate the algorithm. Numerical experiments support the theoretical analysis and demonstrate that the proposed algorithm remains effective in high-dimensional settings.

math.NA

Sharp convergence bounds for sums of POD and SPOD weights

This work analyzes the convergence of sums of the form $S_{\boldsymbolγ}(m)=\sum_{v\subseteq \mathbb{N}}γ_v m^{|v|}$ with product and order dependent (POD) weights $γ_v$. We establish that for a nonnegative sequence $\{Υ_j\mid j\in \mathbb{N}\}$, $$\sum_{v\subseteq \mathbb{N}} |v|! m^{|v|}\prod_{j\in v} Υ_j<\infty \text{ for all } m>0 \text{ if and only if } \sum_{j=1}^\infty Υ_j<\infty.$$ We further characterize the growth of $S_{\boldsymbolγ}(m)$ when $γ_v=(|v|!)^σ\prod_{j\in v}j^{-ρ}$ and prove that $\log S_{\boldsymbolγ}(m)$ is of asymptotic order $m^{1/(ρ-σ)}$ when $ρ>σ\geq 0$. We subsequently generalize both the convergence criterion and the asymptotic order of $\log S_{\boldsymbolγ}(m)$ to smoothness-driven product and order dependent (SPOD) weights, while noting that a full necessary-and-sufficient analogue remains open. Finally, we apply our theory to quasi-Monte Carlo (QMC) integration, showing that interlaced polynomial lattice rules achieve a dimension-independent convergence rate without a commonly imposed assumption in the QMC literature.

math.NA

Quasi-Monte Carlo for SDE Simulation: Error Analysis and Dimensionality Reduction

We investigate the numerical simulation of general stochastic differential equations (SDEs) using Quasi-Monte Carlo (QMC) methods. First, we provide a rigorous theoretical analysis of the QMC method applied to the Euler-Maruyama (EM) scheme, establishing that it significantly accelerates the decay of the sampling error and achieves an asymptotically superior convergence rate over the classical Monte Carlo method. Second, the traditional EM scheme exhibits a slow polynomial decay of the discretization error, which necessitates a large number of time steps and leads to a significantly high integration dimension. To address this issue, we propose a Multilevel Stochastic Time Grid (MSTG) method based on Exact Simulation techniques, and we rigorously establish its convergence rate under randomized QMC sampling, proving that it preserves the high-order convergence of the sampling error. In terms of the overall error, the truncation error of the proposed MSTG method exhibits a remarkably fast super-exponential decay. Consequently, to achieve a given accuracy level, our approach requires significantly fewer discretization steps than the EM scheme, thereby drastically reducing the actual integration dimension of the QMC method. This substantial dimensionality reduction strategy greatly enhances the practical efficiency of the QMC algorithm. Numerical experiments fully corroborate the superiority of the proposed approach.

math.NA

Worst-case $L_p$-approximation of periodic functions using median lattice algorithms

We study the worst-case approximation of multivariate periodic functions from the weighted Korobov space $H_{d,α,γ}$ with smoothness $α>1/2$ in the Lebesgue norm $L_p([0,1]^d)$ for $1\le p\le\infty$. We analyze a \emph{median lattice algorithm} that reconstructs a truncated Fourier series by approximating the coefficients on a hyperbolic-cross-type index set using $R$ rank-1 lattice sampling rules with independent randomly chosen generating vectors, and then aggregating the resulting coefficient estimators via the componentwise median. For an odd number of repetitions $R>1$ and an odd prime lattice size $N$, we prove high-probability error bounds in both $L_\infty$ and $L_2$. Interpolation then yields the result for all $1 \le p\le\infty$. In particular, with a high probability, the algorithm satisfies \[ \mathrm{err}(H_{d,α,γ},L_p,A)\ \le\ C_{d,α,β,\boldsymbolγ,p}\, N^{- α+ (\frac12 - \frac1p)_+ + β}, \qquad 1 \le p\le\infty,\ β>0, \] where $(x)_+ = \max\{x, 0\}$, $N$ is the number of function evaluations, and the weights $\boldsymbolγ$ and the constant $C_{d,α,β,\boldsymbolγ,p}$ are independent of $N$. For $p=\infty$, $C_{d,α,β,\boldsymbolγ,\infty}$ is dimension-independent under the summability condition $\sum_{j=1}^\infty γ_j^{1/(2α)}<\infty$. These results extend recent analyses of median-based lattice approximation in $L_2$ and complement related multiple-shift lattice approaches, showing that median aggregation yields nearly optimal $L_p$-approximation rates (up to logarithmic factors and an arbitrarily small loss) in weighted Korobov spaces.

math.NA

Uncertainty quantification using importance-sampled quasi-Monte Carlo with dimension-independent convergence rates

Quasi-Monte Carlo (QMC) integration over unbounded domains $\mathbb{R}^s$ remains challenging due to the high dimensionality of sampling space and the boundary growth of the integrand. In applications such as uncertainty quantification (UQ), the dimension $s$ can reach hundreds or even thousands. To restore the efficiency of quadrature rules in high dimensions, constructive QMC methods like lattice rules have been successfully developed within the framework of weighted function spaces. In contrast to designing problem-specific quadrature points, this paper proposes transforming the underlying integrand to accommodate the off-the-shelf scrambled nets (a construction-free randomized QMC method) via the boundary-damping importance sampling (BDIS) proposed by Pan et al. (2025). We provide a rigorous analysis of the dimension-independent convergence rate of BDIS-based scrambled nets while covering a broader class of unbounded functions than that in Pan et al. (2025). By exploiting the dimension structure of the parametric input random field, the proposed $n$-point quadrature rule achieves a dimension-independent mean squared error rate of $O(n^{-1-α^*+\varepsilon})$ on standard UQ problems in elliptic partial differential equations (PDEs), where $\varepsilon>0$ is arbitrarily small and $α^*\in (0,1)$ reflects the regularity with respect to the parametric variables. Numerical experiments on elliptic PDEs with high-dimensional parameters further demonstrate the effectiveness of the method.

math.NA

Quasi-Monte Carlo confidence intervals using quantiles of randomized nets

Recent advances in quasi-Monte Carlo integration have shown that for linearly scrambled digital net estimators, the convergence rate can be dramatically improved by taking the median rather than the mean of multiple independent replicates. In this work, we demonstrate that the quantiles of such estimators can be used to construct confidence intervals with asymptotically valid coverage for high-dimensional integrals. By analyzing the error distribution for a class of infinitely differentiable integrands, we prove that as the sample size increases, the integration error decomposes into an asymptotically symmetric component and a vanishing remainder. Consequently, the asymptotic error distribution is symmetric about zero, ensuring that a quantile-based interval constructed from independent replicates captures the true integral with probability converging to a nominal level determined by the binomial distribution.

math.ST

Dimension-independent convergence rates of randomized nets using median-of-means

Recent advances in quasi-Monte Carlo integration demonstrate that the median of linearly scrambled digital net estimators achieves near-optimal convergence rates for high-dimensional integrals without requiring a priori knowledge of the integrand's smoothness. Building on this framework, we prove that the median estimator attains dimension-independent convergence, a property known as strong tractability in complexity theory, under tractability conditions characterized by low effective dimensionality. Using a probabilistic, integrand-specific error criterion, our analysis establishes both faster and dimension-independent convergence under weaker assumptions than previously possible in the worst-case setting.

stat.CO

Quasi-Monte Carlo with one categorical variable

We study randomized quasi-Monte Carlo (RQMC) estimation of a multivariate integral where one of the variables takes only a finite number of values. This problem arises when the variable of integration is drawn from a mixture distribution as is common in importance sampling and also arises in some recent work on transport maps. We find that when integration error decreases at an RQMC rate that it is then important to oversample the smallest mixture components instead of using a proportional allocation. This can even improve the rate of convergence. The optimal allocations depend on the possibly unknown convergence rate. Designing the sample with an incorrect assumption on the rate still attains that convergence rate, with an inferior implied constant. The penalty for using a pessimistic rate is typically higher than for using an optimistic one. We also find that for the most accurate RQMC sampling methods, it is advantageous to arrange that our $n=2^m$ randomized Sobol' points split into subsample sizes that are also powers of $2$.

stat.CO

$L_2$-approximation using median lattice algorithms

In this paper, we study the problem of multivariate $L_2$-approximation of functions belonging to a weighted Korobov space. We propose and analyze a median lattice-based algorithm, inspired by median integration rules, which have attracted significant attention in the theory of quasi-Monte Carlo methods. Our algorithm approximates the Fourier coefficients associated with a suitably chosen frequency index set, where each coefficient is estimated by taking the median over approximations from randomly shifted rank-1 lattice rules with independently chosen generating vectors. We prove that the algorithm achieves, with high probability, a convergence rate of the $L_2$-approximation error that is arbitrarily close to optimal with respect to the number of function evaluations. Furthermore, we show that the error bound depends only polynomially on the dimension, or is even independent of the dimension, under certain summability conditions on the weights. Numerical experiments illustrate the performance of the proposed median lattice-based algorithm.

math.NA

Universal $L_2$-approximation using median lattice algorithms

We study the problem of multivariate $L_2$-approximation of functions in a weighted Korobov space using a median lattice-based algorithm recently proposed by the authors. In the original work, the algorithm requires knowledge of the smoothness and weights of the Korobov space to construct the hyperbolic cross index set, where each coefficient is estimated via the median of approximations obtained from randomly shifted, randomly chosen rank-1 lattice rules. In this paper, we introduce a \emph{universal median lattice-based algorithm}, which eliminates the need for any prior information on smoothness and weights. Although the tractability property of the algorithm slightly deteriorates, we prove that, for individual functions in the Korobov space with arbitrary smoothness and (downward-closed) weights, it achieves an $L_2$-approximation error arbitrarily close to the optimal rate with respect to the number of function evaluations. Numerical experiments are conducted to support our theoretical claim.

math.NA

Quasi-Monte Carlo integration over $\mathbb{R}^s$ with boundary-damping importance sampling

This paper proposes a new importance sampling (IS) that is tailored to quasi-Monte Carlo (QMC) integration over $\mathbb{R}^s$. IS introduces a multiplicative adjustment to the integrand by compensating the sampling from the proposal instead of the target distribution. Improper proposals result in severe adjustment factor for QMC. Our strategy is to first design a adjustment factor to meet desired regularities and then determine a tractable transport map from the standard uniforms to the proposal for using QMC quadrature points as inputs. The transport map has the effect of damping the boundary growth of the resulting integrand so that the effectiveness of QMC can be reclaimed. Under certain conditions on the original integrand, our proposed IS enjoys a fast convergence rate independently of the dimension $s$, making it amenable to high-dimensional problems.

math.NA

Skewness of a randomized quasi-Monte Carlo estimate

Some recent work on confidence intervals for randomized quasi-Monte Carlo (RQMC) sampling found a surprising result: ordinary Student $t$ 95% confidence intervals based on a modest number of replicates were seen to be very effective and even more reliable than some bootstrap $t$ intervals that were expected to be best. One potential explanation is that those RQMC estimates have small skewness. In this paper we give conditions under which the skewness is $O(n^ε)$ for any $ε>0$, so 'almost $O(1)$'. Under a random generator matrix model, we can improve this rate to $O(n^{-1/2+ε})$ with very high probability. We also improve some probabilistic bounds on the distribution of the quality parameter $t$ for a digital net in a prime base under random sampling of generator matrices.

math.NA

Automatic optimal-rate convergence of randomized nets using median-of-means

We study the sample median of independently generated quasi-Monte Carlo estimators based on randomized digital nets and prove it approximates the target integral value at almost the optimal convergence rate for various function spaces. In contrast to previous methods, the algorithm does not require a priori knowledge of underlying function spaces or even an input of pre-designed $(t,m,s)$-digital nets, and is therefore easier to implement. This study provides further evidence that quasi-Monte Carlo estimators are heavy-tailed when applied to smooth integrands and taking the median can significantly improve the error by filtering out the outliers.

math.NA

Computable error bounds for quasi-Monte Carlo using points with non-negative local discrepancy

Let $f:[0,1]^d\to\mathbb{R}$ be a completely monotone integrand as defined by Aistleitner and Dick (2015) and let points $\boldsymbol{x}_0,\dots,\boldsymbol{x}_{n-1}\in[0,1]^d$ have a non-negative local discrepancy (NNLD) everywhere in $[0,1]^d$. We show how to use these properties to get a non-asymptotic and computable upper bound for the integral of $f$ over $[0,1]^d$. An analogous non-positive local discrepancy (NPLD) property provides a computable lower bound. It has been known since Gabai (1967) that the two dimensional Hammersley points in any base $b\ge2$ have non-negative local discrepancy. Using the probabilistic notion of associated random variables, we generalize Gabai's finding to digital nets in any base $b\ge2$ and any dimension $d\ge1$ when the generator matrices are permutation matrices. We show that permutation matrices cannot attain the best values of the digital net quality parameter when $d\ge3$. As a consequence the computable absolutely sure bounds we provide come with less accurate estimates than the usual digital net estimates do in high dimensions. We are also able to construct high dimensional rank one lattice rules that are NNLD. We show that those lattices do not have good discrepancy properties: any lattice rule with the NNLD property in dimension $d\ge2$ either fails to be projection regular or has all its points on the main diagonal. Complete monotonicity is a very strict requirement that for some integrands can be mitigated via a control variate.

math.NA

Gain coefficients for scrambled Halton points

Randomized quasi-Monte Carlo, via certain scramblings of digital nets, produces unbiased estimates of $\int_{[0,1]^d}f(\boldsymbol{x})\,\mathrm{d}\boldsymbol{x}$ with a variance that is $o(1/n)$ for any $f\in L^2[0,1]^d$. It also satisfies some non-asymptotic bounds where the variance is no larger than some $Γ<\infty$ times the ordinary Monte Carlo variance. For scrambled Sobol' points, this quantity $Γ$ grows exponentially in $d$. For scrambled Faure points, $Γ\leqslant \exp(1)\doteq 2.718$ in any dimension, but those points are awkward to use for large $d$. This paper shows that certain scramblings of Halton sequences have gains below an explicit bound that is $O(\log d)$ but not $O( (\log d)^{1-ε})$ for any $ε>0$ as $d\to\infty$. For $6\leqslant d\leqslant 10^6$, the upper bound on the gain coefficient is never larger than $3/2+\log(d/2)$.

math.NA

Super-polynomial accuracy of multidimensional randomized nets using the median-of-means

We study approximate integration of a function $f$ over $[0,1]^s$ based on taking the median of $2r-1$ integral estimates derived from independently randomized $(t,m,s)$-nets in base $2$. The nets are randomized by Matousek's random linear scramble with a digital shift. If $f$ is analytic over $[0,1]^s$, then the probability that any one randomized net's estimate has an error larger than $2^{-cm^2/s}$ times a quantity depending on $f$ is $O(1/\sqrt{m})$ for any $c<3\log(2)/π^2\approx 0.21$. As a result the median of the distribution of these scrambled nets has an error that is $O(n^{-c\log(n)/s})$ for $n=2^m$ function evaluations. The sample median of $2r-1$ independent draws attains this rate too, so long as $r/m^2$ is bounded away from zero as $m\to\infty$. We include results for finite precision estimates and some non-asymptotic comparisons to taking the mean of $2r-1$ independent draws.

math.NA

Super-polynomial accuracy of one dimensional randomized nets using the median-of-means

Let $f$ be analytic on $[0,1]$ with $|f^{(k)}(1/2)|\leq Aα^kk!$ for some constant $A$ and $α<2$. We show that the median estimate of $μ=\int_0^1f(x)\,\mathrm{d}x$ under random linear scrambling with $n=2^m$ points converges at the rate $O(n^{-c\log(n)})$ for any $c< 3\log(2)/π^2\approx 0.21$. We also get a super-polynomial convergence rate for the sample median of $2k-1$ random linearly scrambled estimates, when $k=Ω(m)$. When $f$ has a $p$'th derivative that satisfies a $λ$-Hölder condition then the median-of-means has error $O( n^{-(p+λ)+ε})$ for any $ε>0$, if $k\to\infty$ as $m\to\infty$.

stat.CO