arXiv ScienceSearch

arXiv subjects

Guozhi Zhang

Publications and source records attributed to Guozhi Zhang.

4 recordsLinked to original sources

Shallow neural network approximation in mixed Sobolev spaces

We investigate the best $L_2$ approximation of mixed Sobolev spaces by shallow neural networks with $n$ neurons and general activation functions. We first establish an activation-independent Fourier-block principle: if an activation has univariate approximation order $ρ$ in the sense of the Fourier-block property, then the global approximation rate has algebraic order $\min\{α,ρ\}$ for target functions of mixed smoothness $α$, up to explicit logarithmic factors. To verify this property for concrete activations, we introduce a structured univariate approximation condition that implies the Fourier-block property with explicit parameters. For $\mathrm{ReLU}^k$, a matching algebraic lower bound identifies $\min\{α,k+1\}$ as the optimal algebraic approximation exponent in any dimension, up to logarithmic factors in the upper bound. The framework also yields the exponent $\min\{α,k+1\}$ for cardinal B-splines and soft-$\mathrm{ReLU}^k$, and the full mixed-smoothness exponent $α$ for ELU and cosine activations, again up to logarithmic~factors.

math.NA

Nearly optimal Kolmogorov widths under holomorphic mappings

This paper establishes essentially optimal asymptotic bounds for Kolmogorov widths under holomorphic mappings between complex Banach spaces. Given a compact parameter set whose Kolmogorov widths decay algebraically with rate s, we prove that the widths of its image under a holomorphic mapping decay algebraically with every rate t<s, thereby answering an open question raised by Cohen and DeVore. As an application, we obtain a sharp characterization of the approximability of solution manifolds associated with inf-sup stable parametrized PDEs. We also construct an explicit example showing that the arbitrarily small loss in the algebraic decay exponent is unavoidable. Finally, we provide similar characterization of asymptotic bounds for Kolmogorov widths in the exponentially decaying regime. Our analysis involves multilinear Taylor expansion in Banach spaces and a novel block dyadic expansion-truncation technique.

math.FA

Higher Order Approximation Rates for ReLU CNNs in Korobov Spaces

This paper investigates the $L_p$ approximation error for higher order Korobov functions using deep convolutional neural networks (CNNs) with ReLU activation. For target functions having a mixed derivative of order m+1 in each direction, we improve classical approximation rate of second order to (m+1)-th order (modulo a logarithmic factor) in terms of the depth of CNNs. The key ingredient in our analysis is approximate representation of high-order sparse grid basis functions by CNNs. The results suggest that higher order expressivity of CNNs does not severely suffer from the curse of dimensionality.

cs.LG

Some Super-approximation Rates of ReLU Neural Networks for Korobov Functions

This paper examines the $L_p$ and $W^1_p$ norm approximation errors of ReLU neural networks for Korobov functions. In terms of network width and depth, we derive nearly optimal super-approximation error bounds of order $2m$ in the $L_p$ norm and order $2m-2$ in the $W^1_p$ norm, for target functions with $L_p$ mixed derivative of order $m$ in each direction. The analysis leverages sparse grid finite elements and the bit extraction technique. Our results improve upon classical lowest order $L_\infty$ and $H^1$ norm error bounds and demonstrate that the expressivity of neural networks is largely unaffected by the curse of dimensionality.

cs.LG