arXiv ScienceSearch

arXiv subjects

Charles Fefferman

Publications and source records attributed to Charles Fefferman.

At least 19 recordsLinked to original sources

Denoising data using convex relaxations

We study the problem of denoising observations \(Y_i=X_i+Z_i\), where the latent variables \(X_i\) are sampled from a low-dimensional manifold in \(\mathbb{R}^n\) and the noise variables \(Z_i\) are isotropic Gaussian. We propose a convex-relaxation estimator that first reduces dimension by principal component analysis and then projects the observations onto the convex hull of the projected latent manifold. We construct a statistical oracle that estimates its supporting hyperplanes from empirical Gaussian tail probabilities of the noisy sample. Under a lower-mass condition on the latent distribution, we prove finite-sample guarantees for the oracle and derive error bounds for the resulting denoiser. The analysis combines risk bounds for least-squares projection under convex constraints with entropy bounds for convex hulls. We also verify the assumptions of the framework for a Cryo-Electron Microscopy observation model by establishing suitable covering number and Lipschitz estimates for the associated group action and imaging operators.

stat.ME

Reconstruction of Manifold Distances from Noisy Observations

We consider the problem of reconstructing the intrinsic geometry of a manifold from noisy pairwise distance observations. Specifically, let $M$ denote a diameter 1 d-dimensional manifold and $\mu$ a probability measure on $M$ that is mutually absolutely continuous with the volume measure. Suppose $X_1,\dots,X_N$ are i.i.d. samples of $\mu$ and we observe noisy-distance random variables $d'(X_j, X_k)$ that are related to the true geodesic distances $d(X_j,X_k)$. With mild assumptions on the distributions and independence of the noisy distances, we develop a new framework for recovering all distances between points in a sufficiently dense subsample of $M$. Our framework improves on previous work which assumed i.i.d. additive noise with known moments. Our method is based on a new way to estimate $L_2$-norms of certain expectation-functions $f_x(y)=\mathbb{E}d'(x,y)$ and use them to build robust clusters centered at points of our sample. Using a new geometric argument, we establish that, under mild geometric assumptions--bounded curvature and positive injectivity radius--these clusters allow one to recover the true distances between points in the sample up to an additive error of $O(\varepsilon \log \varepsilon^{-1})$. We develop two distinct algorithms for producing these clusters. The first achieves a sample complexity $N \asymp \varepsilon^{-2d-2}\log(1/\varepsilon)$ and runtime $o(N^3)$. The second introduces novel geometric ideas that warrant further investigation. In the presence of missing observations, we show that a quantitative lower bound on sampling probabilities suffices to modify the cluster construction in the first algorithm and extend all recovery guarantees. Our main technical result also elucidates which properties of a manifold are necessary for the distance recovery, which suggests further extension of our techniques to a broader class of metric probability spaces.

stat.ML

Almost Optimal Agnostic Control of Unknown Linear Dynamics

We consider a simple control problem in which the underlying dynamics depend on a parameter $a$ that is unknown and must be learned. We study three variants of the control problem: Bayesian control, in which we have a prior belief about $a$; bounded agnostic control, in which we have no prior belief about $a$ but we assume that $a$ belongs to a bounded set; and fully agnostic control, in which $a$ is allowed to be an arbitrary real number about which we have no prior belief. In the Bayesian variant, a control strategy is optimal if it minimizes a certain expected cost. In the agnostic variants, a control strategy is optimal if it minimizes a quantity called the worst-case regret. For the Bayesian and bounded agnostic variants above, we produce optimal control strategies. For the fully agnostic variant, we produce almost optimal control strategies, i.e., for any $\varepsilon>0$ we produce a strategy that minimizes the worst-case regret to within a multiplicative factor of $(1+\varepsilon)$.

math.OC

Sobolev extension in a simple case

In this paper, we establish the existence of a bounded, linear extension operator $T: L^{2,p}(E) \to L^{2,p}(\mathbb{R}^2)$ when $1<p<2$ and $E$ is a finite subset of $\mathbb{R}^2$ contained in a line.

math.CA

Fitting a manifold to data in the presence of large noise

We assume that $M_0$ is a $d$-dimensional $C^{2,1}$-smooth submanifold of $R^n$. Let $K_0$ be the convex hull of $M_0,$ and $B^n_1(0)$ be the unit ball. We assume that $ M_0 \subseteq \partial K_0 \subseteq B^n_1(0).$ We also suppose that $M_0$ has volume ($d$-dimensional Hausdorff measure) less or equal to $V$, reach (i.e., normal injectivity radius) greater or equal to $\tau$. Moreover, we assume that $M_0$ is $R$-exposed, that is, tangent to every point $x \in M$ there is a closed ball of radius $R$ that contains $M$. Let $x_1, \dots, x_N$ be independent random variables sampled from uniform distribution on $M_0$ and $\zeta_1, \dots, \zeta_N$ be a sequence of i.i.d Gaussian random variables in $R^n$ that are independent of $x_1, \dots, x_N$ and have mean zero and covariance $\sigma^2 I_n.$ We assume that we are given the noisy sample points $y_i$, given by $$ y_i = x_i + \zeta_i,\quad \hbox{ for }i = 1, 2, \dots,N. $$ Let $\epsilon,\eta>0$ be real numbers and $k\geq 2$. Given points $y_i$, $i=1,2,\dots,N$, we produce a $C^k$-smooth function which zero set is a manifold $M_{rec}\subseteq R^n$ such that the Hausdorff distance between $M_{rec}$ and $M_0$ is at most $ \epsilon$ and $M_{rec}$ has reach that is bounded below by $c\tau/d^6$ with probability at least $1 - \eta.$ Assuming $d < c \sqrt{\log \log n}$ and all the other parameters are positive constants independent of $n$, the number of the needed arithmetic operations is polynomial in $n$. In the present work, we allow the noise magnitude $\sigma$ to be an arbitrarily large constant, thus overcoming a drawback of previous work.

math.ST

Optimal Agnostic Control of Unknown Linear Dynamics in a Bounded Parameter Range

Here and in a follow-on paper, we consider a simple control problem in which the underlying dynamics depend on a parameter $a$ that is unknown and must be learned. In this paper, we assume that $a$ is bounded, i.e., that $|a| \le a_{\text{MAX}}$, and we study two variants of the control problem. In the first variant, Bayesian control, we are given a prior probability distribution for $a$ and we seek a strategy that minimizes the expected value of a given cost function. Assuming that we can solve a certain PDE (the Hamilton-Jacobi-Bellman equation), we produce optimal strategies for Bayesian control. In the second variant, agnostic control, we assume nothing about $a$ and we seek a strategy that minimizes a quantity called the regret. We produce a prior probability distribution $d\text{Prior}(a)$ supported on a finite subset of $[-a_{\text{MAX}},a_{\text{MAX}}]$ so that the agnostic control problem reduces to the Bayesian control problem for the prior $d\text{Prior}(a)$.

math.OC

Controlling Unknown Linear Dynamics with Almost Optimal Regret

Here and in a companion paper, we consider a simple control problem in which the underlying dynamics depend on a parameter $a$ that is unknown and must be learned. In this paper, we assume that $a$ can be any real number and we do not assume that we have a prior belief about $a$. We seek a control strategy that minimizes a quantity called the regret. Given any $\varepsilon>0$, we produce a strategy that minimizes the regret to within a multiplicative factor of $(1+\varepsilon)$.

math.OC

Linear extension operators for Sobolev spaces on radially-symmetric binary trees

Let $1 < p < \infty$ and suppose that we are given a function $f$ defined on the leaves of a weighted tree. We would like to extend $f$ to a function $F$ defined on the entire tree, so as to minimize the weighted $W^{1,p}$-Sobolev norm of the extension. An easy situation is when $p = 2$, where the harmonic extension operator provides such a function $F$. In this note we record our analysis of the particular case of a radially-symmetric binary tree, which is a complete, finite, binary tree with weights that depend only on the distance from the root. Neither the averaging operator nor the harmonic extension operator work here in general. Nevertheless, we prove the existence of a linear extension operator whose norm is bounded by a constant depending solely on $p$. This operator is a variant of the standard harmonic extension operator, and in fact it is harmonic extension with respect to a certain Markov kernel determined by $p$ and by the weights.

math.FA

Classification of implication-closed ideals in certain rings of jets

For a set $E\subset\mathbb{R}^n$ that contains the origin we consider $I^m(E)$ -- the set of all $m^{\text{th}}$ degree Taylor approximations (at the origin) of $C^m$ functions on $\mathbb{R}^n$ that vanish on $E$. This set is a proper ideal in $\mathcal{P}^m(\mathbb{R}^n)$ -- the ring of all $m^{\text{th}}$ degree Taylor approximations of $C^m$ functions on $\mathbb{R}^n$. In [FS] we introduced the notion of a \textit{closed} ideal in $\mathcal{P}^m(\mathbb{R}^n)$, and proved that any ideal of the form $I^m(E)$ is closed. In this paper we classify (up to a natural equivalence relation) all closed ideals in $\mathcal{P}^m(\mathbb{R}^n)$ in all cases in which $m+n\leq5$. We also show that in these cases the converse also holds -- all closed proper ideals in $\mathcal{P}^m(\mathbb{R}^n)$ arise as $I^m(E)$ when $m+n\leq5$. In addition, we prove that in these cases any ideal of the form $I^m(E)$ for some $E\subset\mathbb{R}^n$ that contains the origin already arises as $I^m(V)$ for some semi-algebraic $V\subset\mathbb{R}^n$ that contains the origin. By doing so we prove that a conjecture by N. Zobin holds true in these cases.

math.FA

A property of ideals of jets of functions vanishing on a set

For a set $E\subset\mathbb{R}^n$ that contains the origin we consider $I^m(E)$ -- the set of all $m^{\text{th}}$ degree Taylor approximations (at the origin) of $C^m$ functions on $\mathbb{R}^n$ that vanish on $E$. This set is an ideal in $\mathcal{P}^m(\mathbb{R}^n)$ -- the ring of all $m^{\text{th}}$ degree Taylor approximations of $C^m$ functions on $\mathbb{R}^n$. Which ideals in $\mathcal{P}^m(\mathbb{R}^n)$ arise as $I^m(E)$ for some $E$? In this paper we introduce the notion of a \textit{closed} ideal in $\mathcal{P}^m(\mathbb{R}^n)$, and prove that any ideal of the form $I^m(E)$ is closed. We do not know whether in general any closed ideal is of the form $I^m(E)$ for some $E$, however we prove in [FS] that all closed ideals in $\mathcal{P}^m(\mathbb{R}^n)$ arise as $I^m(E)$ when $m+n\leq5$.

math.FA

Reconstruction and interpolation of manifolds II: Inverse problems with partial data for distances observations and for the heat kernel

We consider how a closed Riemannian manifold $M$ and its metric tensor $g$ can be approximately reconstructed from local distance measurements. Moreover, we consider an inverse problem of determining $(M,g)$ from limited knowledge on the heat kernel. In the part 1 of the paper, we considered the approximate construction of a smooth manifold in the case when one is given the noisy distances $\tilde d(x,y)=d(x,y)+\varepsilon_{x,y}$ for all points $x,y\in X$, where $X$ is a $\delta$-dense subset of $M$ and $|\varepsilon_{x,y}|<\delta$. In this part 2 of the paper, we consider a similar problem with partial data, that is, the approximate construction of the manifold $(M,g)$ when we are given $\tilde d(x,y)$ for $x\in X$ and $y \in U\cap X$, where $U$ is an open subset of $M$. In addition, we consider the inverse problem of determining the manifold $(M,g)$ with non-negative Ricci curvature from noisy observations of the heat kernel $G(y,z,t)$. We show that a manifold approximating $(M,g)$ can be determined in a stable way, when for some unknown source points $z_j$ in $X\setminus U$, we are given the values of the heat kernel $G(y,z_k,t)$ for $y\in X\cap U$ and $t\in (0,1)$ with a multiplicative noise. We also give a uniqueness result for the inverse problem in the case when the data does not contain noise and consider applications in manifold learning. A novel feature of the inverse problem for the heat kernel is that the set $M\setminus U$ containing the sources and the observation set $U$ are disjoint.

math.DG

Controlling Unknown Linear Dynamics with Bounded Multiplicative Regret

We consider a simple control problem in which the underlying dynamics depend on a parameter that is unknown and must be learned. We exhibit a control strategy which is optimal to within a multiplicative constant. While most authors find strategies which are successful as the time horizon tends to infinity, our strategy achieves lowest expected cost up to a constant factor for a fixed time horizon.

math.OC

$C^2$ Interpolation with Range Restriction

Given $ -\infty< \lambda < \Lambda < \infty $, $ E \subset \mathbb{R}^n $ finite, and $ f : E \to [\lambda,\Lambda] $, how can we extend $ f $ to a $ C^m(\mathbb{R}^n) $ function $ F $ such that $ \lambda\leq F \leq \Lambda $ and $ ||F||_{C^m(\mathbb{R}^n)} $ is within a constant multiple of the least possible, with the constant depending only on $ m $ and $ n $? In this paper, we provide the solution to the problem for the case $ m = 2 $. Specifically, we construct a (parameter-dependent, nonlinear) $ C^2(\mathbb{R}^n) $ extension operator that preserves the range $[\lambda,\Lambda]$, and we provide an efficient algorithm to compute such an extension using $ O(N\log N) $ operations, where $ N = #(E) $.

math.CA

Fitting a manifold of large reach to noisy data

Let ${\mathcal M}\subset {\mathbb R}^n$ be a $C^2$-smooth compact submanifold of dimension $d$. Assume that the volume of ${\mathcal M}$ is at most $V$ and the reach (i.e. the normal injectivity radius) of ${\mathcal M}$ is greater than $\tau$. Moreover, let $\mu$ be a probability measure on ${\mathcal M}$ whose density on ${\mathcal M}$ is a strictly positive Lipschitz-smooth function. Let $x_j\in {\mathcal M}$, $j=1,2,\dots,N$ be $N$ independent random samples from distribution $\mu$. Also, let $\xi_j$, $j=1,2,\dots, N$ be independent random samples from a Gaussian random variable in ${\mathbb R}^n$ having covariance $\sigma^2I$, where $\sigma$ is less than a certain specified function of $d, V$ and $\tau$. We assume that we are given the data points $y_j=x_j+\xi_j,$ $j=1,2,\dots,N$, modelling random points of ${\mathcal M}$ with measurement noise. We develop an algorithm which produces from these data, with high probability, a $d$ dimensional submanifold ${\mathcal M}_o\subset {\mathbb R}^n$ whose Hausdorff distance to ${\mathcal M}$ is less than $Cd\sigma^2/\tau$ and whose reach is greater than $c{\tau}/d^6$ with universal constants $C,c > 0$. The number $N$ of random samples required depends almost linearly on $n$, polynomially on $\sigma^{-1}$ and exponentially on $d$.

math.ST

Reconstruction of a Riemannian manifold from noisy intrinsic distances

We consider reconstruction of a manifold, or, invariant manifold learning, where a smooth Riemannian manifold $M$ is determined from intrinsic distances (that is, geodesic distances) of points in a discrete subset of $M$. In the studied problem the Riemannian manifold $(M,g)$ is considered as an abstract metric space with intrinsic distances, not as an embedded submanifold of an ambient Euclidean space. Let $\{X_1,X_2,\dots,X_N\}$ bea set of $N$ sample points sampled randomly from an unknown Riemannian $M$ manifold. We assume that we are given the numbers $D_{jk}=d_M(X_j,X_k)+\eta_{jk}$, where $j,k\in \{1,2,\dots,N\}$. Here, $d_M(X_j,X_k)$ are geodesic distances, $\eta_{jk}$ are independent, identically distributed random variables such that $\mathbb E e^{|\eta_{jk}|}$ is finite. We show that when $N$ is large enough, it is possible to construct an approximation of the Riemannian manifold $(M,g)$ with a large probability. This problem is a generalization of the geometric Whitney problem with random measurement errors. We consider also the case when the information on noisy distance $D_{jk}$ of points $X_j$ and $X_k$ is missing with some probability. In particular, we consider the case when we have no information on points that are far away.

math.PR

Efficient Algorithms for Approximate Smooth Selection

In this paper we provide efficient algorithms for approximate $\mathcal{C}^m(\mathbb{R}^n, \mathbb{R}^D)-$selection. In particular, given a set $E$, constants $M_0 > 0$ and $0 <\tau \leq \tau_{\max}$, and convex sets $K(x) \subset \mathbb{R}^D$ for $x \in E$, we show that an algorithm running in $C(\tau) N \log N$ steps is able to solve the smooth selection problem of selecting a point $y \in (1+\tau)\blacklozenge K(x)$ for $x \in E$ for an appropriate dilation of $K(x)$, $(1+\tau)\blacklozenge K(x)$, and guaranteeing that a function interpolating the points $(x, y)$ will be $\mathcal{C}^m(\mathbb{R}^n, \mathbb{R}^D)$ with norm bounded by $C M_0$.

math.FA

Solutions to a System of Equations for $C^m$ Functions

Fix $m\geq 0$, and let $A=\left( A_{ij}\left( x\right) \right) _{1\leq i\leq N,1\leq j\leq M}$ be a matrix of semialgebraic functions on $\mathbb{R}^{n}$ or on a compact subset $E \subset \mathbb{R}^n$. Given $f=\left( f_{1},\cdots ,f_{N}\right) \in C^{\infty }\left( \mathbb{R}^{n},\mathbb{R}^{N}\right) $, we consider the following system of equations \begin{equation} \sum_{j=1}^{M}A_{ij}\left( x\right) F_{j}\left( x\right) =f_{i}\left( x\right) \text{ }\left( i=1,\cdots ,N\right) \text{.} \end{equation} In this paper, we give algorithms for computing a finite list of linear partial differential operators such that $AF= f$ admits a $C^m(\mathbb{R}^n, \mathbb{R}^M)$ solution $F=(F_1,\cdots, F_M)$ if and only if $f=(f_1,\cdots, f_N)$ is annihilated by the linear partial differential operators.

math.CA