arXiv Science⌕ Search

arXiv subjects

Ryan Murray

Publications and source records attributed to Ryan Murray.

35 records · Page 2Linked to original sources

Distributed Gradient Flow: Nonsmoothness, Nonconvexity, and Saddle Point Evasion

The paper considers distributed gradient flow (DGF) for multi-agent nonconvex optimization. DGF is a continuous-time approximation of distributed gradient descent that is often easier to study than its discrete-time counterpart. The paper has two main contributions. First, the paper considers optimization of nonsmooth, nonconvex objective functions. It is shown that DGF converges to critical points in this setting. The paper then considers the problem of avoiding saddle points. It is shown that if agents' objective functions are assumed to be smooth and nonconvex, then DGF can only converge to a saddle point from a zero-measure set of initial conditions. To establish this result, the paper proves a stable manifold theorem for DGF, which is a fundamental contribution of independent interest. In a companion paper, analogous results are derived for discrete-time algorithms.

math.OC↗

A maximum principle argument for the uniform convergence of graph Laplacian regressors

This paper investigates the use of methods from partial differential equations and the Calculus of variations to study learning problems that are regularized using graph Laplacians. Graph Laplacians are a powerful, flexible method for capturing local and global geometry in many classes of learning problems, and the techniques developed in this paper help to broaden the methodology of studying such problems. In particular, we develop the use of maximum principle arguments to establish asymptotic consistency guarantees within the context of noise corrupted, non-parametric regression with samples living on an unknown manifold embedded in $\mathbb{R}^d$. The maximum principle arguments provide a new technical tool which informs parameter selection by giving concrete error estimates in terms of various regularization parameters. A review of learning algorithms which utilize graph Laplacians, as well as previous developments in the use of differential equation and variational techniques to study those algorithms, is given. In addition, new connections are drawn between Laplacian methods and other machine learning techniques, such as kernel regression and k-nearest neighbor methods.

stat.ML↗

Distributed Gradient Descent: Nonconvergence to Saddle Points and the Stable-Manifold Theorem

The paper studies a distributed gradient descent (DGD) process and considers the problem of showing that in nonconvex optimization problems, DGD typically converges to local minima rather than saddle points. The paper considers unconstrained minimization of a smooth objective function. In centralized settings, the problem of demonstrating nonconvergence to saddle points of gradient descent (and variants) is typically handled by way of the stable-manifold theorem from classical dynamical systems theory. However, the classical stable-manifold theorem is not applicable in distributed settings. The paper develops an appropriate stable-manifold theorem for DGD showing that convergence to saddle points may only occur from a low-dimensional stable manifold. Under appropriate assumptions (e.g., coercivity), this result implies that DGD typically converges to local minima and not to saddle points.

math.OC↗

Neutral competition in a deterministically changing environment: revisiting continuum approaches

Environmental variation can play an important role in ecological competition by influencing the relative advantage between competing species. Here, we consider such effects by extending a classical, competitive Moran model to incorporate an environment that fluctuates periodically in time. We adapt methods from work on these classical models to investigate the effects of the magnitude and frequency of environmental fluctuations on two important population statistics: the probability of fixation and the mean time to fixation. In particular, we find that for small frequencies, the system behaves similar to a system with a constant fitness difference between the two species, and for large frequencies, the system behaves similar to a neutrally competitive model. Most interestingly, the system exhibits nontrivial behavior for intermediate frequencies. We conclude by showing that our results agree quite well with recent theoretical work on competitive models with a stochastically changing environment, and discuss how the methods we develop ease the mathematical analysis required to study such models.

q-bio.PE↗

Regular Potential Games

A fundamental problem with the Nash equilibrium concept is the existence of certain "structurally deficient" equilibria that (i) lack fundamental robustness properties, and (ii) are difficult to analyze. The notion of a "regular" Nash equilibrium was introduced by Harsanyi. Such equilibria are isolated, highly robust, and relatively simple to analyze. A game is said to be regular if all equilibria in the game are regular. In this paper it is shown that almost all potential games are regular. That is, except for a closed subset with Lebesgue measure zero, all potential games are regular. As an immediate consequence of this, the paper also proves an oddness result for potential games: in almost all potential games, the number of Nash equilibrium strategies is finite and odd. Specialized results are given for weighted potential games, exact potential games, and games with identical payoffs. Applications of the results to game-theoretic learning are discussed.

math.OC↗

Spatial extreme values: variational techniques and stochastic integrals

This work employs variational techniques to revisit and expand the construction and analysis of extreme value processes. These techniques permit a novel study of spatial statistics of the location of minimizing events. We develop integral formulas for computing statistics of spatially-biased extremal events, and show that they are analogous to stochastic integrals in the setting of standard stochastic processes. We also establish an asymptotic result in the spirit of the Fisher-Tippett-Gnedenko theory for a broader class of extremal events and discuss some applications of our results.

math.PR↗

Revisiting Normalized Gradient Descent: Fast Evasion of Saddle Points

The note considers normalized gradient descent (NGD), a natural modification of classical gradient descent (GD) in optimization problems. A serious shortcoming of GD in non-convex problems is that GD may take arbitrarily long to escape from the neighborhood of a saddle point. This issue can make the convergence of GD arbitrarily slow, particularly in high-dimensional non-convex problems where the relative number of saddle points is often large. The paper focuses on continuous-time descent. It is shown that, contrary to standard GD, NGD escapes saddle points `quickly.' In particular, it is shown that (i) NGD `almost never' converges to saddle points and (ii) the time required for NGD to escape from a ball of radius $r$ about a saddle point $x^*$ is at most $5\sqrtκr$, where $κ$ is the condition number of the Hessian of $f$ at $x^*$. As an application of this result, a global convergence-time bound is established for NGD under mild assumptions.

math.OC↗

A model for system uncertainty in reinforcement learning

This work provides a rigorous framework for studying continuous time control problems in uncertain environments. The framework considered models uncertainty in state dynamics as a measure on the space of functions. This measure is considered to change over time as agents learn their environment. This model can be seem as a variant of either Bayesian reinforcement learning or adaptive control. We study necessary conditions for locally optimal trajectories within this model, in particular deriving an appropriate dynamic programming principle and Hamilton-Jacobi equations. This model provides one possible framework for studying the tradeoff between exploration and exploitation in reinforcement learning.

math.OC↗

On Best-Response Dynamics in Potential Games

The paper studies the convergence properties of (continuous) best-response dynamics from game theory. Despite their fundamental role in game theory, best-response dynamics are poorly understood in many games of interest due to the discontinuous, set-valued nature of the best-response map. The paper focuses on elucidating several important properties of best-response dynamics in the class of multi-agent games known as potential games---a class of games with fundamental importance in multi-agent systems and distributed control. It is shown that in almost every potential game and for almost every initial condition, the best-response dynamics (i) have a unique solution, (ii) converge to pure-strategy Nash equilibria, and (iii) converge at an exponential rate.

math.OC↗

A Note Regarding Second-Order $Γ$-limits for the Cahn--Hilliard Functional

This note completely resolves the asymptotic development of order $2$ by $Γ$-convergence of the mass-constrained Cahn--Hilliard functional, by showing that one of the critical assumptions of the authors' previous work (Leoni, Murray, Second-order $Γ$-limit for the Cahn--Hilliard functional, Arch. Ration. Mech. Anal. 219, 3, 2016) is unnecessary.

math.AP↗

Cutoff estimates for the Becker-Döring equations

This paper continues the authors' previous study (SIAM J. Math. Anal., 2016) of the trend toward equilibrium of the Becker-Döring equations with subcritical mass, by characterizing certain fine properties of solutions to the linearized equation. In particular, we partially characterize the spectrum of the linearized operator, showing that it contains the entire imaginary axis in polynomially weighted spaces. Moreover, we prove detailed cutoff estimates that establish upper and lower bounds on the lifetime of a class of perturbations to equilibrium.

math.AP↗

Trapping of solute atoms at grain boundaries in GdNi2

Lattice locations of 111In impurity probe atoms in intermetallic GdNi2 were studied as a function of alloy composition and temperature using perturbed angular correlation spectroscopy (PAC). Three nuclear quadrupole interaction signals were detected and their equilibrium site fractions were measured up to 700 oC. Two signals have well-defined electric field gradients (EFGs) and are attributed to In-probes on Gd- and Ni-sites in a well-ordered lattice. A third, inhomogeneously broadened signal was observed at low temperature. This is attributed to trapping, or segregation, of In-probes to lattice sinks such as grain boundaries (GB) that have a large multiplicity of local environments and EFGs. Changes in site fractions were reversible above 300oC. Measurements were made on a pair of samples that were richer and poorer in Gd. Remarkably, the GB-site was populated only in the more Gd-rich sample. This is explained by the hypothesis that excess Gd segregates to the grain boundaries and provides a lower enthalpy environment for In-probe atoms. Observations are discussed in relation to a three-level quantum system. Enthalpy differences between levels were determined from measurements of temperature dependences of ratios of site fractions. The enthalpy of transfer of In-probes from the Gd- to Ni-sublattice was found to be much smaller in the Gd-rich sample. This is attributed to a large temperature-dependence in the degeneracies of levels available to In-solutes in the phase, leading to an effective transfer enthalpy that differs greatly from the difference in site-enthalpies. A possible scenario is discussed. Different segregation enthalpies were measured for In-solute transferring from GB sites to Gd- and Ni-sites, whereas only an average value can be determined through macroscopic measurements.

cond-mat.mtrl-sci↗

A new analytical approach to consistency and overfitting in regularized empirical risk minimization

This work considers the problem of binary classification: given training data $x_1, \dots, x_n$ from a certain population, together with associated labels $y_1,\dots, y_n \in \left\{0,1 \right\}$, determine the best label for an element $x$ not among the training data. More specifically, this work considers a variant of the regularized empirical risk functional which is defined intrinsically to the observed data and does not depend on the underlying population. Tools from modern analysis are used to obtain a concise proof of asymptotic consistency as regularization parameters are taken to zero at rates related to the size of the sample. These analytical tools give a new framework for understanding overfitting and underfitting, and rigorously connect the notion of overfitting with a loss of compactness.

math.ST↗

Solute-solute interactions in intermetallic compounds

Two types of solute-solute interactions are investigated in this work. Quadrupole interactions caused by nearby Ag-solute atoms were measured at nuclei of 111In/Cd solute probe atoms in the binary compound GdAl2 using the method of perturbed angular correlation of gamma rays (PAC). Locations of In-probes and Ag-solutes on both Gd- and Al-sublattices were identified by comparing site fractions in Gd-poor and Gd-rich GdAl2(Ag) samples. Interaction enthalpies between solute-atom pairs were determined from temperature dependences of observed site fractions. Repulsive interactions were observed for close-neighbor complexes In/Gd/+Ag/Gd/ and In/Gd/+Ag/Al/ pairs, whereas a slightly attractive interaction was observed for In/Al/+Ag/Al/. Interaction enthalpies were all in the range +/- 0.15 eV. Temperature dependences of site fractions of In-probes on locally defect-free Gd- and Al-sites yields a transfer enthalpy that was found to be 0.343 eV in a previous study of undoped GdAl2. The corresponding values in GdAl2(Ag) samples are much smaller. This is attributed to competition of In- and Ag-solutes to occupy sites of the same sublattice. While the difference in site-enthalpies of In-solutes on Gd- and Al-sites is temperature independent, it is proposed that the transfer of Ag-solutes from Gd- to Al-sites leads to a large temperature dependence of degeneracies of levels available to In-solutes, resulting in an effective transfer enthalpy that is much smaller than the difference in site-enthalpies.

cond-mat.mtrl-sci↗

Slow motion for the nonlocal Allen-Cahn equation in n-dimensions

The goal of this paper is to study the slow motion of solutions of the nonlocal Allen-Cahn equation in a bounded domain $Ω\subset \mathbb{R}^n$, for $n > 1$. The initial data is assumed to be close to a configuration whose interface separating the states minimizes the surface area (or perimeter); both local and global perimeter minimizers are taken into account. The evolution of interfaces on a time scale $\varepsilon^{-1}$ is deduced, where $\varepsilon$ is the interaction length parameter. The key tool is a second-order $Γ$-convergence analysis of the energy functional, which provides sharp energy estimates. New regularity results are derived for the isoperimetric function of a domain. Slow motion of solutions for the Cahn-Hilliard equation starting close to global perimeter minimizers is proved as well.

math.AP↗

Second-Order $Γ$-limit for the Cahn-Hilliard Functional

The goal of this paper is to solve a long standing open problem, namely, the asymptotic development of order $2$ by $Γ$-convergence of the mass-constrained Cahn-Hilliard functional. This is achieved by introducing a novel rearrangement technique, which works without Dirichlet boundary conditions.

math.AP↗

An electron in the presence of multiple zero range potentials and an external laser field -- exact solutions for photoionization and stimulated bremsstrahlung

The method of zero range potential (ZRP) for one-electron problems is reviewed. In the absence of an external electromagnetic field, the notion of a ZRP is introduced from different points of view and for an arbitrary dimension of space. Then, three-dimensional problems of motion of an electron in the field of several ZRPs and laser radiation are studied. Exact wave functions for the processes of photoionization and stimulated bremsstrahlung in the presence of the laser field of an arbitrary pulse shape are obtained in the form of one-dimensional integral representations.

physics.atom-ph↗