arXiv ScienceSearch

arXiv subjects

Sara Pollock

Publications and source records attributed to Sara Pollock.

At least 19 recordsLinked to original sources

Applying acceleration to Krylov subspace eigenvalue solvers

In this paper, we apply acceleration to the inverse-free preconditioned Krylov subspace method introduced by Golub and Ye, which solves the symmetric generalized eigenvalue problem for the algebraically smallest eigenvalue. As the method is an improvement on steepest descent, we consider acceleration based on Nesterov accelerated steepest descent and Polyak's heavy-ball method. We extend acceleration to the block version of the Krylov subspace method and prove convergence for a more generalized choice of subspace. We present numerical results demonstrating the effect of fixed and safeguarded-adaptive choice of the momentum parameter, which show convergence in fewer outer iterations compared with LOBPCG with the same subspace size and generally fewer iterations than the base method when solving for multiple clustered eigenvalues with small dimension size. We also provide an explanation for the acceleration seen from implementing Polyak's heavy-ball method, including justifying the given parameter range.

math.NA

An Adaptive Method for Optimal Control Problems Constrained by Parabolic Differential Equations

An adaptive direct collocation method is developed for solving optimal control problems constrained by parabolic partial differential equations. The partial differential equation is first reformulated in a variational setting, where the spatial domain is discretized using the hp-Galerkin finite element method. To address nonlinearities in the variational form, a Kirchhoff-like integral transformation is applied to linearize the dynamics. In the temporal dimension, an orthogonal collocation scheme, the hp-flipped Legendre-Gauss-Radau method, is employed to fully discretize the problem, yielding a large, sparse nonlinear programming problem. Upon solving the nonlinear programming problem, solution accuracy is assessed through an implicit residual estimation procedure. This approach evaluates the local error by solving auxiliary residual problems over selected subdomains, providing a novel means of error estimation within an orthogonal collocation framework for optimal control. Based on the computed error estimate, the mesh is adaptively refined or coarsened to meet a prescribed error tolerance. Mesh refinement is guided by the estimated regularity of the solution which is determined via the decay rate of the coefficients of a Legendre polynomial expansion. In overcollocated regions, a mesh reduction strategy is adapted from orthogonal collocation methods for application within the finite element framework. Numerical examples demonstrate that the proposed method can reduce the error by up to five orders of magnitude in both spatial and temporal dimensions.

math.OC

An adaptive framework for first-order gradient methods

Gradient methods are widely used in optimization problems. In practice, while the smoothness parameter can be estimated utilizing techniques such as backtracking, estimating the strong convexity parameter remains a challenge; moreover, even with the optimal parameter choice, convergence can be slow. In this work, we propose a framework for dynamically adapting the step size and momentum parameters in first-order gradient methods for the optimization problem, without prior knowledge of the strong convexity parameter. The main idea is to use the geometric average of the ratios of successive residual norms as an empirical estimate of the upper bound on the convergence rate, which in turn allows us to adaptively update the algorithm parameters. The resulting algorithms are simple to implement, yet efficient in practice, requiring only a few additional computations on existing information. The proposed adaptive gradient methods are shown to converge at least as fast as gradient descent for quadratic optimization problems. Numerical experiments on both quadratic and nonlinear problems validate the effectiveness of the proposed adaptive algorithms. The results show that the adaptive algorithms are comparable to their counterparts using optimal parameters, and in some cases, they capture local information and exhibit improved performance.

math.OC

Momentum accelerated power iterations and the restarted Lanczos method

In this paper we compare two methods for finding extremal eigenvalues and eigenvectors: the restarted Lanczos method and momentum accelerated power iterations. The convergence of both methods is based on ratios of Chebyshev polynomials evaluated at subdominant and dominant eigenvalues; however, the convergence is not the same. Here we compare the theoretical convergence properties of both methods, and determine the relative regimes where each is more efficient. We further introduce a preconditioning technique for the restarted Lanczos method using momentum accelerated power iterations, and demonstrate its effectiveness. The theoretical results are backed up by numerical tests on benchmark problems.

math.NA

Random Walks, Faber Polynomials and Accelerated Power Methods

In this paper, we construct families of polynomials defined by recurrence relations related to mean-zero random walks. We show these families of polynomials can be used to approximate $z^n$ by a polynomial of degree $\sim \sqrt{n}$ in associated radially convex domains in the complex plane. Moreover, we show that the constructed families of polynomials have a useful rapid growth property and a connection to Faber polynomials. Applications to iterative linear algebra are presented, including the development of arbitrary-order dynamic momentum power iteration methods suitable for classes of non-symmetric matrices.

math.NA

Faber polynomials in a deltoid region and power iteration momentum methods

We consider a region in the complex plane enclosed by a deltoid curve inscribed in the unit circle, and define a family of polynomials $P_n$ that satisfy the same recurrence relation as the Faber polynomials for this region. We use this family of polynomials to give a constructive proof that $z^n$ is approximately a polynomial of degree $\sim\sqrt{n}$ within the deltoid region. Moreover, we show that $|P_n| \le 1$ in this deltoid region, and that, if $|z| = 1+\varepsilon$, then the magnitude $|P_n(z)|$ is at least $\frac{1}{3}(1+\sqrt{\varepsilon})^n$, for all $\varepsilon > 0$. We illustrate our polynomial approximation theory with an application to iterative linear algebra. In particular, we construct a higher-order momentum-based method that accelerates the power iteration for certain matrices with complex eigenvalues. We show how the method can be run dynamically when the two dominant eigenvalues are real and positive.

math.NA

Optimal Control of Parabolic Differential Equations Using Radau Collocation

A method is presented for the numerical solution of optimal boundary control problems governed by parabolic partial differential equations. The continuous space-time optimal control problem is transcribed into a sparse nonlinear programming problem through state and control parameterization. In particular, a multi-interval flipped Legendre-Gauss-Radau collocation method is implemented for temporal discretization alongside a Galerkin finite element spatial discretization. The finite element discretization allows for a reduction in problem size and avoids the redefinition of constraints required under a previous method. Further, a generalization of a Kirchoff transformation is performed to handle variational form nonlinearities in the context of numerical optimization. Due to the correspondence between the collocation points and the applied boundary conditions, the multi-interval flipped Legendre-Gauss-Radau collocation method is demonstrated to be preferable over the standard Legendre-Gauss-Radau collocation method for optimal control problems governed by parabolic partial differential equations. The details of the resulting transcription of the optimal control problem into a nonlinear programming problem are provided. Numerical examples demonstrate that the use of a multi-interval flipped Legendre-Gauss-Radau temporal discretization can lead to a reduction in the required number of collocation points to compute accurate values of the optimal objective in comparison to other methods. Lastly, a self-convergence analysis on each test problem illustrates that the error decays exponentially as a function of the mesh size in both the temporal and spatial dimensions.

math.OC

A Finite Element Implementation of the SRTD Algorithm for an Oldroyd 3-Parameter Viscoelastic Fluid Model

In this paper, we discuss a finite element implementation of the SRTD algorithm described by Girault and Scott for the steady-state case of a certain 3-parameter subset of the Oldroyd models. We compare it to the well-known EVSS method, which, though originally described for the upper-convected Maxwell model, can easily accommodate the Oldroyd 3-parameter model. We obtain numerical results for both methods on two benchmark problems: the lid-driven cavity problem and the journal-bearing, or eccentric rotating cylinders, problem. We find that the resulting finite element implementation of SRTD is stable with respect to mesh refinement and is generally faster than EVSS, though is not capable of reaching as high a Weissenberg number as EVSS.

math.NA

Computational analysis of a contraction rheometer for the grade-two fluid model

We explore the possibility of simulating the grade-two fluid model in a geometry related to a contraction rheometer, and we provide details on several key aspects of the computation. We show how the results can be used to determine the viscosity $\nu$ from experimental data. We also explore the identifiability of the grade-two parameters $\alpha_1$ and $\alpha_2$ from experimental data. In particular, as the flow rate varies, force data appears to be nearly the same for certain distinct pairs of values $\alpha_1$ and $\alpha_2$; however we determine a regime for $\alpha_1$ and $\alpha_2$ for which the parameters may be identifiable with a contraction rheometer.

math.NA

Dynamically accelerating the power iteration with momentum

In this paper, we propose, analyze and demonstrate a dynamic momentum method to accelerate power and inverse power iterations with minimal computational overhead. The method can be applied to real diagonalizable matrices, is provably convergent with acceleration in the symmetric case, and does not require a priori spectral knowledge. We review and extend background results on previously developed static momentum accelerations for the power iteration through the connection between the momentum accelerated iteration and the standard power iteration applied to an augmented matrix. We show that the augmented matrix is defective for the optimal parameter choice. We then present our dynamic method which updates the momentum parameter at each iteration based on the Rayleigh quotient and two previous residuals. We present convergence and stability theory for the method by considering a power-like method consisting of multiplying an initial vector by a sequence of augmented matrices. We demonstrate the developed method on a number of benchmark problems, and see that it outperforms both the power iteration and often the static momentum acceleration with optimal parameter choice. Finally, we present and demonstrate an explicit extension of the algorithm to inverse power iterations.

math.NA

Analysis of the Picard-Newton iteration for the Navier-Stokes equations: global stability and quadratic convergence

We analyze and test a simple-to-implement two-step iteration for the incompressible Navier-Stokes equations that consists of first applying the Picard iteration and then applying the Newton iteration to the Picard output. We prove that this composition of Picard and Newton converges quadratically, and our analysis (which covers both the unique solution and non-unique solution cases) also suggests that this solver has a larger convergence basin than usual Newton because of the improved stability properties of Picard-Newton over Newton. Numerical tests show that Picard-Newton dramatically outperforms both the Picard and Newton iterations, especially as the Reynolds number increases. We also consider enhancing the Picard step with Anderson acceleration (AA), and find that the AAPicard-Newton iteration has even better convergence properties on several benchmark test problems.

math.NA

Analysis of an Adaptive Safeguarded Newton-Anderson Algorithm of Depth One with Applications to Fluid Problems

The purpose of this paper is to develop a practical strategy to accelerate Newton's method in the vicinity of singular points. We present an adaptive safeguarding scheme with a tunable parameter, which we call adaptive gamma-safeguarding, that one can use in tandem with Anderson acceleration to improve the performance of Newton's method when solving problems at or near singular points. The key features of adaptive gamma-safeguarding are that it converges locally for singular problems, and it can detect nonsingular problems automatically, in which case the Newton-Anderson iterates are scaled towards a standard Newton step. The result is a flexible algorithm that performs well for singular and nonsingular problems, and can recover convergence from both standard Newton and Newton-Anderson with the right parameter choice. This leads to faster local convergence compared to both Newton's method, and Newton- Anderson without safeguarding, with effectively no additional computational cost. We demonstrate three strategies one can use when implementing Newton-Anderson and gamma-safeguarded Newton-Anderson to solve parameter-dependent problems near singular points. For our benchmark problems, we take two parameter-dependent incompressible flow systems: flow in a channel and Rayleigh-Benard convection.

math.NA

Accelerating the Computation of Tensor $Z$-eigenvalues

Efficient solvers for tensor eigenvalue problems are important tools for the analysis of higher-order data sets. Here we introduce, analyze and demonstrate an extrapolation method to accelerate the widely used shifted symmetric higher order power method for tensor $Z$-eigenvalue problems. We analyze the asymptotic convergence of the method, determining the range of extrapolation parameters that induce acceleration, as well as the parameter that gives the optimal convergence rate. We then introduce an automated method to dynamically approximate the optimal parameter, and demonstrate it's efficiency when the base iteration is run with either static or adaptively set shifts. Our numerical results on both even and odd order tensors demonstrate the theory and show we achieve our theoretically predicted acceleration.

math.NA

Filtering for Anderson acceleration

This work introduces, analyzes and demonstrates an efficient and theoretically sound filtering strategy to ensure the condition of the least-squares problem solved at each iteration of Anderson acceleration. The filtering strategy consists of two steps: the first controls the length disparity between columns of the least-squares matrix, and the second enforces a lower bound on the angles between subspaces spanned by the columns of that matrix. The combined strategy is shown to control the condition number of the least-squares matrix at each iteration. The method is shown to be effective on a range of problems based on discretizations of partial differential equations. It is shown particularly effective for problems where the initial iterate may lie far from the solution, and which progress through distinct preasymptotic and asymptotic phases.

math.NA

Newton-Anderson at Singular Points

In this paper we develop convergence and acceleration theory for Anderson acceleration applied to Newton's method for nonlinear systems in which the Jacobian is singular at a solution. For these problems, the standard Newton algorithm converges linearly in a region about the solution; and, it has been previously observed that Anderson acceleration can substantially improve convergence without additional a priori knowledge, and with little additional computation cost. We present an analysis of the Newton-Anderson algorithm in this context, and introduce a novel and theoretically supported safeguarding strategy. The convergence results are demonstrated with the Chandrasekhar H-equation and some standard benchmark examples.

math.NA

Anderson acceleration for a regularized Bingham model

This paper studies a finite element discretization of the regularized Bingham equations that describe viscoplastic flow. An efficient nonlinear solver for the discrete model is then proposed and analyzed. The solver is based on Anderson acceleration (AA) applied to a Picard iteration, and we show accelerated convergence of the method by applying AA theory (recently developed by the authors) to the iteration, after showing sufficient smoothness properties of the associated fixed point operator. Numerical tests of spatial convergence are provided, as are results of the model for 2D and 3D driven cavity simulations. For each numerical test, the proposed nonlinear solver is also tested and shown to be very effective and robust with respect to the regularization parameter as it goes to zero.

math.NA

Extrapolating the Arnoldi Algorithm To Improve Eigenvector Convergence

We consider extrapolation of the Arnoldi algorithm to accelerate computation of the dominant eigenvalue/eigenvector pair. The basic algorithm uses sequences of Krylov vectors to form a small eigenproblem which is solved exactly. The two dominant eigenvectors output from consecutive Arnoldi steps are then recombined to form an extrapolated iterate, and this accelerated iterate is used to restart the next Arnoldi process. We present numerical results testing the algorithm on a variety of cases and find on most examples it substantially improves the performance of restarted Arnoldi. The extrapolation is a simple post-processing step which has minimal computational cost.

math.NA

A simple extrapolation method for clustered eigenvalues

This paper introduces a simple variant of the power method. It is shown analytically and numerically to accelerate convergence to the dominant eigenvalue/eigenvector pair; and, it is particularly effective for problems featuring a small spectral gap. The introduced method is a one-step extrapolation technique that uses a linear combination of current and previous update steps to form a better approximation of the dominant eigenvector. The provided analysis shows the method converges exponentially with respect to the ratio between the two largest eigenvalues, which is also approximated during the process. An augmented technique is also introduced, and is shown to stabilize the early stages of the iteration. Numerical examples are provided to illustrate the theory and demonstrate the methods.

math.NA