arXiv ScienceSearch

arXiv subjects

Graham Harper

Publications and source records attributed to Graham Harper.

13 recordsLinked to original sources

The unique capabilities of HST for stellar physics: Probing Atmospheric Structure, Chromospheres, and Mass Loss of Evolved Stars

Evolved stars are among the primary sources of chemical enrichment and dust production in galaxies. During the giant phases, stars return a substantial fraction of their mass to the interstellar medium (ISM) through stellar winds, enriching galaxies with newly synthesized elements and dust. However, the atmospheric structure and physical processes that initiate mass loss remain poorly constrained observationally. Understanding the origin, structure, and evolution of stellar chromospheres remains a long-standing problem in stellar astrophysics. While the mechanisms responsible for chromospheric heating and atmospheric dynamics are not fully understood even in the Sun, they become more complex in evolved stars due to pulsation, shocks, convection, extended atmospheres, and possible magnetic activity. Determining the thermal, density, and velocity structure of these extended atmospheres is therefore essential for understanding atmospheric heating, the onset of mass loss, and the late stages of stellar evolution. High-resolution NUV and FUV spectroscopy (R ~ 30,000-100,000) provided by HST/STIS occupies a unique observational parameter space that cannot be replaced by existing facilities. HST/STIS therefore remains essential for understanding the atmospheric physics and mass-loss processes of evolved stars. We highlight the need to preserve and prioritize high-resolution NUV and FUV spectroscopic capabilities with HST. Such programs would provide essential benchmarks for stellar atmosphere modeling, complement ongoing ALMA and optical observations, and help define future UV-optical capabilities for the Habitable Worlds Observatory (HWO).

astro-ph.SR

Multilevel Training for Kolmogorov Arnold Networks

Algorithmic speedup of training common neural architectures is made difficult by the lack of structure guaranteed by the function compositions inherent to such networks. In contrast to multilayer perceptrons (MLPs), Kolmogorov-Arnold networks (KANs) provide more structure by expanding learned activations in a specified basis. This paper exploits this structure to develop practical algorithms and theoretical insights, yielding training speedup via multilevel training for KANs. To do so, we first establish an equivalence between KANs with spline basis functions and multichannel MLPs with power ReLU activations through a linear change of basis. We then analyze how this change of basis affects the geometry of gradient-based optimization with respect to spline knots. The KANs change-of-basis motivates a multilevel training approach, where we train a sequence of KANs naturally defined through a uniform refinement of spline knots with analytic geometric interpolation operators between models. The interpolation scheme enables a ``properly nested hierarchy'' of architectures, ensuring that interpolation to a fine model preserves the progress made on coarse models, while the compact support of spline basis functions ensures complementary optimization on subsequent levels. Numerical experiments demonstrate that our multilevel training approach can achieve orders of magnitude improvement in accuracy over conventional methods to train comparable KANs or MLPs, particularly for physics informed neural networks. Finally, this work demonstrates how principled design of neural networks can lead to exploitable structure, and in this case, multilevel algorithms that can dramatically improve training performance.

cs.LG

Leveraging KANs for Expedient Training of Multichannel MLPs via Preconditioning and Geometric Refinement

Multilayer perceptrons (MLPs) are a workhorse machine learning architecture, used in a variety of modern deep learning frameworks. However, recently Kolmogorov-Arnold Networks (KANs) have become increasingly popular due to their success on a range of problems, particularly for scientific machine learning tasks. In this paper, we exploit the relationship between KANs and multichannel MLPs to gain structural insight into how to train MLPs faster. We demonstrate the KAN basis (1) provides geometric localized support, and (2) acts as a preconditioned descent in the ReLU basis, overall resulting in expedited training and improved accuracy. Our results show the equivalence between free-knot spline KAN architectures, and a class of MLPs that are refined geometrically along the channel dimension of each weight tensor. We exploit this structural equivalence to define a hierarchical refinement scheme that dramatically accelerates training of the multi-channel MLP architecture. We show further accuracy improvements can be had by allowing the $1$D locations of the spline knots to be trained simultaneously with the weights. These advances are demonstrated on a range of benchmark examples for regression and scientific machine learning.

cs.LG

Trilinos: Enabling Scientific Computing Across Diverse Hardware Architectures at Scale

Trilinos is a community-developed, open-source software framework that facilitates building large-scale, complex, multiscale, multiphysics simulation code bases for scientific and engineering problems. Since the Trilinos framework has undergone substantial changes to support new applications and new hardware architectures, this document is an update to ``An Overview of the Trilinos project'' by Heroux et al. (ACM Transactions on Mathematical Software, 31(3):397-423, 2005). It describes the design of Trilinos, introduces its new organization in product areas, and highlights established and new features available in Trilinos. Particular focus is put on the modernized software stack based on the Kokkos ecosystem to deliver performance portability across heterogeneous hardware architectures. This paper also outlines the organization of the Trilinos community and the contribution model to help onboard interested users and contributors.

cs.MS

R-Adaptive Mesh Optimization to Enhance Finite Element Basis Compression

Modern computing systems are capable of exascale calculations, which are revolutionizing the development and application of high-fidelity numerical models in computational science and engineering. While these systems continue to grow in processing power, the available system memory has not increased commensurately, and electrical power consumption continues to grow. A predominant approach to limit the memory usage in large-scale applications is to exploit the abundant processing power and continually recompute many low-level simulation quantities, rather than storing them. However, this approach can adversely impact the throughput of the simulation and diminish the benefits of modern computing architectures. We present two novel contributions to reduce the memory burden while maintaining performance in simulations based on finite element discretizations. The first contribution develops dictionary-based data compression schemes that detect and exploit the structure of the discretization, due to redundancies across the finite element mesh. These schemes are shown to reduce memory requirements by more than 99 percent on meshes with large numbers of nearly identical mesh cells. For applications where this structure does not exist, our second contribution leverages a recently developed augmented Lagrangian sequential quadratic programming algorithm to enable r-adaptive mesh optimization, with the goal of enhancing redundancies in the mesh. Numerical results demonstrate the effectiveness of the proposed methods to detect, exploit and enhance mesh structure on examples inspired by large-scale applications.

math.OC

Entropy-based feature selection for capturing impacts in Earth system models with extreme forcing

This paper presents the development of a new entropy-based feature selection method for identifying and quantifying impacts. Here, impacts are defined as statistically significant differences in spatio-temporal fields when comparing datasets with and without an external forcing in Earth system models. Temporal feature selection is performed by first computing the cross-fuzzy entropy to quantify similarity of patterns between two datasets and then applying changepoint detection to identify regions of statistically constant entropy. The method is used to capture temperate north surface cooling from a 9-member simulation ensemble of the Mt. Pinatubo volcanic eruption, which injected 10 Tg of SO2 into the stratosphere. The results estimate a mean difference decrease in near surface air temperature of -0.560 K with a 99% confidence interval between -0.864 K and -0.257 K between April and November of 1992, one year following the eruption. A sensitivity analysis with decreasing SO2 injection revealed that the impact is statistically significant at 5 Tg but not at 3 Tg. Using identified features, a dependency graph model composed of 68 nodes and 229 edges directly connecting initial aerosol optical depth changes in the tropics to solar flux and temperature changes before the temperate north surface cooling is presented.

stat.AP

In-situ data extraction for pathway analysis in an idealized atmosphere configuration of E3SM

We propose an approach for characterizing source-impact pathways, the interactions of a set of variables in space-time due to an external forcing, in climate models using in-situ analyses that circumvent computationally expensive read/write operations. This approach makes use of a lightweight open-source software library we developed known as CLDERA-Tools. We describe how CLDERA-Tools is linked with the U.S. Department of Energy's Energy Exascale Earth System Model (E3SM) in a minimally invasive way for in-situ extraction of quantities of interested and associated statistics. Subsequently, these quantities are used to represent source-impact pathways with time-dependent directed acyclic graphs (DAGs). The utility of CLDERA-Tools is demonstrated by using the data it extracts in-situ to compute a spatially resolved DAG from an idealized configuration of the atmosphere with a parameterized representation of a volcanic eruption known as HSW-V.

cs.DC

Monolithic Multigrid Preconditioners for High-Order Discretizations of Stokes Equations

This work introduces and assesses the efficiency of a monolithic $ph$MG multigrid framework designed for high-order discretizations of stationary Stokes systems using Taylor-Hood and Scott-Vogelius elements. The proposed approach integrates coarsening in both approximation order ($p$) and mesh resolution ($h$), to address the computational and memory efficiency challenges that are often encountered in conventional high-order numerical simulations. Our numerical results reveal that $ph$MG offers significant improvements over traditional spatial-coarsening-only multigrid ($h$MG) techniques for problems discretized with Taylor-Hood elements across a variety of problem sizes and discretization orders. In particular, the $ph$MG method exhibits superior performance in reducing setup and solve times, particularly when dealing with higher discretization orders and unstructured problem domains. For Scott-Vogelius discretizations, while monolithic $ph$MG delivers low iteration counts and competitive solve phase timings, it exhibits a discernibly slower setup phase when compared to a multilevel (non-monolithic) full-block-factorization (FBF) preconditioner where $ph$MG is employed only for the velocity unknowns. This is primarily due to the setup costs of the larger mixed-field relaxation patches with monolithic $ph$MG versus the patch setup costs with a single unknown type for FBF.

math.NA

Compression and Reduced Representation Techniques for Patch-Based Relaxation

Patch-based relaxation refers to a family of methods for solving linear systems which partitions the matrix into smaller pieces often corresponding to groups of adjacent degrees of freedom residing within patches of the computational domain. The two most common families of patch-based methods are block-Jacobi and Schwarz methods, where the former typically corresponds to non-overlapping domains and the later implies some overlap. We focus on cases where each patch consists of the degrees of freedom within a finite element method mesh cell. Patch methods often capture complex local physics much more effectively than simpler point-smoothers such as Jacobi; however, forming, inverting, and applying each patch can be prohibitively expensive in terms of both storage and computation time. To this end, we propose several approaches for performing analysis on these patches and constructing a reduced representation. The compression techniques rely on either matrix norm comparisons or unsupervised learning via a clustering approach. We illustrate how it is frequently possible to retain/factor less than 5% of all patches and still develop a method that converges with the same number of iterations or slightly more than when all patches are stored/factored.

math.NA

Stars at High Spatial Resolution

We summarize some of the compelling new scientific opportunities for understanding stars and stellar systems that can be enabled by sub-milliarcsec (sub-mas) angular resolution, UV-Optical spectral imaging observations, which can reveal the details of the many dynamic processes (e.g., evolving magnetic fields, accretion, convection, shocks, pulsations, winds, and jets) that affect stellar formation, structure, and evolution. These observations can only be provided by long-baseline interferometers or sparse aperture telescopes in space, since the aperture diameters required are in excess of 500 m (a regime in which monolithic or segmented designs are not and will not be feasible) and since they require observations at wavelengths (UV) not accessible from the ground. Such observational capabilities would enable tremendous gains in our understanding of the individual stars and stellar systems that are the building blocks of our Universe and which serve as the hosts for life throughout the Cosmos.

astro-ph.SR

Fundamental Stellar Properties from Optical Interferometry

High-resolution observations by visible and near-infrared interferometers of both single stars and binaries have made significant contributions to the foundations that underpin many aspects of our knowledge of stellar structure and evolution for cool stars. The CS16 splinter on this topic reviewed contributions of optical interferometry to date, examined highlights of current research, and identified areas for contributions with new observational constraints in the near future.

astro-ph.SR

Mass Transport Processes and their Roles in the Formation, Structure, and Evolution of Stars and Stellar Systems

We summarize some of the compelling new scientific opportunities for understanding stars and stellar systems that can be enabled by sub-mas angular resolution, UV/Optical spectral imaging observations, which can reveal the details of the many dynamic processes (e.g., variable magnetic fields, accretion, convection, shocks, pulsations, winds, and jets) that affect their formation, structure, and evolution. These observations can only be provided by long-baseline interferometers or sparse aperture telescopes in space, since the aperture diameters required are in excess of 500 m - a regime in which monolithic or segmented designs are not and will not be feasible - and since they require observations at wavelengths (UV) not accessible from the ground. Two mission concepts which could provide these invaluable observations are NASA's Stellar Imager (SI; http://hires.gsfc.nasa.gov/si/) interferometer and ESA's Luciola sparse aperture hypertelescope, which each could resolve hundreds of stars and stellar systems. These observatories will also open an immense new discovery space for astrophysical research in general and, in particular, for Active Galactic Nuclei (Kraemer et al. Decadal Survey Science Whitepaper). The technology developments needed for these missions are challenging, but eminently feasible (Carpenter et al. Decadal Survey Technology Whitepaper) with a reasonable investment over the next decade to enable flight in the 2025+ timeframe. That investment would enable tremendous gains in our understanding of the individual stars and stellar systems that are the building blocks of our Universe and which serve as the hosts for life throughout the Cosmos.

astro-ph.SR