arXiv ScienceSearch

arXiv subjects

Juntao Huang

Publications and source records attributed to Juntao Huang.

At least 19 recordsLinked to original sources

A discrete crack-tip theory for nonlinear lattice networks

Crack-tip fields govern deformation localization and failure initiation. Classical continuum fracture mechanics describes these fields through theories such as the Hutchinson-Rice-Rosengren (HRR) field for nonlinear power-law solids. However, continuum descriptions break down near cracks in soft and architected materials, where load is transmitted through discrete chains, fibers, or struts. Here, we develop a discrete crack-tip theory for lattice networks with nonlinear chains. The theory has two central components. First, at large deformation, the strain of a representative chain in layer $i$ depends approximately linearly on the applied macroscopic strain, $\varepsilon_i\approx k_i(\lambda-1)$, defining a layer-dependent strain-amplification factor $k_i$. Second, along topology-selected chain directions, termed discrete HRR lines, the layer-to-layer ratios of $k_i$ follow a two-regime scaling law. Together, for a power-law chain force-strain relation with exponent $p$, the inner discrete regime predicts $\varepsilon_i\sim i^{-1/p}$ and $f_i\sim i^{-1}$, which differs from the classical continuum HRR prediction. The theory also explains why the intrinsic fracture energy approaches a size-independent limit as the network size increases. Photoelastic hydrogel experiments further validate our theory. These results reveal a two-regime crack-tip scaling law in nonlinear lattice networks and provide a framework for predicting chain deformation and intrinsic fracture energy.

cond-mat.soft

AGoQ: Activation and Gradient Quantization for Memory-Efficient Distributed Training of LLMs

Quantization is a key method for reducing the GPU memory requirement of training large language models (LLMs). Yet, current approaches are ineffective for 4-bit activations and 8-bit gradients, which would easily cause slow convergence or accuracy loss. To address this, we introduce AGoQ, incorporating two new techniques: 1) a layer-aware activation quantization algorithm that allocates appropriate bit-widths for activations of various layers based on their types and pipeline stages to achieve near 4-bit activation storage, and 2) a gradient quantization algorithm that reduces memory usage and shortens communication time by employing 8-bit gradient storage and precision-preserving 8-bit All-Reduce communication. We conduct extensive experiments using different sizes of LLMs on two GPU clusters (up to 64 GPUs), and the experimental results show that our AGoQ reduces the memory by up to 52\% and achieves up to 1.34$\times$ improvement of training speed compared to state-of-the-art training systems Megatron-LM (w/ or w/o ZeRO), COAT and DeepSpeed with 8B to 32B LLaMA models, while achieving convergence loss on pretraining and comparable accuracy on downstream tasks with LLaMA architectures.

cs.CL

Machine learning moment closure models for the radiative transfer equation IV: enforcing symmetrizable hyperbolicity in two dimensions

This is our fourth work in the series on machine learning (ML) moment closure models for the radiative transfer equation (RTE). In the first three papers of this series, we considered the RTE in slab geometry in 1D1V (i.e. one dimension in physical space and one dimension in angular space), and introduced a gradient-based ML moment closure [1], then enforced the hyperbolicity through a symmetrizer [2], or together with physical characteristic speeds by learning the eigenvalues of the Jacobian matrix [3]. Here, we extend our framework to the RTE in 2D2V (i.e. two dimensions in physical space and two dimensions in angular space). The main idea is to preserve the leading part of the classical $P_N$ model and modify only the highest-order block row. By analyzing the structural properties of the $P_N$ model, we show that its coefficient matrices are symmetric and admit a block-tridiagonal structure. Then we use this property to introduce a block-diagonal symmetrizer for the ML moment model and derive explicit algebraic conditions on the closure blocks which guarantee the symmetrizable hyperbolicity of the resulting ML system. These conditions lead to a natural parametrization of the closure in terms of a symmetric positive definite matrix together with symmetric closure blocks, which can be learned from data while automatically enforcing symmetrizable hyperbolicity by construction. The numerical results show that the proposed framework improves upon the classical $P_N$ model while maintaining hyperbolicity.

math.NA

Complex Frequency Fingerprint: Interacting Driven Non-Hermitian Skin Effect

The excitation properties of quantum many-body systems are encoded in their response functions. These functions define an associated response Hamiltonian, which is intrinsically non-Hermitian due to the dissipative nature of retarded responses, even in closed systems. By analyzing its eigenvalues and eigenstates, one obtains a unique characterization of the system, referred to as the complex frequency fingerprint. Using this framework, we demonstrate that interactions alone can give rise to both point-gap topology and the non-Hermitian skin effect. Unlike the dissipation-induced skin effect, this interaction-driven phenomenon exhibits pronounced frequency dependence. We further introduce a complex-frequency density of states framework that distinctly separates non-Hermitian skin modes from topological edge modes.

cond-mat.mes-hall

Topological Mechanics of Entangled Networks

Entangled networks are ubiquitous in tissues, polymers, and fabrics. However, their mechanics remain insufficiently understood due to the complexity of the topological constraints at the network level. Here, we develop a mathematical framework that models entangled networks as graphs, capturing topological constraints of entanglements. We prove that entanglements reduce system energy by enabling uniform tension along chains crossing entanglements and by redistributing stress through sliding. Under this framework, we study elasticity and fracture, validated by experiments on entangled fabrics and hydrogels. For elasticity, entanglements increase strength by enabling stress homogeneity in the network. For fracture, entanglements enhance toughness by mitigating stress concentration around crack tips. We discover counterintuitive physical laws governing crack-tip stretch during crack opening: stress deconcentration at small deformation, constitutive-law independence at intermediate deformation, and linear scaling at large deformation. This framework establishes fundamental principles of linking topology to mechanics of entangled networks and offers a foundational tool for designing reconfigurable materials.

cond-mat.soft

Moment-enhanced shallow-water equations with an effective wall closure for no-slip bottoms

Shallow-water equations and low-order shallow-water moment models use vertically coarse representations and therefore cannot, in general, resolve the thin wall-affected region produced by a no-slip bottom. Enforcing the pointwise wall value on a low-order global polynomial reconstruction can introduce stiff relaxation and distort the resolved interior velocity profile. Starting from the incompressible Navier--Stokes equations with Navier bottom friction, we derive a bottom-to-mean relation in a distinguished regular-friction regime and use it to define an endpoint-consistent effective wall-traction closure for the shallow-water equations and the hyperbolic shallow-water moment equations. The closure represents the momentum effect of unresolved near-wall dynamics; it neither resolves the physical boundary layer nor imposes the pointwise no-slip trace on the reconstructed polynomial. It recovers the perfect-slip wall contribution when the friction coefficient vanishes. Because only source terms are changed, the homogeneous principal matrices and their established two-dimensional hyperbolicity classification remain unchanged. We compare the standard and modified reduced models with two-phase incompressible Navier--Stokes computations in OpenFOAM for wet-bed dam-break and three-dimensional collapse tests. In the cases considered, the modified closure reduces the excessive damping of the classical low-order wall source and improves agreement in depth-averaged and resolved-interior velocity diagnostics, but it does not uniformly improve front-propagation speed. The regular-friction asymptotic remainder is not uniform in the large-friction numerical regime; there the effective coefficient is used as a wall-model continuation and assessed empirically.

math.NA

Complex Frequency Detection in a Subsystem

In this study, we systematically explore the non-Hermitian skin effect (NHSE) and its associated complex-frequency detection in the context of a frequency-dependent non-Hermitian Hamiltonian. This Hamiltonian arises from the self-energy correction of the subsystem and can be calculated exactly within our theoretical model, without the need for any approximations. Additionally, complex frequency detection, which encompasses complex frequency excitation, synthesis, and fingerprint, enables us to detect the physical responses induced by the complex driving frequencies. Our calculations reveal that both complex frequency excitation and synthesis are incompatible with the non-Hermitian approximation and are unable to characterize the presence or abscence of the NHSE. In contrast, the complex-frequency fingerprint successfully detects the novel responses induced by the NHSE through the introduction of a double-frequency Green's function. Our work provides a platform for studying non-Hermitian physics and their novel response in quantum systems rigorously without relying on any approximations.

cond-mat.mes-hall

Machine learning-based moment closure model for the linear Boltzmann equation with uncertainties

The Boltzmann equation, a fundamental equation in kinetic theory, serves as a bridge between microscopic particle dynamics and macroscopic continuum mechanics. However, deriving closed macroscopic moment systems from the Boltzmann equation remains a long-standing challenge due to the intrinsic non-closure of the moment hierarchy. In this paper, we propose a machine learning (ML)-based moment closure model for the linear Boltzmann equation, addressing both the deterministic and stochastic settings. Our approach leverages neural networks to learn the spatial gradient of the unclosed highest-order moment, enabling effective training through natural output normalization. For the deterministic problem, to ensure global hyperbolicity and stability, we derive and apply the constraints that enforce symmetrizable hyperbolicity of the system. For the stochastic problem, we adopt the generalized polynomial chaos (gPC)-based stochastic Galerkin method to discretize the random variables, resulting in a system for which the approach in the deterministic case can be used similarly. Several numerical experiments are shown to demonstrate the effectiveness and accuracy of our ML-based moment closure model for the linear Boltzmann equation with or without uncertainties.

math.NA

Complex Frequency Fingerprint: Basic Concept and Theory

We introduce the complex frequency fingerprint (CFF), an experimentally accessible method for detecting the complex frequency Green's function (GF). Unlike the real frequency GF, where $\omega$ is real, this complex frequency GF is shown to play a necessary role in both non-Hermitian and quantum many-body systems. For non-Hermitian systems, we will prove that our method detects complex energy spectra, eigenstates, and complex frequency GFs throughout the complex plane, providing necessary identification of the non-Hermitian skin effect. For quantum many-body systems, our method reveals quasiparticle peaks across the complex plane and intuitively illustrates interaction effects. This information is difficult to obtain with real frequency detection. Our method paves the way for exploring exotic phenomena in both non-Hermitian and quantum many-body systems, bridging theory and experiment across diverse physical areas.

cond-mat.mes-hall

Uniform accuracy of implicit-explicit Runge-Kutta methods for linear hyperbolic relaxation systems

In this paper, we study the uniform accuracy of implicit-explicit (IMEX) Runge-Kutta (RK) schemes for general linear hyperbolic relaxation systems satisfying the structural stability condition proposed in \cite{yong_singular_1999}. We establish the uniform stability and accuracy of a class of IMEX-RK schemes with spatial discretization using a Fourier spectral method. Our results demonstrate that the accuracy of the fully discretized schemes is independent of the relaxation time across all regimes. Numerical experiments on applications in traffic flows and kinetic theory verify our theoretical analysis.

math.NA

Hyperbolic Machine Learning Moment Closures for the BGK Equations

We introduce a hyperbolic closure for the Grad moment expansion of the Bhatnagar-Gross-Krook's (BGK) kinetic model using a neural network (NN) trained on BGK's moment data. This closure is motivated by the exact closure for the free streaming limit that we derived in our paper on closures in transport \cite{Huang2022-RTE1}. The exact closure relates the gradient of the highest moment to the gradient of four lower moments. As with our past work, the model presented here learns the gradient of the highest moment in terms of the coefficients of gradients for all lower ones. By necessity, this means that the resulting hyperbolic system is not conservative in the highest moment. For stability, the output layers of the NN are designed to enforce hyperbolicity and Galilean invariance. This ensures the model can be run outside of the training window of the NN. Unlike our previous work on radiation transport that dealt with linear models, the BGK model's nonlinearity demanded advanced training tools. These comprised an optimal learning rate discovery, one cycle training, batch normalization in each neural layer, and the use of the \texttt{AdamW} optimizer. To address the non-conservative structure of the hyperbolic model, we adopt the FORCE numerical method to achieve robust solutions. This results in a comprehensive computing model combining learned closures with methods for solving hyperbolic models. The proposed model can capture accurate moment solutions across a broad spectrum of Knudsen numbers. Our paper details the multi-scale model construction and is run on a range of test problems.

math.NA

Uniform accuracy of implicit-explicit backward differentiation formulas (IMEX-BDF) for linear hyperbolic relaxation systems

This work is concerned with the uniform accuracy of implicit-explicit backward differentiation formulas for general linear hyperbolic relaxation systems satisfying the structural stability condition proposed previously by the third author. We prove the uniform stability and accuracy of a class of IMEX-BDF schemes discretized spatially by a Fourier spectral method. The result reveals that the accuracy of the fully discretized schemes is independent of the relaxation time in all regimes. It is verified by numerical experiments on several applications to traffic flows, rarefied gas dynamics and kinetic theory.

math.NA

On the rotational invariance and hyperbolicity of shallow water moment equations in two dimensions

In this paper, we investigate the two-dimensional extension of a recently introduced set of shallow water models based on a regularized moment expansion of the incompressible Navier-Stokes equations \cite{kowalski2017moment,koellermeier2020analysis}. We show the rotational invariance of the proposed moment models with two different approaches. The first proof involves the split of the coefficient matrix into the conservative and non-conservative parts and proves the rotational invariance for each part, while the second one relies on the special block structure of the coefficient matrices. With the aid of rotational invariance, the analysis of the hyperbolicity for the moment model in 2D is reduced to the real diagonalizability of the coefficient matrix in 1D. Then we analyze the real diagonalizability by deriving the analytical form of the characteristic polynomial. We find that the moment model in 2D is hyperbolic in most cases and weakly hyperbolic in a degenerate edge case. With a simple modification to the coefficient matrices, we fix this weakly hyperbolicity and propose a new global hyperbolic model. Furthermore, we extend the model to include a more general class of closure relations than the original model and establish that this set of general closure relations retains both rotational invariance and hyperbolicity.

math.NA

Decoding flat bands from compact localized states

The flat band system is an ideal quantum platform to investigate the kaleidoscope created by the electron-electron correlation effects. The central ingredient of realizing a flat band is to find its compact localized states. In this work, we develop a systematic way to generate the compact localized states by designing destructive interference pattern from 1-dimensional chains. A variety of 2-dimensional new flat band systems are constructed with this method. Furthermore, we show that the method can be extended to generate the compact localized states in multi-orbital systems by carefully designing the block hopping scheme, as well as in quasicrystal and disorder systems.

cond-mat.str-el

Bound-preserving discontinuous Galerkin methods with modified Patankar time integrations for chemical reacting flows

In this paper, we develop bound-preserving discontinuous Galerkin (DG) methods for chemical reactive flows. There are several difficulties in constructing suitable numerical schemes. First of all, the density and internal energy are positive, and the mass fraction of each species is between 0 and 1. Secondly, due to the rapid reaction rate, the system may contain stiff sources, and the strong-stability-preserving explicit Runge-Kutta method may result in limited time step sizes. To obtain physically relevant numerical approximations, we apply the bound-preserving technique to the DG methods. For time discretization, we apply the modified Runge-Kutta/multi-step Patankar methods, which are explicit for the flux while implicit for the source. Such methods can handle stiff sources with relatively large time steps, preserve the positivity of the target variables, and keep the summation of the mass fractions to be 1. Finally, it is not straightforward to combine the bound-preserving DG methods and the Patankar time integrations. The positivity-preserving technique for DG method requires positive numerical approximations at the cell interfaces, while Patankar methods can keep the positivity of the pre-selected point-values of the target variables. To match the degree of freedom, we use $Q^k$ polynomials on rectangular meshes for problems in two space dimensions. To evolve in time, we first read the polynomials at the Gaussian points. Then suitable slope limiters can be applied to enforce the positivity of the solutions at those points, which can be preserved by the Patankar methods, leading to positive updated numerical cell averages. In addition, we use another slope limiter to get positive solutions used for the bound-preserving technique for the flux.

math.NA

Adaptive sparse grid discontinuous Galerkin method: review and software implementation

This paper reviews the adaptive sparse grid discontinuous Galerkin (aSG-DG) method for computing high dimensional partial differential equations (PDEs) and its software implementation. The C\texttt{++} software package called AdaM-DG, implementing the aSG-DG method, is available on Github at \url{https://github.com/JuntaoHuang/adaptive-multiresolution-DG}. The package is capable of treating a large class of high dimensional linear and nonlinear PDEs. We review the essential components of the algorithm and the functionality of the software, including the multiwavelets used, assembling of bilinear operators, fast matrix-vector product for data with hierarchical structures. We further demonstrate the performance of the package by reporting numerical error and CPU cost for several benchmark test, including linear transport equations, wave equations and Hamilton-Jacobi equations.

math.NA

Superconvergence and accuracy enhancement of discontinuous Galerkin solutions for Vlasov-Maxwell equations

This paper considers the discontinuous Galerkin (DG) methods for solving the Vlasov-Maxwell (VM) system, a fundamental model for collisionless magnetized plasma. The DG methods provide accurate numerical description with conservation and stability properties. However, to resolve the high dimensional probability distribution function, the computational cost is the main bottleneck even for modern-day supercomputers. This work studies the applicability of a post-processing technique to the DG solution to enhance its accuracy and resolution for the VM system. In particular, we prove the superconvergence of order $(2k+\frac{1}{2})$ in the negative order norm for the probability distribution function and the electromagnetic fields when piecewise polynomial degree $k$ is used. Numerical tests including Landau damping, two-stream instability and streaming Weibel instabilities are considered showing the performance of the post-processor.

math.NA

Coupling conditions for linear hyperbolic relaxation systems in two-scales problems

This work is concerned with coupling conditions for linear hyperbolic relaxation systems with multiple relaxation times. In the region with small relaxation time, an equilibrium system can be used for computational efficiency. Under the assumption that the relaxation system satisfies the structural stability condition and the interface is non-characteristic, we derive a coupling condition at the interface to couple the two systems in a domain decomposition setting. We prove the validity by the energy estimate and Laplace transform, which shows how the error of the domain decomposition method depends on the smaller relaxation time and the boundary layer effects. In addition, we propose a discontinuous Galerkin (DG) scheme for solving the interface problem with the derived coupling condition and prove the L2 stability. We validate our analysis on the linearized Carleman model and the linearized Grad's moment system and show the effectiveness of the DG scheme.

math.NA