arXiv ScienceSearch

arXiv subjects

Timon Rabczuk

Publications and source records attributed to Timon Rabczuk.

At least 19 recordsLinked to original sources

Beyond Residuals: Energy based solutions of partial differential equations using scientific machine learning

Energy-based approaches provide a natural and physically consistent framework for a large class of partial differential equations arising in solid and fluid mechanics, where the governing equations follow from variational principles. In contrast to residual-based physics-informed neural networks (PINNs) and their weak-form variants, which enforce the strong or weak form of the equations through loss minimization, the Deep Energy Method (DEM) directly computes the solution as the minimizer of an energy or incremental potential functional. This eliminates the need for residual weighting, avoids high-order derivatives, and enables the direct enforcement of physical constraints through the variational structure.In this work, we systematically revisit the Deep Energy Method, placing it in the broader context of physics-informed learning and variational modeling. We clarify the relationship between DEM, PINNs, and VPINNs, and identify the class of problems for which energy minimization provides intrinsic advantages in terms of stability, robustness and interpretability. Particular emphasis is placed on incremental variational formulations, which allow DEM to be applied to nonlinear, history-dependent and time-dependent problems, including phase-field fracture and dissipative systems. The variational structure underlying DEM further provides a natural foundation for optimization and inverse problems, where the energy functional acts as a physics-based constraint rather than a residual penalty. Through a series of numerical examples, we demonstrate that DEM offers a principled and effective alternative to residual-based methods for variational problems, highlighting its strengths and limitations relative to existing physics-informed approaches.

physics.comp-ph

A mesh-free multiresolution deep energy method with phase-field modeling of brittle fracture

Phase-field modeling of brittle fracture removes the need to track cracks explicitly by recasting their evolution as the minimization of an energy functional. In return it requires a discretization dense enough to resolve a localization band whose width is set by a regularization length and whose path is not known in advance. We propose a mesh-free discretization in which a single neural network represents the displacement and phase fields and is trained by minimizing the incremental energy directly. The coordinates enter the network through a multiresolution feature encoding built from $C^1$ quadratic B-spline grids, so the finest scale the representation can express is set by choice rather than reached through slow training, and the energy is estimated by stratified Monte Carlo integration on points redrawn at every optimizer iteration. This pairing proves critical, since the crack fails to advance both when the integration points are held fixed and when the encoding is too coarse to represent the band, while each ingredient tolerates a wide range of settings once the other is in place. Because the representation is globally $C^1$, the second- and the fourth-order fracture energy densities run on the identical discretization. Across six problems, from single-edge-notched tension and shear to a thick-walled ring on a single spline patch, the computed load-displacement curves follow staggered finite element references at matched regularization length, with peak loads within about 1% on the single-edge-notched tests and within 8% where the crack pattern changes topology. On a public benchmark dataset of random multi-crack configurations the method classifies the active or dormant state of 90% of the seeded cracks in twenty zero-shot runs, where the deep Ritz baseline of the dataset authors fails.

cs.LG

The role of weak interfaces in the tensile deformation and fracture of particle-filled polymers studied by phase-field model

Weak particle-matrix interfaces play a critical role in the tensile fracture of particle-filled polymer composites, but how they govern progressive debonding, fracture localization, and the resulting changes in macroscopic mechanical properties remains insufficiently understood. In this study, a cohesive-zone phase-field model incorporating a hyperelastic polymer matrix and a smeared interface is employed to investigate the coupled evolution of interfacial debonding and matrix fracture in particle-filled polymer composites. The model is calibrated against and compared with uniaxial tensile responses of particle-filled polyurethane composites and then used to study how interfacial strength, interfacial fracture energy, and matrix fracture properties affect the macroscopic stress-strain response and damage evolution. The results show that weak interfaces can induce an intermediate softening regime in the stress-strain response, characterized by a reduced effective tangent stiffness and associated with distributed interfacial damage. Interfacial strength mainly controls the initiation of debonding, whereas interfacial fracture energy affects whether debonding can develop progressively in a distributed manner or rapidly localizes into a dominant crack band. Comparisons with well-bonded reference systems further demonstrate that weak interfaces may reduce the maximum stress but increase the strain at break by promoting distributed debonding around particles and delaying the formation of a dominant crack band. These findings clarify the dual role of weak interfaces and provide a mechanistic understanding of interface-controlled tensile failure in particle-filled polymer composites.

cond-mat.mtrl-sci

Plasolver: Physics-Informed Neural Operators for Elastoplasticity

Elastoplastic analysis is computationally demanding because its nonlinear, path-dependent constitutive behavior requires incremental loading and repeated iterative solutions. To address this challenge, we propose Plasolver, a physics-informed neural operator framework that combines the efficiency of operator learning with the accuracy and robustness of classical numerical solvers. Plasolver consists of a physics-informed pretraining stage and an optional warm-start stage. During pretraining, the neural operator is trained solely by minimizing the incremental potential energy of elastoplasticity formulated by Simo, without requiring any labeled solution data. It operates directly on unstructured point clouds by encoding spatial coordinates, loading histories, and material properties as unified point-wise prompts. This formulation provides dual invariance to spatial and loading-path discretizations, enabling consistent predictions across different spatial resolutions and different numbers of increments representing the same loading trajectory. The pretrained Plasolver achieves relative errors on the order of 1\% while providing approximately two orders of magnitude acceleration over conventional finite element simulations. In the warm-start stage, the pretrained prediction is supplied as the initial solution to a classical iterative solver, preserving its numerical accuracy, robustness, and convergence properties while substantially accelerating convergence. Numerical results show that Plasolver reduces the required number of iterations by approximately 50\% compared with conventional zero-initialized solvers and converges to solutions at any prescribed tolerance. Plasolver thus provides an efficient, accurate, and discretization-invariant computational framework for nonlinear, path-dependent elastoplastic problems.

physics.comp-ph

Neural Operators for Immersed-Boundary Soft Swimmers Locomotion

High-fidelity immersed-boundary simulation resolves the coupled motion of a deforming swimmer and its surrounding flow, but the resulting cost limits repeated evaluations for engineering design, parameter studies, and control. We develop neural-operator surrogates for temporal prediction of the hydrodynamic fields generated by planar and volumetric eel swimmers. The surrogates are trained on regular-grid fields exported from adaptive fluid--structure simulations and are conditioned on swimmer geometry and Reynolds number. The planar model jointly predicts two velocity components, scalar vorticity, and pressure. On five held-out high-Reynolds-number trajectories, its full-domain global relative L^2 error is 3.51 %. The volumetric formulation uses three target-specific models with a common multichannel input: one model predicts three-dimensional velocity, one predicts vorticity, and one predicts pressure. Their full-domain global relative L^2 errors on five held-out within-range trajectories are 3.44 %, 5.58 %, and 19.2 %. Together, the results demonstrate the feasibility of field-resolved neural surrogates for moving-boundary swimmer flows while identifying pressure accuracy and physical consistency as priorities for further development.

cs.LG

FEVessel: Mesh-Independent Analysis of 3D Pressure Vessels with the Label-Free Pretrained Finite Element Method

Pressure vessel analysis in the chemical, nuclear, and new-energy industries requires solving the same elasticity problem across many materials, geometries, and loads, where mesh quality and repeated solving govern both accuracy and cost. The finite element method (FEM) cannot amortise this repeated cost and fails on degenerate meshes, while the neural operators meant to replace it still need labelled data that FEM must generate. This paper proposes FEVessel, an adaptation of the Pretrained Finite Element Method (PFEM) to three-dimensional (3D) pressure vessels, and validates four capabilities across the two limitations above. FEVessel i) encodes each vessel as a point cloud with coordinate, material, and load channels, ii) pretrains a Transolver operator on the total potential energy instead of FEM labels, and iii) warm-starts iterative solvers with its prediction. A single model generalises across material, geometry, and boundary conditions at a $1.35\%$ relative displacement error, and its $2.07\%$ strain error is about $4.7$ times lower than that of a supervised Fourier neural operator ($9.72\%$), whose structured grid cannot preserve the through-thickness strain. Its warm start cuts algebraic multigrid iterations from $195$ to $18$, a $9.2\times$ end-to-end wall-clock speedup at the $10^{-3}$ engineering tolerance. The model transfers across mesh resolutions without retraining, holding about $3\%$ error at only $30\%$ of the training point density. On inverted and sliver meshes where FEM fails, the error remains below $3.66\%$. To our knowledge, this is the first systematic study of mesh-independent solution on industrially relevant 3D pressure vessels with degenerate meshes. Because training needs no labels, FEVessel works exactly where FEM cannot supply any, removing manual mesh repair from the analysis pipeline.

math.NA

GA-VINO: A Geometry-Aware Variational Physics-informed Neural Operator for Mindlin-Reissner Plates

Plate and shell structures are widely used in engineering fields. Rapid response prediction for such structures under complex geometries, heterogeneous materials, and varying loads is important for engineering design, but conventional numerical methods usually require repeated modeling and solution when the physical configuration changes. To address this issue, this study proposes a geometry-aware variational physics-informed neural operator (GA-VINO) for Mindlin-Reissner plates. GA-VINO represents the plate geometry using boundary point clouds and incorporates a material encoder, a load encoder, and a scalar-parameter branch to handle spatially random material fields, spatially varying pressure loads, and sample-level uniform parameters. Through multi-branch point cloud encoding and cross-attention, GA-VINO fuses geometric, material, loading, and query point information, and predicts the transverse deflection and rotations at arbitrary query locations. Unlike conventional data-driven neural operators, GA-VINO requires no labeled solution data during training. Instead, it minimizes a variational physics-informed loss constructed from the discretized total potential energy of the Mindlin-Reissner plate. Compared with grid-based neural operators, GA-VINO directly processes irregular point clouds and allows different physical fields to be discretized on different point sets, avoiding forced interpolation onto a common grid. The method is validated on multiple examples involving different geometries, material fields, and load distributions. The results show that GA-VINO achieves promising accuracy in deflection, rotation, gradient-sensitive, and energy-based metrics, completes full-field inference for new samples within milliseconds, and exhibits promising cross-geometry generalization capability.

cs.AI

HAMNO: A Hierarchical Adaptive Multi-scale Neural Operator with Physics-Informed Learning for Dynamical Systems

Neural operators provide a powerful framework for learning solution mappings of partial differential equations directly in function space. However, many existing architectures still struggle to represent nonlinear time-dependent systems that involve multi-scale structures, long-range interactions, and stable long-time evolution. In this work, we introduce the Hierarchical Adaptive Multi-scale Neural Operator (HAMNO), a neural-operator architecture that combines local convolutional representations, global spectral operators, and hierarchical encoder-decoder processing. The central component of HAMNO is a data-dependent gating mechanism that adaptively balances local and global information at each spatial location, allowing the model to resolve fine-scale features while preserving long-range dependencies. We further develop a physics-informed extension, PI-HAMNO, based on a multi-objective loss strategy that combines data fitting with strong- and weak-form physics constraints. The strong-form term penalizes the domain-integrated squared PDE residual in physical coordinates, while the weak-form term is constructed by multiplying the governing residual by finite-element test functions and evaluating the resulting element integrals using centroid-based tetrahedral quadrature. The framework is evaluated on non-periodic Allen-Cahn (AC), Cahn-Hilliard (CH), and Swift-Hohenberg (SH) equations defined on cubic domains. Across long-horizon rollout, data-limited training, out-of-distribution initial-condition shifts, and random-seed variations, HAMNO improves predictive accuracy over standard neural-operator baselines, while PI-HAMNO further enhances stability, physical consistency, and data efficiency. The implementation is publicly available at https://github.com/MBamdad/HAMNO .

cs.LG

Dmsh: A Multi-Agent Reinforcement Learning Framework for All-Quad Mesh Generation

Generating high-quality meshes for arbitrary geometries remains a fundamental bottleneck in computational engineering, often demanding heuristic tuning and semi-manual workflows. In this paper, we introduce Dmsh, a first fully automated reinforcement learning pipeline that unifies geometric decomposition and quadrilateral mesh generation within a single learning-based framework. Dmsh decomposes the problem through three coordinated agents handling topology simplification, geometric regularization, and mesh generation. The meshing process is formulated as a Markov Decision Process and solved using a parametric Soft Actor-Critic architecture with decoupled critics, enabling efficient exploration of a hybrid discrete-continuous action space. A curriculum learning strategy ensures scalability from simple domains to highly complex geometries, suppressing seed variance. By design, the recursive decomposition enables parallel meshing of subregions, yielding globally conforming all-quadrilateral meshes without post hoc correction. Across a wide range of benchmarks, Dmsh consistently outperforms existing methods in automation, robustness, and mesh quality, establishing a new paradigm for learning-based mesh generation.

math.NA

Data-Driven Structural Health Monitoring of Short Carbon Fiber-Reinforced Polymer Composites via Multiphysics Phase-Field Simulation

Short carbon fiber-reinforced polymer (SCFRP) composites exploit the intrinsic conductivity of the carbon fiber network for self-sensing, yet no predictive model couples their anisotropic, rate-dependent fracture to piezoresistive damage identification. This work presents a finite deformation multiphysics phase-field framework coupling a viscoelastic-viscoplastic constitutive model, an anisotropic crack resistance formulation, and a piezoresistive conductivity model. The three sub-problems are unified through the second-order fiber orientation tensor, which simultaneously defines fiber family directions, crack resistance anisotropy, and principal conduction paths of the carbon fiber network. A damage-coupled conductivity tensor captures both strain-driven geometric-kinematic resistance changes and irreversible network severance driven by the phase-field variable. The framework is coupled to an eight-electrode electrical impedance tomography configuration, and the normalized inter-electrode conductance ratios serve as inputs to a feedforward artificial neural network that infers normalized crack length and mechanical compliance without mechanical sensing. The network achieves R2 = 0.99 on held-out configurations, confirming generalization across the microstructure space. The framework establishes a physics-based, computationally efficient route for real-time structural health monitoring and inverse damage assessment in SCFRP composites.

cs.CE

WINO: A Weak-Form Physics Informed Neural Operator for Hyperelasticity on Variable Domains

We propose a Weak-form Physics-Informed Neural Operator (WINO), a data-free framework that combines the efficiency of neural operators with the geometric flexibility of the $\varphi$-finite element method ($\varphi$-FEM). $\varphi$-FEM is an unfitted method that accommodates geometric variations without body-fitted meshes, where the domain geometry is represented by the level-set function $\varphi$. To impose the boundary conditions, Dirichlet problems adopt the $\varphi$-FEM lifting so only the homogeneous displacement contribution is learned, whereas traction-driven Neumann problems additionally predict the auxiliary fields necessary for the unfitted weak formulation. Parameters are trained by minimizing squared weak-form residuals aligned with $\varphi$-FEM together with squared penalties on the cut-cell auxiliary equations, which removes the need for large paired datasets of converged reference solutions. When labeled reference data are available, an optional data-augmented variant (WINO+data) can further combine this physics-informed loss with a supervised term. After training, WINO outputs can seed the nonlinear $\varphi$-FEM solvers as neural operator warm starts (NOWS), which reduce iteration counts relative to traditional cold-started solvers. Numerical benchmarks show substantial accuracy of WINO together with total training times of about 15%-70% of those of supervised $\varphi$-FEM-FNO across all cases, without requiring reference-solution generation.

math.NA

Replay-Based Continual Learning for Physics-Informed Neural Operators

Neural operators generally demonstrate strong predictive performance on in-distribution (ID) problems. However, a critical limitation of existing methods is their significant performance degradation when encountering out-of-distribution (OOD) data. To address this issue, this work introduces continual learning into physics-informed neural operators, with particular emphasis on neural operators built upon the Transolver architecture, and proposes a simple yet effective replay-based continual learning strategy. The proposed method is fully physics-informed and does not require labeled data, relying solely on input fields together with physical constraints for training. When new OOD data become available, a small number of past data are incorporated through a distillation-based constraint to preserve previously acquired knowledge and alleviate catastrophic forgetting. Meanwhile, a transfer learning LoRA is employed to enable rapid adaptation to the new data. The proposed framework is systematically validated on three representative physical problems, including the Darcy flow problem in fluid mechanics, a two-dimensional hyperelastic brain tumor problem in biomechanics, and a three-dimensional linear elastic Triply Periodic Minimal Surfaces problem in solid mechanics. The results demonstrate that the proposed method effectively mitigates catastrophic forgetting on previously learned data while maintaining fast adaptability to new data. Compared with conventional joint training strategies, the proposed method significantly improves training efficiency while reducing additional memory usage and computational cost.

cs.LG

An improved PINN framework integrating localized collocation scheme and PIKF

We propose a localized physics-informed kernel function neural network (LPIKFNN), which is an improved physics-informed neural network (PINN) based on physics-informed kernel function (PIKF). In the LPIKFNN framework, the localized collocation scheme discretizes the physical quantities within the local domain, where the physical field is represented as a linear combination of PIKFs. Based on this representation, the multilayer perceptron is trained to iteratively learn the physical quantities. To overcome the computational challenges of conventional PINN in higher-order derivative and high wavenumber problems, the LPIKFNN constructs the loss function using the PIKF and a localized collocation scheme rather than relying on automatic differentiation. As a result, the costly derivative evaluations required to enforce governing equations during iterative training are eliminated, leading to significantly improved computational efficiency and training performance. Moreover, incorporating PIKFs into the loss function enables the proposed LPIKFNN to significantly improve computational accuracy in high-wavenumber problems characterized by highly oscillatory physical fields. To overcome the computational bottleneck of the physics-informed kernel function neural network (PIKFNN) in heterogeneous problems, the LPIKFNN introduces a localized collocation scheme that removes reliance on global PIKFs, enabling accurate predictions where global PIKFs are unavailable. The feasibility and accuracy of the proposed LPIKFNN are demonstrated through a series of benchmark studies, including high wavenumber problems, higher-order derivative problems, nonlinear problems, heterogeneous problems, and potential-based inverse electromyography. The numerical predictions obtained by LPIKFNN show excellent agreement with available analytical solutions and experimental measurements.

cs.CE

A Computational Model for Flexoelectricity-Driven Contact Electrification

Recent theoretical studies show that nanoscale contact on dielectric substrates can induce flexoelectric polarization large enough to drive electron transfer. This has been supported by experimental evidence, indicating that contact electrification is inherently a coupled electromechanical phenomenon. In this work, we develop a computational model for flexoelectricity-driven contact electrification that integrates finite-deformation flexoelectricity with contact mechanics and physically motivated charge transfer. A tunneling transparency function is introduced to regulate the interfacial channel based on the WKB approximation, capturing the irreversible charge trapping during unloading. Three contact scenarios are investigated with specific hypotheses for charge transfer: unbiased metal-dielectric contact driven by surface polarization, biased contact restricted to carriers of a single polarity, and dielectric-dielectric contact where surface states with finite capacity limit the transferable charge. The model is compared with atomic force microscopy measurements on PMMA and PDAP substrates under both biased and unbiased conditions.For contact between identical dielectric materials, we show that geometric asymmetry in surface curvature is sufficient to induce charge separation, with polarity reversal occurring at a critical surface wavenumber. Three-dimensional simulations on random rough surfaces reproduce the mosaic charge distributions observed experimentally, confirming that contact-induced local strain gradient heterogeneity can generate spatially non-uniform charge patterns without introducing any material inhomogeneity.

physics.app-ph

A phase-field framework for anisotropic viscoelastic-viscoplastic fracture in short fiber-reinforced polymers in hygrothermal environments

This work presents a comprehensive phase-field framework for modeling anisotropic viscoelastic-viscoplastic fracture in short fiber-reinforced polymer (SFRP) composites under hygrothermal environments at finite deformation. The constitutive model employs a multiplicative decomposition of the deformation gradient into viscoelastic and viscoplastic components. An anisotropic phase-field formulation is developed using structural tensors to capture orientation-dependent fracture energy induced by multiple fiber families. Hygrothermal effects are incorporated through moisture-dependent swelling, thermal expansion, and temperature- and moisture-sensitive material parameters within the coupled framework. Numerical investigations demonstrate the framework's capability to capture complex fracture phenomena in SFRPs. Results reveal that fiber orientation fundamentally governs the spatial distribution of crack driving force, with maximum energy accumulation along fiber directions persisting throughout viscous relaxation. The anisotropy parameter controlling directional fracture resistance significantly influences crack path deflection. Hygrothermal degradation substantially reduces both peak load and fracture energy, with moisture absorption and elevated temperature each contributing to decreased mechanical performance. The framework captures the influence of fiber mechanical properties on global load-bearing capacity and crack propagation resistance. This unified computational framework advances the predictive modeling of damage evolution in SFRPs subjected to realistic environmental and mechanical loading conditions.

cs.CE

Deep Energy Method with Large Language Model assistance: an open-source Streamlit-based platform for solving variational PDEs

Physics-informed neural networks (PINNs) in energy form, also known as the deep energy method (DEM), offer advantages over strong-form PINNs such as lower-order derivatives and fewer hyperparameters, yet dedicated and user-friendly software for energy-form PINNs remains scarce. To address this gap, we present \textbf{LM-DEM} (Large-Model-assisted Deep Energy Method), an open-source, Streamlit-based platform for solving variational partial differential equations (PDEs) in computational mechanics. LM-DEM integrates large language models (LLMs) for geometry modeling: users can generate Gmsh-compatible geometries directly from natural language descriptions or images, significantly reducing the burden of traditional geometry preprocessing. The solution process is driven by the deep energy method, while finite element solutions can be obtained in parallel. The framework supports built-in problems including Poisson, screened Poisson, linear elasticity, and hyperelasticity in two and three dimensions, as well as user-defined energy functionals analogous to the \texttt{UMAT} interface in Abaqus. The source code is available at https://github.com/yizheng-wang/LMDEM, and a web-based version is accessible at https://ai4m.llmdem.com. LM-DEM aims to lower the barrier for practitioners and beginners to adopt energy-form PINNs for variational PDE problems.

math.NA

Pretrain Finite Element Method: A Pretraining and Warm-start Framework for PDEs via Physics-Informed Neural Operators

We propose a Pretrained Finite Element Method (PFEM),a physics driven framework that bridges the efficiency of neural operator learning with the accuracy and robustness of classical finite element methods (FEM). PFEM consists of a physics informed pretraining stage and an optional finetuning stage. In the pretraining stage, a neural operator based on the Transolver architecture is trained solely from governing partial differential equations, without relying on labeled solution data. The model operates directly on unstructured point clouds, jointly encoding geometric information, material properties, and boundary conditions, and produces physically consistent initial solutions with extremely high computational efficiency. PDE constraints are enforced through explicit finite element, based differentiation, avoiding the overhead associated with automatic differentiation. In the fine-tuning stage, the pretrained prediction is used as an initial guess for conventional FEM solvers, preserving their accuracy, convergence guarantees, and extrapolation capability while substantially reducing the number of iterations required to reach a prescribed tolerance. PFEM is validated on a broad range of benchmark problems, including linear elasticity and nonlinear hyperelasticity with complex geometries, heterogeneous materials, and arbitrary boundary conditions. Numerical results demonstrate strong generalization in the pretraining stage with relative errors on the order of 1\%, and speedups of up to one order of magnitude in the fine-tuning stage compared to FEM with zero initial guesses.

math.NA

Phase-field modeling of multicomponent vesicles in viscoelastic fluid

Multicomponent vesicles suspended in viscoelastic fluids are crucial for understanding a variety of physiological processes. In this work, we develop a continuum surface force (CSF) phase-field model to investigate the hydrodynamics of inextensible multicomponent vesicles in viscoelastic fluid flows with inertial forces. Our model couples a fluid field comprising both Newtonian and Oldroyd-B fluids, a surface concentration field representing the multicomponent distribution on the vesicle membrane, and a phase-field variable governing the membrane evolution. The viscoelasticity effect of extra stress is well incorporated into the full Navier-Stokes equations in the fluid field. The surface concentration field is determined by Cahn-Hilliard equations, while the membrane evolution is governed by a nonlinear advection-diffusion equation. The membrane is coupled to the surrounding fluid through the continuum surface force (CSF) framework. To ensure stable numerical solutions of the highly nonlinear multi-field model, we employ a residual-based variational multiscale (RBVMS) method for the Navier-Stokes equations, a Streamline-Upwind Petrov-Galerkin (SUPG) method for the Oldroyd-B equations, and a standard Galerkin finite element framework for the remaining equations. The system of PDEs is solved using an implicit, monolithic scheme based on the generalized-$\alpha$ time integration method. To enhance spatial accuracy, we employ isogeometric analysis (IGA). We present a series of two-dimensional numerical examples in shear and Poiseuille flows to elucidate the influence of membrane composition and fluid viscoelasticity on the hydrodynamics of multicomponent vesicles.

physics.flu-dyn