arXiv ScienceSearch

arXiv subjects

Markus Kästner

Publications and source records attributed to Markus Kästner.

At least 19 recordsLinked to original sources

Reliable training of neural hyperelastic models via full-field data

We present a systematic investigation of the robustness and limitations of equilibrium gap-based calibrations for hyperelastic physics-augmented neural networks (PANNs), where we consider the special case of isotropic and polyconvex PANNs. In full-field parameterizations, it is commonly assumed that the displacement field is captured with sufficient spatial resolution for an accurate evaluation of the deformation field, and that the specimen is thin enough for plane stress to hold to a good approximation. Since these assumptions are never ideally satisfied in real experiments, we investigate, using synthetically generated data, how severely an under-resolved surface measurement and a non-negligible specimen thickness can affect the model parameterization. Furthermore, we perform calibration on real experimental data for a set of inhomogeneous specimen geometries. We show that the coverage of the admissible deformation states during calibration governs the ability of a model to generalize to unseen geometries and load cases; this ability can be improved further by appropriate combinations of specimens. Accurately depicting the material behavior underlying this rich data, however, requires a sufficiently flexible constitutive model, for which PANNs are well suited. Yet a rich coverage of deformation states alone is not sufficient: unless the calibration data comprise biaxial-tension-like states, models that include the second deformation invariant extrapolate unphysically towards equi-biaxial tension, whereas restricting the PANN to the first invariant remains reliable.

cond-mat.mtrl-sci

Calibration of neural viscoelastic models via full-field data

We propose an unsupervised learning framework for calibrating a physics-augmented neural network (PANN) for small-strain viscoelasticity via full-field data. It only requires quantities that are directly accessible in real experiments for training, namely global reaction forces and surface displacements. The underlying PANN is embedded in the generalized standard materials theory, in which two scalar-valued potentials render the constitutive model thermodynamically consistent by construction, while invariant-based representations of the free energy and the dual dissipation potential additionally ensure material symmetry. Considering a thin specimen under the plane stress assumption, we formulate a constrained optimization problem based on the equilibrium gap method in combination with quasi-Newton optimizers and automatic differentiation. Thereby, the unknown out-of-plane strain follows from the plane stress condition and the evolution of the internal variables is captured by an implicit time integration scheme. The resulting system of nonlinear equations is solved via a local Newton iteration at quadrature point and time step. To drastically reduce the computational cost of training, the backward adjoint method is employed to compute the gradient of the target loss, instead of backpropagating through all Newton iteration steps. The proposed framework is demonstrated for synthetic data, including noisy displacements and forces, showing excellent agreement across a wide range of deformation rates and load paths.

cs.CE

Construction of minimal integrity basis for anisotropic hyperelasticity via structural tensors

We present minimal integrity bases for all common anisotropies in hyperelasticity via the structural tensor concept, which can be used to formulate any algebraic invariant function in the elements of the respective bases. Hence, the provided minimal integrity bases are of great interest for formulating a concise but general anisotropic material model. Our work covers results for the 11 types of anisotropy that arise from the classical 7 crystal systems, as well as findings for 4 additional non-crystal anisotropies derived from the cylindrical, spherical, and icosahedral symmetry systems. By using well-known results from literature about structural tensors, isotropic invariants, and isotropic extension, functional bases are directly determined. A simple analytical-numerical approach is employed to identify polynomial relations between the invariants of these functional bases, thereby enabling the construction of functional bases of reduced cardinality. After that, we show that the determined reduced functional bases are also minimal integrity bases by identifying polynomial relations to known integrity bases from literature. Furthermore, fundamental concepts from invariant theory, including the Hironaka decomposition of invariant rings and the closely related Hilbert series, are employed to further validate the results. Alongside the presented findings, this article also aims to provide an introductory overview of the complex field of modeling anisotropic materials, especially for researchers with an engineering background.

cond-mat.mtrl-sci

On limitations of polyconvexity

Polyconvex constitutive modeling is attractive as it guarantees stability of numerical simulations and can improve the generalization behavior of material models. However, in certain applications, polyconvex formulations perform poorly in reproducing the underlying ground truth material response, which can effectively preclude their practical use. In this work, we address this issue and investigate the limitations of polyconvex constitutive modeling. The main contributions of this paper are as follows: (1) We analyze the theoretical reasons why polyconvexity may, in some cases, impose overly restrictive constraints that limit the achievable accuracy of constitutive models. Thereby, we provide analytical ellipticity guarantees for two non-polyconvex Mooney-Rivlin type potentials. (2) We investigate the practical limitations of polyconvex physics-augmented neural network constitutive models using two representative formulations: models using structural tensor-based invariants and models using signed singular values. Their performance is evaluated on datasets obtained from homogenized microstructured materials, and their predictive capabilities are assessed in finite element simulations. (3) Overall, we provide an overview of benefits, limitations, and mitigation strategies of polyconvex constitutive modeling.

cs.CE

Advances in polyconvex anisotropic hyperelasticity

A key challenge in material theory is the formulation of models that satisfy all common mechanical constitutive conditions while retaining sufficient flexibility. In this context, several important modeling aspects remain unresolved for polyconvex anisotropic hyperelasticity. We address some of these challenges and apply our results for physics-augmented neural network (PANN) constitutive modeling. The main contributions of this paper are as follows: (1) We propose a new polyconvex PANN constitutive model for anisotropic hyperelasticity based on triclinic invariants and group symmetrization. For finite symmetry groups, this model fulfills all common mechanical constitutive conditions a priori. (2) We propose a group symmetrization-based method for the construction of polyconvex invariants for finite symmetry groups. Based on this, we derive a new integrity basis for a tetragonal symmetry group and a new functional basis for a cubic symmetry group. To the best of our knowledge, these are the first polyconvex integrity or functional bases for symmetry groups characterized by structural tensors of order higher than two. (3) We provide an extensive introduction to the construction of polyconvex integrity and functional bases, which form the basis of polyconvex invariant-based constitutive models. We discuss polyconvex bases for triclinic, isotropic, transversely isotropic, monoclinic, rhombic, tetragonal, and cubic symmetry groups. (4) We benchmark the polyconvex PANN constitutive models with highly nonlinear homogenization data of cubic metamaterials.

cs.CE

Generative reconstruction of 2D and 3D polycrystalline microstructures using symmetrized hyperspherical harmonics

Establishing structure-property linkages in polycrystalline materials requires representative two- (2D) and three- (3D) dimensional microstructural inputs for full-field simulations. A core objective of microstructure characterization and reconstruction is the generative synthesis of 2D and 3D microstructures that reflect a target statistical ensemble using limited 2D data as a reference. This work introduces an orientation-based differentiable microstructure characterization and reconstruction framework, implemented in MCRpy, to perform reconstructions of voxelized images. Unit quaternions in combination with symmetrized hyperspherical harmonics are utilized to derive a continuous, symmetry-invariant representation of crystallographic orientations to overcome the numerical singularities and discontinuities associated with traditional Euler-based methods. The descriptor-based reconstructions are driven by a set combining two-point spatial correlations, a novel hybrid three-point variogram, and a mean variation regularizer to capture both global texture and local interfacial topology. The framework's efficiency is demonstrated by reconstructing 3D realizations from 2D orientation data of an aluminum alloy after thermo-mechanical processing, successfully recovering both morphological features and crystallographic distribution. Systematic benchmarking indicates that second-order gradient-based optimization, utilizing the L-BFGS-B algorithm, effectively navigates the complex loss landscape to generate high-fidelity realizations with minimal residuals. This methodology provides a versatile, open-source framework for the digital synthesis of polycrystalline representative volume elements to facilitate the rapid development of microstructure-informed materials design workflows.

cond-mat.mtrl-sci

An Efficient Bayesian Framework for Inverse Problems via Optimization and Inversion: Surrogate Modeling, Parameter Inference, and Uncertainty Quantification

The present paper proposes a Bayesian framework for inverse problems that seamlessly integrates optimization and inversion to enable rapid surrogate modeling, accurate parameter inference, and rigorous uncertainty quantification. Bayesian optimization is employed to adaptively construct accurate Gaussian process surrogate models using a minimal number of high-fidelity model evaluations, strategically focusing sampling in regions of high predictive uncertainty. The trained surrogate model is then leveraged within a Bayesian inversion scheme to infer optimal parameter values by combining prior knowledge with observed quantities of interest, resulting in posterior distributions that rigorously characterize epistemic uncertainty. The framework is theoretically grounded, computationally efficient, and particularly suited for engineering applications in which high-fidelity models -- whether arising from numerical simulations or physical experiments -- are computationally expensive, analytically intractable, or difficult to replicate, and data availability is limited. Furthermore, the combined use of Bayesian optimization and inversion outperforms their separate application, highlighting the synergistic benefits of unifying the two approaches. The performance of the proposed Bayesian framework is demonstrated on a suite of one- and two-dimensional analytical benchmarks, including the Mixed Gaussian-Periodic, Lévy, Griewank, Forrester, and Rosenbrock functions, which provide a controlled setting to assess surrogate modeling accuracy, parameter inference robustness, and uncertainty quantification. The results demonstrate the framework's effectiveness in efficiently solving inverse problems while providing informative uncertainty quantification and supporting reliable engineering decision-making at reduced computational cost.

cs.CE

Precise, efficient and flexible modeling of crystallizing elastomers based on physics-augmented neural networks

We propose a precise and efficient physics-augmented neural network (PANN) to model strain-induced crystallization in rubbery polymers. We demonstrate that the model can be flexibly employed for both unfilled and filled natural rubber (NR). The approach is based on a two potential framework, similar to the concept of generalized standard materials (GSMs). To describe the material behavior, neural network-based free energy and dissipation potentials are employed. The evolution of crystallinity is derived from the two potentials. To ensure boundedness of the crystallinity, a novel constrained GSM-type evolution problem is proposed. To this end, two additional Lagrange multipliers together with the corresponding Karush-Kuhn-Tucker conditions are introduced. As a result, it is guaranteed that crystallinity can be interpreted as a variable of concentration type. The neural network-based potentials ensure all physically desirable properties by construction. Most importantly, objectivity, material symmetry and thermodynamic consistency are automatically fulfilled. In addition, an alternative derivation of the governing model equations in time-discrete form is presented based on an incremental variational framework, which also serves as the basis for a finite element implementation. We demonstrate the predictive capability of the PANN using three different experimental data sets from literature, considering both stress and crystallinity evolution at material point level as well as the corresponding field distributions in a notched specimen. Moreover, we show that model parameterization is also possible when experimental crystallinity data is not available, still enabling suitable stress predictions.

cond-mat.mtrl-sci

A physics-augmented neural network framework for finite strain incompressible viscoelasticity

We propose a physics-augmented neural network (PANN) framework for finite strain incompressible viscoelasticity within the generalized standard materials theory. The formulation is based on the multiplicative decomposition of the deformation gradient and enforces unimodularity of the inelastic deformation part throughout the evolution. Invariant-based representations of the free energy and the dual dissipation potential by monotonic and fully input-convex neural networks ensure thermodynamic consistency, objectivity, and material symmetry by construction. The evolution of the internal variables during training is handled by solving the evolution equations using an implicit exponential time integrator. In addition, a trainable gate layer combined with lp regularization automatically identifies the required number of internal variables during training. The PANN is calibrated with synthetic and experimental data, showing excellent agreement for a wide range of deformation rates and different load paths. We also show that the proposed model achieves excellent interpolation as well as plausible and accurate extrapolation behaviors. In addition, we demonstrate consistency of the PANN with linear viscoelasticity by linearization of the full model.

cs.CE

A data-driven multiscale scheme for anisotropic finite strain magneto-elasticity

In this work, we develop a neural network-based, data-driven, decoupled multiscale scheme for the modeling of structured magnetically soft magnetorheological elastomers (MREs). On the microscale, sampled magneto-mechanical loading paths are imposed on a representative volume element containing spherical particles and an elastomer matrix, and the resulting boundary value problem is solved using a mixed finite element formulation. The computed microscale responses are homogenized to construct a database for the training and testing of a macroscopic physics-augmented neural network model. The proposed model automatically detects the material's preferred direction during training and enforces key physical principles, including objectivity, material symmetry, thermodynamic consistency, and the normalization of free energy, stress, and magnetization. Within the range of the training data, the model enables accurate predictions of magnetization, mechanical stress, and total stress. For larger magnetic fields, the model yields plausible results. Finally, we apply the model to investigate the magnetostrictive behavior of a macroscopic spherical MRE sample, which exhibits contraction along the magnetic field direction when aligned with the material's preferred direction.

cs.CE

Data-efficient inverse design of spinodoid metamaterials

We create an data-efficient and accurate surrogate model for structure-property linkages of spinodoid metamaterials with only 75 data points -- far fewer than the several thousands used in prior works -- and demonstrate its use in multi-objective inverse design. The inverse problem of finding a material microstructure that leads to given bulk properties is of great interest in mechanics and materials science. These inverse design tasks often require a large dataset, which can become unaffordable when considering material behavior that requires more expensive simulations or experiments. We generate a data-efficient surrogate for the mapping between the characteristics of the local material structure and the effective elasticity tensor and use it to inversely design structures with multiple objectives simultaneously. The presented neural network-based surrogate model achieves its data efficiency by inherently satisfying certain requirements, such as equivariance with respect to permutations of structure parameters, which avoids having to learn them from data. The resulting surrogate of the forward model is differentiable, allowing its direct use in gradient-based optimization for the inverse design problem. We demonstrate in three inverse design tasks of varying complexity that this approach yields reliable results while requiring significantly less training data than previous approaches based on neural-network surrogates. This paves the way for inverse design involving nonlinear mechanical behavior, where data efficiency is currently the limiting factor.

cs.CE

A dual-stage constitutive modeling framework based on finite strain data-driven identification and physics-augmented neural networks

In this contribution, we present a novel consistent dual-stage approach for the automated generation of hyperelastic constitutive models which only requires experimentally measurable data. To generate input data for our approach, an experiment with full-field measurement has to be conducted to gather testing force and corresponding displacement field of the sample. Then, in the first step of the dual-stage framework, a new finite strain Data-Driven Identification (DDI) formulation is applied. This method enables to identify tuples consisting of stresses and strains by only prescribing the applied boundary conditions and the measured displacement field. In the second step, the data set is used to calibrate a Physics-Augmented Neural Network (PANN), which fulfills all common conditions of hyperelasticity by construction and is very flexible at the same time. We demonstrate the applicability of our approach by several descriptive examples. Two-dimensional synthetic data are exemplarily generated in virtual experiments by using a reference constitutive model. The calibrated PANN is then applied in 3D Finite Element simulations. In addition, a real experiment including noisy data is mimicked.

cs.CE

When invariants matter: The role of I1 and I2 in neural network models of incompressible hyperelasticity

For the formulation of machine learning-based material models, the usage of invariants of deformation tensors is attractive, since this can a priori guarantee objectivity and material symmetry. In this work, we consider incompressible, isotropic hyperelasticity, where two invariants I1 and I2 are required for depicting a deformation state. First, we aim at enhancing the understanding of the invariants. We provide an explicit representation of the set of invariants that are admissible, i.e. for which (I1, I2) a deformation state does indeed exist. Furthermore, we prove that uniaxial and equi-biaxial deformation correspond to the boundary of the set of admissible invariants. Second, we study how the experimentally-observed behaviour of different materials can be captured by means of neural network models of incompressible hyperelasticity, depending on whether both I1 and I2 or solely one of the invariants, i.e. either only I1 or only I2, are taken into account. To this end, we investigate three different experimental data sets from the literature. In particular, we demonstrate that considering only one invariant, either I1 or I2, can allow for good agreement with experiments in case of small deformations. In contrast, it is necessary to consider both invariants for precise models at large strains, for instance when rubbery polymers are deformed. Moreover, we show that multiaxial experiments are strictly required for the parameterisation of models considering I2. Otherwise, if only data from uniaxial deformation is available, significantly overly stiff responses could be predicted for general deformation states. On the contrary, I1-only models can make qualitatively correct predictions for multiaxial loadings even if parameterised only from uniaxial data, whereas I2-only models are completely incapable in even qualitatively capturing experimental stress data at large deformations.

cond-mat.mtrl-sci

Fatigue monitoring and maneuver identification for vehicle fleets using a virtual sensing approach

Extensive monitoring comes at a prohibitive cost, limiting Predictive Maintenance strategies for vehicle fleets. This paper presents a measurement-based virtual sensing technique where local strain gauges are only required for few reference vehicles, while the remaining fleet relies exclusively on accelerometers. The scattering transform is used to perform feature extraction, while principal component analysis provides a reduced, low dimensional data representation. This enables direct fatigue damage regression, parameterized from unlabeled usage data. Identification measurements allow for a physical interpretation of the reduced representation. The approach is demonstrated using experimental data from a sensor equipped eBike, which is made publicly available.

eess.SP

Phase-field models for ductile fatigue fracture

Fatigue fracture is one of the main causes of failure in structures. However, the simulation of fatigue crack growth is computationally demanding due to the large number of load cycles involved. Metals in the low cycle fatigue range often show significant plastic zones at the crack tip, calling for elastic-plastic material models, which increase the computation time even further. In pursuit of a more efficient model, we propose a simplified phase-field model for ductile fatigue fracture, which indirectly accounts for plasticity within the fatigue damage accumulation. Additionally, a cycle-skipping approach is inherent to the concept, reducing computation time by up to several orders of magnitude. Essentially, the proposed model is a simplification of a phase-field model with elastic-plastic material behavior. As a reference, we therefore implement a conventional elastic-plastic phase-field fatigue model with nonlinear hardening and a fatigue variable based on the strain energy density, and compare the simplified model to it. Its approximation of the stress-strain behavior, the neglect of the plastic crack driving force and consequential range of applicability are discussed. Since in fact the novel efficient model is similar in its structure to a phase-field fatigue model we published in the past, we include this older version in the comparison, too. Compared to this model variant, the novel model improves the approximation of the plastic strains and corresponding stresses and refines the damage computation based on the Local Strain Approach. For all model variants, experimentally determined values for elastic, plastic, fracture and fatigue properties of AA2024 T351 aluminum sheet material are employed.

physics.comp-ph

Neural networks meet anisotropic hyperelasticity: A framework based on generalized structure tensors and isotropic tensor functions

We present a data-driven framework for the multiscale modeling of anisotropic finite strain elasticity based on physics-augmented neural networks (PANNs). Our approach allows the efficient simulation of materials with complex underlying microstructures which reveal an overall anisotropic and nonlinear behavior on the macroscale. By using a set of invariants as input, an energy-type output and by adding several correction terms to the overall energy density functional, the model fulfills multiple physical principles by construction. The invariants are formed from the right Cauchy-Green deformation tensor and fully symmetric 2nd, 4th or 6th order structure tensors which enables to describe a wide range of symmetry groups. Besides the network parameters, the structure tensors are simultaneously calibrated during training so that the underlying anisotropy of the material is reproduced most accurately. In addition, sparsity of the model with respect to the number of invariants is enforced by adding a trainable gate layer and using lp regularization. Our approach works for data containing tuples of deformation, stress and material tangent, but also for data consisting only of tuples of deformation and stress, as is the case in real experiments. The developed approach is exemplarily applied to several representative examples, where necessary data for the training of the PANN surrogate model are collected via computational homogenization. We show that the proposed model achieves excellent interpolation and extrapolation behaviors. In addition, the approach is benchmarked against an NN model based on the components of the right Cauchy-Green deformation tensor.

cs.CE

Inverse design of spinodoid structures using Bayesian optimization

Tailoring materials to achieve a desired behavior in specific applications is of significant scientific and industrial interest as design of materials is a key driver to innovation. Overcoming the rather slow and expertise-bound traditional forward approaches of trial and error, inverse design is attracting substantial attention. Targeting a property, the design model proposes a candidate structure with the desired property. This concept can be particularly well applied to the field of architected materials as their structures can be directly tuned. The bone-like spinodoid materials are a specific class of architected materials. They are of considerable interest thanks to their non-periodicity, smoothness, and low-dimensional statistical description. Previous work successfully employed machine learning (ML) models for inverse design. The amount of data necessary for most ML approaches poses a severe obstacle for broader application, especially in the context of inelasticity. That is why we propose an inverse-design approach based on Bayesian optimization to operate in the small-data regime. Necessitating substantially less data, a small initial data set is iteratively augmented by in silico generated data until a structure with the targeted properties is found. The application to the inverse design of spinodoid structures of desired elastic properties demonstrates the framework's potential for paving the way for advance in inverse design.

cond-mat.mtrl-sci

Viscoelasticty with physics-augmented neural networks: Model formulation and training methods without prescribed internal variables

We present an approach for the data-driven modeling of nonlinear viscoelastic materials at small strains which is based on physics-augmented neural networks (NNs) and requires only stress and strain paths for training. The model is built on the concept of generalized standard materials and is therefore thermodynamically consistent by construction. It consists of a free energy and a dissipation potential, which can be either expressed by the components of their tensor arguments or by a suitable set of invariants. The two potentials are described by fully/partially input convex neural networks. For training of the NN model by paths of stress and strain, an efficient and flexible training method based on a recurrent cell, particularly a long short-term memory cell, is developed to automatically generate the internal variable(s) during the training process. The proposed method is benchmarked and thoroughly compared with existing approaches. These include a method that obtains the internal variable by integrating the evolution equation over the entire sequence, while the other method uses an an auxiliary feedforward neural network for the internal variable(s). Databases for training are generated by using a conventional nonlinear viscoelastic reference model, where 3D and 2D plane strain data with either ideal or noisy stresses are generated. The coordinate-based and the invariant-based formulation are compared and the advantages of the latter are demonstrated. Afterwards, the invariant-based model is calibrated by applying the three training methods using ideal or noisy stress data. All methods yield good results, but differ in computation time and usability for large data sets. The presented training method based on a recurrent cell turns out to be particularly robust and widely applicable and thus represents a promising approach for the calibration of other types of models as well.

cs.CE