arXiv Science⌕ Search

arXiv subjects

Karl A. Kalina

Publications and source records attributed to Karl A. Kalina.

At least 19 recordsLinked to original sources

Reliable training of neural hyperelastic models via full-field data

We present a systematic investigation of the robustness and limitations of equilibrium gap-based calibrations for hyperelastic physics-augmented neural networks (PANNs), where we consider the special case of isotropic and polyconvex PANNs. In full-field parameterizations, it is commonly assumed that the displacement field is captured with sufficient spatial resolution for an accurate evaluation of the deformation field, and that the specimen is thin enough for plane stress to hold to a good approximation. Since these assumptions are never ideally satisfied in real experiments, we investigate, using synthetically generated data, how severely an under-resolved surface measurement and a non-negligible specimen thickness can affect the model parameterization. Furthermore, we perform calibration on real experimental data for a set of inhomogeneous specimen geometries. We show that the coverage of the admissible deformation states during calibration governs the ability of a model to generalize to unseen geometries and load cases; this ability can be improved further by appropriate combinations of specimens. Accurately depicting the material behavior underlying this rich data, however, requires a sufficiently flexible constitutive model, for which PANNs are well suited. Yet a rich coverage of deformation states alone is not sufficient: unless the calibration data comprise biaxial-tension-like states, models that include the second deformation invariant extrapolate unphysically towards equi-biaxial tension, whereas restricting the PANN to the first invariant remains reliable.

cond-mat.mtrl-sci↗

Calibration of neural viscoelastic models via full-field data

We propose an unsupervised learning framework for calibrating a physics-augmented neural network (PANN) for small-strain viscoelasticity via full-field data. It only requires quantities that are directly accessible in real experiments for training, namely global reaction forces and surface displacements. The underlying PANN is embedded in the generalized standard materials theory, in which two scalar-valued potentials render the constitutive model thermodynamically consistent by construction, while invariant-based representations of the free energy and the dual dissipation potential additionally ensure material symmetry. Considering a thin specimen under the plane stress assumption, we formulate a constrained optimization problem based on the equilibrium gap method in combination with quasi-Newton optimizers and automatic differentiation. Thereby, the unknown out-of-plane strain follows from the plane stress condition and the evolution of the internal variables is captured by an implicit time integration scheme. The resulting system of nonlinear equations is solved via a local Newton iteration at quadrature point and time step. To drastically reduce the computational cost of training, the backward adjoint method is employed to compute the gradient of the target loss, instead of backpropagating through all Newton iteration steps. The proposed framework is demonstrated for synthetic data, including noisy displacements and forces, showing excellent agreement across a wide range of deformation rates and load paths.

cs.CE↗

Construction of minimal integrity basis for anisotropic hyperelasticity via structural tensors

We present minimal integrity bases for all common anisotropies in hyperelasticity via the structural tensor concept, which can be used to formulate any algebraic invariant function in the elements of the respective bases. Hence, the provided minimal integrity bases are of great interest for formulating a concise but general anisotropic material model. Our work covers results for the 11 types of anisotropy that arise from the classical 7 crystal systems, as well as findings for 4 additional non-crystal anisotropies derived from the cylindrical, spherical, and icosahedral symmetry systems. By using well-known results from literature about structural tensors, isotropic invariants, and isotropic extension, functional bases are directly determined. A simple analytical-numerical approach is employed to identify polynomial relations between the invariants of these functional bases, thereby enabling the construction of functional bases of reduced cardinality. After that, we show that the determined reduced functional bases are also minimal integrity bases by identifying polynomial relations to known integrity bases from literature. Furthermore, fundamental concepts from invariant theory, including the Hironaka decomposition of invariant rings and the closely related Hilbert series, are employed to further validate the results. Alongside the presented findings, this article also aims to provide an introductory overview of the complex field of modeling anisotropic materials, especially for researchers with an engineering background.

cond-mat.mtrl-sci↗

On limitations of polyconvexity

Polyconvex constitutive modeling is attractive as it guarantees stability of numerical simulations and can improve the generalization behavior of material models. However, in certain applications, polyconvex formulations perform poorly in reproducing the underlying ground truth material response, which can effectively preclude their practical use. In this work, we address this issue and investigate the limitations of polyconvex constitutive modeling. The main contributions of this paper are as follows: (1) We analyze the theoretical reasons why polyconvexity may, in some cases, impose overly restrictive constraints that limit the achievable accuracy of constitutive models. Thereby, we provide analytical ellipticity guarantees for two non-polyconvex Mooney-Rivlin type potentials. (2) We investigate the practical limitations of polyconvex physics-augmented neural network constitutive models using two representative formulations: models using structural tensor-based invariants and models using signed singular values. Their performance is evaluated on datasets obtained from homogenized microstructured materials, and their predictive capabilities are assessed in finite element simulations. (3) Overall, we provide an overview of benefits, limitations, and mitigation strategies of polyconvex constitutive modeling.

cs.CE↗

Advances in polyconvex anisotropic hyperelasticity

A key challenge in material theory is the formulation of models that satisfy all common mechanical constitutive conditions while retaining sufficient flexibility. In this context, several important modeling aspects remain unresolved for polyconvex anisotropic hyperelasticity. We address some of these challenges and apply our results for physics-augmented neural network (PANN) constitutive modeling. The main contributions of this paper are as follows: (1) We propose a new polyconvex PANN constitutive model for anisotropic hyperelasticity based on triclinic invariants and group symmetrization. For finite symmetry groups, this model fulfills all common mechanical constitutive conditions a priori. (2) We propose a group symmetrization-based method for the construction of polyconvex invariants for finite symmetry groups. Based on this, we derive a new integrity basis for a tetragonal symmetry group and a new functional basis for a cubic symmetry group. To the best of our knowledge, these are the first polyconvex integrity or functional bases for symmetry groups characterized by structural tensors of order higher than two. (3) We provide an extensive introduction to the construction of polyconvex integrity and functional bases, which form the basis of polyconvex invariant-based constitutive models. We discuss polyconvex bases for triclinic, isotropic, transversely isotropic, monoclinic, rhombic, tetragonal, and cubic symmetry groups. (4) We benchmark the polyconvex PANN constitutive models with highly nonlinear homogenization data of cubic metamaterials.

cs.CE↗

Precise, efficient and flexible modeling of crystallizing elastomers based on physics-augmented neural networks

We propose a precise and efficient physics-augmented neural network (PANN) to model strain-induced crystallization in rubbery polymers. We demonstrate that the model can be flexibly employed for both unfilled and filled natural rubber (NR). The approach is based on a two potential framework, similar to the concept of generalized standard materials (GSMs). To describe the material behavior, neural network-based free energy and dissipation potentials are employed. The evolution of crystallinity is derived from the two potentials. To ensure boundedness of the crystallinity, a novel constrained GSM-type evolution problem is proposed. To this end, two additional Lagrange multipliers together with the corresponding Karush-Kuhn-Tucker conditions are introduced. As a result, it is guaranteed that crystallinity can be interpreted as a variable of concentration type. The neural network-based potentials ensure all physically desirable properties by construction. Most importantly, objectivity, material symmetry and thermodynamic consistency are automatically fulfilled. In addition, an alternative derivation of the governing model equations in time-discrete form is presented based on an incremental variational framework, which also serves as the basis for a finite element implementation. We demonstrate the predictive capability of the PANN using three different experimental data sets from literature, considering both stress and crystallinity evolution at material point level as well as the corresponding field distributions in a notched specimen. Moreover, we show that model parameterization is also possible when experimental crystallinity data is not available, still enabling suitable stress predictions.

cond-mat.mtrl-sci↗

A physics-augmented neural network framework for finite strain incompressible viscoelasticity

We propose a physics-augmented neural network (PANN) framework for finite strain incompressible viscoelasticity within the generalized standard materials theory. The formulation is based on the multiplicative decomposition of the deformation gradient and enforces unimodularity of the inelastic deformation part throughout the evolution. Invariant-based representations of the free energy and the dual dissipation potential by monotonic and fully input-convex neural networks ensure thermodynamic consistency, objectivity, and material symmetry by construction. The evolution of the internal variables during training is handled by solving the evolution equations using an implicit exponential time integrator. In addition, a trainable gate layer combined with lp regularization automatically identifies the required number of internal variables during training. The PANN is calibrated with synthetic and experimental data, showing excellent agreement for a wide range of deformation rates and different load paths. We also show that the proposed model achieves excellent interpolation as well as plausible and accurate extrapolation behaviors. In addition, we demonstrate consistency of the PANN with linear viscoelasticity by linearization of the full model.

cs.CE↗

A data-driven multiscale scheme for anisotropic finite strain magneto-elasticity

In this work, we develop a neural network-based, data-driven, decoupled multiscale scheme for the modeling of structured magnetically soft magnetorheological elastomers (MREs). On the microscale, sampled magneto-mechanical loading paths are imposed on a representative volume element containing spherical particles and an elastomer matrix, and the resulting boundary value problem is solved using a mixed finite element formulation. The computed microscale responses are homogenized to construct a database for the training and testing of a macroscopic physics-augmented neural network model. The proposed model automatically detects the material's preferred direction during training and enforces key physical principles, including objectivity, material symmetry, thermodynamic consistency, and the normalization of free energy, stress, and magnetization. Within the range of the training data, the model enables accurate predictions of magnetization, mechanical stress, and total stress. For larger magnetic fields, the model yields plausible results. Finally, we apply the model to investigate the magnetostrictive behavior of a macroscopic spherical MRE sample, which exhibits contraction along the magnetic field direction when aligned with the material's preferred direction.

cs.CE↗

A dual-stage constitutive modeling framework based on finite strain data-driven identification and physics-augmented neural networks

In this contribution, we present a novel consistent dual-stage approach for the automated generation of hyperelastic constitutive models which only requires experimentally measurable data. To generate input data for our approach, an experiment with full-field measurement has to be conducted to gather testing force and corresponding displacement field of the sample. Then, in the first step of the dual-stage framework, a new finite strain Data-Driven Identification (DDI) formulation is applied. This method enables to identify tuples consisting of stresses and strains by only prescribing the applied boundary conditions and the measured displacement field. In the second step, the data set is used to calibrate a Physics-Augmented Neural Network (PANN), which fulfills all common conditions of hyperelasticity by construction and is very flexible at the same time. We demonstrate the applicability of our approach by several descriptive examples. Two-dimensional synthetic data are exemplarily generated in virtual experiments by using a reference constitutive model. The calibrated PANN is then applied in 3D Finite Element simulations. In addition, a real experiment including noisy data is mimicked.

cs.CE↗

When invariants matter: The role of I1 and I2 in neural network models of incompressible hyperelasticity

For the formulation of machine learning-based material models, the usage of invariants of deformation tensors is attractive, since this can a priori guarantee objectivity and material symmetry. In this work, we consider incompressible, isotropic hyperelasticity, where two invariants I1 and I2 are required for depicting a deformation state. First, we aim at enhancing the understanding of the invariants. We provide an explicit representation of the set of invariants that are admissible, i.e. for which (I1, I2) a deformation state does indeed exist. Furthermore, we prove that uniaxial and equi-biaxial deformation correspond to the boundary of the set of admissible invariants. Second, we study how the experimentally-observed behaviour of different materials can be captured by means of neural network models of incompressible hyperelasticity, depending on whether both I1 and I2 or solely one of the invariants, i.e. either only I1 or only I2, are taken into account. To this end, we investigate three different experimental data sets from the literature. In particular, we demonstrate that considering only one invariant, either I1 or I2, can allow for good agreement with experiments in case of small deformations. In contrast, it is necessary to consider both invariants for precise models at large strains, for instance when rubbery polymers are deformed. Moreover, we show that multiaxial experiments are strictly required for the parameterisation of models considering I2. Otherwise, if only data from uniaxial deformation is available, significantly overly stiff responses could be predicted for general deformation states. On the contrary, I1-only models can make qualitatively correct predictions for multiaxial loadings even if parameterised only from uniaxial data, whereas I2-only models are completely incapable in even qualitatively capturing experimental stress data at large deformations.

cond-mat.mtrl-sci↗

Inverse design of anisotropic microstructures using physics-augmented neural networks

Composite materials often exhibit mechanical anisotropy owing to the material properties or geometrical configurations of the microstructure. This makes their inverse design a two-fold problem. First, we must learn the type and orientation of anisotropy and then find the optimal design parameters to achieve the desired mechanical response. In our work, we solve this challenge by first training a forward surrogate model based on the macroscopic stress-strain data obtained via computational homogenization for a given multiscale material. To this end, we use partially Input Convex Neural Networks (pICNNs) to obtain a polyconvex representation of the strain energy in terms of the invariants of the Cauchy-Green deformation tensor. The network architecture and the strain energy function are modified to incorporate, by construction, physics and mechanistic assumptions into the framework. While training the neural network, we find the type of anisotropy, if any, along with the preferred directions. Once the model is trained, we solve the inverse problem using an evolution strategy to obtain the design parameters that give a desired mechanical response. We test the framework against synthetic macroscale and also homogenized data. For cases where polyconvexity might be violated during the homogenization process, we present viable alternate formulations. The trained model is also integrated into a finite element framework to invert design parameters that result in a desired macroscopic response. We show that the invariant-based model is able to solve the inverse problem for a stress-strain dataset with a different preferred direction than the one it was trained on and is able to not only learn the polyconvex potentials of hyperelastic materials but also recover the correct parameters for the inverse design problem.

cs.CE↗

Neural networks meet anisotropic hyperelasticity: A framework based on generalized structure tensors and isotropic tensor functions

We present a data-driven framework for the multiscale modeling of anisotropic finite strain elasticity based on physics-augmented neural networks (PANNs). Our approach allows the efficient simulation of materials with complex underlying microstructures which reveal an overall anisotropic and nonlinear behavior on the macroscale. By using a set of invariants as input, an energy-type output and by adding several correction terms to the overall energy density functional, the model fulfills multiple physical principles by construction. The invariants are formed from the right Cauchy-Green deformation tensor and fully symmetric 2nd, 4th or 6th order structure tensors which enables to describe a wide range of symmetry groups. Besides the network parameters, the structure tensors are simultaneously calibrated during training so that the underlying anisotropy of the material is reproduced most accurately. In addition, sparsity of the model with respect to the number of invariants is enforced by adding a trainable gate layer and using lp regularization. Our approach works for data containing tuples of deformation, stress and material tangent, but also for data consisting only of tuples of deformation and stress, as is the case in real experiments. The developed approach is exemplarily applied to several representative examples, where necessary data for the training of the PANN surrogate model are collected via computational homogenization. We show that the proposed model achieves excellent interpolation and extrapolation behaviors. In addition, the approach is benchmarked against an NN model based on the components of the right Cauchy-Green deformation tensor.

cs.CE↗

Inverse design of spinodoid structures using Bayesian optimization

Tailoring materials to achieve a desired behavior in specific applications is of significant scientific and industrial interest as design of materials is a key driver to innovation. Overcoming the rather slow and expertise-bound traditional forward approaches of trial and error, inverse design is attracting substantial attention. Targeting a property, the design model proposes a candidate structure with the desired property. This concept can be particularly well applied to the field of architected materials as their structures can be directly tuned. The bone-like spinodoid materials are a specific class of architected materials. They are of considerable interest thanks to their non-periodicity, smoothness, and low-dimensional statistical description. Previous work successfully employed machine learning (ML) models for inverse design. The amount of data necessary for most ML approaches poses a severe obstacle for broader application, especially in the context of inelasticity. That is why we propose an inverse-design approach based on Bayesian optimization to operate in the small-data regime. Necessitating substantially less data, a small initial data set is iteratively augmented by in silico generated data until a structure with the targeted properties is found. The application to the inverse design of spinodoid structures of desired elastic properties demonstrates the framework's potential for paving the way for advance in inverse design.

cond-mat.mtrl-sci↗

Viscoelasticty with physics-augmented neural networks: Model formulation and training methods without prescribed internal variables

We present an approach for the data-driven modeling of nonlinear viscoelastic materials at small strains which is based on physics-augmented neural networks (NNs) and requires only stress and strain paths for training. The model is built on the concept of generalized standard materials and is therefore thermodynamically consistent by construction. It consists of a free energy and a dissipation potential, which can be either expressed by the components of their tensor arguments or by a suitable set of invariants. The two potentials are described by fully/partially input convex neural networks. For training of the NN model by paths of stress and strain, an efficient and flexible training method based on a recurrent cell, particularly a long short-term memory cell, is developed to automatically generate the internal variable(s) during the training process. The proposed method is benchmarked and thoroughly compared with existing approaches. These include a method that obtains the internal variable by integrating the evolution equation over the entire sequence, while the other method uses an an auxiliary feedforward neural network for the internal variable(s). Databases for training are generated by using a conventional nonlinear viscoelastic reference model, where 3D and 2D plane strain data with either ideal or noisy stresses are generated. The coordinate-based and the invariant-based formulation are compared and the advantages of the latter are demonstrated. Afterwards, the invariant-based model is calibrated by applying the three training methods using ideal or noisy stress data. All methods yield good results, but differ in computation time and usability for large data sets. The presented training method based on a recurrent cell turns out to be particularly robust and widely applicable and thus represents a promising approach for the calibration of other types of models as well.

cs.CE↗

Neural network-based multiscale modeling of finite strain magneto-elasticity with relaxed convexity criteria

We present a framework for the multiscale modeling of finite strain magneto-elasticity based on physics-augmented neural networks (NNs). By using a set of problem specific invariants as input, an energy functional as the output and by adding several non-trainable expressions to the overall total energy density functional, the model fulfills multiple physical principles by construction, e.g., thermodynamic consistency and material symmetry. Three NN-based models with varying requirements in terms of an extended polyconvexity condition of the magneto-elastic potential are presented. First, polyconvexity, which is a global concept, is enforced via input convex neural networks (ICNNs). Afterwards, we formulate a relaxed local version of the polyconvexity and fulfill it in a weak sense by adding a tailored loss term. As an alternative, a loss term to enforce the weaker requirement of strong ellipticity locally is proposed, which can be favorable to obtain a better trade-off between compatibility with data and physical constraints. Databases for training of the models are generated via computational homogenization for both compressible and quasi-incompressible magneto-active polymers (MAPs). Thereby, to reduce the computational cost, 2D statistical volume elements and an invariant-based sampling technique for the pre-selection of relevant states are used. All models are calibrated by using the database, whereby interpolation and extrapolation are considered separately. Furthermore, the performance of the NN models is compared to a conventional model from the literature. The numerical study suggests that the proposed physics-augmented NN approach is advantageous over the conventional model for MAPs. Thereby, the two more flexible NN models in combination with the weakly enforced local polyconvexity lead to good results, whereas the model based only on ICNNs has proven to be too restrictive.

cs.CE↗

Neural networks meet hyperelasticity: A guide to enforcing physics

In the present work, a hyperelastic constitutive model based on neural networks is proposed which fulfills all common constitutive conditions by construction, and in particular, is applicable to compressible material behavior. Using different sets of invariants as inputs, a hyperelastic potential is formulated as a convex neural network, thus fulfilling symmetry of the stress tensor, objectivity, material symmetry, polyconvexity, and thermodynamic consistency. In addition, a physically sensible stress behavior of the model is ensured by using analytical growth terms, as well as normalization terms which ensure the undeformed state to be stress free and with zero energy. In particular, polyconvex, invariant-based stress normalization terms are formulated for both isotropic and transversely isotropic material behavior. By fulfilling all of these conditions in an exact way, the proposed physics-augmented model combines a sound mechanical basis with the extraordinary flexibility that neural networks offer. Thus, it harmonizes the theory of hyperelasticity developed in the last decades with the up-to-date techniques of machine learning. Furthermore, the non-negativity of the hyperelastic neural network-based potentials is numerically examined by sampling the space of admissible deformations states, which, to the best of the authors' knowledge, is the only possibility for the considered nonlinear compressible models. For the isotropic neural network model, the sampling space required for that is reduced by analytical considerations. In addition, a proof for the non-negativity of the compressible Neo-Hooke potential is presented. The applicability of the model is demonstrated by calibrating it on data generated with analytical potentials, which is followed by an application of the model to finite element simulations. In addition, an adaption of the model to noisy data is shown and its [...]

cs.CE↗

Fast reconstruction of microstructures with ellipsoidal inclusions using analytical descriptors

Microstructure reconstruction is an important and emerging aspect of computational materials engineering and multiscale modeling and simulation. Despite extensive research and fast progress in the field, the application of descriptor-based reconstruction remains limited by computational resources. Common methods for increasing the computational feasibility of descriptor-based microstructure reconstruction lie in approximating the microstructure by simple geometrical shapes and by utilizing differentiable descriptors to enable gradient-based optimization. The present work combines these two ideas for structures composed of non-overlapping ellipsoidal inclusions such as magnetorheological elastomers. This requires to express the descriptors as a function of the microstructure parametrization. Deriving these relations leads to analytical solutions that further speed up the reconstruction procedure. Based on these descriptors, microstructure reconstruction is formulated as a multi-stage optimization procedure. The developed algorithm is validated by means of different numerical experiments and advantages and limitations are discussed in detail.

cond-mat.mtrl-sci↗

A comparative study on different neural network architectures to model inelasticity

The mathematical formulation of constitutive models to describe the path-dependent, i.e., inelastic, behavior of materials is a challenging task and has been a focus in mechanics research for several decades. There have been increased efforts to facilitate or automate this task through data-driven techniques, impelled in particular by the recent revival of neural networks (NNs) in computational mechanics. However, it seems questionable to simply not consider fundamental findings of constitutive modeling originating from the last decades research within NN-based approaches. Herein, we propose a comparative study on different feedforward and recurrent neural network architectures to model inelasticity. Within this study, we divide the models into three basic classes: black box NNs, NNs enforcing physics in a weak form, and NNs enforcing physics in a strong form. Thereby, the first class of networks can learn constitutive relations from data while the underlying physics are completely ignored, whereas the latter two are constructed such that they can account for fundamental physics, where special attention is paid to the second law of thermodynamics in this work. Conventional linear and nonlinear viscoelastic as well as elastoplastic models are used for training data generation and, later on, as reference. After training with random walk time sequences containing information on stress, strain, and, for some models, internal variables, the NN-based models are compared to the reference solution, whereby interpolation and extrapolation are considered. Besides the quality of the stress prediction, the related free energy and dissipation rate are analyzed to evaluate the models. Overall, the presented study enables a clear recording of the advantages and disadvantages of different NN architectures to model inelasticity and gives guidance on how to train and apply these models.

cs.CE↗