arXiv ScienceSearch

arXiv subjects

Dunhui Xiao

Publications and source records attributed to Dunhui Xiao.

12 recordsLinked to original sources

Gappy probabilistic manifold decomposition for nonlinear field reconstruction

This paper proposes gappy probabilistic manifold decomposition (Gappy PMD), a nonlinear method for reconstructing high-dimensional fields from extremely sparse measurements. Gappy PMD reconstructs the field on the nonlinear manifold learned by probabilistic manifold decomposition (PMD). We further propose a differentiable point selection method for reduced-order model (ROM)-based field reconstruction (DPS). Using differentiable meshless interpolation within the ROM-based reconstruction framework, DPS makes the full-field reconstruction error differentiable with respect to the sampling locations and directly optimizes these locations. In addition, a theoretical error analysis for Gappy PMD is also given. It splits the squared reconstruction error into two orthogonal parts: one normal to the reconstruction manifold and the other induced by sparse sampling and observation noise. Under a stability condition on the sampling operator, this error vanishes with the PMD approximation error and the noise. The Gappy PMD is evaluated on three numerical test cases: flow past a cylinder, lid-driven cavity flow, and backward-facing step flow. For the same reduced dimension and sampling points, Gappy PMD attains mean relative $L^2$ errors one to two orders of magnitude below Gappy POD. Optimizing the sampling points with DPS further improves reconstruction accuracy and robustness.

math.NA

Convergence analysis of Parametric Probabilistic Manifold Decomposition

This paper presents a convergence analysis for a newly developed nonlinear model reduction method: parametric probabilistic manifold decomposition (PPMD)~\cite{guo2026parametric}. In addition, existing analyzes of nonlinear reduced order models typically treat subspace reduction, manifold representation, regression, and nonlinear reconstruction as separate components and often remain at the level of discrete state vectors. To the best of our knowledge, no theory tracks the complete error propagation in a data-dependent model whose basis, residual geometry, spectral coordinates, parameter maps, and lifting operator are all learned from the same numerical solution data. We develop a coupled perturbation analysis for the entire PPMD procedure. A trajectory geometry induced by the spatial discretization and temporal quadrature connects discrete trajectory vectors isometrically with the corresponding PDE norm. Population spectral objects are introduced to align the empirical residual coordinates and derive a uniform coordinate error estimate, whose propagation through the Hilbert-valued kernel lifting estimator is then quantified. Combining these results with the full order discretization error, weighted low-rank approximation, parameter regression, and residual representation defect yields deterministic and high-probability trajectory error bounds and consistency in probability in the continuous PDE trajectory space. The theory identifies how the principal errors interact and which components limit the accuracy of the nonlinear reduced order model.

math.NA

Accurate identification and measurement of the precipitate area by two-stage deep neural networks in novel chromium-based alloys

The performance of advanced materials for extreme environments is underpinned by their microstructure, including the size and distribution of reinforcing phases. Chromium-based superalloys are a recently proposed alternative to conventional face-centred-cubic superalloys for high-temperature applications, such as Concentrated Solar Power, and their development requires efficient measurement of precipitate volume fraction and size distribution from electron microscopy images. Traditional fixed-threshold image processing is sensitive to background noise, generalises poorly across materials, and requires substantial manual measurement effort. To address these bottlenecks, this study proposes DT-SegNet, an end-to-end two-stage deep learning scheme based on YOLOv5 and SegFormer for object detection and segmentation in electron microscopy images. The approach combines the training efficiency of convolutional neural networks at the detection stage with the segmentation accuracy of a Vision Transformer. Numerical experiments show that DT-SegNet substantially outperforms state-of-the-art segmentation tools offered by Weka and ilastik across metrics including accuracy, precision, recall, and F1-score. The model provides a useful tool for alloy-development microstructure examinations and helps address the large datasets associated with high-throughput alloy development.

cs.CV

Coupling-Informed Transport Maps for Bayesian Filtering in Nonlinear Dynamical Systems

A likelihood-free transport filtering method is proposed based on the couplings between state and observation variables. By exploiting a block-triangular structure in the transport map, the analysis step of filtering is reformulated as the minimization of the maximum mean discrepancy (MMD) between the true joint measure and its transport-based approximation. To circumvent the non-convexity in the MMD optimization, we introduce a training-free transport filter method via gradient flows, which leads to an analytic computation for the transport map that implies the steepest descent direction of the MMD. The proposed approach accurately approximates non-Gaussian filtering posteriors and avoids particle collapse. We provide a convergence analysis for the expectation of the MMD between the approximated posterior and the truth posterior. Finally, we extend the method to high-dimensional problems through domain localization. Numerical examples demonstrate the superior performance of our approach over conventional filtering methods in nonlinear, non-Gaussian scenarios.

stat.ML

Parametric Probabilistic Manifold Decomposition for Nonlinear Model Reduction

Probabilistic Manifold Decomposition (PMD)\cite{doi:10.1137/25M1738863}, developed in our earlier work, provides a nonlinear model reduction by embedding high-dimensional dynamics onto low-dimensional probabilistic manifolds. The PMD has demonstrated strong performance for time-dependent systems. However, its formulation is for temporal dynamics and does not directly accommodate parametric variability, which limits its applicability to tasks such as design optimization, control, and uncertainty quantification. In order to address the limitations, a \emph{Parametric Probabilistic Manifold Decomposition} (PPMD) is presented to deal with parametric problems. The central advantage of PPMD is its ability to construct continuous, high-fidelity parametric surrogates while retaining the transparency and non-intrusive workflow of PMD. By integrating probabilistic-manifold embeddings with parameter-aware latent learning, PPMD enables smooth predictions across unseen parameter values (such as different boundary or initial conditions). To validate the proposed method, a comprehensive convergence analysis is established for PPMD, covering the approximation of the linear principal subspace, the geometric recovery of the nonlinear solution manifold, and the statistical consistency of the kernel ridge regression used for latent learning. The framework is then numerically demonstrated on two classic flow configurations: flow past a cylinder and backward-facing step flow. Results confirm that PPMD achieves superior accuracy and generalization beyond the training parameter range compared to the conventional proper orthogonal decomposition with Gaussian process regression (POD+GPR) method.

math.NA

Quantum machine learning for efficient reduced order modelling of turbulent flows

Accurately predicting turbulent flows remains a central challenge in fluid dynamics due to their high dimensionality and intrinsic nonlinearity. Recent developments in quantum algorithms and machine learning offer new opportunities for overcoming the computational barriers inherent in turbulence modeling. Here we present a new hybrid quantum-classical framework that enables faster-than-real-time turbulence prediction by integrating machine learning, quantum computation, and fluid dynamics modeling, in particular, the reduced-order modeling. The novel framework combines quantum proper orthogonal decomposition (QPOD) with a quantum-enhanced deep kernel learning (QDKL) approach. QPOD employs quantum circuits to perform efficient eigenvalue decomposition for low-rank flow reconstruction, while QDKL exploits quantum entanglement and nonlinear mappings to enhance kernel expressivity and dynamic prediction accuracy. The new method is demonstrated on three benchmark turbulent flows, our architecture achieves significantly improved predictive accuracy at reduced model ranks, with training speeds up to 10 times faster and parameter counts reduced by a factor of 1/N compared to classical counterparts, where N is the input dimensionality. Although constrained by current noisy intermediate-scale quantum (NISQ) hardware, our results demonstrate the potential of quantum machine learning to transform turbulence simulation and lay a solid foundation for scalable, real-time quantum fluid modeling in future quantum computers.

physics.flu-dyn

Nonlinear Model Reduction by Probabilistic Manifold Decomposition

This paper presents a novel non-linear model reduction method: Probabilistic Manifold Decomposition (PMD), which provides a powerful framework for constructing non-intrusive reduced-order models (ROMs) by embedding a high-dimensional system into a low-dimensional probabilistic manifold and predicting the dynamics. Through explicit mappings, PMD captures both linearity and non-linearity of the system. A key strength of PMD lies in its predictive capabilities, allowing it to generate stable dynamic states based on embedded representations. The method also offers a mathematically rigorous approach to analyze the convergence of linear feature matrices and low-dimensional probabilistic manifolds, ensuring that sample-based approximations converge to the true data distributions as sample sizes increase. These properties, combined with its computational efficiency, make PMD a versatile tool for applications requiring high accuracy and scalability, such as fluid dynamics simulations and other engineering problems. By preserving the geometric and probabilistic structures of the high-dimensional system, PMD achieves a balance between computational speed, accuracy, and predictive capabilities, positioning itself as a robust alternative to the traditional model reduction method.

math.NA

Machine learning for modelling unstructured grid data in computational physics: a review

Unstructured grid data are essential for modelling complex geometries and dynamics in computational physics. Yet, their inherent irregularity presents significant challenges for conventional machine learning (ML) techniques. This paper provides a comprehensive review of advanced ML methodologies designed to handle unstructured grid data in high-dimensional dynamical systems. Key approaches discussed include graph neural networks, transformer models with spatial attention mechanisms, interpolation-integrated ML methods, and meshless techniques such as physics-informed neural networks. These methodologies have proven effective across diverse fields, including fluid dynamics and environmental simulations. This review is intended as a guidebook for computational scientists seeking to apply ML approaches to unstructured grid data in their domains, as well as for ML researchers looking to address challenges in computational physics. It places special focus on how ML methods can overcome the inherent limitations of traditional numerical techniques and, conversely, how insights from computational physics can inform ML development. To support benchmarking, this review also provides a summary of open-access datasets of unstructured grid data in computational physics. Finally, emerging directions such as generative models with unstructured data, reinforcement learning for mesh generation, and hybrid physics-data-driven paradigms are discussed to inspire future advancements in this evolving field.

cs.LG

Parametric Taylor series based latent dynamics identification neural networks

Numerical solving parameterised partial differential equations (P-PDEs) is highly practical yet computationally expensive, driving the development of reduced-order models (ROMs). Recently, methods that combine latent space identification techniques with deep learning algorithms (e.g., autoencoders) have shown great potential in describing the dynamical system in the lower dimensional latent space, for example, LaSDI, gLaSDI and GPLaSDI. In this paper, a new parametric latent identification of nonlinear dynamics neural networks, P-TLDINets, is introduced, which relies on a novel neural network structure based on Taylor series expansion and ResNets to learn the ODEs that govern the reduced space dynamics. During the training process, Taylor series-based Latent Dynamic Neural Networks (TLDNets) and identified equations are trained simultaneously to generate a smoother latent space. In order to facilitate the parameterised study, a $k$-nearest neighbours (KNN) method based on an inverse distance weighting (IDW) interpolation scheme is introduced to predict the identified ODE coefficients using local information. Compared to other latent dynamics identification methods based on autoencoders, P-TLDINets remain the interpretability of the model. Additionally, it circumvents the building of explicit autoencoders, avoids dependency on specific grids, and features a more lightweight structure, which is easy to train with high generalisation capability and accuracy. Also, it is capable of using different scales of meshes. P-TLDINets improve training speeds nearly hundred times compared to GPLaSDI and gLaSDI, maintaining an $L_2$ error below $2\%$ compared to high-fidelity models.

cs.LG

Deep-learning assisted reduced order model for high-dimensional flow prediction from sparse data

The reconstruction and prediction of full-state flows from sparse data are of great scientific and engineering significance yet remain challenging, especially in applications where data are sparse and/or subjected to noise. To this end, this study proposes a deep-learning assisted non-intrusive reduced order model (named DCDMD) for high-dimensional flow prediction from sparse data. Based on the compressed sensing (CS)-Dynamic Mode Decomposition (DMD), the DCDMD model is distinguished by two novelties. Firstly, a sparse matrix is defined to overcome the strict random distribution condition of sensor locations in CS, thus allowing flexible sensor deployments and requiring very few sensors. Secondly, a deep-learning-based proxy is invoked to acquire coherent flow modes from the sparse data of high-dimensional flows, thereby addressing the issue of defining sparsity and the stringent incoherence condition in the conventional CSDMD. The two advantageous features, combined with the fact that the model retains flow physics in the online stage, lead to significant enhancements in accuracy and efficiency, as well as superior insensitivity to data noises (i.e., robustness), in both reconstruction and prediction of full-state flows. These are demonstrated by three benchmark examples, i.e., cylinder wake, weekly-mean sea surface temperature and isotropic turbulence in a periodic square area.

physics.flu-dyn

Machine learning with data assimilation and uncertainty quantification for dynamical systems: a review

Data Assimilation (DA) and Uncertainty quantification (UQ) are extensively used in analysing and reducing error propagation in high-dimensional spatial-temporal dynamics. Typical applications span from computational fluid dynamics (CFD) to geoscience and climate systems. Recently, much effort has been given in combining DA, UQ and machine learning (ML) techniques. These research efforts seek to address some critical challenges in high-dimensional dynamical systems, including but not limited to dynamical system identification, reduced order surrogate modelling, error covariance specification and model error correction. A large number of developed techniques and methodologies exhibit a broad applicability across numerous domains, resulting in the necessity for a comprehensive guide. This paper provides the first overview of the state-of-the-art researches in this interdisciplinary field, covering a wide range of applications. This review aims at ML scientists who attempt to apply DA and UQ techniques to improve the accuracy and the interpretability of their models, but also at DA and UQ experts who intend to integrate cutting-edge ML approaches to their systems. Therefore, this article has a special focus on how ML methods can overcome the existing limits of DA and UQ, and vice versa. Some exciting perspectives of this rapidly developing research field are also discussed.

cs.LG

Egret Swarm Optimization Algorithm: An Evolutionary Computation Approach for Model Free Optimization

A novel meta-heuristic algorithm, Egret Swarm Optimization Algorithm (ESOA), is proposed in this paper, which is inspired by two egret species' (Great Egret and Snowy Egret) hunting behavior. ESOA consists of three primary components: Sit-And-Wait Strategy, Aggressive Strategy as well as Discriminant Conditions. The performance of ESOA on 36 benchmark functions as well as 2 engineering problems are compared with Particle Swarm Optimization (PSO), Genetic Algorithm (GA), Differential Evolution (DE), Grey Wolf Optimizer (GWO), and Harris Hawks Optimization (HHO). The result proves the superior effectiveness and robustness of ESOA. The source code used in this work can be retrieved from https://github.com/Knightsll/Egret_Swarm_Optimization_Algorithm; https://ww2.mathworks.cn/matlabcentral/fileexchange/115595-egret-swarm-optimization-algorithm-esoa.

cs.NE