arXiv ScienceSearch

arXiv subjects

Guowei He

Publications and source records attributed to Guowei He.

At least 19 recordsLinked to original sources

Geometry-Preserving Reduced-Order Modeling via Immersed Tensor Decomposition (ITD)

Body-fitted finite-element methods deliver high-order accuracy but hinge on a clean, watertight, conforming mesh, a requirement that breaks down for the geometrically imperfect CAD assemblies, image-based volumetric data, and voxel-native designs that pervade biomedical engineering and additive manufacturing, where mesh generation has become the dominant cost of the analysis cycle. Immersed methods on regular background Cartesian grids sidestep body-fitted meshing, but classical implementations integrate over irregular cut subdomains, destroying the tensor-product structure that enables separable, reduced-order methods such as tensor decomposition. In this paper we propose the \emph{Immersed Tensor Decomposition} (ITD) framework, which couples a mesh-free geometric representation via body-fitted function with the separable C-HiDeNN-TD reduced-order solver to enable large-scale simulation directly on regular background voxel meshes. The geometry is encoded in three steps: a signed-distance function represents the boundary, a body-fitted function $\Phi$ approximates it with controllable error, and a low-rank Tucker decomposition provides model-order reduction; for a fixed grid spacing $h$, accuracy is improved by raising the approximation order of C-HiDeNN interpolation up to degree $p$ with a linear background mesh. The central contribution is an exact Dirichlet formulation that enforces the boundary condition strongly by multiplying the trial function with $\Phi$, so that $u=g$ holds by construction without any variational penalty or interface quadrature. We establish an a priori error estimate for the formulation and assess it on canonical 2D/3D domains, demonstrating optimal convergence and robustness on non-Cartesian geometries discretized by regular voxel meshes.

math.NA

Balance-Guided Sparse Identification of Multiscale Nonlinear PDEs with Small-coefficient Terms

Data-driven discovery of governing equations has advanced significantly in recent years; however, existing methods often struggle in multiscale systems where dynamically significant terms may have small coefficients. Therefore, we propose Balance-Guided SINDy (BG-SINDy) inspired by the principle of dominant balance, which reformulates $\ell_0$-constrained sparse regression as a term-level $\ell_{2,0}$-regularized problem and solves it using a progressive pruning strategy. Terms are ranked according to their relative contributions to the governing equation balance rather than their absolute coefficient magnitudes. Based on this criterion, BG-SINDy alternates between least-squares regression and elimination of negligible terms, thereby preserving dynamically significant terms even when their coefficients are small. Numerical experiments on the Korteweg--de Vries equation with a small dispersion coefficient, a modified Burgers equation with vanishing hyperviscosity, a modified Kuramoto--Sivashinsky equation with multiple small-coefficient terms, and a two-dimensional reaction--diffusion system demonstrate the validity of BG-SINDy in discovering small-coefficient terms. The proposed method thus provides an efficient approach for discovering governing equations that contain small-coefficient terms.

cs.LG

Space-time correlations of passive scalars in colored-noise flows

The space-time correlation of a passive scalar advected by a Gaussian colored-noise velocity with wavenumber-dependent correlation times and power-law spatial spectra is investigated in the present paper. Within the inertial-convective subrange, we derive an analytical solution for the space-time correlation. This solution validates the elliptic approximation (EA) model [He and Zhang, Phys. Rev. E 73, 055303(R) (2006)], demonstrating that the iso-correlation contours are self-similar in the co-moving space-time frame $(r-U\tau, V\tau)$, with a universal spatial-to-temporal intercept ratio of 1.55. Unlike the classic Kraichnan white-noise model, our formulation simultaneously recovers the Obukhov--Corrsin scaling for spatial correlations (when the velocity obeys Kolmogorov scaling) and reproduces the random-sweeping mechanism, yielding Gaussian (rather than exponential) temporal decorrelation of scalar Fourier modes. Our results clarify the underlying decorrelation mechanism of passive scalars: mean-flow advection and large-scale sweeping dominate temporal decorrelation, and small-scale distortion dominates spatial decorrelation.

physics.flu-dyn

Data-driven detached-eddy simulations based on explicit algebraic stress expressions for turbulent flows

This work proposes a data-driven explicit algebraic stress-based detached-eddy simulation (DES) method. Despite the widespread use of data-driven methods in model development for both Reynolds-averaged Navier-Stokes (RANS) and large-eddy simulations (LES), their applications to DES remain limited. The challenge mainly lies in the absence of modelled stress data, the requirement for proper length scales in RANS and LES branches, and the maintenance of a reasonable switching behaviour. The data-driven DES method is constructed based on the algebraic stress equation. The control of RANS/LES switching is achieved through the eddy viscosity in the linear part of the modelled stress, under the $\ell^2-\omega$ DES framework. Three model coefficients associated with the pressure-strain terms and the LES length scale are represented by a neural network as functions of scalar invariants of velocity gradient. The neural network is trained using velocity data with the ensemble Kalman method, thereby circumventing the requirement for modelled stress data. Moreover, the baseline coefficient values are incorporated as additional reference data to ensure reasonable switching behaviour. The proposed approach is evaluated on two challenging turbulent flows, i.e., the secondary flow in a square duct and the separated flow over a bump. The trained model achieves significant improvements in predicting mean flow statistics compared to the baseline model. This is attributed to improved predictions of the modelled stress. The trained model also exhibits reasonable switching behaviour, enlarging the LES region to resolve more turbulent structures. Furthermore, the model shows satisfactory generalization capabilities for both cases in similar flow configurations.

physics.flu-dyn

K2-V2: A 360-Open, Reasoning-Enhanced LLM

We introduce K2-V2, a 360-open LLM built from scratch as a superior base for reasoning adaptation, in addition to functions such as conversation and knowledge retrieval from general LLMs. It stands as the strongest fully open model, rivals open-weight leaders in its size class, outperforms Qwen2.5-72B and approaches the performance of Qwen3-235B. We actively infuse domain knowledge, reasoning, long-context, and tool use throughout the training process. This explicitly prepares the model for complex reasoning tasks. We demonstrate this potential using simple supervised fine-tuning, establishing a strong baseline that indicates significant headroom for advanced alignment. By releasing the full training history and data composition, we maximize the effectiveness of continuous training, a key open source production scenario. We release the model weights and signature LLM360 artifacts, such as complete training data, to empower the community with a capable, reasoning-centric foundation.

cs.LG

Toward end-to-end quantum simulation of rapidly distorted turbulence

We propose an end-to-end quantum algorithm to simulate rapidly distorted turbulence via linear combination of Hamiltonian (LCHS). The algorithm comprises three primary stages: the efficient preparation of an initial turbulent state with a prescribed energy spectrum, its subsequent time evolution via LCHS, and the direct measurement of key turbulence statistics. Our analysis indicates that the algorithm can offer a practical quantum speedup over the classical simulation methods for a sufficiently large computational grid. We evaluate the quantum resource requirements for simulating a minimal instance of non-trivial turbulence with classical validation. The numerical results show excellent agreement with ground-truth solutions, capturing both the qualitative evolution of turbulent fields and the quantitative behavior of statistics, including the Reynolds stresses and the fluctuating velocity spectrum. Despite its linearity, rapidly distorted turbulence captures essential turbulence mechanisms and may inform the development of quantum algorithms for the Navier-Stokes equations. Our work establishes a foundation for addressing more complex turbulent phenomena on future fault-tolerant quantum computers.

physics.flu-dyn

Shape optimization for trailing-edge noise reduction using large-eddy simulation and ensemble-based method

In this work, the trailing-edge shape of an airfoil is optimized to reduce the acoustic noise based on large-eddy simulation (LES). It is achieved by the ensemble Kalman method, which can enhance the optimization efficiency by using the gradient of cost function approximated with sample covariances. Moreover, the update scheme is reformulated to impose smoothness regularization and enable simultaneous reduction in the trailing edge noise and the drag-to-lift ratio. The trailing edge is optimized with a reduced bevel angle based on the ensemble Kalman method. The flow field near the optimal trailing edge shows that the flow separation and vortex shedding are suppressed compared to the baseline shape, indicating a significant decrease in the drag-to-lift ratio and noise generation. Also, the spectral proper orthogonal decomposition method is used to analyze the flow structure around the trailing edge, identifying that the optimal shape achieves acoustic noise reduction by disrupting large-scale flow structures. Further, the spectrum of Lighthill stress reveals that the optimal trailing edge suppresses the high-frequency noise through the nonlinear interaction of reduced low-frequency velocity fluctuations.

physics.flu-dyn

K2-Think: A Parameter-Efficient Reasoning System

K2-Think is a reasoning system that achieves state-of-the-art performance with a 32B parameter model, matching or surpassing much larger models like GPT-OSS 120B and DeepSeek v3.1. Built on the Qwen2.5 base model, our system shows that smaller models can compete at the highest levels by combining advanced post-training and test-time computation techniques. The approach is based on six key technical pillars: Long Chain-of-thought Supervised Finetuning, Reinforcement Learning with Verifiable Rewards (RLVR), Agentic planning prior to reasoning, Test-time Scaling, Speculative Decoding, and Inference-optimized Hardware, all using publicly available open-source datasets. K2-Think excels in mathematical reasoning, achieving state-of-the-art scores on public benchmarks for open-source models, while also performing strongly in other areas such as Code and Science. Our results confirm that a more parameter-efficient model like K2-Think 32B can compete with state-of-the-art systems through an integrated post-training recipe that includes long chain-of-thought training and strategic inference-time enhancements, making open-source reasoning systems more accessible and affordable. K2-Think is freely available at k2think.ai, offering best-in-class inference speeds of over 2,000 tokens per second per request via the Cerebras Wafer-Scale Engine.

cs.LG

Onset of vortex shedding in flow past Rankine ovals

The Rankine oval is a classical geometry in potential flow, formed by superimposing a uniform stream with velocity U and a source-sink pair separated by distance 2a with strength m, resulting in a closed stagnation streamline whose shape is governed by the dimensionless parameter Ua/m. Although the Rankine body serves as a cornerstone for the classical theory of potential flow, its behavior in viscous flow remains unexplored. The Rankine oval is streamlined in inviscid flow but behaves as a bluff body in viscous flow. The onset of vortex shedding is a critical phenomenon in flows past a bluff body, mapping the transition from steady to periodic wakes. This study systematically investigates the onset of vortex shedding in Rankine oval flows and its associated fluid dynamics by performing direct numerical simulations of incompressible flow past Rankine ovals over Reynolds numbers from 10 to 200 and Ua/m from 0 to 1. The investigation reveals a linear relationship between Ua/m and the critical Reynolds number. This study further characterizes the lift and drag coefficients and Strouhal number, analyzes the vortex formation, and performs a data-driven dimensional analysis. This analysis identifies the dimensionless quantities and empirical formula that determine St and the friction drag coefficient as a function of Re, independent of Ua/m. For sufficiently large Ua/m, the pressure drag can be estimated using potential flow solutions, enabling reliable predictions of the total drag without numerical simulations. These conclusions collectively provide insights into the fluid dynamics of Rankine ovals across diverse flow conditions.

physics.flu-dyn

Re4: Scientific Computing Agent with Rewriting, Resolution, Review and Revision

Large language models (LLMs) serve as an active and promising field of generative artificial intelligence and have demonstrated abilities to perform complex tasks in multiple domains, including mathematical and scientific reasoning. In this work, we construct a novel agent framework for solving representative problems in scientific computing. The proposed agent, incorporating a "rewriting-resolution-review-revision" logical chain via three reasoning LLMs (functioning as the Consultant, Reviewer, and Programmer, respectively), is integrated in a collaborative and interactive manner. The Consultant module endows the agent with knowledge transfer capabilities to link problems to professional domain insights, thereby rewriting problem descriptions through text augmentation. The Programmer module is responsible for generating and executing well-structured code to deliver the problem resolution. The Reviewer module equips the agent with the capacity for self-debugging and self-refinement through interactive feedback with code runtime outputs. By leveraging the end-to-end review mechanism, the executable code provided by the Programmer attains the iterative revision. A comprehensive evaluation is conducted on the performance of the proposed agent framework in solving partial differential equations (PDEs), ill-conditioned linear systems, and data-driven physical analysis problems. Compared to single-model, this collaborative framework significantly improves the bug-free code generation rate and reduces the occurrence of non-physical solutions, thereby establishing a highly reliable framework for autonomous code generation based on natural language descriptions. The review mechanism improved the average execution success rate of the modern reasoning models. Our code is available at https://github.com/ChengAo21/Re4_Sci_Agent

cs.AI

CFDagent: A Language-Guided, Zero-Shot Multi-Agent System for Complex Flow Simulation

We introduce CFDagent, a zero-shot, multi-agent system that enables fully autonomous computational fluid dynamics (CFD) simulations from natural language prompts. CFDagent integrates three specialized LLM-driven agents: (i) the Preprocessing Agent that generates 3D geometries from textual or visual inputs using a hybrid text-to-3D diffusion model (Point-E) and automatically meshes the geometries; (ii) the Solver Agent that configures and executes an immersed boundary flow solver; and (iii) the Postprocessing Agent that analyzes and visualizes the results, including multimodal renderings. These agents are interactively guided by GPT-4o via conversational prompts, enabling intuitive and user-friendly interaction. We validate CFDagent by reproducing canonical sphere flows at Reynolds numbers of 100 and 300 using three distinct inputs: a simple text prompt (i.e., "sphere"), an image-based input, and a standard sphere model. The computed drag and lift coefficients from meshes produced by each input approach closely match available data. The proposed system enables synthesization of flow simulations and photorealistic visualizations for complex geometries. Through extensive tests on canonical and realistic scenarios, we demonstrate the robustness, versatility, and practical applicability of CFDagent. By bridging generative AI with high-fidelity simulations, CFDagent significantly lowers barriers to expert-level CFD, unlocking broad opportunities in education, scientific research, and practical engineering applications.

physics.flu-dyn

Lagrangian-Eulerian learning of flow field and trajectories with TrajectoryFlowNet

Predicting particle transport in complex flows is traditionally achieved by solving the Navier-Stokes equations. While various numerical and experimental methods exist, they typically require deep physical insights and incur high computational costs. Machine learning offers an alternative by learning predictive patterns directly from data, avoiding explicit physical modeling. However, purely data-driven approaches often lack interpretability, physical consistency, and generalizability in sparse data regimes. To this end, we propose TrajectoryFlowNet, a Lagrangian-Eulerian physics-informed neural network architecture, for fluid flow velocimetry and imaging via learning to predict spatiotemporal flow fields and long-range particle trajectories. The salient features of our model include its ability to handle complex flow patterns with irregular boundaries, predict the full-field flows, image the long-range flow trajectory of any arbitrary particle, and ensure physical consistency in predictions based only on very scarce measurement of flow trajectories. We validate TrajectoryFlowNet via both numerical examples (e.g., lid-driven cavity flow and complex cylinder flow) and experimental test cases (e.g., aortic and ventricle blood flows) across diverse flow scenarios. The results demonstrate our model's effectiveness in capturing intricate particle-laden flow dynamics, enabling long-range tracking of particles and accurate construction of flow fields in real-world applications.

physics.flu-dyn

A framework for learning symbolic turbulence models from indirect observation data via neural networks and feature importance analysis

Learning symbolic turbulence models from indirect observation data is of significant interest as it not only improves the accuracy of posterior prediction but also provides explicit model formulations with good interpretability. However, it typically resorts to gradient-free evolutionary algorithms, which can be relatively inefficient compared to gradient-based approaches, particularly when the Reynolds-averaged Navier-Stokes (RANS) simulations are involved in the training process. In view of this difficulty, we propose a framework that uses neural networks and the associated feature importance analysis to improve the efficiency of symbolic turbulence modeling. In doing so, the gradient-based method can be used to efficiently learn neural network-based representations of Reynolds stress from indirect data, which is further transformed into simplified mathematical expressions with symbolic regression. Moreover, feature importance analysis is introduced to accelerate the convergence of symbolic regression by excluding insignificant input features. The proposed training strategy is tested in the flow in a square duct, where it correctly learns underlying analytic models from indirect velocity data. Further, the method is applied in the flow over the periodic hills, demonstrating that the feature importance analysis can significantly improve the training efficiency and learn symbolic turbulence models with satisfactory generalizability.

physics.flu-dyn

MegaMath: Pushing the Limits of Open Math Corpora

Mathematical reasoning is a cornerstone of human intelligence and a key benchmark for advanced capabilities in large language models (LLMs). However, the research community still lacks an open, large-scale, high-quality corpus tailored to the demands of math-centric LLM pre-training. We present MegaMath, an open dataset curated from diverse, math-focused sources through following practices: (1) Revisiting web data: We re-extracted mathematical documents from Common Crawl with math-oriented HTML optimizations, fasttext-based filtering and deduplication, all for acquiring higher-quality data on the Internet. (2) Recalling Math-related code data: We identified high quality math-related code from large code training corpus, Stack-V2, further enhancing data diversity. (3) Exploring Synthetic data: We synthesized QA-style text, math-related code, and interleaved text-code blocks from web data or code data. By integrating these strategies and validating their effectiveness through extensive ablations, MegaMath delivers 371B tokens with the largest quantity and top quality among existing open math pre-training datasets.

cs.CL

LLM360 K2: Building a 65B 360-Open-Source Large Language Model from Scratch

We detail the training of the LLM360 K2-65B model, scaling up our 360-degree OPEN SOURCE approach to the largest and most powerful models under project LLM360. While open-source LLMs continue to advance, the answer to "How are the largest LLMs trained?" remains unclear within the community. The implementation details for such high-capacity models are often protected due to business considerations associated with their high cost. This lack of transparency prevents LLM researchers from leveraging valuable insights from prior experience, e.g., "What are the best practices for addressing loss spikes?" The LLM360 K2 project addresses this gap by providing full transparency and access to resources accumulated during the training of LLMs at the largest scale. This report highlights key elements of the K2 project, including our first model, K2 DIAMOND, a 65 billion-parameter LLM that surpasses LLaMA-65B and rivals LLaMA2-70B, while requiring fewer FLOPs and tokens. We detail the implementation steps and present a longitudinal analysis of K2 DIAMOND's capabilities throughout its training process. We also outline ongoing projects such as TXT360, setting the stage for future models in the series. By offering previously unavailable resources, the K2 project also resonates with the 360-degree OPEN SOURCE principles of transparency, reproducibility, and accessibility, which we believe are vital in the era of resource-intensive AI research.

cs.LG

LLM360: Towards Fully Transparent Open-Source LLMs

The recent surge in open-source Large Language Models (LLMs), such as LLaMA, Falcon, and Mistral, provides diverse options for AI practitioners and researchers. However, most LLMs have only released partial artifacts, such as the final model weights or inference code, and technical reports increasingly limit their scope to high-level design choices and surface statistics. These choices hinder progress in the field by degrading transparency into the training of LLMs and forcing teams to rediscover many details in the training process. We present LLM360, an initiative to fully open-source LLMs, which advocates for all training code and data, model checkpoints, and intermediate results to be made available to the community. The goal of LLM360 is to support open and collaborative AI research by making the end-to-end LLM training process transparent and reproducible by everyone. As a first step of LLM360, we release two 7B parameter LLMs pre-trained from scratch, Amber and CrystalCoder, including their training code, data, intermediate checkpoints, and analyses (at https://www.llm360.ai). We are committed to continually pushing the boundaries of LLMs through this open-source effort. More large-scale and stronger models are underway and will be released in the future.

cs.CL

Physical interpretation of neural network-based nonlinear eddy viscosity models

Neural network-based turbulence modeling has gained significant success in improving turbulence predictions by incorporating high--fidelity data. However, the interpretability of the learned model is often not fully analyzed, which has been one of the main criticism of neural network-based turbulence modeling. Therefore, it is increasingly demanding to provide physical interpretation of the trained model, which is of significant interest for guiding the development of interpretable and unified turbulence models. The present work aims to interpret the predictive improvement of turbulence flows based on the behavior of the learned model, represented with tensor basis neural networks. The ensemble Kalman method is used for model learning from sparse observation data due to its ease of implementation and high training efficiency. Two cases, i.e., flow over the S809 airfoil and flow in a square duct, are used to demonstrate the physical interpretation of the ensemble-based turbulence modeling. For the flow over the S809 airfoil, our results show that the ensemble Kalman method learns an optimal linear eddy viscosity model, which improves the prediction of the aerodynamic lift by reducing the eddy viscosity in the upstream boundary layer and promoting the early onset of flow separation. For the square duct case, the method provides a nonlinear eddy viscosity model, which predicts well secondary flows by capturing the imbalance of the Reynolds normal stresses. The flexibility of the ensemble-based method is highlighted to capture characteristics of the flow separation and secondary flow by adjusting the nonlinearity of the turbulence model.

physics.flu-dyn

Combining direct and indirect sparse data for learning generalizable turbulence models

Learning turbulence models from observation data is of significant interest in discovering a unified model for a broad range of practical flow applications. Either the direct observation of Reynolds stress or the indirect observation of velocity has been used to improve the predictive capacity of turbulence models. In this work, we propose combining the direct and indirect sparse data to train neural network-based turbulence models. The backpropagation technique and the observation augmentation approach are used to train turbulence models with different observation data in a unified ensemble-based framework. These two types of observation data can explore synergy to constrain the model training in different observation spaces, which enables learning generalizable models from very sparse data. The present method is tested in secondary flows in a square duct and separated flows over periodic hills. Both cases demonstrate that combining direct and indirect observations is able to improve the generalizability of the learned model in similar flow configurations, compared to using only indirect data. The ensemble-based method can serve as a practical tool for model learning from different types of observations due to its non-intrusive and derivative-free nature.

physics.flu-dyn