arXiv ScienceSearch

arXiv · 2006.07964

Accelerating CFD simulation with high order finite difference method on curvilinear coordinates for modern GPU clusters

Abstract

A high fidelity flow simulation for complex geometries for high Reynolds number ($Re$) flow is still very challenging, which requires more powerful computational capability of HPC system. However, the development of HPC with traditional CPU architecture suffers bottlenecks due to its high power consumption and technical difficulties. Heterogeneous architecture computation is raised to be a promising solution of difficulties of HPC development. GPU accelerating technology has been utilized in low order scheme CFD solvers on structured grid and high order scheme solvers on unstructured meshes. The high order finite difference methods on structured grid possess many advantages, e.g. high efficiency, robustness and low storage, however, the strong dependence among points for a high order finite difference scheme still limits its application on GPU platform. In present work, we propose a set of hardware-aware technology to optimize the efficiency of data transfer between CPU and GPU, and efficiency of communication between GPUs. An in-house multi-block structured CFD solver with high order finite difference methods on curvilinear coordinates is ported onto GPU platform, and obtain satisfying performance with speedup maximum around 2000x over a single CPU core. This work provides efficient solution to apply GPU computing in CFD simulation with certain high order finite difference methods on current GPU heterogeneous computers. The test shows that significant accelerating effects can been achieved for different GPUs.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Chuangchao Ye, Pengjunyi Zhang, Rui Yan, Dejun Sun, Zhenhua Wan. 2022-03-02. Accelerating CFD simulation with high order finite difference method on curvilinear coordinates for modern GPU clusters. https://doi.org/10.1186/s42774-021-00098-3

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

Spin disorder competing with positional symmetry breaking governs the metal-insulator behavior in oxide paramagnets

Numerous transition-metal oxides have low-temperature, long-range-ordered antiferromagnetic (AFM) states that are generally insulating, and high-temperature, disordered paramagnetic (PM) phases. The latter can be either insulating (predicted here for NaFeO3), or metallic (predicted here and previously observed in NaOsO3). Similar distinctions have been traditionally affected in strongly correlated models by the value used for Coulomb repulsion U. Here we show an alternative, strong-correlation-free (U=0) view suggesting that the distinction between insulating and metallic PM phases is governed by the competition between local magnetic moment disorder and the polymorphous distribution of off-center atomic displacements. Such parameter-free, energy-lowering symmetry breaking density functional calculations provide a framework for understanding metal-insulator behaviors across different quantum materials in terms of measurable local structural and magnetic parameters.

physics.comp-ph

Digital Twin of an Argon-Hydrogen Plasma Reactor

The principal proof of concept revolves around an argon-hydrogen plasma reactor that melts, reduces, atomizes and quenches critical raw material in one step, with premium spherical powder as the deliverable output and control of the composition chemistry. Each usage of the reactor is monitored through thermocouples and pressure sensors, which provide a daily data source of the real-world experiments. The reactor is modeled through COMSOL Multiphysics, which represents the core solver used to provide multiphysics simulations. The usage of COMSOL is complemented with Artificial Intelligence (AI) models, to enable seamless data assimilation and optimization. This paper presents the COMSOL twin of the reaction chamber and converging-diverging nozzle, together with a custom phase-change particle-tracing layer validated on Ti-6Al-4V (Ti64). Moreover, we highlight how the synergy between COMSOL simulations and AI-based digital surrogates can be leveraged to build self-consistent optimization loops geared toward (i) fully autonomous live control of the reactor and (ii) optimization of the process.

physics.comp-ph

Efficient calculation of inductive coupling for arrays of wire ring resonators

Generalization of the inductance to the case of non-quasistatic electromagnetic field oscillations appears to be fruitful when considering wireless power transfer and RF metamaterials consisting of thin wire loop meta-atoms. When dealing with large systems of interacting loops carrying currents, efficiency and precision of calculation in the presence of retardation is crucial. In this work, we derive a series expansion of such generalized inductance and propose a way for its efficient numerical approximation. Illustrative examples are provided both for inductance convergence of a pair of two loops and extinction efficiency for scattering by metamaterial samples.

physics.comp-ph