arXiv · 2211.15716
Distributed Parallelization of xPU Stencil Computations in Julia
Abstract
We present a straightforward approach for distributed parallelization of stencil-based xPU applications on a regular staggered grid, which is instantiated in the package ImplicitGlobalGrid.jl. The approach allows to leverage remote direct memory access and enables close to ideal weak scaling of real-world applications on thousands of GPUs. The communication costs can be easily hidden behind computation.
Explore related subjects
Keep this discovery
Samuel Omlin, Ludovic Räss, Ivan Utkin. 2022-11-28. Distributed Parallelization of xPU Stencil Computations in Julia. https://arxiv.org/abs/2211.15716
Cite the original work for its findings. Save a collection to share your selection of sources.