arXiv · 2305.07030
Using Hierarchical Parallelism to Accelerate the Solution of Many Small Partial Differential Equations
Abstract
This paper presents efforts to improve the hierarchical parallelism of a two scale simulation code. Two methods to improve the GPU parallel performance were developed and compared. The first used the NVIDIA Multi-Process Service and the second moved the entire sub-problem loop into a single kernel using Kokkos hierarchical parallelism and a PackedView data structure. Both approaches improved parallel performance with the second method providing the greatest improvements.
Explore related subjects
Keep this discovery
Jacob Merson, Mark S. Shephard. 2023-05-05. Using Hierarchical Parallelism to Accelerate the Solution of Many Small Partial Differential Equations. https://arxiv.org/abs/2305.07030
Cite the original work for its findings. Save a collection to share your selection of sources.