arXiv Science⌕ Search

arXiv · 2610.09766

Performance Portable $\mathrm{SU}(N)$ Lattice Gauge Theory Simulation with Kokkos

Abstract

The increasing diversity of high performance computing systems makes separate, architecture specific implementations of lattice gauge theory algorithms costly to maintain. We present \texttt{kwqft}, a performance portable Kokkos implementation of Wilson pure gauge Monte Carlo simulation for $\mathrm{SU}(N)$ Yang-Mills theory in an arbitrary number of space-time dimensions. The gauge group order $N$ and the dimension $D$ are compile time parameters. A single source targets the Serial, OpenMP, CUDA, HIP, and SYCL execution spaces, with MPI halo exchange overlapped with interior updates. The implementation reproduces the exact two-dimensional plaquette and published three and four dimensional values for gauge groups up to $\mathrm{SU}(17)$. On an NVIDIA A100 the Kokkos CUDA backend is competitive with a native CUDA code, SIMD acceleration improves the OpenMP path on Armv9 processors, and a large scale speedup is demonstrated for an $\mathrm{SU}(4)$ lattice on the LineShine supercomputer, currently ranked first on the TOP500 list.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Wei Sun. 2026-10-07. Performance Portable $\mathrm{SU}(N)$ Lattice Gauge Theory Simulation with Kokkos. https://arxiv.org/abs/2610.09766

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

Bjorken sum rule down to the elastic region from the lattice

We present a lattice determination of the Bjorken sum rule via a Feynman-Hellmann calculation of the polarised forward Compton amplitude. We evaluate the lowest moment of the nucleon's polarised structure function, $g_1$, which we calculate over the range $0.5\lesssim Q^2 \lesssim 8$ GeV$^2$, covering both non-perturbative and perturbative regions. We compare our calculated moments with the Bjorken sum rule, and discuss implications for the calculation of higher-twist effects, as well as evaluations of $α_s(Q^2)$ from hadronic quantities to complement phenomenological and other lattice methods.

hep-lat↗

Finite-size scaling analysis of three dimensional Z(2) and O(2) spin models with non-vanishing symmetry breaking parameter

We present results from a detailed finite-size scaling analysis of the $3$-$d$, $Z(2)$ and $O(2)$ spin models in an external field $H$. Using high statistics Monte Carlo data, we obtain the leading finite-size scaling correction to the infinite volume scaling functions. We show that these corrections are proportional to $\widetilde{z}_L^2=1/(H L^{3/(1+1/δ))})^{2}$. This provides the parametric form for finite-size corrections to bulk thermodynamic observables as well as the pseudo-critical temperatures determined at non-vanishing $H$. In particular, it allows to eliminate systematic errors in the analysis of the chiral phase transition temperature in (2+1)-flavor QCD, that arose from the previously not well-controlled ansatz for infinite volume extrapolations. We also establish the validity range of this leading order correction and point out that there are significant differences between the $3$-$d$, $O(2)$ and $Z(2)$ universality classes.

hep-lat↗

Mass as pairing: an explicit Majorana mass on half of a generation does not seesaw in symmetric mass generation

A fermion mass pairs a field with a partner of conjugate charge: elementary (Higgs-Dirac), self-conjugate (seesaw) or a three-fermion composite (symmetric mass generation, SMG). We give half of a Spin(10)-like generation an explicit Majorana mass in a four-dimensional SMG model and ask what the other half does in a phase with no condensate. With the source on, every mass-type term of the unsourced half averages to zero under the symmetry group of the sourced action, so a light mass term, and a seesaw's light-heavy mixing, requires a spontaneously broken symmetry; exact on the lattice and a general proposition. The source is colour-democratic (a Delta_R-type mass on every flavour's right-handed taste doublet); the single-nu_R split lies outside the positivity class used but inside the symmetry statement. Measured up to 8^4: (i) in the symmetric-gap phase, with the source screened to at most 1.1% of free, the unsourced half's single-fermion readout stays at 0.3% of free while its pair susceptibility rises from 0.2% to 1.2-1.7% at h=2; on like mode sets it falls with the volume, on (4^4,8^4) no inequivalent light mass channel grows faster than free, and symmetry breaking is not resolved at L<=8; (ii) at the critical point the explicit mass collapses the critical sigma channel and moves the light half toward free as a saturating crossover, not a power law; direction tests (MIXED by their pre-registered letter) are consistent with a critical coupling rising with the mass; y_c(h) is not located. Pole versus zero of the light propagator at p->0 is not measured. The outcomes are readings after pre-registered tests that mostly returned no verdict. Also: no first-order signature of the merged transition at L<=8, a Z_4-odd response handed to a lattice composite, and flavour democracy of the positivity class. Every claim has an evidence class and a runnable check.

hep-lat↗