English

Wilson matrix kernel for lattice QCD on A64FX architecture

Distributed, Parallel, and Cluster Computing 2023-03-16 v1 High Energy Physics - Lattice

Abstract

We study the implementation of the even-odd Wilson fermion matrix for lattice QCD simulations on the A64FX architecture. Efficient coding of the stencil operation is investigated for two-dimensional packing to SIMD vectors. We measure the sustained performance on the supercomputer Fugaku at RIKEN R-CCS and show the profiler result of our code, which may signal an unexpected source of slow-down in addition to the detailed efficiency of each part of the code.

Cite

@article{arxiv.2303.08609,
  title  = {Wilson matrix kernel for lattice QCD on A64FX architecture},
  author = {Issaku Kanamori and Keigo Nitadori and Hideo Matsufuru},
  journal= {arXiv preprint arXiv:2303.08609},
  year   = {2023}
}

Comments

10 pages, contribtuion to the International Workshop on Arm-based HPC: Practice and Experience (IWAHPCE-2023), held in conjunction with The International Conference on High Performance Computing in Asia-Pacific Region (HPC Asia 2023), Singapore, Feb 27 - March 2, 2023

R2 v1 2026-06-28T09:18:27.998Z