Performance Optimization of Baryon-block Construction in the Stochastic LapH Method
December 07, 2022
Implementations of measurement kernels in high-level Lattice QCD frameworks enable rapid prototyping, but can leave hardware capabilities significantly underutilized. This is an acceptable tradeoff if the time spent in unoptimized routines is generally small. The computational cost of modern spectroscopy projects however can be comparable to or even exceed the cost of generating gauge configurations and computing solutions of the Dirac equation. One such key kernel in the stochastic LapH method is the computation of baryon blocks; we discuss several implementation strategies and achieve a 7.2x speedup over the current implementation on a system with Intel® Xeon® Platinum 8358 processors, formerly Ice Lake.
How to cite
Metadata are provided both in "article" format (very similar to INSPIRE) as this helps creating
very compact bibliographies which can be beneficial to authors and
readers, and in "proceeding" format
which is more detailed and complete.