English
Related papers

Related papers: Two-link Staggered Quark Smearing in QUDA

200 papers

We review our work done to optimize the staggered conjugate gradient (CG) algorithm in the MILC code for use with the Intel Knights Landing (KNL) architecture. KNL is the second gener- ation Intel Xeon Phi processor. It is capable of…

High Energy Physics - Lattice · Physics 2016-11-04 Carleton DeTar , Douglas Doerfler , Steven Gottlieb , Ashish Jha , Dhiraj Kalamkar , Ruizi Li , Doug Toussaint

We discuss the CUDA approach to the simulation of pure gauge Lattice SU(2). CUDA is a hardware and software architecture developed by NVIDIA for computing on the GPU. We present an analysis and performance comparison between the GPU and CPU…

High Energy Physics - Lattice · Physics 2011-01-27 Nuno Cardoso , Pedro Bicudo

Our progress in computing the spectrum of excited baryons and mesons in lattice QCD is described. Sets of spatially-extended hadron operators with a variety of different momenta are used. A new method of stochastically estimating the…

High Energy Physics - Lattice · Physics 2015-05-27 C. Morningstar , A. Bell , J. Bulava , J. Foley , K. J. Juge , D. Lenkner , C. H. Wong

Spin squeezing is a powerful resource for quantum metrology, and recent hardware platforms based on interacting qubits provide multiple possible architectures to generate and reverse squeezing during a sensing protocol. In this work, we…

Quantum Physics · Physics 2025-12-11 Nickholas Gutierrez , Rodrigo Araiza Bravo , Susanne Yelin

We present and compare new types of algorithms for lattice QCD with staggered fermions in the limit of infinite gauge coupling. These algorithms are formulated on a discrete spatial lattice but with continuous Euclidean time. They make use…

High Energy Physics - Lattice · Physics 2012-12-03 Wolfgang Unger , Philippe de Forcrand

We give details of our precise determination of the light quark masses m_{ud}=(m_u+m_d)/2 and m_s in 2+1 flavor QCD, with simulated pion masses down to 120 MeV, at five lattice spacings, and in large volumes. The details concern the action…

High Energy Physics - Lattice · Physics 2015-05-20 S. Durr , Z. Fodor , C. Hoelbling , S. D. Katz , S. Krieg , T. Kurth , L. Lellouch , T. Lippert , K. K. Szabo , G. Vulvert

Noisy hardware forms one of the main hurdles to the realization of a near-term quantum internet. Distillation protocols allows one to overcome this noise at the cost of an increased overhead. We consider here an experimentally relevant…

Graphics Processing Units (GPUs) consisting of Streaming Multiprocessors (SMs) achieve high throughput by running a large number of threads and context switching among them to hide execution latencies. The number of thread blocks, and hence…

Hardware Architecture · Computer Science 2015-06-08 Vishwesh Jatala , Jayvant Anantpur , Amey Karkare

In this paper, we develop a new parallel auxiliary grid algebraic multigrid (AMG) method to leverage the power of graphic processing units (GPUs). In the construction of the hierarchical coarse grid, we use a simple and fixed coarsening…

Numerical Analysis · Mathematics 2012-12-07 Lu Wang , Xiaozhe Hu , Jonathan Cohen , Jinchao Xu

Tunable couplers enable high-fidelity two-qubit gates leveraging high on/off coupling ratios and reduced crosstalk within a single design. We investigate a galvanically connected direct-current superconducting quantum interference device…

Large Language Models (LLMs) have gained popularity in recent years, driving up the demand for inference. LLM inference is composed of two phases with distinct characteristics: a compute-bound prefill phase followed by a memory-bound decode…

Hardware Architecture · Computer Science 2025-10-10 Hengrui Zhang , Pratyush Patel , August Ning , David Wentzlaff

We present preliminary results from exploring the phase diagram of finite temperature QCD with three degenerate flavors and with two light flavors and the mass of the third held approximately at the strange quark mass. We use an order…

High Energy Physics - Lattice · Physics 2008-11-26 C. Bernard , T. Burch , S. Datta , T. A. DeGrand , C. E. DeTar , Steven Gottlieb , U. M. Heller , K. Orginos , R. L. Sugar , D. Toussaint

We simulate Quantum Chromodynamics in four Euclidean dimensions with two (degenerate mass) flavors of dynamical quarks. The Dirac operator is the so-called chirally improved operator that has been studied so far in quenched calculations. We…

High Energy Physics - Lattice · Physics 2007-05-23 C. B. Lang , Pushan Majumdar , Wolfgang Ortner

Persistent homology is a crucial invariant that is used in many areas to understand data. The $O(N^4)$ run time is a hindrance to its use on most large datasets. We give a parallelization method to utilize multi-core machines and clusters.…

Distributed, Parallel, and Cluster Computing · Computer Science 2022-03-10 Michael G. Rawson

While lattice QCD allows for reliable results at small momentum transfers (large quark separations), perturbative QCD is restricted to large momentum transfers (small quark separations). The latter is determined up to a reference momentum…

High Energy Physics - Phenomenology · Physics 2018-12-18 Felix Karbstein , Marc Wagner , Michelle Weber

We report on a calculation of $B_c$ ground state and radial excitation energies, obtained from heavy-charm highly improved staggered quark (HISQ) correlators computed on MILC gauge ensembles, with lattice spacings down to $a=0.044$ fm.…

High Energy Physics - Lattice · Physics 2018-11-26 Andrew Lytle , Brian Colquhoun , Christine Davies , Jonna Koponen

We have extended our program of QCD simulations with an improved Kogut-Susskind quark action to a smaller lattice spacing, approximately 0.09 fm. Also, the simulations with a approximately 0.12 fm have been extended to smaller quark masses.…

High Energy Physics - Lattice · Physics 2008-11-26 C. Aubin , C. Bernard , C. DeTar , Steven Gottlieb , E. B. Gregory , U. M. Heller , J. E. Hetrick , J. Osborn , R. Sugar , D. Toussaint

Analysis of processing time and similarity of images generated between CPU and GPU architectures and sequential and parallel programming. For image processing a computer with AMD FX-8350 processor and an Nvidia GTX 960 Maxwell GPU was used,…

We report on progress in our study of high temperature QCD with three flavors of improved staggered quarks. Simulations are being carried out with three degenerate quarks with masses less than or equal to the strange quark mass, $m_s$, and…

We present LBcuda, a GPU accelerated version of LBsoft, our open-source MPI-based software for the simulation of multi-component colloidal flows. We describe the design principles, the optimization and the resulting performance as compared…