English
Related papers

Related papers: Scaling SU(2) to 1000 GPUs using HiRep

200 papers

The increasing scale and wealth of inter-connected data, such as those accrued by social network applications, demand the design of new techniques and platforms to efficiently derive actionable knowledge from large-scale graphs. However,…

Distributed, Parallel, and Cluster Computing · Computer Science 2014-12-08 Abdullah Gharaibeh , Tahsin Reza , Elizeu Santos-Neto , Lauro Beltrao Costa , Scott Sallinen , Matei Ripeanu

The Monte Carlo method is a powerful technique for computing thermodynamic magnetic states of otherwise unsolvable spin Hamiltonians, but the method becomes computationally prohibitive with increasing number of spins and the simulation of…

Computational Physics · Physics 2021-06-22 Michalis Charilaou

Simulating the real-time dynamics of lattice gauge theories, underlying the Standard Model of particle physics, is a notoriously difficult problem where quantum simulators can provide a practical advantage over classical approaches. In this…

Quantum Physics · Physics 2023-10-18 Torsten V. Zache , Daniel González-Cuadra , Peter Zoller

We present a hybrid numerical approach to simulate quantum many body problems on two spatial dimensional quantum lattice models via the non-Abelian ab initio version of the density matrix renormalization group method on state-of-the-art…

Strongly Correlated Electrons · Physics 2024-06-05 Andor Menczer , Kornél Kapás , Miklós Antal Werner , Örs Legeza

Scaling up the sparse matrix-vector multiplication kernel on modern Graphics Processing Units (GPU) has been at the heart of numerous studies in both academia and industry. In this article we present a novel non-parametric, self-tunable,…

Numerical Analysis · Computer Science 2012-12-24 Xintian Yang , Srinivasan Parthasarathy , Ponnuswamy Sadayappan

It has been proposed to abandon the requirement that parallel transporters in gauge theories are unitary (or pseudoorthogonal). This leads to a geometric interpretation of Vierbein fields as parts of gauge fields, and nonunitary parallel…

High Energy Physics - Lattice · Physics 2009-11-10 Claudia Lehmann , Gerhard Mack

Training large transformers is slow, but recent innovations on GPU architecture give us an advantage. NVIDIA Ampere GPUs can execute a fine-grained 2:4 sparse matrix multiplication twice as fast as its dense equivalent. In the light of this…

Machine Learning · Computer Science 2024-10-29 Yuezhou Hu , Kang Zhao , Weiyu Huang , Jianfei Chen , Jun Zhu

Neural networks have become indispensable for a wide range of applications, but they suffer from high computational- and memory-requirements, requiring optimizations from the algorithmic description of the network to the hardware…

Signal Processing · Electrical Eng. & Systems 2020-05-05 Andreas Toftegaard Kristensen , Robert Giterman , Alexios Balatsoukas-Stimming , Andreas Burg

Over the past decade there has been a growing interest in the development of parallel hardware systems for simulating large-scale networks of spiking neurons. Compared to other highly-parallel systems, GPU-accelerated solutions have the…

Neurons and Cognition · Quantitative Biology 2021-02-22 Bruno Golosio , Gianmarco Tiddia , Chiara De Luca , Elena Pastorelli , Francesco Simula , Pier Stanislao Paolucci

Surrogate models driven by sizeable datasets and scientific machine-learning methods have emerged as an attractive microstructure simulation tool with the potential to deliver predictive microstructure evolution dynamics with huge savings…

Materials Science · Physics 2024-01-22 Shaoxun Fan , Andrew L. Hitt , Ming Tang , Babak Sadigh , Fei Zhou

We study the performance of a cloud-based GPU-accelerated inference server to speed up event reconstruction in neutrino data batch jobs. Using detector data from the ProtoDUNE experiment and employing the standard DUNE grid job submission…

High Energy Physics - Experiment · Physics 2023-10-31 Tejin Cai , Kenneth Herner , Tingjun Yang , Michael Wang , Maria Acosta Flechas , Philip Harris , Burt Holzman , Kevin Pedro , Nhan Tran

We propose a generic algorithmic building block to accelerate training of machine learning models on heterogeneous compute systems. Our scheme allows to efficiently employ compute accelerators such as GPUs and FPGAs for the training of…

Machine Learning · Computer Science 2017-11-08 Celestine Dünner , Thomas Parnell , Martin Jaggi

Deep Gaussian processes (DGPs) are increasingly popular as predictive models in machine learning (ML) for their non-stationary flexibility and ability to cope with abrupt regime changes in training data. Here we explore DGPs as surrogates…

Methodology · Statistics 2021-08-27 Annie Sauer , Robert B. Gramacy , David Higdon

Intel Max GPUs are a new option available to CGYRO fusion simulation users. This paper outlines the changes that were needed to successfully run CGYRO on Intel Max 1550 GPUs on TACC's Stampede3 HPC system and presents benchmark results…

Distributed, Parallel, and Cluster Computing · Computer Science 2024-10-10 Igor Sfiligoi , Jeff Candy , Emily A. Belli

The open source HIP platform for GPU computing provides an uniform framework to support both the NVIDIA and AMD GPUs, and also the possibility to porting the CUDA code to the HIP- compatible one. We present the porting progress on the…

High Energy Physics - Lattice · Physics 2020-01-28 Yu-Jiang Bi , Yi Xiao , Ming Gong , Wei-Yi Guo , Peng Sun , Shun Xu , Yi-Bo Yang

We report on simulations with two flavors of O(a) improved degenerate Wilson fermions with Schroedinger functional boundary conditions. The algorithm which is used is Hybrid Monte Carlo with two pseudo-fermion fields as proposed by M.…

High Energy Physics - Lattice · Physics 2009-11-10 M. Della Morte , F. Knechtli , J. Rolf , R. Sommer , I. Wetzorke , U. Wolff

Realistic simulations of detailed, biophysics-based, multi-scale models require very high resolution and, thus, large-scale compute facilities. Existing simulation environments, especially for biomedical applications, are designed to allow…

Computational Engineering, Finance, and Science · Computer Science 2018-02-12 Chris Bradley , Nehzat Emamy , Thomas Ertl , Dominik Göddeke , Andreas Hessenthaler , Thomas Klotz , Aaron Krämer , Michael Krone , Benjamin Maier , Miriam Mehl , Tobias Rau , Oliver Röhrle

We explore aspects of the phase structure of SU(2) and SU(3) lattice gauge theories at strong coupling with many flavours $N_f$ of Wilson fermions in the fundamental representation. The pseudoscalar meson mass as a function of hopping…

High Energy Physics - Lattice · Physics 2010-04-30 Kei-ichi Nagai , Georgina Carrillo-Ruiz , Gergana Koleva , Randy Lewis

One significant advantage of superconducting processors is their extensive design flexibility, which encompasses various types of qubits and interactions. Given the large number of tunable parameters of a processor, the ability to perform…

Quantum Physics · Physics 2025-04-25 Ziang Wang , Feng Wu , Hui-Hai Zhao , Xin Wan , Xiaotong Ni

We present preliminary results of lattice simulations of SU(2) gauge theory with two Wilson fermions in the adjoint representation. This theory has recently attracted considerable attention because it might possess an infrared fixed point…

High Energy Physics - Lattice · Physics 2010-01-21 Ari Hietanen , Jarno Rantaharju , Kari Rummukainen , Kimmo Tuominen