English
Related papers

Related papers: Development of Lattice QCD Tool Kit on Cell Broadb…

200 papers

In this work we explore the performance of CUDA in quenched lattice SU(2) simulations. CUDA, NVIDIA Compute Unified Device Architecture, is a hardware and software architecture developed by NVIDIA for computing on the GPU. We present an…

High Energy Physics - Lattice · Physics 2015-03-17 Nuno Cardoso , Pedro Bicudo

Quantum computers (QCs) have the potential to solve critical problems significantly faster than today's most advanced supercomputers. One major challenge in realizing this technology is designing robust electrostatic pulses to realize…

Quantum Physics · Physics 2025-11-18 Dhilan Nag , Suhun Kim , Cole Johnson , Collin Sumrell

A computational system for lattice QCD with exact chiral symmetry is described. The platform is a home-made Linux PC cluster, built with off-the-shelf components. At present this system constitutes of 64 nodes, with each node consisting of…

High Energy Physics - Lattice · Physics 2011-02-16 Ting-Wai Chiu , Tung-Han Hsieh , Chao-Hsi Huang , Tsung-Ren Huang

The exponentially growing model size drives the continued success of deep learning, but it brings prohibitive computation and memory cost. From the algorithm perspective, model sparsification and quantization have been studied to alleviate…

Distributed, Parallel, and Cluster Computing · Computer Science 2023-05-09 Shigang Li , Kazuki Osawa , Torsten Hoefler

Given a quantum gate circuit, how does one execute it in a fault-tolerant architecture with as little overhead as possible? In this paper, we discuss strategies for surface-code quantum computing on small, intermediate and large scales.…

Quantum Physics · Physics 2019-03-07 Daniel Litinski

We present the results of a precision computation of B_K with Wilson fermions. Simulations are performed at different lattice spacings, enabling continuum limit extrapolations. Two different twisted mass QCD (tmQCD) regularisations are…

High Energy Physics - Lattice · Physics 2007-05-23 P. Dimopoulos , J. Heitger , C. Pena , S. Sint , A. Vladikas

We develop an automated framework for proving lower bounds on the bilinear complexity of matrix multiplication over finite fields. Our approach systematically combines orbit classification of the restricted first matrix and dynamic…

Computational Complexity · Computer Science 2026-05-19 Chengu Wang

Low-bit quantized neural networks are of great interest in practical applications because they significantly reduce the consumption of both memory and computational resources. Binary neural networks are memory and computationally efficient…

Machine Learning · Computer Science 2022-05-20 Anton Trusov , Elena Limonova , Dmitry Nikolaev , Vladimir V. Arlazarov

An overview is given of the QCDOC architecture, a massively parallel and highly scalable computer optimized for lattice QCD using system-on-a-chip technology. The heart of a single node is the PowerPC-based QCDOC ASIC, developed in…

High Energy Physics - Lattice · Physics 2007-05-23 P. A. Boyle , C. Jung , T. Wettig

The fast proliferation of extreme-edge applications using Deep Learning (DL) based algorithms required dedicated hardware to satisfy extreme-edge applications' latency, throughput, and precision requirements. While inference is achievable…

Hardware Architecture · Computer Science 2022-04-26 Yvan Tortorella , Luca Bertaccini , Davide Rossi , Luca Benini , Francesco Conti

Modern graphics hardware is designed for highly parallel numerical tasks and promises significant cost and performance benefits for many scientific applications. One such application is lattice quantum chromodyamics (lattice QCD), where the…

High Energy Physics - Lattice · Physics 2010-12-06 M. A. Clark , R. Babich , K. Barros , R. C. Brower , C. Rebbi

We calculate---for the first time in three-flavor lattice QCD---the hadronic matrix elements of all five local operators that contribute to neutral $B^0$- and $B_s$-meson mixing in and beyond the Standard Model. We present a complete error…

We study the feasibility of a PC-based parallel computer for medium to large scale lattice QCD simulations. The E\"otv\"os Univ., Inst. Theor. Phys. cluster consists of 137 Intel P4-1.7GHz nodes with 512 MB RDRAM. The 32-bit, single…

High Energy Physics - Lattice · Physics 2009-11-07 Z. Fodor , S. D. Katz , G. Papp

Latency and efficiency issues are often overlooked when evaluating IR models based on Pretrained Language Models (PLMs) in reason of multiple hardware and software testing scenarios. Nevertheless, efficiency is an important part of such…

Information Retrieval · Computer Science 2022-07-11 Carlos Lassance , Stéphane Clinchant

We present a 625 MHz clocked coherent one-way quantum key distribution (QKD) system which continuously distributes secret keys over an optical fibre link. To support high secret key rates, we implemented a fast hardware key distillation…

Sparse data structures are commonly used in neural networks to reduce the memory footprint. These data structures are compact but cause irregularities such as random memory accesses, which prevent efficient use of the memory hierarchy. GPUs…

Programming Languages · Computer Science 2025-06-19 Hossein Albakri , Kazem Cheshmi

We calculate, in the continuum limit of quenched lattice QCD, the form factor that enters in the decay rate of the semileptonic decay B --> D l nu. Making use of the step scaling method (SSM), previously introduced to handle two scale…

High Energy Physics - Lattice · Physics 2008-11-26 G. M. de Divitiis , E. Molinaro , R. Petronzio , N. Tantalo

In the near-future noisy intermediate-scale quantum (NISQ) era of quantum computing technology, applications of quantum computing will be limited to calculations of very modest scales in terms of the number of qubits used. The need to…

Quantum Physics · Physics 2019-07-03 Daniel C. Hackett , Kiel Howe , Ciaran Hughes , William Jay , Ethan T. Neil , James N. Simone

Modern Neural Network (NN) architectures heavily rely on vast numbers of multiply-accumulate arithmetic operations, constituting the predominant computational cost. Therefore, this paper proposes a high-throughput, scalable and energy…

Hardware Architecture · Computer Science 2024-07-09 Xuqi Zhu , Huaizhi Zhang , JunKyu Lee , Jiacheng Zhu , Chandrajit Pal , Sangeet Saha , Klaus D. McDonald-Maier , Xiaojun Zhai

Recent experimental progress in realizing surface code on hardware, including demonstrations of break-even logical memory on devices with up to hundreds of physical qubits, has materially advanced the prospects for fault-tolerant quantum…