English
Related papers

Related papers: apeNEXT: A multi-TFlops Computer for Simulations i…

200 papers

We introduce CORTEX, an algorithmic framework designed for large-scale brain simulation. Leveraging the computational capacity of the Fugaku Supercomputer, CORTEX maximizes available problem size and processing performance. Our primary…

Distributed, Parallel, and Cluster Computing · Computer Science 2024-06-07 Tianxiang Lyu , Mitsuhisa Sato , Shigeki Aoki , Ryutaro Himeno , Zhe Sun

Large-scale distributed graph-parallel computing is challenging. On one hand, due to the irregular computation pattern and lack of locality, it is hard to express parallelism efficiently. On the other hand, due to the scale-free nature,…

Distributed, Parallel, and Cluster Computing · Computer Science 2013-10-22 Jie Yan , Guangming Tan , Ninghui Sun

We define a benchmark suite for lattice QCD and report on benchmark results from several computer platforms. The platforms considered are apeNEXT, CRAY T3E, Hitachi SR8000, IBM p690, PC-Clusters, and QCDOC.

High Energy Physics - Lattice · Physics 2009-11-10 M. Hasenbusch , K. Jansen , D. Pleiter , H. St"uben , P. Wegner , T. Wettig , H. Wittig

Amplitude Estimation (AE) is a critical subroutine in many quantum algorithms, allowing for a quadratic speedup in various applications like those involving estimating statistics of various functions as in financial Monte Carlo simulations.…

Quantum Physics · Physics 2022-01-28 Salvatore Certo , Anh Dung Pham , Daniel Beaulieu

AMD Xilinx's new Versal Adaptive Compute Acceleration Platform (ACAP) is an FPGA architecture combining reconfigurable fabric with other on-chip hardened compute resources. AI engines are one of these and, by operating in a highly…

Distributed, Parallel, and Cluster Computing · Computer Science 2023-01-31 Nick Brown

In this article, we review some of the recent developments towards the future goal of quantum computing or quantum simulating lattice QCD. This includes a novel theoretical framework developed for non-Abelian gauge theories that is the…

High Energy Physics - Lattice · Physics 2021-07-21 Indrakshi Raychowdhury

High-energy physics (HEP) experiments have developed millions of lines of code over decades that are optimized to run on traditional x86 CPU systems. However, we are seeing a rapidly increasing fraction of floating point computing power in…

Developed by the APE group, APENet is a new high speed, low latency, 3-dimensional interconnect architecture optimized for PC clusters running LQCD-like numerical applications. The hardware implementation is based on a single PCI-X 133MHz…

High Energy Physics - Lattice · Physics 2009-11-10 R. Ammendola , M. Guagnelli , G. Mazza , F. Palombi , R. Petronzio , D. Rossetti , A. Salamon , P. Vicini

Deploying large language models (LLMs) for online inference is often constrained by limited GPU memory, particularly due to the growing KV cache during auto-regressive decoding. Hybrid GPU-CPU execution has emerged as a promising solution…

Distributed, Parallel, and Cluster Computing · Computer Science 2026-01-16 Jiakun Fan , Yanglin Zhang , Xiangchen Li , Dimitrios S. Nikolopoulos

The infinite projected entangled-pair state (iPEPS) ansatz is a powerful tensor-network approximation of an infinite two-dimensional quantum many-body state. Tensor-based calculations are particularly well-suited to utilize the high…

Strongly Correlated Electrons · Physics 2025-03-19 Addison D. S. Richards , Erik S. Sørensen

The CP-PACS is a massively parallel computer dedicated for calculations in computational physics and will be in operation in the spring of 1996 at Center for Computational Physics, University of Tsukuba. In this article, we describe the…

High Energy Physics - Lattice · Physics 2008-11-26 T. Yoshie

We present a parallel FFT algorithm for SIMD systems following the `Transpose Algorithm' approach. The method is based on the assignment of the data field onto a 1-dimensional ring of systolic cells. The systolic array can be universally…

High Energy Physics - Lattice · Physics 2015-06-25 Thomas Lippert , Klaus Schilling , Federico Toschi , Sven Trentmann , Raffaele Tripiccione

We consider the implementation of a parallel Monte Carlo code for high-performance simulations on PC clusters with MPI. We carry out tests of speedup and efficiency. The code is used for numerical simulations of pure SU(2) lattice gauge…

High Energy Physics - Lattice · Physics 2007-05-23 Attilio Cucchieri , Tereza Mendes , Gonzalo Travieso , Andre R. Taurines

This paper proposes a novel analytical framework, termed the Multiport Analytical Pixel Electromagnetic Simulator (MAPES). MAPES enables efficient and accurate prediction of the electromagnetic (EM) performance of arbitrary pixel-based…

Signal Processing · Electrical Eng. & Systems 2025-12-15 Junhui Rao , Yi Liu , Jichen Zhang , Zhaoyang Ming , Tianrui Qiao , Yujie Zhang , Chi Yuk Chiu , Hua Wang , Ross Murch

First results on the autocorrelation behaviour of a recently proposed fermion algorithm by M. L\"uscher are presented and discussed. The occurence of unexpected large autocorrelation times is explained. Possible improvements are discussed.

High Energy Physics - Lattice · Physics 2009-10-28 B. Jegerlehner

Approximate Bayes Computations (ABC) are used for parameter inference when the likelihood function of the model is expensive to evaluate but relatively cheap to sample from. In particle ABC, an ensemble of particles in the product space of…

Computation · Statistics 2016-04-15 Carlo Albert , Hans R. Kuensch , Andreas Scheidegger

We discuss the state of art of Lattice Boltzmann (LB) computing, with special focus on prospective LB schemes capable of meeting the forthcoming Exascale challenge. After reviewing the basic notions of LB computing, we discuss current…

Computational Physics · Physics 2020-06-14 Sauro Succi , Giorgio Amati , Massimo Bernaschi , Giacomo Falcucci , Marco Lauricella , Andrea Montessori

In view of the tremendous computing power jump of modern RISC processors the interest in parallel computing seems to be thinning out. Why use a complicated system of parallel processors, if the problem can be solved by a single powerful…

comp-gas · Physics 2008-02-03 G. Odor , F. Rohrbach , G. Vesztergombi , G. Varga , F. Tatrai

As modern analogue/mixed-signal design increasingly relies on optimization-in-the-loop flows, such as AI and LLM-based sizing agents that repeatedly invoke SPICE-efficient, accurate high-performance simulators have become an indispensable…

Hardware Architecture · Computer Science 2026-04-06 Xuanhao Bao , Danial Chitnis

Quantum computers, with parallel computing and entanglement effects, excel in cryptography analysis and big data processing. However, they are not fully developed yet, and their performance needs further evaluation. Traditional computer…

Quantum Physics · Physics 2024-09-11 Zili Chen