English
Related papers

Related papers: Status of the apeNEXT project

200 papers

Automatic Post-Editing (APE) aims to correct systematic errors in a machine translated text. This is primarily useful when the machine translation (MT) system is not accessible for improvement, leaving APE as a viable option to improve…

Computation and Language · Computer Science 2019-10-22 Rajen Chatterjee

The deployment of the next generation computing platform at ExaFlops scale requires to solve new technological challenges mainly related to the impressive number (up to 10^6) of compute elements required. This impacts on system power…

The rise of power-efficient embedded computers based on highly-parallel accelerators opens a number of opportunities and challenges for researchers and engineers, and paved the way to the era of edge computing. At the same time, advances in…

Distributed, Parallel, and Cluster Computing · Computer Science 2023-10-13 Paolo Burgio , Gianluca Brilli

We define some of the programming and system-level challenges facing the application of quantum processing to high-performance computing. Alongside barriers to physical integration, prominent differences in the execution of quantum and…

Quantum Physics · Physics 2017-12-06 Keith A. Britt , Fahd A. Mohiyaddin , Travis S. Humble

Power efficiency is becoming an ever more important metric for both high performance and high throughput computing. Over the course of next decade it is expected that flops/watt will be a major driver for the evolution of computer…

Computational Physics · Physics 2014-10-24 David Abdurachmanov , Peter Elmer , Giulio Eulisse , Shahzad Muzaffar

Future applications demand more performance, but technology advances have been faltering. A promising approach to further improve computer system performance under energy constraints is to employ hardware accelerators. Already today, mobile…

Distributed, Parallel, and Cluster Computing · Computer Science 2021-11-25 Mark D. Hill , Vijay Janapa Reddi

While the HPC community is working towards the development of the first Exaflop computer (expected around 2020), after reaching the Petaflop milestone in 2008 still only few HPC applications are able to fully exploit the capabilities of…

Distributed, Parallel, and Cluster Computing · Computer Science 2015-06-10 Erika Abraham , Costas Bekas , Ivona Brandic , Samir Genaim , Einar Broch Johnsen , Ivan Kondov , Sabri Pllana , Achim Streit

The matrix element (ME) calculation in any Monte Carlo physics event generator is an ideal fit for implementing data parallelism with lockstep processing on GPUs and vector CPUs. For complex physics processes where the ME calculation is the…

In this paper we describe the current status of the GRAPE-6 project to develop a special-purpose computer with a peak speed exceeding 100 Tflops for the simulation of astrophysical N-body problems. One of the main targets of the GRAPE-6…

Astrophysics · Physics 2007-05-23 Junichiro Makino

One of the outstanding challenges in contemporary science and technology is building a quantum computer that is useful in applications. By starting from an estimate of the algorithm success rate, we can explicitly connect gate fidelity to…

Quantum Physics · Physics 2026-03-20 R. Barends , F. K. Wilhelm

Turning the current experimental plasma accelerator state-of-the-art from a promising technology into mainstream scientific tools depends critically on high-performance, high-fidelity modeling of complex processes that develop over a wide…

Accelerator Physics · Physics 2018-12-26 J. -L. Vay , A. Almgren , J. Bell , L. Ge , D. P. Grote , M. Hogan , O. Kononenko , R. Lehe , A. Myers , C. Ng , J. Park , R. Ryne , O. Shapoval , M. Thevenet , W. Zhang

We report on recent improvements of the Atacama Pathfinder Experiment Control System (APECS) to cope with the ever increasing data rates and volumes. Also the very wide bandwidths of current instruments required switching to vectorized…

Instrumentation and Methods for Astrophysics · Physics 2021-12-16 Dirk Muders , Carsten König , Reinhold Schaaf , Felipe Mac-Auliffe , Juan-Pablo Pérez-Beaupuits

Parallel computers dedicated to lattice field theories are reviewed with emphasis on the three recent projects, the Teraflops project in the US, the CP-PACS project in Japan and the 0.5-Teraflops project in the US. Some new commercial…

High Energy Physics - Lattice · Physics 2009-10-22 Y. Iwasaki

The infinite projected entangled-pair state (iPEPS) ansatz is a powerful tensor-network approximation of an infinite two-dimensional quantum many-body state. Tensor-based calculations are particularly well-suited to utilize the high…

Strongly Correlated Electrons · Physics 2025-03-19 Addison D. S. Richards , Erik S. Sørensen

Agile is one of the terms with which software professionals are quite familiar. Agile models promote fast development to develop high quality software. XP process model is one of the most widely used and most documented agile models. XP…

Software Engineering · Computer Science 2014-08-28 M. Rizwan Jameel Qureshi

APEmille is a SIMD parallel processor under development at the Italian National Institute for Nuclear Physics (INFN). APEmille is very well suited for Lattice QCD applications, both for its hardware characteristics and for its software and…

High Energy Physics - Lattice · Physics 2009-10-28 Emanuele Panizzi

Mobile edge computing (MEC) is an emerging communication scheme that aims at reducing latency. In this paper, we investigate a green MEC system under the existence of an eavesdropper. We use computation efficiency, which is defined as the…

Networking and Internet Architecture · Computer Science 2020-04-14 Haijian Sun , Qun Wang , Xiang Ma , Yongjun Xu , Rose Qingyang Hu

Project X is a multi-megawatt proton facility being developed to support a world-leading program in Intensity Frontier physics at Fermilab. The facility will support programs in elementary particle and nuclear physics, with the potential…

Accelerator Physics · Physics 2014-09-23 S. D. Holmes , M. Kaducak , R. Kephart , I. Kourbanis , V. Lebedev , S. Mishra , S. Nagaitsev , N. Solyak , R. Tschirhart

Deploying large language models (LLMs) for online inference is often constrained by limited GPU memory, particularly due to the growing KV cache during auto-regressive decoding. Hybrid GPU-CPU execution has emerged as a promising solution…

Distributed, Parallel, and Cluster Computing · Computer Science 2026-01-16 Jiakun Fan , Yanglin Zhang , Xiangchen Li , Dimitrios S. Nikolopoulos

A considerable amount of research and engineering went into designing proxy applications, which represent common high-performance computing workloads, to co-design and evaluate the current generation of supercomputers, e.g., RIKEN's…

Distributed, Parallel, and Cluster Computing · Computer Science 2022-04-18 Satoshi Matsuoka , Jens Domke , Mohamed Wahib , Aleksandr Drozd , Ray Bair , Andrew A. Chien , Jeffrey S. Vetter , John Shalf