English
Related papers

Related papers: Extreme Scale-out SuperMUC Phase 2 - lessons learn…

200 papers

Modern large language model workloads put increasing demands on parallel compute capability and on-chip memory capacity, while also stressing fine-grained data movement and synchronization. These trends motivate exploring and designing…

Hardware Architecture · Computer Science 2026-05-11 Yinrong Li , Zexin Fu , Yichao Zhang , Germain Haugou , Chi Zhang , Marco Bertuletti , Bowen Wang , Luca Benini

New challenges in Astronomy and Astrophysics (AA) are urging the need for a large number of exceptionally computationally intensive simulations. "Exascale" (and beyond) computational facilities are mandatory to address the size of…

ATLAS, a general-purpose experiment at the Large Hadron Collider (LHC), makes use of a large internationally-distributed computing infrastructure, including over $10^6$ TB of managed data on disk and tape and almost one million…

High Energy Physics - Experiment · Physics 2026-04-16 ATLAS Collaboration

We present the design, implementation and engineering experience in building and deploying MegaScale, a production system for training large language models (LLMs) at the scale of more than 10,000 GPUs. Training LLMs at this scale brings…

The ATLAS and CMS experiments at the Large Hadron Collider (LHC) at CERN have searched for signals of new physics, in particular for supersymmetry. The data collected until 2012 at center-of-mass energies of 7 and 8 TeV and integrated…

High Energy Physics - Experiment · Physics 2016-09-07 Christian Autermann

A memory leak in an application deployed on the cloud can affect the availability and reliability of the application. Therefore, identifying and ultimately resolve it quickly is highly important. However, in the production environment…

Distributed, Parallel, and Cluster Computing · Computer Science 2021-06-17 Anshul Jindal , Paul Staab , Pooja Kulkarni , Jorge Cardoso , Michael Gerndt , Vladimir Podolskiy

We report on the 3-year INFN ATLAS-CMS joint research activity in collaboration with FBK, started in 2014, and aimed at the development of new thin pixel detectors for the High Luminosity LHC Phase-2 upgrades. The program is concerned with…

Memory-bound algorithms show complex performance and energy consumption behavior on multicore processors. We choose the lattice-Boltzmann method (LBM) on an Intel Sandy Bridge cluster as a prototype scenario to investigate if and how…

Performance · Computer Science 2015-05-25 Markus Wittmann , Georg Hager , Thomas Zeiser , Jan Treibig , Gerhard Wellein

The CERN LHC provided proton and heavy ion collisions during its Run 2 operation period from 2015 to 2018. Proton-proton collisions reached a peak instantaneous luminosity of 2.1 $\times$ 10$^{34}$ cm$^{-2}$s$^{-1}$, twice the initial…

Instrumentation and Detectors · Physics 2024-11-26 CMS Collaboration

The energy consumption of an exascale High-Performance Computing (HPC) supercomputer rivals that of tens of thousands of people in terms of electricity demand. Given the substantial energy footprint of exascale HPC systems and the…

Optimization and Control · Mathematics 2024-04-05 Luc Angelelli , Danilo Carastan-Santos , Pierre-François Dutot

Data scouting, introduced by CMS in 2011, is the use of specialized data streams based on reduced event content, enabling LHC experiments to record unprecedented numbers of proton-proton collision events that would otherwise be rejected by…

High Energy Physics - Experiment · Physics 2018-08-03 Javier Duarte

A high-performance implementation of a multiphase lattice Boltzmann method based on the conservative Allen-Cahn model supporting high-density ratios and high Reynolds numbers is presented. Metaprogramming techniques are used to generate…

Fluid Dynamics · Physics 2020-12-14 Markus Holzer , Martin Bauer , Ulrich Rüde

This paper presents the first, 15-PetaFLOP Deep Learning system for solving scientific pattern classification problems on contemporary HPC architectures. We develop supervised convolutional architectures for discriminating signals in…

After successfully completing Phase I upgrades during LHC Long Shutdown 2, the ATLAS detector is back in operation with several upgrades implemented. The most important and challenging upgrade is in the Muon Spectrometer, where the two…

High Energy Physics - Experiment · Physics 2025-03-24 Simone Francescato

Phase-change materials (PCMs) are increasingly recognised as promising platforms for tunable photonic devices due to their ability to modulate optical properties through solid-state phase transitions. Ultrathin and low-loss PCMs are highly…

We present the simulation tools developed to aid the design phase of the Cassegrain U-Band Efficient Spectrograph (CUBES) for the Very Large Telescope (VLT), exploring aspects of the system design and evaluating the performance for…

Turning the current experimental plasma accelerator state-of-the-art from a promising technology into mainstream scientific tools depends critically on high-performance, high-fidelity modeling of complex processes that develop over a wide…

Accelerator Physics · Physics 2018-12-26 J. -L. Vay , A. Almgren , J. Bell , L. Ge , D. P. Grote , M. Hogan , O. Kononenko , R. Lehe , A. Myers , C. Ng , J. Park , R. Ryne , O. Shapoval , M. Thevenet , W. Zhang

The DZERO experiment located at Fermilab has recently started RunII with an upgraded detector. The RunII physics program requires the Data Acquisition to readout the detector at a rate of 1 KHz. Events fragments, totaling 250 KB, are…

Physical phenomena such as chemical reactions, bond breaking, and phase transition require molecular dynamics (MD) simulation with ab initio accuracy ranging from milliseconds to microseconds. However, previous state-of-the-art neural…

Distributed, Parallel, and Cluster Computing · Computer Science 2025-04-22 Jianxiong Li , Boyang Li , Zhuoqiang Guo , Mingzhen Li , Enji Li , Lijun Liu , Guojun Yuan , Zhan Wang , Guangming Tan , Weile Jia

This article reports on first results of the KONWIHR-II project OMI4papps at the Leibniz Supercomputing Centre (LRZ). The first part describes Apex-MAP, a tunable synthetic benchmark designed to simulate the performance of typical…

Performance · Computer Science 2010-02-01 Volker Weinberg , Matthias Brehm , Iris Christadler