English
Related papers

Related papers: HEP Benchmark Suite: Enhancing Efficiency and Sust…

200 papers

System monitoring is an established tool to measure the utilization and health of HPC systems. Usually system monitoring infrastructures make no connection to job information and do not utilize hardware performance monitoring (HPM) data. To…

Distributed, Parallel, and Cluster Computing · Computer Science 2019-02-19 Thomas Röhl , Jan Eitzinger , Georg Hager , Gerhard Wellein

Facing an environment increasingly complex, uncertain and changing, even in crisis, organizations are driven to be agile in order to survive. Agility, at the core heart of business strategy, represents the ability to grow in a competitive…

Software Engineering · Computer Science 2021-09-16 Yassine Rdiouat , Samir Bahsani , Mouhsine Lakhdissi , Alami Semma

The selection, development, or comparison of machine learning methods in data mining can be a difficult task based on the target problem and goals of a particular study. Numerous publicly available real-world and simulated benchmark…

Machine Learning · Computer Science 2017-03-03 Randal S. Olson , William La Cava , Patryk Orzechowski , Ryan J. Urbanowicz , Jason H. Moore

Evaluative claims about LLM infrastructure -- ``workload X is fastest on hardware Y with software Z'' -- depend on a complex configuration space spanning hardware accelerators, interconnect bandwidth, software frameworks, parallelism plans,…

Distributed, Parallel, and Cluster Computing · Computer Science 2026-05-08 Eric Ding , Byungsoo Oh , Bhaskar Kataria , Kaiwen Guo , Jelena Gvero , Abhishek Vijaya Kumar , Arjun Devraj , Lindsey Bowen , Atharv Sonwane , Emaad Manzoor , Rachee Singh

The shift towards high-bandwidth networks driven by AI workloads in data centers and HPC clusters has unintentionally aggravated network latency, adversely affecting the performance of communication-intensive HPC applications. As…

Distributed, Parallel, and Cluster Computing · Computer Science 2024-04-23 Siyuan Shen , Langwen Huang , Marcin Chrapek , Timo Schneider , Jai Dayal , Manisha Gajbe , Robert Wisniewski , Torsten Hoefler

The convergence of IoT, Edge, Cloud, and HPC technologies creates a compute continuum that merges cloud scalability and flexibility with HPC's computational power and specialized optimizations. However, integrating cloud and HPC resources…

Distributed, Parallel, and Cluster Computing · Computer Science 2025-05-20 Aasish Kumar Sharma , Christian Boehme , Patrick Gelß , Ramin Yahyapour , Julian Kunkel

Common and community software packages, such as ROOT, Geant4 and event generators have been a key part of the LHC's success so far and continued development and optimisation will be critical in the future. The challenges are driven by an…

Computational Physics · Physics 2020-09-01 HEP Software Foundation , : , Thea Aarrestad , Simone Amoroso , Markus Julian Atkinson , Joshua Bendavid , Tommaso Boccali , Andrea Bocci , Andy Buckley , Matteo Cacciari , Paolo Calafiura , Philippe Canal , Federico Carminati , Taylor Childers , Vitaliano Ciulli , Gloria Corti , Davide Costanzo , Justin Gage Dezoort , Caterina Doglioni , Javier Mauricio Duarte , Agnieszka Dziurda , Peter Elmer , Markus Elsing , V. Daniel Elvira , Giulio Eulisse , Javier Fernandez Menendez , Conor Fitzpatrick , Rikkert Frederix , Stefano Frixione , Krzysztof L Genser , Andrei Gheata , Francesco Giuli , Vladimir V. Gligorov , Hadrien Benjamin Grasland , Heather Gray , Lindsey Gray , Alexander Grohsjean , Christian Gütschow , Stephan Hageboeck , Philip Coleman Harris , Benedikt Hegner , Lukas Heinrich , Burt Holzman , Walter Hopkins , Shih-Chieh Hsu , Stefan Höche , Philip James Ilten , Vladimir Ivantchenko , Chris Jones , Michel Jouvin , Teng Jian Khoo , Ivan Kisel , Kyle Knoepfel , Dmitri Konstantinov , Attila Krasznahorkay , Frank Krauss , Benjamin Edward Krikler , David Lange , Paul Laycock , Qiang Li , Kilian Lieret , Miaoyuan Liu , Vladimir Loncar , Leif Lönnblad , Fabio Maltoni , Michelangelo Mangano , Zachary Louis Marshall , Pere Mato , Olivier Mattelaer , Joshua Angus McFayden , Samuel Meehan , Alaettin Serhan Mete , Ben Morgan , Stephen Mrenna , Servesh Muralidharan , Ben Nachman , Mark S. Neubauer , Tobias Neumann , Jennifer Ngadiuba , Isobel Ojalvo , Kevin Pedro , Maurizio Perini , Danilo Piparo , Jim Pivarski , Simon Plätzer , Witold Pokorski , Adrian Alan Pol , Stefan Prestel , Alberto Ribon , Martin Ritter , Andrea Rizzi , Eduardo Rodrigues , Stefan Roiser , Holger Schulz , Markus Schulz , Marek Schönherr , Elizabeth Sexton-Kennedy , Frank Siegert , Andrzej Siódmok , Graeme A Stewart , Malik Sudhir , Sioni Paris Summers , Savannah Jennifer Thais , Nhan Viet Tran , Andrea Valassi , Marc Verderi , Dorothea Vom Bruch , Gordon T. Watts , Torre Wenaus , Efe Yazgan

Many tools and libraries employ hardware performance monitoring (HPM) on modern processors, and using this data for performance assessment and as a starting point for code optimizations is very popular. However, such data is only useful if…

Performance · Computer Science 2013-02-20 Jan Treibig , Georg Hager , Gerhard Wellein

As large language models (LLMs) continue to scale and new GPUs are released even more frequently, there is an increasing demand for LLM post-training in heterogeneous environments to fully leverage underutilized mid-range or…

Distributed, Parallel, and Cluster Computing · Computer Science 2026-04-14 Yongjun He , Shuai Zhang , Jiading Gai , Xiyuan Zhang , Boran Han , Bernie Wang , Huzefa Rangwala , George Karypis

The high-performance computing (HPC) community has adopted incentive structures to motivate reproducible research, with major conferences awarding badges to papers that meet reproducibility requirements. Yet, many papers do not meet such…

Distributed, Parallel, and Cluster Computing · Computer Science 2025-09-01 Valérie Hayot-Sasson , Nathaniel Hudson , André Bauer , Maxime Gonthier , Ian Foster , Kyle Chard

Can cloud computing infrastructures provide HPC-competitive performance for scientific applications broadly? Despite prolific related literature, this question remains open. Answers are crucial for designing future systems and democratizing…

Distributed, Parallel, and Cluster Computing · Computer Science 2021-03-09 Giulia Guidi , Marquita Ellis , Aydin Buluc , Katherine Yelick , David Culler

Simulating physical systems is a core component of scientific computing, encompassing a wide range of physical domains and applications. Recently, there has been a surge in data-driven methods to complement traditional numerical simulations…

Machine Learning · Computer Science 2021-08-19 Karl Otness , Arvi Gjoka , Joan Bruna , Daniele Panozzo , Benjamin Peherstorfer , Teseo Schneider , Denis Zorin

With the growing scale and complexity of high-performance computing (HPC) systems, resilience solutions that ensure continuity of service despite frequent errors and component failures must be methodically designed to balance the…

Distributed, Parallel, and Cluster Computing · Computer Science 2017-10-10 Saurabh Hukerikar , Christian Engelmann

Driven by artificial intelligence, data science, and high-resolution simulations, I/O workloads and hardware on high-performance computing (HPC) systems have become increasingly complex. This complexity can lead to large I/O overheads and…

Distributed, Parallel, and Cluster Computing · Computer Science 2025-01-03 Hammad Ather , Jean Luca Bez , Chen Wang , Hank Childs , Allen D. Malony , Suren Byna

The recent boom of big data, coupled with the challenges of its processing and storage gave rise to the development of distributed data processing and storage paradigms like MapReduce, Spark, and NoSQL databases. With the advent of cloud…

Distributed, Parallel, and Cluster Computing · Computer Science 2017-11-30 Sheriffo Ceesay , Adam Barker , Blesson Varghese

The SPEC Power benchmark offers valuable insights into the energy efficiency of server systems, allowing comparisons across various hardware and software configurations. Benchmark results are publicly available for hundreds of systems from…

Hardware Architecture · Computer Science 2024-11-12 Hannes Tröpgen , Robert Schöne , Thomas Ilsche , Daniel Hackenberg

Cloud systems have rapidly expanded worldwide in the last decade, shifting computational tasks to cloud servers where clients submit their requests. Among cloud workloads, latency-critical applications -- characterized by high-percentile…

Distributed, Parallel, and Cluster Computing · Computer Science 2025-05-07 Zhilin Li , Lucia Pons , Salvador Petit , Julio Sahuquillo , Julio Pons

Resource disaggregation is a promising technique for improving the efficiency of large-scale computing systems. However, this comes at the cost of increased memory access latency due to the need to rely on the network fabric to transfer…

The demand in computing power has never stopped growing over the years. Today, the performance of the most powerful systems exceeds the exascale and the number of petascale systems continues to grow. Unfortunately, this growth also goes…

Computers and Society · Computer Science 2024-03-27 Abdessalam Benhari , Denis Trystram , Fanny Dufossé , Yves Denneulin , Frédéric Desprez

The plethora of complex artificial intelligence (AI) algorithms and available high performance computing (HPC) power stimulates the expeditious development of AI components with heterogeneous designs. Consequently, the need for cross-stack…

Distributed, Parallel, and Cluster Computing · Computer Science 2021-03-16 Zhixiang Ren , Yongheng Liu , Tianhui Shi , Lei Xie , Yue Zhou , Jidong Zhai , Youhui Zhang , Yunquan Zhang , Wenguang Chen
‹ Prev 1 8 9 10 Next ›