Hardware Architecture · Computer Science
Evaluating Modern GPU Interconnect: PCIe, NVLink, NV-SLI, NVSwitch and GPUDirect
Ang Li, Shuaiwen Leon Song, Jieyang Chen, Jiajia Li +3
2019-08-26
Distributed, Parallel, and Cluster Computing · Computer Science
To Use or Not to Use: CPUs' Cache Optimization Techniques on GPGPUs
Vajira Thambawita, Roshan G. Ragel, Dhammike Elkaduwe
2018-10-10
Distributed, Parallel, and Cluster Computing · Computer Science
GPUs as Storage System Accelerators
Samer Al-Kiswany, Abdullah Gharaibeh, Matei Ripeanu
2016-11-18
Distributed, Parallel, and Cluster Computing · Computer Science
Low Overhead Instruction Latency Characterization for NVIDIA GPGPUs
Yehia Arafa, Abdel-Hameed Badawy, Gopinath Chennupati, Nandakishore Santhi +1
2019-09-04
Distributed, Parallel, and Cluster Computing · Computer Science
Analyzing the HCP Datasets using GPUs: The Anatomy of a Science Engagement
John-Paul Robinson, Thomas Anthony, Ravi Tripathi, Sara A. Sims +2
2019-09-10
Distributed, Parallel, and Cluster Computing · Computer Science
Performance Analysis and Efficient Execution on Systems with multi-core CPUs, GPUs and MICs
George Teodoro, Tahsin Kurc, Guilherme Andrade, Jun Kong +2
2015-05-15
Distributed, Parallel, and Cluster Computing · Computer Science
The Landscape of GPU-Centric Communication
Didem Unat, Ilyas Turimbetov, Mohammed Kefah Taha Issa, Doğan Sağbili +3
2026-04-24
Distributed, Parallel, and Cluster Computing · Computer Science
Understanding Data Movement in AMD Multi-GPU Systems with Infinity Fabric
Gabin Schieffer, Ruimin Shi, Stefano Markidis, Andreas Herten +2
2024-10-02
Distributed, Parallel, and Cluster Computing · Computer Science
Apple Silicon Performance in Scientific Computing
Connor Kenyon, Collin Capano
2022-11-03
Distributed, Parallel, and Cluster Computing · Computer Science
Comparative Performance Analysis of Intel Xeon Phi, GPU, and CPU
George Teodoro, Tahsin Kurc, Jun Kong, Lee Cooper +1
2013-11-05
Distributed, Parallel, and Cluster Computing · Computer Science
The anachronism of whole-GPU accounting
Igor Sfiligoi, David Schultz, Frank Würthwein, Benedikt Riedel +1
2022-07-12
Distributed, Parallel, and Cluster Computing · Computer Science
10 Years Later: Cloud Computing is Closing the Performance Gap
Giulia Guidi, Marquita Ellis, Aydin Buluc, Katherine Yelick +1
2021-03-09
Distributed, Parallel, and Cluster Computing · Computer Science
Taking GPU Programming Models to Task for Performance Portability
Joshua H. Davis, Pranav Sivaraman, Joy Kitson, Konstantinos Parasyris +4
2025-09-08
Databases · Computer Science
Vortex: Overcoming Memory Capacity Limitations in GPU-Accelerated Large-Scale Data Analytics
Yichao Yuan, Advait Iyer, Lin Ma, Nishil Talati
2025-02-14
Distributed, Parallel, and Cluster Computing · Computer Science
Scalable GPU Performance Variability Analysis framework
Ankur Lahiry, Ayush Pokharel, Seth Ockerman, Amal Gueroudji +2
2025-06-27
Distributed, Parallel, and Cluster Computing · Computer Science
Preparing for Performance Analysis at Exascale
Jonathon Anderson, Yumeng Liu, John Mellor-Crummey
2022-03-11
Distributed, Parallel, and Cluster Computing · Computer Science
Not All GPUs Are Created Equal: Characterizing Variability in Large-Scale, Accelerator-Rich Systems
Prasoon Sinha, Akhil Guliani, Rutwik Jain, Brandon Tran +2
2022-11-10
Distributed, Parallel, and Cluster Computing · Computer Science
Intra-node Memory Safe GPU Co-Scheduling
Carlos Reano, Federico Silla, Dimitrios S. Nikolopoulos, Blesson Varghese
2017-12-14
Distributed, Parallel, and Cluster Computing · Computer Science
Dissecting the NVIDIA Blackwell Architecture with Microbenchmarks
Aaron Jarmusch, Nathan Graddon, Sunita Chandrasekaran
2025-07-23