Distributed, Parallel, and Cluster Computing · Computer Science
Multi-GPU Graph Analytics
Yuechao Pan, Yangzihao Wang, Yuduo Wu, Carl Yang +1
2017-03-02
Machine Learning · Computer Science
Scalable Graph Embedding LearningOn A Single GPU
Azita Nouri, Philip E. Davis, Pradeep Subedi, Manish Parashar
2022-01-21
Hardware Architecture · Computer Science
GPU Domain Specialization via Composable On-Package Architecture
Yaosheng Fu, Evgeny Bolotin, Niladrish Chatterjee, David Nellans +1
2021-04-07
Distributed, Parallel, and Cluster Computing · Computer Science
Performance Analysis of Deep Learning Workloads on a Composable System
Kauotar El Maghraoui, Lorraine M. Herger, Chekuri Choudary, Kim Tran +2
2021-03-22
Hardware Architecture · Computer Science
AMOEBA: A Coarse Grained Reconfigurable Architecture for Dynamic GPU Scaling
Xianwei Cheng, Hui Zhao, Mahmut Kandemir, Beilei Jiang +1
2019-11-11
Distributed, Parallel, and Cluster Computing · Computer Science
A GPU-Accelerated Distributed Algorithm for Optimal Power Flow in Distribution Systems
Minseok Ryu, Geunyeong Byeon, Kibaek Kim
2025-01-15
Distributed, Parallel, and Cluster Computing · Computer Science
Beyond Desktop Computation: Challenges in Scaling a GPU Infrastructure
Martin Uray, Eduard Hirsch, Gerold Katzinger, Michael Gadermayr
2021-10-12
Distributed, Parallel, and Cluster Computing · Computer Science
Adaptive GPU Resource Allocation for Multi-Agent Collaborative Reasoning in Serverless Environments
Guilin Zhang, Wulan Guo, Ziqi Tan
2026-01-05
Distributed, Parallel, and Cluster Computing · Computer Science
Lightning: Scaling the GPU Programming Model Beyond a Single GPU
Stijn Heldens, Pieter Hijma, Ben van Werkhoven, Jason Maassen +1
2022-03-03
Distributed, Parallel, and Cluster Computing · Computer Science
Scalable GPU Performance Variability Analysis framework
Ankur Lahiry, Ayush Pokharel, Seth Ockerman, Amal Gueroudji +2
2025-06-27
Hardware Architecture · Computer Science
Evolution, Challenges, and Optimization in Computer Architecture: The Role of Reconfigurable Systems
Jefferson Ederhion, Festus Zindozin, Hillary Owusu, Chukwurimazu Ozoemezim +4
2024-12-30
Distributed, Parallel, and Cluster Computing · Computer Science
Optimizing Hardware Resource Partitioning and Job Allocations on Modern GPUs under Power Caps
Eishi Arima, Minjoon Kang, Issa Saba, Josef Weidendorfer +2
2024-05-08
Hardware Architecture · Computer Science
Scalable and Efficient Intra- and Inter-node Interconnection Networks for Post-Exascale Supercomputers and Data centers
Joaquin Tarraga-Moreno, Daniel Barley, Francisco J. Andujar Munoz, Jesus Escudero-Sahuquillo +4
2025-11-07
Distributed, Parallel, and Cluster Computing · Computer Science
Towards a Dynamic Composability Approach for using Heterogeneous Systems in Remote Sensing
Ilkay Altintas, Ismael Perez, Dmitry Mishin, Adrien Trouillaud +7
2022-11-15
Distributed, Parallel, and Cluster Computing · Computer Science
Overview of the IBM Neural Computer Architecture
Pritish Narayanan, Charles E. Cox, Alexis Asseman, Nicolas Antoine +3
2020-03-26
Distributed, Parallel, and Cluster Computing · Computer Science
Contention-Aware GPU Partitioning and Task-to-Partition Allocation for Real-Time Workloads
Houssam-Eddine Zahaf, Ignacio Sanudo Olmedo, Jayati Singh, Nicola Capodieci +1
2021-05-24
Mathematical Software · Computer Science
A Scalable and Modular Software Architecture for Finite Elements on Hierarchical Hybrid Grids
Nils Kohl, Dominik Thönnes, Daniel Drzisga, Dominik Bartuschat +1
2018-05-28
Distributed, Parallel, and Cluster Computing · Computer Science
Flex-MIG: Enabling Distributed Execution on MIG
Myeongsu Kim, Ikjun Yeom, Younghoon Kim
2025-11-14
Databases · Computer Science
Revisiting Query Performance in GPU Database Systems
Jiashen Cao, Rathijit Sen, Matteo Interlandi, Joy Arulraj +1
2023-02-03
Distributed, Parallel, and Cluster Computing · Computer Science
Hierarchical Resource Partitioning on Modern GPUs: A Reinforcement Learning Approach
Urvij Saroliya, Eishi Arima, Dai Liu, Martin Schulz
2024-05-15
Distributed, Parallel, and Cluster Computing · Computer Science
Overview of Swallow --- A Scalable 480-core System for Investigating the Performance and Energy Efficiency of Many-core Applications and Operating Systems
Simon J. Hollis, Steve Kerrison
2015-04-27