Distributed, Parallel, and Cluster Computing · Computer Science
Stencil Computations on Cerebras Wafer-Scale Engine
Elia Belli, Daniele De Sensi
2026-05-11
Distributed, Parallel, and Cluster Computing · Computer Science
TensorFlow as a DSL for stencil-based computation on the Cerebras Wafer Scale Engine
Nick Brown, Brandon Echols, Justs Zarins, Tobias Grosser
2022-10-11
Distributed, Parallel, and Cluster Computing · Computer Science
An MLIR Lowering Pipeline for Stencils at Wafer-Scale
Nicolai Stawinoga, David Katz, Anton Lydike, Justs Zarins +3
2026-01-27
Distributed, Parallel, and Cluster Computing · Computer Science
Efficient Algorithms for Monte Carlo Particle Transport on AI Accelerator Hardware
John Tramm, Bryce Allen, Kazutomo Yoshii, Andrew Siegel +1
2023-11-08
Distributed, Parallel, and Cluster Computing · Computer Science
Wafer-Scale Fast Fourier Transforms
Marcelo Orenes-Vera, Ilya Sharapov, Robert Schreiber, Mathias Jacquelin +2
2023-06-26
Distributed, Parallel, and Cluster Computing · Computer Science
Near-Optimal Wafer-Scale Reduce
Piotr Luczynski, Lukas Gianinazzi, Patrick Iff, Leighton Wilson +2
2024-09-04
Distributed, Parallel, and Cluster Computing · Computer Science
Disruptive Changes in Field Equation Modeling: A Simple Interface for Wafer Scale Engines
Mino Woo, Terry Jordan, Robert Schreiber, Ilya Sharapov +4
2022-09-30
Hardware Architecture · Computer Science
Record Acceleration of the Two-Dimensional Ising Model Using High-Performance Wafer Scale Engine
Dirk Van Essendelft, Hayl Almolyki, Wei Shi, Terry Jordan +2
2024-05-03
Neural and Evolutionary Computing · Computer Science
Demonstrating the Advantages of Analog Wafer-Scale Neuromorphic Hardware
Hartmut Schmidt, Andreas Grübl, José Montes, Eric Müller +2
2024-12-04
Neural and Evolutionary Computing · Computer Science
Trackable Island-model Genetic Algorithms at Wafer Scale
Matthew Andres Moreno, Connor Yang, Emily Dolson, Luis Zaman
2024-05-07
Distributed, Parallel, and Cluster Computing · Computer Science
Whale: Efficient Giant Model Training over Heterogeneous GPUs
Xianyan Jia, Le Jiang, Ang Wang, Wencong Xiao +8
2022-06-07
Distributed, Parallel, and Cluster Computing · Computer Science
Breaking the mold: overcoming the time constraints of molecular dynamics on general-purpose hardware
Danny Perez, Aidan Thompson, Stan Moore, Tomas Oppelstrup +10
2025-02-20
Hardware Architecture · Computer Science
Wafer-scale Computing: Advancements, Challenges, and Future Perspectives
Yang Hu, Xinhan Lin, Huizheng Wang, Zhen He +11
2023-10-17
Computational Physics · Physics
Breaking the Molecular Dynamics Timescale Barrier Using a Wafer-Scale System
Kylee Santos, Stan Moore, Tomas Oppelstrup, Amirali Sharifian +10
2024-12-30
Hardware Architecture · Computer Science
The xPU-athalon: Quantifying the Competition of AI Acceleration
Alicia Golden, Carole-Jean Wu, Gu-Yeon Wei, David Brooks
2026-04-14
Distributed, Parallel, and Cluster Computing · Computer Science
Fast Stencil-Code Computation on a Wafer-Scale Processor
Kamil Rocki, Dirk Van Essendelft, Ilya Sharapov, Robert Schreiber +6
2020-10-09
Emerging Technologies · Computer Science
DarwinWafer: A Wafer-Scale Neuromorphic Chip
Xiaolei Zhu, Xiaofei Jin, Ziyang Kang, Chonghui Sun +10
2025-09-23
Machine Learning · Computer Science
WaferLLM: Large Language Model Inference at Wafer Scale
Congjie He, Yeqi Huang, Pei Mu, Ziming Miao +4
2025-06-02
Emerging Technologies · Computer Science
From Clean Room to Machine Room: Commissioning of the First-Generation BrainScaleS Wafer-Scale Neuromorphic System
Hartmut Schmidt, José Montes, Andreas Grübl, Maurice Güttler +8
2023-03-23
Distributed, Parallel, and Cluster Computing · Computer Science
Insights into DeepSeek-V3: Scaling Challenges and Reflections on Hardware for AI Architectures
Chenggang Zhao, Chengqi Deng, Chong Ruan, Damai Dai +11
2025-12-24
Hardware Architecture · Computer Science
Sustainable AI Training via Hardware-Software Co-Design on NVIDIA, AMD, and Emerging GPU Architectures
Yashasvi Makin, Rahul Maliakkal
2025-08-20
Distributed, Parallel, and Cluster Computing · Computer Science
MoEntwine: Unleashing the Potential of Wafer-scale Chips for Large-scale Expert Parallel Inference
Xinru Tang, Jingxiang Hou, Dingcheng Jiang, Taiquan Wei +8
2025-10-30
Neural and Evolutionary Computing · Computer Science
BrainFuse: a unified infrastructure integrating realistic biological modeling and core AI methodology
Baiyu Chen, Yujie Wu, Siyuan Xu, Peng Qu +8
2026-01-30
Distributed, Parallel, and Cluster Computing · Computer Science
WarpSpeed: A High-Performance Library for Concurrent GPU Hash Tables
Hunter McCoy, Prashant Pandey
2025-10-24
Hardware Architecture · Computer Science
Addressing memory bandwidth scalability in vector processors for streaming applications
Jordi Altayo, Paul Delestrac, David Novo, Simey Yang +2
2025-05-20