English
Related papers

Related papers: Switch-Less Dragonfly on Wafers: A Scalable Interc…

200 papers

Transformer-based large language models are increasingly constrained by data movement as communication bandwidth drops sharply beyond the chip boundary. Wafer-scale integration using wafer-on-wafer hybrid bonding alleviates this limitation…

Hardware Architecture · Computer Science 2026-03-25 Patrick Iff , Tommaso Bonato , Maciej Besta , Luca Benini , Torsten Hoefler

Much like classical supercomputers, scaling up quantum computers requires an optical interconnect. However, signal attenuation leads to irreversible qubit loss, making quantum interconnect design guidelines and metrics different from…

Multichip systems with memory stacks and various processing chips are at the heart of platform based designs such as servers and embedded systems. Full utilization of the benefits of these integrated multichip systems need a seamless, and…

Hardware Architecture · Computer Science 2017-09-25 Md Shahriar Shamim , M Meraj Ahmed , Naseef Mansoor , Amlan Ganguly

Neuromorphic computing promises brain-like efficiency, yet today's multi-chip systems scale over PCBs and incur orders-of-magnitude penalties in bandwidth, latency, and energy, undermining biological algorithms and system efficiency. We…

Emerging Technologies · Computer Science 2025-09-23 Xiaolei Zhu , Xiaofei Jin , Ziyang Kang , Chonghui Sun , Junjie Feng , Dingwen Hu , Zengyi Wang , Hanyue Zhuang , Qian Zheng , Huajin Tang , Shi Gu , Xin Du , De Ma , Gang Pan

Dragonfly interconnect is a crucial network technology for supercomputers. To support exascale systems, network resources are shared such that links and routers are not dedicated to any node pair. While link utilization is increased,…

Networking and Internet Architecture · Computer Science 2024-04-05 Yao Kang , Xin Wang , Zhiling Lan

Distributed Deep Neural Network (DNN) training is a technique to reduce the training overhead by distributing the training tasks into multiple accelerators, according to a parallelization strategy. However, high-performance compute and…

Hardware Architecture · Computer Science 2025-06-10 Saeed Rashidi , William Won , Sudarshan Srinivasan , Puneet Gupta , Tushar Krishna

The need for wireless communication has driven the communication systems to high performance. However, the main bottleneck that affects the communication capability is the Fast Fourier Transform (FFT), which is the core of most modulators.…

Signal Processing · Electrical Eng. & Systems 2018-08-09 Rozita Teymourzadeh , Yazan Samir , Masuri Othman , Mok Vee Hong

Multi-plane architectures have become increasingly prevalent in the Fat-Tree networks of AI data centers. By leveraging multiple ports on a single network interface card (NIC) or multiple NICs within a scale-up domain, each port or NIC is…

Networking and Internet Architecture · Computer Science 2026-04-28 Ziyu Wang , Fei Lei , Dezun Dong

Hierarchical ring networks, which hierarchically connect multiple levels of rings, have been proposed in the past to improve the scalability of ring interconnects, but past hierarchical ring designs sacrifice some of the key benefits of…

Distributed, Parallel, and Cluster Computing · Computer Science 2016-02-22 Rachata Ausavarungnirun , Chris Fallin , Xiangyao Yu , Kevin Kai-Wei Chang , Greg Nazario , Reetuparna Das , Gabriel H. Loh , Onur Mutlu

There are increasing number of works addressing the design challenges of fast, scalable solutions for the growing number of new type of applications. Recently, many of the solutions aimed at improving processing element capabilities to…

Hardware Architecture · Computer Science 2019-12-16 Somnath Mazumdar , Alberto Scionti

Novel low-diameter network topologies such as Slim Fly (SF) offer significant cost and power advantages over the established Fat Tree, Clos, or Dragonfly. To spearhead the adoption of low-diameter networks, we design, implement, deploy, and…

To interconnect their growing number of servers, current supercomputers and data centers are starting to adopt low-diameter networks, such as HyperX, Dragonfly and Dragonfly+. These emergent topologies require balancing the load over their…

Hardware Architecture · Computer Science 2024-02-02 Alejandro Cano , Cristóbal Camarero , Carmen Martínez , Ramón Beivide

The Dragonfly topology is currently one of the most popular network topologies in high-performance parallel systems. The interconnection networks of many of these systems are built from components based on the InfiniBand specification.…

Networking and Internet Architecture · Computer Science 2025-02-04 German Maglione-Mathey , Jesus Escudero-Sahuquillo , Pedro Javier Garcia , Francisco J. Quiles , Eitan Zahavi

The interconnect is one of the most critical components in large scale computing systems, and its impact on the performance of applications is going to increase with the system size. In this paper, we will describe Slingshot, an…

Distributed, Parallel, and Cluster Computing · Computer Science 2022-10-28 Daniele De Sensi , Salvatore Di Girolamo , Kim H. McMahon , Duncan Roweth , Torsten Hoefler

Virtual channel flow control is the de facto choice for modern networks-on-chip to allow better utilization of the link bandwidth through buffering and packet switching, which are also the sources of large power footprint and long per-hop…

Hardware Architecture · Computer Science 2020-05-19 Yuan He , Jinyu Jiao , Thang Cao , Masaaki Kondo

Increasingly large AI workloads are calling for hyper-scale infrastructure; however, traditional interconnection network architecture is neither scalable nor cost-effective enough. Tree-based topologies such as the \textit{Rail-optimized}…

Hardware Architecture · Computer Science 2025-07-28 Yinxiao Feng , Tiancheng Chen , Yuchen Wei , Siyuan Shen , Shiju Wang , Wei Li , Kaisheng Ma , Torsten Hoefler

The main design principles in computer architecture have recently shifted from a monolithic scaling-driven approach to the development of heterogeneous architectures that tightly co-integrate multiple specialized processor and memory…

We introduce a high-performance cost-effective network topology called Slim Fly that approaches the theoretically optimal network diameter. Slim Fly is based on graphs that approximate the solution to the degree-diameter problem. We analyze…

Networking and Internet Architecture · Computer Science 2020-07-01 Maciej Besta , Torsten Hoefler

Conventional hybrid beamforming architectures are often compared with one another and with the fully-digital architecture under the same \emph{radiated} antenna power. However, the physically relevant budget is the power injected by the…

Signal Processing · Electrical Eng. & Systems 2026-03-19 Nikola Zlatanov , Damir Salakhov

In today's data centers, the performance of interconnects plays a pivotal role. However, many of the underlying technologies for these interconnects have a history of several decades and existed long before data centers came into being.To…

Networking and Internet Architecture · Computer Science 2023-11-03 Qianfeng Shen , Paul Chow
‹ Prev 1 2 3 10 Next ›