English
Related papers

Related papers: A Low-Dissipation and Scalable GEMM Accelerator wi…

200 papers

Silicon nitride (SiN) photonics platform has attributes of ultra-low linear and nonlinear propagation losses and CMOS-compatible fabrication process, promising large-scale multifunctional photonic circuits. However, the centrosymmetric…

Optics · Physics 2022-06-08 Binbin Wang , Yafei Ji , Linpeng Gu , Liang Fang , Xuetao Gan , Jianlin Zhao

The use of the Silicon-on-Insulator (SOI) platform has been prominent for realizing CMOS-compatible, high-performance photonic integrated circuits (PICs). But in recent years, the silicon-nitride-on-silicon-dioxide (SiN-on-SiO$_2$) platform…

Silicon nitride (SiN) has emerged as a promising platform for integrated nonlinear photonics because of its low propagation loss, wide transparency window, and CMOS compatibility. Nonlinear processes arising from photon-electron…

The use of the Silicon-on-Insulator (SOI) platform has been prominent for realizing CMOS-compatible, high-performance photonic integrated circuits (PICs). But in recent years, the silicon-nitride-on-silicon-dioxide (SiN-on-SiO$_2$) platform…

Several microring resonator (MRR) based analog photonic architectures have been proposed to accelerate general matrix-matrix multiplications (GEMMs) in deep neural networks with exceptional throughput and energy efficiency. To implement…

Hardware Architecture · Computer Science 2024-02-20 Sairam Sri Vatsavai , Venkata Sai Praneeth Karempudi , Oluwaseun Adewunmi Alo , Ishan Thakkar

Photonic integrated circuits are emerging as a promising platform for accelerating matrix multiplications in deep learning, leveraging the inherent parallel nature of light. Although various schemes have been proposed and demonstrated to…

Emerging Technologies · Computer Science 2024-06-04 Rui Tang , Shuhei Ohno , Ken Tanizawa , Kazuhiro Ikeda , Makoto Okano , Kasidit Toprasertpong , Shinichi Takagi , Mitsuru Takenaka

Deep Neural Networks (DNNs) predominantly rely on General Matrix Multiply (GEMM) kernels, which are often accelerated using specialized hardware architectures. Recently, analog photonic GEMM accelerators have emerged as a promising…

Hardware Architecture · Computer Science 2024-07-09 Oluwaseun Adewunmi Alo , Sairam Sri Vatsavai , Ishan Thakkar

A system-on-chip (SoC) photonic-electronic linear-algebra accelerator with the features of wavelength-division-multiplexing (WDM) based broadband photodetections and high-dimensional matrix-inversion operations fabricated in advanced…

Systems and Control · Electrical Eng. & Systems 2024-02-14 Tzu-Chien Hsueh , Yeshaiahu Fainman , Bill Lin

Sparse neural networks can greatly facilitate the deployment of neural networks on resource-constrained platforms as they offer compact model sizes while retaining inference accuracy. Because of the sparsity in parameter matrices, sparse…

Machine Learning · Computer Science 2021-09-10 Febin Sunny , Mahdi Nikdast , Sudeep Pasricha

In-memory computing (IMC) on a monolithic chip for deep learning faces dramatic challenges on area, yield, and on-chip interconnection cost due to the ever-increasing model sizes. 2.5D integration or chiplet-based architectures interconnect…

Machine Learning · Computer Science 2021-08-23 Gokul Krishnan , Sumit K. Mandal , Manvitha Pannala , Chaitali Chakrabarti , Jae-sun Seo , Umit Y. Ogras , Yu Cao

The number of parameters in deep neural networks (DNNs) is scaling at about 5$\times$ the rate of Moore's Law. To sustain this growth, photonic computing is a promising avenue, as it enables higher throughput in dominant general…

The computing wall and data movement challenges of deep neural networks (DNNs) have exposed the limitations of conventional CMOS-based DNN accelerators. Furthermore, the deep structure and large model size will make DNNs prohibitive to…

Signal Processing · Electrical Eng. & Systems 2019-12-12 Geng Yuan , Xiaolong Ma , Sheng Lin , Zhengang Li , Caiwen Ding

Sparse Ternary General Matrix-Matrix Multiplication (GEMM) remains under-optimized in existing libraries for Apple Silicon CPUs. We present a Sparse Ternary GEMM kernel optimized specifically for Apple's M-series processors. We propose a…

Performance · Computer Science 2025-10-15 Baraq Lipshitz , Alessio Melone , Charalampos Maraziaris , Muhammed Bilal

Current Artificial Intelligence (AI) computation systems face challenges, primarily from the memory-wall issue, limiting overall system-level performance, especially for Edge devices with constrained battery budgets, such as smartphones,…

Hardware Architecture · Computer Science 2024-10-15 Lucas Huijbregts , Liu Hsiao-Hsuan , Paul Detterer , Said Hamdioui , Amirreza Yousefzadeh , Rajendra Bishnoi

General matrix multiplication (GeMM) is a core operation in virtually all AI applications. Systolic array (SA) based architectures have shown great promise as GeMM hardware accelerators thanks to their speed and energy efficiency.…

Hardware Architecture · Computer Science 2025-01-13 Md Mizanur Rahaman Nayan , Ritik Raj , Gouse Basha Shaik , Tushar Krishna , Azad J Naeemi

There is a growing interest in custom spatial accelerators for machine learning applications. These accelerators employ a spatial array of processing elements (PEs) interacting via custom buffer hierarchies and networks-on-chip. The…

Distributed, Parallel, and Cluster Computing · Computer Science 2021-06-22 Gordon E. Moon , Hyoukjun Kwon , Geonhwa Jeong , Prasanth Chatarasi , Sivasankaran Rajamanickam , Tushar Krishna

Silicon nitride (SiN) waveguides with ultra-low optical loss enable integrated photonic applications including low noise, narrow linewidth lasers, chip-scale nonlinear photonics, and microwave photonics. Lasers are key components to SiN…

Broadband active materials are pivotal for advancing emerging technologies spanning on-chip optical interconnects, artificial intelligence, quantum systems and precision metrology. Current semiconductor gain media face bandwidth…

Optics · Physics 2026-02-09 Ken Liu

The currently dominant AI/ML workloads, such as Large Language Models (LLMs), rely on the efficient execution of General Matrix-Matrix Multiplication (GEMM) operations. Thus, most systems are equipped with dedicated matrix hardware…

Hardware Architecture · Computer Science 2026-04-01 Luigi Altamura , Alessio Cicero , Mateo Vázquez Maceiras , Mohammad Ali Maleki , Pedro Trancoso

Monolithic integration of solid-state color centers with photonic elements of the same material is a promising approach to overcome the constraints of fabrication complexity and coupling losses in traditional hybrid integration approaches.…

‹ Prev 1 2 3 10 Next ›