English
Related papers

Related papers: A Non-Volatile All-Spin Non-Binary Matrix Multipli…

200 papers

Large-scale artificial neural networks have shown significant promise in addressing a wide range of classification and recognition applications. However, their large computational requirements stretch the capabilities of computing…

Neural and Evolutionary Computing · Computer Science 2017-11-13 Syed Shakib Sarwar , Swagath Venkataramani , Anand Raghunathan , Kaushik Roy

On-chip implementation of optical nonlinear activation functions (NAFs) is essential for realizing large-scale photonic neural chips. To implement different neural processing and machine learning tasks with optimal performances, different…

Multiplication is a fundamental operation in many applications, and multipliers are widely adopted in various circuits. However, optimizing multipliers is challenging due to the extensive design space. In this paper, we propose a multiplier…

Hardware Architecture · Computer Science 2024-12-30 Dongsheng Zuo , Jiadong Zhu , Yikang Ouyang , Yuzhe Ma

Probabilistic machine learning enabled by the Bayesian formulation has recently gained significant attention in the domain of automated reasoning and decision-making. While impressive strides have been made recently to scale up the…

Emerging Technologies · Computer Science 2020-04-22 Kezhou Yang , Akul Malhotra , Sen Lu , Abhronil Sengupta

Fast combinational multipliers with large bit widths can occupy significant silicon area, which also drives up power consumption. Area can be reduced through resource sharing (i.e., folding) at the expense of lower throughput, which is…

Hardware Architecture · Computer Science 2025-09-03 Ahmad Houraniah , H. Fatih Ugurdag , C. Emre Dedeagac

Artificial intelligence workloads, especially transformer models, exhibit emergent sparsity in which computations perform selective sparse access to dense data. The workloads are inefficient on hardware designed for dense computations and…

Data Structures and Algorithms · Computer Science 2024-02-23 Brian Wheatman , Meghana Madhyastha , Randal Burns

Deep Convolutional Neural Networks (CNNs) have become state-of-the art for computer vision and other signal processing tasks due to their superior accuracy. In recent years, large efforts have been made to reduce the computational costs of…

Hardware Architecture · Computer Science 2021-04-13 Mario Fischer , Juergen Wassner

Linear-scaling electronic-structure techniques, also called O(N) techniques, rely heavily on the multiplication of sparse matrices, where the sparsity arises from spatial cut-offs. In order to treat very large systems, the calculations must…

Materials Science · Physics 2009-10-31 D. R. Bowler , T. Miyazaki , M. J. Gillan

In-memory associative processor architectures are offered as a great candidate to overcome memory-wall bottleneck and to enable vector/parallel arithmetic operations. In this paper, we extend the functionality of the associative processor…

Hardware Architecture · Computer Science 2021-10-20 Mira Hout , Mohammed E. Fouda , Rouwaida Kanj , Ahmed M. Eltawil

Specialized computational units that perform small matrix multiplications as primitive operations are typically present in modern AI accelerators. However, these Matrix Multiplication Units (MMUs) are often underutilized for many…

Data Structures and Algorithms · Computer Science 2025-09-25 Aleksandros Sobczyk , Giuseppe Sorrentino , Anastasios Zouzias

The growth of connected intelligent devices in the Internet of Things has created a pressing need for real-time processing and understanding of large volumes of analogue data. The difficulty in boosting the computing speed renders digital…

The memristor is promising to be the basic cell of next-generation computation systems. Compared to the traditional MOSFET device, the memristor is efficient over energy and area. But one of the biggest challenges faced with researchers is…

Emerging Technologies · Computer Science 2016-11-22 Junyi Li , Fulin Peng , Fan Yang , Xuan Zeng

Recent artificial neural network architectures improve performance and power dissipation by leveraging resistive devices to store and multiply synaptic weights with input data. Negative and positive synaptic weights are stored on the…

Emerging Technologies · Computer Science 2019-07-29 Krishna Prasad Gnawali , Seyed Nima Mozaffari , Spyros Tragoudas

Naturally random devices that exploit ambient thermal noise have recently attracted attention as hardware primitives for accelerating probabilistic computing applications. One such approach is to use a low barrier nanomagnet as the free…

Mesoscale and Nanoscale Physics · Physics 2021-05-05 Kerem Y. Camsari , Mustafa Mert Torunbalci , William A. Borders , Hideo Ohno , Shunsuke Fukami

This work presents a method to maximize power-efficiency of fixed point multiplier units by decomposing them into sub-components. First, an encoder block converts the operands from a two's complement to a sign magnitude representation,…

Neural and Evolutionary Computing · Computer Science 2025-07-25 Felix Arnold , Maxence Bouvier , Ryan Amaudruz , Renzo Andri , Lukas Cavigelli

Non-volatile memristors offer a salient platform for artificial neural network (ANN), but the integration of different function blocks into one hardware system remains challenging. Here we demonstrate the implementation of brain-like…

Mesoscale and Nanoscale Physics · Physics 2023-05-22 Puyang Huang , Xinqi Liu , Yue Xin , Yu Gu , Albert Lee , Zhuo Xu , Peng Chen , Yu Zhang , Weijie Deng , Guoqiang Yu , Zhongkai Liu , Qi Yao , Yumeng Yang , Zhifeng Zhu , Xufeng Kou

This paper introduces a novel oscillator that combines the tunability of spin Hall-driven nano oscillators with the high quality factor (Q) of high overtone bulk acoustic wave resonators (HBAR), integrating both reference and tunable…

Mesoscale and Nanoscale Physics · Physics 2018-01-22 Mustafa Mert Torunbalci , Tanay A. Gosavi , Kerem Y. Camsari , Sunil A. Bhave

In this paper, we propose a general framework to accelerate significantly the algorithms for nonnegative matrix factorization (NMF). This framework is inspired from the extrapolation scheme used to accelerate gradient methods in convex…

Numerical Analysis · Computer Science 2020-01-14 Andersen Man Shun Ang , Nicolas Gillis

Two promising strategies for achieving efficient control of magnetization in future magnetic memory and non-volatile spin logic devices are spin transfer torque from spin polarized currents and voltage-controlled magnetic anisotropy (VCMA).…

Materials Science · Physics 2012-09-06 Luqiao Liu , Chi-Feng Pai , D. C. Ralph , R. A. Buhrman

Matrix multiplication is the foundation from much of the success from high performance technologies like deep learning, scientific simulations, and video graphics. High level programming languages like Python and R rely on highly optimized…

Performance · Computer Science 2025-09-08 Ethan Davis
‹ Prev 1 4 5 6 7 8 10 Next ›