English
Related papers

Related papers: Compact and High-Performance TCAM Based on Scaled …

200 papers

Ternary quantization has emerged as a powerful technique for reducing both computational and memory footprint of large language models (LLM), enabling efficient real-time inference deployment without significantly compromising model…

Hardware Architecture · Computer Science 2025-09-18 Zhirui Huang , Rui Ma , Shijie Cao , Ran Shu , Ian Wang , Ting Cao , Chixiao Chen , Yongqiang Xiong

The rapid growth of deep neural network (DNN) workloads has significantly increased the demand for large-capacity on-chip SRAM in machine learning (ML) applications, with SRAM arrays now occupying a substantial fraction of the total die…

Hardware Architecture · Computer Science 2025-12-30 Subhradip Chakraborty , Ankur Singh , Xuming Chen , Gourav Datta , Akhilesh R. Jaiswal

By providing highly efficient one-sided communication with globally shared memory space, Partitioned Global Address Space (PGAS) has become one of the most promising parallel computing models in high-performance computing (HPC). Meanwhile,…

Distributed, Parallel, and Cluster Computing · Computer Science 2022-07-12 Yashael Faith Arthanto , David Ojika , Joo-Young Kim

Transformer neural networks achieve state-of-the-art accuracy across language and vision tasks, but their deployment on embedded hardware is hindered by stringent area, latency, and energy constraints. During inference, performance and…

Hardware Architecture · Computer Science 2026-04-09 Jan Klhufek , Alberto Marchisio , Vojtech Mrazek , Lukas Sekanina , Muhammad Shafique

Gating is a key technique used for integrating information from multiple sources by long short-term memory (LSTM) models and has recently also been applied to other models such as the highway network. Although gating is powerful, it is…

Computation and Language · Computer Science 2018-06-19 Chao Zhang , Philip Woodland

Rapid single-flux quantum (RSFQ), a leading cryogenic superconductive electronics (SCE) technology, offers extremely low power dissipation and high speed. However, implementing RSFQ systems at VLSI complexity faces challenges, such as…

Emerging Technologies · Computer Science 2024-03-12 Rassul Bairamkulov , Mingfei Yu , Giovanni De Micheli

Cyber-security for 5G networks is drawing notable attention due to an increase in complex jamming attacks that could target the critical 5G Radio Frequency (RF) domain. These attacks pose a significant risk to heterogeneous network (HetNet)…

Cryptography and Security · Computer Science 2025-09-30 Samhita Kuili , Mohammadreza Amini , Burak Kantarci

In this letter, we demonstrate a non-volatile memory device in a graphene FET structure using ferroelectric gating. The binary information, i.e. "1" and "0", is represented by the high and low resistance states of the graphene working…

Mesoscale and Nanoscale Physics · Physics 2009-04-23 Yi Zheng , Guang-Xin Ni , Chee-Tat Toh , Ming-Gang Zeng , Shu-Ting Chen , Kui Yao , Barbaros Ozyilmaz

Thickness engineered tunneling field-effect transistors (TE-TFET) as a high performance ultra-scaled steep transistor is proposed. This device exploits a specific property of 2D materials: layer thickness dependent energy bandgap (Eg).…

Mesoscale and Nanoscale Physics · Physics 2017-03-08 Fan W. Chen , Hesameddin Ilatikhameneh , Tarek A. Ameen , Gerhard Klimeck , Rajib Rahman

The continuous shift of computational bottlenecks to the memory access and data transfer, especially for AI applications, poses the urgent needs of re-engineering the computer architecture fundamentals. Many edge computing applications,…

Systems and Control · Electrical Eng. & Systems 2025-01-31 Georgios Papandroulidakis , Shady Agwa , Ahmet Cirakoglu , Themis Prodromakis

The rapid adoption of low-precision arithmetic in artificial intelligence and edge computing has created a strong demand for energy-efficient and flexible floating-point multiply-accumulate (MAC) units. This paper presents a dual-precision…

Hardware Architecture · Computer Science 2026-04-10 Shubham Kumar , Vijay Pratap Sharma , Vaibhav Neema , Santosh Kumar Vishvakarma

This work demonstrates a novel energy-efficient tunnel FET (TFET)-CMOS hybrid foundry platform for ultralow-power AIoT applications. By utilizing the proposed monolithic integration process, the novel complementary n and p-type Si TFET…

In this letter, we quantify the impact of device limitations on the classification accuracy of an artificial neural network, where the synaptic weights are implemented in a Ferroelectric FET (FeFET) based in-memory processing architecture.…

Emerging Technologies · Computer Science 2019-08-22 Insik Yoon , Matthew Jerry , Suman Datta , Arijit Raychowdhury

InGaZnO (IGZO) channel FeFETs have attracted notable interest thanks to their advances in endurance. This work evaluates the viability of NOR-type IGZO FeFETs for readcentric AI inference workloads via design-technology cooptimization…

Fast Fourier Transform (FFT) is an essential tool in scientific and engineering computation. The increasing demand for mixed-precision FFT has made it possible to utilize half-precision floating-point (FP16) arithmetic for faster speed and…

Distributed, Parallel, and Cluster Computing · Computer Science 2021-04-26 Binrui Li , Shenggan Cheng , James Lin

The exponential emergence of Field Programmable Gate Array (FPGA) has accelerated the research of hardware implementation of Deep Neural Network (DNN). Among all DNN processors, domain specific architectures, such as, Google's Tensor…

Hardware Architecture · Computer Science 2022-02-15 Rourab Paul , Sreetama Sarkar , Suman Sau , Koushik Chakraborty , Sanghamitra Roy , Amlan Chakrabarti

The Bicameral Cache is a cache organization proposal for a vector architecture that segregates data according to their access type, distinguishing scalar from vector references. Its aim is to avoid both types of references from interfering…

Hardware Architecture · Computer Science 2025-03-28 Susana Rebolledo , Borja Perez , Jose Luis Bosque , Peter Hsu

This paper analyzes the performance and energy efficiency of Netcast, a recently proposed optical neural-network architecture designed for edge computing. Netcast performs deep neural network inference by dividing the computational task…

Emerging Technologies · Computer Science 2022-07-06 Ryan Hamerly , Alexander Sludds , Saumil Bandyopadhyay , Zaijun Chen , Zhizhen Zhong , Liane Bernstein , Dirk Englund

We present a study based on numerical simulations and comparative analysis of recent experimental data concerning the operation and design of FeFETs. Our results show that a proper consideration of charge trapping in the…

Emerging Technologies · Computer Science 2022-09-13 Daniel Lizzit , David Esseni

In this paper we present a comprehensive design and benchmarking study of Content Addressable Memory (CAM) at the 7nm technology node in the context of similarity search applications. We design CAM cells based on SRAM, spin-orbit torque,…

Emerging Technologies · Computer Science 2025-01-07 Siri Narla , Piyush Kumar , Mohammad Adnaan , Azad Naeemi
‹ Prev 1 4 5 6 7 8 10 Next ›