English
Related papers

Related papers: Ternary Weight Networks

200 papers

Binary Neural Networks (BNNs) enable efficient deep learning by saving on storage and computational costs. However, as the size of neural networks continues to grow, meeting computational requirements remains a challenge. In this work, we…

Machine Learning · Computer Science 2024-07-18 Matt Gorbett , Hossein Shirazi , Indrakshi Ray

Convolutional neural networks (CNNs) have been used in many machine learning fields. In practical applications, the computational cost of convolutional neural networks is often high with the deepening of the network and the growth of data…

Computer Vision and Pattern Recognition · Computer Science 2021-07-08 Shiqing Fan , Liu Liying , Ye Luo

Deep neural networks have achieved impressive results in computer vision and machine learning. Unfortunately, state-of-the-art networks are extremely compute and memory intensive which makes them unsuitable for mW-devices such as IoT…

Distributed, Parallel, and Cluster Computing · Computer Science 2019-03-15 Renzo Andri , Lukas Cavigelli , Davide Rossi , Luca Benini

We introduce Virtual Width Networks (VWN), a framework that delivers the benefits of wider representations without incurring the quadratic cost of increasing the hidden size. VWN decouples representational width from backbone width,…

Machine Learning · Computer Science 2025-11-18 Seed , Baisheng Li , Banggu Wu , Bole Ma , Bowen Xiao , Chaoyi Zhang , Cheng Li , Chengyi Wang , Chengyin Xu , Chi Zhang , Chong Hu , Daoguang Zan , Defa Zhu , Dongyu Xu , Du Li , Faming Wu , Fan Xia , Ge Zhang , Guang Shi , Haobin Chen , Hongyu Zhu , Hongzhi Huang , Huan Zhou , Huanzhang Dou , Jianhui Duan , Jianqiao Lu , Jianyu Jiang , Jiayi Xu , Jiecao Chen , Jin Chen , Jin Ma , Jing Su , Jingji Chen , Jun Wang , Jun Yuan , Juncai Liu , Jundong Zhou , Kai Hua , Kai Shen , Kai Xiang , Kaiyuan Chen , Kang Liu , Ke Shen , Liang Xiang , Lin Yan , Lishu Luo , Mengyao Zhang , Ming Ding , Mofan Zhang , Nianning Liang , Peng Li , Penghao Huang , Pengpeng Mu , Qi Huang , Qianli Ma , Qiyang Min , Qiying Yu , Renming Pang , Ru Zhang , Shen Yan , Shen Yan , Shixiong Zhao , Shuaishuai Cao , Shuang Wu , Siyan Chen , Siyu Li , Siyuan Qiao , Tao Sun , Tian Xin , Tiantian Fan , Ting Huang , Ting-Han Fan , Wei Jia , Wenqiang Zhang , Wenxuan Liu , Xiangzhong Wu , Xiaochen Zuo , Xiaoying Jia , Ximing Yang , Xin Liu , Xin Yu , Xingyan Bin , Xintong Hao , Xiongcai Luo , Xujing Li , Xun Zhou , Yanghua Peng , Yangrui Chen , Yi Lin , Yichong Leng , Yinghao Li , Yingshuan Song , Yiyuan Ma , Yong Shan , Yongan Xiang , Yonghui Wu , Yongtao Zhang , Yongzhen Yao , Yu Bao , Yuehang Yang , Yufeng Yuan , Yunshui Li , Yuqiao Xian , Yutao Zeng , Yuxuan Wang , Zehua Hong , Zehua Wang , Zengzhi Wang , Zeyu Yang , Zhengqiang Yin , Zhenyi Lu , Zhexi Zhang , Zhi Chen , Zhi Zhang , Zhiqi Lin , Zihao Huang , Zilin Xu , Ziyun Wei , Zuo Wang

Significant success has been reported recently using deep neural networks for classification. Such large networks can be computationally intensive, even after training is over. Implementing these trained networks in hardware chips with a…

Machine Learning · Statistics 2013-10-25 Daniel Soudry , Ron Meir

Contemporary state-of-the-art neural networks have increasingly large numbers of parameters, which prevents their deployment on devices with limited computational power. Pruning is one technique to remove unnecessary weights and reduce…

Machine Learning · Computer Science 2023-08-15 Sahel Mohammad Iqbal , Subhankar Mishra

Binary neural networks have attracted tremendous attention due to the efficiency for deploying them on mobile devices. Since the weak expression ability of binary weights and features, their accuracy is usually much lower than that of…

Machine Learning · Computer Science 2019-09-18 Mingzhu Shen , Kai Han , Chunjing Xu , Yunhe Wang

The state-of-the-art approaches employ approximate computing to reduce the energy consumption of DNN hardware. Approximate DNNs then require extensive retraining afterwards to recover from the accuracy loss caused by the use of approximate…

Neural and Evolutionary Computing · Computer Science 2020-01-31 Vojtech Mrazek , Zdenek Vasicek , Lukas Sekanina , Muhammad Abdullah Hanif , Muhammad Shafique

We introduce a method to train Quantized Neural Networks (QNNs) --- neural networks with extremely low precision (e.g., 1-bit) weights and activations, at run-time. At train-time the quantized weights and activations are used for computing…

Neural and Evolutionary Computing · Computer Science 2016-09-23 Itay Hubara , Matthieu Courbariaux , Daniel Soudry , Ran El-Yaniv , Yoshua Bengio

In contrast to traditional weight optimization in a continuous space, we demonstrate the existence of effective random networks whose weights are never updated. By selecting a weight among a fixed set of random values for each individual…

Machine Learning · Computer Science 2021-06-09 Maxwell Mbabilla Aladago , Lorenzo Torresani

We propose two efficient approximations to standard convolutional neural networks: Binary-Weight-Networks and XNOR-Networks. In Binary-Weight-Networks, the filters are approximated with binary values resulting in 32x memory saving. In…

Computer Vision and Pattern Recognition · Computer Science 2016-08-04 Mohammad Rastegari , Vicente Ordonez , Joseph Redmon , Ali Farhadi

Low-bit quantized neural networks are of great interest in practical applications because they significantly reduce the consumption of both memory and computational resources. Binary neural networks are memory and computationally efficient…

Machine Learning · Computer Science 2022-05-20 Anton Trusov , Elena Limonova , Dmitry Nikolaev , Vladimir V. Arlazarov

Binary Neural Networks enable smart IoT devices, as they significantly reduce the required memory footprint and computational complexity while retaining a high network performance and flexibility. This paper presents ChewBaccaNN, a 0.7…

Signal Processing · Electrical Eng. & Systems 2021-03-01 Renzo Andri , Geethan Karunaratne , Lukas Cavigelli , Luca Benini

We propose a novel fine-grained quantization (FGQ) method to ternarize pre-trained full precision models, while also constraining activations to 8 and 4-bits. Using this method, we demonstrate a minimal loss in classification accuracy on…

Machine Learning · Computer Science 2017-05-31 Naveen Mellempudi , Abhisek Kundu , Dheevatsa Mudigere , Dipankar Das , Bharat Kaul , Pradeep Dubey

Deep convolutional neural network (CNN) inference requires significant amount of memory and computation, which limits its deployment on embedded devices. To alleviate these problems to some extent, prior research utilize low precision…

Machine Learning · Computer Science 2017-03-10 Liangzhen Lai , Naveen Suda , Vikas Chandra

The proliferation of Artificial Neural Networks (ANNs) has led to increased energy consumption, raising concerns about their sustainability. Spiking Neural Networks (SNNs), which are inspired by biological neural systems and operate using…

Neural and Evolutionary Computing · Computer Science 2024-09-25 Lucas Deckers , Benjamin Vandersmissen , Ing Jyh Tsang , Werner Van Leekwijck , Steven Latré

Convolutional Neural Networks (CNNs) are pivotal in image classification tasks due to their robust feature extraction capabilities. However, their high computational and memory requirements pose challenges for deployment in…

Computer Vision and Pattern Recognition · Computer Science 2025-01-28 Nathan Isong

This paper introduces a tensor neural network (TNN) to address nonparametric regression problems, leveraging its distinct sub-network structure to effectively facilitate variable separation and enhance the approximation of complex,…

Machine Learning · Statistics 2024-09-16 Yongxin Li , Yifan Wang , Zhongshuo Lin , Hehu Xie

BitNet b1.58 (Ma et al., 2024) demonstrates that large language models can operate entirely on ternary weights {-1, 0, +1}, yet no native binary wire format exists for such models. NativeTernary closes this gap. Benchmarked against GGUF on…

Machine Learning · Computer Science 2026-04-09 Maharshi Savdhariya

In a recently published paper [1], it is shown that deep neural networks (DNNs) with random Gaussian weights preserve the metric structure of the data, with the property that the distance shrinks more when the angle between the two data…

Machine Learning · Statistics 2019-04-02 Talha Cihad Gulcu , Alper Gungor
‹ Prev 1 4 5 6 7 8 10 Next ›