中文
相关论文

相关论文: Adaptive Quantization for Key Generation in Low-Po…

200 篇论文

Large pre-trained models, such as large language models (LLMs), present significant resource challenges for fine-tuning due to their extensive parameter sizes, especially for applications in mobile systems. To address this, Low-Rank…

机器学习 · 计算机科学 2024-07-18 Yuzhu Mao , Siqi Ping , Zihao Zhao , Yang Liu , Wenbo Ding

The deployment of LoRa networks necessitates joint performance optimization, including packet delivery rate, energy efficiency, and throughput. Additionally, multiple LoRa parameters for packet transmission must be dynamically configured to…

网络与互联网体系结构 · 计算机科学 2025-04-24 Ruiqi Wang , Tongyu Song , Jing Ren , Xiong Wang , Shizhong Xu , Sheng Wang

Deploying Large Language Models (LLMs) on resource-constrained edge devices like the Raspberry Pi presents challenges in computational efficiency, power consumption, and response latency. This paper explores quantization-based optimization…

机器学习 · 计算机科学 2025-04-04 Mahsa Ardakani , Jinendra Malekar , Ramtin Zand

LoRaWAN is a promising low power long range wireless communications technology for the Internet of Things. An important feature of LoRaWAN gateways is related to so-called capture effect: under some conditions the gateway may correctly…

网络与互联网体系结构 · 计算机科学 2020-08-10 Dmitry Bankov , Evgeny Khorov , Andrey Lyakhov

Meticulous modelling and performance analysis of Low-Power Wide-Area (LPWA) networks are essential for large scale dense Internet-of-Things (IoT) deployments. As Long Range (LoRa) is currently one of the most prominent LPWA technologies, we…

信号处理 · 电气工程与系统科学 2020-08-28 Yathreb Bouazizi , Fatma Benkhelifa , Julie McCann

The scarcity of high-quality residential load data can pose obstacles for decarbonizing the residential sector as well as effective grid planning and operation. The above challenges have motivated research into generating synthetic load…

机器学习 · 计算机科学 2025-04-28 Xinyu Liang , Hao Wang

The use of low-rank adaptation (LoRA) with frozen pretrained language models (PLMs) has become increasing popular as a mainstream, resource-efficient modeling approach for memory-constrained hardware. In this study, we first explore how to…

This work presents generalized low-rank signal decompositions with the aid of switching techniques and adaptive algorithms, which do not require eigen-decompositions, for space-time adaptive processing. A generalized scheme is proposed to…

信息论 · 计算机科学 2013-04-09 R. C. de Lamare

Post-training quantization (PTQ) has emerged as a prevailing technique for deploying large language models (LLMs) efficiently in terms of both memory and computation, across edge devices and server platforms. Existing PTQ methods primarily…

机器学习 · 计算机科学 2026-03-10 Yeonsik Park , Hyeonseong Kim , Seungkyu Choi

Low-Rank Adaptation (LoRA) enables parameter-efficient fine-tuning of large language models by decomposing weight updates into low-rank matrices, significantly reducing storage and computational overhead. While effective, standard LoRA…

机器学习 · 计算机科学 2025-09-03 Patryk Marszałek , Klaudia Bałazy , Jacek Tabor , Tomasz Kuśmierczyk

Recently, the promising aspects of compressive sensing have inspired new circuit-level approaches for their efficient realization within the literature. However, most of these recent advances involving novel sampling techniques have been…

新兴技术 · 计算机科学 2019-03-14 Soheil Salehi , Ramtin Zand , Alireza Zaeemzadeh , Nazanin Rahnavard , Ronald F. DeMara

LoRaWAN is one of the leading Low Power Wide Area Network (LPWAN) architectures. It was originally designed for systems consisting of static sensor or Internet of Things (IoT) devices and static gateways. It was recently updated to…

网络与互联网体系结构 · 计算机科学 2020-04-15 Po-Yu Chen , Laksh Bhatia , Roman Kolcun , David Boyle , Julie A. McCann

The Key-Value (KV) cache is a crucial component in serving transformer-based autoregressive large language models (LLMs), enabling faster inference by storing previously computed KV vectors. However, its memory consumption scales linearly…

机器学习 · 计算机科学 2024-10-07 Rongzhi Zhang , Kuang Wang , Liyuan Liu , Shuohang Wang , Hao Cheng , Chao Zhang , Yelong Shen

In this work we propose an adaptive buffer-aided space-time coding scheme for cooperative wireless networks. A maximum likelihood receiver and adjustable code vectors are considered subject to a power constraint with an amplify-and-forward…

信息论 · 计算机科学 2015-07-07 T. Peng , R. C. de Lamare

There have been many research efforts in the area of localization in recent years. Especially within the Internet of Things (IoT), the knowledge of position information for individual components is of great interest, for example, in asset…

信号处理 · 电气工程与系统科学 2024-01-23 Thomas Maul , Joerg Robert , Sebastian Klob

Low power long-range networks like LoRa have become increasingly mainstream for Internet of Things deployments. Given the versatility of applications that these protocols enable, they support many data rates and bandwidths. Yet, for a given…

网络与互联网体系结构 · 计算机科学 2021-09-14 Zerina Kapetanovic , Deepak Vasisht , Tusher Chakraborty , Joshua R. Smith , Ranveer Chandra

Low-rank adaptation (LoRA) is a predominant parameter-efficient finetuning method for adapting large language models (LLMs) to downstream tasks. Meanwhile, Compute-in-Memory (CIM) architectures demonstrate superior energy efficiency due to…

计算与语言 · 计算机科学 2026-03-10 Taiqiang Wu , Chenchen Ding , Wenyong Zhou , Yuxin Cheng , Xincheng Feng , Shuqi Wang , Wendong Xu , Chufan Shi , Zhengwu Liu , Ngai Wong

Low-rank adaptation (LoRA) has shifted the paradigm of adapting pre-trained Vision Transformers (ViT), achieving great efficiency by updating only a subset of tailored parameters to approximate weight updates. However, the multi-head design…

计算机视觉与模式识别 · 计算机科学 2024-10-10 Yibo Zhong , Yao Zhou

Post-training Quantization (PTQ) has become a widely used technique for improving inference efficiency of large language models (LLMs). However, existing PTQ methods generally suffer from crucial limitations such as heavy calibration data…

机器学习 · 计算机科学 2025-11-03 Yongyi Yang , Jianyang Gao , Wei Hu

State-of-the-art neural language models represented by Transformers are becoming increasingly complex and expensive for practical applications. Low-bit deep neural network quantization techniques provides a powerful solution to dramatically…

计算与语言 · 计算机科学 2021-12-23 Junhao Xu , Shoukang Hu , Jianwei Yu , Xunying Liu , Helen Meng
‹ 上一页 1 8 9 10 下一页 ›