English
Related papers

Related papers: Proximal Policy Optimization for Integrated Sensin…

200 papers

Distributed Opportunistic Scheduling (DOS) techniques have been recently proposed to improve the throughput performance of wireless networks. With DOS, each station contends for the channel with a certain access probability. If a contention…

Networking and Internet Architecture · Computer Science 2014-12-16 Andres Garcia-Saavedra , Albert Banchs , Pablo Serrano , Joerg Widmer

Utilizing millimeter-wave (mmWave) frequencies for wireless communication in \emph{mobile} systems is challenging since it requires continuous tracking of the beam direction. Recently, beam tracking techniques based on channel sparsity…

Signal Processing · Electrical Eng. & Systems 2020-01-07 Daoud Burghal , Naveed A. Abbasi , Andreas F. Molisch

In this paper, adaptive pilot beam sequence design for channel estimation in large millimeter-wave (mmWave) MIMO systems is considered. By exploiting the sparsity of mmWave MIMO channels with the virtual channel representation and imposing…

Information Theory · Computer Science 2016-11-15 Junyeong Seo , Youngchul Sung , Gilwon Lee , Donggun Kim

Millimeter-wave (mmWave) networks offer the potential for high-speed data transfer and precise localization, leveraging large antenna arrays and extensive bandwidths. However, these networks are challenged by significant path loss and…

Social and Information Networks · Computer Science 2023-12-29 Wan-Ting Shih , Chao-Kai Wen , Shang-Ho Tsai , Shi Jin , Chau Yuen

In millimeter wave communications, beam training is an effective way to achieve beam alignment. Traditional beam training method allocates training resources equally to each beam in the pre-designed beam training codebook. The performance…

Information Theory · Computer Science 2018-10-25 Zihan Tang , Jun Wang , Jintao Wang , Jian Song

Reinforcement learning (RL) has shown extraordinary potential in aligning diffusion models to downstream tasks, yet most of them still suffer from significant reward hacking, which degrades generative diversity and quality by inducing…

Machine Learning · Computer Science 2026-05-14 Jiaming Li , Chenyu Zhu , Nanxi Yi , Youjun Bao , Li Sun , Quanying Lv , Xiang Fang , Daizong Liu , Jianjun Li , Kun He , Bowen Zhou , Zhiyuan Ma

The use of higher frequencies in mobile communication systems leads to smaller cell sizes, resulting in the deployment of more base stations and an increase in handovers to support user mobility. This can lead to frequent radio link…

Networking and Internet Architecture · Computer Science 2025-03-28 Johannes Voigt , Peter Jiacheng Gu , Peter Rost

Direct preference optimization (DPO) is widely used as a simple and stable method for aligning large language models (LLMs) with human preferences. This paper investigates a generalized DPO loss that enables a policy model to match the…

Machine Learning · Computer Science 2025-10-28 Yeongmin Kim , Heesun Bae , Byeonghu Na , Il-Chul Moon

In this paper, we investigate the problem of beam alignment in millimeter wave (mmWave) systems, and design an optimal algorithm to reduce the overhead. Specifically, due to directional communications, the transmitter and receiver beams…

Information Theory · Computer Science 2017-12-29 Morteza Hashemi , Ashutosh Sabharwal , C. Emre Koksal , Ness B. Shroff

The allocation of scarce spectral resources to support as many user applications as possible while maintaining reasonable quality of service is a fundamental problem in wireless communication. We argue that the problem is best formulated in…

Networking and Internet Architecture · Computer Science 2007-05-23 Zygmunt Haas , Joseph Y. Halpern , Li Li , Stephen B. Wicker

Location information offered by external positioning systems, e.g., satellite navigation, can be used as prior information in the process of beam alignment and channel parameter estimation for reconfigurable intelligent surface (RIS)-aided…

Signal Processing · Electrical Eng. & Systems 2021-03-30 Jiguang He , Henk Wymeersch , Markku Juntti

A pinching-antenna system (PASS)-enhanced mobile edge computing (MEC) architecture is investigated to improve the task offloading efficiency and latency performance in dynamic wireless environments. By leveraging dielectric waveguides and…

Signal Processing · Electrical Eng. & Systems 2025-10-28 Zhaoming Hu , Ruikang Zhong , Xidong Mu , Dengao Li , Yuanwei Liu

Location-aided beam alignment has been proposed recently as a potential approach for fast link establishment in millimeter wave (mmWave) massive MIMO (mMIMO) communications. However, due to mobility and other imperfections in the estimation…

Information Theory · Computer Science 2017-08-29 Flavio Maschietti , David Gesbert , Paul de Kerret , Henk Wymeersch

In recent years, reinforcement learning (RL) has gained increasing attention in control engineering. Especially, policy gradient methods are widely used. In this work, we improve the tracking performance of proximal policy optimization…

Machine Learning · Computer Science 2021-07-21 Jana Mayer , Johannes Westermann , Juan Pedro Gutiérrez H. Muriedas , Uwe Mettin , Alexander Lampe

Multi-Agent Proximal Policy Optimization (MAPPO) is a variant of the Proximal Policy Optimization (PPO) algorithm, specifically tailored for multi-agent reinforcement learning (MARL). MAPPO optimizes cooperative multi-agent settings by…

Machine Learning · Computer Science 2026-05-14 Changha Lee , Gyusang Cho

Direct preference optimization (DPO), a widely adopted offline preference optimization algorithm, aims to align large language models (LLMs) with human-desired behaviors using pairwise preference data. However, the generation of the winning…

Computation and Language · Computer Science 2025-02-19 Yuxin Jiang , Bo Huang , Yufei Wang , Xingshan Zeng , Liangyou Li , Yasheng Wang , Xin Jiang , Lifeng Shang , Ruiming Tang , Wei Wang

Proximal policy optimization (PPO) is one of the most popular deep reinforcement learning (RL) methods, achieving state-of-the-art performance across a wide range of challenging tasks. However, as a model-free RL method, the success of PPO…

Machine Learning · Computer Science 2019-11-11 Yuhui Wang , Hao He , Xiaoyang Tan , Yaozhong Gan

In this paper, we develop an efficient training beam sequence design approach for millimeter wave MISO tracking systems. We impose a discrete state Markov process assumption on the evolution of the angle of departure and introduce the…

Information Theory · Computer Science 2022-09-19 Deyou Zhang , Ming Xiao , Mikael Skoglund

Accurate parameter estimation such as angle of arrival (AOA) is essential to enhance the performance of integrated sensing and communication (ISAC) in mmWave multiple-input multiple-output (MIMO) systems. This work presents a sensing-aided…

Information Theory · Computer Science 2025-03-05 Ngoc-Son Duong , Khac-Hoang Ngo , Thai-Mai Dinh , Van-Linh Nguyen

Fast and precise beam alignment is crucial for high-quality data transmission in millimeter-wave (mmWave) communication systems, where large-scale antenna arrays are utilized to overcome the severe propagation loss. To tackle the…

Signal Processing · Electrical Eng. & Systems 2023-08-24 Junyi Yang , Weifeng Zhu , Meixia Tao , Shu Sun
‹ Prev 1 3 4 5 6 7 10 Next ›