English
Related papers

Related papers: Proximal Policy Optimization for Integrated Sensin…

200 papers

Proximal policy optimization (PPO) is one of the most successful deep reinforcement-learning methods, achieving state-of-the-art performance across a wide range of challenging tasks. However, its optimization behavior is still far from…

Machine Learning · Computer Science 2020-01-15 Yuhui Wang , Hao He , Chao Wen , Xiaoyang Tan

Optical computing holds promise for high-speed, energy-efficient information processing, with diffractive optical networks emerging as a flexible platform for implementing task-specific transformations. A challenge, however, is the…

Machine Learning · Computer Science 2026-01-05 Yuhang Li , Shiqi Chen , Tingyu Gong , Aydogan Ozcan

We investigate the performance of a multiple reconfigurable intelligence surface (RIS)-aided millimeter wave (mmWave) beamspace multiple-input multiple-output (MIMO) system with multiple users (UEs). We focus on a challenging scenario in…

Information Theory · Computer Science 2025-05-21 Zaid Abdullah , Mario R. Camana , Abuzar B. M. Adam , Chandan K. Sheemar , Eva Lagunas , Symeon Chatzinotas

Proximal policy optimization (PPO) is one of the most popular state-of-the-art on-policy algorithms that has become a standard baseline in modern reinforcement learning with applications in numerous fields. Though it delivers stable…

Machine Learning · Computer Science 2025-02-25 Qisai Liu , Zhanhong Jiang , Hsin-Jung Yang , Mahsa Khosravi , Joshua R. Waite , Soumik Sarkar

Proximal Policy Optimization (PPO) has become the predominant algorithm for on-policy reinforcement learning due to its scalability and empirical robustness across domains. However, there is a significant disconnect between the underlying…

Emerging wireless communication systems will be characterized by a tight coupling between communication and positioning. This is particularly apparent in millimeter-wave (mm-wave) communications, where devices use a large number of antennas…

Information Theory · Computer Science 2017-05-17 Gabriel E. Garcia , Gonzalo Seco-Granados , Eleftherios Karipidis , Henk Wymeersch

Problem Definition: Managing inpatient flow in large hospital systems is challenging due to the complexity of assigning randomly arriving patients -- either waiting for primary units or being overflowed to alternative units. Current…

Optimization and Control · Mathematics 2026-05-08 Jingjing Sun , Jim Dai , Pengyi Shi

According to the physical phenomena of atmospheric channels and wave propagation, performance of wireless communication systems can be optimized by simply adjusting its parameters. This way is more economically favorable than consuming…

Signal Processing · Electrical Eng. & Systems 2018-02-23 Mohammad Ali Amirabadi , Vahid Tabataba Vakili

Millimeter wave communications are essential for modern wireless networks. It supports high data rates but suffers from severe path loss, which requires precise beam alignment to maintain reliable links. This beam management is particularly…

Signal Processing · Electrical Eng. & Systems 2025-11-05 Ailton Oliveira , Amir Khatibi , Daniel Suzuki , Ilan Correa , José Rezende , Aldebaro Klautau

In order to cope with the severe path loss, millimeter-wave (mm-wave) systems exploit highly directional communication. As a consequence, even a slight beam misalignment between two communicating devices (for example, due to mobility) can…

Networking and Internet Architecture · Computer Science 2016-12-26 Joan Palacios , Danilo De Donno , Joerg Widmer

Wireless extended reality (XR) teleoperation provides embodied interaction capability for collecting humanoid robot demonstrations, but the large-scale adoption is restricted by the overhead of high-frequency motion transmission. This paper…

Information Theory · Computer Science 2026-05-20 Caolu Xu , Zhiyong Chen , Meixia Tao , Li Song , Feng Yang , Wenjun Zhang

Reinforcement Learning (RL) has made significant strides in various domains, and policy gradient methods like Proximal Policy Optimization (PPO) have gained popularity due to their balance in performance, training stability, and…

Machine Learning · Computer Science 2025-05-21 Andrei Cozma , Landon Harris , Hairong Qi

Proximal Policy Optimization (PPO) is among the most widely used algorithms in reinforcement learning, which achieves state-of-the-art performance in many challenging problems. The keys to its success are the reliable policy updates through…

Machine Learning · Computer Science 2021-07-02 Mónika Farsang , Luca Szegletes

Proximal Policy Optimization (PPO) is a widely used reinforcement learning algorithm that heavily relies on accurate advantage estimates for stable and efficient training. However, raw advantage signals can exhibit significant variance,…

Machine Learning · Computer Science 2025-05-22 Soham Sane

Proximal Policy Optimization (PPO) is a highly popular model-free reinforcement learning (RL) approach. However, we observe that in a continuous action space, PPO can prematurely shrink the exploration variance, which leads to slow progress…

Machine Learning · Computer Science 2020-11-04 Perttu Hämäläinen , Amin Babadi , Xiaoxiao Ma , Jaakko Lehtinen

Proximal policy optimization (PPO) algorithm is a deep reinforcement learning algorithm with outstanding performance, especially in continuous control tasks. But the performance of this method is still affected by its exploration ability.…

Machine Learning · Computer Science 2020-11-12 Junwei Zhang , Zhenghao Zhang , Shuai Han , Shuai Lü

Huge overhead of beam training poses a significant challenge to mmWave communications. To address this issue, beam tracking has been widely investigated whereas existing methods are hard to handle serious multipath interference and…

Signal Processing · Electrical Eng. & Systems 2021-02-09 Ke Ma , Dongxuan He , Hancun Sun , Zhaocheng Wang

Coverage and capacity are the important metrics for performance evaluation in wireless networks, while the coverage and capacity have several conflicting relationships, e.g. high transmit power contributes to large coverage but high…

Information Theory · Computer Science 2022-04-14 Xinyu Gao , Wenqiang Yi , Alexandros Agapitos , Hao Wang , Yuanwei Liu

Next generation communication systems require accurate beam alignment to counteract the impairments that characterize propagation in high-frequency bands. The overhead of the pilot sequences required to select the best beam pair is…

Signal Processing · Electrical Eng. & Systems 2024-02-27 Tien Ngoc Ha , Daniel Romero , Roberto López-Valcarce

Future intelligent indoor wireless environments require fast and reliable beam alignment to sustain high-throughput links under mobility and blockage. Exhaustive beam training achieves optimal performance but is prohibitively costly. In…

Networking and Internet Architecture · Computer Science 2026-02-19 Parth Ashokbhai Shiroya , Amod Ashtekar , Swarnagowri Shashidhar , Mohammed E. Eltayeb