English
Related papers

Related papers: A Deep Reinforcement Learning-based Approach for A…

200 papers

While traditional handovers (THOs) have served as a backbone for mobile connectivity, they increasingly suffer from failures and delays, especially in dense deployments and high-frequency bands. To address these limitations, 3GPP introduced…

Networking and Internet Architecture · Computer Science 2025-12-29 Michail Kalntis , George Iosifidis , José Suárez-Varela , Andra Lutu , Fernando A. Kuipers

Real-time Three-dimensional (3D) scene representation is a foundational element that supports a broad spectrum of cutting-edge applications, including digital manufacturing, Virtual, Augmented, and Mixed Reality (VR/AR/MR), and the emerging…

Computer Vision and Pattern Recognition · Computer Science 2025-09-23 Xiangmin Xu , Zhen Meng , Kan Chen , Jiaming Yang , Emma Li , Philip G. Zhao , David Flynn

A limitation of bandwidth in the wireless network and the exponential rise in the high data rate requirement prompted the development of Massive Multiple-Input-Multiple-Output (MIMO) technique in 5G. Using this method the ever rising data…

Signal Processing · Electrical Eng. & Systems 2024-06-11 Mariea Sharaf Anzum , Moontasir Rafique , Md. Asif Ishrak Sarder , Fehima Tajrian , Abdullah Bin Shams

Reinforcement learning (RL) has become a cornerstone for fine-tuning Large Language Models (LLMs), with Proximal Policy Optimization (PPO) serving as the de facto standard algorithm. Despite its ubiquity, we argue that the core ratio…

Machine Learning · Computer Science 2026-05-27 Penghui Qi , Xiangxin Zhou , Zichen Liu , Tianyu Pang , Chao Du , Min Lin , Wee Sun Lee

This paper elaborates on Conditional Handover (CHO) modelling, aimed at maximizing the use of contention free random access (CFRA) during mobility. This is a desirable behavior as CFRA increases the chance of fast and successful handover.…

Networking and Internet Architecture · Computer Science 2023-07-28 Jedrzej Stanczak , Umur Karabulut , Ahmad Awada

Traditional risk factors like beta, size/value, and momentum often lag behind market dynamics in measuring and predicting stock return volatility. Statistical models like PCA and factor analysis fail to capture hidden nonlinear…

Computational Engineering, Finance, and Science · Computer Science 2025-09-23 Wenyan Xu , Jiayu Chen , Dawei Xiang , Chen Li , Yonghong Hu , Zhonghua Lu

We demonstrate a distributed and a centralized 4G/5G compliant approach to minimize signaling and latency related to user mobility in cellular networks. This is crucial due to the densification of networks and the additional signaling…

Networking and Internet Architecture · Computer Science 2022-01-05 Merim Dzaferagic , Nicola Marchetti , Irene Macaluso

Optimizing radio transmission power and user data rates in wireless systems via power control requires an accurate and instantaneous knowledge of the system model. While this problem has been extensively studied in the literature, an…

Optimization and Control · Mathematics 2016-11-22 Euhanna Ghadimi , Francesco Davide Calabrese , Gunnar Peters , Pablo Soldati

Previous studies that have formulated multi-agent reinforcement learning (RL) algorithms for adaptive traffic signal control have primarily used value-based RL methods. However, recent literature has shown that policy-based methods may…

Multiagent Systems · Computer Science 2025-07-03 Dickness Kakitahi Kwesiga , Angshuman Guin , Michael Hunter

Direct Preference Optimization (DPO) has emerged as a predominant alignment method for diffusion models, facilitating off-policy training without explicit reward modeling. However, its reliance on large-scale, high-quality human preference…

Computer Vision and Pattern Recognition · Computer Science 2026-02-09 Khiem Pham , Quang Nguyen , Tung Nguyen , Jingsen Zhu , Michele Santacatterina , Dimitris Metaxas , Ramin Zabih

Direct Preference Optimization (DPO) has emerged as a simple and effective method for aligning large language models. However, its reliance on a fixed temperature parameter leads to suboptimal training on diverse preference data, causing…

Machine Learning · Computer Science 2025-10-08 Hyung Gyu Rho

Recent studies have shown the great potential of diffusion models in improving reinforcement learning (RL) by modeling complex policies, expressing a high degree of multi-modality, and efficiently handling high-dimensional continuous…

Robotics · Computer Science 2025-05-14 Huiyun Jiang , Zhuang Yang

We consider a dynamic millimeter-wave network with integrated access and backhaul, where mobile relay nodes move to auto-reconfigure the wireless backhaul. Specifically, we focus on in-band relaying networks, which conduct access and…

Multiagent Systems · Computer Science 2023-02-16 Mohamed Sana , Benoit Miscopein

With the extensive applications of machine learning models, automatic hyperparameter optimization (HPO) has become increasingly important. Motivated by the tuning behaviors of human experts, it is intuitive to leverage auxiliary knowledge…

Machine Learning · Computer Science 2022-06-07 Yang Li , Yu Shen , Huaijun Jiang , Wentao Zhang , Zhi Yang , Ce Zhang , Bin Cui

Proximal Policy Optimization (PPO) is among the most widely used algorithms in reinforcement learning, which achieves state-of-the-art performance in many challenging problems. The keys to its success are the reliable policy updates through…

Machine Learning · Computer Science 2021-07-02 Mónika Farsang , Luca Szegletes

Direct Preference Optimization (DPO) and its variants have become increasingly popular for aligning language models with human preferences. These methods aim to teach models to better distinguish between chosen (or preferred) and rejected…

Computation and Language · Computer Science 2025-06-09 Xiliang Yang , Feng Jiang , Qianen Zhang , Lei Zhao , Xiao Li

Recent mobile equipment (as well as the norm IEEE 802.21) now offers the possibility for users to switch from one technology to another (vertical handover). This allows flexibility in resource assignments and, consequently, increases the…

Computer Science and Game Theory · Computer Science 2009-09-07 Pierre Coucheney , Corinne Touati , Bruno Gaujal

This paper introduces a novel approach to radio resource allocation in multi-cell wireless networks using a fully scalable multi-agent reinforcement learning (MARL) framework. A distributed method is developed where agents control…

Multiagent Systems · Computer Science 2024-09-19 Yiming Zhang , Dongning Guo

The construction of Low Earth Orbit (LEO) satellite constellations has recently attracted tremendous attention from both academia and industry. The 5G and 6G standards have identified LEO satellite networks as a key component of future…

Networking and Internet Architecture · Computer Science 2025-07-11 Jiasheng Wu , Shaojie Su , Wenjun Zhu , Xiong Wang , Jingjing Zhang , Xingqiu He , Yue Gao

Direct Preference Optimization (DPO) has gained attention as an efficient alternative to reinforcement learning from human feedback (RLHF) for aligning large language models (LLMs) with human preferences. Despite its advantages, DPO suffers…

Computation and Language · Computer Science 2025-02-21 Ruichen Shao , Bei Li , Gangao Liu , Yang Chen , Xiang Zhou , Jingang Wang , Xunliang Cai , Peng Li