English
Related papers

Related papers: Group Relative Policy Optimization for Robust Blin…

200 papers

This letter proposes a secure beamforming design for downlink non-orthogonal multiple access (NOMA) systems utilizing fluid antenna systems (FAS). We consider a setup where a base station (BS) with $M$ fluid antennas (FAs) communicates to a…

Signal Processing · Electrical Eng. & Systems 2024-11-14 Lifeng Mai , Junteng Yao , Jie Tang , Tuo Wu , Kai-Kit Wong , Hyundong Shin , Fumiyuki Adachi

Proximal Policy Optimization (PPO) has become the predominant algorithm for on-policy reinforcement learning due to its scalability and empirical robustness across domains. However, there is a significant disconnect between the underlying…

Reinforcement learning (RL) has emerged as an effective approach for enhancing the reasoning capabilities of large language models (LLMs), especially in scenarios where supervised fine-tuning (SFT) falls short due to limited…

Machine Learning · Computer Science 2026-04-15 Jian Xiong , Jingbo Zhou , Jingyong Ye , Qiang Huang , Dejing Dou

Group Relative Policy Optimization (GRPO), recently introduced by DeepSeek, is a critic-free reinforcement learning algorithm for fine-tuning large language models. GRPO replaces the value function in Proximal Policy Optimization (PPO) with…

Machine Learning · Computer Science 2026-03-24 Lei Pang , Jun Luo , Ruinan Jin

Optimizing communication topology is fundamental to the efficiency and effectiveness of Large Language Model (LLM)-based Multi-Agent Systems (MAS). While recent approaches utilize reinforcement learning to dynamically construct…

Computation and Language · Computer Science 2026-03-04 Yueyang Cang , Xiaoteng Zhang , Erlu Zhao , Zehua Ji , Yuhang Liu , Yuchen He , Zhiyuan Ning , Chen Yijun , Wenge Que , Li Shi

Fluid Antenna Systems (FAS) introduce a new degree of freedom for wireless networks by enabling the physical antenna position to adapt dynamically to changing radio conditions. While existing studies primarily emphasize physical-layer…

Networking and Internet Architecture · Computer Science 2026-03-24 Ian F. Akyildiz , Tuğçe Bilen

With the evolution of mobile communication systems toward large-scale arrays, high-frequency operation, and reconfigurable antenna architectures, fluid antenna systems (FAS) operating in the near-field (NF) regime provide new degrees of…

Information Theory · Computer Science 2026-05-12 Peng Zhang , Jian Dang , Miaowen Wen , Ziyang Liu , Chen Zhao , Huaifeng Shi , Chengsheng Pan , Zaichen Zhang

The fluid antenna (FA) index modulation (IM)-enabled multiple-input multiple-output (MIMO) system, referred to as FA-IM, significantly enhances spectral efficiency (SE) compared to the conventional FA-assisted MIMO system. To improve…

Information Theory · Computer Science 2024-12-31 Xinghao Guo , Yin Xu , Dazhi He , Cixiao Zhang , Hanjiang Hong , Kai-Kit Wong , Wenjun Zhang , Yiyan Wu

This paper proposes a novel fluid antenna system (FAS)-enabled architecture to improve energy efficiency (EE) without sacrificing capacity. Specifically, we integrate FAS into cell-free massive MIMO systems to counteract low-resolution…

Information Theory · Computer Science 2026-05-28 Jun Qian , Ross Murch , Khaled B. Letaief

Recent Progress in post-training flow matching for text-to-image (T2I) generation with Group Relative Policy Optimization (GRPO) has demonstrated strong potential. However, it is hindered by a critical limitation: inaccurate advantage…

Computer Vision and Pattern Recognition · Computer Science 2026-05-21 Yifu Luo , Haoyuan Sun , Xinhao Hu , Penghui Du , Keyu Fan , Bo Li , Sinan Du , Xu Wan , Zhiyu Chen , Bo Xia , Tiantian Zhang , Yongzhe Chang , Changqian Yu , Kun Gai , Xueqian Wang

Reinforcement Learning with Verifiable Rewards (RLVR) has emerged as a promising paradigm for post-training reasoning models. However, group-based methods such as Group Relative Policy Optimization (GRPO) face a critical dilemma in…

Machine Learning · Computer Science 2026-04-07 Yuning Wu , Ke Wang , Devin Chen , Kai Wei

This paper presents a novel framework for enhancing physical-layer security in integrated sensing and communication (ISAC) systems by leveraging the reconfigurability of fluid antenna systems (FAS). We propose a joint precoding and port…

Signal Processing · Electrical Eng. & Systems 2025-10-01 Abdelhamid Salem , Hao Xu , Kai-Kit Wong , Chan-Byoung Chae , Yangyang Zhang

Speech Recognition has seen a dramatic shift towards adopting Large Language Models (LLMs). This shift is partly driven by good scalability properties demonstrated by LLMs, ability to leverage large amounts of labelled, unlabelled speech…

Audio and Speech Processing · Electrical Eng. & Systems 2025-09-03 Prashanth Gurunath Shivakumar , Yile Gu , Ankur Gandhe , Ivan Bulyko

The fluid antenna concept represents shape-flexible and position-flexible antenna technologies designed to enhance wireless communication applications. In this paper, we apply this concept to reconfigurable intelligent surfaces (RISs),…

Signal Processing · Electrical Eng. & Systems 2025-02-25 Abdelhamid Salem , Kai-Kit Wong , George Alexandropoulos , Chan-Byoung Chae , Ross Murch

As an emerging antenna technology, a fluid antenna system (FAS) enhances spatial diversity to improve both sensing and communication performance by shifting the active antennas among available ports. In this letter, we study the potential…

Signal Processing · Electrical Eng. & Systems 2024-05-12 Jiaqi Zou , Hao Xu , Chao Wang , Lvxin Xu , Songlin Sun , Kaitao Meng , Christos Masouros , Kai-Kit Wong

We propose a blind interference alignment (BIA) through staggered antenna switching scheme with no ideal channel assumption. Contrary to the ideal assumption that channels remain constant during BIA symbol extension period, when the…

Information Theory · Computer Science 2016-11-17 Heecheol Yang , Wonjae Shin , Jungwoo Lee

While reinforcement learning methods such as Group Relative Preference Optimization (GRPO) have significantly enhanced Large Language Models, adapting them to diffusion models remains challenging. In particular, GRPO demands a stochastic…

Machine Learning · Computer Science 2025-10-10 Yihong Luo , Tianyang Hu , Jing Tang

Fluid antennas (FAs) and movable antennas (MAs) have drawn increasing attention in wireless communications recently due to their ability to create favorable channel conditions via local antenna movement within a confined region. In this…

Signal Processing · Electrical Eng. & Systems 2024-08-26 Xin Wei , Weidong Mei , Dong Wang , Boyu Ning , Zhi Chen

The flexibility and reconfigurability at the radio frequency (RF) front-end offered by the fluid antenna system (FAS) make this technology promising for providing remarkable diversity gains in networks with small and constrained devices.…

This letter studies the performance of reconfigurable intelligent surface (RIS)-aided communications for a fluid antenna system (FAS) enabled receiver. Specifically, a fixed singleantenna base station (BS) transmits information through a…

Information Theory · Computer Science 2024-02-27 Farshad Rostami Ghadi , Kai-Kit Wong , Wee Kiat New , Hao Xu , Ross Murch , Yangyang Zhang