English

Over-the-air Federated Policy Gradient

Machine Learning 2024-02-27 v3 Distributed, Parallel, and Cluster Computing Signal Processing

Abstract

In recent years, over-the-air aggregation has been widely considered in large-scale distributed learning, optimization, and sensing. In this paper, we propose the over-the-air federated policy gradient algorithm, where all agents simultaneously broadcast an analog signal carrying local information to a common wireless channel, and a central controller uses the received aggregated waveform to update the policy parameters. We investigate the effect of noise and channel distortion on the convergence of the proposed algorithm, and establish the complexities of communication and sampling for finding an ϵ\epsilon-approximate stationary point. Finally, we present some simulation results to show the effectiveness of the algorithm.

Keywords

Cite

@article{arxiv.2310.16592,
  title  = {Over-the-air Federated Policy Gradient},
  author = {Huiwen Yang and Lingying Huang and Subhrakanti Dey and Ling Shi},
  journal= {arXiv preprint arXiv:2310.16592},
  year   = {2024}
}

Comments

To appear at IEEE ICC 2024

R2 v1 2026-06-28T13:01:30.667Z