English
Related papers

Related papers: Unified continuous-time q-learning for mean-field …

200 papers

This paper studies a new class of dynamic optimization problems of large-population (LP) system which consists of a large number of negligible and coupled agents. The most significant feature in our setup is the dynamics of individual…

Optimization and Control · Mathematics 2014-03-18 Jianhui Huang , Shujun Wang , Hua Xiao

This paper studies the statistical theory of batch data reinforcement learning with function approximation. Consider the off-policy evaluation problem, which is to estimate the cumulative value of a new target policy from logged history…

Machine Learning · Computer Science 2020-02-25 Yaqi Duan , Mengdi Wang

Recent advances in mean-field game literature enable the reduction of large-scale multi-agent problems to tractable interactions between a representative agent and a population distribution. However, existing approaches typically assume a…

Multiagent Systems · Computer Science 2026-02-17 Bhavini Jeloka , Yue Guan , Panagiotis Tsiotras

This paper investigates a linear-quadratic mean field games problem with common noise, where the drift term and diffusion term of individual state equations are coupled with both the state, control, and mean field terms of the state, and we…

Optimization and Control · Mathematics 2025-08-12 Wenyu Cong , Jingtao Shi , Bingchang Wang

This paper concerns imitation learning (IL) (i.e, the problem of learning to mimic expert behaviors from demonstrations) in cooperative multi-agent systems. The learning problem under consideration poses several challenges, characterized by…

Machine Learning · Computer Science 2023-10-11 The Viet Bui , Tien Mai , Thanh Hong Nguyen

In recent years, reinforcement learning and its multi-agent analogue have achieved great success in solving various complex control problems. However, multi-agent reinforcement learning remains challenging both in its theoretical analysis…

Robotics · Computer Science 2023-02-10 Kai Cui , Mengguang Li , Christian Fabian , Heinz Koeppl

In this paper, we explore the susceptibility of the independent Q-learning algorithms (a classical and widely used multi-agent reinforcement learning method) to strategic manipulation of sophisticated opponents in normal-form games played…

Computer Science and Game Theory · Computer Science 2024-07-17 Yuksel Arslantas , Ege Yuceel , Muhammed O. Sayin

We introduce a one-step generative policy for offline reinforcement learning that maps noise directly to actions via a residual reformulation of MeanFlow, making it compatible with Q-learning. While one-step Gaussian policies enable fast…

Machine Learning · Computer Science 2025-11-18 Zeyuan Wang , Da Li , Yulin Chen , Ye Shi , Liang Bai , Tianyuan Yu , Yanwei Fu

We study mean-field control (MFC) problems with common noise using the control randomisation framework, where we substitute the control process with an independent Poisson point process, controlling its intensity instead. To address the…

Optimization and Control · Mathematics 2024-12-31 Robert Denkert , Idris Kharroubi , Huyên Pham

Quantum federated learning (QFL) is a combination of distributed quantum computing and federated machine learning, integrating the strengths of both to enable privacy-preserving decentralized learning with quantum-enhanced capabilities. It…

Machine Learning · Computer Science 2025-08-25 Dinh C. Nguyen , Md Raihan Uddin , Shaba Shaon , Ratun Rahman , Octavia Dobre , Dusit Niyato

This paper introduces a multi-agent approach to adjust traffic lights based on traffic situation in order to reduce average delay time. In the traffic model, lights of each intersection are controlled by an autonomous agent. Since decision…

Multiagent Systems · Computer Science 2019-05-07 Abolghasem Daeichian , Amir Haghani

This paper is concerned with uniform stabilization and social optimality for general mean field linear quadratic control systems, where subsystems are coupled via individual dynamics and costs, and the state weight is not assumed with the…

Optimization and Control · Mathematics 2020-03-02 Bing-Chang Wang , Huanshui Zhang , Ji-Feng Zhang

In this paper, we investigate a new model of a linear-quadratic mean-field stochastic Stackelberg differential game with one leader and two followers, in which the leader is allowed to stop her strategy at a random time. Our overarching…

Optimization and Control · Mathematics 2021-06-08 Zhun Gou , Nan-jing Huang , Ming-hui Wang

Model-free reinforcement learning has been successfully applied to a range of challenging problems, and has recently been extended to handle large neural network policies and value functions. However, the sample complexity of model-free…

Machine Learning · Computer Science 2016-03-03 Shixiang Gu , Timothy Lillicrap , Ilya Sutskever , Sergey Levine

The use of target networks is a common practice in deep reinforcement learning for stabilizing the training; however, theoretical understanding of this technique is still limited. In this paper, we study the so-called periodic Q-learning…

Machine Learning · Computer Science 2020-02-25 Donghwan Lee , Niao He

Q-learning is a promising method for solving optimal control problems for uncertain systems without the explicit need for system identification. However, approaches for continuous-time Q-learning have limited provable safety guarantees,…

Systems and Control · Electrical Eng. & Systems 2024-01-30 Soutrik Bandyopadhyay , Shubhendu Bhasin

Recent reinforcement learning (RL) methods have achieved success in various domains. However, multi-agent RL (MARL) remains a challenge in terms of decentralization, partial observability and scalability to many agents. Meanwhile,…

Machine Learning · Computer Science 2024-02-26 Kai Cui , Sascha Hauck , Christian Fabian , Heinz Koeppl

In this work, we focus on an infinite horizon mean-field linear-quadratic stochastic control problem with jumps. Firstly, the infinite horizon linear mean-field stochastic differential equations and backward stochastic differential…

Optimization and Control · Mathematics 2023-11-14 Qingmeng Wei , Yaqi Xu , Zhiyong Yu

Q-learning has long been one of the most popular reinforcement learning algorithms, and theoretical analysis of Q-learning has been an active research topic for decades. Although researches on asymptotic convergence analysis of Q-learning…

Artificial Intelligence · Computer Science 2022-07-26 Han-Dong Lim , Donghwan Lee

Many practical reinforcement learning environments have a discrete factored action space that induces a large combinatorial set of actions, thereby posing significant challenges. Existing approaches leverage the regular structure of the…

Machine Learning · Computer Science 2025-05-01 Junkyu Lee , Tian Gao , Elliot Nelson , Miao Liu , Debarun Bhattacharjya , Songtao Lu