English
Related papers

Related papers: Stackelberg Meta-Learning Based Control for Guided…

200 papers

In this paper, we consider a discrete-time stochastic Stackelberg game with a single leader and multiple followers. Both the followers and the leader together have conditionally independent private types, conditioned on action and previous…

Optimization and Control · Mathematics 2022-09-21 Deepanshu Vasal

In this work, we propose a novel memory-based multi-agent meta-learning architecture and learning procedure that allows for learning of a shared communication policy that enables the emergence of rapid adaptation to new and unseen…

Open ad hoc teamwork is the problem of training a single agent to efficiently collaborate with an unknown group of teammates whose composition may change over time. A variable team composition creates challenges for the agent, such as the…

Multiagent Systems · Computer Science 2023-10-31 Arrasy Rahman , Ignacio Carlucho , Niklas Höpner , Stefano V. Albrecht

Recent years have witnessed a rapid proliferation of smart Internet of Things (IoT) devices. IoT devices with intelligence require the use of effective machine learning paradigms. Federated learning can be a promising solution for enabling…

Distributed, Parallel, and Cluster Computing · Computer Science 2020-09-08 Latif U. Khan , Shashi Raj Pandey , Nguyen H. Tran , Walid Saad , Zhu Han , Minh N. H. Nguyen , Choong Seon Hong

Batch reinforcement learning (RL) defines the task of learning from a fixed batch of data lacking exhaustive exploration. Worst-case optimality algorithms, which calibrate a value-function model class from logged experience and perform some…

Machine Learning · Statistics 2023-10-03 Wenzhuo Zhou , Annie Qu

The cooperation among AI systems, and between AI systems and humans is becoming increasingly important. In various real-world tasks, an agent needs to cooperate with unknown partner agent types. This requires the agent to assess the…

Machine Learning · Computer Science 2021-10-05 Antti Keurulainen , Isak Westerlund , Ariel Kwiatkowski , Samuel Kaski , Alexander Ilin

To be helpful assistants, AI agents must be aware of their own capabilities and limitations. This includes knowing when to answer from parametric knowledge versus using tools, when to trust tool outputs, and when to abstain or hedge. Such…

Machine Learning · Computer Science 2025-09-01 Jacob Eisenstein , Reza Aghajani , Adam Fisch , Dheeru Dua , Fantine Huot , Mirella Lapata , Vicky Zayats , Jonathan Berant

Existing methods for learning Stackelberg equilibria typically assume that the followers' (variational, generalized) Nash equilibrium is unique. However, in the presence of multiple equilibria, without a selection convention, the problem…

Optimization and Control · Mathematics 2026-04-30 Silvia Cianchi , Anibal Sanjab , Sergio Grammatico

Model-based reinforcement learning (MBRL) has recently gained immense interest due to its potential for sample efficiency and ability to incorporate off-policy data. However, designing stable and efficient MBRL algorithms using rich…

Machine Learning · Computer Science 2021-03-12 Aravind Rajeswaran , Igor Mordatch , Vikash Kumar

The rapid advancement of chat-based language models has led to remarkable progress in complex task-solving. However, their success heavily relies on human input to guide the conversation, which can be challenging and time-consuming. This…

Artificial Intelligence · Computer Science 2023-11-03 Guohao Li , Hasan Abed Al Kader Hammoud , Hani Itani , Dmitrii Khizbullin , Bernard Ghanem

In this paper, we study the transmission strategy adaptation problem in an RF-powered cognitive radio network, in which hybrid secondary users are able to switch between the harvest-then-transmit mode and the ambient backscatter mode for…

Networking and Internet Architecture · Computer Science 2018-04-10 Wenbo Wang , Dinh Thai Hoang , Dusit Niyato , Ping Wang , Dong In Kim

When deploying autonomous agents in the real world, we need effective ways of communicating objectives to them. Traditional skill learning has revolved around reinforcement and imitation learning, each with rigid constraints on the format…

Artificial Intelligence · Computer Science 2019-11-21 Mark Woodward , Chelsea Finn , Karol Hausman

This paper is concerned with a Stackelberg stochastic differential game with asymmetric noisy observation, with one follower and one leader. In our model, the follower cannot observe the state process directly, but could observe a noisy…

Optimization and Control · Mathematics 2020-07-14 Yueyang Zheng , Jingtao Shi

This paper investigates a robust incentive Stackelberg stochastic differential game problem for a linear-quadratic mean field system, where the model uncertainty appears in the drift term of the leader's state equation. Moreover, both the…

Optimization and Control · Mathematics 2026-03-31 Na Xiang , Jingtao Shi

Multi-agent coordination under partial observability requires agents to share complementary private information. While recent methods optimize messages for intermediate objectives (e.g., reconstruction accuracy or mutual information),…

Machine Learning · Computer Science 2026-05-14 Benjamin Amoh , Geoffrey Parker , Wesley Marrero

Large scale systems are forecasted to greatly impact our future lives thanks to their wide ranging applications including cooperative robotics, mobility on demand, resource allocation, supply chain management. While technological…

Optimization and Control · Mathematics 2024-12-20 Dario Paccagnan

The aim of this paper is to perform a Stackelberg strategy to control parabolic equations. We have one control, \textit{the leader}, that is responsible for a null controllability property; additionally, we have a control \textit{the…

Optimization and Control · Mathematics 2016-10-20 Víctor Hernández-Santamaría , Luz de Teresa

The success of federated learning (FL) ultimately depends on how strategic participants behave under partial observability, yet most formulations still treat FL as a static optimization problem. We instead view FL deployments as governed…

Machine Learning · Computer Science 2026-03-03 Dongseok Kim , Hyoungsun Choi , Mohamed Jismy Aashik Rasool , Gisung Oh

Federated learning (FL) rests on the notion of training a global model in a decentralized manner. Under this setting, mobile devices perform computations on their local data before uploading the required updates to improve the global model.…

Machine Learning · Computer Science 2020-05-07 Shashi Raj Pandey , Nguyen H. Tran , Mehdi Bennis , Yan Kyaw Tun , Aunas Manzoor , Choong Seon Hong

In this paper, we investigate a new model of a linear-quadratic mean-field stochastic Stackelberg differential game with one leader and two followers, in which the leader is allowed to stop her strategy at a random time. Our overarching…

Optimization and Control · Mathematics 2021-06-08 Zhun Gou , Nan-jing Huang , Ming-hui Wang