中文
相关论文

相关论文: Online distributed algorithms for mixed equilibriu…

200 篇论文

In this paper, we focus on the question of the extent to which online learning can benefit from distributed computing. We focus on the setting in which $N$ agents online-learn cooperatively, where each agent only has access to its own data.…

机器学习 · 计算机科学 2019-08-17 Hua Ouyang , Alexander Gray

We study agents communicating over an underlying network by exchanging messages, in order to optimize their individual regret in a common nonstochastic multi-armed bandit problem. We derive regret minimization algorithms that guarantee for…

机器学习 · 计算机科学 2019-11-19 Yogev Bar-On , Yishay Mansour

As e-commerce expands, delivering real-time personalized recommendations from vast catalogs poses a critical challenge for retail platforms. Maximizing revenue requires careful consideration of both individual customer characteristics and…

信息检索 · 计算机科学 2026-02-16 Seong Jin Lee , Will Wei Sun , Yufeng Liu

In this paper we consider a general, challenging distributed optimization set-up arising in several important network control applications. Agents of a network want to minimize the sum of local cost functions, each one depending on a local…

系统与控制 · 计算机科学 2018-06-15 Ivano Notarnicola , Giuseppe Notarstefano

The dueling bandit problem, an essential variation of the traditional multi-armed bandit problem, has become significantly prominent recently due to its broad applications in online advertising, recommendation systems, information…

机器学习 · 计算机科学 2025-04-08 Bongsoo Yi , Yue Kang , Yao Li

In this work we consider a generalization of the well-known multivehicle routing problem: given a network, a set of agents occupying a subset of its nodes, and a set of tasks, we seek a minimum cost sequence of movements subject to the…

分布式、并行与集群计算 · 计算机科学 2024-02-27 Jamison W. Weber , Dhanush R. Giriyan , Devendra R. Parkar , Dimitri P. Bertsekas , Andréa W. Richa

To deal with changing environments, a new performance measure -- adaptive regret, defined as the maximum static regret over any interval, was proposed in online learning. Under the setting of online convex optimization, several algorithms…

机器学习 · 计算机科学 2021-05-17 Lijun Zhang , Guanghui Wang , Wei-Wei Tu , Zhi-Hua Zhou

Constrained submodular set function maximization problems often appear in multi-agent decision-making problems with a discrete feasible set. A prominent example is the problem of multi-agent mobile sensor placement over a discrete domain.…

最优化与控制 · 数学 2020-12-01 Navid Rezazadeh , Solmaz S. Kia

This paper proposes two nonlinear dynamics to solve constrained distributed optimization problem for resource allocation over a multi-agent network. In this setup, coupling constraint refers to resource-demand balance which is preserved at…

系统与控制 · 电气工程与系统科学 2023-10-30 Mohammadreza Doostmohammadian , Alireza Aghasi , Maria Vrakopoulou , Hamid R. Rabiee , Usman A. Khan , Themistoklis Charalambou

We study nonconvex distributed optimization in multiagent networks where the communications between nodes is modeled as a time-varying sequence of arbitrary digraphs. We introduce a novel broadcast-based distributed algorithmic framework…

分布式、并行与集群计算 · 计算机科学 2016-12-16 Ying Sun , Gesualdo Scutari , Daniel Palomar

We study distributed optimization problems over a network when the communication between the nodes is constrained, and so information that is exchanged between the nodes must be quantized. This imperfect communication poses a fundamental…

最优化与控制 · 数学 2018-10-30 Thinh T. Doan , Siva Theja Maguluri , Justin Romberg

This paper studies the decentralized online convex optimization problem for heterogeneous linear multi-agent systems. Agents have access to their time-varying local cost functions related to their own outputs, and there are also…

最优化与控制 · 数学 2022-07-05 Yang Yu , Xiuxian Li , Li Li , Lihua Xie

In this paper, we propose a new framework to study distributed optimization problems with stochastic gradients by employing a multi-agent system with continuous-time dynamics. Here the goal of the agents is to cooperatively minimize the sum…

系统与控制 · 电气工程与系统科学 2026-02-10 Jianhua Sun , Kaihong Lu , Xin Yu

Multi-agent optimization problems with many objective functions have drawn much interest over the past two decades. Many works on the subject minimize the sum of objective functions, which implicitly carries a decision about the problem…

系统与控制 · 电气工程与系统科学 2020-03-05 Maude J. Blondin , Matthew Hale

Classical paradigms for distributed learning, such as federated or decentralized gradient descent, employ consensus mechanisms to enforce homogeneity among agents. While these strategies have proven effective in i.i.d. scenarios, they can…

机器学习 · 计算机科学 2023-04-18 Shreya Wadehra , Roula Nassif , Stefan Vlaski

This paper considers a distributed multi-agent optimization problem, with the global objective consisting of the sum of local objective functions of the agents. The agents solve the optimization problem using local computation and…

分布式、并行与集群计算 · 计算机科学 2017-11-07 Shripad Gade , Nitin H. Vaidya

Bilevel optimization have gained growing interests, with numerous applications found in meta learning, minimax games, reinforcement learning, and nested composition optimization. This paper studies the problem of distributed bilevel…

机器学习 · 统计学 2022-06-23 Shuoguang Yang , Xuezhou Zhang , Mengdi Wang

In this paper, we consider a distributed constrained optimization problem with delayed subgradient information over the time-varying communication network, where each agent can only communicate with its neighbors and the communication…

最优化与控制 · 数学 2021-06-16 Jie Liu , Zhan Yu , Daniel W. C. Ho

The distributed coordination problem of multi-agent systems is addressed in this paper under the assumption of intermittent communication between agents in the presence of time-varying communication delays. Specifically, we consider the…

最优化与控制 · 数学 2016-09-20 Abdelkader Abdessameud , Ilia G. Polushin , Abdelhamid Tayebi

We consider the problem where M agents collaboratively interact with an instance of a stochastic K-armed contextual bandit, where K>>M. The goal of the agents is to simultaneously minimize the cumulative regret over all the agents over a…

机器学习 · 计算机科学 2022-11-16 Jiabin Lin , Shana Moothedath