English
Related papers

Related papers: Multiple Approximate-Response Agents (MARA): Fast …

200 papers

Multidimensional data have become ubiquitous and are frequently encountered in situations where the information is aggregated over multiple data atoms. The aggregation can be over time or other features, such as geographical location. We…

Machine Learning · Computer Science 2020-04-14 Faisal M. Almutairi , Charilaos I. Kanatsoulis , Nicholas D. Sidiropoulos

Constrained multiagent reinforcement learning (C-MARL) is gaining importance as MARL algorithms find new applications in real-world systems ranging from energy systems to drone swarms. Most C-MARL algorithms use a primal-dual approach to…

Systems and Control · Electrical Eng. & Systems 2023-04-28 Daniel Tabas , Ahmed S. Zamzam , Baosen Zhang

The problem of constrained Markov decision process is considered. An agent aims to maximize the expected accumulated discounted reward subject to multiple constraints on its costs (the number of constraints is relatively small). A new dual…

Optimization and Control · Mathematics 2022-10-21 Egor Gladin , Maksim Lavrik-Karmazin , Karina Zainullina , Varvara Rudenko , Alexander Gasnikov , Martin Takáč

In tabular multi-agent reinforcement learning with average-cost criterion, a team of agents sequentially interacts with the environment and observes local incentives. We focus on the case that the global reward is a sum of local rewards,…

Optimization and Control · Mathematics 2021-10-26 Alec Koppel , Amrit Singh Bedi , Bhargav Ganguly , Vaneet Aggarwal

We propose decentralized primal-dual methods for cooperative multi-agent consensus optimization problems over both static and time-varying communication networks, where only local communications are allowed. The objective is to minimize the…

Optimization and Control · Mathematics 2022-02-23 Erfan Yazdandoost Hamedani , Necdet Serhat Aybat

In convex optimization, duality theory can sometimes lead to simpler solution methods than those resulting from direct primal analysis. In this paper, this principle is applied to a class of composite variational problems arising in…

Optimization and Control · Mathematics 2010-06-22 Patrick L. Combettes , Dinh Dung , Bang Cong Vu

We present two new methods for multivariate exponential analysis. In [7], we developed a new algorithm for reconstruction of univariate exponential sums by exploiting the rational structure of their Fourier coefficients and reconstructing…

Numerical Analysis · Mathematics 2025-04-29 Nadiia Derevianko , Lennart Aljoscha Hübner

We consider a class of multi-agent cooperative consensus optimization problems with local nonlinear convex constraints where only those agents connected by an edge can directly communicate, hence, the optimal consensus decision lies in the…

Optimization and Control · Mathematics 2023-02-23 Nazanin Abolfazli , Afrooz Jalilzadeh , Erfan Yazdandoost Hamedani

Heterogeneous networks comprise agents with varying capabilities in terms of computation, storage, and communication. In such settings, it is crucial to factor in the operating characteristics in allowing agents to choose appropriate…

Optimization and Control · Mathematics 2022-09-07 Yichuan Li , Petros Voulgaris , Nikolaos M. Freris

We introduce a new framework for optimal routing and arbitrage in AMM driven markets. This framework improves on the original best-practice convex optimization by restricting the search to the boundary of the optimal space. We can…

Mathematical Finance · Quantitative Finance 2025-02-13 Stefan Loesch , Mark Bentley Richardson

We consider a class of multi-agent optimization problems, where each agent has a local objective function that depends on its own decision variables and the aggregate of others, and is willing to cooperate with other agents to minimize the…

Systems and Control · Electrical Eng. & Systems 2021-10-04 Yuanhanqing Huang , Jianghai Hu

This study investigates how Multi-Agent Reinforcement Learning (MARL) can improve dynamic pricing strategies in supply chains, particularly in contexts where traditional ERP systems rely on static, rule-based approaches that overlook…

Machine Learning · Computer Science 2025-07-04 Thomas Hazenberg , Yao Ma , Seyed Sahand Mohammadi Ziabari , Marijn van Rijswijk

Demand response (DR) leverages demand-side flexibility, offering a promising approach to enhance market conditions like mitigating wholesale price spikes. However, poorly chosen DR locations can inadvertently increase electricity prices.…

Systems and Control · Electrical Eng. & Systems 2024-08-06 Yufan Zhang , Honglin Wen , Tao Feng , Yize Chen

Offline Reinforcement Learning (RL) aims to learn a near-optimal policy from a fixed dataset of transitions collected by another policy. This problem has attracted a lot of attention recently, but most existing methods with strong…

Machine Learning · Computer Science 2023-05-23 Germano Gabbianelli , Gergely Neu , Nneka Okolo , Matteo Papini

Primal-Dual Interior-Point methods are capable of solving constrained convex optimization problems to tight tolerances in a fast and robust manner. The derivatives of the primal-dual solution with respect to the problem matrices can be…

Optimization and Control · Mathematics 2024-06-21 Kevin Tracy , Zachary Manchester

We consider a general class of two-stage distributionally robust optimization (DRO) problems where the ambiguity set is constrained by fixed marginal probability laws that are not necessarily discrete. We derive primal and dual formulations…

Optimization and Control · Mathematics 2025-10-17 Ariel Neufeld , Qikun Xiang

Reward models are pivotal for aligning Large Language Models (LLMs) with human preferences. Existing approaches face two key limitations: Discriminative reward models require large-scale annotated data, as they cannot exploit the preference…

Computation and Language · Computer Science 2026-02-03 Yongfu Xue

This paper studies efficient distributed optimization methods for multi-agent networks. Specifically, we consider a convex optimization problem with a globally coupled linear equality constraint and local polyhedra constraints, and develop…

Systems and Control · Computer Science 2016-11-15 Tsung-Hui Chang

In contrast with many other convex optimization classes, state-of-the-art semidefinite programming solvers are yet unable to efficiently solve large scale instances. This work aims to reduce this scalability gap by proposing a novel…

Optimization and Control · Mathematics 2018-12-20 Mario Souto , Joaquim D. Garcia , Alvaro Veiga

Optimizing the advertiser's cumulative value of winning impressions under budget constraints poses a complex challenge in online advertising, under the paradigm of AI-Generated Bidding (AIGB). Advertisers often have personalized objectives…

Artificial Intelligence · Computer Science 2026-01-22 Mingxuan Song , Yusen Huo , Bohan Zhou , Shenglin Yin , Zhen Xiao , Jieyi Long , Zhilin Zhang , Chuan Yu