English
Related papers

Related papers: A Strong Duality Result for Constrained POMDPs wit…

200 papers

Partially-observable Markov decision processes (POMDPs) with discounted-sum payoff are a standard framework to model a wide range of problems related to decision making under uncertainty. Traditionally, the goal has been to obtain policies…

Artificial Intelligence · Computer Science 2018-05-01 Krishnendu Chatterjee , Adrián Elgyütt , Petr Novotný , Owen Rouillé

Privacy has traditionally been a major motivation for distributed problem solving. Distributed Constraint Satisfaction Problem (DisCSP) as well as Distributed Constraint Optimization Problem (DCOP) are fundamental models used to solve…

In POMDPs, information about the hidden state, delivered through observations, is both valuable to the agent, allowing it to base its actions on better informed internal states, and a "curse", exploding the size and diversity of the…

Machine Learning · Computer Science 2015-12-31 Roy Fox , Naftali Tishby

We consider the problem of optimizing the expected logarithmic utility of the value of a portfolio in a binomial model with proportional transaction costs with a long time horizon. By duality methods, we can find expressions for the…

Portfolio Management · Quantitative Finance 2012-09-25 Christian Bayer , Bezirgen Veliyev

This paper considers the problem of finding a solution to the finite horizon constrained Markov decision processes (CMDP) where the objective as well as constraints are sum of additive and multiplicative utilities. Towards solving this, we…

Optimization and Control · Mathematics 2023-03-16 Uday Kumar M , Sanjay P Bhat , Veeraruna Kavitha , Nandyala Hemachandra

Specialization and hierarchical organization are important features of efficient collaboration in economical, artificial, and biological systems. Here, we investigate the hypothesis that both features can be explained by the fact that each…

Multiagent Systems · Computer Science 2018-09-19 Sebastian Gottwald , Daniel A. Braun

In this work, we consider solving a distributed optimization problem in a multi-agent network with multiple clusters. In each cluster, the involved agents cooperatively optimize a separable composite function with a common decision…

Optimization and Control · Mathematics 2022-03-03 Jianzheng Wang , Guoqiang Hu

We propose a pseudo-market solution to resource allocation problems subject to constraints. Our treatment of constraints is general: including bihierarchical constraints due to considerations of diversity in school choice, or scheduling in…

Theoretical Economics · Economics 2020-11-09 Federico Echenique , Antonio Miralles , Jun Zhang

Partially observable Markov Decision Processes (POMDPs) are a standard model for agents making decisions in uncertain environments. Most work on POMDPs focuses on synthesizing strategies based on the available capabilities. However, system…

Artificial Intelligence · Computer Science 2024-07-12 Alyzia-Maria Konsta , Alberto Lluch Lafuente , Christoph Matheja

The problem of constrained Markov decision process is considered. An agent aims to maximize the expected accumulated discounted reward subject to multiple constraints on its costs (the number of constraints is relatively small). A new dual…

Optimization and Control · Mathematics 2022-10-21 Egor Gladin , Maksim Lavrik-Karmazin , Karina Zainullina , Varvara Rudenko , Alexander Gasnikov , Martin Takáč

We establish a variant of Monge--Kantorovich duality for a constrained optimal transport problem with a continuum of agents, a finite set of alternatives, and general linear constraints. As an application, we revisit the large-market model…

Theoretical Economics · Economics 2026-04-06 Koji Yokote

A problem of the erroneous duality gap caused by the presence of symmetries is solved in this paper utilizing point group theory. The optimization problems are first divided into two classes based on their predisposition to suffer from this…

Computational Physics · Physics 2021-06-23 Miloslav Capek , Lukas Jelinek , Michal Masek

We consider nonlinear model predictive control (MPC) with multiple competing cost functions. This leads to the formulation of multiobjective optimal control problems (MO OCPs). Since the design of MPC algorithms for directly solving…

Optimization and Control · Mathematics 2022-11-23 Lars Grüne , Lisa Krügel , Matthias A. Müller

The incorporation of macro-actions (temporally extended actions) into multi-agent decision problems has the potential to address the curse of dimensionality associated with such decision problems. Since macro-actions last for stochastic…

Artificial Intelligence · Computer Science 2019-05-30 Kunal Menda , Yi-Chun Chen , Justin Grana , James W. Bono , Brendan D. Tracey , Mykel J. Kochenderfer , David Wolpert

We consider the problem of finding the best memoryless stochastic policy for an infinite-horizon partially observable Markov decision process (POMDP) with finite state and action spaces with respect to either the discounted or mean reward…

Optimization and Control · Mathematics 2022-05-02 Johannes Müller , Guido Montúfar

Decentralized partially observable Markov decision processes with communication (Dec-POMDP-Com) provide a framework for multiagent decision making under uncertainty, but the NEXP-complete complexity for finite-horizon problems renders…

Multiagent Systems · Computer Science 2025-11-18 Dylan M. Asmar , Mykel J. Kochenderfer

The aim of this short note is to establish a limit theorem for the optimal trading strategies in the setup of the utility maximization problem with proportional transaction costs. This limit theorem resolves the open question from [4]. The…

Mathematical Finance · Quantitative Finance 2021-09-28 Erhan Bayraktar , Christoph Czichowsky , Leonid Dolinskyi , Yan Dolinsky

Recent advances in our understanding of higher derived limits carry multiple implications in the fields of condensed and pyknotic mathematics, as well as for the study of strong homology. These implications are thematically diverse,…

Algebraic Topology · Mathematics 2025-08-12 Jeffrey Bergfalk , Chris Lambie-Hanson

This paper presents a pseudo-spectral method for Dynamic Optimization Problems (DOPs) that allows for tight polynomial bounds to be achieved via flexible sub-intervals. The proposed method not only rigorously enforces inequality…

Optimization and Control · Mathematics 2026-04-08 Eduardo M. G. Vila , Eric C. Kerrigan , Paul Bruce

Multi-environment POMDPs (ME-POMDPs) extend standard POMDPs with discrete model uncertainty. ME-POMDPs represent a finite set of POMDPs that share the same state, action, and observation spaces, but may arbitrarily vary in their transition,…

Artificial Intelligence · Computer Science 2025-10-29 Eline M. Bovy , Caleb Probine , Marnix Suilen , Ufuk Topcu , Nils Jansen