English
Related papers

Related papers: A Strong Duality Result for Constrained POMDPs wit…

200 papers

Existing work on linear constrained Markov decision processes (CMDPs) has primarily focused on stochastic settings, where the losses and costs are either fixed or drawn from fixed distributions. However, such formulations are inherently…

Machine Learning · Computer Science 2026-05-13 Kihyun Yu , Seoungbin Bae , Dabeen Lee

This paper studies the utility maximization on the terminal wealth with random endowments and proportional transaction costs. To deal with unbounded random payoffs from some illiquid claims, we propose to work with the acceptable portfolios…

Mathematical Finance · Quantitative Finance 2018-08-27 Erhan Bayraktar , Xiang Yu

We study synthesis problems with constraints in partially observable Markov decision processes (POMDPs), where the objective is to compute a strategy for an agent that is guaranteed to satisfy certain safety and performance specifications.…

We consider a finite-state partially observable Markov decision problem (POMDP) with an infinite horizon and a discounted cost, and we propose a new method for computing a cost function approximation that is based on features and…

Systems and Control · Electrical Eng. & Systems 2025-07-08 Yuchao Li , Kim Hammar , Dimitri Bertsekas

Many realistic decision-making problems in networked scenarios, such as formation control and collaborative task offloading, often involve complicatedly entangled local decisions, which, however, have not been sufficiently investigated yet.…

Optimization and Control · Mathematics 2025-11-20 Dandan Wang , Xuyang Wu , Zichong Ou , Jie Lu

We revisit the duality theorem for multimarginal optimal transportation problems. In particular, we focus on the Coulomb cost. We use a discrete approximation to prove equality of the extremal values and some careful estimates of the…

Analysis of PDEs · Mathematics 2015-05-08 Luigi De Pascale

In this paper we extend the duality theory of the multi-marginal optimal transport problem for cost functions depending on a decreasing function of the distance (not necessarily bounded). This class of cost functions appears in the context…

Analysis of PDEs · Mathematics 2018-05-03 Augusto Gerolin , Anna Kausamo , Tapio Rajala

Minimizing divergence measures under a constraint is an important problem. We derive a sufficient condition that binary divergence measures provide lower bounds for symmetric divergence measures under a given triangular discrimination or…

Information Theory · Computer Science 2022-11-11 Tomohiro Nishiyama

The ability to compute reward-optimal policies for given and known finite Markov decision processes (MDPs) underpins a variety of applications across planning, controller synthesis, and verification. However, we often want policies (1) to…

Logic in Computer Science · Computer Science 2025-11-18 Linus Heck , Filip Macák , Milan Češka , Sebastian Junges

We study the convex duality method for robust utility maximization in the presence of a random endowment. When the underlying price process is a locally bounded semimartingale, we show that the fundamental duality relation holds true for a…

Computational Finance · Quantitative Finance 2015-03-17 Keita Owari

In these notes, we examine certain implications of Sion's minimax theorem for compact quadratically constrained quadratic programs (QCQPs), particularly QCQPs arising in the context of optimizing wave scattering, in relation to Lagrangian…

Optimization and Control · Mathematics 2022-03-04 Sean Molesky , Pengning Chao , Alejandro W. Rodriguez

Dynamics of many supersymmetric monopoles are studied in the low energy approximation. A conjecture for the exact moduli space metric is given for all collections of fundamental monopoles of distinct type, and various partial confirmations…

High Energy Physics - Theory · Physics 2007-05-23 Piljin Yi

A distributed nonsmooth robust resource allocation problem with cardinality constrained uncertainty is investigated in this paper. The global objective is consisted of local objectives, which are convex but nonsmooth. Each agent is…

Optimization and Control · Mathematics 2019-11-05 Yue Wei , Shuxin Ding , Hao Fang , Xianlin Zeng , Qingkai Yang , Bin Xin

We consider the problem of learning the optimal policy for infinite-horizon Markov decision processes (MDPs). For this purpose, some variant of Stochastic Mirror Descent is proposed for convex programming problems with Lipschitz-continuous…

Optimization and Control · Mathematics 2022-03-01 Daniil Tiapkin , Alexander Gasnikov

This paper proposes two nonlinear dynamics to solve constrained distributed optimization problem for resource allocation over a multi-agent network. In this setup, coupling constraint refers to resource-demand balance which is preserved at…

Systems and Control · Electrical Eng. & Systems 2023-10-30 Mohammadreza Doostmohammadian , Alireza Aghasi , Maria Vrakopoulou , Hamid R. Rabiee , Usman A. Khan , Themistoklis Charalambou

Despite the significant progress in multiagent teamwork, existing research does not address the optimality of its prescriptions nor the complexity of the teamwork problem. Without a characterization of the optimality-complexity tradeoffs,…

Artificial Intelligence · Computer Science 2011-06-24 D. V. Pynadath , M. Tambe

We analyze entropic uncertainty relations in a finite dimensional Hilbert space and derive several strong bounds for the sum of two entropies obtained in projective measurements with respect to any two orthogonal bases. We improve the…

Quantum Physics · Physics 2015-06-30 Łukasz Rudnicki , Zbigniew Puchała , Karol Życzkowski

We study an approximation method for partially observed Markov decision processes (POMDPs) with continuous spaces. Belief MDP reduction, which has been the standard approach to study POMDPs requires rigorous approximation methods for…

Optimization and Control · Mathematics 2025-01-20 Ali Devran Kara , Erhan Bayraktar , Serdar Yuksel

We consider a resource allocation problem over an undirected network of agents, where edges of the network define communication links. The goal is to minimize the sum of agent-specific convex objective functions, while the agents' decisions…

Optimization and Control · Mathematics 2019-11-27 Goran Banjac , Felix Rey , Paul Goulart , John Lygeros

Uncertain partially observable Markov decision processes (uPOMDPs) allow the probabilistic transition and observation functions of standard POMDPs to belong to a so-called uncertainty set. Such uncertainty, referred to as epistemic…

Artificial Intelligence · Computer Science 2021-11-02 Murat Cubuktepe , Nils Jansen , Sebastian Junges , Ahmadreza Marandi , Marnix Suilen , Ufuk Topcu