English
Related papers

Related papers: A Strong Duality Result for Constrained POMDPs wit…

200 papers

We study a general class of dynamic multi-agent decision problems with asymmetric information and non-strategic agents, which includes dynamic teams as a special case. When agents are non-strategic, an agent's strategy is known to the other…

Multiagent Systems · Computer Science 2018-12-05 Hamidreza Tavafoghi , Yi Ouyang , Demosthenis Teneketzis

We consider finite model approximations of discrete-time partially observed Markov decision processes (POMDPs) under the discounted cost criterion. After converting the original partially observed stochastic control problem to a fully…

Systems and Control · Computer Science 2017-10-20 Naci Saldi , Serdar Yüksel , Tamás Linder

We consider the problem of finding good finite-horizon policies for POMDPs under the expected reward metric. The policies considered are {em free finite-memory policies with limited memory}; a policy is a mapping from the space of…

Artificial Intelligence · Computer Science 2013-01-30 Christopher Lusena , Tong Li , Shelia Sittinger , Chris Wells , Judy Goldsmith

Planning for distributed agents with partial state information is considered from a decision- theoretic perspective. We describe generalizations of both the MDP and POMDP models that allow for decentralized control. For even a small number…

Artificial Intelligence · Computer Science 2013-01-18 Daniel S Bernstein , Shlomo Zilberstein , Neil Immerman

In this paper we consider the problem of distributed nonlinear optimisation of a separable convex cost function over a graph subject to cone constraints. We show how to generalise, using convex analysis, monotone operator theory and…

Distributed, Parallel, and Cluster Computing · Computer Science 2024-05-16 Richard Heusdens , Guoqiang Zhang

A finite horizon optimal tracking problem is considered for linear dynamical systems subject to parametric uncertainties in the state-space matrices and exogenous disturbances. A suboptimal solution is proposed using a model predictive…

Optimization and Control · Mathematics 2022-02-08 Anilkumar Parsi , Andrea Iannelli , Roy S. Smith

We study a class of convex-concave min-max problems in which the coupled component of the objective is linear in at least one of the two decision vectors. We identify such problem structure as interpolating between the bilinearly and…

Optimization and Control · Mathematics 2025-07-10 Ronak Mehta , Jelena Diakonikolas , Zaid Harchaoui

We consider cooperative multi-agent consensus optimization problems over both static and time-varying communication networks, where only local communications are allowed. The objective is to minimize the sum of agent-specific possibly…

Optimization and Control · Mathematics 2017-06-27 Erfan Yazdandoost Hamedani , Necdet Serhat Aybat

Correlated equilibria enable a coordinator to influence the self-interested agents by recommending actions that no player has an incentive to deviate from. However, the effectiveness of this mechanism relies on accurate knowledge of the…

Computer Science and Game Theory · Computer Science 2026-05-18 Jaehan Im , Ufuk Topcu , David Fridovich-Keil

This paper studies discrete-time average-cost infinite-horizon Markov decision processes (MDPs) with Borel state and action sets. It introduces new sufficient conditions for { the} validity of optimality inequalities and optimality…

Optimization and Control · Mathematics 2025-01-28 Eugene A. Feinberg , Pavlo O. Kasyanov , Liliia S. Paliichuk

Partially observable Markov decision processes (POMDPs) provide an elegant mathematical framework for modeling complex decision and planning problems in stochastic domains in which states of the system are observable only indirectly, via a…

Artificial Intelligence · Computer Science 2011-06-02 M. Hauskrecht

This article is devoted to investigate a nonsmooth/nonconvex uncertain multiobjective optimization problem with composition fields (CUP) for brevity) over arbitrary Asplund spaces. Employing some advanced techniques of variational analysis…

Optimization and Control · Mathematics 2024-03-12 Maryam Saadati , Morteza Oveisiha

Autonomous systems often have logical constraints arising, for example, from safety, operational, or regulatory requirements. Such constraints can be expressed using temporal logic specifications. The system state is often partially…

Artificial Intelligence · Computer Science 2024-06-21 Krishna C. Kalagarla , Dhruva Kartik , Dongming Shen , Rahul Jain , Ashutosh Nayyar , Pierluigi Nuzzo

Doubly robust methods hold considerable promise for off-policy evaluation in Markov decision processes (MDPs) under sequential ignorability: They have been shown to converge as $1/\sqrt{T}$ with the horizon $T$, to be statistically…

Machine Learning · Statistics 2025-09-30 Mohammad Mehrabi , Stefan Wager

Sequential incentive marketing is an important approach for online businesses to acquire customers, increase loyalty and boost sales. How to effectively allocate the incentives so as to maximize the return (e.g., business objectives) under…

Artificial Intelligence · Computer Science 2023-03-03 Shuai Xiao , Le Guo , Zaifan Jiang , Lei Lv , Yuanbo Chen , Jun Zhu , Shuang Yang

This paper investigates the limit behavior of Markov Decision Processes (MDPs) made of independent particles evolving in a common environment, when the number of particles goes to infinity. In the finite horizon case or with a discounted…

Probability · Mathematics 2009-06-10 Nicolas Gast , Bruno Gaujal

This paper is concerned with the properties of the sets of strategic measures induced by admissible team policies in decentralized stochastic control and the convexity properties in dynamic team problems. To facilitate a convex analytical…

Optimization and Control · Mathematics 2016-11-01 Serdar Yüksel , Naci Saldi

Cooperatively optimizing a vast number of agents that are connected over a large-scale network brings unprecedented scalability challenges. This paper revolves around problems optimizing coupled objective functions under coupled…

Optimization and Control · Mathematics 2020-10-14 Xiang Huo , Mingxi Liu

We present a numerical iterative optimization algorithm for the minimization of a cost function consisting of a linear combination of three convex terms, one of which is differentiable, a second one is prox-simple and the third one is the…

Optimization and Control · Mathematics 2024-10-04 Ignace Loris , Simone Rebegoldi

Convex duality for two two different super--replication problems in a continuous time financial market with proportional transaction cost is proved. In this market, static hedging in a finite number of options, in addition to usual dynamic…

Mathematical Finance · Quantitative Finance 2015-10-20 Yan Dolinsky , H. Mete Soner
‹ Prev 1 4 5 6 7 8 10 Next ›