中文
相关论文

相关论文: Unbounded Markov Dynamic Programming with Weighted…

200 篇论文

We address the stability problem for linear switching systems with mode-dependent restrictions on the switching intervals. Their lengths can be bounded as from below (the guaranteed dwell-time) as from above. The upper bounds make this…

最优化与控制 · 数学 2022-06-01 Vladimir Yu. Protasov , Rinat Kamalov

The article poses a general model for optimal control subject to information constraints, motivated in part by recent work of Sims and others on information-constrained decision-making by economic agents. In the average-cost optimal control…

最优化与控制 · 数学 2016-02-24 Ehsan Shafieepoorfard , Maxim Raginsky , Sean P. Meyn

This paper studies function approximation for finite horizon discrete time Markov decision processes under certain convexity assumptions. Uniform convergence of these approximations on compact sets is proved under several sampling schemes…

最优化与控制 · 数学 2018-02-21 Jeremy Yee

In this work we address the problem of finding feasible policies for Constrained Markov Decision Processes under probability one constraints. We argue that stationary policies are not sufficient for solving this problem, and that a rich…

机器学习 · 计算机科学 2023-02-14 Agustin Castellano , Hancheng Min , Juan Bazerque , Enrique Mallada

This study considers an optimal reinsurance, investment, and dividend strategy control problem for insurance companies in a regulated Markov regime-switching environment, intending to maximize long-run average reward. Unlike existing single…

最优化与控制 · 数学 2025-12-18 Lingjia Zeng , Manman Li

This paper extends the core results of discrete time infinite horizon dynamic programming to the case of state-dependent discounting. We obtain a condition on the discount factor process under which all of the standard optimality results…

综合经济学 · 经济学 2020-10-15 John Stachurski , Junnan Zhang

We prove a functional limit theorem for Markov chains that, in each step, move up or down by a possibly state dependent constant with probability $1/2$, respectively. The theorem entails that the law of every one-dimensional regular…

概率论 · 数学 2020-05-13 Stefan Ankirchner , Thomas Kruse , Mikhail Urusov

We introduce a contractive abstract dynamic programming framework and related policy iteration algorithms, specifically designed for sequential zero-sum games and minimax problems with a general structure. Aside from greater generality, the…

计算机科学与博弈论 · 计算机科学 2021-10-22 Dimitri Bertsekas

This paper proposes a computationally tractable algorithm for learning infinite-horizon average-reward linear Markov decision processes (MDPs) and linear mixture MDPs under the Bellman optimality condition. While guaranteeing computational…

机器学习 · 计算机科学 2024-09-25 Woojin Chae , Dabeen Lee

We derive a variational formula for the optimal growth rate of reward in the infinite horizon risk-sensitive control problem for discrete time Markov decision processes with compact metric state and action spaces, extending a formula of…

最优化与控制 · 数学 2015-01-06 Venkatachalam Anantharam , Vivek Shripad Borkar

We propose a novel randomized linear programming algorithm for approximating the optimal policy of the discounted Markov decision problem. By leveraging the value-policy duality and binary-tree data structures, the algorithm adaptively…

最优化与控制 · 数学 2019-06-04 Mengdi Wang

In this paper, we present an optimal control problem for stochastic differential games under Markov regime-switching forward-backward stochastic differential equations with jumps and partial information. First, we prove a sufficient maximum…

最优化与控制 · 数学 2014-10-14 Olivier Menoukeu Pamen , Romual Herve Momeya

The first motivation of our paper is to explore further the idea that, in risk control problems, it may be profitable to base decisions both on the position of the underlying process Xt and on its supremum Xt := sup 0$\le$s$\le$t Xs.…

最优化与控制 · 数学 2019-11-15 Florin Avram , Dan Goreac

In this paper we establish spatial central limit theorems for a large class of supercritical branching Markov processes with general spatial-dependent branching mechanisms. These are generalizations of the spatial central limit theorems…

概率论 · 数学 2013-05-06 Y. -X. Ren , R. Song , R. Zhang

This paper analyzes finite state Markov Decision Processes (MDPs) with uncertain parameters in compact sets and re-examines results from robust MDP via set-based fixed point theory. To this end, we generalize the Bellman and policy…

机器学习 · 计算机科学 2023-08-09 Sarah H. Q. Li , Assalé Adjé , Pierre-Loïc Garoche , Behçet Açıkmeşe

Previous work has separately addressed different forms of action, state and action-state entropy regularization, pure exploration and space occupation. These problems have become extremely relevant for regularization, generalization,…

机器学习 · 计算机科学 2023-02-03 Dmytro Grytskyy , Jorge Ramírez-Ruiz , Rubén Moreno-Bote

The optimization problems with simple bounds are an important class of problems. To facilitate the computation of such problems, an unconstrained-like dynamic method, motivated by the Lyapunov control principle, is proposed. This method…

最优化与控制 · 数学 2021-10-19 Sheng Zhang , Xin Du , Fang-Fang Hu , Jiang-Tao Huang

In this paper, we consider a large class of constrained non-cooperative stochastic Markov games with countable state spaces and discounted cost criteria. In one-player case, i.e., constrained discounted Markov decision models, it is…

最优化与控制 · 数学 2021-12-16 Anna Jaśkiewicz , Andrzej S. Nowak

Two algorithms are proposed, analyzed, and tested for solving continuous optimization problems with nonlinear equality constraints. Each is an extension of a stochastic momentum-based method from the unconstrained setting to the setting of…

最优化与控制 · 数学 2026-01-21 Qi Wang , Christian Piermarini , Yunlang Zhu , Frank E. Curtis

In this work, we investigate the optimal control problem for continuous-time Markov decision processes with the random impact of the environment. We provide conditions to show the existence of optimal controls under finite-horizon criteria.…

最优化与控制 · 数学 2020-06-23 Jinghai Shao , Kun Zhao