中文
相关论文

相关论文: Proper Policies in Infinite-State Stochastic Short…

200 篇论文

We consider a broad class of dynamic programming (DP) problems that involve a partially linear structure and some positivity properties in their system equation and cost function. We address deterministic and stochastic problems, possibly…

最优化与控制 · 数学 2026-04-21 Yuchao Li , Dimitri Bertsekas

The main objective of this paper is to develop a martingale-type solution to optimal consumption--investment choice problems ([Merton, 1969] and [Merton, 1971]) under time-varying incomplete preferences driven by externalities such as…

数理金融 · 定量金融 2025-01-14 Weixuan Xia

The infinite horizon setting is widely adopted for problems of reinforcement learning (RL). These invariably result in stationary policies that are optimal. In many situations, finite horizon control problems are of interest and for such…

机器学习 · 计算机科学 2025-03-21 Soumyajit Guin , Shalabh Bhatnagar

We study the optimal control of path-dependent piecewise deterministic processes. An appropriate dynamic programming principle is established. We prove that the associated value function is the unique minimax solution of the corresponding…

概率论 · 数学 2025-10-28 Elena Bandini , Christian Keller

Policy optimization methods are one of the most widely used classes of Reinforcement Learning (RL) algorithms. However, theoretical understanding of these methods remains insufficient. Even in the episodic (time-inhomogeneous) tabular…

机器学习 · 计算机科学 2022-12-06 Tianhao Wu , Yunchang Yang , Han Zhong , Liwei Wang , Simon S. Du , Jiantao Jiao

Learned action policies are increasingly popular in sequential decision-making, but suffer from a lack of safety guarantees. Recent work introduced a pipeline for testing the safety of such policies under initial-state and action-outcome…

人工智能 · 计算机科学 2026-03-17 Johannes Schmalz , Chaahat Jain

We propose a machine learning algorithm for solving finite-horizon stochastic control problems based on a deep neural network representation of the optimal policy functions. The algorithm has three features: (1) It can solve…

综合经济学 · 经济学 2024-12-09 Xianhua Peng , Steven Kou , Lekang Zhang

Probabilistic guarantees of safety and performance are important in constrained dynamical systems with stochastic uncertainty. We consider the stochastic reachability problem, which maximizes the probability that the state remains within…

最优化与控制 · 数学 2020-12-01 Abraham P. Vinod , Meeko M. K. Oishi

We study an agent's lifecycle portfolio choice problem with stochastic labor income, borrowing constraints and a finite retirement date. Similarly to arXiv:2002.00201, wages evolve in a path-dependent way, but the presence of a finite…

最优化与控制 · 数学 2024-02-27 Sara Biagini , Enrico Biffis , Fausto Gozzi , Margherita Zanella

We present a dynamic programming-based solution to a stochastic optimal control problem up to a hitting time for a discrete-time Markov control process. Firstly, we determine an optimal control policy to steer the process toward a compact…

最优化与控制 · 数学 2009-09-28 Debasish Chatterjee , Eugenio Cinquemani , Giorgos Chaloulos , John Lygeros

Optimization of decision problems in stochastic environments is usually concerned with maximizing the probability of achieving the goal and minimizing the expected episode length. For interacting agents in time-critical applications,…

人工智能 · 计算机科学 2007-05-23 Balint Takacs , Istvan Szita , Andras Lorincz

We study a fundamental stochastic selection problem involving $n$ independent random variables, each of which can be queried at some cost. Given a tolerance level $\delta$, the goal is to find a value that is $\delta$-approximately minimum…

数据结构与算法 · 计算机科学 2025-04-25 Hessa Al-Thani , Viswanath Nagarajan

In this article, we provide a numerical method based on fitted finite volume method to approximate the Hamilton-Jacobi-Bellman (HJB) equation coming from stochastic optimal control problems. The computational challenge is due to the nature…

数值分析 · 数学 2020-02-21 Christelle Dleuna Nyoumbi , Antoine Tambue

In this paper, we consider a class of continuous-time, continuous-space stochastic optimal control problems. Building upon recent advances in Markov chain approximation methods and sampling-based algorithms for deterministic path planning,…

机器人学 · 计算机科学 2012-02-27 Vu Anh Huynh , Sertac Karaman , Emilio Frazzoli

In this paper, we demonstrate that policy iteration, introduced in the context of HJB equations in [Forsyth & Labahn, 2007], is an extremely simple generic algorithm for solving linear complementarity problems resulting from the finite…

计算金融 · 定量金融 2012-06-19 Christoph Reisinger , Jan Hendrik Witte

We consider a pathwise stochastic optimal control problem and study the associated (not necessarily adapted) Hamilton-Jacobi-Bellman stochastic partial differential equation. We show that the value process is the unique solution of this…

概率论 · 数学 2023-11-02 Neeraj Bhauryal , Ana Bela Cruzeiro , Carlos Oliveira

In this work we address the problem of finding feasible policies for Constrained Markov Decision Processes under probability one constraints. We argue that stationary policies are not sufficient for solving this problem, and that a rich…

机器学习 · 计算机科学 2023-02-14 Agustin Castellano , Hancheng Min , Juan Bazerque , Enrique Mallada

In this study, we consider an optimal control problem driven by a stochastic differential system with a stopping time terminal cost functional. We establish the stochastic maximum principle for this new kind of an optimal control problem by…

最优化与控制 · 数学 2018-12-11 Shuzhen Yang

Time-consistency is an essential requirement in risk sensitive optimal control problems to make rational decisions. An optimization problem is time consistent if its solution policy does not depend on the time sequence of solving the…

最优化与控制 · 数学 2015-03-26 Yinlam Chow , Marco Pavone

We revisit the incremental autonomous exploration problem proposed by Lim & Auer (2012). In this setting, the agent aims to learn a set of near-optimal goal-conditioned policies to reach the $L$-controllable states: states that are…

机器学习 · 计算机科学 2022-05-24 Haoyuan Cai , Tengyu Ma , Simon Du
‹ 上一页 1 8 9 10 下一页 ›