中文
相关论文

相关论文: Proper Policies in Infinite-State Stochastic Short…

200 篇论文

The two-stage stochastic unit commitment problem has become an important tool to support decision-making under uncertainty in power systems. Representing the uncertainty by a large number of scenarios guarantees accurate results but…

最优化与控制 · 数学 2025-12-23 Yannick Werner , Juan Miguel Morales , Salvador Pineda , Line Roald , Sonja Wogrin

We consider a class of two-player zero-sum stochastic games with finite state and compact control spaces, which we call stochastic shortest path (SSP) games. They are undiscounted total cost stochastic dynamic games that have a cost-free…

最优化与控制 · 数学 2014-12-31 Huizhen Yu

Bounded policy iteration is an approach to solving infinite-horizon POMDPs that represents policies as stochastic finite-state controllers and iteratively improves a controller by adjusting the parameters of each node using linear…

人工智能 · 计算机科学 2012-06-18 Eric A. Hansen

The path-integral control, which stems from the stochastic Hamilton-Jacobi-Bellman equation, is one of the methods to control stochastic nonlinear systems. This paper gives a new insight into nonlinear stochastic optimal control problems…

最优化与控制 · 数学 2021-09-14 Jun Ohkubo

Standard algorithms for finding the shortest path in a graph require that the cost of a path be additive in edge costs, and typically assume that costs are deterministic. We consider the problem of uncertain edge costs, with potential…

人工智能 · 计算机科学 2013-02-21 Michael P. Wellman , Matthew Ford , Kenneth Larson

In this paper, we investigate dynamic optimization problems featuring both stochastic control and optimal stopping in a finite time horizon. The paper aims to develop new methodologies, which are significantly different from those of mixed…

投资组合管理 · 定量金融 2014-06-27 Xiongfei Jian , Xun Li , Fahuai Yi

This paper considers the infinite horizon optimal control problem for nonlinear systems. Under the condition of nonlinear controllability of the system to any terminal set containing the origin and forward invariance of the terminal set, we…

最优化与控制 · 数学 2026-02-17 Mohamed Naveed Gul Mohamed , Abhijeet , Aayushman Sharma , Raman Goyal , Suman Chakravorty

Models of many real-life applications, such as queuing models of communication networks or computing systems, have a countably infinite state-space. Algorithmic and learning procedures that have been developed to produce optimal policies…

系统与控制 · 电气工程与系统科学 2024-03-19 Saghar Adler , Vijay Subramanian

An optimal control problem is studied for a linear mean-field stochastic differential equation with a quadratic cost functional. The coefficients and the weighting matrices in the cost functional are all assumed to be deterministic.…

最优化与控制 · 数学 2016-02-26 Xun Li , Jingrui Sun , Jiongmin Yong

This paper studies chance-constrained stochastic optimization problems with finite support. It presents an iterative method that solves reduced-size chance-constrained models obtained by partitioning the scenario set. Each reduced problem…

最优化与控制 · 数学 2024-11-26 Marius Roland , Alexandre Forel , Thibaut Vidal

We study the problem of optimal portfolio selection under stochastic volatility within a continuous time reinforcement learning framework with portfolio constraints. Exploration is modeled through entropy-regularized relaxed controls, where…

数理金融 · 定量金融 2026-04-27 Thai Nguyen , Pertiny Nkuize

We study a constrained stochastic control problem with jumps; the jump times of the controlled process are given by a Poisson process. The cost functional comprises quadratic components for an absolutely continuous control and the…

最优化与控制 · 数学 2013-04-29 Peter Kratz

In this article we approach a class of stochastic reachability problems with state constraints from an optimal control perspective. Preceding approaches to solving these reachability problems are either confined to the deterministic setting…

最优化与控制 · 数学 2017-11-27 Peyman Mohajerin Esfahani , Debasish Chatterjee , John Lygeros

In this paper, we study a Markov decision process with a non-linear discount function and with a Borel state space. We define a recursive discounted utility, which resembles non-additive utility functions considered in a number of models in…

最优化与控制 · 数学 2025-10-16 Nicole Bäuerle , Anna Jaśkiewicz , Andrzej S. Nowak

The paper deals with finite-state Markov decision processes (MDPs) with integer weights assigned to each state-action pair. New algorithms are presented to classify end components according to their limiting behavior with respect to the…

计算机科学中的逻辑 · 计算机科学 2018-05-01 Christel Baier , Nathalie Bertrand , Clemens Dubslaff , Daniel Gburek , Ocan Sankur

In this paper, problems of optimal control are considered where in the objective function, in addition to the control cost there is a tracking term that measures the distance to a desired stationary state. The tracking term is given by some…

最优化与控制 · 数学 2020-06-15 Martin Gugat , Michael Schuster , Enrique Zuazua

Path integral control solves a class of stochastic optimal control problems with a Monte Carlo (MC) method for an associated Hamilton-Jacobi-Bellman (HJB) equation. The MC approach avoids the need for a global grid of the domain of the HJB…

最优化与控制 · 数学 2014-08-26 Insoon Yang , Matthias Morzfeld , Claire J. Tomlin , Alexandre J. Chorin

In a classical optimal stopping problem the aim is to maximize the expected value of a functional of a diffusion evaluated at a stopping time. This note considers optimal stopping problems beyond this paradigm. We study problems in which…

概率论 · 数学 2017-08-04 Vicky Henderson , David Hobson , Matthew Zeng

We study a specific class of finite-horizon mean field optimal stopping problems by means of the dynamic programming approach. In particular, we consider problems where the state process is not affected by the stopping time. Such problems…

最优化与控制 · 数学 2025-03-07 Andrea Cosso , Laura Perelli

This paper studies a class of continuous-time scalar-state stochastic Linear-Quadratic (LQ) optimal control problem with the linear control constraints. Applying the state separation theorem induced from its special structure, we develop…

投资组合管理 · 定量金融 2018-06-12 Weiping Wu , Jianjun Gao , Junguo Lu , Xun Li