English
Related papers

Related papers: Resolvent-Techniques For Multiple Exercise Problem…

200 papers

In many engineered systems, optimization is used for decision making at time-scales ranging from real-time operation to long-term planning. This process often involves solving similar optimization problems over and over again with slightly…

Optimization and Control · Mathematics 2019-01-18 Sidhant Misra , Line Roald , Yeesian Ng

In the literature on optimal stopping, the problem of maximizing the expected discounted reward over all stopping times has been explicitly solved for some special reward functions (including $(x^+)^{\nu}$, $(e^x-K)^+$, $(K-e^{-x})^+$,…

Probability · Mathematics 2017-10-13 Yi-Shen Lin , Yi-Ching Yao

Many potential applications of reinforcement learning (RL) require guarantees that the agent will perform well in the face of disturbances to the dynamics or reward function. In this paper, we prove theoretically that maximum entropy…

Machine Learning · Computer Science 2022-05-06 Benjamin Eysenbach , Sergey Levine

In this paper, the reinforcement learning (RL)-based optimal control problem is studied for multiplicative-noise systems, where input delay is involved and partial system dynamics is unknown. To solve a variant of Riccati-ZXL equations,…

Optimization and Control · Mathematics 2023-01-10 Hongxia Wang , Fuyu Zhao , Zhaorong Zhang , Juanjuan Xu , Xun Li

The purpose of this paper is to consider the exit-time problem for a finite-range Markov jump process, i.e, the distance the particle can jump is bounded independent of its location. Such jump diffusions are expedient models for anomalous…

Probability · Mathematics 2015-01-29 Nathanial Burch , Marta D'Elia , R. B. Lehoucq

In this paper, we study the asymptotic of exit problem for controlled Markov diffusion processes with random jumps and vanishing diffusion terms, where the random jumps are introduced in order to modify the evolution of the controlled…

Dynamical Systems · Mathematics 2018-02-08 Getachew K. Befekadu

We consider the problem of computing optimal policies in average-reward Markov decision processes. This classical problem can be formulated as a linear program directly amenable to saddle-point optimization methods, albeit with a number of…

Optimization and Control · Mathematics 2020-01-13 Joan Bas-Serrano , Gergely Neu

We study the optimal stopping problem of McKean-Vlasov diffusions when the criterion is a function of the law of the stopped process. A remarkable new feature in this setting is that the stopping time also impacts the dynamics of the…

Probability · Mathematics 2023-01-18 Mehdi Talbi , Nizar Touzi , Jianfeng Zhang

This paper is devoted to studying the average optimality in continuous-time Markov decision processes with fairly general state and action spaces. The criterion to be maximized is expected average rewards. The transition rates of underlying…

Probability · Mathematics 2007-05-23 Xianping Guo , Ulrich Rieder

Via operator theoretic methods, we formalize the concentration phenomenon for a given observable `$r$' of a discrete time Markov chain with `$\mu_{\pi}$' as invariant ergodic measure, possibly having support on an unbounded state space. The…

Machine Learning · Computer Science 2023-06-01 Muhammad Abdullah Naeem , Miroslav Pajic

We investigate the optimal stopping problems involving the supremum of a diffusion. The starting point is the link between works of Peskir and Meilijson, which we describe in a unified manner. The description developped follows mainly the…

Probability · Mathematics 2007-05-23 Jan Obloj

We study the local regularity and multifractal nature of the sample paths of jump diffusion processes, which are solutions to a class of stochastic differential equations with jumps. This article extends the recent work of Barral {\it et…

Probability · Mathematics 2017-09-06 Xiaochuan Yang

Random tensors can be used to produce random matrices. This idea is, for instance, very natural when one studies random quantum states with the aim of exploring properties that are generically true, or true with some probability. We hereby…

Mathematical Physics · Physics 2019-07-22 Stephane Dartois

Let $D\subset R^d$ be a bounded domain and denote by $\mathcal P(D)$ the space of probability measures on $D$. Let \begin{equation*} L=\frac12\nabla\cdot a\nabla +b\nabla \end{equation*} be a second order elliptic operator. Let…

Probability · Mathematics 2011-05-19 Ross G. Pinsky

This paper studies the continuous-time reinforcement learning (RL) for optimal switching problems across multiple regimes. We consider a type of exploratory formulation under entropy regularization where the agent randomizes both the timing…

Optimization and Control · Mathematics 2025-12-23 Yijie Huang , Mengge Li , Xiang Yu , Zhou Zhou

We extend to multi-dimensions the work of [1], where new fully explicit kinetic methods were built for the approximation of linear and non-linear convection-diffusion problems. The fundamental principles from the earlier work are retained:…

Numerical Analysis · Mathematics 2023-12-29 Gauthier Wissocq , Rémi Abgrall

We use the geometry of suitably generalised potentials to solve risk-sensitive Markovian optimal stopping problems. As in the linear case due to Dynkin and Yushkievich (1967), the value function is the pointwise infimum of those functions…

Optimization and Control · Mathematics 2025-06-12 Tomasz Kosmala , John Moriarty

We provide resolvent asymptotics as well as various operator-norm estimates for the system of linear partial differential equations describing the thin infinite elastic rod with material coefficients which periodically highly oscillate…

Analysis of PDEs · Mathematics 2023-04-12 Kirill Cherednichenko , Igor Velčić , Josip Žubrinić

By the recent advances in computer technology leading to the invention of more powerful processors, the importance of creating models using data training is even greater than ever. Given the significance of this issue, this work tries to…

Optimization and Control · Mathematics 2023-12-27 Saman Khoramian

This paper considers optimization over multiple renewal systems coupled by time average constraints. These systems act asynchronously over variable length frames. For each system, at the beginning of each renewal frame, it chooses an action…

Optimization and Control · Mathematics 2018-05-23 Xiaohan Wei , Michael J. Neely
‹ Prev 1 8 9 10 Next ›