Related papers: Weak Convergence Methods for Approximation of Path…
Stochastic policies (also known as relaxed controls) are widely used in continuous-time reinforcement learning algorithms. However, executing a stochastic policy and evaluating its performance in a continuous-time environment remain open…
We propose a new weak convergence theorem for martingales, under gentler conditions than the usual convergence in probability of the sequence of associated quadratic variations. Its proof requires the combined use of Skorohod's…
This article characterizes conjugates and subdifferentials of convex integral functionals over linear spaces of cadlag stochastic processes. The approach is based on new measurability results on the Skorokhod space and new interchange rules…
We provide a criterion for establishing lower bounds on the rate of convergence in $f$-variation of a continuous-time ergodic Markov process to its invariant measure. The criterion consists of novel super- and submartingale conditions for…
The authors present a new simple algorithm to approximate weakly stochastic differential equations in the spirit of [1] and [2]. They apply it to the problem of pricing Asian options under the Heston stochastic volatility model, and compare…
An improved version of the functional limit theorem is proved establishing weak convergence of random walks generated by compound doubly stochastic Poisson processes (compound Cox processes) to L{\'e}vy processes in the Skorokhod space…
This article is concerned with the existence of solution to the stochastic Degasperis-Procesi equation on $\mathbb{R}$ with an infinite dimensional multiplicative noise and integrable initial data. Writing the equation as a system composed…
This paper explores the well known approximation approach to decide weak bisimilarity of Basic Parallel Processes. We look into how different refinement functions can be used to prove weak bisimilarity decidable for certain subclasses. We…
Our aim is to find sufficient conditions for weak convergence of stochastic integrals with respect to the state occupation measure of a Markov chain. First, we study properties of the state indicator function and the state occupation…
We study the problem of minimizing a $m$-weakly convex and possibly nonsmooth function. Weak convexity provides a broad framework that subsumes convex, smooth, and many composite nonconvex functions. In this work, we propose a…
In this article a stochastic particle system approximation to the parametric sensitivity in the Smoluchowski coagulation equation is introduced. The parametric sensitivity is the derivative of the solution to the equation with respect to…
We consider the emphatic temporal-difference (TD) algorithm, ETD($\lambda$), for learning the value functions of stationary policies in a discounted, finite state and action Markov decision process. The ETD($\lambda$) algorithm was recently…
We describe an asynchronous parallel stochastic proximal coordinate descent algorithm for minimizing a composite objective function, which consists of a smooth convex function plus a separable convex function. In contrast to previous…
We study large deviation properties of systems of weakly interacting particles modeled by It\^{o} stochastic differential equations (SDEs). It is known under certain conditions that the corresponding sequence of empirical measures…
We prove an invariance principle for non-stationary random processes and establish a rate of convergence under a new type of mixing condition. The dependence is exponentially decaying in the gap between the past and the future and is…
Recent empirical studies suggest that the volatilities associated with financial time series exhibit short-range correlations. This entails that the volatility process is very rough and its autocorrelation exhibits sharp decay at the…
An approach for the description of stochastic systems is derived. Some of the variables in the system are studied forward in time, others backward in time. The approach is based on a perturbation expansion in the strength of the coupling…
We develop a practical approach to establish the stability, that is, the recurrence in a given set, of a large class of controlled Markov chains. These processes arise in various areas of applied science and encompass important numerical…
In this paper, weak convergences of marked empirical processes in $L^2(\mathbb{R},\nu)$ and their applications to statistical goodness-of-fit tests are provided, where $L^2(\mathbb{R},\nu)$ is the set of equivalence classes of the square…
The limits of scaled relative entropies between probability distributions associated with N-particle weakly interacting Markov processes are considered. The convergence of such scaled relative entropies is established in various settings.…