English
Related papers

Related papers: A random measure approach to reinforcement learnin…

200 papers

Inspired by the ubiquitous use of differential equations to model continuous dynamics across diverse scientific and engineering domains, we propose a novel and intuitive approach to continuous sequence modeling. Our method interprets…

Machine Learning · Computer Science 2025-02-03 Macheng Shen , Chen Cheng

In sequential change detection, existing performance measures differ significantly in the way they treat the time of change. By modeling this quantity as a random time, we introduce a general framework capable of capturing and better…

Statistics Theory · Mathematics 2008-12-18 George V. Moustakides

For the stochastic differential equation (SDE) which has piecewise continuous arguments (PCAs), is driven by multiplicative noises and its drift coefficients are dissipative, we show that the solution at integer time is a Markov chain and…

Numerical Analysis · Mathematics 2024-09-23 Chuchu Chen , Jialin Hong , Yulan Lu

Diffusion with stochastic transport is investigated here when the random driving process is a very general Gaussian process, including Fractional Brownian motion. The purpose is the comparison with a deterministic PDE, which in certain…

Probability · Mathematics 2026-04-20 Franco Flandoli , Francesco Russo

This paper proposes a novel approach to controller design for MR-damped vehicle suspension system. This approach is predicated on the premise that the optimal control strategy can be learned through real-world or simulated experiments…

Systems and Control · Electrical Eng. & Systems 2023-09-06 AmirReza BabaAhmadi , Masoud ShariatPanahi , Moosa Ayati

For linear systems, many data-driven control methods rely on the behavioral framework, using historical data of the system to predict the future trajectories. However, measurement noise introduces errors in predictions. When the noise is…

Optimization and Control · Mathematics 2023-08-29 Baiwei Guo , Yuning Jiang , Colin N. Jones , Giancarlo Ferrari-Trecate

We propose a novel framework to solve risk-sensitive reinforcement learning (RL) problems where the agent optimises time-consistent dynamic spectral risk measures. Based on the notion of conditional elicitability, our methodology constructs…

Machine Learning · Computer Science 2023-05-02 Anthony Coache , Sebastian Jaimungal , Álvaro Cartea

Non-uniform sampling arises when an experimenter does not have full control over the sampling characteristics of the process under investigation. Moreover, it is introduced intentionally in algorithms such as Bayesian optimization and…

Machine Learning · Statistics 2020-07-03 Stijn de Waele

Recent years have witnessed significant progress in developing effective training and fast sampling techniques for diffusion models. A remarkable advancement is the use of stochastic differential equations (SDEs) and their…

Computer Vision and Pattern Recognition · Computer Science 2024-08-26 Defang Chen , Zhenyu Zhou , Jian-Ping Mei , Chunhua Shen , Chun Chen , Can Wang

We propose and analyze a randomization scheme for a general class of impulse control problems. The solution to this randomized problem is characterized as the fixed point of a compound operator which consists of a regularized nonlocal…

Optimization and Control · Mathematics 2026-05-26 Haoyang Cao , Yuchao Dong , Zhouhao Yang

Randomized smoothing is a defensive technique to achieve enhanced robustness against adversarial examples which are small input perturbations that degrade the performance of neural network models. Conventional randomized smoothing adds…

Machine Learning · Computer Science 2024-07-17 Ryo Hase , Ye Wang , Toshiaki Koike-Akino , Jing Liu , Kieran Parsons

In this paper, by introducing a new type asymptotic coupling by reflection, we explore the long time behavior of random probability measure flows associated with a large class of one-dimensional McKean-Vlasov SDEs with common noise.…

Probability · Mathematics 2024-01-17 Bao Jianhai , Wang Jian

This paper studies the continuous-time reinforcement learning (RL) for optimal switching problems across multiple regimes. We consider a type of exploratory formulation under entropy regularization where the agent randomizes both the timing…

Optimization and Control · Mathematics 2025-12-23 Yijie Huang , Mengge Li , Xiang Yu , Zhou Zhou

Stochastic differential equations (SDEs) or diffusions are continuous-valued continuous-time stochastic processes widely used in the applied and mathematical sciences. Simulating paths from these processes is usually an intractable problem,…

Computation · Statistics 2020-05-27 Qi Wang , Vinayak Rao , Yee Whye Teh

We present a framework and algorithms to learn controlled dynamics models using neural stochastic differential equations (SDEs) -- SDEs whose drift and diffusion terms are both parametrized by neural networks. We construct the drift term to…

Machine Learning · Computer Science 2023-10-17 Franck Djeumou , Cyrus Neary , Ufuk Topcu

Stochastic evolution equations with compensated Poisson noise are considered in the variational approach with monotone and coercive coefficients. Here the Poisson noise is assumed to be time-homogeneous with $\sigma$-finite intensity…

Probability · Mathematics 2022-04-20 Sima Mehri , Erfan Salavati , Bijan Z. Zangeneh

We present two elegant solutions for modeling continuous-time dynamics, in a novel model-based reinforcement learning (RL) framework for semi-Markov decision processes (SMDPs), using neural ordinary differential equations (ODEs). Our models…

Machine Learning · Computer Science 2020-10-27 Jianzhun Du , Joseph Futoma , Finale Doshi-Velez

In reinforcement learning (RL), an autonomous agent learns to perform complex tasks by maximizing an exogenous reward signal while interacting with its environment. In real-world applications, test conditions may differ substantially from…

Robotics · Computer Science 2019-10-30 Matteo Turchetta , Andreas Krause , Sebastian Trimpe

We consider a backward stochastic differential equation with jumps (BSDEJ) which is driven by a Brownian motion and a Poisson random measure. We present two candidate-approximations to this BSDEJ and we prove that the solution of each…

Probability · Mathematics 2013-12-19 Giulia Di Nunno , Asma Khedher , Michele Vanmaele

Reinforcement Learning (RL) is a computational approach to reward-driven learning in sequential decision problems. It implements the discovery of optimal actions by learning from an agent interacting with an environment rather than from…

Methodology · Statistics 2022-10-06 Mauricio Tec , Yunshan Duan , Peter Müller