English
Related papers

Related papers: A random measure approach to reinforcement learnin…

200 papers

Measuring learning progress is essential for curiosity-driven exploration in reinforcement learning, but widely used signals such as prediction error often fail to distinguish meaningful, learnable patterns from random noise. This paper…

Machine Learning · Computer Science 2026-05-08 Samuel Blad , Martin Längkvist , Amy Loutfi

We extend temporal-difference (TD) learning in order to obtain risk-sensitive, model-free reinforcement learning algorithms. This extension can be regarded as modification of the Rescorla-Wagner rule, where the (sigmoidal) stimulus is taken…

Machine Learning · Computer Science 2021-11-05 Grégoire Delétang , Jordi Grau-Moya , Markus Kunesch , Tim Genewein , Rob Brekelmans , Shane Legg , Pedro A. Ortega

Safe exploration remains a fundamental challenge in reinforcement learning (RL), limiting the deployment of RL agents in the real world. We propose Sampling-Based Safe Reinforcement Learning (SBSRL), a model-based RL algorithm that…

Machine Learning · Computer Science 2026-05-20 Luca Vignola , Bruce D. Lee , Manish Prajapat , Manuel Wendl , Melanie Zeilinger , Andreas Krause , Yarden As

Imagine a patient in critical condition. What and when should be measured to forecast detrimental events, especially under the budget constraints? We answer this question by deep reinforcement learning (RL) that jointly minimizes the…

Machine Learning · Computer Science 2019-06-11 Chun-Hao Chang , Mingjie Mai , Anna Goldenberg

In this paper, we investigate the stochastic evolution equations (SEEs) driven by $\log$-Whittle-Mat$\acute{{\mathrm{e}}}$rn (W-M) random diffusion coefficient field and $Q$-Wiener multiplicative force noise. First, the well-posedness of…

Numerical Analysis · Mathematics 2022-07-05 X. Qi , M. Azaiez , C. Huang , C. Xu

We present a supervised learning framework of training generative models for density estimation. Generative models, including generative adversarial networks, normalizing flows, variational auto-encoders, are usually considered as…

Machine Learning · Computer Science 2023-10-24 Yanfang Liu , Minglei Yang , Zezhong Zhang , Feng Bao , Yanzhao Cao , Guannan Zhang

Diffusion models have emerged as a dominant framework for generative modeling, but their mathematical foundations are often presented separately through diffusion probabilistic models, score-based modeling, stochastic differential…

Machine Learning · Computer Science 2026-05-29 Jiayi Fu , Yuxia Wang

In this thesis, we extend the recently introduced theory of stochastic modified equations (SMEs) for stochastic gradient optimization algorithms. In Ch. 3 we study time-inhomogeneous SDEs driven by Brownian motion. For certain SDEs we prove…

Probability · Mathematics 2025-11-26 Stefan Perko

We prove a general existence result in stochastic optimal control in discrete time where controls take values in conditional metric spaces, and depend on the current state and the information of past decisions through the evolution of a…

Optimization and Control · Mathematics 2018-12-19 Asgar Jamneshan , Michael Kupper , José Miguel Zapata

In this paper we study Backward Stochastic Differential Equations with two reflecting right continuous with left limits obstacles (or barriers) when the noise is given by Brownian motion and a Poisson random measure mutually independent.…

Probability · Mathematics 2008-12-10 S. Hamadéne , H. Wang

Sampling is ubiquitous in machine learning methodologies. Due to the growth of large datasets and model complexity, we want to learn and adapt the sampling process while training a representation. Towards achieving this grand goal, a…

Machine Learning · Computer Science 2022-12-14 Jason Xiaotian Dou , Alvin Qingkai Pan , Runxue Bao , Haiyi Harry Mao , Lei Luo , Zhi-Hong Mao

We study stochastic optimal control of rough stochastic differential equations (RSDEs). This is in the spirit of the pathwise control problem (Lions--Souganidis 1998, Buckdahn--Ma 2007; also Davis--Burstein 1992), with renewed interest and…

Probability · Mathematics 2025-10-24 Peter K. Friz , Khoa Lê , Huilin Zhang

Stochastic differential equations (SDEs) provide a natural framework for modelling intrinsic stochasticity inherent in many continuous-time physical processes. When such processes are observed in multiple individuals or experimental units,…

Computation · Statistics 2016-05-19 Gavin A. Whitaker , Andrew Golightly , Richard J. Boys , Chris Sherlock

This paper is concerned with optimal control of systems driven by G-stochastic differential equations (G-SDEs), with controlled jump term. We study the relaxed problem, in which admissible controls are measurevalued processes and the state…

Optimization and Control · Mathematics 2021-11-04 Hanane Ben Gherbal , Amel Redjil , Omar Kebiri

We consider the control of semilinear stochastic partial differential equations (SPDEs) via deterministic controls. In the case of multiplicative noise, existence of optimal controls and necessary conditions for optimality are derived. In…

Optimization and Control · Mathematics 2021-10-28 Wilhelm Stannat , Lukas Wessels

We consider a class of dissipative stochastic differential equations (SDE's) with time-periodic coefficients in finite dimension, and the response of time-asymptotic probability measures induced by such SDE's to sufficiently regular, small…

Probability · Mathematics 2022-01-04 Michal Branicki , Kenneth Uda

Reinforcement learning with human-annotated data has boosted chain-of-thought reasoning in large reasoning models, but these gains come at high costs in labeled data while faltering on harder tasks. A natural next step is experience-driven…

Computation and Language · Computer Science 2025-10-03 Zhaoning Yu , Will Su , Leitian Tao , Haozhu Wang , Aashu Singh , Hanchao Yu , Jianyu Wang , Hongyang Gao , Weizhe Yuan , Jason Weston , Ping Yu , Jing Xu

Noisy dynamical models are employed to describe a wide range of phenomena. Since exact modeling of these phenomena requires access to their microscopic dynamics, whose time scales are typically much shorter than the observable time scales,…

Statistical Mechanics · Physics 2015-11-18 Giovanni Volpe , Jan Wehr

Revisiting the continuous-time Mean-Variance (MV) Portfolio Optimization problem, we model the market dynamics with a jump-diffusion process and apply Reinforcement Learning (RL) techniques to facilitate informed exploration within the…

Portfolio Management · Quantitative Finance 2025-12-11 Yuling Max Chen , Bin Li , David Saunders

In real-world scenarios, the observation data for reinforcement learning with continuous control is commonly noisy and part of it may be dynamically missing over time, which violates the assumption of many current methods developed for…

Machine Learning · Computer Science 2019-02-18 Yuhui Wang , Hao He , Xiaoyang Tan
‹ Prev 1 4 5 6 7 8 10 Next ›