English
Related papers

Related papers: A random measure approach to reinforcement learnin…

200 papers

Numerous heuristics and advanced approaches have been proposed for exploration in different settings for deep reinforcement learning. Noise-based exploration generally fares well with dense-shaped rewards and bonus-based exploration with…

Machine Learning · Computer Science 2025-10-22 Sebastian Griesbach , Carlo D'Eramo

Discrete adversarial attacks are symbolic perturbations to a language input that preserve the output label but lead to a prediction error. While such attacks have been extensively explored for the purpose of evaluating model robustness,…

Machine Learning · Computer Science 2021-11-02 Maor Ivgi , Jonathan Berant

Reinforcement learning (RL) algorithms allow artificial agents to improve their selection of actions to increase rewarding experiences in their environments. Temporal Difference (TD) Learning -- a model-free RL method -- is a leading…

Machine Learning · Computer Science 2019-09-05 Jacob Rafati , David C. Noelle

Reinforcement learning (RL) with sparse and deceptive rewards is challenging because non-zero rewards are rarely obtained. Hence, the gradient calculated by the agent can be stochastic and without valid information. Recent studies that…

Machine Learning · Computer Science 2024-02-08 Guojian Wang , Faguo Wu , Xiao Zhang , Jianxiang Liu

We introduce a discrete time reflected scheme to solve doubly reflected Backward Stochastic Differential Equations with jumps (in short DRBSDEs), driven by a Brownian motion and an independent compensated Poisson process. As in…

Probability · Mathematics 2015-11-11 Roxana Dumitrescu , Céline Labart

Stochastic differential equations (SDEs) are a staple of mathematical modelling of temporal dynamics. However, a fundamental limitation has been that such models have typically been relatively inflexible, which recent work introducing…

Machine Learning · Computer Science 2021-05-12 Patrick Kidger , James Foster , Xuechen Li , Harald Oberhauser , Terry Lyons

Reward design remains a significant bottleneck in applying reinforcement learning (RL) to real-world problems. A popular alternative is reward learning, where reward functions are inferred from human feedback rather than manually specified.…

Machine Learning · Computer Science 2026-01-16 Chaitanya Kharyal , Calarina Muslimani , Matthew E. Taylor

Local perturbations in conservative particle systems can have a non-local influence on the stationary measure. To capture this phenomenon, we analyze in this paper two toy models. We study the symmetric exclusion process on a countable set…

Probability · Mathematics 2024-10-25 Frank Redig , Ellen Saada

We study an optimal control problem on infinite horizon for a controlled stochastic differential equation driven by Brownian motion, with a discounted reward functional. The equation may have memory or delay effects in the coefficients,…

Optimization and Control · Mathematics 2017-10-19 F. Confortola , A. Cosso , M. Fuhrman

We consider an optimal control problem for piecewise deterministic Markov processes (PDMPs) on a bounded state space. The control problem under study is very general: a pair of controls acts continuously on the deterministic flow and on the…

Optimization and Control · Mathematics 2018-02-14 Elena Bandini

In this paper, we present a novel method for achieving dexterous manipulation of complex objects, while simultaneously securing the object without the use of passive support surfaces. We posit that a key difficulty for training such…

In the Multiple Measurements Vector (MMV) model, measurement vectors are connected to unknown, jointly sparse signal vectors through a linear regression model employing a single known measurement matrix (or dictionary). Typically, the…

Methodology · Statistics 2024-08-05 Esa Ollila

Reinforcement learning (RL) is a control approach that can handle nonlinear stochastic optimal control problems. However, despite the promise exhibited, RL has yet to see marked translation to industrial practice primarily due to its…

Machine Learning · Computer Science 2021-04-15 Elton Pan , Panagiotis Petsagkourakis , Max Mowbray , Dongda Zhang , Antonio del Rio-Chanona

This paper addresses continuous-time reinforcement learning (CTRL) where the system dynamics are governed by an unknown stochastic differential equation, and only discrete-time observations are available. Existing approaches face…

Optimization and Control · Mathematics 2025-10-14 Yuhua Zhu , Yuming Zhang , Haoyu Zhang

We study the problem of optimal control for mean-field stochastic partial differential equations (stochastic evolution equations) driven by a Brownian motion and an independent Poisson random measure, in the case of \textit{partial…

Optimization and Control · Mathematics 2017-04-12 Roxana Dumitrescu , Bernt Øksendal , Agnès Sulem

We present a statistical learning framework for robust identification of partial differential equations from noisy spatiotemporal data. Extending previous sparse regression approaches for inferring PDE models from simulated data, we address…

Numerical Analysis · Mathematics 2019-07-19 Suryanarayana Maddu , Bevan L. Cheeseman , Ivo F. Sbalzarini , Christian L. Müller

This paper studies a recent proposal to use randomized value functions to drive exploration in reinforcement learning. These randomized value functions are generated by injecting random noise into the training data, making the approach…

Machine Learning · Computer Science 2024-09-23 Daniel Russo

The existence of random attractors for a large class of stochastic partial differential equations (SPDE) driven by general additive noise is established. The main results are applied to various types of SPDE, as e.g. stochastic…

Analysis of PDEs · Mathematics 2011-07-21 Benjamin Gess , Wei Liu , Michael Roeckner

Gradient optimization algorithms using epochs, that is those based on stochastic gradient descent without replacement (SGDo), are predominantly used to train machine learning models in practice. However, the mathematical theory of SGDo and…

Machine Learning · Computer Science 2025-12-05 Stefan Perko

While Bayesian-based exploration often demonstrates superior empirical performance compared to bonus-based methods in model-based reinforcement learning (RL), its theoretical understanding remains limited for model-free settings. Existing…

Machine Learning · Computer Science 2026-02-05 He Wang , Xingyu Xu , Yuejie Chi
‹ Prev 1 8 9 10 Next ›