English
Related papers

Related papers: Stochastic processes with competing reinforcements

200 papers

Reinforcement learning has traditionally focused on learning state-dependent policies to solve optimal control problems in a closed-loop fashion. In this work, we introduce the paradigm of open-loop reinforcement learning where a fixed…

Machine Learning · Computer Science 2025-04-23 Onno Eberhard , Claire Vernade , Michael Muehlebach

We introduce a class of learning problems where the agent is presented with a series of tasks. Intuitively, if there is relation among those tasks, then the information gained during execution of one task has value for the execution of…

Machine Learning · Computer Science 2012-09-06 Christos Dimitrakakis

Markov jump processes are continuous-time stochastic processes with a wide range of applications in both natural and social sciences. Despite their widespread use, inference in these models is highly non-trivial and typically proceeds via…

Machine Learning · Computer Science 2023-06-01 Patrick Seifner , Ramses J. Sanchez

Inspired by applications in sports where the skill of players or teams competing against each other varies over time, we propose a probabilistic model of pairwise-comparison outcomes that can capture a wide range of time dynamics. We…

Machine Learning · Statistics 2019-05-20 Lucas Maystre , Victor Kristof , Matthias Grossglauser

We study reinforcement learning from human feedback in general Markov decision processes, where agents learn from trajectory-level preference comparisons. A central challenge in this setting is to design algorithms that select informative…

Machine Learning · Computer Science 2025-12-05 Andreas Schlaginhaufen , Reda Ouhamma , Maryam Kamgarpour

In this paper, we proposed a stochastic model which describes two species of particles moving in counterflow. The model generalizes the theoretical framework describing the transport in random systems since particles can work as mobile…

Soft Condensed Matter · Physics 2017-08-09 Eduardo Velasco Stock , Roberto da Silva , Henrique Almeida Fernandes

The influence of a time-periodic forcing on stochastic processes can essentially be emphasized in the large time behaviour of their paths. The statistics of transition in a simple Markov chain model permits to quantify this influence. In…

Probability · Mathematics 2013-03-27 Samuel Herrmann , Damien Landon

Bioprocesses have received a lot of attention to produce clean and sustainable alternatives to fossil-based materials. However, they are generally difficult to optimize due to their unsteady-state operation modes and stochastic behaviours.…

We present a framework for hedging a portfolio of derivatives in the presence of market frictions such as transaction costs, market impact, liquidity constraints or risk limits using modern deep reinforcement machine learning methods. We…

Computational Finance · Quantitative Finance 2018-02-12 Hans Bühler , Lukas Gonon , Josef Teichmann , Ben Wood

Model-Free Reinforcement Learning has achieved meaningful results in stable environments but, to this day, it remains problematic in regime changing environments like financial markets. In contrast, model-based RL is able to capture some…

Machine Learning · Computer Science 2021-04-23 Eric Benhamou , David Saltiel , Serge Tabachnik , Sui Kai Wong , François Chareyron

Deep reinforcement learning has made significant strides in various robotic tasks. However, employing deep reinforcement learning methods to tackle multi-stage tasks still a challenge. Reinforcement learning algorithms often encounter…

Robotics · Computer Science 2025-03-06 Jiechao Deng , Ning Tan

The influence of an external random field on the competition process in a nonlinear open spatially extended system is analyzed numerically. A three-component model is chosen as the competition model in which a "weak" species can move in…

Pattern Formation and Solitons · Physics 2015-12-02 S. E. Kurushina , V. V. Maximov , E. A. Shapovalova , Yu. M. Romanovskii , I. P. Zavershinskii , D. S. Garipov

Monotonicity is a key qualitative prediction of a wide array of economic models derived via robust comparative statics. It is therefore important to design effective and practical econometric methods for testing this prediction in empirical…

Statistics Theory · Mathematics 2019-07-10 Denis Chetverikov

We introduce a reinforcement learning method for a class of non-Markov systems; our approach extends the actor-critic framework given by Rose et al. [New J. Phys. 23 013013 (2021)] for obtaining scaled cumulant generating functions…

Statistical Mechanics · Physics 2026-03-09 Venkata D. Pamulaparthy , Rosemary J. Harris

Machine-learning techniques are emerging as a valuable tool in experimental physics, and among them, reinforcement learning offers the potential to control high-dimensional, multistage processes in the presence of fluctuating environments.…

We propose a robust optimization approach for constructing confidence bands for stochastic processes using a finite number of simulated sample paths. Our approach can be used to quantify uncertainty in realizations of stochastic processes…

Optimization and Control · Mathematics 2025-08-13 Timothy Chan , Jangwon Park , Vahid Sarhangian

Causal understanding is important in many disciplines of science and engineering, where we seek to understand how different factors in the system causally affect an experiment or situation and pave a pathway towards creating effective or…

Robotics · Computer Science 2025-05-14 Miguel Arana-Catania , Weisi Guo

We consider a class of reinforcement processes, called WARMs, on tree graphs. These processes involve a parameter $\alpha$ which governs the strength of the reinforcement, and a collection of Poisson processes indexed by the vertices of the…

Probability · Mathematics 2020-09-17 Christian Hirsch , Mark Holmes , Victor Kleptsyn

This work provides a Deep Reinforcement Learning approach to solving a periodic review inventory control system with stochastic vendor lead times, lost sales, correlated demand, and price matching. While this dynamic program has…

Machine Learning · Computer Science 2022-11-30 Dhruv Madeka , Kari Torkkola , Carson Eisenach , Anna Luo , Dean P. Foster , Sham M. Kakade

Recent developments in sequential experimental design look to construct a policy that can efficiently navigate the design space, in a way that maximises the expected information gain. Whilst there is work on achieving tractable policies for…

Machine Learning · Computer Science 2025-08-20 Yasir Zubayr Barlas , Kizito Salako