English
Related papers

Related papers: A survey of random processes with reinforcement

200 papers

OpenAI o1 has shown that applying reinforcement learning to integrate reasoning steps directly during inference can significantly improve a model's reasoning capabilities. This result is exciting as the field transitions from the…

Artificial Intelligence · Computer Science 2025-02-18 Jun Wang

This article studies vertex reinforced random walks that are non-backtracking (denoted VRNBW), i.e. U-turns forbidden. With this last property and for a strong reinforcement, the emergence of a path may occur with positive probability.…

Probability · Mathematics 2017-08-02 Line C. Le Goff , Olivier Raimond

Exponential random graph models (ERGMs) are a widely used framework for network data, enabling hypothesis testing on the structural mechanisms underlying observed networks. Bayesian ERGMs provide principled uncertainty quantification and…

Methodology · Statistics 2026-05-26 Alberto Caimo , Isabella Gollini

Deep reinforcement learning is the combination of reinforcement learning (RL) and deep learning. This field of research has been able to solve a wide range of complex decision-making tasks that were previously out of reach for a machine.…

Machine Learning · Computer Science 2018-12-04 Vincent Francois-Lavet , Peter Henderson , Riashat Islam , Marc G. Bellemare , Joelle Pineau

Working in combinatorial model $\mathrm{W_{co}}(d)$, $d=1,2,\dots$, of P\'olya's random walker in $\mathbb{Z}^d$, we prove two theorems on recurrence to a vertex. We obtain an effective version of the first theorem if $d=2$. Using a…

Probability · Mathematics 2025-05-29 Martin Klazar , Richard Horský

Reinforcement Learning (RL) algorithms suffer from the dependency on accurately engineered reward functions to properly guide the learning agents to do the required tasks. Preference-based reinforcement learning (PbRL) addresses that by…

Artificial Intelligence · Computer Science 2024-08-23 Youssef Abdelkareem , Shady Shehata , Fakhri Karray

We present a general construction for dependent random measures based on thinning Poisson processes on an augmented space. The framework is not restricted to dependent versions of a specific nonparametric model, but can be applied to all…

Machine Learning · Statistics 2012-11-21 Nicholas J. Foti , Joseph D. Futoma , Daniel N. Rockmore , Sinead Williamson

Social robot navigation is an evolving research field that aims to find efficient strategies to safely navigate dynamic environments populated by humans. A critical challenge in this domain is the accurate modeling of human motion, which…

Human-Computer Interaction · Computer Science 2025-07-01 Tommaso Van Der Meer , Andrea Garulli , Antonio Giannitrapani , Renato Quartullo

This article is devoted to methods of construction and study of stochastic models based on Monte Carlo method. A model of Brownian motion, the construction and processing which brings to a world of random numbers and mathematical…

Physics Education · Physics 2018-09-18 Illia O. Teplytskyi , Serhiy O. Semerikov

Evolution is a fundamental process that shapes the biological world we inhabit, and reinforcement learning is a powerful tool used in artificial intelligence to develop intelligent agents that learn from their environment. In recent years,…

Neural and Evolutionary Computing · Computer Science 2023-06-19 Taboubi Ahmed

Reinforcement Learning (RL) is a powerful machine learning paradigm that has been applied in various fields such as robotics, natural language processing and game playing achieving state-of-the-art results. Targeted to solve sequential…

Artificial Intelligence · Computer Science 2023-10-31 Simon Schindler , Martin Uray , Stefan Huber

The randomized play-the-winner (RPW) model is a generalized P\'olya Urn process with broad applications ranging from clinical trials to molecular evolution. We derive an exact expression for the variance of the RPW model by transforming the…

Applications · Statistics 2024-01-02 Ivan Specht , Michael Mitzenmacher

We study reinforcement learning from human feedback in general Markov decision processes, where agents learn from trajectory-level preference comparisons. A central challenge in this setting is to design algorithms that select informative…

Machine Learning · Computer Science 2025-12-05 Andreas Schlaginhaufen , Reda Ouhamma , Maryam Kamgarpour

Recent progress on the understanding of the Random Conductance Model is reviewed and commented. A particular emphasis is on the results on the scaling limit of the random walk among random conductances for almost every realization of the…

Probability · Mathematics 2012-01-04 Marek Biskup

Recent times are witnessing rapid development in machine learning algorithm systems, especially in reinforcement learning, natural language processing, computer and robot vision, image processing, speech, and emotional processing and…

The need for algorithms able to solve Reinforcement Learning (RL) problems with few trials has motivated the advent of model-based RL methods. The reported performance of model-based algorithms has dramatically increased within recent…

Machine Learning · Computer Science 2022-03-22 Giacomo Arcieri , David Wölfle , Eleni Chatzi

A step-reinforced random walk is a discrete-time non-Markovian process with long range memory. At each step, with a fixed probability p, the positively step-reinforced random walk repeats one of its preceding steps chosen uniformly at…

Probability · Mathematics 2023-11-28 Zhishui Hu , Yiting Zhang

We introduce the use of reinforcement learning for indirect mechanisms, working with the existing class of sequential price mechanisms, which generalizes both serial dictatorship and posted price mechanisms and essentially characterizes all…

Computer Science and Game Theory · Computer Science 2021-05-07 Gianluca Brero , Alon Eden , Matthias Gerstgrasser , David C. Parkes , Duncan Rheingans-Yoo

We apply reinforcement learning (RL) to robotics tasks. One of the drawbacks of traditional RL algorithms has been their poor sample efficiency. One approach to improve the sample efficiency is model-based RL. In our model-based RL…

Machine Learning · Computer Science 2023-05-16 Adithya Ramesh , Balaraman Ravindran

In reinforcement learning (RL), agents often operate in partially observed and uncertain environments. Model-based RL suggests that this is best achieved by learning and exploiting a probabilistic model of the world. 'Active inference' is…

Machine Learning · Computer Science 2019-11-26 Alexander Tschantz , Manuel Baltieri , Anil. K. Seth , Christopher L. Buckley