English
Related papers

Related papers: A survey of random processes with reinforcement

200 papers

Consider a finite undirected graph and place an urn with balls of two colours at each vertex. At every discrete time step, for each urn, a fixed number of balls are drawn from that same urn with probability $p$, and from a randomly chosen…

Probability · Mathematics 2024-08-29 Yogesh Dahiya , Neeraja Sahasrabudhe

This survey (re)introduces reinforcement learning methods to economists. The curse of dimensionality limits how far exact dynamic programming can be effectively applied, forcing us to rely on suitably "small" problems or our ability to…

General Economics · Economics 2026-03-25 Pranjal Rawat

This manuscript surveys reinforcement learning from the perspective of optimization and control with a focus on continuous control applications. It surveys the general formulation, terminology, and typical experimental implementations of…

Optimization and Control · Mathematics 2018-11-13 Benjamin Recht

We consider a class of multi-particle reinforced interacting random walks. In this model, there are some (finite or infinite) particles performing random walks on a given (finite or infinite) connected graph, so that each particle has…

Probability · Mathematics 2013-03-26 Jun Chen

These lecture notes give a statistical perspective on the foundations of reinforcement learning and interactive decision making. We present a unifying framework for addressing the exploration-exploitation dilemma using frequentist and…

Machine Learning · Computer Science 2023-12-29 Dylan J. Foster , Alexander Rakhlin

In recent years some researchers have explored the use of reinforcement learning (RL) algorithms as key components in the solution of various natural language processing tasks. For instance, some of these algorithms leveraging deep neural…

Computation and Language · Computer Science 2026-04-29 Victor Uc-Cetina , Nicolas Navarro-Guerrero , Anabel Martin-Gonzalez , Cornelius Weber , Stefan Wermter

The generalized P\`olya urn (GPU) models and their variants have been investigated in several disciplines. However, typical assumptions made with respect to the GPU do not include urn models with diagonal replacement matrix, which arise in…

Probability · Mathematics 2015-02-24 Andrea Ghiglietti , Anand N. Vidyashankar , William F. Rosenberger

Reinforcement Learning (RL) has emerged as a highly effective technique for addressing various scientific and applied problems. Despite its success, certain complex tasks remain challenging to be addressed solely with a single model and…

Machine Learning · Computer Science 2023-12-14 Yanjie Song , P. N. Suganthan , Witold Pedrycz , Junwei Ou , Yongming He , Yingwu Chen , Yutong Wu

Randomized experiments (a.k.a. A/B tests) are a powerful tool for estimating treatment effects, to inform decisions making in business, healthcare and other applications. In many problems, the treatment has a lasting effect that evolves…

Machine Learning · Computer Science 2022-10-17 Ziyang Tang , Yiheng Duan , Stephanie Zhang , Lihong Li

Reinforcement learning is one of the core components in designing an artificial intelligent system emphasizing real-time response. Reinforcement learning influences the system to take actions within an arbitrary environment either having…

Artificial Intelligence · Computer Science 2020-02-03 Amit Kumar Mondal

We introduce a new framework for the analysis of the dynamics of networks, based on randomly reinforced urn (RRU) processes, in which the weight of the edges is determined by a reinforcement mechanism. We rigorously explain the empirical…

Physics and Society · Physics 2012-03-08 Guido Caldarelli , Alessandro Chessa , Irene Crimaldi , Fabio Pammolli

Simulating trajectories of virtual crowds is a commonly encountered task in Computer Graphics. Several recent works have applied Reinforcement Learning methods to animate virtual agents, however they often make different design choices when…

Machine Learning · Computer Science 2022-09-21 Ariel Kwiatkowski , Vicky Kalogeiton , Julien Pettré , Marie-Paule Cani

Sufficient conditions are developed for a class of generalized Polya urn schemes ensuring exchangeability. The extended class includes the Blackwell-MacQueen Polya urn and the urn schemes for the two-parameter Poisson-Dirichlet process and…

Probability · Mathematics 2007-05-23 Hemant Ishwaran , Mahmoud Zarepour

This manuscript gives a big-picture, up-to-date overview of the field of (deep) reinforcement learning and sequential decision making, covering value-based methods, policy-based methods, model-based methods, multi-agent RL, LLMs and RL, and…

Artificial Intelligence · Computer Science 2025-12-03 Kevin Murphy

Deep reinforcement learning enables algorithms to learn complex behavior, deal with continuous action spaces and find good strategies in environments with high dimensional state spaces. With deep reinforcement learning being an active area…

Machine Learning · Computer Science 2018-10-17 Winfried Lötzsch

Governments around the world aspire to ground decision-making on evidence. Many of the foundations of policy making - e.g. sensing patterns that relate to societal needs, developing evidence-based programs, forecasting potential outcomes of…

Artificial Intelligence · Computer Science 2023-12-12 Theodore Wolf , Nantas Nardelli , John Shawe-Taylor , Maria Perez-Ortiz

As a subfield of machine learning, reinforcement learning (RL) aims at empowering one's capabilities in behavioural decision making by using interaction experience with the world and an evaluative feedback. Unlike traditional supervised…

Machine Learning · Computer Science 2020-04-27 Chao Yu , Jiming Liu , Shamim Nemati

Reinforced Galton--Watson processes describe the dynamics of a population where reproduction events are reinforced, in the sense that offspring numbers of forebears can be repeated randomly by descendants. More specifically, the evolution…

Probability · Mathematics 2025-02-24 Jean Bertoin , Bastien Mallein

Reinforcement learning (RL) is concerned with how intelligence agents take actions in a given environment to maximize the cumulative reward they receive. In healthcare, applying RL algorithms could assist patients in improving their health…

Machine Learning · Statistics 2025-04-21 Chengchun Shi

A number of papers have examined various aspects of "random random" walks on finite groups; the purpose of this article is to provide a survey of this work and to show, bring together, and discuss some of the arguments and results in this…

Probability · Mathematics 2007-05-23 Martin Hildebrand