中文
相关论文

相关论文: A survey of random processes with reinforcement

200 篇论文

OpenAI o1 has shown that applying reinforcement learning to integrate reasoning steps directly during inference can significantly improve a model's reasoning capabilities. This result is exciting as the field transitions from the…

人工智能 · 计算机科学 2025-02-18 Jun Wang

This article studies vertex reinforced random walks that are non-backtracking (denoted VRNBW), i.e. U-turns forbidden. With this last property and for a strong reinforcement, the emergence of a path may occur with positive probability.…

概率论 · 数学 2017-08-02 Line C. Le Goff , Olivier Raimond

Exponential random graph models (ERGMs) are a widely used framework for network data, enabling hypothesis testing on the structural mechanisms underlying observed networks. Bayesian ERGMs provide principled uncertainty quantification and…

统计方法学 · 统计学 2026-05-26 Alberto Caimo , Isabella Gollini

Deep reinforcement learning is the combination of reinforcement learning (RL) and deep learning. This field of research has been able to solve a wide range of complex decision-making tasks that were previously out of reach for a machine.…

机器学习 · 计算机科学 2018-12-04 Vincent Francois-Lavet , Peter Henderson , Riashat Islam , Marc G. Bellemare , Joelle Pineau

Working in combinatorial model $\mathrm{W_{co}}(d)$, $d=1,2,\dots$, of P\'olya's random walker in $\mathbb{Z}^d$, we prove two theorems on recurrence to a vertex. We obtain an effective version of the first theorem if $d=2$. Using a…

概率论 · 数学 2025-05-29 Martin Klazar , Richard Horský

Reinforcement Learning (RL) algorithms suffer from the dependency on accurately engineered reward functions to properly guide the learning agents to do the required tasks. Preference-based reinforcement learning (PbRL) addresses that by…

人工智能 · 计算机科学 2024-08-23 Youssef Abdelkareem , Shady Shehata , Fakhri Karray

We present a general construction for dependent random measures based on thinning Poisson processes on an augmented space. The framework is not restricted to dependent versions of a specific nonparametric model, but can be applied to all…

机器学习 · 统计学 2012-11-21 Nicholas J. Foti , Joseph D. Futoma , Daniel N. Rockmore , Sinead Williamson

Social robot navigation is an evolving research field that aims to find efficient strategies to safely navigate dynamic environments populated by humans. A critical challenge in this domain is the accurate modeling of human motion, which…

人机交互 · 计算机科学 2025-07-01 Tommaso Van Der Meer , Andrea Garulli , Antonio Giannitrapani , Renato Quartullo

This article is devoted to methods of construction and study of stochastic models based on Monte Carlo method. A model of Brownian motion, the construction and processing which brings to a world of random numbers and mathematical…

物理教育 · 物理学 2018-09-18 Illia O. Teplytskyi , Serhiy O. Semerikov

Evolution is a fundamental process that shapes the biological world we inhabit, and reinforcement learning is a powerful tool used in artificial intelligence to develop intelligent agents that learn from their environment. In recent years,…

神经与进化计算 · 计算机科学 2023-06-19 Taboubi Ahmed

Reinforcement Learning (RL) is a powerful machine learning paradigm that has been applied in various fields such as robotics, natural language processing and game playing achieving state-of-the-art results. Targeted to solve sequential…

人工智能 · 计算机科学 2023-10-31 Simon Schindler , Martin Uray , Stefan Huber

The randomized play-the-winner (RPW) model is a generalized P\'olya Urn process with broad applications ranging from clinical trials to molecular evolution. We derive an exact expression for the variance of the RPW model by transforming the…

应用统计 · 统计学 2024-01-02 Ivan Specht , Michael Mitzenmacher

We study reinforcement learning from human feedback in general Markov decision processes, where agents learn from trajectory-level preference comparisons. A central challenge in this setting is to design algorithms that select informative…

机器学习 · 计算机科学 2025-12-05 Andreas Schlaginhaufen , Reda Ouhamma , Maryam Kamgarpour

Recent progress on the understanding of the Random Conductance Model is reviewed and commented. A particular emphasis is on the results on the scaling limit of the random walk among random conductances for almost every realization of the…

概率论 · 数学 2012-01-04 Marek Biskup

Recent times are witnessing rapid development in machine learning algorithm systems, especially in reinforcement learning, natural language processing, computer and robot vision, image processing, speech, and emotional processing and…

The need for algorithms able to solve Reinforcement Learning (RL) problems with few trials has motivated the advent of model-based RL methods. The reported performance of model-based algorithms has dramatically increased within recent…

机器学习 · 计算机科学 2022-03-22 Giacomo Arcieri , David Wölfle , Eleni Chatzi

A step-reinforced random walk is a discrete-time non-Markovian process with long range memory. At each step, with a fixed probability p, the positively step-reinforced random walk repeats one of its preceding steps chosen uniformly at…

概率论 · 数学 2023-11-28 Zhishui Hu , Yiting Zhang

We introduce the use of reinforcement learning for indirect mechanisms, working with the existing class of sequential price mechanisms, which generalizes both serial dictatorship and posted price mechanisms and essentially characterizes all…

计算机科学与博弈论 · 计算机科学 2021-05-07 Gianluca Brero , Alon Eden , Matthias Gerstgrasser , David C. Parkes , Duncan Rheingans-Yoo

We apply reinforcement learning (RL) to robotics tasks. One of the drawbacks of traditional RL algorithms has been their poor sample efficiency. One approach to improve the sample efficiency is model-based RL. In our model-based RL…

机器学习 · 计算机科学 2023-05-16 Adithya Ramesh , Balaraman Ravindran

In reinforcement learning (RL), agents often operate in partially observed and uncertain environments. Model-based RL suggests that this is best achieved by learning and exploiting a probabilistic model of the world. 'Active inference' is…

机器学习 · 计算机科学 2019-11-26 Alexander Tschantz , Manuel Baltieri , Anil. K. Seth , Christopher L. Buckley