English
Related papers

Related papers: A survey of random processes with reinforcement

200 papers

This paper considers a class of reinforcement-based learning (namely, perturbed learning automata) and provides a stochastic-stability analysis in repeatedly-played, positive-utility, finite strategic-form games. Prior work in this class of…

Computer Science and Game Theory · Computer Science 2019-01-29 Georgios C. Chasparis

The organising principles underlying the structure of phenomenologically viable string vacua can be accessed by sampling such vacua. In many cases this is prohibited by the computational cost of standard sampling methods in the high…

High Energy Physics - Theory · Physics 2021-07-12 Sven Krippendorf , Rene Kroepsch , Marc Syvaeri

The deployment of reinforcement learning (RL) in the real world comes with challenges in calibrating user trust and expectations. As a step toward developing RL systems that are able to communicate their competencies, we present a method of…

Machine Learning · Computer Science 2020-11-19 Aastha Acharya , Rebecca Russell , Nisar R. Ahmed

Reinforcement learning (RL) algorithms find applications in inventory control, recommender systems, vehicular traffic management, cloud computing and robotics. The real-world complications of many tasks arising in these domains makes them…

Machine Learning · Computer Science 2021-06-03 Sindhu Padakandla

In extreme value statistics, the peaks-over-threshold method is widely used. The method is based on the generalized Pareto distribution characterizing probabilities of exceedances over high thresholds in $\mathbb {R}^d$. We present a…

Probability · Mathematics 2014-10-17 Ana Ferreira , Laurens de Haan

Reinforcement learning (RL) has become a foundational approach for enabling intelligent robotic behavior in dynamic and uncertain environments. This work presents an in-depth review of RL principles, advanced deep reinforcement learning…

Robotics · Computer Science 2026-03-17 Kumater Ter , Abolanle Adetifa , Daniel Udekwe

We present an AI-based ecosystem simulator that uses three-dimensional models of the terrain and animal models controlled by deep reinforcement learning. The simulations take place in a game engine environment, which enables continuous…

Multiagent Systems · Computer Science 2023-03-22 Claes Strannegård , Niklas Engsner , Rasmus Lindgren , Simon Olsson , John Endler

We exploit a bijection between plane recursive trees and Stirling permutations; this yields the equivalence of some results previously proven separately by different methods for the two types of objects as well as some new results. We also…

Combinatorics · Mathematics 2008-03-10 Svante Janson

We consider a self-attracting random walk in dimension d=1, in presence of a field of strength s, which biases the walker toward a target site. We focus on the dynamic case (true reinforced random walk), where memory effects are implemented…

Statistical Mechanics · Physics 2015-06-05 Elena Agliari , Raffaella Burioni , Guido Uguzzoni

In this survey, we propose an overview on Lyapunov functions for a variety of compartmental models in epidemiology. We exhibit the most widely employed functions, together with a commentary on their use. Our aim is to provide a…

Dynamical Systems · Mathematics 2023-06-09 Nicolò Cangiotti , Marco Capolli , Mattia Sensi , Sara Sottile

In practical applications, we can rarely assume full observability of a system's environment, despite such knowledge being important for determining a reactive control system's precise interaction with its environment. Therefore, we propose…

Machine Learning · Computer Science 2022-06-24 Edi Muskardin , Martin Tappler , Bernhard K. Aichernig , Ingo Pill

Rapid advances of hardware-based technologies during the past decades have opened up new possibilities for Life scientists to gather multimodal data in various application domains (e.g., Omics, Bioimaging, Medical Imaging, and…

Machine Learning · Computer Science 2018-01-09 Mufti Mahmud , M. Shamim Kaiser , Amir Hussain , Stefano Vassanelli

Efficiently tackling multiple tasks within complex environment, such as those found in robot manipulation, remains an ongoing challenge in robotics and an opportunity for data-driven solutions, such as reinforcement learning (RL).…

Robotics · Computer Science 2024-04-03 Carlos Plou , Ana C. Murillo , Ruben Martinez-Cantin

We consider a version of the classical P\'olya urn scheme which incorporates innovations. The space $S$ of colors is an arbitrary measurable set. After each sampling of a ball in the urn, one returns $C$ balls of the same color and…

Probability · Mathematics 2022-11-17 Jean Bertoin

Vertex-reinforced random walk is defined in Pemantle's (1988) thesis; it is a random walk that is biased to visit sites it has already visited a lot. We show that this reinforcement scheme, in contrast to the scheme of edge-reinforcement,…

Probability · Mathematics 2016-09-07 Robin Pemantle , Stanislav Volkov

Consistently checking the statistical significance of experimental results is the first mandatory step towards reproducible science. This paper presents a hitchhiker's guide to rigorous comparisons of reinforcement learning algorithms.…

Methodology · Statistics 2022-08-30 Cédric Colas , Olivier Sigaud , Pierre-Yves Oudeyer

A general random effects model is proposed that allows for continuous as well as discrete distributions of the responses. Responses can be unrestricted continuous, bounded continuous, binary, ordered categorical or given in the form of…

Methodology · Statistics 2024-04-30 Gerhard Tutz

Understanding pedestrian behavior is crucial for the safe deployment of Autonomous Vehicles (AVs) in urban environments. Traditional pedestrian behavior models often fall into two categories: mechanistic models, which do not generalize well…

Human-Computer Interaction · Computer Science 2024-09-24 Yueyang Wang , Aravinda Ramakrishnan Srinivasan , Yee Mun Lee , Gustav Markkula

Extreme shock models have been introduced in Gut and H\"usler (1999) to study systems that at random times are subject to shock of random magnitude. These systems break down when some shock overcomes a given resistance level. In this paper…

Other Statistics · Statistics 2010-10-19 Pasquale Cirillo , Jürg Hüsler

Process control is widely discussed in the manufacturing process, especially for semiconductor manufacturing. Due to unavoidable disturbances in manufacturing, different process controllers are proposed to realize variation reduction. Since…

Systems and Control · Electrical Eng. & Systems 2021-10-25 Yanrong Li , Juan Du , Wei Jiang
‹ Prev 1 8 9 10 Next ›