中文
相关论文

相关论文: A survey of random processes with reinforcement

200 篇论文

This paper considers a class of reinforcement-based learning (namely, perturbed learning automata) and provides a stochastic-stability analysis in repeatedly-played, positive-utility, finite strategic-form games. Prior work in this class of…

计算机科学与博弈论 · 计算机科学 2019-01-29 Georgios C. Chasparis

The organising principles underlying the structure of phenomenologically viable string vacua can be accessed by sampling such vacua. In many cases this is prohibited by the computational cost of standard sampling methods in the high…

高能物理 - 理论 · 物理学 2021-07-12 Sven Krippendorf , Rene Kroepsch , Marc Syvaeri

The deployment of reinforcement learning (RL) in the real world comes with challenges in calibrating user trust and expectations. As a step toward developing RL systems that are able to communicate their competencies, we present a method of…

机器学习 · 计算机科学 2020-11-19 Aastha Acharya , Rebecca Russell , Nisar R. Ahmed

Reinforcement learning (RL) algorithms find applications in inventory control, recommender systems, vehicular traffic management, cloud computing and robotics. The real-world complications of many tasks arising in these domains makes them…

机器学习 · 计算机科学 2021-06-03 Sindhu Padakandla

In extreme value statistics, the peaks-over-threshold method is widely used. The method is based on the generalized Pareto distribution characterizing probabilities of exceedances over high thresholds in $\mathbb {R}^d$. We present a…

概率论 · 数学 2014-10-17 Ana Ferreira , Laurens de Haan

Reinforcement learning (RL) has become a foundational approach for enabling intelligent robotic behavior in dynamic and uncertain environments. This work presents an in-depth review of RL principles, advanced deep reinforcement learning…

机器人学 · 计算机科学 2026-03-17 Kumater Ter , Abolanle Adetifa , Daniel Udekwe

We present an AI-based ecosystem simulator that uses three-dimensional models of the terrain and animal models controlled by deep reinforcement learning. The simulations take place in a game engine environment, which enables continuous…

多智能体系统 · 计算机科学 2023-03-22 Claes Strannegård , Niklas Engsner , Rasmus Lindgren , Simon Olsson , John Endler

We exploit a bijection between plane recursive trees and Stirling permutations; this yields the equivalence of some results previously proven separately by different methods for the two types of objects as well as some new results. We also…

组合数学 · 数学 2008-03-10 Svante Janson

We consider a self-attracting random walk in dimension d=1, in presence of a field of strength s, which biases the walker toward a target site. We focus on the dynamic case (true reinforced random walk), where memory effects are implemented…

统计力学 · 物理学 2015-06-05 Elena Agliari , Raffaella Burioni , Guido Uguzzoni

In this survey, we propose an overview on Lyapunov functions for a variety of compartmental models in epidemiology. We exhibit the most widely employed functions, together with a commentary on their use. Our aim is to provide a…

动力系统 · 数学 2023-06-09 Nicolò Cangiotti , Marco Capolli , Mattia Sensi , Sara Sottile

In practical applications, we can rarely assume full observability of a system's environment, despite such knowledge being important for determining a reactive control system's precise interaction with its environment. Therefore, we propose…

机器学习 · 计算机科学 2022-06-24 Edi Muskardin , Martin Tappler , Bernhard K. Aichernig , Ingo Pill

Rapid advances of hardware-based technologies during the past decades have opened up new possibilities for Life scientists to gather multimodal data in various application domains (e.g., Omics, Bioimaging, Medical Imaging, and…

机器学习 · 计算机科学 2018-01-09 Mufti Mahmud , M. Shamim Kaiser , Amir Hussain , Stefano Vassanelli

Efficiently tackling multiple tasks within complex environment, such as those found in robot manipulation, remains an ongoing challenge in robotics and an opportunity for data-driven solutions, such as reinforcement learning (RL).…

机器人学 · 计算机科学 2024-04-03 Carlos Plou , Ana C. Murillo , Ruben Martinez-Cantin

We consider a version of the classical P\'olya urn scheme which incorporates innovations. The space $S$ of colors is an arbitrary measurable set. After each sampling of a ball in the urn, one returns $C$ balls of the same color and…

概率论 · 数学 2022-11-17 Jean Bertoin

Vertex-reinforced random walk is defined in Pemantle's (1988) thesis; it is a random walk that is biased to visit sites it has already visited a lot. We show that this reinforcement scheme, in contrast to the scheme of edge-reinforcement,…

概率论 · 数学 2016-09-07 Robin Pemantle , Stanislav Volkov

Consistently checking the statistical significance of experimental results is the first mandatory step towards reproducible science. This paper presents a hitchhiker's guide to rigorous comparisons of reinforcement learning algorithms.…

统计方法学 · 统计学 2022-08-30 Cédric Colas , Olivier Sigaud , Pierre-Yves Oudeyer

A general random effects model is proposed that allows for continuous as well as discrete distributions of the responses. Responses can be unrestricted continuous, bounded continuous, binary, ordered categorical or given in the form of…

统计方法学 · 统计学 2024-04-30 Gerhard Tutz

Understanding pedestrian behavior is crucial for the safe deployment of Autonomous Vehicles (AVs) in urban environments. Traditional pedestrian behavior models often fall into two categories: mechanistic models, which do not generalize well…

人机交互 · 计算机科学 2024-09-24 Yueyang Wang , Aravinda Ramakrishnan Srinivasan , Yee Mun Lee , Gustav Markkula

Extreme shock models have been introduced in Gut and H\"usler (1999) to study systems that at random times are subject to shock of random magnitude. These systems break down when some shock overcomes a given resistance level. In this paper…

其他统计学 · 统计学 2010-10-19 Pasquale Cirillo , Jürg Hüsler

Process control is widely discussed in the manufacturing process, especially for semiconductor manufacturing. Due to unavoidable disturbances in manufacturing, different process controllers are proposed to realize variation reduction. Since…

系统与控制 · 电气工程与系统科学 2021-10-25 Yanrong Li , Juan Du , Wei Jiang
‹ 上一页 1 8 9 10 下一页 ›