中文
相关论文

相关论文: A survey of random processes with reinforcement

200 篇论文

A P\'olya urn process is a Markov chain that models the evolution of an urn containing some coloured balls, the set of possible colours being $\{1,\ldots,d\}$ for $d\in \mathbb{N}$. At each time step, a random ball is chosen uniformly in…

概率论 · 数学 2017-03-13 Cécile Mailler , Jean-François Marckert

The processes of the averaged regression quantiles and of their modifications provide useful tools in the regression models when the covariates are not fully under our control. As an application we mention the probabilistic risk assessment…

统计理论 · 数学 2017-10-19 Jana Jurečková , Martin Schindler , Jan Picek

As Evolutionary Dynamics moves from the realm of theory into application, algorithms are needed to move beyond simple models. Yet few such methods exist in the literature. Ecological and physiological factors are known to be central to…

种群与进化 · 定量生物学 2025-05-20 Bryce Allen Bagley , Navin Khoshnan , Claudia K Petritsch

We consider a system of urns of Polya-type, with balls of two colors; the reinforcement of each urn depends both on the content of the same urn and on the average content of all urns. We show that the urns synchronize almost surely, in the…

概率论 · 数学 2016-03-08 Paolo Dai Pra , Pierre-Yves Louis , Ida G. Minelli

Simulations play important and diverse roles in statistical workflows, for example, in model specification, checking, validation, and even directly in model inference. Over the past decades, the application areas and overall potential of…

统计计算 · 统计学 2025-08-27 Paul-Christian Bürkner , Marvin Schmitt , Stefan T. Radev

Reinforcement Learning (RL) has become a critical tool for optimization challenges within automation, leading to significant advancements in several areas. This review article examines the current landscape of RL within automation, with a…

机器学习 · 计算机科学 2025-03-05 Ahmad Farooq , Kamran Iqbal

In reinforcement learning (RL), agents continually interact with the environment and use the feedback to refine their behavior. To guide policy optimization, reward models are introduced as proxies of the desired objectives, such that when…

机器学习 · 计算机科学 2025-06-19 Rui Yu , Shenghua Wan , Yucen Wang , Chen-Xiao Gao , Le Gan , Zongzhang Zhang , De-Chuan Zhan

In recent years, a specific machine learning method called deep learning has gained huge attraction, as it has obtained astonishing results in broad applications such as pattern recognition, speech recognition, computer vision, and natural…

机器学习 · 计算机科学 2018-06-26 Seyed Sajad Mousavi , Michael Schukat , Enda Howley

This review article provides an overview of recent work in the modeling and analysis of recurrent events arising in engineering, reliability, public health, biomedicine and other areas. Recurrent event modeling possesses unique facets…

统计方法学 · 统计学 2007-08-03 Edsel A. Peña

We analyze the Brownian Motion limit of a prototypical unit step reinforced random-walk on the half line. A reinforced random walk is one which changes the weight of any edge (or vertex) visited to increase the frequency of return visits.…

概率论 · 数学 2013-10-02 Jerome K. Percus , Ora E. Percus

The edge-reinforced random walk (ERRW) is a random process on the vertices of a graph that is more likely to cross the edges it has visited in the past. Depending on the strength of the reinforcement, the ERRW of a single particle can…

概率论 · 数学 2025-09-23 Giordano Giambartolomei , Nadia Sidorova

We present the first rigorous quantitative analysis of once-reinforced random walks (ORRW) on general graphs, based on a novel change of measure formula.~This enables us to prove large deviations estimates for the range of the walk to have…

概率论 · 数学 2025-09-05 Andrea Collevecchio , Pierre Tarrès

Regression has attracted immense interest lately due to its effectiveness in tasks like predicting values. And Regression is of widespread use in multiple fields such as Economics, Finance, Business, Biology and so on. While considerable…

机器学习 · 计算机科学 2021-04-27 Yunpeng Tai

The challenge of spatial resource allocation is pervasive across various domains such as transportation, industry, and daily life. As the scale of real-world issues continues to expand and demands for real-time solutions increase,…

机器学习 · 计算机科学 2024-03-08 Di Zhang , Moyang Wang , Joseph Mango , Xiang Li , Xianrui Xu

Reinforcement learning has recently experienced increased prominence in the machine learning community. There are many approaches to solving reinforcement learning problems with new techniques developed constantly. When solving problems…

机器学习 · 计算机科学 2020-12-14 Belinda Stapelberg , Katherine M. Malan

With large-scale integration of renewable generation and distributed energy resources, modern power systems are confronted with new operational challenges, such as growing complexity, increasing uncertainty, and aggravating volatility.…

机器学习 · 计算机科学 2022-02-28 Xin Chen , Guannan Qu , Yujie Tang , Steven Low , Na Li

We revisit an unpublished paper of Vervoort (2002) on the once reinforced random walk, and prove that this process is recurrent on any graph of the form $\mathbb{Z}\times \Gamma$, with $\Gamma$ a finite graph, for sufficiently large…

概率论 · 数学 2018-07-20 Daniel Kious , Bruno Schapira , Arvind Singh

Reinforcement learning (RL) has achieved remarkable success in real-world decision-making across diverse domains, including gaming, robotics, online advertising, public health, and natural language processing. Despite these advances, a…

应用统计 · 统计学 2026-01-23 Asim H. Gazi , Yongyi Guo , Daiqi Gao , Ziping Xu , Kelly W. Zhang , Susan A. Murphy

Reinforcement Learning (RL) has emerged as a powerful paradigm in Artificial Intelligence (AI), enabling agents to learn optimal behaviors through interactions with their environments. Drawing from the foundations of trial and error, RL…

人工智能 · 计算机科学 2025-02-04 Majid Ghasemi , Amir Hossein Moosavi , Dariush Ebrahimi

Computer simulations of amphiphilic systems are reviewed. Research areas cover a wide range of length and time scales, and a whole hierarchy of models and methods has been developed to address them all. They range from atomistically…

软凝聚态物质 · 物理学 2009-09-25 Friederike Schmid