English
Related papers

Related papers: Positive reinforced generalized time-dependent P\'…

200 papers

This is the second part of a two-part investigation. We continue the study of a class of balanced urn schemes on balls of two colors (white and black). At each drawing, a sample of size $m\ge 1$ is drawn from the urn and ball addition rules…

Probability · Mathematics 2015-10-01 Markus Kuba , Hosam M. Mahmoud

This article presents a short and concise description of stochastic approximation algorithms in reinforcement learning of Markov decision processes. The algorithms can also be used as a suboptimal method for partially observed Markov…

Optimization and Control · Mathematics 2015-12-25 Vikram Krishnamurthy

Our main result is to prove almost-sure convergence of a stochastic-approximation algorithm defined on the space of measures on a non-compact space. Our motivation is to apply this result to measure-valued P\'olya processes (MVPPs, also…

Probability · Mathematics 2020-01-22 Cécile Mailler , Denis Villemonais

In the apportionment problem, a fixed number of seats must be distributed among parties in proportion to the number of voters supporting each party. We study a generalization of this setting, in which voters can support multiple parties by…

Computer Science and Game Theory · Computer Science 2022-03-31 Markus Brill , Paul Gölz , Dominik Peters , Ulrike Schmidt-Kraepelin , Kai Wilker

Competing urns refers to the random experiment where m balls are dropped, randomly and independently, into urns 1,...,n. Formally, we have a random map $\sigma$ from {1,...,m} to {1,...,n} with the $\sigma(i)$'s i.i.d. With $x_j$ the…

Probability · Mathematics 2010-01-06 Jeff Kahn , Michael Neiman

We introduce a new model for contagion spread using a network of interacting finite memory two-color P\'{o}lya urns, which we refer to as the finite memory interacting P\'{o}lya contagion network. The urns interact in the sense that the…

Dynamical Systems · Mathematics 2021-12-22 Somya Singh , Fady Alajaji , Bahman Gharesifard

We consider a version of the classical Ehrenfest urn model with two urns and two types of balls: regular and heavy. Each ball is selected independently according to a Poisson process having rate $1$ for regular balls and rate…

Probability · Mathematics 2024-07-12 Matteo Quattropani

We present a distributional approach to theoretical analyses of reinforcement learning algorithms for constant step-sizes. We demonstrate its effectiveness by presenting simple and unified proofs of convergence for a variety of…

Machine Learning · Computer Science 2020-03-30 Philip Amortila , Doina Precup , Prakash Panangaden , Marc G. Bellemare

Consider an urn containing balls labeled with integer values. Define a discrete-time random process by drawing two balls, one at a time and with replacement, and noting the labels. Add a new ball labeled with the sum of the two drawn…

Probability · Mathematics 2023-06-22 Mackenzie Simper

We study first passage statistics of the Polya urn model. In this random process, the urn contains two types of balls. In each step, one ball is drawn randomly from the urn, and subsequently placed back into the urn together with an…

Statistical Mechanics · Physics 2010-07-12 Tibor Antal , E. Ben-Naim , P. L. Krapivsky

The paper deals with the problem of finding the best alternatives on the basis of pairwise comparisons when these comparisons need not be transitive. In this setting, we study a reinforcement urn model. We prove convergence to the optimal…

Optimization and Control · Mathematics 2013-01-25 Benoit Laslier , Jean-Francois Laslier

A basic experiment in probability theory is drawing without replacement from an urn filled with multiple balls of different colours. Clearly, it is physically impossible to overdraw, that is, to draw more balls from the urn than it…

Probability · Mathematics 2023-12-21 Bart Jacobs , Dario Stein

Reinforcement learning (RL) problems are fundamental in online decision-making and have been instrumental in finding an optimal policy for Markov decision processes (MDPs). Function approximations are usually deployed to handle large or…

Machine Learning · Computer Science 2025-05-20 Jiashuo Jiang , Yiming Zong , Yinyu Ye

Value-function approximation methods that operate in batch mode have foundational importance to reinforcement learning (RL). Finite sample guarantees for these methods often crucially rely on two types of assumptions: (1) mild distribution…

Machine Learning · Computer Science 2019-05-02 Jinglin Chen , Nan Jiang

Consider throwing $n$ balls at random into $m$ urns, each ball landing in urn $i$ with probability $p_i$. Let $S$ be the resulting number of singletons, i.e., urns containing just one ball. We give an error bound for the Kolmogorov distance…

Probability · Mathematics 2009-01-23 Mathew D. Penrose

We study the phase transition and the critical properties of a nonlinear P\'{o}lya urn, which is a simple binary stochastic process $X(t)\in \{0,1\},t=1,\cdots$ with a feedback mechanism. Let $f$ be a continuous function from the unit…

Statistical Mechanics · Physics 2021-07-21 Kazuaki Nakayama , Shintaro Mori

We study a discrete-time Markov process $X_n\in\mathbb{R}^d$, for which the distribution of the future increments depends only on the relative ranking of its components (descending order by value). We endow the process with a…

Probability · Mathematics 2021-05-04 Pantelis P. Analytis , Alexandros Gelastopoulos , Hrvoje Stojic

The asymptotic behaviour of a generalised P\'olya--Eggenberger urn is well--known to depend on the spectrum of its replacement matrix: If its dominant eigenvalue $r$ is simple and no other eigenvalue is `large' in the sense that its real…

Probability · Mathematics 2019-03-13 Noela Müller

Consider the multicolored urn model where, after every draw, balls of the different colors are added to the urn in a proportion determined by a given stochastic replacement matrix. We consider some special replacement matrices which are not…

Probability · Mathematics 2009-02-09 Arup Bose , Amites Dasgupta , Krishanu Maulik

Using P\'{o}lya's urn model with negative replacement we introduce a new Bernstein-type operator and we show that the new operator improves upon the known estimates for the classical Bernstein operator. We also provide numerical evidence…

Classical Analysis and ODEs · Mathematics 2017-10-25 Mihai N. Pascu , Nicolae R. Pascu , Florenţa Tripşa