English
Related papers

Related papers: Positive reinforced generalized time-dependent P\'…

200 papers

This paper presents the link between stochastic approximation and clinical trials based on randomized urn models investigated in Bai and Hu (1999,2005) and Bai, Hu and Shen (2002). We reformulate the dynamics of both the urn composition and…

Probability · Mathematics 2017-01-19 Sophie Laruelle , Gilles Pagès

This thesis examines edge-reinforced random walks with some modifications to the standard definition. An overview of known results relating to the standard model is given and the proof of recurrence for the standard linearly edge-reinforced…

Probability · Mathematics 2023-09-07 Fabian Michel

We propose a reinforcement learning (RL) framework under a broad class of risk objectives, characterized by convex scoring functions. This class covers many common risk measures, such as variance, Expected Shortfall, entropic Value-at-Risk,…

Mathematical Finance · Quantitative Finance 2025-05-16 Shanyu Han , Yang Liu , Xiang Yu

In this work we introduce a new type of urn model with infinite but countable many colors indexed by an appropriate infinite set. We mainly consider the indexing set of colors to be the $d$-dimensional integer lattice and consider balanced…

Probability · Mathematics 2018-01-09 Antar Bandyopadhyay , Debleena Thacker

We introduce a multi-colour multi-urn generalisation of the Bernoulli-Laplace urn model, consisting of $d$ urns, $m$ colours, and $dmn$ balls, with $dn$ balls of each colour and $mn$ balls in each urn. At each step, one ball is drawn…

Probability · Mathematics 2025-11-14 Ritesh Goenka , Jonathan Hermon , Dominik Schmid

Given a stochastic structure with a filtration $\mathbb{F}$, the class of all random times whose conditional distribution functions are differentiable with respect to some $\mathbb{F}$ adapted non decreasing processes is considered. The…

Probability · Mathematics 2013-12-20 Shiqi Song

We introduce a model of graph-constrained dynamic choice with reinforcement modeled by positively $\alpha$-homogeneous rewards. We show that its empirical process, which can be written as a stochastic approximation recursion with Markov…

Optimization and Control · Mathematics 2021-07-27 Konstantin Avrachenkov , Vivek S. Borkar , Sharayu Moharir , Suhail M. Shah

A functional central limit theorem is established for weighted occupancy processes of the Karlin model. The weighted occupancy processes take the form of, with $D_{n,j}$ denoting the number of urns with $j$-balls after the first $n$…

Probability · Mathematics 2025-04-22 Jaime Garza , Yizao Wang

We propose an automata-theoretic approach for reinforcement learning (RL) under complex spatio-temporal constraints with time windows. The problem is formulated using a Markov decision process under a bounded temporal logic constraint.…

Artificial Intelligence · Computer Science 2023-08-01 Xiaoshan Lin , Abbasali Koochakzadeh , Yasin Yazicioglu , Derya Aksaray

We consider a collection of weighted Euclidian random balls in R^d distributed according a determinantal point process. We perform a zoom-out procedure by shrinking the radii while increasing the number of balls. We observe that the…

Probability · Mathematics 2019-07-24 Adrien Clarenne

We introduce a simple but powerful technique to study processes driven by two or more reinforcement mechanisms in competition. We apply our method to two types of models: to non conservative zero range processes on finite graphs, and to…

Probability · Mathematics 2022-06-30 Dirk Erhard , Guilherme Reis

Sufficient conditions are developed for a class of generalized Polya urn schemes ensuring exchangeability. The extended class includes the Blackwell-MacQueen Polya urn and the urn schemes for the two-parameter Poisson-Dirichlet process and…

Probability · Mathematics 2007-05-23 Hemant Ishwaran , Mahmoud Zarepour

A succesful method to describe the asymptotic behavior of a discrete time stochastic process governed by some recursive formula is to relate it to the limit sets of a well chosen mean differential equation. Under an attainability condition,…

Probability · Mathematics 2011-01-19 Mathieu Faure , Gregory Roth

We develop a new class of model-free deep reinforcement learning algorithms for data-driven, learning-based control. Our Generalized Policy Improvement algorithms combine the policy improvement guarantees of on-policy methods with the…

Machine Learning · Computer Science 2024-10-15 James Queeney , Ioannis Ch. Paschalidis , Christos G. Cassandras

Urn models play an important role to express various basic ideas in probability theory. Here we extend this urn model with tubes. An urn contains coloured balls, which can be drawn with probabilities proportional to the numbers of balls of…

Probability · Mathematics 2024-08-07 Bart Jacobs

We consider predictive inference using a class of temporally dependent Dirichlet processes driven by Fleming--Viot diffusions, which have a natural bearing in Bayesian nonparametrics and lend the resulting family of random probability…

Methodology · Statistics 2020-01-28 Filippo Ascolani , Antonio Lijoi , Matteo Ruggiero

The randomized play-the-winner (RPW) model is a generalized P\'olya Urn process with broad applications ranging from clinical trials to molecular evolution. We derive an exact expression for the variance of the RPW model by transforming the…

Applications · Statistics 2024-01-02 Ivan Specht , Michael Mitzenmacher

Generalized Polya urn models can describe the dynamics of finite populations of interacting genotypes. Three basic questions these models can address are: Under what conditions does a population exhibit growth? On the event of growth, at…

Probability · Mathematics 2007-05-23 Michel Benaim , Sebastian J. Schreiber , Pierre Tarres

We construct an independent increments Gaussian process associated to a class of multicolor urn models. The construction uses random variables from the urn model which are different from the random variables for which central limit theorems…

Probability · Mathematics 2007-05-23 Gopal K Basak , Amites Dasgupta

Generalized probability distributions for Maxwell-Boltzmann, Bose-Einstein and Fermi-Dirac statistics, with unequal source probabilities $q_i$ for each level $i$, are obtained by combinatorial reasoning. For equiprobable degenerate…

Statistical Mechanics · Physics 2008-08-18 Robert K. Niven , Marian Grendar
‹ Prev 1 4 5 6 7 8 10 Next ›