English
Related papers

Related papers: Positive reinforced generalized time-dependent P\'…

200 papers

We give a central limit theorem, which has applications to Bayesian statistics and urn problems. The latter are investigated, by paying special attention to multicolor randomly reinforced generalized Polya urns.

Probability · Mathematics 2009-04-27 Patrizia Berti , Irene Crimaldi , Luca Pratelli , Pietro Rigo

This paper explores how deliberate mutations of reward function in reinforcement learning can produce diversified skill variations in robotic manipulation tasks, examined with a liquid pouring use case. To this end, we developed a new…

Robotics · Computer Science 2025-09-24 Jannick van Buuren , Roberto Giglio , Loris Roveda , Luka Peternel

Classical voting rules assume that ballots are complete preference orders over candidates. However, when the number of candidates is large enough, it is too costly to ask the voters to rank all candidates. We suggest to fix a rank k, to ask…

Computer Science and Game Theory · Computer Science 2020-02-17 Manel Ayadi , Nahla Ben amor , Jérôme Lang

We introduce and discuss a special type of feedback interacting urn model with deterministic interaction. This is a generalisation of the very well known Eggenberger and Polya (1923) urn model. In our model, balls are added to a particular…

Probability · Mathematics 2022-11-15 Krishanu Maulik , Manit Paul

A step-reinforced random walk is a discrete-time stochastic process with long-range dependence. At each step, with a fixed probability $\alpha$, the so-called positively step-reinforced random walk repeats one of its previous steps, chosen…

Probability · Mathematics 2025-05-01 Rafik Aguech , Samir Ben Hariz , Mohamed El Machkouri , Youssef Faouzi

Stochastic Taylor expansions of the expectation of functionals applied to diffusion processes which are solutions of stochastic differential equation systems are introduced. Taylor formulas w.r.t. increments of the time are presented for…

Probability · Mathematics 2013-10-24 Andreas Rößler

This thesis rigorously studies fundamental reinforcement learning (RL) methods in modern practical considerations, including robust RL, distributional RL, and offline RL with neural function approximation. The thesis first prepares the…

Machine Learning · Computer Science 2022-03-04 Thanh Nguyen-Tang

It is a classical result in rational approximation theory that certain non-smooth or singular functions, such as $|x|$ and $x^{1/p}$, can be efficiently approximated using rational functions with root-exponential convergence in terms of…

Numerical Analysis · Mathematics 2025-06-27 Kingsley Yeon , Steven B. Damelin

In this paper we show how to extend the Sample-Path Large Deviation Principle for the urn model of Hill, Lane and Sudderth to the case in which the increment of the urn is not a binary variable. In particular, we sketch how to modify the…

Probability · Mathematics 2025-11-19 Simone Franchini

We revisit the random allocation model in which $n$ balls are independently placed into $N$ boxes with probabilities $q_1,\ldots,q_N$. A classical asymptotic result due to Kolchin, Sevastyanov, and Chistyakov for the expectations,…

Probability · Mathematics 2026-04-28 Serik Sagitov

For the interacting urn model with polynomial reinforcement, it has been conjectured that almost surely one color monopolizes all the urns if the interaction parameter $p>0$. We disprove the conjecture. For the case $p=1$, we give a…

Probability · Mathematics 2024-10-08 Shuo Qin

Motivated by engineering applications such as resource allocation in networks and inventory systems, we consider average-reward Reinforcement Learning with unbounded state space and reward function. Recent works studied this problem in the…

Machine Learning · Computer Science 2025-11-10 Shaan Ul Haque , Siva Theja Maguluri

Consider a coin tossing experiment which consists of tossing one of two coins at a time, according to a renewal process. The first coin is fair and the second has probability $1/2 + \theta$, $\theta \in [-1/2,1/2]$, $\theta$ unknown but…

Probability · Mathematics 2019-03-25 Diego Marcondes , Cláudia Peixoto

We study the computational complexity of approximating general constrained Markov decision processes. Our primary contribution is the design of a polynomial time $(0,\epsilon)$-additive bicriteria approximation algorithm for finding optimal…

Data Structures and Algorithms · Computer Science 2025-02-12 Jeremy McMahan

A step-reinforced random walk is a discrete-time non-Markovian process with long range memory. At each step, with a fixed probability p, the positively step-reinforced random walk repeats one of its preceding steps chosen uniformly at…

Probability · Mathematics 2023-11-28 Zhishui Hu , Yiting Zhang

Conditional identity in distribution (Berti et al. (2004)) is a new type of dependence for random variables, which generalizes the well-known notion of exchangeability. In this paper, a class of random sequences, called Generalized Species…

Probability · Mathematics 2008-06-18 Federico Bassetti , Irene Crimaldi , Fabrizio Leisen

Eldan's stochastic localization is a probabilistic construction that has proved instrumental to modern breakthroughs in high-dimensional geometry and the design of sampling algorithms. Motivated by sampling under non-Euclidean geometries…

Probability · Mathematics 2026-03-18 Anming Gu , Bobby Shi , Kevin Tian

We analyse the balls in bins process with feedback with primary focus on the power law feedback function $f(\omega)=\eta \omega^{\gamma}\,$, $\eta>0\,$ $\gamma \geq0\,$. Using the recursive solution to the master equation we find for power…

Probability · Mathematics 2023-08-22 Samuel Forbes

Covariate-adaptive randomization (CAR) procedures are frequently used in comparative studies to increase the covariate balance across treatment groups. However, because randomization inevitably uses the covariate information when forming…

Statistics Theory · Mathematics 2022-07-08 Wei Ma , Yichen Qin , Yang Li , Feifang Hu

Generating random variates from high-dimensional distributions is often done approximately using Markov chain Monte Carlo. In certain cases, perfect simulation algorithms exist that allow one to draw exactly from the stationary…

Data Structures and Algorithms · Computer Science 2017-01-05 Mark Huber
‹ Prev 1 8 9 10 Next ›