English
Related papers

Related papers: Strongly reinforced P\'olya urns with graph-based …

200 papers

Behavioral experiments on the ultimatum game (UG) reveal that we humans prefer fair acts, which contradicts the prediction made in orthodox Economics. Existing explanations, however, are mostly attributed to exogenous factors within the…

Machine Learning · Computer Science 2026-02-04 Guozhong Zheng , Jiqiang Zhang , Xin Ou , Shengfeng Deng , Li Chen

Growing concerns regarding the operational usage of AI models in the real-world has caused a surge of interest in explaining AI models' decisions to humans. Reinforcement Learning is not an exception in this regard. In this work, we propose…

Machine Learning · Computer Science 2023-10-06 Omid Davoodi , Majid Komeili

Deep reinforcement learning has shown remarkable success in the past few years. Highly complex sequential decision making problems from game playing and robotics have been solved with deep model-free methods. Unfortunately, the sample…

Machine Learning · Computer Science 2021-07-20 Aske Plaat , Walter Kosters , Mike Preuss

Reinforcement learning is a powerful paradigm for learning optimal policies from experimental data. However, to find optimal policies, most reinforcement learning algorithms explore all possible actions, which may be harmful for real-world…

Machine Learning · Statistics 2017-11-15 Felix Berkenkamp , Matteo Turchetta , Angela P. Schoellig , Andreas Krause

The subject of this paper is reinforcement learning. Policies are considered here that produce actions based on states and random elements autocorrelated in subsequent time instants. Consequently, an agent learns from experiments that are…

Machine Learning · Computer Science 2020-09-11 Marcin Szulc , Jakub Łyskawa , Paweł Wawrzyński

We consider Reinforced Random Walks where transition probabilities are a function of the proportion of times the walk has traversed an edge. We give conditions for recurrence or transience. A phase transition is observed, similar to…

Probability · Mathematics 2009-07-15 Olivier Raimond , Bruno Schapira

We examine the question of when and how parametric models are most useful in reinforcement learning. In particular, we look at commonalities and differences between parametric models and experience replay. Replay-based learning algorithms…

Machine Learning · Computer Science 2019-09-18 Hado van Hasselt , Matteo Hessel , John Aslanides

We consider a linearly edge-reinforced random walk on a class of two-dimensional graphs with constant initial weights. The graphs are obtained from $\mathbb{Z}^2$ by replacing every edge by a sufficiently large, but fixed number of edges in…

Probability · Mathematics 2009-10-13 Franz Merkl , Silke W. W. Rolles

Consider two urns, $A$ and $B$, where initially $A$ contains a large number $n$ of balls and $B$ is empty. At each step, with equal probability, either we pick a ball at random in $A$ and place it in $B$, or vice-versa (provided of course…

Probability · Mathematics 2010-07-26 Jean Bertoin

Recently, a novel machine learning model has emerged in the field of reinforcement learning known as deep Q-learning. This model is capable of finding the best possible solution in systems consisting of millions of choices, without ever…

Image and Video Processing · Electrical Eng. & Systems 2018-10-26 Iman Sajedian , Trevon Badloe , Junsuk Rho

Human beings are particularly good at reasoning and inference from just a few examples. When facing new tasks, humans will leverage knowledge and skills learned before, and quickly integrate them with the new task. In addition to learning…

Artificial Intelligence · Computer Science 2019-09-30 Hua Huang , Adrian Barbu

Abstraction is key to scaling up reinforcement learning (RL). However, autonomously learning abstract state and action representations to enable transfer and generalization remains a challenging open problem. This paper presents a novel…

Artificial Intelligence · Computer Science 2024-12-24 Rashmeet Kaur Nayyar , Siddharth Srivastava

As Evolutionary Dynamics moves from the realm of theory into application, algorithms are needed to move beyond simple models. Yet few such methods exist in the literature. Ecological and physiological factors are known to be central to…

Populations and Evolution · Quantitative Biology 2025-05-20 Bryce Allen Bagley , Navin Khoshnan , Claudia K Petritsch

An edge-colouring is {\em strong} if every colour class is an induced matching. In this work we give a formulae that determines either the optimal or the optimal plus one strong chromatic index of bipartite outerplanar graphs. Further, we…

Discrete Mathematics · Computer Science 2013-12-20 Valentin Borozan , Leandro Montero , Narayanan Narayanan

We provide a framework for exploring physics beyond the Standard Model with reinforcement learning using graph representations of new physics theories. The graph structure allows for model-building without a priori specifying definite…

High Energy Physics - Phenomenology · Physics 2024-07-11 George N. Wojcik , Shu Tian Eu , Lisa L. Everett

In this paper we introduce the subdistribution beta-Stacy process, a novel Bayesian nonparametric process prior for subdistribution functions useful for the analysis of competing risks data. In particular, we i) characterize this process…

Statistics Theory · Mathematics 2018-11-30 Andrea Arfé , Stefano Peluso , Pietro Muliere

In Reinforcement Learning we look for meaning in the flow of input/output information. If we do not find meaning, the information flow is not more than noise to us. Before we are able to find meaning, we should first learn how to discover…

Machine Learning · Computer Science 2019-06-27 Dimiter Dobrev

In the first part of this paper, we consider weighted domination in the case where the vertices of the complete graph on~\(n\) vertices are equipped with independent and identically distributed (i.i.d.) weights. We use the probabilistic…

Probability · Mathematics 2023-01-16 Ghurumuruhan Ganesan

The significance of network structures in promoting group cooperation within social dilemmas has been widely recognized. Prior studies attribute this facilitation to the assortment of strategies driven by spatial interactions. Although…

Multiagent Systems · Computer Science 2024-08-20 Tianyu Ren , Xiao-Jun Zeng

In this short paper, we consider the Once-reinforced random walk with reinforcement parameter $a$ on trees with bounded degree which are transient for the simple random walk. On each of these trees, we prove that there exists an explicit…

Probability · Mathematics 2017-08-23 Daniel Kious , Vladas Sidoravicius