English
Related papers

Related papers: Reinforced Loop Soup via Wilson's Algorithm

200 papers

We develop a tree boosting algorithm for collider measurements of multiple Wilson coefficients in effective field theories describing phenomena beyond the standard model of particle physics. The design of the discriminant exploits per-event…

High Energy Physics - Phenomenology · Physics 2022-05-27 Suman Chatterjee , Stefan Rohshap , Robert Schöfbeck , Dennis Schwarz

We define a linearly reinforced process called the *-Edge-Reinforced Random Walk (*-ERRW ) which can be seen as a Yaglom reversible, hence non-reversible, extension of the Edge-Reinforced Random Walk (ERRW) introduced by Coppersmith and…

Probability · Mathematics 2023-11-30 Sergio Bacallado , Christophe Sabot , Pierre Tarrès

We propose randomized least-squares value iteration (RLSVI) -- a new reinforcement learning algorithm designed to explore and generalize efficiently via linearly parameterized value functions. We explain why versions of least-squares value…

Machine Learning · Statistics 2016-02-16 Ian Osband , Benjamin Van Roy , Zheng Wen

We consider a one-dimensional simple random walk surviving among a field of static soft traps : each time it meets a trap the walk is killed with probability 1--e --$\beta$ , where $\beta$ is a positive and fixed parameter. The positions of…

Probability · Mathematics 2018-10-02 Julien Poisat , François Simenhaus

We present a comparative study of several algorithms for an in-plane random walk with a variable step. The goal is to check the efficiency of the algorithm in the case where the random walk terminates at some boundary. We recently found…

Statistical Mechanics · Physics 2019-04-17 Olga Klimenkova , Anton Yu. Menshutin , Lev N. Shchur

In this paper, we explore deep reinforcement learning algorithms for vision-based robotic grasping. Model-free deep reinforcement learning (RL) has been successfully applied to a range of challenging environments, but the proliferation of…

Robotics · Computer Science 2018-03-30 Deirdre Quillen , Eric Jang , Ofir Nachum , Chelsea Finn , Julian Ibarz , Sergey Levine

In this paper we introduce a new simple but powerful general technique for the study of edge- and vertex-reinforced processes with super-linear reinforcement, based on the use of order statistics for the number of edge, respectively of…

Probability · Mathematics 2016-06-03 Codina Cotar , Debleena Thacker

We develop an algorithm for efficiently computing recursively defined functions on posets. We illustrate this algorithm by disproving conjectures about the game Subset Takeaway (Chomp on a hypercube) and computing the number of linear…

Combinatorics · Mathematics 2017-07-11 Andries E. Brouwer , J. Daniel Christensen

We propose an exact iterative algorithm for minimization of a class of continuous cell-wise linear convex functions on a hyperplane arrangement. Our particular setup is motivated by evaluation of so-called rank estimators used in robust…

Optimization and Control · Mathematics 2020-01-01 Michal Černý , Milan Hladík , Miroslav Rada

Recently, several works have shown that natural modifications of the classical conditional gradient method (aka Frank-Wolfe algorithm) for constrained convex optimization, provably converge with a linear rate when: i) the feasible set is a…

Optimization and Control · Mathematics 2016-05-23 Dan Garber , Ofer Meshi

We propose a long-term memory design for artificial general intelligence based on Solomonoff's incremental machine learning methods. We use R5RS Scheme and its standard library with a few omissions as the reference machine. We introduce a…

Artificial Intelligence · Computer Science 2015-03-19 Eray Özkural

We introduce an algorithm that conjectures the structure of a permutation class in the form of a disjoint cover of "rules"; similar to generalized grid classes. The cover is usually easily verified by a human and translated into an…

Combinatorics · Mathematics 2017-05-12 Christian Bean , Bjarki Gudmundsson , Henning Ulfarsson

Inverse reinforcement learning aims to infer the reward function that explains expert behavior observed through trajectories of state--action pairs. A long-standing difficulty in classical IRL is the non-uniqueness of the recovered reward:…

Machine Learning · Statistics 2025-12-09 Denis Belomestny , Alexey Naumov , Sergey Samsonov

This paper introduces an algorithm for discovering implicit and delayed causal relations between events observed by a robot at arbitrary times, with the objective of improving data-efficiency and interpretability of model-based…

Machine Learning · Computer Science 2020-08-05 Junchi Liang , Abdeslam Boularias

In this note we propose a new approach towards solving numerically optimal stopping problems via reinforced regression based Monte Carlo algorithms. The main idea of the method is to reinforce standard linear regression algorithms in each…

Numerical Analysis · Mathematics 2019-07-02 Denis Belomestny , John Schoenmakers , Vladimir Spokoiny , Bakhyt Zharkynbay

We consider the Reinforcement Learning problem of controlling an unknown dynamical system to maximise the long-term average reward along a single trajectory. Most of the literature considers system interactions that occur in discrete time…

Artificial Intelligence · Computer Science 2023-09-07 Lorenzo Croissant , Marc Abeille , Bruno Bouchard

Reinforced Galton-Watson processes have been introduced in arxiv:2306.02476 as population models with non-overlapping generations, such that reproduction events along genealogical lines can be repeated at random. We investigate here some of…

Probability · Mathematics 2024-08-15 Jean Bertoin , Bastien Mallein

We prove that the restriction of the vertex-reinforced jump process to a subset of the vertex set is a mixture of vertex-reinforced jump processes. A similar statement holds for the non-linear hyperbolic supersymmetric sigma model. This is…

Probability · Mathematics 2024-11-12 Margherita Disertori , Franz Merkl , Silke W. W. Rolles

We describe and analyze how reinforced random walks can eventually localize, i.e. only visit finitely many sites. After introducing vertex and edge self-interacting walks on a discrete graph in a general setting, and stating the main…

Probability · Mathematics 2011-03-30 Pierre Tarrès

Motivated by the question of optimal functional approximation via compressed sensing, we propose generalizations of the Iterative Hard Thresholding and the Compressive Sampling Matching Pursuit algorithms able to promote sparse in levels…

Information Theory · Computer Science 2021-11-01 Ben Adcock , Simone Brugiapaglia , Matthew King-Roskamp
‹ Prev 1 4 5 6 7 8 10 Next ›