English
Related papers

Related papers: Reinforced Loop Soup via Wilson's Algorithm

200 papers

To overcome the curses of dimensionality and modeling of Dynamic Programming (DP) methods to solve Markov Decision Process (MDP) problems, Reinforcement Learning (RL) methods are adopted in practice. Contrary to traditional RL algorithms…

Machine Learning · Computer Science 2021-08-24 Arghyadip Roy , Vivek Borkar , Abhay Karandikar , Prasanna Chaporkar

Neural networks have shown state-of-the-art performances in various classification and regression tasks. Rectified linear units (ReLU) are often used as activation functions for the hidden layers in a neural network model. In this article,…

Machine Learning · Computer Science 2026-01-12 Shufei Ge , Shijia Wang , Lloyd Elliott

We study the mixing time of a non-Markovian process, the step-reinforced random walk (SRRW) on a finite group. This process differs from a classical random walk in that at each integer time, with probability $\alpha$ the next step is chosen…

Probability · Mathematics 2026-04-29 Yuval Peres , Shuo Qin

The vertex-reinforced jump process (VRJP) is a form of self-interacting random walk in which the walker is biased towards returning to previously visited vertices with the bias depending linearly on the local time at these vertices. We…

Probability · Mathematics 2021-05-17 Gady Kozma , Ron Peled

For a linear equality constrained convex optimization problem involving two objective functions with a ``nonsmooth" + ``nonsmooth" composite structure, we study two algorithms derived from a mixed-order dynamical system which incorporates…

Optimization and Control · Mathematics 2026-03-25 Geng-Hua Li , Hai-Yi Zhao , Xiangkai Sun

We consider a multi-particle generalization of linear edge-reinforced random walk (ERRW). We observe that in absence of exchangeability, new techniques are needed in order to study the multi-particle model. We describe an unusual coupling…

Probability · Mathematics 2007-05-23 Yevgeniy Kovchegov

The Bellman equation and its continuous form, the Hamilton-Jacobi-Bellman equation, are ubiquitous in reinforcement learning and control theory. However, these equations become intractable for high-dimensional or nonlinear systems. This…

Artificial Intelligence · Computer Science 2026-05-04 Preston Rozwood , Edward Mehrez , Ludger Paehler , Wen Sun , Steven L. Brunton

This paper concerns the Vertex Reinforced Jump Process (VRJP) and its representations as a Markov process in random environment. We show that all possible representations of the VRJP as a mixture of Markov processes can be expressed in a…

Probability · Mathematics 2019-03-26 Thomas Gerard

We propose a reinforcement-learning algorithm to tackle the challenge of reconstructing phylogenetic trees. The search for the tree that best describes the data is algorithmically challenging, thus all current algorithms for phylogeny…

Populations and Evolution · Quantitative Biology 2023-03-14 Dana Azouri , Oz Granit , Michael Alburquerque , Yishay Mansour , Tal Pupko , Itay Mayrose

Given a random walk $(S_n)$ with typical step distributed according to some fixed law and a fixed parameter $p \in (0,1)$, the associated positively step-reinforced random walk is a discrete-time process which performs at each step, with…

Probability · Mathematics 2022-10-19 Marco Bertenghi , Alejandro Rosales-Ortiz

In this work, we investigate a novel setting of Markovian loop measures and introduce a new class of loop measures called Bosonic loop measures. Namely, we consider loop soups with varying intensity $ \mu\le 0 $ (chemical potential in…

Probability · Mathematics 2019-06-21 Stefan Adams , Quirin Vogel

We present tournament results and several powerful strategies for the Iterated Prisoner's Dilemma created using reinforcement learning techniques (evolutionary and particle swarm algorithms). These strategies are trained to perform well…

Computer Science and Game Theory · Computer Science 2018-02-07 Marc Harper , Vincent Knight , Martin Jones , Georgios Koutsovoulos , Nikoleta E. Glynatsi , Owen Campbell

Building on previous work using reinforcement learning (RL) focused on identification of exfiltration paths, this work expands the methodology to include protocol and payload considerations. The former approach to exfiltration path…

Cryptography and Security · Computer Science 2023-10-06 Riddam Rishu , Akshay Kakkar , Cheng Wang , Abdul Rahman , Christopher Redino , Dhruv Nandakumar , Tyler Cody , Ryan Clark , Daniel Radke , Edward Bowen

Deep reinforcement learning is an increasingly popular technique for synthesising policies to control an agent's interaction with its environment. There is also growing interest in formally verifying that such policies are correct and…

Artificial Intelligence · Computer Science 2022-06-02 Edoardo Bacci , David Parker

Lifted samplers form a class of Markov chain Monte Carlo methods which has drawn a lot attention in recent years due to superior performance in challenging Bayesian applications. A canonical example of lifted samplers is the one that is…

Computation · Statistics 2026-05-01 Philippe Gagnon , Florian Maire

Consider a sequence of Poisson point processes of non-trivial loops with certain intensity measures $(\mu^{(n)})_n$, where each $\mu^{(n)}$ is explicitly determined by transition probabilities $p^{(n)}$ of a random walk on a finite state…

Probability · Mathematics 2025-06-23 Yinshan Chang

A weighted regression procedure is proposed for regression type problems where the innovations are heavy-tailed. This method approximates the least absolute regression method in large samples, and the main advantage will be if the sample is…

Computation · Statistics 2018-11-06 J. Martin van Zyl

Markov jump processes are continuous-time stochastic processes with a wide range of applications in both natural and social sciences. Despite their widespread use, inference in these models is highly non-trivial and typically proceeds via…

Machine Learning · Computer Science 2023-06-01 Patrick Seifner , Ramses J. Sanchez

We prove a combinatorial lemma about the distribution of directed currents in a complex "loop soup" and use it to give a new proof of the isomorphism relating loop measures and complex Gaussian fields.

Probability · Mathematics 2018-04-04 Gregory F. Lawler , Petr Panov

This thesis examines edge-reinforced random walks with some modifications to the standard definition. An overview of known results relating to the standard model is given and the proof of recurrence for the standard linearly edge-reinforced…

Probability · Mathematics 2023-09-07 Fabian Michel