English
Related papers

Related papers: Reinforced Loop Soup via Wilson's Algorithm

200 papers

We introduce a variation of the step-reinforced random walk with general memory. For the diffusive regime, we establish a functional invariance principle and show that, given suitable conditions on the memory sequence, the arising limiting…

Probability · Mathematics 2024-02-15 Marco Bertenghi , Lucile Laulin

In this paper, we present an extension to the recursive Gaussian Process (RGP) regression that enables the satisfaction of inequality constraints and is well suited for a real-time execution in control applications. The soft inequality…

Systems and Control · Electrical Eng. & Systems 2025-10-30 Ricus Husmann , Sven Weishaupt , Harald Aschemann

Reinforcement learning has traditionally focused on learning state-dependent policies to solve optimal control problems in a closed-loop fashion. In this work, we introduce the paradigm of open-loop reinforcement learning where a fixed…

Machine Learning · Computer Science 2025-04-23 Onno Eberhard , Claire Vernade , Michael Muehlebach

This paper considers the problem of iterative Bayesian smoothing in nonlinear state-space models with additive noise using Gaussian approximations. Iterative methods are known to improve smoothed estimates but are not guaranteed to…

Optimization and Control · Mathematics 2025-02-11 Jakob Lindqvist , Simo Särkkä , Ángel F. García-Fernández , Matti Raitoharju , Lennart Svensson

On-policy reinforcement learning (RL) methods widely used for language model post-training, like Group Relative Policy Optimization (GRPO), often suffer from limited exploration and early saturation due to low sampling diversity. While…

Computation and Language · Computer Science 2026-01-30 Lei Yang , Wei Bi , Chenxi Sun , Renren Jin , Deyi Xiong

Reinforced random walks are random walks on graphs whose transition probabilities along edges from a vertex are proportional to the weights of those edges, but where the weight of an edge evolves in a way that depends on the past traversals…

Information Theory · Computer Science 2026-05-22 Qinghua , Ding , Venkat Anantharam

Reinforcement learning (RL) problems are fundamental in online decision-making and have been instrumental in finding an optimal policy for Markov decision processes (MDPs). Function approximations are usually deployed to handle large or…

Machine Learning · Computer Science 2025-05-20 Jiashuo Jiang , Yiming Zong , Yinyu Ye

This paper develops the first class of algorithms that enable unbiased estimation of steady-state expectations for multidimensional reflected Brownian motion. In order to explain our ideas, we first consider the case of compound Poisson…

Probability · Mathematics 2015-10-27 Jose Blanchet , Xinyun Chen

Various methods for solving the inverse reinforcement learning (IRL) problem have been developed independently in machine learning and economics. In particular, the method of Maximum Causal Entropy IRL is based on the perspective of entropy…

Machine Learning · Computer Science 2021-03-05 Navyata Sanghvi , Shinnosuke Usami , Mohit Sharma , Joachim Groeger , Kris Kitani

The standard quantum annealing algorithm tries to approach the ground state of a classical system by slowly decreasing the hopping rates of a quantum random walk in the configuration space of the problem, where the on-site energies are…

Disordered Systems and Neural Networks · Physics 2018-12-26 A. Ramezanpour

This paper studies the joint support recovery of similar sparse vectors on the basis of a limited number of noisy linear measurements, i.e., in a multiple measurement vector (MMV) model. The additive noise signals on each measurement vector…

Information Theory · Computer Science 2015-06-18 J. F. Determe , J. Louveaux , L. Jacques , F. Horlin

Let (S, BS) be the log-pair associated with a compactification of a given smooth quasi-projective surface V . Under the assumption that the boundary BS is irreducible, we propose an algorithm, in the spirit of the (log) Sarkisov program, to…

Algebraic Geometry · Mathematics 2009-02-11 Adrien Dubouloz , Stéphane Lamy

The fast marching algorithm, and its variants, solves numerically the generalized eikonal equation associated to an underlying riemannian metric. A major challenge for these algorithms is the non-isotropy of the riemannian metric.…

Numerical Analysis · Mathematics 2012-05-25 J. -M. Mirebeau

We present an off-policy actor-critic algorithm for Reinforcement Learning (RL) that combines ideas from gradient-free optimization via stochastic search with learned action-value function. The result is a simple procedure consisting of…

Reinforced Galton--Watson processes describe the dynamics of a population where reproduction events are reinforced, in the sense that offspring numbers of forebears can be repeated randomly by descendants. More specifically, the evolution…

Probability · Mathematics 2025-02-24 Jean Bertoin , Bastien Mallein

Solving linear systems of equations is a fundamental problem with a wide variety of applications across many fields of science, and there is increasing effort to develop quantum linear solver algorithms. [Suba\c{s}i et al., Phys. Rev. Lett.…

Quantum Physics · Physics 2026-01-09 David Jennings , Matteo Lostaglio , Sam Pallister , Andrew T Sornborger , Yiğit Subaşı

Randomized numerical linear algebra is proved to bridge theoretical advancements to offer scalable solutions for approximating tensor decomposition. This paper introduces fast randomized algorithms for solving the fixed Tucker-rank problem…

Numerical Analysis · Mathematics 2025-06-06 Maolin Che , Yimin Wei , Chong Wu , Hong Yan

The four-roll mill, wherein four identical cylinders undergo rotation of identical magnitude but alternate signs, was originally proposed by GI Taylor to create local extensional flows and study their ability to deform small liquid drops.…

Fluid Dynamics · Physics 2021-12-08 Marco Vona , Eric Lauga

We develop sampling methods, which consist of Gaussian invariant versions of random walk Metropolis (RWM), Metropolis adjusted Langevin algorithm (MALA) and second order Hessian or Manifold MALA. Unlike standard RWM and MALA we show that…

Machine Learning · Statistics 2025-06-27 Michalis K. Titsias , Angelos Alexopoulos , Siran Liu , Petros Dellaportas

We consider reinforcement learning in changing Markov Decision Processes where both the state-transition probabilities and the reward functions may vary over time. For this problem setting, we propose an algorithm using a sliding window…

Machine Learning · Computer Science 2018-05-28 Pratik Gajane , Ronald Ortner , Peter Auer