English
Related papers

Related papers: Rugged Metropolis Sampling with Simultaneous Updat…

200 papers

We study the relaxation of the Metropolis Monte Carlo algorithm corresponding to a single particle trapped in a one-dimensional confining potential, with even jump distributions that ensure that the dynamics verifies detailed balance.…

Statistical Mechanics · Physics 2025-02-03 A. Patrón , A. D. Chepelianskii , A. Prados , E. Trizac

Randomized iterative methods have gained recent interest in machine learning and signal processing for solving large-scale linear systems. One such example is the randomized Douglas-Rachford (RDR) method, which updates the iterate by…

Numerical Analysis · Mathematics 2025-06-13 Liqi Guo , Ruike Xiang , Deren Han , Jiaxin Xie

This paper proposes a novel robust reinforcement learning framework for discrete-time linear systems with model mismatch that may arise from the sim-to-real gap. A key strategy is to invoke advanced techniques from control theory. Using the…

Systems and Control · Electrical Eng. & Systems 2023-12-07 Leilei Cui , Tamer Başar , Zhong-Ping Jiang

There has been a recent surge of interest in coupling methods for Markov chain Monte Carlo algorithms: they facilitate convergence quantification and unbiased estimation, while exploiting embarrassingly parallel computing capabilities.…

Computation · Statistics 2025-09-03 Tamás P. Papp , Chris Sherlock

In model-based reinforcement learning (MBRL), most algorithms rely on simulating trajectories from one-step dynamics models learned on data. A critical challenge of this approach is the compounding of one-step prediction errors as length of…

Machine Learning · Computer Science 2023-10-12 Abdelhakim Benechehab , Giuseppe Paolo , Albert Thomas , Maurizio Filippone , Balázs Kégl

Stochastic hydrodynamics provides a dynamical framework for the evolution of fluctuations in heavy-ion collisions, but poses significant challenges in numerical simulations. We present an algorithm for the simulation of non-relativistic…

Nuclear Theory · Physics 2026-02-03 Mattis Harhoff , Sören Schlichting , Lorenz von Smekal

Entropy regularized algorithms such as Soft Q-learning and Soft Actor-Critic, recently showed state-of-the-art performance on a number of challenging reinforcement learning (RL) tasks. The regularized formulation modifies the standard RL…

Machine Learning · Statistics 2019-10-15 Elena Smirnova , Elvis Dohmatob

This paper proposes a simulation-based reinforcement learning algorithm for controlling systems with uncertain and varying system parameters. While simulators are useful for safely learning control policies, the reality gap remains a major…

Systems and Control · Electrical Eng. & Systems 2026-05-14 Junya Ikemoto

Bayesian methods are critical for quantifying the behaviors of systems. They capture our uncertainty about a system's behavior using probability distributions and update this understanding as new information becomes available. Probabilistic…

Computation · Statistics 2018-04-25 Thomas A. Catanach , James L. Beck

We study the performance of an automated hybrid Monte Carlo (HMC) approach for conditional simulation of a recently proposed, single-parameter Gibbs Markov random field (Gibbs MRF). The MRF is based on a modified version of the planar…

Computational Physics · Physics 2020-07-08 Milan Žukovič , Dionissios T. Hristopulos

Designing and analyzing model-based RL (MBRL) algorithms with guaranteed monotonic improvement has been challenging, mainly due to the interdependence between policy optimization and model learning. Existing discrepancy bounds generally…

Machine Learning · Computer Science 2023-11-09 Tianying Ji , Yu Luo , Fuchun Sun , Mingxuan Jing , Fengxiang He , Wenbing Huang

Robust reinforcement learning is essential for deploying reinforcement learning algorithms in real-world scenarios where environmental uncertainty predominates. Traditional robust reinforcement learning often depends on rectangularity…

Machine Learning · Computer Science 2024-06-13 Adil Zouitine , David Bertoin , Pierre Clavier , Matthieu Geist , Emmanuel Rachelson

Deploying reinforcement learning (RL) systems requires robustness to uncertainty and model misspecification, yet prior robust RL methods typically only study noise introduced independently across time. However, practical sources of…

Massively parallel GPU simulation environments have accelerated reinforcement learning (RL) research by enabling fast data collection for on-policy RL algorithms like Proximal Policy Optimization (PPO). To maximize throughput, it is common…

Machine Learning · Computer Science 2025-11-27 Sid Bharthulwar , Stone Tao , Hao Su

Reinforcement Learning has drawn huge interest as a tool for solving optimal control problems. Solving a given problem (task or environment) involves converging towards an optimal policy. However, there might exist multiple optimal policies…

Machine Learning · Computer Science 2023-02-16 Simo Alami. C , Fernando Llorente , Rim Kaddah , Luca Martino , Jesse Read

A sensible application of the Hybrid Monte Carlo (HMC) method is often hindered by the presence of large - or even infinite - potential barriers. These potential barriers separate the configuration space into distinct sectors and can lead…

Strongly Correlated Electrons · Physics 2025-08-06 Finn L. Temmen , Evan Berkowitz , Anthony Kennedy , Thomas Luu , Johann Ostmeyer , Xinhao Yu

In MCMC methods, such as the Metropolis-Hastings (MH) algorithm, the Gibbs sampler, or recent adaptive methods, many different strategies can be proposed, often associated in practice to unknown rates of convergence. In this paper we…

Statistics Theory · Mathematics 2007-06-13 Didier Chauveau , Pierre Vandekerkhove

Power systems that need to integrate renewables at a large scale must account for the high levels of uncertainty introduced by these power sources. This can be accomplished with a system of many distributed grid-level storage devices.…

Optimization and Control · Mathematics 2020-02-04 Joseph L. Durante , Juliana Nascimento , Warren B. Powell

Powerful ideas recently appeared in the literature are adjusted and combined to design improved samplers for Bayesian exponential random graph models. Different forms of adaptive Metropolis-Hastings proposals (vertical, horizontal and…

Computation · Statistics 2014-09-18 Alberto Caimo , Antonietta Mira

Momentum methods for convex optimization often rely on precise choices of algorithmic parameters, based on knowledge of problem parameters, in order to achieve fast convergence, as well as to prevent oscillations that could severely…

Systems and Control · Electrical Eng. & Systems 2021-03-24 Justin H. Le , Andrew R. Teel