English
Related papers

Related papers: Accelerating Nonconvex Learning via Replica Exchan…

200 papers

The exact estimation of latent variable models with big data is known to be challenging. The latents have to be integrated out numerically, and the dimension of the latent variables increases with the sample size. This paper develops a…

Econometrics · Economics 2023-06-27 Ruben Loaiza-Maya , Didier Nibbering , Dan Zhu

The InfoNCE loss in contrastive learning depends critically on a temperature parameter, yet its dynamics under fixed versus annealed schedules remain poorly understood. We provide a theoretical analysis by modeling embedding evolution under…

Machine Learning · Computer Science 2026-03-16 Faris Chaudhry

Discontinuous transitions into absorbing states require an effective mechanism that prevents the stabilization of low density states. They can be found in different systems, such as lattice models or stochastic differential equations (e.g.…

Statistical Mechanics · Physics 2015-08-12 Salete Pianegonda , Carlos E. Fiore

Machine unlearning has raised significant interest with the adoption of laws ensuring the ``right to be forgotten''. Researchers have provided a probabilistic notion of approximate unlearning under a similar definition of Differential…

Machine Learning · Computer Science 2025-09-25 Eli Chien , Haoyu Wang , Ziang Chen , Pan Li

This paper introduces and analyses interacting underdamped Langevin algorithms, termed Kinetic Interacting Particle Langevin Monte Carlo (KIPLMC) methods, for statistical inference in latent variable models. We propose a diffusion process…

Computation · Statistics 2026-04-17 Paul Felix Valsecchi Oliva , O. Deniz Akyildiz

We present an improved analysis of the Euler-Maruyama discretization of the Langevin diffusion. Our analysis does not require global contractivity, and yields polynomial dependence on the time horizon. Compared to existing approaches, we…

Probability · Mathematics 2019-11-05 Wenlong Mou , Nicolas Flammarion , Martin J. Wainwright , Peter L. Bartlett

This paper addresses a distributed nonconvex optimization problem over multi-agent networks, where each agent exchanges its local information solely with its neighbors. Given that most existing distributed nonconvex optimization algorithms…

Optimization and Control · Mathematics 2026-02-27 Zichong Ou , Jie Lu

We study the simulated annealing algorithm based on the kinetic Langevin dynamics, in order to find the global minimum of a non-convex potential function. For both the continuous time formulation and a discrete time analogue, we obtain the…

Probability · Mathematics 2022-06-14 Xuedong He , Xiaolu Tan , Ruocheng Wu

This paper presents a diffusion process with a novel resetting mechanism in which the amplitude of the process is instantaneously converted to a proportion of its value at random times. This model is described by a Langevin equation with…

Statistical Mechanics · Physics 2022-04-18 J. Kevin Pierce

We propose a new discretization of the mirror-Langevin diffusion and give a crisp proof of its convergence. Our analysis uses relative convexity/smoothness and self-concordance, ideas which originated in convex optimization, together with a…

Statistics Theory · Mathematics 2021-10-26 Kwangjun Ahn , Sinho Chewi

We describe a stochastic, dynamical system capable of inference and learning in a probabilistic latent variable model. The most challenging problem in such models - sampling the posterior distribution over latent variables - is proposed to…

Machine Learning · Statistics 2022-07-26 Michael Y. -S. Fang , Mayur Mudigonda , Ryan Zarcone , Amir Khosrowshahi , Bruno A. Olshausen

We propose {\it HumanDiffusion,} a diffusion model trained from humans' perceptual gradients to learn an acceptable range of data for humans (i.e., human-acceptable distribution). Conventional HumanGAN aims to model the human-acceptable…

Human-Computer Interaction · Computer Science 2023-06-22 Yota Ueda , Shinnosuke Takamichi , Yuki Saito , Norihiro Takamune , Hiroshi Saruwatari

In order to solve tasks like uncertainty quantification or hypothesis tests in Bayesian imaging inverse problems, we often have to draw samples from the arising posterior distribution. For the usually log-concave but high-dimensional…

Computation · Statistics 2025-01-23 Matthias J. Ehrhardt , Lorenz Kuger , Carola-Bibiane Schönlieb

It is well known in many settings that reversible Langevin diffusions in confining potentials converge to equilibrium exponentially fast. Adding irreversible perturbations to the drift of a Langevin diffusion that maintain the same…

Methodology · Statistics 2019-07-02 Michela Ottobre , Natesh S. Pillai , Konstantinos Spiliopoulos

This paper proposes and analyzes a communication-efficient distributed optimization framework for general nonconvex nonsmooth signal processing and machine learning problems under an asynchronous protocol. At each iteration, worker machines…

Optimization and Control · Mathematics 2020-07-15 Jineng Ren , Jarvis Haupt

The probability distribution effectively sampled by a complex Langevin process for theories with a sign problem is not known a priori and notoriously hard to understand. Diffusion models, a class of generative AI, can learn distributions…

High Energy Physics - Lattice · Physics 2024-12-05 Diaa E. Habibi , Gert Aarts , Lingxiao Wang , Kai Zhou

We study the Langevin dynamics of diffusive particles with regular pairwise interactions under mean-field scaling. By approximating empirical distributions with conditional distributions, we establish coercive and contractive properties for…

Probability · Mathematics 2026-05-28 Songbo Wang

Nesterov's Accelerated Gradient (NAG) for optimization has better performance than its continuous time limit (noiseless kinetic Langevin) when a finite step-size is employed \citep{shi2021understanding}. This work explores the sampling…

Machine Learning · Computer Science 2022-06-22 Ruilin Li , Hongyuan Zha , Molei Tao

This paper presents a new Metropolis-adjusted Langevin algorithm (MALA) that uses convex analysis to simulate efficiently from high-dimensional densities that are log-concave, a class of probability distributions that is widely used in…

Methodology · Statistics 2015-04-06 Marcelo Pereyra

Training a very deep neural network is a challenging task, as the deeper a neural network is, the more non-linear it is. We compare the performances of various preconditioned Langevin algorithms with their non-Langevin counterparts for the…

Machine Learning · Computer Science 2023-01-02 Pierre Bras
‹ Prev 1 4 5 6 7 8 10 Next ›