English

Analysis of Primal-Dual Langevin Algorithms

Optimization and Control 2024-11-06 v3

Abstract

We analyze a recently proposed class of algorithms for the problem of sampling from probability distributions μ\mu^\ast in Rd\mathbb{R}^d with a Lebesgue density of the form μ(x)exp(f(Kx)g(x))\mu^\ast(x) \propto \exp(-f(Kx)-g(x)), where KK is a linear operator and f,gf,g convex and non-smooth. The method is a generalization of the primal-dual hybrid gradient optimization algorithm to a sampling scheme. We give the iteration's continuous time limit, a stochastic differential equation in the joint primal-dual variable, and its mean field limit Fokker-Planck equation. Under mild conditions, the scheme converges to a unique stationary state in continuous and discrete time. Contrary to purely primal overdamped Langevin diffusion, the stationary state in continuous time does not have μ\mu^\ast as its primal marginal. Thus, further analysis is carried out to bound the bias induced by the partial dualization, and potentially correct for it in the diffusion. Time discretizations of the diffusion lead to implementable algorithms, but, as is typical in Langevin Monte Carlo methods, introduce further bias. We prove bounds for these discretization errors, which allow to give convergence results relating the produced samples to the target. We demonstrate our findings numerically first on small-scale examples in which we can exactly verify the theoretical results, and subsequently on typical examples of larger scale from Bayesian imaging inverse problems.

Keywords

Cite

@article{arxiv.2405.18098,
  title  = {Analysis of Primal-Dual Langevin Algorithms},
  author = {Martin Burger and Matthias J. Ehrhardt and Lorenz Kuger and Lukas Weigand},
  journal= {arXiv preprint arXiv:2405.18098},
  year   = {2024}
}
R2 v1 2026-06-28T16:43:43.492Z