English
Related papers

Related papers: On Deterministically Approximating Total Variation…

200 papers

The 2-Wasserstein distance (or RMS distance) is a useful measure of similarity between probability distributions that has exciting applications in machine learning. For discrete distributions, the problem of computing this distance can be…

Computational Geometry · Computer Science 2020-07-17 Nathaniel Lahn , Sharath Raghvendra

A striking result of [Acharya et al. 2017] showed that to estimate symmetric properties of discrete distributions, plugging in the distribution that maximizes the likelihood of observed multiset of frequencies, also known as the profile…

Statistics Theory · Mathematics 2020-11-03 Yanjun Han , Kirankumar Shiragur

This paper addresses the task of estimating a covariance matrix under a patternless sparsity assumption. In contrast to existing approaches based on thresholding or shrinkage penalties, we propose a likelihood-based method that regularizes…

Methodology · Statistics 2021-09-13 Jason Xu , Kenneth Lange

Stochastic natural gradient variational inference (NGVI) is a popular and efficient algorithm for Bayesian inference. Despite empirical success, the convergence of this method is still not fully understood. In this work, we define and study…

Methodology · Statistics 2026-04-02 Thomas Guilmeau , Hadrien Hendrikx , Florence Forbes

In this article, we consider products of ergodic Markov chains and discuss their cutoffs in the total variation. Through a new inequality relating the total variation and the Hellinger distance, we may identify the total variation cutoffs…

Probability · Mathematics 2017-02-13 Guan-Yu Chen , Takashi Kumagai

Real-world complex systems often comprise many distinct types of elements as well as many more types of networked interactions between elements. When the relative abundances of types can be measured well, we often observe heavy-tailed…

Physics and Society · Physics 2025-03-17 P. S. Dodds , J. R. Minot , M. V. Arnold , T. Alshaabi , J. L. Adams , A. J. Reagan , C. M. Danforth

In both Tweedie and geometric Tweedie models, the common power parameter $p\notin(0,1)$ works as an automatic distribution selection. It mainly separates two subclasses of semicontinuous ($1<p<2$) and positive continuous ($p\geq 2$)…

Methodology · Statistics 2020-01-30 Rahma Abid , Célestin C. Kokonendji

The maximum product of spacings (MPS) is employed in the estimation of the Generalized Extreme Value Distribution (GEV) and the Generalized Pareto Distribution (GPD). Efficient estimators are obtained by the MPS for all $\gamma$. This…

Statistics Theory · Mathematics 2007-06-13 T. S. T. Wong , W. K. Li

In this paper, we aim at establishing an approximation theory and a learning theory of distribution regression via a fully connected neural network (FNN). In contrast to the classical regression methods, the input variables of distribution…

Machine Learning · Statistics 2023-07-10 Zhongjie Shi , Zhan Yu , Ding-Xuan Zhou

We introduce an iterative optimization scheme for convex objectives consisting of a linear loss and a non-separable penalty, based on the expectation-consistent approximation and the vector approximate message-passing (VAMP) algorithm.…

Machine Learning · Statistics 2018-09-18 Andre Manoel , Florent Krzakala , Gaël Varoquaux , Bertrand Thirion , Lenka Zdeborová

This paper studies the problem of distribution matching (DM), which is a fundamental machine learning problem seeking to robustly align two probability distributions. Our approach is established on a relaxed formulation, called partial…

Machine Learning · Computer Science 2025-05-27 Zi-Ming Wang , Nan Xue , Ling Lei , Rebecka Jörnsten , Gui-Song Xia

We provide a variable metric stochastic approximation theory. In doing so, we provide a convergence theory for a large class of online variable metric methods including the recently introduced online versions of the BFGS algorithm and its…

Data Analysis, Statistics and Probability · Physics 2009-08-26 Peter Sunehag , Jochen Trumpf , S. V. N. Vishwanathan , Nicol Schraudolph

In this paper, we consider a backward problem for a time-space fractional diffusion process. For this problem, we propose to construct the initial data by minimizing data residual error in fourier space domain and variable total variation…

Numerical Analysis · Mathematics 2016-05-24 Junxiong Jia , Jigen Peng , Jinghuai Gao , Yujiao Li

Thompson sampling (TS) is a class of algorithms for sequential decision-making, which requires maintaining a posterior distribution over a model. However, calculating exact posterior distributions is intractable for all but the simplest…

Machine Learning · Statistics 2019-02-21 Ruiyi Zhang , Zheng Wen , Changyou Chen , Lawrence Carin

We study computational and statistical aspects of learning Latent Markov Decision Processes (LMDPs). In this model, the learner interacts with an MDP drawn at the beginning of each epoch from an unknown mixture of MDPs. To sidestep known…

Machine Learning · Computer Science 2024-06-13 Fan Chen , Constantinos Daskalakis , Noah Golowich , Alexander Rakhlin

In reinforcement learning, temporal difference (TD) is the most direct algorithm to learn the value function of a policy. For large or infinite state spaces, exact representations of the value function are usually not available, and it must…

Machine Learning · Computer Science 2018-05-03 Yann Ollivier

We introduce a new decomposition technique for random variables that maps a generic instance of the prophet inequalities problem to a new instance where all but a constant number of variables have a tractable structure that we refer to as…

Data Structures and Algorithms · Computer Science 2020-11-10 Allen Liu , Renato Paes Leme , Martin Pal , Jon Schneider , Balasubramanian Sivan

We propose a new sufficient dimension reduction approach designed deliberately for high-dimensional classification. This novel method is named maximal mean variance (MMV), inspired by the mean variance index first proposed by Cui, Li and…

Methodology · Statistics 2018-12-11 Xin Chen , Jingjing Wu , Zhigang Yao , Jia Zhang

We study the revenue maximization problem with an imprecisely estimated distribution of a single buyer or several independent and identically distributed buyers given that this estimation is not far away from the true distribution. We use…

Computer Science and Game Theory · Computer Science 2019-03-05 Yingkai Li , Pinyan Lu , Haoran Ye

We consider total variation minimization for manifold valued data. We propose a cyclic proximal point algorithm and a parallel proximal point algorithm to minimize TV functionals with $\ell^p$-type data terms in the manifold case. These…

Optimization and Control · Mathematics 2014-12-12 Andreas Weinmann , Laurent Demaret , Martin Storath