English
Related papers

Related papers: How good is Good-Turing for Markov samples?

200 papers

Learning a Gaussian mixture model (GMM) is a fundamental problem in machine learning, learning theory, and statistics. One notion of learning a GMM is proper learning: here, the goal is to find a mixture of $k$ Gaussians $\mathcal{M}$ that…

Data Structures and Algorithms · Computer Science 2015-06-04 Jerry Li , Ludwig Schmidt

Positive continuous outcomes with a point mass at zero are prevalent in biomedical research. To model the point mass at zero and to provide marginalized covariate effect estimates, marginalized two part models (MTP) have been developed for…

This paper considers maximum likelihood (ML) estimation in a large class of models with hidden Markov regimes. We investigate consistency of the ML estimator and local asymptotic normality for the models under general conditions which allow…

Statistics Theory · Mathematics 2021-12-07 Demian Pouzo , Zacharias Psaradakis , Martin Sola

Transit timing variations (TTVs) are a valuable tool to determine the masses and orbits of transiting planets in multi-planet systems. TTVs can be readily modeled given knowledge of the interacting planets' orbital configurations and…

Instrumentation and Methods for Astrophysics · Physics 2019-01-30 Noah W. Tuchow , Eric B. Ford , Theodore Papamarkou , Alexey Lindo

We consider $n\times n$ random matrices $M_{n}=\sum_{\alpha =1}^{m}{\tau _{\alpha }}\mathbf{y}_{\alpha }\otimes \mathbf{y}_{\alpha }$, where $\tau _{\alpha }\in \mathbb{R}$, $\{\mathbf{y}_{\alpha }\}_{\alpha =1}^{m}$ are i.i.d. isotropic…

Probability · Mathematics 2013-12-02 O. Guédon , A. Lytova , A. Pajor , L. Pastur

The telegraph process $X(t)$, $t>0$, (Goldstein, 1951) and the geometric telegraph process $S(t) = s_0 \exp\{(\mu -\frac12\sigma^2)t + \sigma X(t)\}$ with $\mu$ a known constant and $\sigma>0$ a parameter are supposed to be observed at…

Statistics Theory · Mathematics 2007-06-13 Alessandro De Gregorio , Stefano M. Iacus

The classical asymptotic theory for parametric $M$-estimators guarantees that, in the limit of infinite sample size, the excess risk has a chi-square type distribution, even in the misspecified case. We demonstrate how self-concordance of…

Statistics Theory · Mathematics 2020-12-01 Dmitrii Ostrovskii , Francis Bach

Given a Gaussian Markov random field, we consider the problem of selecting a subset of variables to observe which minimizes the total expected squared prediction error of the unobserved variables. We first show that finding an exact…

Machine Learning · Computer Science 2012-09-27 Satyaki Mahalanabis , Daniel Stefankovic

In this paper we consider the problem of sampling from the low-temperature exponential random graph model (ERGM). The usual approach is via Markov chain Monte Carlo, but Bhamidi et al. showed that any local Markov chain suffers from an…

Probability · Mathematics 2022-10-05 Guy Bresler , Dheeraj Nagaraj , Eshaan Nichani

The problem of estimating discovery probabilities originated in the context of statistical ecology, and in recent years it has become popular due to its frequent appearance in challenging applications arising in genetics, bioinformatics,…

Methodology · Statistics 2015-06-17 Stefano Favaro , Bernardo Nipoti , Yee Whye Teh

In this paper, we analyze the convergence rate of a collapsed Gibbs sampler for crossed random effects models. Our results apply to a substantially larger range of models than previous works, including models that incorporate missingness…

Computation · Statistics 2021-10-22 Swarnadip Ghosh , Chenyang Zhong

Gibbs samplers are preeminent Markov chain Monte Carlo algorithms used in computational physics and statistical computing. Yet, their most fundamental properties, such as relations between convergence characteristics of their various…

Computation · Statistics 2024-07-11 Iwona Chlebicka , Krzysztof Łatuszyński , Błażej Miasojedow

This article studies the expected occupancy probabilities on an alphabet. Unlike the standard situation, where observations are assumed to be independent and identically distributed (iid), we assume that they follow a regime switching…

Probability · Mathematics 2020-05-19 Michael Grabchak , Mark Kelbert , Quentin Paris

Nested sampling is a simulation method for approximating marginal likelihoods proposed by Skilling (2006). We establish that nested sampling has an approximation error that vanishes at the standard Monte Carlo rate and that this error is…

Computation · Statistics 2010-10-11 Nicolas Chopin , Christian Robert

Iteration of randomly chosen quadratic maps defines a Markov process: X_{n+1}=\epsilon_{n+1}X_n(1-X_n), where \epsilon_n are i.i.d. with values in the parameter space [0,4] of quadratic maps F_{\theta}(x)=\theta x(1-x). Its study is of…

Probability · Mathematics 2007-05-23 Rabi Bhattacharya , Mukul Majumdar

Consider a sequence of continuous-time irreducible reversible Markov chains and a sequence of initial distributions, $\mu_n$. The sequence is said to exhibit $\mu_n$-cutoff if the convergence to stationarity in total variation distance is…

Probability · Mathematics 2018-02-27 Jonathan Hermon

Motivated by broad applications in reinforcement learning and machine learning, this paper considers the popular stochastic gradient descent (SGD) when the gradients of the underlying objective function are sampled from Markov processes.…

Optimization and Control · Mathematics 2020-04-02 Thinh T. Doan , Lam M. Nguyen , Nhan H. Pham , Justin Romberg

In this work, we investigate Gaussian Mixture Models ({\it abbrv} GMM) and the related problem of non parametric maximum likelihood estimation ({\it abbrv} NPMLE) from the perspective of statistical mechanics. In particular, we establish…

Statistics Theory · Mathematics 2026-03-25 Subhroshekhar Ghosh , Adityanand Guntuboyina , Satyaki Mukherjee , Hoang-Son Tran

We consider the problem of predicting the next observation given a sequence of past observations, and consider the extent to which accurate prediction requires complex algorithms that explicitly leverage long-range dependencies. Perhaps…

Machine Learning · Computer Science 2018-06-29 Vatsal Sharan , Sham Kakade , Percy Liang , Gregory Valiant

We find the exact typical error exponent of constant composition generalized random Gilbert-Varshamov (RGV) codes over DMCs channels with generalized likelihood decoding. We show that the typical error exponent of the RGV ensemble is equal…

Information Theory · Computer Science 2022-11-23 Lan V. Truong , Albert Guillén i Fàbregas