中文
相关论文

相关论文: High-dimensional unsupervised classification via p…

200 篇论文

A common approach to analyze a covariate-sample count matrix, an element of which represents how many times a covariate appears in a sample, is to factorize it under the Poisson likelihood. We show its limitation in capturing the tendency…

统计方法学 · 统计学 2017-10-06 Mingyuan Zhou

We propose a semi-supervised generative model, SeGMA, which learns a joint probability distribution of data and their classes and which is implemented in a typical Wasserstein auto-encoder framework. We choose a mixture of Gaussians as a…

机器学习 · 计算机科学 2020-08-28 Marek Śmieja , Maciej Wołczyk , Jacek Tabor , Bernhard C. Geiger

Posterior computation for high-dimensional data with many parameters can be challenging. This article focuses on a new method for approximating posterior distributions of a low- to moderate-dimensional parameter in the presence of a…

统计计算 · 统计学 2022-04-08 Willem van den Boom , Galen Reeves , David B. Dunson

Graphical model has been widely used to investigate the complex dependence structure of high-dimensional data, and it is common to assume that observed data follow a homogeneous graphical model. However, observations usually come from…

统计方法学 · 统计学 2016-01-01 Kevin Lee , Lingzhou Xue

High-dimensional changepoint analysis is a growing area of research and has applications in a wide range of fields. The aim is to accurately and efficiently detect changepoints in time series data when both the number of time points and…

统计方法学 · 统计学 2020-04-01 Thomas Grundy , Rebecca Killick , Gueorgui Mihaylov

This article carries out a large dimensional analysis of standard regularized discriminant analysis classifiers designed on the assumption that data arise from a Gaussian mixture model with different means and covariances. The analysis…

Deep Gaussian Processes learn probabilistic data representations for supervised learning by cascading multiple Gaussian Processes. While this model family promises flexible predictive distributions, exact inference is not tractable.…

机器学习 · 统计学 2020-10-23 Jakob Lindinger , David Reeb , Christoph Lippert , Barbara Rakitsch

Most of previous works and applications of Bayesian factor model have assumed the normal likelihood regardless of its validity. We propose a Bayesian factor model for heavy-tailed high-dimensional data based on multivariate Student-$t$…

统计方法学 · 统计学 2020-12-10 Jaejoon Lee , Jaeyong Lee

We present the Gaussian process density sampler (GPDS), an exchangeable generative model for use in nonparametric Bayesian density estimation. Samples drawn from the GPDS are consistent with exact, independent samples from a distribution…

统计计算 · 统计学 2009-12-25 Ryan Prescott Adams , Iain Murray , David J. C. MacKay

Gaussian Graphical Models (GGMs) are widely used in high-dimensional data analysis to synthesize the interaction between variables. In many applications, such as genomics or image analysis, graphical models rely on sparsity and clustering…

机器学习 · 统计学 2026-03-25 Do Edmond Sanou , Christophe Ambroise , Geneviève Robin

High dimensional data analysis is known to be as a challenging problem. In this article, we give a theoretical analysis of high dimensional classification of Gaussian data which relies on a geometrical analysis of the error measure. It…

统计理论 · 数学 2008-07-10 Robin Girard

Modern biomedical datasets are increasingly high dimensional and exhibit complex correlation structures. Generalized Linear Mixed Models (GLMMs) have long been employed to account for such dependencies. However, proper specification of the…

统计方法学 · 统计学 2024-04-18 Hillary M. Heiling , Naim U. Rashid , Quefeng Li , Xianlu L. Peng , Jen Jen Yeh , Joseph G. Ibrahim

We propose a novel approach to estimating the precision matrix of multivariate Gaussian data that relies on decomposing them into a low-rank and a diagonal component. Such decompositions are very popular for modeling large covariance…

统计方法学 · 统计学 2022-08-18 Noirrit Kiran Chandra , Peter Mueller , Abhra Sarkar

In this paper, we study the problem of learning one-dimensional Gaussian mixture models (GMMs) with a specific focus on estimating both the model order and the mixing distribution from independent and identically distributed (i.i.d.)…

机器学习 · 统计学 2026-02-24 Xinyu Liu , Hai Zhang

We consider in this paper a contamined regression model where the distribution of the contaminating component is known when the Eu- clidean parameters of the regression model, the noise distribution, the contamination ratio and the…

统计理论 · 数学 2011-11-10 Pierre Vandekerkhove

Anomaly detection aims to identify observations that deviate from the typical pattern of data. Anomalous observations may correspond to financial fraud, health risks, or incorrectly measured data in practice. We show detecting anomalies in…

机器学习 · 统计学 2020-05-26 Matthew Davidow , David S. Matteson

Deep learning is a hierarchical inference method formed by subsequent multiple layers of learning able to more efficiently describe complex relationships. In this work, Deep Gaussian Mixture Models are introduced and discussed. A Deep…

机器学习 · 统计学 2017-11-21 Cinzia Viroli , Geoffrey J. McLachlan

We consider the problem of clustering data points in high dimensions, i.e. when the number of data points may be much smaller than the number of dimensions. Specifically, we consider a Gaussian mixture model (GMM) with non-spherical…

统计理论 · 数学 2014-06-10 Martin Azizyan , Aarti Singh , Larry Wasserman

We introduce a novel class of Bayesian mixtures for normal linear regression models which incorporates a further Gaussian random component for the distribution of the predictor variables. The proposed cluster-weighted model aims to…

统计方法学 · 统计学 2026-05-26 Panagiotis Papastamoulis , Konstantinos Perrakis

We present a new nonparametric mixture-of-experts model for multivariate regression problems, inspired by the probabilistic k-nearest neighbors algorithm. Using a conditionally specified model, predictions for out-of-sample inputs are based…

机器学习 · 统计学 2022-08-05 Tianfang Zhang , Rasmus Bokrantz , Jimmy Olsson