中文
相关论文

相关论文: High-dimensional unsupervised classification via p…

200 篇论文

It has become increasingly common to collect high-dimensional binary response data; for example, with the emergence of new sampling techniques in ecology. In smaller dimensions, multivariate probit (MVP) models are routinely used for…

统计方法学 · 统计学 2022-10-26 Antik Chakraborty , Rihui Ou , David B. Dunson

This paper studies the distribution estimation of contaminated data by the MoM-GAN method, which combines generative adversarial net (GAN) and median-of-mean (MoM) estimation. We use a deep neural network (DNN) with a ReLU activation…

机器学习 · 统计学 2022-12-29 Fang Xie , Lihu Xu , Qiuran Yao , Huiming Zhang

The dependency structure of multivariate data can be analyzed using the covariance matrix $\Sigma$. In many fields the precision matrix $\Sigma^{-1}$ is even more informative. As the sample covariance estimator is singular in…

统计方法学 · 统计学 2015-06-04 Viktoria Öllerer , Christophe Croux

Inferring causal relationships or related associations from observational data can be invalidated by the existence of hidden confounding. We focus on a high-dimensional linear regression setting, where the measured covariates are affected…

统计方法学 · 统计学 2021-07-22 Zijian Guo , Domagoj Ćevid , Peter Bühlmann

Massive data analysis calls for distributed algorithms and theories. We design a multi-round distributed algorithm for canonical correlation analysis. We construct principal directions through the convex formulation of canonical correlation…

统计计算 · 统计学 2024-12-24 Canyi Chen , Liping Zhu

The factor analysis model is a statistical model where a certain number of hidden random variables, called factors, affect linearly the behaviour of another set of observed random variables, with additional random noise. The main assumption…

统计理论 · 数学 2023-12-06 Muhammad Ardiyansyah , Luca Sodomaco

Count data with complex features arise in many disciplines, including ecology, agriculture, criminology, medicine, and public health. Zero inflation, spatial dependence, and non-equidispersion are common features in count data. There are…

统计方法学 · 统计学 2024-05-14 Bokgyeong Kang , John Hughes , Murali Haran

We consider the problem of estimating the conditional probability distribution of missing values given the observed ones. We propose an approach, which combines the flexibility of deep neural networks with the simplicity of Gaussian mixture…

机器学习 · 计算机科学 2020-11-20 Marcin Przewięźlikowski , Marek Śmieja , Łukasz Struski

We introduce efficient MCMC algorithms for Bayesian inference for single-factor models with correlated residuals where the residuals' distribution is a Gaussian graphical model. We call this family of models single-factor graphical models.…

统计方法学 · 统计学 2025-10-17 David Marcano , Adrian Dobra

We propose a method for inference on moderately high-dimensional, nonlinear, non-Gaussian, partially observed Markov process models for which the transition density is not analytically tractable. Markov processes with intractable transition…

统计方法学 · 统计学 2020-04-02 Joonha Park , Edward L. Ionides

We present a novel approach for the analysis of multivariate case-control georeferenced data using Bayesian inference in the context of disease mapping, where the spatial distribution of different types of cancers is analyzed. Extending…

The mixture of Gaussian distributions, a soft version of k-means , is considered a state-of-the-art clustering algorithm. It is widely used in computer vision for selecting classes, e.g., color, texture, and shapes. In this algorithm, each…

机器学习 · 统计学 2016-12-30 Mahajabin Rahman , Davi Geiger

Generalised linear models for multi-class classification problems are one of the fundamental building blocks of modern machine learning tasks. In this manuscript, we characterise the learning of a mixture of $K$ Gaussians with generic means…

Bayesian hierarchical models have been demonstrated to provide efficient algorithms for finding sparse solutions to ill-posed inverse problems. The models comprise typically a conditionally Gaussian prior model for the unknown, augmented by…

数值分析 · 数学 2023-03-31 Daniela Calvetti , Erkki Somersalo

We describe and analyze a broad class of mixture models for real-valued multivariate data in which the probability density of observations within each component of the model is represented as an arbitrary combination of basis functions.…

统计方法学 · 统计学 2025-02-28 M. E. J. Newman

We study the task of high-dimensional entangled mean estimation in the subset-of-signals model. Specifically, given $N$ independent random points $x_1,\ldots,x_N$ in $\mathbb{R}^D$ and a parameter $\alpha \in (0, 1)$ such that each $x_i$ is…

数据结构与算法 · 计算机科学 2025-01-10 Ilias Diakonikolas , Daniel M. Kane , Sihan Liu , Thanasis Pittas

Unmeasured or latent variables are often the cause of correlations between multivariate measurements, which are studied in a variety of fields such as psychology, ecology, and medicine. For Gaussian measurements, there are classical tools…

机器学习 · 计算机科学 2022-01-28 Łukasz Kidziński , Francis K. C. Hui , David I. Warton , Trevor Hastie

Anomaly detection is a field of intense research. Identifying low probability events in data/images is a challenging problem given the high-dimensionality of the data, especially when no (or little) information about the anomaly is…

机器学习 · 计算机科学 2022-04-13 José A. Padrón-Hidalgo , Valero Laparra , Gustau Camps-Valls

The radiological characterization of contaminated elements (walls, grounds, objects) from nuclear facilities often suffers from a too small number of measurements. In order to determine risk prediction bounds on the level of contamination,…

应用统计 · 统计学 2017-05-30 Géraud Blatman , Thibault Delage , Bertrand Iooss , Nadia Pérot

Distributed Gaussian process (DGP) is a popular approach to scale GP to big data which divides the training data into some subsets, performs local inference for each partition, and aggregates the results to acquire global prediction. To…

机器学习 · 计算机科学 2022-02-08 Hamed Jalali , Gjergji Kasneci
‹ 上一页 1 8 9 10 下一页 ›