中文
相关论文

相关论文: Shared Subspace Models for Multi-Group Covariance …

200 篇论文

This paper investigates the asymptotics of eigenstructure of sample covariance matrix under the spiked covariance matrix model in ultra-high-dimensional settings, where the dimensionality can grow much faster than the sample size with $ p…

统计理论 · 数学 2026-04-30 Wonjun Seo

A problem that is frequently encountered in a variety of mathematical contexts, is to find the common invariant subspaces of a single, or set of matrices. A new method is proposed that gives a definitive answer to this problem. The key idea…

综合数学 · 数学 2024-08-29 Ahmad Y. Al-Dweik , Ryad Ghanam , Gerard Thompson , Hassan Azad

Covariance estimation becomes challenging in the regime where the number p of variables outstrips the number n of samples available to construct the estimate. One way to circumvent this problem is to assume that the covariance matrix is…

概率论 · 数学 2012-06-14 Richard Y. Chen , Alex Gittens , Joel A. Tropp

Until recently obtaining data on populations of networks was typically rare. However, with the advancement of automatic monitoring devices and the growing social and scientific interest in networks, such data has become more widely…

统计方法学 · 统计学 2020-01-22 Mirko Signorelli , Ernst Wit

In this paper, I outline several conceptual and methodological issues related to modeling individual and group processes embedded in clustered/hierarchical data structures. We position multilevel modeling techniques within a broader set of…

统计方法学 · 统计学 2022-12-29 Amira Ibrahim El-Desokey

Finite Gaussian mixture models are widely used for model-based clustering of continuous data. Nevertheless, since the number of model parameters scales quadratically with the number of variables, these models can be easily…

统计方法学 · 统计学 2018-09-25 Michael Fop , Thomas Brendan Murphy , Luca Scrucca

High-dimensional variable selection, with many more covariates than observations, is widely documented in standard regression models, but there are still few tools to address it in non-linear mixed-effects models where data are collected…

Factor models are widely used for dimension reduction in the analysis of multivariate data. This is achieved through decomposition of a p x p covariance matrix into the sum of two components. Through a latent factor representation, they can…

统计方法学 · 统计学 2024-07-01 Sarah Elizabeth Heaps , Ian Hyla Jermyn

Covariance matrix estimation is an important problem in multivariate data analysis, both from theoretical as well as applied points of view. Many simple and popular covariance matrix estimators are known to be severely affected by model…

统计方法学 · 统计学 2025-11-21 Soumya Chakraborty , Ayanendranath Basu , Abhik Ghosh

High dimensional data often contain multiple facets, and several clustering patterns can co-exist under different variable subspaces, also known as the views. While multi-view clustering algorithms were proposed, the uncertainty…

机器学习 · 统计学 2019-10-09 Leo L Duan

Divergence is not only an important mathematical concept in information theory, but also applied to machine learning problems such as low-dimensional embedding, manifold learning, clustering, classification, and anomaly detection. We…

统计计算 · 统计学 2016-11-22 Kun Yang , Hao Su , Wing Hung Wong

Multidimensional network data can have different levels of complexity, as nodes may be characterized by heterogeneous individual-specific features, which may vary across the networks. This paper introduces a class of models for…

统计方法学 · 统计学 2021-04-01 Silvia D'Angelo , Marco Alfò , Thomas Brendan Murphy

The multivariate hypergeometric distribution describes sampling without replacement from a discrete population of elements divided into multiple categories. Addressing a gap in the literature, we tackle the challenge of estimating discrete…

机器学习 · 计算机科学 2024-06-11 Liam Hodgson , Danilo Bzdok

Computer Vision practitioners must thoroughly understand their model's performance, but conditional evaluation is complex and error-prone. In biometric verification, model performance over continuous covariates---real-number attributes of…

机器学习 · 计算机科学 2020-09-22 Mel McCurrie , Hamish Nicholson , Walter J. Scheirer , Samuel Anthony

The major sources of abundant data are constantly expanding with the available data collection methodologies in various applications - medical, insurance, scientific, bio-informatics and business. These data sets may be distributed…

分布式、并行与集群计算 · 计算机科学 2016-06-24 Aruna Govada , Sanjay K. Sahay

We develop an empirical Bayes (EB) G-modeling framework for short-panel linear models with nonparametric prior for the random intercepts, slopes, dynamics, and non-spherical error variances. We establish identification and consistency of…

计量经济学 · 经济学 2026-02-13 Myunghyun Song , Sokbae Lee , Serena Ng

A common task in high-throughput biology is to screen for associations across thousands of units of interest, e.g., genes or proteins. Often, the data for each unit are modeled as Gaussian measurements with unknown mean and variance and are…

统计理论 · 数学 2024-10-01 Nikolaos Ignatiadis , Bodhisattva Sen

How to estimate heterogeneity, e.g. the effect of some variable differing across observations, is a key question in political science. Methods for doing so make simplifying assumptions about the underlying nature of the heterogeneity to…

统计方法学 · 统计学 2021-03-31 Max Goplerud

Bayesian model comparison (BMC) offers a principled probabilistic approach to study and rank competing models. In standard BMC, we construct a discrete probability distribution over the set of possible models, conditional on the observed…

机器学习 · 统计学 2023-02-22 Marvin Schmitt , Stefan T. Radev , Paul-Christian Bürkner

The posterior probability distribution for a set of model parameters encodes all that the data have to tell us in the context of a given model; it is the fundamental quantity for Bayesian parameter estimation. In order to infer the…

天体物理仪器与方法 · 物理学 2015-06-16 Rupert Allison , Joanna Dunkley