中文
相关论文

相关论文: Matrix factorization and prediction for high dimen…

200 篇论文

Recurrent event time data arise in many studies, including biomedicine, public health, marketing, and social media analysis. High-dimensional recurrent event data involving many event types and observations have become prevalent with…

统计方法学 · 统计学 2025-04-02 Fangyi Chen , Yunxiao Chen , Zhiliang Ying , Kangjie Zhou

Recent work on overfitting Bayesian mixtures of distributions offers a powerful framework for clustering multivariate data using a latent Gaussian model which resembles the factor analysis model. The flexibility provided by overfitting…

统计方法学 · 统计学 2019-08-29 Panagiotis Papastamoulis

High dimensionality comparable to sample size is common in many statistical problems. We examine covariance matrix estimation in the asymptotic framework that the dimensionality $p$ tends to $\infty$ as the sample size $n$ increases.…

统计理论 · 数学 2007-06-13 Jianqing Fan , Yingying Fan , Jinchi Lv

Weighted networks encode not only the presence of interactions but also their strength. Existing methods for weighted network community detection often rely on Poisson models, which can be restrictive for overdispersed data and make…

统计方法学 · 统计学 2026-04-28 Fumiya Iwashige

Advancements in data collection techniques and the heterogeneity of data resources can yield high percentages of missing observations on variables, such as block-wise missing data. Under missing-data scenarios, traditional methods such as…

统计方法学 · 统计学 2022-05-17 Wei Lan , Xuerong Chen , Tao Zou , Chih-Ling Tsai

This paper proposes a computationally efficient Bayesian factor model for multiple grouped count data. Adopting the link function approach, the proposed model can capture the association within and between the at-risk probabilities and…

统计方法学 · 统计学 2024-05-13 Genya Kobayashi , Yuta Yamauchi

Inferring causal relationships or related associations from observational data can be invalidated by the existence of hidden confounding. We focus on a high-dimensional linear regression setting, where the measured covariates are affected…

统计方法学 · 统计学 2021-07-22 Zijian Guo , Domagoj Ćevid , Peter Bühlmann

Many popular statistical models, such as factor and random effects models, give arise a certain type of covariance structures that is a summation of low rank and sparse matrices. This paper introduces a penalized approximation framework to…

统计方法学 · 统计学 2015-03-19 Xi Luo

High-dimensional data analysis using traditional models suffers from overparameterization. Two types of techniques are commonly used to reduce the number of parameters - regularization and dimension reduction. In this project, we combine…

统计方法学 · 统计学 2026-03-26 Xialu Liu , Xin Wang

This paper introduces a simple and effective form of data augmentation for recommender systems. A paraphrase similarity model is applied to widely available textual data, such as reviews and product descriptions, yielding new semantic…

计算与语言 · 计算机科学 2021-09-21 Federico López , Martin Scholz , Jessica Yung , Marie Pellat , Michael Strube , Lucas Dixon

We develop a factor analysis for mixed continuous and binary observed variables. To this end, we utilized a recently developed multivariate probability distribution for mixed-type random variables, the Gaussian-Grassmann distribution. In…

统计方法学 · 统计学 2025-12-12 Takashi Arai

We describe and analyze a broad class of mixture models for real-valued multivariate data in which the probability density of observations within each component of the model is represented as an arbitrary combination of basis functions.…

统计方法学 · 统计学 2025-02-28 M. E. J. Newman

Sequencing-based technologies provide an abundance of high-dimensional biological datasets with skewed and zero-inflated measurements. Classification of such data with linear discriminant analysis leads to poor performance due to the…

统计方法学 · 统计学 2022-08-09 Hee Cheol Chung , Yang Ni , Irina Gaynanova

Biological signals of interest in high-dimensional data are often masked by dominant variation shared across conditions. This variation, arising from baseline biological structure or technical effects, can prevent standard dimensionality…

机器学习 · 计算机科学 2026-02-27 Yixuan Li , Archer Y. Yang , Yue Li

To estimate casual treatment effects, we propose a new matching approach based on the reduced covariates obtained from sufficient dimension reduction. Compared to the original covariates and the propensity score, which are commonly used for…

统计方法学 · 统计学 2017-02-03 Wei Luo , Yeying Zhu

In this paper, we estimate the high dimensional precision matrix under the weak sparsity condition where many entries are nearly zero. We revisit the sparse column-wise inverse operator (SCIO) estimator \cite{liu2015fast} and derive its…

统计理论 · 数学 2022-10-21 Zeyu Wu , Cheng Wang , Weidong Liu

The mixture of factor analyzers (MFA) model provides a powerful tool for analyzing high-dimensional data as it can reduce the number of free parameters through its factor-analytic representation of the component covariance matrices. This…

统计方法学 · 统计学 2013-07-09 Tsung-I Lin , Geoffrey J. McLachlan , Sharon X. Lee

Linear mixed-effects models are widely used in analyzing clustered or repeated measures data. We propose a quasi-likelihood approach for estimation and inference of the unknown parameters in linear mixed-effects models with high-dimensional…

统计方法学 · 统计学 2021-03-10 Sai Li , Tony T. Cai , Hongzhe Li

While graphical models for continuous data (Gaussian graphical models) and discrete data (Ising models) have been extensively studied, there is little work on graphical models linking both continuous and discrete variables (mixed data),…

机器学习 · 统计学 2016-08-22 Jie Cheng , Tianxi Li , Elizaveta Levina , Ji Zhu

Recent literature provides many computational and modeling approaches for covariance matrices estimation in a penalized Gaussian graphical models but relatively little study has been carried out on the choice of the tuning parameter. This…

统计方法学 · 统计学 2009-09-08 Heng Lian
‹ 上一页 1 8 9 10 下一页 ›