中文
相关论文

相关论文: Zero-Truncated Poisson Regression for Sparse Multi…

200 篇论文

A frequent challenge encountered with compositional ecological data is how to interpret and model data with a high proportion of zeros and $N$'s. Such data frequently occur in ecological applications where counts of species are collected…

统计方法学 · 统计学 2025-08-04 James Sweeney , John Haslett , Dipankar Bandyopadhyay , Michael Fop , Andrew C. Parnell

Modeling data with multivariate count responses is a challenging problem due to the discrete nature of the responses. Existing methods for univariate count responses cannot be easily extended to the multivariate case since the dependency…

统计方法学 · 统计学 2016-08-15 Hao Wu , Xinwei Deng , Naren Ramakrishnan

We study a new class of codes for lossy compression with the squared-error distortion criterion, designed using the statistical framework of high-dimensional linear regression. Codewords are linear combinations of subsets of columns of a…

信息论 · 计算机科学 2015-12-21 Ramji Venkataramanan , Antony Joseph , Sekhar Tatikonda

The zero-truncated Poisson distributions are certain discrete probability distributions whose supports are the set of positive integers, which are also known as the conditional Poisson distributions or the positive Poisson distributions. In…

概率论 · 数学 2023-01-11 Taekyun Kim , Dae san Kim , Si-Hyeon Lee , Seong-Ho Park , Lee-Chae jang

Tensor completion is a fundamental tool for incomplete data analysis, where the goal is to predict missing entries from partial observations. However, existing methods often make the explicit or implicit assumption that the observed entries…

机器学习 · 统计学 2022-03-18 Yuning Qiu , Guoxu Zhou , Qibin Zhao , Shengli Xie

We propose a novel sparse tensor decomposition method, namely Tensor Truncated Power (TTP) method, that incorporates variable selection into the estimation of decomposition components. The sparsity is achieved via an efficient truncation…

机器学习 · 统计学 2016-05-04 Will Wei Sun , Junwei Lu , Han Liu , Guang Cheng

Truncated linear regression is a classical challenge in Statistics, wherein a label, $y = w^T x + \varepsilon$, and its corresponding feature vector, $x \in \mathbb{R}^k$, are only observed if the label falls in some subset $S \subseteq…

统计方法学 · 统计学 2022-08-26 Constantinos Daskalakis , Patroklos Stefanou , Rui Yao , Manolis Zampetakis

Regression for count data is widely performed by models such as Poisson, negative binomial (NB) and zero-inflated regression. A challenge often faced by practitioners is the selection of the right model to take into account dispersion,…

统计方法学 · 统计学 2018-08-02 Hadeel S. Klakattawi , Veronica Vinciotti , Keming Yu

The observations in many applications consist of counts of discrete events, such as photons hitting a detector, which cannot be effectively modeled using an additive bounded or Gaussian noise model, and instead require a Poisson noise…

最优化与控制 · 数学 2011-10-13 Zachary T. Harmany , Roummel F. Marcia , Rebecca M. Willett

Variable selection methods are required in practical statistical modeling, to identify and include only the most relevant predictors, and then improving model interpretability. Such variable selection methods are typically employed in…

This paper studies a tensor-structured linear regression model with a scalar response variable and tensor-structured predictors, such that the regression parameters form a tensor of order $d$ (i.e., a $d$-fold multiway array) in…

机器学习 · 计算机科学 2020-11-26 Talal Ahmed , Haroon Raja , Waheed U. Bajwa

We consider the problem of low-rank decomposition of incomplete multiway tensors. Since many real-world data lie on an intrinsically low dimensional subspace, tensor low-rank decomposition with missing entries has applications in many data…

数值分析 · 计算机科学 2016-08-24 Linxiao Yang , Jun Fang , Hongbin Li , Bing Zeng

In this paper, we investigate right-truncated count data models incorporating cavariates into the parameters. A regression method is proposed to model right-truncated count data exibiting high heterogeneity. The study encompasses the…

统计方法学 · 统计学 2025-03-11 Babagnidé François Koladjo , Ricardo Anderson Donte , Epiphane Sodjinou

We study the problem of low-rank tensor factorization in the presence of missing data. We ask the following question: how many sampled entries do we need, to efficiently and exactly reconstruct a tensor with a low-rank orthogonal…

机器学习 · 统计学 2014-06-12 Prateek Jain , Sewoong Oh

A wide range of problems in computational science and engineering require estimation of sparse eigenvectors for high dimensional systems. Here, we propose two variants of the Truncated Orthogonal Iteration to compute multiple leading…

数值分析 · 数学 2021-03-26 Hexuan Liu , Aleksandr Aravkin

This paper proposes a fast and accurate method for sparse regression in the presence of missing data. The underlying statistical model encapsulates the low-dimensional structure of the incomplete data matrix and the sparsity of the…

机器学习 · 统计学 2015-03-31 Ravi Ganti , Rebecca M. Willett

Systematic under-counting effects are observed in data collected across many disciplines, e.g., epidemiology and ecology. Under-counted tensor completion (UC-TC) is well-motivated for many data analytics tasks, e.g., inferring the case…

机器学习 · 计算机科学 2023-06-07 Shahana Ibrahim , Xiao Fu , Rebecca Hutchinson , Eugene Seo

This paper describes a compound Poisson-based random effects structure for modeling zero-inflated data. Data with large proportion of zeros are found in many fields of applied statistics, for example in ecology when trying to model and…

应用统计 · 统计学 2009-07-29 Marie-Pierre Etienne , Eric Parent , Benoit Hugues , Bernier Jacques

The problem of sparse linear regression is relevant in the context of linear system identification from large datasets. When data are collected from real-world experiments, measurements are always affected by perturbations or low-precision…

最优化与控制 · 数学 2020-04-01 S. M. Fosson , V. Cerone , D. Regruto

Bimodal truncated count distributions are frequently observed in aggregate survey data and in user ratings when respondents are mixed in their opinion. They also arise in censored count data, where the highest category might create an…

统计方法学 · 统计学 2014-01-24 Pragya Sur , Galit Shmueli , Smarajit Bose , Paromita Dubey