English
Related papers

Related papers: An Improvement on the Hotelling $T^2$ Test Using t…

200 papers

We propose a flexible dual functional factor model for modelling high-dimensional functional time series. In this model, a high-dimensional fully functional factor parametrisation is imposed on the observed functional processes, whereas a…

Econometrics · Economics 2024-01-15 Chenlei Leng , Degui Li , Hanlin Shang , Yingcun Xia

In large-scale few-shot learning for classification problems, often there are a large number of classes and few high-dimensional observations per class. Previous model-based methods, such as Fisher's linear discriminant analysis (LDA),…

Methodology · Statistics 2025-04-16 Andrew Simpson , Semhar Michael

The classic likelihood ratio test for testing the equality of two covariance matrices breakdowns due to the singularity of the sample covariance matrices when the data dimension $p$ is larger than the sample size $n$. In this paper, we…

Methodology · Statistics 2015-11-06 Tung-Lung Wu , Ping Li

The parametric Welch $t$-test and the non-parametric Wilcoxon-Mann-Whitney test are the most commonly used two independent sample means tests. More recent testing approaches include the non-parametric, empirical likelihood and exponential…

Methodology · Statistics 2019-10-08 Michail Tsagris , Abdulaziz Alenazi , Kleio-Maria Verrou , Nikolaos Pandis

The classic censored regression model (tobit model) has been widely used in the economic literature. This model assumes normality for the error distribution and is not recommended for cases where positive skewness is present. Moreover, in…

Methodology · Statistics 2021-03-09 Danúbia R. Cunha , Jose A. Divino , Helton Saulo

Both genetic drift and natural selection cause the frequencies of alleles in a population to vary over time. Discriminating between these two evolutionary forces, based on a time series of samples from a population, remains an outstanding…

Populations and Evolution · Quantitative Biology 2013-12-09 Alison Feder , Sergey Kryazhimskiy , Joshua B. Plotkin

Recent literature provides many computational and modeling approaches for covariance matrices estimation in a penalized Gaussian graphical models but relatively little study has been carried out on the choice of the tuning parameter. This…

Methodology · Statistics 2009-09-08 Heng Lian

Under Markovian assumptions, we leverage a Central Limit Theorem (CLT) for the empirical measure in the test statistic of the composite hypothesis Hoeffding test so as to establish weak convergence results for the test statistic, and,…

Systems and Control · Computer Science 2018-02-14 Jing Zhang , Ioannis Ch. Paschalidis

In this paper, we give necessary and sufficient conditions for weighted $L^2$ estimates with matrix-valued measures of well localized operators. Namely, we seek estimates of the form: \[ \| T(\mathbf{W} f)\|_{L^2(\mathbf{V})} \le…

Functional Analysis · Mathematics 2016-11-22 Kelly Bickel , Amalia Culiuc , Sergei Treil , Brett D. Wick

In real life we often deal with independent but not identically distributed observations (i.n.i.d.o), for which the most well-known statistical model is the multiple linear regression model (MLRM) without random covariates. While the…

Statistics Theory · Mathematics 2021-02-25 Elena Castilla , Maria Jaenada , Leandro Pardo

Non-deterministic measurements are common in real-world scenarios: the performance of a stochastic optimization algorithm or the total reward of a reinforcement learning agent in a chaotic environment are just two examples in which…

Machine Learning · Statistics 2022-08-31 Etor Arza , Josu Ceberio , Ekhiñe Irurozki , Aritz Pérez

In many real-world problems, complex dependencies are present both among samples and among features. The Kronecker sum or the Cartesian product of two graphs, each modeling dependencies across features and across samples, has been used as…

Machine Learning · Statistics 2021-05-21 Jun Ho Yoon , Seyoung Kim

Count data play a critical role in medical research, such as heart disease. The Poisson regression model is a common technique for evaluating the impact of a set of covariates on the count responses. The mixture of Poisson regression models…

Methodology · Statistics 2023-09-13 Elsayed Ghanem , Moein Yoosefi , Armin Hatefi

We propose novel methodology for testing equality of model parameters between two high-dimensional populations. The technique is very general and applicable to a wide range of models. The method is based on sample splitting: the data is…

Methodology · Statistics 2013-01-17 Nicolas Städler , Sach Mukherjee

Neuron-level firing data is believed to be governed by latent activation patterns during task completion. Analysing repeated trials of a task allows us to study these patterns, typically by averaging in-vivo neural spikes across trials.…

Methodology · Statistics 2026-04-07 Angel Garcia de la Garza , Britton Sauerbrei , Jeff Goldsmith

We consider here estimation of an unknown probability density s belonging to L2(mu) where mu is a probability measure. We have at hand n i.i.d. observations with density s and use the squared L2-norm as our loss function. The purpose of…

Statistics Theory · Mathematics 2013-01-22 Lucien Birgé

This paper studies the impact of bootstrap procedure on the eigenvalue distributions of the sample covariance matrix under a high-dimensional factor structure. We provide asymptotic distributions for the top eigenvalues of bootstrapped…

Statistics Theory · Mathematics 2023-11-21 Long Yu , Peng Zhao , Wang Zhou

We consider the problem of estimating the covariance matrix of a random vector by observing i.i.d samples and each entry of the sampled vector is missed with probability $p$. Under the standard $L_4-L_2$ moment equivalence assumption, we…

Statistics Theory · Mathematics 2024-06-17 Pedro Abdalla

One of the goals in scaling sequential machine learning methods pertains to dealing with high-dimensional data spaces. A key related challenge is that many methods heavily depend on obtaining the inverse covariance matrix of the data. It is…

Computation · Statistics 2017-07-28 Tomer Lancewicki

Solutions of the bivariate, linear errors-in-variables estimation problem with unspecified errors are expected to be invariant under interchange and scaling of the coordinates. The appealing model of normally distributed true values and…

Statistics Theory · Mathematics 2012-02-07 David Leonard