English
Related papers

Related papers: Concentration inequalities for the sample correlat…

200 papers

The random coefficients model is an extension of the linear regression model that allows for unobserved heterogeneity in the population by modeling the regression coefficients as random variables. Given data from this model, the statistical…

Methodology · Statistics 2018-03-15 Fabian Dunker , Konstantin Eckle , Katharina Proksch , Johannes Schmidt-Hieber

Subsampling is an efficient method to deal with massive data. In this paper, we investigate the optimal subsampling for linear quantile regression when the covariates are functions. The asymptotic distribution of the subsampling estimator…

Numerical Analysis · Mathematics 2022-05-06 Qian Yan , Hanyu Li , Chengmei Niu

For many important problems the quantity of interest is an unknown function of the parameters, which is a random vector with known statistics. Since the dependence of the output on this random vector is unknown, the challenge is to identify…

Machine Learning · Statistics 2021-04-28 Themistoklis P. Sapsis

Composite endpoints are widely used in cardiovascular clinical trials to improve statistical efficiency while preserving clinical relevance. The Win Ratio (WR) measure and more general frameworks of Win Statistics have emerged as…

Methodology · Statistics 2025-11-24 Yunhan Mou , Fan Li , Denise Esserman , Yuan Huang

A common goal in observational research is to estimate marginal causal effects in the presence of confounding variables. One solution to this problem is to use the covariate distribution to weight the outcomes such that the data appear…

Methodology · Statistics 2020-08-18 Kevin P. Josey , Elizabeth Juarez-Colunga , Fan Yang , Debashis Ghosh

Computing ratios of normalizing constants plays an important role in statistical modeling. Two important examples are hypothesis testing in latent variables models, and model comparison in Bayesian statistics. In both examples, the…

Applications · Statistics 2024-08-26 Tom Guédon , Charlotte Baey , Estelle Kuhn

Cochran's $Q$ statistic is routinely used for testing heterogeneity in meta-analysis. Its expected value is also used for estimation of between-study variance $\tau^2$. Cochran's $Q$, or $Q_{IV}$, uses estimated inverse-variance weights…

Methodology · Statistics 2021-03-08 Ilyas Bakbergenuly , David C. Hoaglin , Elena Kulinskaya

We consider the standard non-parametric regression model with Gaussian errors but where the data consist of different samples. The question to be answered is whether the samples can be adequately represented by the same regression function.…

Statistics Theory · Mathematics 2008-09-17 A. Kovac , P. L. Davies

As context windows in large language models continue to expand, it is essential to characterize how attention behaves at extreme sequence lengths. We introduce token-sample complexity: the rate at which attention computed on $n$ tokens…

Machine Learning · Computer Science 2026-03-24 Léa Bohbot , Cyril Letrouit , Gabriel Peyré , François-Xavier Vialard

Testing differences between a treatment and control group is common practice in biomedical research like randomized controlled trials (RCT). The standard two-sample t-test relies on null hypothesis significance testing (NHST) via p-values,…

Methodology · Statistics 2020-05-18 Riko Kelter

Let $Y$ be a nonnegative random variable with mean $\mu$ and finite positive variance $\sigma^2$, and let $Y^s$, defined on the same space as $Y$, have the $Y$ size biased distribution, that is, the distribution characterized by…

Probability · Mathematics 2011-06-20 Subhankar Ghosh , Larry Goldstein

We derive a Gaussian approximation result for the maximum of a sum of random vectors under $(2+\iota)$-th moments. Our main theorem is abstract and nonasymptotic, and can be applied to a variety of statistical learning problems. The proof…

Statistics Theory · Mathematics 2019-05-28 Qiang Sun

In a range of genomic applications, it is of interest to quantify the evidence that the signal at site~$i$ is active given conditionally independent replicate observations summarized by the sample mean and variance $(\bar Y, s^2)$ at each…

Statistics Theory · Mathematics 2023-12-20 Micol Tresoldi , Daniel Xiang , Peter McCullagh

This paper is about vector autoregressive-moving average (VARMA) models with time-dependent coefficients to represent non-stationary time series. Contrarily to other papers in the univariate case, the coefficients depend on time but not on…

Statistics Theory · Mathematics 2015-06-05 Abdelkamel Alj , Christophe Ley , Guy Mélard

We propose an empirically stable and asymptotically efficient covariate-balancing approach to the problem of estimating survival causal effects in data with conditionally-independent censoring. This addresses a challenge often encountered…

In comparative studies, such as in causal inference and clinical trials, balancing important covariates is often one of the most important concerns for both efficient and credible comparison. However, chance imbalance still exists in many…

Methodology · Statistics 2018-07-30 Yichen Qin , Yang Li , Wei Ma , Feifang Hu

We consider a random variable $X$ that takes values in a (possibly infinite-dimensional) topological vector space $\mathcal{X}$. We show that, with respect to an appropriate "normal distance" on $\mathcal{X}$, concentration inequalities for…

Probability · Mathematics 2010-09-27 Timothy John Sullivan , Houman Owhadi

Following [1], the aim of this paper is to analyze the relative weighted entropy involving the central moments weight functions. We compare the standard relative entropy with the weighted case in two particular forms of Gaussian…

Information Theory · Computer Science 2015-06-23 Salimeh Yasaei Sekeh , Adriano Polpo

The Gaussian product inequality is an important conjecture concerning the moments of Gaussian random vectors. While all attempts to prove the Gaussian product inequality in full generality have been unsuccessful to date, numerous partial…

Probability · Mathematics 2022-04-26 Dominic Edelmann , Donald Richards , Thomas Royen

Covariance matrix estimation concerns the problem of estimating the covariance matrix from a collection of samples, which is of extreme importance in many applications. Classical results have shown that $O(n)$ samples are sufficient to…

Information Theory · Computer Science 2019-03-19 Wei Cui , Xu Zhang , Yulong Liu