English
Related papers

Related papers: A regularized MANOVA test for semicontinuous high-…

200 papers

In this note we derive large-scale regularity properties of solutions to second-order linear elliptic equations with random coefficients on the half- space with homogeneous Neumann boundary data; it is a companion to arXiv:1604.02717 in…

Analysis of PDEs · Mathematics 2017-03-14 Claudia Raithel

We consider Markov decision processes (MDPs) in which the transition probabilities and rewards belong to an uncertainty set parametrized by a collection of random variables. The probability distributions for these random parameters are…

Logic in Computer Science · Computer Science 2020-02-26 Murat Cubuktepe , Nils Jansen , Sebastian Junges , Joost-Pieter Katoen , Ufuk Topcu

This paper considers testing a covariance matrix $\Sigma$ in the high dimensional setting where the dimension $p$ can be comparable or much larger than the sample size $n$. The problem of testing the hypothesis $H_0:\Sigma=\Sigma_0$ for a…

Statistics Theory · Mathematics 2013-12-18 T. Tony Cai , Zongming Ma

In randomized controlled trials, covariate adjustment can improve statistical power and reduce the required sample size compared with unadjusted estimators. Several regulatory agencies have released guidance on covariate adjustment, which…

Methodology · Statistics 2025-08-27 Takumi Kanata , Yasuhiro Hagiwara , Koji Oba

Models that perform well on a training domain often fail to generalize to out-of-domain (OOD) examples. Data augmentation is a common method used to prevent overfitting and improve OOD generalization. However, in natural language, it is…

Computation and Language · Computer Science 2020-10-06 Nathan Ng , Kyunghyun Cho , Marzyeh Ghassemi

We propose a high dimensional mean test framework for shrinking random variables, where the underlying random variables shrink to zero as the sample size increases. By pooling observations across overlapping subsets of dimensions, we…

Methodology · Statistics 2026-02-11 Liujun Chen , Chen Zhou

Establishing a low-dimensional representation of the data leads to efficient data learning strategies. In many cases, the reduced dimension needs to be explicitly stated and estimated from the data. We explore the estimation of dimension in…

Methodology · Statistics 2022-02-10 Wei Q. Deng , Radu V. Craiu

We propose a test for a change in the mean for a sequence of functional observations that are only partially observed on subsets of the domain, with no information available on the complement. The framework accommodates important scenarios,…

Methodology · Statistics 2025-10-10 Šárka Hudecová , Claudia Kirch

The computational complexity of simultaneous inference methods in high-dimensional linear regression models quickly increases with the number variables. This paper proposes a computationally efficient method based on the Moore-Penrose…

Statistics Theory · Mathematics 2021-02-02 Tom Boot , Didier Nibbering

We present a novel technique to parametrize experimental data, based on the construction of a probability measure in the space of functions, which retains the full experimental information on errors and correlations. This measure is…

High Energy Physics - Phenomenology · Physics 2007-05-23 Joan Rojo

We study theoretical properties of regularized robust M-estimators, applicable when data are drawn from a sparse high-dimensional linear model and contaminated by heavy-tailed distributions and/or outliers in the additive errors and…

Statistics Theory · Mathematics 2015-01-05 Po-Ling Loh

We devise survey-weighted pseudo posterior distribution estimators under two-stage informative sampling of both primary clusters and secondary nested units for a one-way analysis of variance (ANOVA) population generating model as a simple…

Methodology · Statistics 2023-05-16 Terrance D. Savitsky , Matthew R. Williams , Sanvesh Srivastava

When the outcome of interest is semicontinuous and collected longitudinally, efficient testing can be difficult. Daily rainfall data is an excellent example which we use to illustrate the various challenges. Even under the simplest…

Methodology · Statistics 2018-07-10 Harlan Campbell

Semi-supervised classification, where unlabeled data are massive but labeled data are limited, often arises in machine learning applications. We address this challenge under high-dimensional data by leveraging the manifold and cluster…

Machine Learning · Statistics 2026-04-28 Ruoxu Tan , Yiming Zang

This paper considers the problem of testing temporal homogeneity of $p$-dimensional population mean vectors from the repeated measurements of $n$ subjects over $T$ times. To cope with the challenges brought by high-dimensional longitudinal…

Methodology · Statistics 2016-08-29 Ping-Shou Zhong , Jun Li

Simultaneously addressing multiple objectives is becoming increasingly important in modern machine learning. At the same time, data is often high-dimensional and costly to label. For a single objective such as prediction risk, conventional…

Machine Learning · Statistics 2025-03-13 Tobias Wegel , Filip Kovačević , Alexandru Ţifrea , Fanny Yang

Statistical hypothesis testing and effect size measurement are routine parts of quantitative research. Advancements in computer processing power have greatly improved the capability of statistical inference through the availability of…

Methodology · Statistics 2024-01-18 Michael J. Crosse , John J. Foxe , Sophie Molholm

The development of data acquisition systems is facilitating the collection of data that are apt to be modelled as functional data. In some applications, the interest lies in the identification of significant differences in group functional…

Performances of the Multivariate Kurtosis are investigated when applied to colored data, with or without Auto-Regressive pre-whitening, and with or without projection onto a lower-dimensional random subspace. Computer experiments…

Methodology · Statistics 2022-06-15 Sara Elbouch , Olivier Michel , Pierre Comon

We present semi-supervised models with data augmentation (SMDA), a semi-supervised text classification system to classify interactive affective responses. SMDA utilizes recent transformer-based models to encode each sentence and employs…

Computation and Language · Computer Science 2020-04-24 Jiaao Chen , Yuwei Wu , Diyi Yang
‹ Prev 1 8 9 10 Next ›