English
Related papers

Related papers: Cross Sectional Regression with Cluster Dependence…

200 papers

Machine learning systems increasingly depend on pipelines of multiple algorithms to provide high quality and well structured predictions. This paper argues interaction effects between clustering and prediction (e.g. classification,…

Machine Learning · Statistics 2019-01-01 Matt Barnes , Artur Dubrawski

Conditional copula models allow dependence structures to vary with observed covariates while preserving a separation between marginal behavior and association. We study the uniform asymptotic behavior of kernel-weighted local likelihood…

Statistics Theory · Mathematics 2026-01-06 Mathias Nthiani Muia

Semisupervised methods inevitably invoke some assumption that links the marginal distribution of the features to the regression function of the label. Most commonly, the cluster or manifold assumptions are used which imply that the…

Statistics Theory · Mathematics 2011-12-02 Martin Azizyan , Aarti Singh , Larry Wasserman

Sequential data collection has emerged as a widely adopted technique for enhancing the efficiency of data gathering processes. Despite its advantages, such data collection mechanism often introduces complexities to the statistical inference…

Statistics Theory · Mathematics 2023-11-09 Mufang Ying , Koulik Khamaru , Cun-Hui Zhang

To measure the degree of agreement between two observers that independently classify $n$ subjects within $K$ categories, it is common to use different kappa type coefficients, the most common of which is the $\kappa_C$ coefficient (Cohen's…

Statistics Theory · Mathematics 2026-02-24 A. Martín Andrés , M. Álvarez Hernández

It is well known that if the power spectral density of a continuous time stationary stochastic process does not have a compact support, data sampled from that process at any uniform sampling rate leads to biased and inconsistent spectrum…

Statistics Theory · Mathematics 2010-06-09 Radhendushka Srivastava , Debasis Sengupta

We consider a robust linear regression model $y=X\beta^* + \eta$, where an adversary oblivious to the design $X\in \mathbb{R}^{n\times d}$ may choose $\eta$ to corrupt all but an $\alpha$ fraction of the observations $y$ in an arbitrary…

Machine Learning · Computer Science 2021-05-26 Tommaso d'Orsi , Gleb Novikov , David Steurer

The split-plot design assigns different interventions at the whole-plot and sub-plot levels, respectively, and induces a group structure on the final treatment assignments. A common strategy is to use the OLS fit of the outcome on the…

Methodology · Statistics 2021-10-25 Anqi Zhao , Peng Ding

Inference in linear panel data models is complicated by the presence of fixed effects when (some of) the regressors are not strictly exogenous. Under asymptotics where the number of cross-sectional observations and time periods grow at the…

Econometrics · Economics 2025-02-13 Ayden Higgins , Koen Jochmans

Misspecified models often provide useful information about the true data generating distribution. For example, if $y$ is a non-linear function of $x$ the least squares estimator $\hat{\beta}$ is an estimate of $\beta$, the slope of the best…

Methodology · Statistics 2017-05-17 James P. Long

A rank-invariant clustering of variables is introduced that is based on the predictive strength between groups of variables, i.e., two groups are assigned a high similarity if the variables in the first group contain high predictive…

Methodology · Statistics 2023-12-29 Sebastian Fuchs , Yuping Wang

Indirect Inference (I-I) estimation of structural parameters $\theta$ {{requires matching observed and simulated statistics, which are most often generated using an auxiliary model that depends on instrumental parameters $\beta$.}} {The…

Statistics Theory · Mathematics 2019-08-21 David T. Frazier , Eric Renault

This paper introduces a new fixed effects estimator for linear panel data models with clustered time patterns of unobserved heterogeneity. The method avoids non-convex and combinatorial optimization by combining a preliminary consistent…

Econometrics · Economics 2025-04-21 Martin Mugnier

In a cluster-randomized experiment, treatment is assigned to clusters of individual units of interest--households, classrooms, villages, etc.--instead of the units themselves. The number of clusters sampled and the number of units sampled…

Methodology · Statistics 2020-02-20 Yeng Xiong , Michael J. Higgins

A novel non-parametric estimator of the correlation between grouped measurements of a quantity is proposed in the presence of noise. This work is primarily motivated by functional brain network construction from fMRI data, where brain…

Methodology · Statistics 2023-02-16 Hanâ Lbath , Alexander Petersen , Wendy Meiring , Sophie Achard

The premise of independence among subjects in the same cluster/group often fails in practice, and models that rely on such untenable assumption can produce misleading results. To overcome this severe deficiency, we introduce a new…

Methodology · Statistics 2022-02-22 Jussiane Nader Gonçalves , Wagner Barreto-Souza , Hernando Ombao

We examine asymptotic properties of the OLS estimator when the values of the regressor of interest are assigned randomly and independently of other regressors. We find that the OLS variance formula in this case is often simplified,…

Econometrics · Economics 2023-03-21 Denis Chetverikov , Jinyong Hahn , Zhipeng Liao , Andres Santos

Estimating causal effects with propensity scores relies upon the availability of treated and untreated units observed at each value of the estimated propensity score. In settings with strong confounding, limited so-called "overlap" in…

Methodology · Statistics 2017-10-25 Corwin M Zigler , Matthew Cefalu

Causal inference necessarily relies upon untestable assumptions; hence, it is crucial to assess the robustness of obtained results to violations of identification assumptions. However, such sensitivity analysis is only occasionally…

Methodology · Statistics 2025-05-19 Tobias Freidling , Qingyuan Zhao

We provide various norm-based definitions of different types of cross-sectional dependence and the relations between them. These definitions facilitate to comprehend and to characterize the various forms of cross-sectional dependence, such…

Methodology · Statistics 2018-04-24 Gopal K Basak , Samarjit Das