English
Related papers

Related papers: Weak (Proxy) Factors Robust Hansen-Jagannathan Dis…

200 papers

Large-scale multiple testing under static factor models is widely used to detect sparse signals in high-dimensional data. However, static factor models are arguably too stringent because they ignore serial correlation, which seriously…

Statistics Theory · Mathematics 2025-04-04 Xinxin Yang , Lilun Du

We consider the estimation of approximate factor models for time series data, where strong serial and cross-sectional correlations amongst the idiosyncratic component are present. This setting comes up naturally in many applications, but…

Methodology · Statistics 2019-12-10 Jiahe Lin , George Michailidis

The linear instrumental variable (IV) model is widely used in observational studies, yet its validity hinges on strong assumptions. Classical specification tests such as the Sargan-Hansen J test are limited to overidentified settings and…

Methodology · Statistics 2026-04-21 Cyrill Scheidegger , Malte Londschien , Peter Bühlmann

In many randomized experiments, the treatment effect of the long-term metric (i.e. the primary outcome of interest) is often difficult or infeasible to measure. Such long-term metrics are often slow to react to changes and sufficiently…

Since the weak convergence for stochastic processes does not account for the growth of information over time which is represented by the underlying filtration, a slightly erroneous stochastic model in weak topology may cause huge loss in…

Methodology · Statistics 2024-05-27 Jiajie Tao , Hao Ni , Chong Liu

Factor models are a class of powerful statistical models that have been widely used to deal with dependent measurements that arise frequently from various applications from genomics and neuroscience to economics and finance. As data are…

Methodology · Statistics 2018-08-14 Jianqing Fan , Kaizheng Wang , Yiqiao Zhong , Ziwei Zhu

We consider random walks on the cone of $m \times m$ positive definite matrices, where the underlying random matrices have orthogonally invariant distributions on the cone and the Riemannian metric is the measure of distance on the cone. By…

Probability · Mathematics 2022-06-22 Armine Bagyan , Donald Richards

Weak signal identification and inference are very important in the area of penalized model selection, yet they are under-developed and not well-studied. Existing inference procedures for penalized estimators are mainly focused on strong…

Methodology · Statistics 2016-11-16 Peibei Shi , Annie Qu

Pervasive cross-section dependence is increasingly recognized as a characteristic of economic data and the approximate factor model provides a useful framework for analysis. Assuming a strong factor structure where $\Lop\Lo/N^\alpha$ is…

Econometrics · Economics 2023-03-07 Jushan Bai , Serena Ng

Evaluating fairness can be challenging in practice because the sensitive attributes of data are often inaccessible due to privacy constraints. The go-to approach that the industry frequently adopts is using off-the-shelf proxy models to…

Machine Learning · Computer Science 2023-02-01 Zhaowei Zhu , Yuanshun Yao , Jiankai Sun , Hang Li , Yang Liu

In this article, we address the challenge of identifying skilled mutual funds among a large pool of candidates, utilizing the linear factor pricing model. Assuming observable factors with a weak correlation structure for the idiosyncratic…

Methodology · Statistics 2024-11-22 Hongfei Wang , Long Feng , Ping Zhao , Zhaojun Wang

This paper proposes a new method for determining similarity and anomalies between time series, most practically effective in large collections of (likely related) time series, by measuring distances between structural breaks within such a…

Machine Learning · Computer Science 2020-12-01 Nick James , Max Menzies , Lamiae Azizi , Jennifer Chan

This paper studies optimal estimation of large-dimensional nonlinear factor models. The key challenge is that the observed variables are possibly nonlinear functions of some latent variables where the functional forms are left unspecified.…

Statistics Theory · Mathematics 2023-11-14 Yingjie Feng

We develop a factor analysis for mixed continuous and binary observed variables. To this end, we utilized a recently developed multivariate probability distribution for mixed-type random variables, the Gaussian-Grassmann distribution. In…

Methodology · Statistics 2025-12-12 Takashi Arai

In this paper, we use the results in Andrews and Cheng (2012), extended to allow for parameters to be near or at the boundary of the parameter space, to derive the asymptotic distributions of the two test statistics that are used in the…

Econometrics · Economics 2022-10-21 Philipp Ketz

Exogeneity is key for IV estimators, which can assessed via overidentification (OID) tests. We discuss the Kleibergen-Paap (KP) rank test as a heteroskedasticity-robust OID test and compare to the typical J-test. We derive the…

Econometrics · Economics 2025-09-26 Stuart Lane , Frank Windmeijer

BACKGROUND: Random-effects meta-analysis is commonly performed by first deriving an estimate of the between-study variation, the heterogeneity, and subsequently using this as the basis for combining results, i.e., for estimating the effect,…

Methodology · Statistics 2015-11-18 Christian Röver , Guido Knapp , Tim Friede

We propose a weak-identification-robust test for linear instrumental variable (IV) regressions with high-dimensional instruments, whose number is allowed to exceed the sample size. In addition, our test is robust to general error…

Econometrics · Economics 2025-07-01 Qu Feng , Sombut Jaidee , Wenjie Wang

Information Value (IV) is a widely used technique for feature selection prior to the modeling phase, particularly in credit scoring and related domains. However, conventional IV-based practices rely on fixed empirical thresholds, which lack…

Statistics Theory · Mathematics 2026-01-28 Helder Rojas , Cirilo Alvarez , Nilton Rojas

The search for lightweight authentication protocols suitable for low-cost RFID tags constitutes an active and challenging research area. In this context, a family of protocols based on the LPN problem has been proposed: the so-called…

Cryptography and Security · Computer Science 2015-05-20 Jose Carrijo , Rafael Tonicelli , Anderson C. A. Nascimento