中文
相关论文

相关论文: Three-quarter Sibling Regression for Denoising Obs…

200 篇论文

We propose the orthogonal random forest, an algorithm that combines Neyman-orthogonality to reduce sensitivity with respect to estimation error of nuisance parameters with generalized random forests (Athey et al., 2017)--a flexible…

机器学习 · 计算机科学 2019-09-27 Miruna Oprescu , Vasilis Syrgkanis , Zhiwei Steven Wu

Missing values are unavoidable in many applications of machine learning and present challenges both during training and at test time. When variables are missing in recurring patterns, fitting separate pattern submodels have been proposed as…

机器学习 · 计算机科学 2023-11-27 Lena Stempfle , Ashkan Panahi , Fredrik D. Johansson

Detecting and measuring confounding effects from data is a key challenge in causal inference. Existing methods frequently assume causal sufficiency, disregarding the presence of unobserved confounding variables. Causal sufficiency is both…

人工智能 · 计算机科学 2024-09-27 Abbavaram Gowtham Reddy , Vineeth N Balasubramanian

We study methods for simultaneous analysis of many noisy and biased estimates, each paired with an even noisier estimate of its own bias. The analyst's goal is to construct short calibrated intervals for each parameter. The standard…

统计方法学 · 统计学 2026-05-11 Wanyi Ling , Sida Li , Junming Guan , Nikolaos Ignatiadis

Data-driven, model-free analytics are natural choices for discovery and forecasting of complex, nonlinear systems. Methods that operate in the system state-space require either an explicit multidimensional state-space, or, one approximated…

机器学习 · 统计学 2021-03-15 Joseph Park , Gerald M Pao , Erik Stabenau , George Sugihara , Thomas Lorimer

Managers, employers, policymakers, and others often seek to understand whether decisions are biased against certain groups. One popular analytic strategy is to estimate disparities after adjusting for observed covariates, typically with a…

应用统计 · 统计学 2024-01-29 Jongbin Jung , Sam Corbett-Davies , Johann D. Gaebler , Ravi Shroff , Sharad Goel

We study variable selection (also called support recovery) in high-dimensional sparse linear regression when one has external information on which variables are likely to be associated with the response. Consistent recovery is only possible…

统计理论 · 数学 2026-02-16 Paul Rognon-Vael , David Rossell , Piotr Zwiernik

In many applications of causal inference, the treatment received by one unit may influence the outcome of another, a phenomenon referred to as interference. Although there are several frameworks for conducting causal inference in the…

统计方法学 · 统计学 2025-11-27 Matvey Ortyashov , AmirEmad Ghassami

Observational and/or astrophysical systematics modulating the observed number of luminous tracers can constitute a major limitation in the cosmological exploitation of surveys of the large scale structure of the universe. Part of this…

This paper considers learning the hidden causal network of a linear networked dynamical system (NDS) from the time series data at some of its nodes -- partial observability. The dynamics of the NDS are driven by colored noise that generates…

机器学习 · 计算机科学 2024-02-13 Augusto Santos , Diogo Rente , Rui Seabra , José M. F. Moura

Educational disparities are rooted in and perpetuate social inequalities across multiple dimensions such as race, socioeconomic status, and geography. To reduce disparities, most intervention strategies focus on a single domain and…

统计方法学 · 统计学 2026-04-17 Soojin Park , Su Yeon Kim , Xinyao Zheng , Chioun Lee

One of the fundamental challenges found throughout the data sciences is to explain why things happen in specific ways, or through which mechanisms a certain variable $X$ exerts influences over another variable $Y$. In statistics and machine…

统计方法学 · 统计学 2023-06-09 Drago Plecko , Elias Bareinboim

Separating signals from an additive mixture may be an unnecessarily hard problem when one is only interested in specific properties of a given signal. In this work, we tackle simpler "statistical component separation" problems that focus on…

机器学习 · 统计学 2024-03-01 Bruno Régaldo-Saint Blancard , Michael Eickenberg

Observed associations in a database may be due in whole or part to variations in unrecorded (latent) variables. Identifying such variables and their causal relationships with one another is a principal goal in many scientific and practical…

机器学习 · 计算机科学 2012-12-12 Ricardo Silva , Richard Scheines , Clark Glymour , Peter L. Spirtes

When employing non-linear methods to characterise complex systems, it is important to determine to what extent they are capturing genuine non-linear phenomena that could not be assessed by simpler spectral methods. Specifically, we are…

统计方法学 · 统计学 2021-09-22 Pedro A. M. Mediano , Fernando E. Rosas , Adam B. Barrett , Daniel Bor

This paper develops a new framework, called modular regression, to utilize auxiliary information -- such as variables other than the original features or additional data sets -- in the training process of linear models. At a high level, our…

统计方法学 · 统计学 2023-11-27 Ying Jin , Dominik Rothenhäusler

The monotonic ordinal classification has increased the interest of researchers and practitioners within machine learning community in the last years. In real applications, the problems with monotonicity constraints are very frequent. To…

人工智能 · 计算机科学 2018-10-23 José-Ramón Cano , Julián Luengo , Salvador García

Noise in various interferometer systems can sometimes couple non-linearly to create excess noise in the gravitational wave (GW) strain data. Third-order statistics, such as bicoherence and biphase, can identify these couplings and help…

广义相对论与量子宇宙学 · 物理学 2024-07-29 Bernard Hall , Sudhagar Suyamprakasam , Nairwita Mazumder , Anupreeta More , Sukanta Bose

Regression method has been widely used to explore relationship between dependent and independent variables. In practice, data issues such as censoring and missing data often exist. When the response variable is (fixed) censored, Tobit…

统计方法学 · 统计学 2021-07-06 Hailin Huang

We argue that the selective inclusion of data points based on latent objectives is common in practical situations, such as music sequences. Since this selection process often distorts statistical analysis, previous work primarily views it…

机器学习 · 计算机科学 2024-07-02 Yujia Zheng , Zeyu Tang , Yiwen Qiu , Bernhard Schölkopf , Kun Zhang