中文
相关论文

相关论文: Matching in Selective and Balanced Representation …

200 篇论文

When investigators seek to estimate causal effects, they often assume that selection into treatment is based only on observed covariates. Under this identification strategy, analysts must adjust for observed confounders. While basic…

应用统计 · 统计学 2019-01-09 Luke Keele , Dylan Small

Selective inference aims at providing valid inference after a data-driven selection of models or hypotheses. It is essential to avoid overconfident results and replicability issues. While significant advances have been made in this area for…

统计方法学 · 统计学 2025-03-14 Matteo D'Alessandro , Magne Thoresen

The amount of data for machine learning (ML) applications is constantly growing. Not only the number of observations, especially the number of measured variables (features) increases with ongoing digitization. Selecting the most appropriate…

机器学习 · 计算机科学 2021-11-25 Konstantin Hopf , Sascha Reifenrath

Matching is one of the most widely used causal inference frameworks in observational studies. However, all the existing matching-based causal inference methods are designed for either a single treatment with general treatment types (e.g.,…

统计方法学 · 统计学 2025-12-23 Jianan Zhu , Tianruo Zhang , Diana Silver , Ellicott Matthay , Omar El-Shahawy , Hyunseung Kang , Siyu Heng

Matching estimators for average treatment effects are widely used in the binary treatment setting, in which missing potential outcomes are imputed as the average of observed outcomes of all matches for each unit. With more than two…

统计方法学 · 统计学 2019-04-29 Anthony D. Scotina , Francesca L. Beaudoin , Roee Gutman

Raman spectroscopy is an effective, low-cost, non-intrusive technique often used for chemical identification. Typical approaches are based on matching observations to a reference database, which requires careful preprocessing, or supervised…

机器学习 · 计算机科学 2022-10-12 Bo Li , Mikkel N. Schmidt , Tommy S. Alstrøm

We propose a matching method for observational data that matches units with others in unit-specific, hyper-box-shaped regions of the covariate space. These regions are large enough that many matches are created for each unit and small…

统计方法学 · 统计学 2020-08-11 Marco Morucci , Vittorio Orlandi , Sudeepa Roy , Cynthia Rudin , Alexander Volfovsky

We propose a matching method that recovers direct treatment effects from randomized experiments where units are connected in an observed network, and units that share edges can potentially influence each others' outcomes. Traditional…

统计方法学 · 统计学 2020-05-12 M. Usaid Awan , Marco Morucci , Vittorio Orlandi , Sudeepa Roy , Cynthia Rudin , Alexander Volfovsky

Understanding effect modification -- how treatment effects vary across subpopulations -- is practically important in observational studies, as it helps identify which subgroups are likely to benefit from a given treatment. In this paper, we…

统计方法学 · 统计学 2026-05-12 Yu Gui , Dylan S Small , Zhimei Ren

We study aleatoric and epistemic uncertainty estimation in a learned regressive system dynamics model. Disentangling aleatoric uncertainty (the inherent randomness of the system) from epistemic uncertainty (the lack of data) is crucial for…

机器学习 · 计算机科学 2025-03-21 Zhiyu An , Zhibo Hou , Wan Du

The challenges in feature selection, particularly in balancing model accuracy, interpretability, and computational efficiency, remain a critical issue in advancing machine learning methodologies. To address these complexities, this study…

机器学习 · 计算机科学 2026-01-06 Nachiket Kapure , Harsh Joshi , Parul Kumari , Rajeshwari Mistri , Manasi Mali

Regression models are used for inference and prediction in a wide range of applications providing a powerful scientific tool for researchers and analysts from different fields. In many research fields the amount of available data as well as…

统计方法学 · 统计学 2018-06-08 Aliaksandr Hubin , Geir Storvik , Florian Frommlet

To estimate casual treatment effects, we propose a new matching approach based on the reduced covariates obtained from sufficient dimension reduction. Compared to the original covariates and the propensity score, which are commonly used for…

统计方法学 · 统计学 2017-02-03 Wei Luo , Yeying Zhu

Estimating the average treatment effect (ATE) from observational data is challenging due to selection bias. Existing works mainly tackle this challenge in two ways. Some researchers propose constructing a score function that satisfies the…

机器学习 · 计算机科学 2022-09-07 Yiyan Huang , Cheuk Hang Leung , Shumin Ma , Qi Wu , Dongdong Wang , Zhixiang Huang

We study the problem of feature selection in general machine learning (ML) context, which is one of the most critical subjects in the field. Although, there exist many feature selection methods, however, these methods face challenges such…

机器学习 · 计算机科学 2024-06-18 Mehmet Y. Turali , Mehmet E. Lorasdagi , Ali T. Koc , Suleyman S. Kozat

High-dimensional variable selection, with many more covariates than observations, is widely documented in standard regression models, but there are still few tools to address it in non-linear mixed-effects models where data are collected…

The fundamental problem in treatment effect estimation from observational data is confounder identification and balancing. Most of the previous methods realized confounder balancing by treating all observed pre-treatment variables as…

统计方法学 · 统计学 2021-10-13 Anpeng Wu , Kun Kuang , Junkun Yuan , Bo Li , Runze Wu , Qiang Zhu , Yueting Zhuang , Fei Wu

Support vector machine (SVM) is one of the most popular classification algorithms in the machine learning literature. We demonstrate that SVM can be used to balance covariates and estimate average causal effects under the unconfoundedness…

统计方法学 · 统计学 2021-07-02 Alexander Tarr , Kosuke Imai

In many medical and business applications, researchers are interested in estimating individualized treatment effects using data from a randomized experiment. For example in medical applications, doctors learn the treatment effects from…

统计方法学 · 统计学 2022-03-01 Kevin Wu Han , Han Wu

Selection of covariates is crucial in the estimation of average treatment effects given observational data with high or even ultra-high dimensional pretreatment variables. Existing methods for this problem typically assume sparse linear…

统计方法学 · 统计学 2023-03-20 Juan Chen , Yingchun Zhou