中文
相关论文

相关论文: Smooth Multi-Policy Causal Effect Estimation in Lo…

200 篇论文

We introduce Compositional Imitation Learning and Execution (CompILE): a framework for learning reusable, variable-length segments of hierarchically-structured behavior from demonstration data. CompILE uses a novel unsupervised,…

This paper studies the synthesis of control policies for heterogeneous and interconnected multi-agent systems that collaborate through data exchange over a communication network to minimize a collective cost. We propose a distributed…

信号处理 · 电气工程与系统科学 2026-05-15 Mohammadreza Barzegaran , Kemeng Han , Hamid Jafarkhani

Linear mixed-effects model (LMM) is a cornerstone of longitudinal data analysis, but is limited to adeptly make heterogeneous analyses predictable under both group-specific fixed effects and subject-specific random effects. To address this…

统计方法学 · 统计学 2026-03-10 Xinkai Yue , Xiaodong Yan , Haohui Han , Liya Fu

Randomized experiments are widely used to estimate causal effects across a variety of domains. However, classical causal inference approaches rely on critical independence assumptions that are violated by network interference, when the…

统计方法学 · 统计学 2022-10-18 Mayleen Cortez , Matthew Eichhorn , Christina Lee Yu

Pretrained encoders for mathematical texts have achieved significant improvements on various tasks such as formula classification and information retrieval. Yet they remain limited in representing and capturing student strategies for entire…

计算机与社会 · 计算机科学 2026-04-13 Siddhartha Pradhan , Ethan Prihar , Erin Ottmar

In recent years, precision treatment strategy have gained significant attention in medical research, particularly for patient care. We propose a novel framework for estimating conditional average treatment effects (CATE) in time-to-event…

统计方法学 · 统计学 2024-07-29 Runjia Li , Victor B. Talisa , Chung-Chou H. Chang

Individual Treatment Effect (ITE) prediction is an important area of research in machine learning which aims at explaining and estimating the causal impact of an action at the granular level. It represents a problem of growing interest in…

The analysis of randomized controlled trials is often complicated by intercurrent events (IEs) -- events that occur after treatment initiation and affect either the interpretation or existence of outcome measurements. Examples include…

统计方法学 · 统计学 2026-04-07 Sizhu Lu , Yanyao Yi , Yongming Qu , Huayu Karen Liu , Ting Ye , Peng Ding

Reinforcement learning (RL) is a classical tool to solve network control or policy optimization problems in unknown environments. The original Q-learning suffers from performance and complexity challenges across very large networks. Herein,…

机器学习 · 计算机科学 2024-09-02 Talha Bozkus , Urbashi Mitra

Non-linear mixed effects modeling and simulation (NLME M&S) is evaluated to be used for standardization with longitudinal data in presence of confounders. Standardization is a well-known method in causal inference to correct for confounding…

统计方法学 · 统计学 2024-05-01 Christian Bartels , Martina Scauda , Neva Coello , Thomas Dumortier , Bjoern Bornkamp , Giusi Moffa

Medical treatments often involve a sequence of decisions, each informed by previous outcomes. This process closely aligns with reinforcement learning (RL), a framework for optimizing sequential decisions to maximize cumulative rewards under…

机器学习 · 计算机科学 2024-10-15 Ali Shirali , Alexander Schubert , Ahmed Alaa

Treatment non-compliance, where individuals deviate from their assigned experimental conditions, frequently complicates the estimation of causal effects. To address this, we introduce a novel learning framework based on a mixture of experts…

统计方法学 · 统计学 2025-06-25 François Grolleau , Céline Béji , Raphaël Porcher , François Petit

In recent years, there has been growing interest in causal machine learning estimators for quantifying subject-specific effects of a binary treatment on time-to-event outcomes. Estimation approaches have been proposed which attenuate the…

统计方法学 · 统计学 2026-03-30 Matthew Pryce , Karla Diaz-Ordaz , Ruth H. Keogh , Stijn Vansteelandt

Large language models (LLMs) demonstrate strong reasoning abilities in solving complex real-world problems. Yet, the internal mechanisms driving these complex reasoning behaviors remain opaque. Existing interpretability approaches targeting…

人工智能 · 计算机科学 2026-02-04 Changming Li , Kaixing Zhang , Haoyun Xu , Yingdong Shi , Zheng Zhang , Kaitao Song , Kan Ren

In reinforcement learning, off-policy evaluation (OPE) is the problem of estimating the expected return of an evaluation policy given a fixed dataset that was collected by running one or more different policies. One of the more empirically…

机器学习 · 计算机科学 2023-10-31 Brahma S. Pavse , Josiah P. Hanna

Estimating long-term causal effects by combining long-term observational and short-term experimental data is a crucial but challenging problem in many real-world scenarios. In existing methods, several ideal assumptions, e.g. latent…

机器学习 · 计算机科学 2025-05-12 Ruichu Cai , Junjie Wan , Weilin Chen , Zeqin Yang , Zijian Li , Peng Zhen , Jiecheng Guo

As estimation of Heterogeneous Treatment Effect (HTE) is increasingly adopted across a wide range of scientific and industrial applications, the treatment action space can naturally expand, from a binary treatment variable to a structured…

机器学习 · 计算机科学 2025-07-09 Jennifer Y. Zhang , Shuyang Du , Will Y. Zou

Routinely collected data from electronic health records (EHR) provide opportunities to study effects of longitudinal treatment strategies in real-world clinical settings. A challenge presented by EHR data is that frequency of covariate…

应用统计 · 统计学 2026-04-14 Leah Pirondini , Karla Diaz-Ordaz , Edward Palmer , Ruth H. Keogh

This paper develops a unified framework for the identification, estimation, and uniform inference of local treatment effects (LTEs) in sharp regression kink designs (RKDs). These LTEs quantify the effect of a marginal change in the…

计量经济学 · 经济学 2025-11-03 Zhixin Wang , Zhengyu Zhang

We show that the popular reinforcement learning (RL) strategy of estimating the state-action value (Q-function) by minimizing the mean squared Bellman error leads to a regression problem with confounding, the inputs and output noise being…

机器学习 · 计算机科学 2022-12-01 Yutian Chen , Liyuan Xu , Caglar Gulcehre , Tom Le Paine , Arthur Gretton , Nando de Freitas , Arnaud Doucet