中文
相关论文

相关论文: Markovian Interference in Experiments

200 篇论文

For many open quantum systems, a master equation approach employing the Markov approximation cannot reliably describe the dynamical behaviour. This is the case, for example, in a number of solid state or biological systems, and it has…

量子物理 · 物理学 2015-06-16 V. Venkataraman , A. D. K. Plato , Tommaso Tufarelli , M. S. Kim

Experimentation with interference poses a significant challenge in contemporary online platforms. Prior research on experimentation with interference has concentrated on the final output of a policy. The cumulative performance, while…

机器学习 · 计算机科学 2024-07-17 Su Jia , Peter Frazier , Nathan Kallus

One of the most basic problems in reinforcement learning (RL) is policy evaluation: estimating the long-term return, i.e., value function, corresponding to a given fixed policy. The celebrated Temporal Difference (TD) learning algorithm…

机器学习 · 计算机科学 2025-02-10 Sreejeet Maity , Aritra Mitra

We study the policy testing problem in discounted Markov decision processes (MDPs) in the fixed-confidence setting under a generative model with static sampling. The goal is to decide whether the value of a given policy exceeds a specified…

机器学习 · 统计学 2026-04-21 Kaito Ariu , Po-An Wang , Alexandre Proutiere , Kenshi Abe

In auction and matching markets, estimating the welfare effects of demand-side treatments is challenging because of spillovers through the mechanism. We develop a quasi-experimental approach that avoids parametric assumptions typically…

计量经济学 · 经济学 2026-03-03 Evan Munro

Motivated by wide-ranging applications such as video delivery over networks using Multiple Description Codes, congestion control, and inventory management, we study the state-tracking of a Markovian random process with a known transition…

信息论 · 计算机科学 2017-03-06 Parisa Mansourifard , Tara Javidi , Bhaskar Krishnamachari

Poor sample efficiency is a major limitation of deep reinforcement learning in many domains. This work presents an attention-based method to project neural network inputs into an efficient representation space that is invariant under…

机器学习 · 计算机科学 2020-03-23 John Mern , Dorsa Sadigh , Mykel J. Kochenderfer

This paper discusses difference-in-differences (DID) estimation when there exist many control variables, potentially more than the sample size. In this case, traditional estimation methods, which require a limited number of variables, do…

综合经济学 · 经济学 2019-01-09 Neng-Chieh Chang

A key methodological challenge in observational studies with interference between units is twofold: (1) each unit's outcome may depend on many others' treatments, and (2) treatment assignments may exhibit complex dependencies across units.…

统计方法学 · 统计学 2025-12-17 Souhardya Sengupta , Kosuke Imai , Georgia Papadogeorgou

Understanding the pathways whereby an intervention has an effect on an outcome is a common scientific goal. A rich body of literature provides various decompositions of the total intervention effect into pathway specific effects.…

统计方法学 · 统计学 2020-01-20 David Benkeser

We consider the sequential decision-making problem of making proactive request assignment and rejection decisions for a profit-maximizing operator of an autonomous mobility on demand system. We formalize this problem as a Markov decision…

机器学习 · 计算机科学 2023-05-11 Tobias Enders , James Harrison , Marco Pavone , Maximilian Schiffer

Although it may seem The Delayed Choice experiments contradict causality and one could construct an experiment which could possibly affect the past, using Many World interpretation we prove it is not possible. We also find a mathematical…

量子物理 · 物理学 2019-07-16 Dominik Šafránek

No man is an island, as individuals interact and influence one another daily in our society. When social influence takes place in experiments on a population of interconnected individuals, the treatment on a unit may affect the outcomes of…

统计方法学 · 统计学 2017-08-30 Edward K. Kao

By reusing data throughout training, off-policy deep reinforcement learning algorithms offer improved sample efficiency relative to on-policy approaches. For continuous action spaces, the most popular methods for off-policy learning include…

机器学习 · 计算机科学 2023-12-01 Jared Markowitz , Jesse Silverberg , Gary Collins

We extend the classical setting of an optimal stopping problem under full information to include for problems with an unknown state. The framework allows the unknown state to influence (i) the drift of the underlying process, (ii) the…

概率论 · 数学 2024-05-08 Erik Ekström , Yuqiong Wang

To unbiasedly evaluate multiple target policies, the dominant approach among RL practitioners is to run and evaluate each target policy separately. However, this evaluation method is far from efficient because samples are not shared across…

机器学习 · 计算机科学 2024-12-30 Shuze Daniel Liu , Claire Chen , Shangtong Zhang

In this article, we study the effect of vector-valued interventions in votes under a binary voter model, where each voter expresses their vote as a $0-1$ valued random variable to choose between two candidates. We assume that the outcome is…

应用统计 · 统计学 2022-10-17 Manit Paul , Rishideep Roy , Soudeep Deb

We derive an unbiased estimator for expectations over discrete random variables based on sampling without replacement, which reduces variance as it avoids duplicate samples. We show that our estimator can be derived as the…

机器学习 · 计算机科学 2020-02-17 Wouter Kool , Herke van Hoof , Max Welling

A growing number of researchers are conducting randomized experiments to analyze causal relationships in network settings where units influence one another. A dominant methodology for analyzing these experiments is design-based, leveraging…

统计方法学 · 统计学 2024-07-30 Ambarish Chattopadhyay , Kosuke Imai , Jose R. Zubizarreta

Off-policy evaluation is a key component of reinforcement learning which evaluates a target policy with offline data collected from behavior policies. It is a crucial step towards safe reinforcement learning and has been used in…

机器学习 · 计算机科学 2020-12-01 Jinlin Lai , Lixin Zou , Jiaxing Song
‹ 上一页 1 8 9 10 下一页 ›