中文
相关论文

相关论文: Reading policies for joins: An asymptotic analysis

200 篇论文

Two main procedures characterize the way in which social actors evaluate the qualities of the options in decision-making processes: they either seek to evaluate their intrinsic qualities (individual learners), or they rely on the opinion of…

物理与社会 · 物理学 2024-07-31 Arkadiusz Jędrzejewski , Laura Hernández

Empirical research shows that individuals' responses to treatments vary along latent characteristics, such as innate ability or motivation. Therefore, a policymaker seeking to maximize welfare may consider designing policies based on…

计量经济学 · 经济学 2026-05-06 Giacomo Opocher

The purpose of this article is to examine the greedy adaptive measurement policy in the context of a linear Guassian measurement model with an optimization criterion based on information gain. In the special case of sequential scalar…

信息论 · 计算机科学 2012-08-20 Entao Liu , Edwin K. P. Chong , Louis L. Scharf

In "Recognizing the Maximum of a Sequence", Gilbert and Mosteller analyze a full information game where n measurements from an uniform distribution are drawn and a player (knowing n) must decide at each draw whether or not to choose that…

概率论 · 数学 2018-05-30 Marcos Costa Santos Carreira

We study behavior of the restricted maximum likelihood (REML) estimator under a misspecified linear mixed model (LMM) that has received much attention in recent gnome-wide association studies. The asymptotic analysis establishes consistency…

统计理论 · 数学 2014-04-10 Jiming Jiang , Cong Li , Debashis Paul , Can Yang , Hongyu Zhao

We consider optimal stopping problems, in which a sequence of independent random variables is drawn from a known continuous density. The objective of such problems is to find a procedure which maximizes the expected reward; this is often…

概率论 · 数学 2020-12-07 Hugh Entwistle , Christopher Lustri , Georgy Sofronov

In many machine learning applications, one needs to interactively select a sequence of items (e.g., recommending movies based on a user's feedback) or make sequential decisions in a certain order (e.g., guiding an agent through a series of…

机器学习 · 计算机科学 2019-06-21 Marko Mitrovic , Ehsan Kazemi , Moran Feldman , Andreas Krause , Amin Karbasi

This paper studies a 2-class, 2-server parallel server system under the recently introduced extended heavy traffic condition, which states that the underlying 'static allocation' linear program (LP) is critical, but does not require that it…

最优化与控制 · 数学 2022-07-19 Rami Atar , Eyal Castiel , Marty Reiman

Recent literature on policy learning has primarily focused on regret bounds of the learned policy. We provide a new perspective by developing a unified semiparametric efficiency framework for policy learning, allowing for general treatments…

计量经济学 · 经济学 2026-02-10 Yue Fang , Geert Ridder , Haitian Xie

Empirical evidence shows that wealthy households have substantially higher saving rates and markedly lower marginal propensity to consume (MPC) than other groups. Existing theory cannot account for this pattern unless under restrictive…

理论经济学 · 经济学 2026-01-21 Qingyin Ma , Xinxi Song , Alexis Akira Toda

Consider a network of agents that all want to guess the correct value of some ground truth state. In a sequential order, each agent makes its decision using a single private signal which has a constant probability of error, as well as…

社会与信息网络 · 计算机科学 2024-10-08 Kevin Lu , Jordan Chong , Matt Lu , Jie Gao

We consider a multi-hypothesis testing problem involving a K-armed bandit. Each arm's signal follows a distribution from a vector exponential family. The actual parameters of the arms are unknown to the decision maker. The decision maker…

信息论 · 计算机科学 2022-06-13 Gayathri R Prabhu , Srikrishna Bhashyam , Aditya Gopalan , Rajesh Sundaresan

Let $X_n,...,X_1$ be i.i.d. random variables with distribution function $F$. A statistician, knowing $F$, observes the $X$ values sequentially and is given two chances to choose $X$'s using stopping rules. The statistician's goal is to stop…

概率论 · 数学 2007-06-13 David Assaf , Larry Goldstein , Ester Samuel-Cahn

Data-driven decision making plays an important role even in high stakes settings like medicine and public policy. Learning optimal policies from observed data requires a careful formulation of the utility function whose expected value is…

机器学习 · 统计学 2023-11-29 Eli Ben-Michael , Kosuke Imai , Zhichao Jiang

We study the problem of active nonparametric sequential two-sample testing over multiple heterogeneous data sources. In each time slot, a decision-maker adaptively selects one of $K$ data sources and receives a paired sample generated from…

统计理论 · 数学 2025-12-30 Chia-Yu Hsu , Shubhanshu Shekhar

This work considers the sample and computational complexity of obtaining an $\epsilon$-optimal policy in a discounted Markov Decision Process (MDP), given only access to a generative model. In this work, we study the effectiveness of the…

机器学习 · 计算机科学 2020-04-07 Alekh Agarwal , Sham Kakade , Lin F. Yang

This paper considers the uplink of a distributed Massive MIMO network where $N$ base stations (BSs), each equipped with $M$ antennas, receive data from $K=2$ users. We study the asymptotic spectral efficiency (as $M\to \infty$) with spatial…

信息论 · 计算机科学 2018-11-09 Luca Sanguinetti , Emil Bjornson , Jakob Hoydis

Despite empirical success, the theory of reinforcement learning (RL) with value function approximation remains fundamentally incomplete. Prior work has identified a variety of pathological behaviours that arise in RL algorithms that combine…

机器学习 · 计算机科学 2020-10-30 Kenny Young , Richard S. Sutton

In this paper we investigate the asymptotic optimality property of a randomized sampling based motion planner, namely RRT. We prove that a RRT planner is not an asymptotically optimal motion planner. Our result, while being consistent with…

机器人学 · 计算机科学 2017-07-14 Titas Bera , Debasish Ghose , Sundaram Suresh

We consider a queueing system composed of a dispatcher that routes deterministically jobs to a set of non-observable queues working in parallel. In this setting, the fundamental problem is which policy should the dispatcher implement to…

性能 · 计算机科学 2025-02-23 Jonatha Anselmi , Bruno Gaujal , Tommaso Nesti