English
Related papers

Related papers: Reading policies for joins: An asymptotic analysis

200 papers

Two main procedures characterize the way in which social actors evaluate the qualities of the options in decision-making processes: they either seek to evaluate their intrinsic qualities (individual learners), or they rely on the opinion of…

Physics and Society · Physics 2024-07-31 Arkadiusz Jędrzejewski , Laura Hernández

Empirical research shows that individuals' responses to treatments vary along latent characteristics, such as innate ability or motivation. Therefore, a policymaker seeking to maximize welfare may consider designing policies based on…

Econometrics · Economics 2026-05-06 Giacomo Opocher

The purpose of this article is to examine the greedy adaptive measurement policy in the context of a linear Guassian measurement model with an optimization criterion based on information gain. In the special case of sequential scalar…

Information Theory · Computer Science 2012-08-20 Entao Liu , Edwin K. P. Chong , Louis L. Scharf

In "Recognizing the Maximum of a Sequence", Gilbert and Mosteller analyze a full information game where n measurements from an uniform distribution are drawn and a player (knowing n) must decide at each draw whether or not to choose that…

Probability · Mathematics 2018-05-30 Marcos Costa Santos Carreira

We study behavior of the restricted maximum likelihood (REML) estimator under a misspecified linear mixed model (LMM) that has received much attention in recent gnome-wide association studies. The asymptotic analysis establishes consistency…

Statistics Theory · Mathematics 2014-04-10 Jiming Jiang , Cong Li , Debashis Paul , Can Yang , Hongyu Zhao

We consider optimal stopping problems, in which a sequence of independent random variables is drawn from a known continuous density. The objective of such problems is to find a procedure which maximizes the expected reward; this is often…

Probability · Mathematics 2020-12-07 Hugh Entwistle , Christopher Lustri , Georgy Sofronov

In many machine learning applications, one needs to interactively select a sequence of items (e.g., recommending movies based on a user's feedback) or make sequential decisions in a certain order (e.g., guiding an agent through a series of…

Machine Learning · Computer Science 2019-06-21 Marko Mitrovic , Ehsan Kazemi , Moran Feldman , Andreas Krause , Amin Karbasi

This paper studies a 2-class, 2-server parallel server system under the recently introduced extended heavy traffic condition, which states that the underlying 'static allocation' linear program (LP) is critical, but does not require that it…

Optimization and Control · Mathematics 2022-07-19 Rami Atar , Eyal Castiel , Marty Reiman

Recent literature on policy learning has primarily focused on regret bounds of the learned policy. We provide a new perspective by developing a unified semiparametric efficiency framework for policy learning, allowing for general treatments…

Econometrics · Economics 2026-02-10 Yue Fang , Geert Ridder , Haitian Xie

Empirical evidence shows that wealthy households have substantially higher saving rates and markedly lower marginal propensity to consume (MPC) than other groups. Existing theory cannot account for this pattern unless under restrictive…

Theoretical Economics · Economics 2026-01-21 Qingyin Ma , Xinxi Song , Alexis Akira Toda

Consider a network of agents that all want to guess the correct value of some ground truth state. In a sequential order, each agent makes its decision using a single private signal which has a constant probability of error, as well as…

Social and Information Networks · Computer Science 2024-10-08 Kevin Lu , Jordan Chong , Matt Lu , Jie Gao

We consider a multi-hypothesis testing problem involving a K-armed bandit. Each arm's signal follows a distribution from a vector exponential family. The actual parameters of the arms are unknown to the decision maker. The decision maker…

Information Theory · Computer Science 2022-06-13 Gayathri R Prabhu , Srikrishna Bhashyam , Aditya Gopalan , Rajesh Sundaresan

Let $X_n,...,X_1$ be i.i.d. random variables with distribution function $F$. A statistician, knowing $F$, observes the $X$ values sequentially and is given two chances to choose $X$'s using stopping rules. The statistician's goal is to stop…

Probability · Mathematics 2007-06-13 David Assaf , Larry Goldstein , Ester Samuel-Cahn

Data-driven decision making plays an important role even in high stakes settings like medicine and public policy. Learning optimal policies from observed data requires a careful formulation of the utility function whose expected value is…

Machine Learning · Statistics 2023-11-29 Eli Ben-Michael , Kosuke Imai , Zhichao Jiang

We study the problem of active nonparametric sequential two-sample testing over multiple heterogeneous data sources. In each time slot, a decision-maker adaptively selects one of $K$ data sources and receives a paired sample generated from…

Statistics Theory · Mathematics 2025-12-30 Chia-Yu Hsu , Shubhanshu Shekhar

This work considers the sample and computational complexity of obtaining an $\epsilon$-optimal policy in a discounted Markov Decision Process (MDP), given only access to a generative model. In this work, we study the effectiveness of the…

Machine Learning · Computer Science 2020-04-07 Alekh Agarwal , Sham Kakade , Lin F. Yang

This paper considers the uplink of a distributed Massive MIMO network where $N$ base stations (BSs), each equipped with $M$ antennas, receive data from $K=2$ users. We study the asymptotic spectral efficiency (as $M\to \infty$) with spatial…

Information Theory · Computer Science 2018-11-09 Luca Sanguinetti , Emil Bjornson , Jakob Hoydis

Despite empirical success, the theory of reinforcement learning (RL) with value function approximation remains fundamentally incomplete. Prior work has identified a variety of pathological behaviours that arise in RL algorithms that combine…

Machine Learning · Computer Science 2020-10-30 Kenny Young , Richard S. Sutton

In this paper we investigate the asymptotic optimality property of a randomized sampling based motion planner, namely RRT. We prove that a RRT planner is not an asymptotically optimal motion planner. Our result, while being consistent with…

Robotics · Computer Science 2017-07-14 Titas Bera , Debasish Ghose , Sundaram Suresh

We consider a queueing system composed of a dispatcher that routes deterministically jobs to a set of non-observable queues working in parallel. In this setting, the fundamental problem is which policy should the dispatcher implement to…

Performance · Computer Science 2025-02-23 Jonatha Anselmi , Bruno Gaujal , Tommaso Nesti