English
Related papers

Related papers: The speed of sequential asymptotic learning

200 papers

We consider sequential selection of an alternating subsequence from a sequence of independent, identically distributed, continuous random variables, and we determine the exact asymptotic behavior of an optimal sequentially selected…

Probability · Mathematics 2011-08-15 Alessandro Arlotto , Robert W. Chen , Lawrence A. Shepp , J. Michael Steele

We study non-Bayesian social learning on random directed graphs and show that under mild connectivity assumptions, all the agents almost surely learn the true state of the world asymptotically in time if the sequence of the associated…

Optimization and Control · Mathematics 2021-08-02 Rohit Parasnis , Massimo Franceschetti , Behrouz Touri

The problem of detecting the presence of a signal that can lead to a disaster is studied. A decision-maker collects data sequentially over time. At some point in time, called the change point, the distribution of data changes. This change…

Signal Processing · Electrical Eng. & Systems 2023-03-07 Tim Brucks , Taposh Banerjee , Rahul Mishra

Learning in games has been widely used to solve many cooperative multi-agent problems such as coverage control, consensus, self-reconfiguration or vehicle-target assignment. One standard approach in this domain is to formulate the problem…

Systems and Control · Electrical Eng. & Systems 2022-09-07 Abbasali Koochakzadeh , Yasin Yazıcıoğlu

Learning how to learn efficiently is a fundamental challenge for biological agents and a growing concern for artificial ones. To learn effectively, an agent must regulate its learning speed, balancing the benefits of rapid improvement…

Machine Learning · Computer Science 2026-01-13 Valentina Njaradi , Rodrigo Carrasco-Davis , Peter E. Latham , Andrew Saxe

Use-dependent bias is a phenomenon in human sensorimotor behavior whereby movements become biased towards previously repeated actions. Despite being well-documented, the reason why this phenomenon occurs is not yet clearly understood. Here,…

Neurons and Cognition · Quantitative Biology 2024-08-19 Hokin Deng , Adrian Haith

We study the convergence speed of distributed iterative algorithms for the consensus and averaging problems, with emphasis on the latter. We first consider the case of a fixed communication topology. We show that a simple adaptation of a…

Optimization and Control · Mathematics 2011-06-13 Alex Olshevsky , John N. Tsitsiklis

This work considers the problem of detecting signals from multiple sequentially observed data streams, where only one stream can be observed at every time instant. The goal is to detect signals as quickly as possible while controlling the…

Methodology · Statistics 2026-04-07 Yiming Xing , Georgios Fellouris

To what extent can an external observer observing an equilibrium action distribution in an incomplete information game infer the underlying information structure? We investigate this issue in a general linear-quadratic-Gaussian framework. A…

Theoretical Economics · Economics 2024-03-19 Masaki Miyashita

In the human brain, internal states are often correlated over time (due to local recurrence and other intrinsic circuit properties), punctuated by abrupt transitions. At first glance, temporal smoothness of internal states presents a…

Machine Learning · Computer Science 2023-05-24 Shima Rahimi Moghaddam , Fanjun Bu , Christopher J. Honey

The effects of policy sharing between agents in a multi-agent dynamical system has not been studied extensively. I simulate a system of agents optimizing the same task using reinforcement learning, to study the effects of different…

Multiagent Systems · Computer Science 2008-12-10 Jake Ellowitz

We study Bayesian coordination games where agents receive noisy private information over the game's payoff structure, and over each others' actions. If private information over actions is precise, we find that agents can coordinate on…

General Economics · Economics 2019-04-25 Dominik Grafenhofer , Wolfgang Kuhle

I study a social learning model in which the object to learn is a strategic player's endogenous actions rather than an exogenous state. A patient seller faces a sequence of buyers and decides whether to build a reputation for supplying high…

Theoretical Economics · Economics 2020-11-03 Harry Pei

Humans can learn several tasks in succession with minimal mutual interference but perform more poorly when trained on multiple tasks at once. The opposite is true for standard deep neural networks. Here, we propose novel computational…

Neurons and Cognition · Quantitative Biology 2022-09-07 Timo Flesch , David G. Nagy , Andrew Saxe , Christopher Summerfield

Selective attention allows to process stimuli which are behaviorally relevant, while attenuating distracting information. However, it is an open question what mechanisms implement selective routing, and how they are engaged in dependence on…

Neurons and Cognition · Quantitative Biology 2023-05-24 Maik Schünemann , Udo Ernst

I study endogenous learning dynamics for people who misperceive intertemporal correlations in random sequences. Biased agents face an optimal-stopping problem. They are uncertain about the underlying distribution and learn its parameters…

Economics · Quantitative Finance 2022-11-15 Kevin He

People often interact repeatedly: with relatives, through file sharing, in politics, etc. Many such interactions are reciprocal: reacting to the actions of the other. In order to facilitate decisions regarding reciprocal interactions, we…

Computer Science and Game Theory · Computer Science 2016-03-01 Gleb Polevoy , Mathijs de Weerdt , Catholijn Jonker

We formalize trust calibration for agentic tool use (deciding when an automated agent's proposed action may execute autonomously versus require human approval) as a preference-learning problem. A policy gateway maintains a Gaussian-process…

Artificial Intelligence · Computer Science 2026-05-20 Changkun Ou

Traditional models of active learning assume a learner can directly manipulate or query a covariate $X$ in order to study its relationship with a response $Y$. However, if $X$ is a feature of a complex system, it may be possible only to…

Statistics Theory · Mathematics 2023-01-24 Shashank Singh

Learning from small data sets is critical in many practical applications where data collection is time consuming or expensive, e.g., robotics, animal experiments or drug design. Meta learning is one way to increase the data efficiency of…

Machine Learning · Statistics 2018-07-10 Steindór Sæmundsson , Katja Hofmann , Marc Peter Deisenroth