中文
相关论文

相关论文: Sequential Naive Learning

200 篇论文

Our main goal is to study a class of processes whose increments are generated via a cellular automata rule. Given the increments of a simple biased random walk, a new sequence of (dependent) Bernoulli random variables is produced. It is…

概率论 · 数学 2017-10-24 Andrea Collevecchio , Kais Hamza , Yunxuan Liu

In nonstationary bandit learning problems, the decision-maker must continually gather information and adapt their action selection as the latent state of the environment evolves. In each time period, some latent optimal action maximizes…

机器学习 · 计算机科学 2023-12-27 Seungki Min , Daniel Russo

Although evidence integration to the boundary model has successfully explained a wide range of behavioral and neural data in decision making under uncertainty, how animals learn and optimize the boundary remains unresolved. Here, we propose…

神经与进化计算 · 计算机科学 2024-08-13 Jamal Esmaily , Rani Moran , Yasser Roudi , Bahador Bahrami

We study the problem of learning a good set of policies, so that when combined together, they can solve a wide variety of unseen reinforcement learning tasks with no or very little new data. Specifically, we consider the framework of…

机器学习 · 计算机科学 2022-03-16 Safa Alver , Doina Precup

Strategic classification studies the problem where self-interested individuals or agents manipulate their response to obtain favorable decision outcomes made by classifiers, typically turning to dishonest actions when they are less costly…

机器学习 · 计算机科学 2026-05-07 Ziyuan Huang , Lina Alkarmi , Mingyan Liu

An evolving population, in which individual members (`agents') adapt their behaviour according to past experience, is of central importance to many disciplines. Because of their limited knowledge and capabilities, agents are forced to make…

凝聚态物理 · 物理学 2009-10-31 Neil F. Johnson , Pak Ming Hui , Rob Jonson , Ting Shek Lo

In this paper we present an optimization-based view of distributed parameter estimation and observational social learning in networks. Agents receive a sequence of random, independent and identically distributed (i.i.d.) signals, each of…

机器学习 · 计算机科学 2013-09-11 Shahin Shahrampour , Ali Jadbabaie

We study the design of information acquisition games-environments where a designer contracts their action on Sender's choice of experiment and the realized signals about some state-and identify which predictions can be made absent knowledge…

理论经济学 · 经济学 2026-01-22 Eric Gao , Daniel Luo

We consider sequential hypothesis testing between two quantum states using adaptive and non-adaptive strategies. In this setting, samples of an unknown state are requested sequentially and a decision to either continue or to accept one of…

量子物理 · 物理学 2023-03-07 Yonglong Li , Vincent Y. F. Tan , Marco Tomamichel

Reinforcement learning agents tend to develop habits that are effective only under specific policies. Following an initial exploration phase where agents try out different actions, they eventually converge onto a particular policy. As this…

机器学习 · 计算机科学 2024-06-25 Miguel Suau , Matthijs T. J. Spaan , Frans A. Oliehoek

We consider a distributed learning setup where a network of agents sequentially access realizations of a set of random variables with unknown distributions. The network objective is to find a parametrized distribution that best describes…

最优化与控制 · 数学 2016-05-10 Angelia Nedić , Alex Olshevsky , César Uribe

In this paper, we propose a general framework for combining evidence of varying quality to estimate underlying binary latent variables in the presence of restrictions imposed to respect the scientific context. The resulting algorithms…

统计方法学 · 统计学 2018-08-28 Zhenke Wu , Livia Casciola-Rosen , Antony Rosen , Scott L. Zeger

As autonomous agents become increasingly sophisticated, validating their sequential behavior presents a significant challenge. Traditional testing approaches require manual specification, exact sequence matching, or thousands of training…

人工智能 · 计算机科学 2026-05-06 Reshabh K Sharma , Gaurav Mittal , Yu Hu

We consider the forecast aggregation problem in repeated settings, where the forecasts are done on a binary event. At each period multiple experts provide forecasts about an event. The goal of the aggregator is to aggregate those forecasts…

机器学习 · 计算机科学 2018-02-21 Yakov Babichenko , Dan Garber

In many machine learning applications, one needs to interactively select a sequence of items (e.g., recommending movies based on a user's feedback) or make sequential decisions in a certain order (e.g., guiding an agent through a series of…

机器学习 · 计算机科学 2019-06-21 Marko Mitrovic , Ehsan Kazemi , Moran Feldman , Andreas Krause , Amin Karbasi

While many multiagent algorithms are designed for homogeneous systems (i.e. all agents are identical), there are important applications which require an agent to coordinate its actions without knowing a priori how the other agents behave.…

人工智能 · 计算机科学 2019-07-17 Stefano V. Albrecht , Subramanian Ramamoorthy

We propose a learning dynamics to model how strategic agents repeatedly play a continuous game while relying on an information platform to learn an unknown payoff-relevant parameter. In each time step, the platform updates a belief estimate…

多智能体系统 · 计算机科学 2023-11-02 Manxi Wu , Saurabh Amin , Asuman Ozdaglar

In our previous work, we introduced the rule-based Bayesian Regression, a methodology that leverages two concepts: (i) Bayesian inference, for the general framework and uncertainty quantification and (ii) rule-based systems for the…

机器学习 · 统计学 2022-03-01 Themistoklis Botsas , Lachlan R. Mason , Omar K. Matar , Indranil Pan

A learner does not only fit data; it also determines how strongly the training sample may shape its output and how much distortion it can hedge. We study this relation as a bounded-rational decision problem whose primitive object is the…

机器学习 · 计算机科学 2026-05-18 Pedro A. Ortega

An analyst observes the frequency with which a decision maker (DM) takes actions, but not the frequency conditional on payoff-relevant states. We ask when the analyst can rationalize the DM's choices as if the DM first learns something…

理论经济学 · 经济学 2025-06-18 Laura Doval , Ran Eilat , Tianhao Liu , Yangfan Zhou