中文
相关论文

相关论文: From Agreement to Asymptotic Learning

200 篇论文

We explore unconstrained natural language feedback as a learning signal for artificial agents. Humans use rich and varied language to teach, yet most prior work on interactive learning from language assumes a particular form of input (e.g.,…

人工智能 · 计算机科学 2021-07-06 Theodore R. Sumers , Mark K. Ho , Robert D. Hawkins , Karthik Narasimhan , Thomas L. Griffiths

Autonomous and learning agents increasingly participate in markets - setting prices, placing bids, ordering inventory. Such agents are not just aiming to optimize in an uncertain environment; they are making decisions in a game-theoretical…

计算机科学与博弈论 · 计算机科学 2025-06-24 Martin Bichler , Julius Durmann , Matthias Oberlechner

We study a dynamic model of Bayesian persuasion in sequential decision-making settings. An informed principal observes an external parameter of the world and advises an uninformed agent about actions to take over time. The agent takes…

计算机科学与博弈论 · 计算机科学 2022-05-25 Jiarui Gan , Rupak Majumdar , Goran Radanovic , Adish Singla

We study a symmetric collaborative dialogue setting in which two agents, each with private knowledge, must strategically communicate to achieve a common goal. The open-ended dialogue state in this setting poses new challenges for existing…

计算与语言 · 计算机科学 2017-04-25 He He , Anusha Balakrishnan , Mihail Eric , Percy Liang

We study how a consensus emerges in a finite population of like-minded individuals who are asymmetrically informed about the realization of the true state of the world. Agents observe a private signal about the state and then start…

理论经济学 · 经济学 2022-02-14 Michele Crescenzi

We study a simple learning model based on the Hebb rule to cope with "delayed", unspecific reinforcement. In spite of the unspecific nature of the information-feedback, convergence to asymptotically perfect generalization is observed, with…

统计力学 · 物理学 2009-10-31 Reimer Kuehn , Ion-Olimpiu Stamatescu

A universal feature of human societies is the adoption of systems of rules and norms in the service of cooperative ends. How can we build learning agents that do the same, so that they may flexibly cooperate with the human institutions they…

人工智能 · 计算机科学 2024-02-23 Ninell Oldenburg , Tan Zhi-Xuan

An AI agent might surprisingly find she has reached an unknown state which she has never been aware of -- an unknown unknown. We mathematically ground this scenario in reinforcement learning: an agent, after taking an action calculated from…

机器学习 · 计算机科学 2025-09-04 Juntian Zhu , Miguel de Carvalho , Zhouwang Yang , Fengxiang He

People increasingly use large language models (LLMs) to explore ideas, gather information, and make sense of the world. In these interactions, they encounter agents that are overly agreeable. We argue that this sycophancy poses a unique…

计算机与社会 · 计算机科学 2026-02-17 Rafael M. Batista , Thomas L. Griffiths

In this paper we consider a network scenario in which agents can evaluate each other according to a score graph that models some physical or social interaction. The goal is to design a distributed protocol, run by the agents, allowing them…

最优化与控制 · 数学 2017-06-14 Francesco Sasso , Angelo Coluccia , Giuseppe Notarstefano

In the classical Bayesian persuasion model an informed player and an uninformed one engage in a static interaction. The informed player, the sender, knows the state of nature, while the uninformed one, the receiver, does not. The informed…

理论经济学 · 经济学 2025-12-16 Ehud Lehrer , Dimitry Shaiderman

We consider the inverse reinforcement learning problem, that is, the problem of learning from, and then predicting or mimicking a controller based on state/action data. We propose a statistical model for such data, derived from the…

机器学习 · 统计学 2012-11-27 Sumeetpal S. Singh , Nicolas Chopin , Nick Whiteley

Bayesian persuasion, a central model in information design, studies how a sender, who privately observes a state drawn from a prior distribution, strategically sends a signal to influence a receiver's action. A key assumption is that both…

计算机科学与博弈论 · 计算机科学 2025-05-23 Jingwu Tang , Jiahao Zhang , Fei Fang , Zhiwei Steven Wu

Continual learning is often motivated by the idea, known as the big world hypothesis, that "the world is bigger" than the agent. Recent problem formulations capture this idea by explicitly constraining an agent relative to the environment.…

人工智能 · 计算机科学 2025-12-30 Alex Lewandowski , Adtiya A. Ramesh , Edan Meyer , Dale Schuurmans , Marlos C. Machado

We study the following repeated non-atomic routing game. In every round, nature chooses a state in an i.i.d. manner according to a publicly known distribution, which influences link latency functions. The system planner makes private route…

系统与控制 · 电气工程与系统科学 2022-07-26 Yixian Zhu , Ketan Savla

Most of the works on planning and learning, e.g., planning by (model based) reinforcement learning, are based on two main assumptions: (i) the set of states of the planning domain is fixed; (ii) the mapping between the observations from the…

人工智能 · 计算机科学 2018-11-27 Luciano Serafini , Paolo Traverso

We consider a network scenario in which agents can evaluate each other according to a score graph that models some interactions. The goal is to design a distributed protocol, run by the agents, that allows them to learn their unknown state…

系统与控制 · 计算机科学 2018-06-05 Francesco Sasso , Angelo Coluccia , Giuseppe Notarstefano

Methods for learning optimal policies in autonomous agents often assume that the way the domain is conceptualised---its possible states and actions and their causal structure---is known in advance and does not change during learning. This…

人工智能 · 计算机科学 2018-01-11 Craig Innes , Alex Lascarides , Stefano V Albrecht , Subramanian Ramamoorthy , Benjamin Rosman

In the classical herding model, asymptotic learning refers to situations where individuals eventually take the correct action regardless of their private information. Classical results identify classes of information structures for which…

计算机科学与博弈论 · 计算机科学 2020-02-14 Itay Kavaler

Contemporary focus on selective inference has renewed interest in the theory of selection models. In this paper, we analyze the asymptotic properties of selection models built on independent and identically distributed observations. We show…

统计理论 · 数学 2026-03-16 Daniel G. Rasines , G. Alastair Young