中文
相关论文

相关论文: From Agreement to Asymptotic Learning

200 篇论文

Artificial general intelligence aims to create agents capable of learning to solve arbitrary interesting problems. We define two versions of asymptotic optimality and prove that no agent can satisfy the strong version while in some cases,…

人工智能 · 计算机科学 2012-02-10 Tor Lattimore , Marcus Hutter

Suppose two Bayesian agents each learn a generative model of the same environment. We will assume the two have converged on the predictive distribution, i.e. distribution over some observables in the environment, but may have different…

概率论 · 数学 2025-09-05 John Wentworth , David Lorell

Agents interacting with an incompletely known world need to be able to reason about the effects of their actions, and to gain further information about that world they need to use sensors of some sort. Unfortunately, both the effects of…

人工智能 · 计算机科学 2007-05-23 Fahiem Bacchus , Joseph Y. Halpern , Hector J. Levesque

We consider learning by fictitious play in a large population of agents engaged in single-play, two-person rounds of a symmetric game, and derive a mean-filed type model for the corresponding stochastic process. Using this model, we…

计算机科学与博弈论 · 计算机科学 2019-01-11 Misha Perepelitsa

We demonstrate that a wide array of machine learning algorithms are specific instances of one single paradigm: reciprocal learning. These instances range from active learning over multi-armed bandits to self-training. We show that all these…

机器学习 · 统计学 2024-11-05 Julian Rodemann , Christoph Jansen , Georg Schollmeyer

We develop a learning-based algorithm for the distributed formation control of networked multi-agent systems governed by unknown, nonlinear dynamics. Most existing algorithms either assume certain parametric forms for the unknown dynamic…

系统与控制 · 电气工程与系统科学 2022-01-13 Christos K. Verginis , Zhe Xu , Ufuk Topcu

We consider the setting where a collection of time series, modeled as random processes, evolve in a causal manner, and one is interested in learning the graph governing the relationships of these processes. A special case of wide interest…

机器学习 · 计算机科学 2016-08-30 Hossein Hosseini , Sreeram Kannan , Baosen Zhang , Radha Poovendran

In this article, we work towards the goal of developing agents that can learn to act in complex worlds. We develop a probabilistic, relational planning rule representation that compactly models noisy, nondeterministic action effects, and…

机器学习 · 计算机科学 2011-10-12 L. P. Kaelbling , H. M. Pasula , L. S. Zettlemoyer

We propose the use of Bayesian networks, which provide both a mean value and an uncertainty estimate as output, to enhance the safety of learned control policies under circumstances in which a test-time input differs significantly from the…

机器学习 · 计算机科学 2019-02-18 Keuntaek Lee , Kamil Saigol , Evangelos A. Theodorou

The technology for autonomous vehicles is close to replacing human drivers by artificial systems endowed with high-level decision-making capabilities. In this regard, systems must learn about the usual vehicle's behavior to predict imminent…

图像与视频处理 · 电气工程与系统科学 2020-04-22 Mahdyar Ravanbakhsh , Mohamad Baydoun , Damian Campo , Pablo Marin , David Martin , Lucio Marcenaro , andCarlo Regazzoni

Success-driven social learning, in which individuals preferentially adopt the ideas and methods that appear most successful, is a foundational principle of collective behavior across systems ranging from ant colonies to scientific…

物理与社会 · 物理学 2026-05-01 Avery W. Louis , Marina Dubova

We apply recent advances in deep generative modeling to the task of imitation learning from biological agents. Specifically, we apply variations of the variational recurrent neural network model to a multi-agent setting where we learn…

机器学习 · 计算机科学 2020-07-02 Michael Teng , Tuan Anh Le , Adam Scibior , Frank Wood

Consider the finite state graph that results from a simple, discrete, dynamical system in which an agent moves in a rectangular grid picking up and dropping packages. Can the state variables of the problem, namely, the agent location and…

人工智能 · 计算机科学 2022-07-13 Blai Bonet , Hector Geffner

What can be learned about causality and experimentation from passive data? This question is salient given recent successes of passively-trained language models in interactive domains such as tool use. Passive learning is inherently limited.…

机器学习 · 计算机科学 2023-10-03 Andrew Kyle Lampinen , Stephanie C Y Chan , Ishita Dasgupta , Andrew J Nam , Jane X Wang

Non-Bayesian social learning is a framework for distributed hypothesis testing aimed at learning the true state of the environment. Traditionally, the agents are assumed to receive observations conditioned on the same true state, although…

社会与信息网络 · 计算机科学 2024-06-26 Valentina Shumovskaia , Mert Kayaalp , Ali H. Sayed

We study continuous-time consensus dynamics for multi-agent systems with undirected switching interaction graphs. We establish a necessary and sufficient condition for exponential asymptotic consensus based on the classical theory of…

系统与控制 · 计算机科学 2016-08-10 Brian D. O. Anderson , Guodong Shi , Jochen Trumpf

We present a novel Bayesian approach to semiotic dynamics, which is a cognitive analogue of the naming game model restricted to two conventions. The one-shot learning that characterizes the agent dynamics in the basic naming game is…

物理与社会 · 物理学 2020-06-30 Gionni Marchetti , Marco Patriarca , Els Heinsalu

Learning to cooperate with other agents is challenging when those agents also possess the ability to adapt to our own behavior. Practical and theoretical approaches to learning in cooperative settings typically assume that other agents'…

计算机科学与博弈论 · 计算机科学 2022-11-29 Robert Loftin , Frans A. Oliehoek

We study Bayesian learning in episodic, finite-horizon zero-sum Markov games with unknown transition and reward models. We investigate a posterior algorithm in which each player maintains a Bayesian posterior over the game model,…

机器学习 · 计算机科学 2026-03-24 Chang-Wei Yueh , Andy Zhao , Ashutosh Nayyar , Rahul Jain

Inferring the laws of interaction between particles and agents in complex dynamical systems from observational data is a fundamental challenge in a wide variety of disciplines. We propose a non-parametric statistical learning approach to…

机器学习 · 计算机科学 2022-06-08 Fei Lu , Mauro Maggioni , Sui Tang , Ming Zhong