中文
相关论文

相关论文: Bayesian Inverse Reinforcement Learning for Collec…

200 篇论文

This paper investigates how a Bayesian reinforcement learning method can be used to create a tactical decision-making agent for autonomous driving in an intersection scenario, where the agent can estimate the confidence of its recommended…

机器人学 · 计算机科学 2020-11-04 Carl-Johan Hoel , Tommy Tram , Jonas Sjöberg

We propose a probabilistic modeling framework for learning the dynamic patterns in the collective behaviors of social agents and developing profiles for different behavioral groups, using data collected from multiple information sources.…

机器学习 · 统计学 2016-06-28 Lin Li , Ananthram Swami , Anna Scaglione

In this paper we propose a novel gradient algorithm to learn a policy from an expert's observed behavior assuming that the expert behaves optimally with respect to some unknown reward function of a Markovian Decision Problem. The…

机器学习 · 计算机科学 2012-06-26 Gergely Neu , Csaba Szepesvari

Inverse Reinforcement Learning (IRL) describes the problem of learning an unknown reward function of a Markov Decision Process (MDP) from observed behavior of an agent. Since the agent's behavior originates in its policy and MDP policies…

人工智能 · 计算机科学 2016-04-14 Michael Herman , Tobias Gindele , Jörg Wagner , Felix Schmitt , Wolfram Burgard

Computational level explanations based on optimal feedback control with signal-dependent noise have been able to account for a vast array of phenomena in human sensorimotor behavior. However, commonly a cost function needs to be assumed for…

机器学习 · 计算机科学 2021-10-22 Matthias Schultheis , Dominik Straub , Constantin A. Rothkopf

A diversity of decision-making systems has been observed in animal collectives. In some species, choices depend on the differences of the numbers of animals that have chosen each of the available options, while in other species on the…

种群与进化 · 定量生物学 2012-12-13 Sara Arganda , Alfonso Pérez-Escudero , Gonzalo G. de Polavieja

Training a high-dimensional simulated agent with an under-specified reward function often leads the agent to learn physically infeasible strategies that are ineffective when deployed in the real world. To mitigate these unnatural behaviors,…

人工智能 · 计算机科学 2022-03-30 Alejandro Escontrela , Xue Bin Peng , Wenhao Yu , Tingnan Zhang , Atil Iscen , Ken Goldberg , Pieter Abbeel

Classical Bayesian persuasion studies how a sender influences receivers through carefully designed signaling policies within a single strategic interaction. In many real-world environments, such interactions are repeated across multiple…

计算机科学与博弈论 · 计算机科学 2026-03-24 Ata Poyraz Turna , Asrin Efe Yorulmaz , Tamer Başar

In the realm of the cooperative control of multi-agent systems (MASs) with unknown dynamics, Gaussian process (GP) regression is widely used to infer the uncertainties due to its modeling flexibility of nonlinear functions and the existence…

系统与控制 · 电气工程与系统科学 2024-09-18 Xiaobing Dai , Zewen Yang , Sihua Zhang , Di-Hua Zhai , Yuanqing Xia , Sandra Hirche

A key requirement for the current generation of artificial decision-makers is that they should adapt well to changes in unexpected situations. This paper addresses the situation in which an AI for aerial dog fighting, with tunable…

机器学习 · 统计学 2016-12-14 Brett W. Israelsen , Nisar Ahmed , Kenneth Center , Roderick Green , Winston Bennett

Reinforcement learning (RL) allows to solve complex tasks such as Go often with a stronger performance than humans. However, the learned behaviors are usually fixed to specific tasks and unable to adapt to different contexts. Here we…

机器学习 · 计算机科学 2020-04-21 Chris Reinke

Adaptive Informative Path Planning with Multimodal Sensing (AIPPMS) considers the problem of an agent equipped with multiple sensors, each with different sensing accuracy and energy costs. The agent's goal is to explore the environment and…

人工智能 · 计算机科学 2022-09-19 Joshua Ott , Edward Balaban , Mykel J. Kochenderfer

There is increasing interest to develop Bayesian inferential algorithms for point process models with intractable likelihoods. A purpose of this paper is to illustrate the utility of using simulation based strategies, including Approximate…

统计计算 · 统计学 2026-02-02 Chaoyi Lu , Nial Friel

We consider a problem of learning the reward and policy from expert examples under unknown dynamics. Our proposed method builds on the framework of generative adversarial networks and introduces the empowerment-regularized maximum-entropy…

机器学习 · 计算机科学 2019-02-26 Ahmed H. Qureshi , Byron Boots , Michael C. Yip

Contextual policy search allows adapting robotic movement primitives to different situations. For instance, a locomotion primitive might be adapted to different terrain inclinations or desired walking speeds. Such an adaptation is often…

机器学习 · 统计学 2015-11-17 Jan Hendrik Metzen

Collective migration of animals in a cohesive group is rendered possible by a strategic distribution of tasks among members: some track the travel route, which is time and energy-consuming, while the others follow the group by interacting…

最优化与控制 · 数学 2015-08-05 Benedetto Piccoli , Nastassia Pouradier Duteil , Benjamin Scharf

We propose a bio-inspired, agent-based approach to describe the natural phenomenon of group chasing in both two and three dimensions. Using a set of local interaction rules we created a continuous-space and discrete-time model with time…

生物物理 · 物理学 2017-05-24 Milán Janosov , Csaba Virágh , Gábor Vásárhelyi , Tamás Vicsek

We propose a novel actor-critic, model-free reinforcement learning algorithm which employs a Bayesian method of parameter space exploration to solve environments. A Gaussian process is used to learn the expected return of a policy given the…

机器学习 · 计算机科学 2020-03-03 Ashish Rao , Bidipta Sarkar , Tejas Narayanan

Robotic agents must adopt existing social conventions in order to be effective teammates. These social conventions, such as driving on the right or left side of the road, are arbitrary choices among optimal policies, but all agents on a…

人工智能 · 计算机科学 2020-10-09 Mycal Tucker , Yilun Zhou , Julie Shah

Collective motion is ubiquitous in nature; groups of animals, such as fish, birds, and ungulates appear to move as a whole, exhibiting a rich behavioral repertoire that ranges from directed movement to milling to disordered swarming.…

适应与自组织系统 · 物理学 2024-05-15 Conor Heins , Beren Millidge , Lancelot da Costa , Richard Mann , Karl Friston , Iain Couzin