中文
相关论文

相关论文: 'Indifference' methods for managing agent rewards

200 篇论文

Active inference proposes expected free energy as an objective for planning and decision-making to adequately balance exploitative and explorative drives in learning agents. The exploitative drive, or what an agent wants to achieve, is…

人工智能 · 计算机科学 2025-12-04 Filippo Torresan , Ryota Kanai , Manuel Baltieri

Rewards typically express desirabilities or preferences over a set of alternatives. Here we propose that rewards can be defined for any probability distribution based on three desiderata, namely that rewards should be real-valued, additive…

人工智能 · 计算机科学 2009-12-31 Pedro A. Ortega , Daniel A. Braun

In practice, most mechanisms for selling, buying, matching, voting, and so on are not incentive compatible. We present techniques for estimating how far a mechanism is from incentive compatible. Given samples from the agents' type…

计算机科学与博弈论 · 计算机科学 2023-12-12 Maria-Florina Balcan , Tuomas Sandholm , Ellen Vitercik

One of the main research areas in Artificial Intelligence is the coding of agents (programs) which are able to learn by themselves in any situation. This means that agents must be useful for purposes other than those they were created for,…

人工智能 · 计算机科学 2011-02-04 Javier Insa-Cabrera , Jose Hernandez-Orallo

Financial institutions mostly deal with people. Therefore, characterizing different kinds of human behavior can greatly help institutions for improving their relation with customers and with regulatory offices. In many of such interactions,…

人工智能 · 计算机科学 2020-11-06 Daniel Borrajo , Manuela Veloso

Imitation is a key component of human social behavior, and is widely used by both children and adults as a way to navigate uncertain or unfamiliar situations. But in an environment populated by multiple heterogeneous agents pursuing…

神经元与认知 · 定量生物学 2023-05-15 Max Taylor-Davies , Stephanie Droop , Christopher G. Lucas

This work focuses on the indifference pricing of American call option underlying a non-traded stock, which may be partially hedgeable by another traded stock. Under the exponential forward measure, the indifference price is formulated as a…

证券定价 · 定量金融 2012-01-04 Xiaoshan Chen , Qingshuo Song , Fahuai Yi , George Yin

Participants in socio-economic systems are often ranked based on their performance. Rankings conveniently reduce the complexity of such systems to ordered lists. Yet, it has been shown in many contexts that those who reach the top are not…

物理与社会 · 物理学 2024-01-30 Federica De Domenico , Fabio Caccioli , Giacomo Livan , Guido Montagna , Oreste Nicrosini

Modeling multi-agent systems requires understanding how agents interact. Such systems are often difficult to model because they can involve a variety of types of interactions that layer together to drive rich social behavioral dynamics.…

机器学习 · 计算机科学 2023-01-26 Fan-Yun Sun , Isaac Kauvar , Ruohan Zhang , Jiachen Li , Mykel Kochenderfer , Jiajun Wu , Nick Haber

Intrinsically motivated reinforcement learning aims to address the exploration challenge for sparse-reward tasks. However, the study of exploration methods in transition-dependent multi-agent settings is largely absent from the literature.…

机器学习 · 计算机科学 2019-12-30 Tonghan Wang , Jianhao Wang , Yi Wu , Chongjie Zhang

We propose an incentive scheme based on intervention to sustain cooperation among self-interested users. In the proposed scheme, an intervention device collects imperfect signals about the actions of the users for a test period, and then…

计算机科学与博弈论 · 计算机科学 2010-12-09 Jaeok Park , Mihaela van der Schaar

Inverse Reinforcement Learning (IRL) describes the problem of learning an unknown reward function of a Markov Decision Process (MDP) from observed behavior of an agent. Since the agent's behavior originates in its policy and MDP policies…

人工智能 · 计算机科学 2016-04-14 Michael Herman , Tobias Gindele , Jörg Wagner , Felix Schmitt , Wolfram Burgard

Differences-in-differences (DiD) is a causal inference method for observational longitudinal data that assumes parallel expected potential outcome trajectories between treatment groups under the counterfactual scenario where all units…

统计方法学 · 统计学 2026-05-12 Michael Jetsupphasuk , Didong Li , Michael G. Hudgens

Modelling other agents' behaviors plays an important role in decision models for interactions among multiple agents. To optimise its own decisions, a subject agent needs to model what other agents act simultaneously in an uncertain…

人工智能 · 计算机科学 2022-03-08 Yinghui Pan , Hanyi Zhang , Yifeng Zeng , Biyang Ma , Jing Tang , Zhong Ming

A reinforcement learning agent that needs to pursue different goals across episodes requires a goal-conditional policy. In addition to their potential to generalize desirable behavior to unseen goals, such policies may also enable…

机器学习 · 计算机科学 2019-02-21 Paulo Rauber , Avinash Ummadisingu , Filipe Mutz , Juergen Schmidhuber

To operate reliably under changing conditions, complex systems require feedback on how effectively they use resources, not just whether objectives are met. Current AI systems process vast information to produce sophisticated predictions,…

人工智能 · 计算机科学 2026-03-10 Wael Hafez , Chenan Wei , Rodrigo Pena , Amir Nazeri , Cameron Reid

The problem of controlling multi-agent systems under different models of information sharing among agents has received significant attention in the recent literature. In this paper, we consider a setup where rather than committing to a…

最优化与控制 · 数学 2021-04-23 Sagar Sudhakara , Dhruva Kartik , Rahul Jain , Ashutosh Nayyar

Reinforcement learning agents learn by encouraging behaviours which maximize their total reward, usually provided by the environment. In many environments, however, the reward is provided after a series of actions rather than each single…

人工智能 · 计算机科学 2022-01-04 Mohammad Reza Bonyadi , Rui Wang , Maryam Ziaei

Active inference is an ambitious theory that treats perception, inference and action selection of autonomous agents under the heading of a single principle. It suggests biologically plausible explanations for many cognitive phenomena,…

人工智能 · 计算机科学 2018-06-22 Martin Biehl , Christian Guckelsberger , Christoph Salge , Simón C. Smith , Daniel Polani

Federated learning promises significant sample-efficiency gains by pooling data across multiple agents, yet incentive misalignment is an obstacle: each update is costly to the contributor but boosts every participant. We introduce a…

计算机科学与博弈论 · 计算机科学 2026-02-02 Ariel D. Procaccia , Han Shao , Itai Shapira
‹ 上一页 1 8 9 10 下一页 ›