English
Related papers

Related papers: Conditional Recall

200 papers

Iterated Prisoner's Dilemma(IPD) is a well-known benchmark for studying the long term behaviors of rational agents, such as how cooperation can emerge among selfish and unrelated agents that need to co-exist over long term. Many well-known…

Computer Science and Game Theory · Computer Science 2017-12-19 Shiheng Wang , Fangzhen Lin

When subjected to automated decision-making, decision subjects may strategically modify their observable features in ways they believe will maximize their chances of receiving a favorable decision. In many practical situations, the…

Computer Science and Game Theory · Computer Science 2022-10-10 Keegan Harris , Valerie Chen , Joon Sik Kim , Ameet Talwalkar , Hoda Heidari , Zhiwei Steven Wu

Deception plays a key role in adversarial or strategic interactions for the purpose of self-defence and survival. This paper introduces a general framework and solution to address deception. Most existing approaches for deception consider…

Artificial Intelligence · Computer Science 2019-04-26 Bo Wu , Murat Cubuktepe , Suda Bharadwaj , Ufuk Topcu

Reinforcement learning (RL) systems can be complex and non-interpretable, making it challenging for non-AI experts to understand or intervene in their decisions. This is due in part to the sequential nature of RL in which actions are chosen…

Artificial Intelligence · Computer Science 2025-04-16 Amal Alabdulkarim , Madhuri Singh , Gennie Mansi , Kaely Hall , Upol Ehsan , Mark O. Riedl

Interacting with a significant number of individuals on a daily basis is commonplace for many professionals, which can lead to challenges in recalling specific details: Who is this person? What did we talk about last time? The advant of…

Human-Computer Interaction · Computer Science 2025-05-20 Raphaël A. El Haddad , Zeyu Wang , Yeonsu Shin , Ranyi Liu , Yuntao Wang , Chun Yu

We discuss the two moments of human cognition, namely, apprehension (A), whereby a coherent perception emerges from the recruitment of neuronal groups, and judgment(B),that entails the comparison of two apprehensions acquired at different…

Neurons and Cognition · Quantitative Biology 2018-02-28 F. Tito Arecchi

Social dilemmas, where mutual cooperation can lead to high payoffs but participants face incentives to cheat, are ubiquitous in multi-agent interaction. We wish to construct agents that cooperate with pure cooperators, avoid exploitation by…

Artificial Intelligence · Computer Science 2019-05-27 Alexander Peysakhovich , Adam Lerer

Recurrent Neural Networks (RNNs) have become increasingly popular for the task of language understanding. In this task, a semantic tagger is deployed to associate a semantic label to each word in an input sequence. The success of RNN may be…

Computation and Language · Computer Science 2015-06-02 Baolin Peng , Kaisheng Yao

It is the purpose of the present article to collect arguments for, that there should exist in fact -- although not necessarily yet found -- some law, which imply an adjustment to special features to occur in the future. In our own "complex…

General Physics · Physics 2015-03-31 Holger Bech Nielsen

We consider a number of questions related to tradeoffs between reward and regret in repeated gameplay between two agents. To facilitate this, we introduce a notion of $\textit{generalized equilibrium}$ which allows for asymmetric regret…

Computer Science and Game Theory · Computer Science 2023-12-19 William Brown , Jon Schneider , Kiran Vodrahalli

The iterated prisoner's dilemma is a game that produces many counter-intuitive and complex behaviors in a social environment, based on very simple basic rules. It illustrates that cooperation can be a good thing even in a competitive world,…

Computer Science and Game Theory · Computer Science 2020-09-07 Robert Prentner

Contextual bandit learning is a reinforcement learning problem where the learner repeatedly receives a set of features (context), takes an action and receives a reward based on the action and context. We consider this problem under a…

Machine Learning · Computer Science 2012-03-05 Alekh Agarwal , Miroslav Dudík , Satyen Kale , John Langford , Robert E. Schapire

Humans spend a remarkable fraction of waking life engaged in acts of "mental time travel". We dwell on our actions in the past and experience satisfaction or regret. More than merely autobiographical storytelling, we use these event…

Artificial Intelligence · Computer Science 2018-12-24 Chia-Chun Hung , Timothy Lillicrap , Josh Abramson , Yan Wu , Mehdi Mirza , Federico Carnevale , Arun Ahuja , Greg Wayne

Promoting cooperation is an intellectual challenge in the social sciences, for which the iterated Prisoners' Dilemma (IPD) is a fundamental framework. The traditional view that there exists no simple ultimatum strategy whereby one player…

Physics and Society · Physics 2015-05-12 Bin Xu , Yanran Zhou , Jaimie W. Lien , Jie Zheng , Zhijian Wang

This work investigates the case of a network of agents that attempt to learn some unknown state of the world amongst the finitely many possibilities. At each time step, agents all receive random, independently distributed private signals…

Applications · Statistics 2016-11-29 M. Amin Rahimian , Ali Jadbabaie

Recently, Frazier et al. proposed a natural model for crowdsourced exploration of different a priori unknown options: a principal is interested in the long-term welfare of a population of agents who arrive one by one in a multi-armed bandit…

Computer Science and Game Theory · Computer Science 2015-12-29 Li Han , David Kempe , Ruixin Qiang

Comment on [R.L. Ingraham, Phys. Rev. A 50, 4502 (1994)]. Ingraham suggested ``a delayed-choice experiment with partial, controllable memory erasing''. It is shown that he cannot be right since his predictions contradict relativistic…

Quantum Physics · Physics 2016-09-08 Y. Aharonov , S. Popsecu , L. Vaidman

Reasoning at multiple levels of temporal abstraction is one of the key attributes of intelligence. In reinforcement learning, this is often modeled through temporally extended courses of actions called options. Options allow agents to make…

Machine Learning · Computer Science 2023-04-13 Marlos C. Machado , Andre Barreto , Doina Precup , Michael Bowling

In this paper we consider finite conditional random quantities and conditional previsions assessments in the setting of coherence. We use a suitable representation for conditional random quantities; in particular the indicator of a…

Probability · Mathematics 2015-04-13 Angelo Gilio , Giuseppe Sanfilippo

We study a dynamic contracting problem with multiple agents and limited commitment. A principal seeks to screen efficient agents using one-period contracts, but is tempted to revise contract terms upon knowing an agent's type. Alterations…

Theoretical Economics · Economics 2025-03-24 Mehmet Ekmekci , Lucas Maestri , Dong Wei