中文
相关论文

相关论文: A Formal Solution to the Grain of Truth Problem

200 篇论文

When inferring the goals that others are trying to achieve, people intuitively understand that others might make mistakes along the way. This is crucial for activities such as teaching, offering assistance, and deciding between blame or…

人工智能 · 计算机科学 2021-06-28 Arwa Alanqary , Gloria Z. Lin , Joie Le , Tan Zhi-Xuan , Vikash K. Mansinghka , Joshua B. Tenenbaum

Models of economic decision makers often include idealized assumptions, such as rationality, perfect foresight, and access to all relevant pieces of information. These assumptions often assure the models' internal validity, but, at the same…

综合经济学 · 经济学 2021-07-09 Patrick Reinwald , Stephan Leitner , Friederike Wall

We generalise the problem of inverse reinforcement learning to multiple tasks, from multiple demonstrations. Each one may represent one expert trying to solve a different task, or as different experts trying to solve the same task. Our main…

机器学习 · 统计学 2012-09-04 Christos Dimitrakakis , Constantin Rothkopf

In many prediction problems, the predictive model affects the distribution of the prediction target. This phenomenon is known as performativity and is often caused by the behavior of individuals with vested interests in the outcome of the…

机器学习 · 统计学 2024-06-03 Seamus Somerstep , Ya'acov Ritov , Yuekai Sun

We consider the problem of finding optimal classifiers in an adversarial setting where the class-1 data is generated by an attacker whose objective is not known to the defender -- an aspect that is key to realistic applications but has so…

计算机科学与博弈论 · 计算机科学 2021-10-26 Patrick Loiseau , Benjamin Roussillon

Bayesian priors offer a compact yet general means of incorporating domain knowledge into many learning tasks. The correctness of the Bayesian analysis and inference, however, largely depends on accuracy and correctness of these priors.…

机器学习 · 计算机科学 2012-02-20 Mahdi MIlani Fard , Joelle Pineau , Csaba Szepesvari

A Bayesian optimization algorithm for the nurse scheduling problem is presented, which involves choosing a suitable scheduling rule from a set for each nurses assignment. Unlike our previous work that used Gas to implement implicit…

神经与进化计算 · 计算机科学 2010-07-05 Jingpeng Li , Uwe Aickelin

In bipartite matching problems, agents on two sides of a graph want to be paired according to their preferences. The stability of a matching depends on these preferences, which in uncertain environments also reflect agents' beliefs about…

计算机科学与博弈论 · 计算机科学 2025-11-10 Jonathan Shaki , Jiarui Gan , Sarit Kraus

We propose and analyze a framework for mean-field Markov games under model uncertainty. In this framework, a state-measure flow describing the collective behavior of a population affects the given reward function as well as the unknown…

最优化与控制 · 数学 2024-10-16 Johannes Langner , Ariel Neufeld , Kyunghyun Park

In the impartial selection problem, a subset of agents up to a fixed size $k$ among a group of $n$ is to be chosen based on votes cast by the agents themselves. A selection mechanism is impartial if no agent can influence its own chance of…

计算机科学与博弈论 · 计算机科学 2024-08-06 Javier Cembrano , Svenja M. Griesbach , Maximilian J. Stahlberg

The question of what global information must distributed rational agents a-priori know about the network in order for equilibrium to be possible is researched here. Until now, distributed algorithms with rational agents have assumed that…

分布式、并行与集群计算 · 计算机科学 2018-04-10 Yehuda Afek , Shaked Rafaeli , Moshe Sulamy

Non-Bayesian social learning is a framework for distributed hypothesis testing aimed at learning the true state of the environment. Traditionally, the agents are assumed to receive observations conditioned on the same true state, although…

社会与信息网络 · 计算机科学 2024-06-26 Valentina Shumovskaia , Mert Kayaalp , Ali H. Sayed

When observing the actions of others, humans make inferences about why they acted as they did, and what this implies about the world; humans also use the fact that their actions will be interpreted in this manner, allowing them to act…

Multi-agent reinforcement learning methods have shown remarkable potential in solving complex multi-agent problems but mostly lack theoretical guarantees. Recently, mean field control and mean field games have been established as a…

机器学习 · 计算机科学 2021-12-20 Kai Cui , Anam Tahir , Mark Sinzger , Heinz Koeppl

This paper investigates how an autonomous agent can transmit information through its motion in an adversarial setting. We consider scenarios where an agent must reach its goal while deceiving an intelligent observer about its destination.…

系统与控制 · 电气工程与系统科学 2025-06-17 Violetta Rostobaya , James Berneburg , Yue Guan , Michael Dorothy , Daigo Shishika

Wisdom of the crowd revealed a striking fact that the majority answer from a crowd is often more accurate than any individual expert. We observed the same story in machine learning--ensemble methods leverage this idea to combine multiple…

机器学习 · 计算机科学 2019-10-01 Tianyi Luo , Yang Liu

We consider distributed learning problem in games with an unknown cost-relevant parameter, and aim to find the Nash equilibrium while learning the true parameter. Inspired by the social learning literature, we propose a distributed…

最优化与控制 · 数学 2023-03-14 Shijie Huang , Jinlong Lei , Yiguang Hong

The optimal policy of a reinforcement learning problem is often discontinuous and non-smooth. I.e., for two states with similar representations, their optimal policies can be significantly different. In this case, representing the entire…

机器学习 · 计算机科学 2020-02-10 Zhimin Hou , Kuangen Zhang , Yi Wan , Dongyu Li , Chenglong Fu , Haoyong Yu

In network formation games, agents form edges with each other to maximize their utility. Each agent's utility depends on its private beliefs and its edges in the network. Strategic agents can misrepresent their beliefs to get a better…

最优化与控制 · 数学 2024-09-04 Akhil Jalan , Deepayan Chakrabarti

We propose a solution and a mechanism for two-agent social choice problems with large (infinite) policy spaces. Our solution is an efficient compromise rule between the two agents, built on a common cardinalization of their preferences. Our…

理论经济学 · 经济学 2026-02-03 Federico Echenique , Matías Núñez
‹ 上一页 1 8 9 10 下一页 ›