中文
相关论文

相关论文: Neural Network-based Information Set Weighting for…

200 篇论文

An increasing number of domains are providing us with detailed trace data on human decisions in settings where we can evaluate the quality of these decisions via an algorithm. Motivated by this development, an emerging line of work has…

人工智能 · 计算机科学 2016-06-17 Ashton Anderson , Jon Kleinberg , Sendhil Mullainathan

Many natural processes rely on optimizing the success ratio of a search process. We use an experimental setup consisting of a simple online game in which players have to find a target hidden on a board, to investigate the how the rounds are…

生物物理 · 物理学 2017-05-19 Ricardo Martinez-Garcia , Justin M. Calabrese , Cristobal Lopez

Players are statistical learners who learn about payoffs from data. They may interpret the same data differently, but have common knowledge of a class of learning procedures. I propose a metric for the analyst's "confidence" in a strategic…

理论经济学 · 经济学 2020-07-13 Annie Liang

We consider the problem of prediction by a machine learning algorithm, called learner, within an adversarial learning setting. The learner's task is to correctly predict the class of data passed to it as a query. However, along with queries…

机器学习 · 计算机科学 2020-02-11 Prithviraj Dasgupta , Joseph B. Collins , Michael McCarrick

In this paper, we extend the Descent framework, which enables learning and planning in the context of two-player games with perfect information, to the framework of stochastic games. We propose two ways of doing this, the first way…

人工智能 · 计算机科学 2023-02-10 Quentin Cohen-Solal , Tristan Cazenave

Agents trained in simulation may make errors in the real world due to mismatches between training and execution environments. These mistakes can be dangerous and difficult to discover because the agent cannot predict them a priori. We…

机器学习 · 计算机科学 2018-05-24 Ramya Ramakrishnan , Ece Kamar , Debadeepta Dey , Julie Shah , Eric Horvitz

In this work, we proposed a new $N$-person game in which the players can bet on two options, for example represented by two boxers. Some of the players have privileged information about the boxers and part of them can provide this…

统计力学 · 物理学 2018-04-04 Roberto da Silva , Henrique A. Fernandes

Current research in distributed Nash equilibrium (NE) seeking in the partial information setting assumes that information is exchanged between agents that are "truthful". However, in general noncooperative games agents may consider sending…

最优化与控制 · 数学 2021-11-30 Dian Gadjov , Lacra Pavel

We study strategic interaction in linear-quadratic network games where agents act on subjective, misspecified models of their environment. Agents observe noisy aggregate signals generated by local network externalities and interpret them…

计算机科学与博弈论 · 计算机科学 2026-03-19 Quanyan Zhu , Zhengye Han

Search in test time is often used to improve the performance of reinforcement learning algorithms. Performing theoretically sound search in fully adversarial two-player games with imperfect information is notoriously difficult and requires…

计算机科学与博弈论 · 计算机科学 2025-01-30 Ondrej Kubicek , Neil Burch , Viliam Lisy

We introduce an algorithm where the individual bits representing the weights of a neural network are learned. This method allows training weights with integer values on arbitrary bit-depths and naturally uncovers sparse networks, without…

机器学习 · 计算机科学 2022-02-22 Cristian Ivan

In settings with incomplete information, players can find it difficult to coordinate to find states with good social welfare. For example, in financial settings, if a collection of financial firms have limited information about each other's…

计算机科学与博弈论 · 计算机科学 2014-09-01 Avrim Blum , Jamie Morgenstern , Ankit Sharma , Adam Smith

We consider a class of interdependent security games on networks where each node chooses a personal level of security investment. The attack probability experienced by a node is a function of her own investment and the investment by her…

计算机科学与博弈论 · 计算机科学 2016-08-16 Ashish R. Hota , Shreyas Sundaram

A network of cognitive transmitters is considered. Each transmitter has to decide his power control policy in order to maximize energy-efficiency of his transmission. For this, a transmitter has two actions to take. He has to decide whether…

计算机科学与博弈论 · 计算机科学 2012-10-25 Maël Le Treust , Yezekael Hayel , Samson Lasaulce , Mérouane Debbah

Humans rapidly learn abstract knowledge when encountering novel environments and flexibly deploy this knowledge to guide efficient and intelligent action. Can modern AI systems learn and plan in a similar way? We study this question using a…

We examine settings in which agents choose behaviors and care about their neighbors' behaviors, but have incomplete information about the network in which they are embedded. We develop a model in which agents use local knowledge of their…

理论经济学 · 经济学 2024-12-04 Promit K. Chaudhuri , Matthew O. Jackson , Sudipta Sarangi , Hector Tzavellas

The problem of state estimation for unobservable distribution systems is considered. A deep learning approach to Bayesian state estimation is proposed for real-time applications. The proposed technique consists of distribution learning of…

机器学习 · 统计学 2019-02-26 Kursat Rasim Mestav , Jaime Luengo-Rozas , Lang Tong

Understanding how people behave in strategic settings--where they make decisions based on their expectations about the behavior of others--is a long-standing problem in the behavioral sciences. We conduct the largest study to date of…

综合经济学 · 经济学 2024-08-16 Jian-Qiao Zhu , Joshua C. Peterson , Benjamin Enke , Thomas L. Griffiths

Infinite games with imperfect information are known to be undecidable unless the information flow is severely restricted. One fundamental decidable case occurs when there is a total ordering among players, such that each player has access…

计算机科学与博弈论 · 计算机科学 2016-07-19 Dietmar Berwanger , Anup Basil Mathew , Marie van den Bogaard

Desensitization addresses safe optimal planning under parametric uncertainties by providing sensitivity function-based risk estimates. This paper expands upon the existing work on desensitization in optimal control to address safe planning…

系统与控制 · 电气工程与系统科学 2024-02-08 Vinodhini Comandur , Tulasi Ram Vechalapu , Venkata Ramana Makkapati , Panagiotis Tsiotras , Seth Hutchinson