中文
相关论文

相关论文: Scalable Board Expansion within a General Game Sys…

200 篇论文

Decision making in modern large-scale and complex systems such as communication networks, smart electricity grids, and cyber-physical systems motivate novel game-theoretic approaches. This paper investigates big strategic (non-cooperative)…

计算机科学与博弈论 · 计算机科学 2016-09-22 Tansu Alpcan , Benjamin I. P. Rubinstein , Christopher Leckie

The drivers of compositionality in artificial languages that emerge when two (or more) agents play a non-visual referential game has been previously investigated using approaches based on the REINFORCE algorithm and the (Neural) Iterated…

计算与语言 · 计算机科学 2020-12-22 Kevin Denamganaï , James Alfred Walker

The real world unfolds along a single set of physics laws, yet human intelligence demonstrates a remarkable capacity to generalize experiences from this singular physical existence into a multiverse of games, each governed by entirely…

计算机视觉与模式识别 · 计算机科学 2026-05-13 Kuan Zhang , Dongchen Liu , Qiyue Zhao , Tianyu Xin , Yue Su , Haisheng Wang , Han Yin , Hongbo Ma , Peize Li , Tianjun Gu , Xiangnan Wu , Xinran Zhang , Yongxuan Li , Zirong Chen , Yiming Li

Self-play is a technique for machine learning in multi-agent systems where a learning algorithm learns by interacting with copies of itself. Self-play is useful for generating large quantities of data for learning, but has the drawback that…

计算机科学与博弈论 · 计算机科学 2023-11-30 Revan MacQueen , James R. Wright

Playing text-based games requires skills in processing natural language and sequential decision making. Achieving human-level performance on text-based games remains an open challenge, and prior research has largely relied on hand-crafted…

In this paper, we consider continuous-time semi-decentralized dynamics for the equilibrium computation in a class of aggregative games. Specifically, we propose a scheme where decentralized projected-gradient dynamics are driven by an…

最优化与控制 · 数学 2018-03-29 Claudio De Persis , Sergio Grammatico

The Gaussian expansion has been developed since early 80s as a powerful analytical method, which enables nonperturbative studies of various systems using `perturbative' calculations. Recently the method has been used to suggest that 4d…

高能物理 - 理论 · 物理学 2009-11-10 Jun Nishimura , Toshiyuki Okubo , Fumihiko Sugino

In this work, we propose, for the first time, a reinforcement learning framework specifically designed for zero-sum linear-quadratic stochastic differential games. This approach offers a generalized solution for scenarios in which accurate…

最优化与控制 · 数学 2026-02-10 Yiyuan Wang

We study online reinforcement learning in average-reward stochastic games (SGs). An SG models a two-player zero-sum game in a Markov environment, where state transitions and one-step payoffs are determined simultaneously by a learner and an…

机器学习 · 计算机科学 2017-12-05 Chen-Yu Wei , Yi-Te Hong , Chi-Jen Lu

Games have been the perfect test-beds for artificial intelligence research for the characteristics that widely exist in real-world scenarios. Learning and optimisation, decision making in dynamic and uncertain environments, game theory,…

人工智能 · 计算机科学 2024-06-05 Chengpeng Hu , Yunlong Zhao , Ziqi Wang , Haocheng Du , Jialin Liu

Large Language Models' (LLMs) programming capabilities enable their participation in open-source games: a game-theoretic setting in which players submit computer programs in lieu of actions. These programs offer numerous advantages,…

计算机科学与博弈论 · 计算机科学 2025-12-02 Swadesh Sistla , Max Kleiman-Weiner

Mean-field game theory relies on approximating games that are intractable to model due to a very large to infinite population of players. While these kinds of games can be solved analytically via the associated system of partial…

机器学习 · 计算机科学 2026-04-16 Anna C. M. Thöni , Yoram Bachrach , Tal Kachman

In multi-player card games such as Skat or Bridge, the early stages of the game, such as bidding, game selection, and initial card selection, are often more critical to the success of the play than refined middle- and end-game play. At the…

人工智能 · 计算机科学 2025-12-18 Stefan Edelkamp

According to evolutionary game theory, cooperation in public goods games is eliminated by free-riders, yet in nature, cooperation is ubiquitous. Artificial models resolve this contradiction via the mechanism of network reciprocity. However,…

计算机科学与博弈论 · 计算机科学 2016-05-10 Steve Miller , Joshua Knowles

This paper studies a large class of two-player perfect-information turn-based parity games on infinite graphs, namely those generated by collapsible pushdown automata. The main motivation for studying these games comes from the connections…

形式语言与自动机理论 · 计算机科学 2020-10-14 Christopher H. Broadbent , Arnaud Carayol , Matthew Hague , Andrzej S. Murawski , C. -H. Luke Ong , Olivier Serre

Software game is a kind of application that is used not only for entertainment, but also for serious purposes that can be applicable to different domains such as education, business, and health care. Although the game development process…

软件工程 · 计算机科学 2017-11-27 Saiqa Aleem , Luiz Fernando Capretz , Faheem Ahmed

The objective of this book is to give a comprehensive presentation of the research field concerned with infinite duration games on graphs. Historically, these game models appeared in the study of automata and logic, and they later became…

In the empirical approach to game-theoretic analysis (EGTA), the model of the game comes not from declarative representation, but is derived by interrogation of a procedural description of the game environment. The motivation for developing…

计算机科学与博弈论 · 计算机科学 2025-02-21 Michael P. Wellman , Karl Tuyls , Amy Greenwald

The direct purpose of this paper - as its title suggests - is to present how the visual evaluator extension is implemented in the GRASP programming system. The indirect purpose is to provide a tutorial around the design of GRASP, and in…

人机交互 · 计算机科学 2025-08-08 Panicz Maciej Godek

While current General Game Playing (GGP) systems facilitate useful research in Artificial Intelligence (AI) for game-playing, they are often somewhat specialised and computationally inefficient. In this paper, we describe the "ludemic"…