中文
相关论文

相关论文: An improved lion strategy for the lion and man pro…

200 篇论文

Planning safe robot motions in the presence of humans requires reliable forecasts of future human motion. However, simply predicting the most likely motion from prior interactions does not guarantee safety. Such forecasts fail to model the…

人工智能 · 计算机科学 2023-10-23 Kushal Kedia , Prithwish Dan , Sanjiban Choudhury

An universal primal-dual approach of description equilibriums in large class of hierarchical congestion population games is proposed. At the very core of the approach is hierarchy of enclosed to each other transport networks. In different…

最优化与控制 · 数学 2016-03-09 Alexander Gasnikov , Evgenia Gasnikova , Sergey Matsievsky , Inna Usik

While Large Language Models (LLMs) excel in certain reasoning tasks, they struggle in multi-agent games where the final outcome depends on the joint strategies of all agents. In multi-agent games, the non-stationarity of other agents brings…

人工智能 · 计算机科学 2026-05-26 Yidong He , Yutao Lai , Pengxu Yang , Jiarui Gan , Jiexin Wang , Yi Cai , Mengchen Zhao

Learning in strategy games (e.g. StarCraft, poker) requires the discovery of diverse policies. This is often achieved by iteratively training new policies against existing ones, growing a policy population that is robust to exploit. This…

人工智能 · 计算机科学 2022-02-16 Siqi Liu , Luke Marris , Daniel Hennes , Josh Merel , Nicolas Heess , Thore Graepel

This paper combines the idea of a hierarchical distributed genetic algorithm with different inter-agent partnering strategies. Cascading clusters of sub-populations are built from bottom up, with higher-level sub-populations optimising…

神经与进化计算 · 计算机科学 2010-07-05 Uwe Aickelin

There has been a recent explosion in the capabilities of game-playing artificial intelligence. Many classes of tasks, from video games to motor control to board games, are now solvable by fairly generic algorithms, based on deep learning…

人工智能 · 计算机科学 2018-10-18 Vlad Firoiu , Tina Ju , Josh Tenenbaum

This paper considers a game-theoretic formulation of the covert communications problem with finite blocklength, where the transmitter (Alice) can randomly vary her transmit power in different blocks, while the warden (Willie) can randomly…

信息论 · 计算机科学 2020-05-28 Alex S. Leong , Daniel E. Quevedo , Subhrakanti Dey

This paper presents a distributed rule-based Lloyd algorithm (RBL) for multi-robot motion planning and control. The main limitations of the basic Loyd-based algorithm (LB) concern deadlock issues and the failure to address dynamic…

机器人学 · 计算机科学 2025-05-07 Manuel Boldrer , Alvaro Serra-Gomez , Lorenzo Lyons , Vit Kratky , Javier Alonso-Mora , Laura Ferranti

The classic paper of Shapley and Shubik \cite{Shapley1971assignment} characterized the core of the assignment game using ideas from matching theory and LP-duality theory and their highly non-trivial interplay. Whereas the core of this game…

计算机科学与博弈论 · 计算机科学 2021-07-19 Vijay V. Vazirani

Situations of conflict giving rise to social dilemmas are widespread in society and game theory is one major way in which they can be investigated. Starting from the observation that individuals in society interact through networks of…

物理与社会 · 物理学 2010-11-24 Enea Pestelacci , Marco Tomassini , Leslie Luthi

We present a simple game which mimics the complex dynamics found in most natural and social systems. Intelligent players modify their strategies periodically, depending on their performances. We propose that the agents use hybridized…

统计力学 · 物理学 2009-11-07 Marko Sysi-Aho , Anirban Chakraborti , Kimmo Kaski

We introduce a primal-dual stochastic gradient oracle method for distributed convex optimization problems over networks. We show that the proposed method is optimal in terms of communication steps. Additionally, we propose a new analysis…

最优化与控制 · 数学 2019-11-28 Darina Dvinskikh , Eduard Gorbunov , Alexander Gasnikov , Pavel Dvurechensky , Cesar A. Uribe

Optimization methods are at the core of many problems in signal/image processing, computer vision, and machine learning. For a long time, it has been recognized that looking at the dual of an optimization problem may drastically simplify…

数值分析 · 计算机科学 2014-12-04 Nikos Komodakis , Jean-Christophe Pesquet

We consider distributed computation of generalized Nash equilibrium (GNE) over networks, in games with shared coupling constraints. Existing methods require that each player has full access to opponents' decisions. In this paper, we assume…

最优化与控制 · 数学 2024-10-30 Lacra Pavel

The paper develops a technique for solving a linear equation $Ax=b$ with a square and nonsingular matrix $A$, using a decentralized gradient algorithm. In the language of control theory, there are $n$ agents, each storing at time $t$ an…

系统与控制 · 计算机科学 2015-09-16 Brian D. O. Anderson , Shaoshuai Mou , A. Stephen Morse , Uwe Helmke

We study the capture of a diffusing "lamb" by diffusing "lions" in one dimension. The capture dynamics is exactly soluble by probabilistic techniques when the number of lions is very small, and is tractable by extreme statistics…

统计力学 · 物理学 2009-10-31 S. Redner , P. L. Krapivsky

Models of adaptive bet-hedging commonly adopt insights from Kelly's famous work on optimal gambling strategies and the financial value of information. In particular, such models seek evolutionary solutions that maximize long term average…

种群与进化 · 定量生物学 2020-03-18 Omri Tal , Tat Dat Tran

We study variants of a stochastic game inspired by backgammon where players may propose to double the stake, with the game state dictated by a one-dimensional random walk. Our variants allow for different numbers of proposals and different…

Generalized from the concept of consensus, this paper considers a group of edge agreements, i.e. constraints defined for neighboring agents, in which each pair of neighboring agents is required to satisfy one edge agreement constraint. Edge…

最优化与控制 · 数学 2023-12-04 Zehui Lu , Shaoshuai Mou

We study the classic divide-and-choose method for equitably allocating divisible goods between two players who are rational, self-interested Bayesian agents. The players have additive values for the goods. The prior distributions on those…

计算机科学与博弈论 · 计算机科学 2024-10-22 Jamie Tucker-Foltz , Richard Zeckhauser
‹ 上一页 1 8 9 10 下一页 ›