中文
相关论文

相关论文: An improved lion strategy for the lion and man pro…

200 篇论文

The pursuit-evasion game with two persons is considered. Both players are moving in a metric space, have equal maximum speeds and complete information about the location of each other. We study the sufficient conditions for a capture (with…

最优化与控制 · 数学 2017-12-15 Olga Yufereva

In this paper, we consider groups of agents in a network that select actions in order to satisfy a set of constraints that vary arbitrarily over time and minimize a time-varying function of which they have only local observations. The…

最优化与控制 · 数学 2020-07-15 Santiago Paternain , Soomin Lee , Michael M. Zavlanos , Alejandro Ribeiro

This article adopts game theory to build a model for explaining the predation behavior of animals.We assume that both the prey and the preydator have two stratigies in this game,the active one and the passive one.By calculating the outcome…

种群与进化 · 定量生物学 2007-05-23 Shi Chen , Sheng Bao , Ling Yan , Cheng Huang

Adversarial online linear optimization (OLO) is essentially about making performance tradeoffs with respect to the unknown difficulty of the adversary. In the setting of one-dimensional fixed-time OLO on a bounded domain, it has been…

机器学习 · 统计学 2026-02-09 Zhiyu Zhang , Aaditya Ramdas

This work develops a fully decentralized multi-agent algorithm for policy evaluation. The proposed scheme can be applied to two distinct scenarios. In the first scenario, a collection of agents have distinct datasets gathered following…

机器学习 · 计算机科学 2019-08-13 Lucas Cassano , Kun Yuan , Ali H. Sayed

The divide and conquer strategy, which breaks a massive data set into a se- ries of manageable data blocks, and then combines the independent results of data blocks to obtain a final decision, has been recognized as a state-of-the-art…

机器学习 · 计算机科学 2016-03-15 Xiangyu Chang , Shaobo Lin , Yao Wang

Game theory provides the gold standard for analyzing adversarial engagements, offering strong optimality guarantees. However, these guarantees often become brittle when assumptions such as perfect information are violated. Reinforcement…

机器学习 · 计算机科学 2026-03-18 Goutam Das , Michael Dorothy , Kyle Volle , Daigo Shishika

This paper presents a reinforced genetic approach to a defined d-resource system optimization problem. The classical evolution schema was ineffective due to a very strict feasibility function in the studied problem. Hence, the presented…

神经与进化计算 · 计算机科学 2025-11-07 Leszek Sliwko

We propose a general agent population learning system, and on this basis, we propose lineage evolution reinforcement learning algorithm. Lineage evolution reinforcement learning is a kind of derivative algorithm which accords with the…

神经与进化计算 · 计算机科学 2020-10-29 Zeyu Zhang , Guisheng Yin

Humans exhibit remarkable abilities to coordinate in groups. As large language models (LLMs) become more capable, it remains an open question whether they can demonstrate comparable adaptive coordination and whether they use the same…

多智能体系统 · 计算机科学 2026-04-06 Sahaj Singh Maini , Robert L. Goldstone , Zoran Tiganj

In this paper, we study a theoretical math problem of game theory and calculus of variations in which we minimize a functional involving two players. A general relationship between the optimal strategies for both players is presented,…

最优化与控制 · 数学 2024-11-05 Grace Luo , Christopher Boyer , Siddharth Penmetsa

Distributed optimization and Nash equilibrium (NE) seeking problems have drawn much attention in the control community recently. This paper studies a class of non-cooperative games, known as N-cluster game, which subsumes both cooperative…

最优化与控制 · 数学 2023-03-01 Yipeng Pang , Guoqiang Hu

Balancing game difficulty in video games is a key task to create interesting gaming experiences for players. Mismatching the game difficulty and a player's skill or commitment results in frustration or boredom on the player's side, and…

人工智能 · 计算机科学 2024-08-14 Ronja Fuchs , Robin Gieseke , Alexander Dockhorn

This work proposes a novel distributed approach for computing a Nash equilibrium in convex games with merely monotone and restricted strongly monotone pseudo-gradients. By leveraging the idea of the centralized operator extrapolation method…

最优化与控制 · 数学 2025-07-18 Tatiana Tatarenko , Angelia Nedich

Stochastic games are a popular framework for studying multi-agent reinforcement learning (MARL). Recent advances in MARL have focused primarily on games with finitely many states. In this work, we study multi-agent learning in stochastic…

机器学习 · 计算机科学 2024-03-28 Awni Altabaa , Bora Yongacoglu , Serdar Yüksel

This paper introduces a systematic methodological framework to design and analyze distributed algorithms for optimization and games over networks. Starting from a centralized method, we identify an aggregation function involving all the…

最优化与控制 · 数学 2025-05-26 Guido Carnevale , Nicola Mimmo , Giuseppe Notarstefano

We present two improved algorithms for weighted discrete $p$-center problem for tree networks with $n$ vertices. One of our proposed algorithms runs in $O(n \log n + p \log^2 n \log(n/p))$ time. For all values of $p$, our algorithm thus…

数据结构与算法 · 计算机科学 2016-04-27 Aritra Banik , Binay Bhattacharya , Sandip Das , Tsunehiko Kameda , Zhao Song

We provide a polynomial algorithm to find the value and an optimal strategy for a generalization of the Pig game. Modeled as a competitive Markov decision process, the corresponding Bellman equations can be decoupled leading to systems of…

概率论 · 数学 2018-08-22 Fabián Crocce , Ernesto Mordecki

The LION (evoLved sIgn mOmeNtum) optimizer for deep neural network training was found by Google via program search, with the simple sign update yet showing impressive performance in training large scale networks. Although previous studies…

机器学习 · 计算机科学 2024-11-13 Yiming Dong , Huan Li , Zhouchen Lin

Recent advances in game AI, such as AlphaZero and Ath\'enan, have achieved superhuman performance across a wide range of board games. While highly powerful, these agents are ill-suited for human-AI interaction, as they consistently…

人工智能 · 计算机科学 2026-03-25 Quentin Cohen-Solal , Tristan Cazenave