中文
相关论文

相关论文: An improved lion strategy for the lion and man pro…

200 篇论文

In this paper, we consider a simple discrete-time optimal betting problem using the celebrated Kelly criterion, which calls for maximization of the expected logarithmic growth of wealth. While the classical Kelly betting problem can be…

最优化与控制 · 数学 2021-03-11 Chung-Han Hsieh

Motivated by applications in job scheduling, queuing networks, and load balancing in cyber-physical systems, we develop and analyze a game-theoretic framework to balance the load among servers in static and dynamic settings. In these…

计算机科学与博弈论 · 计算机科学 2025-12-24 Fatemeh Fardno , S. Rasoul Etesami

We present a formal model of human decision-making in explore-exploit tasks using the context of multi-armed bandit problems, where the decision-maker must choose among multiple options with uncertain rewards. We address the standard…

机器学习 · 计算机科学 2019-12-23 Paul Reverdy , Vaibhav Srivastava , Naomi E. Leonard

This paper explores an idealized dynamic population sizing strategy for solving additive decomposable problems of uniform scale. The method is designed on top of the foundations of existing population sizing theory for this class of…

神经与进化计算 · 计算机科学 2011-04-15 Fernando G. Lobo

We introduce the notion of Local Computation Mechanism Design - designing game theoretic mechanisms which run in polylogarithmic time and space. Local computation mechanisms reply to each query in polylogarithmic time and space, and the…

计算机科学与博弈论 · 计算机科学 2014-06-10 Avinatan Hassidim , Yishay Mansour , Shai Vardi

This paper considers the design of fully distributed Nash equilibrium seeking strategies for multi-agent games. To develop fully distributed seeking strategies, two adaptive control laws, including a node-based control law and an edge-based…

最优化与控制 · 数学 2019-12-03 Maojiao Ye , Guoqiang Hu

This paper mainly conducts further research to alleviate the issue of limit cycling behavior in training generative adversarial networks (GANs) through the proposed predictive centripetal acceleration algorithm (PCAA). Specifically, we…

机器学习 · 统计学 2023-08-14 Li Keke , Yang Xinmin

In this paper, we deal with the equilibrium selection problem, which amounts to steering a population of individuals engaged in strategic game-theoretic interactions to a desired collective behavior. In the literature, this problem has been…

系统与控制 · 电气工程与系统科学 2025-11-11 Lorenzo Zino , Mengbin Ye , Giuseppe Carlo Calafiore , Alessandro Rizzo

We solve the classical "Game of Pure Strategy" using linear programming. We notice an intricate even-odd behavior in the results of our computations, that seems to encourage odd or maximal bids.

计算机科学与博弈论 · 计算机科学 2016-06-28 Glenn C. Rhoads , Laurent Bartholdi

In 1980 Steven Smale introduced a class of strategies for the Iterated Prisoner's Dilemma which used as data the running average of the previous payoff pairs. This approach is quite different from the Markov chain approach, common before…

动力系统 · 数学 2017-04-18 Ethan Akin

We propose a decentralized solution for a pursuit-evasion game involving a heterogeneous group of rational (selfish) pursuers and a single evader based on the framework of potential games. In the proposed game, the evader aims to delay (or,…

系统与控制 · 电气工程与系统科学 2021-08-19 Yoonjae Lee , Efstathios Bakolas

The policy iteration method is a classical algorithm for solving optimal control problems. In this paper, we introduce a policy iteration method for Mean Field Games systems, and we study the convergence of this procedure to a solution of…

偏微分方程分析 · 数学 2021-07-12 Simone Cacace , Fabio Camilli , Alessandro Goffi

This paper proposes a novel distributed approach for solving a cooperative Constrained Multi-agent Reinforcement Learning (CMARL) problem, where agents seek to minimize a global objective function subject to shared constraints. Unlike…

系统与控制 · 电气工程与系统科学 2026-05-08 Ali Kahe , Hamed Kebriaei

Enormous successes have been made by quantum algorithms during the last decade. In this paper, we combine the quantum game with the problem of data clustering, and then develop a quantum-game-based clustering algorithm, in which data points…

机器学习 · 计算机科学 2015-05-13 Qiang Li , Yan He , Jing-ping Jiang

We construct a semi-Lagrangian scheme for first-order, time-dependent, and non-local Mean Field Games. The convergence of the scheme to a weak solution of the system is analyzed by exploiting a key monotonicity property. To solve the…

数值分析 · 数学 2026-05-12 Elisabetta Carlini , Valentina Coscetti

Capability planning problems are pervasive throughout many areas of human interest with prominent examples found in defense and security. Planning provides a unique context for optimization that has not been explored in great detail and…

神经与进化计算 · 计算机科学 2009-07-03 James M. Whitacre , Hussein A. Abbass , Ruhul Sarker , Axel Bender , Stephen Baker

We propose EAGLE update rule, a novel optimization method that accelerates loss convergence during the early stages of training by leveraging both current and previous step parameter and gradient values. The update algorithm estimates…

机器学习 · 计算机科学 2025-02-04 Takumi Fujimoto , Hiroaki Nishi

Strategic reasoning enables agents to cooperate, communicate, and compete with other agents in diverse situations. Existing approaches to solving strategic games rely on extensive training, yielding strategies that do not generalize to new…

人工智能 · 计算机科学 2023-05-31 Kanishk Gandhi , Dorsa Sadigh , Noah D. Goodman

In this paper we analyze, based on an interplay between ideas and techniques from logic and geometric analysis, a pursuit-evasion game. More precisely, we focus on a uniform betweenness property and use it in the study of a discrete lion…

度量几何 · 数学 2020-11-13 Ulrich Kohlenbach , Genaro López-Acedo , Adriana Nicolae

The paper proposes a novel upper confidence bound (UCB) procedure for identifying the arm with the largest mean in a multi-armed bandit game in the fixed confidence setting using a small number of total samples. The procedure cannot be…

机器学习 · 统计学 2013-12-30 Kevin Jamieson , Matthew Malloy , Robert Nowak , Sébastien Bubeck