中文
相关论文

相关论文: Lead distance under a pickoff limit in Major Leagu…

200 篇论文

Two-player games on graphs is central in many problems in formal verification and program analysis such as synthesis and verification of open systems. In this work we consider solving recursive game graphs (or pushdown game graphs) that can…

计算机科学中的逻辑 · 计算机科学 2016-05-17 Krishnendu Chatterjee , Yaron Velner

Recent successes of game-theoretic formulations in ML have caused a resurgence of research interest in differentiable games. Overwhelmingly, that research focuses on methods and upper bounds on their speed of convergence. In this work, we…

机器学习 · 计算机科学 2020-09-16 Adam Ibrahim , Waïss Azizian , Gauthier Gidel , Ioannis Mitliagkas

We study two-player reachability games on finite graphs. At each state the interaction between the players is concurrent and there is a stochastic Nature. Players also play stochastically. The literature tells us that 1) Player B, who wants…

计算机科学与博弈论 · 计算机科学 2021-10-29 Benjamin Bordais , Patricia Bouyer , Stéphane Le Roux

In Major League Baseball, every ballpark is different, with different dimensions and climates. These differences make some ballparks more conducive to hitting home runs than others. Several factors conspire to make estimation of these…

应用统计 · 统计学 2025-06-30 Jason A. Osborne , Richard A. Levine

We consider a computing system where a master processor assigns tasks for execution to worker processors through the Internet. We model the workers decision of whether to comply (compute the task) or not (return a bogus result to save the…

分布式、并行与集群计算 · 计算机科学 2015-08-25 Antonio Fernández Anta , Chryssis Georgiou , Miguel A. Mosteiro , Daniel Pareja

Problem definition: Professional sports leagues may be suspended due to various reasons such as the recent COVID-19 pandemic. A critical question the league must address when re-opening is how to appropriately select a subset of the…

最优化与控制 · 数学 2024-04-02 Ali Hassanzadeh , Mojtaba Hosseini , John G. Turner

In this paper, we investigate a new model of a linear-quadratic mean-field stochastic Stackelberg differential game with one leader and two followers, in which the leader is allowed to stop her strategy at a random time. Our overarching…

最优化与控制 · 数学 2021-06-08 Zhun Gou , Nan-jing Huang , Ming-hui Wang

In the paper we present a model of discrete-time mean-field game with several populations of players. Mean-field games with multiple populations of the players have only been studied in the literature in the continuous-time setting. The…

最优化与控制 · 数学 2023-04-07 Piotr Więcek

The speed-accuracy tradeoffs are prevalent in a wide range of physical systems. In this paper, we demonstrate speed-accuracy tradeoffs in the game of cricket, where 'batters' score runs on the balls bowled by the 'bowlers'. It is shown that…

物理与社会 · 物理学 2025-12-17 Mohd Suhail Rizvi

Most historical National Football League (NFL) analysis, both mainstream and academic, has relied on public, play-level data to generate team and player comparisons. Given the number of oft omitted variables that impact on-field results,…

应用统计 · 统计学 2020-05-14 Michael J. Lopez

In Major League Baseball, strategy and planning are major factors in determining the outcome of a game. Previous studies have aided this by building machine learning models for predicting the winning team of any given game. We extend this…

机器学习 · 计算机科学 2025-11-05 Morgan Allen , Paul Savala

We study payoff manipulation in repeated multi-objective Stackelberg games, where a leader may strategically influence a follower's deterministic best response, e.g., by offering a share of their own payoff. We assume that the follower's…

计算机科学与博弈论 · 计算机科学 2025-08-27 Phurinut Srisawad , Juergen Branke , Long Tran-Thanh

We consider the Bilevel Knapsack with Interdiction Constraints, an extension of the classic 0-1 knapsack problem formulated as a Stackelberg game with two agents, a leader and a follower, that choose items from a common set and hold their…

计算机科学与博弈论 · 计算机科学 2018-11-13 Federico Della Croce , Rosario Scatamacchia

We study Markov decision processes (MDPs) with multiple limit-average (or mean-payoff) functions. We consider two different objectives, namely, expectation and satisfaction objectives. Given an MDP with k limit-average functions, in the…

计算机科学与博弈论 · 计算机科学 2015-07-01 Tomáš Brázdil , Václav Brožek , Krishnendu Chatterjee , Vojtěch Forejt , Antonín Kučera

In this paper, we settle the sampling complexity of solving discounted two-player turn-based zero-sum stochastic games up to polylogarithmic factors. Given a stochastic game with discount factor $\gamma\in(0,1)$ we provide an algorithm that…

机器学习 · 计算机科学 2019-08-30 Aaron Sidford , Mengdi Wang , Lin F. Yang , Yinyu Ye

The paper proposes a natural measure space of zero-sum perfect information games with upper semicontinuous payoffs. Each game is specified by the game tree, and by the assignment of the active player and of the capacity to each node of the…

计算机科学与博弈论 · 计算机科学 2021-04-22 János Flesch , Arkadi Predtetchinski , Ville Suomala

This paper is concerned with the closed-loop Stackelberg strategy for linear-quadratic leader-follower game. Completely different from the open-loop and feedback Stackelberg strategy, the solvability of the closed-loop solution even the…

最优化与控制 · 数学 2024-03-19 Hongdan Li , Juanjuan Xu , Hunashui Zhang

We address two-player general-sum stochastic Stackelberg games (SSGs), where the leader's policy is optimized considering the best-response follower whose policy is optimal for its reward under the leader. Existing policy gradient and value…

计算机科学与博弈论 · 计算机科学 2026-03-17 Mikoto Kudo , Youhei Akimoto

Consider the problem of modeling memory effects in discrete-state random walks using higher-order Markov chains. This paper explores cross validation and information criteria as proxies for a model's predictive accuracy. Our objective is to…

统计方法学 · 统计学 2019-03-22 Joshua C. Chang

This paper reframes approachability theory within the context of population games. Thus, whilst one player aims at driving her average payoff to a predefined set, her opponent is not malevolent but rather extracted randomly from a…

最优化与控制 · 数学 2014-07-16 Dario Bauso , Thomas W L Norman