中文
相关论文

相关论文: A Markov process approach to untangling intention …

200 篇论文

This paper considers the impact of unforced errors in sport. Although the proposed methods are applicable to various sports, we demonstrate the approach in the context of professional tennis. The value of the approach is that we can provide…

应用统计 · 统计学 2024-07-30 Hashan Peiris , Nirodha Epasinghege Dona , Tim Swartz

Sports tracking data are the high-resolution spatiotemporal observations of a competitive event. The growing collection of these data in professional sport allows us to address a fundamental problem of modern sport: how to attribute value…

应用统计 · 统计学 2020-05-27 Stephanie Kovalchik , Martin Ingram , Kokum Weeratunga , Cagatay Goncu

In many competitive settings, from education to politics, rules do not reward effort evenly, and thresholds (e.g., grade cutoffs or electoral majorities) make some moments disproportionately important. Success thus depends on efficiently…

计算机科学与博弈论 · 计算机科学 2026-01-23 Masatsugu Yoshizawa , Yuta Kawamoto , Daisuke Takeshita

This paper examines a trade execution game for two large traders in a generalized price impact model. We incorporate a stochastic and sequentially dependent factor that exogenously affects the market price into financial markets. Our model…

交易与市场微观结构 · 定量金融 2024-05-14 Masamitsu Ohnishi , Makoto Shimoshimizu

In this paper, we model one-day international cricket games as Markov processes, applying forward and inverse Reinforcement Learning (RL) to develop three novel tools for the game. First, we apply Monte-Carlo learning to fit a nonlinear…

机器学习 · 计算机科学 2021-03-09 Manohar Vohra , George S. D. Gordon

A Markov decision process can be parameterized by a transition kernel and a reward function. Both play essential roles in the study of reinforcement learning as evidenced by their presence in the Bellman equations. In our inquiry of various…

机器学习 · 计算机科学 2023-09-04 Falcon Z. Dai

The dynamics in games involving multiple players, who adaptively learn from their past experience, is not yet well understood. We analyzed a class of stochastic games with Markov strategies in which players choose their actions…

概率论 · 数学 2018-04-30 Shohei Hidaka

We study how individuals trade off outcome ("what") and process ("how") utility in high-stakes strategic decisions, namely professional tennis. Using optimality conditions and the second-service rule, we derive a sufficient condition for…

计量经济学 · 经济学 2026-05-25 Arnaud Dupuy

To take advantage of strategy commitment, a useful tactic of playing games, a leader must learn enough information about the follower's payoff function. However, this leaves the follower a chance to provide fake information and influence…

计算机科学与博弈论 · 计算机科学 2023-06-14 Yurong Chen , Xiaotie Deng , Yuhao Li

We consider a discrete-time Markov decision process with Borel state and action spaces. The performance criterion is to maximize a total expected {utility determined by unbounded return function. It is shown the existence of optimal…

概率论 · 数学 2018-10-08 François Dufour , Alexandre Genadot

Markov decision processes (MDPs) are standard models for probabilistic systems with non-deterministic behaviours. Mean payoff (or long-run average reward) provides a mathematically elegant formalism to express performance related…

性能 · 计算机科学 2017-09-08 Jan Křetínský , Tobias Meggendorfer

For decades, National Football League (NFL) coaches' observed fourth down decisions have been largely inconsistent with prescriptions based on statistical models. In this paper, we develop a framework to explain this discrepancy using an…

应用统计 · 统计学 2026-03-06 Nathan Sandholtz , Lucas Wu , Martin Puterman , Timothy C. Y. Chan

AI systems are increasingly used to assist humans in sequential decision-making tasks, yet determining when and how an AI assistant should intervene remains a fundamental challenge. A potential baseline is to recommend the optimal action…

人工智能 · 计算机科学 2026-04-17 Saumik Narayanan , Raja Panjwani , Siddhartha Sen , Chien-Ju Ho

Many works in the domain of artificial intelligence in games focus on board or video games due to the ease of reimplementing their mechanics. Decision-making problems in real-world sports share many similarities to such domains.…

人工智能 · 计算机科学 2024-08-13 Carlo Nübel , Alexander Dockhorn , Sanaz Mostaghim

We consider reinforcement learning in changing Markov Decision Processes where both the state-transition probabilities and the reward functions may vary over time. For this problem setting, we propose an algorithm using a sliding window…

机器学习 · 计算机科学 2018-05-28 Pratik Gajane , Ronald Ortner , Peter Auer

In dynamic programming and reinforcement learning, the policy for the sequential decision making of an agent in a stochastic environment is usually determined by expressing the goal as a scalar reward function and seeking a policy that…

人工智能 · 计算机科学 2025-02-26 Simon Dima , Simon Fischer , Jobst Heitzig , Joss Oliver

This paper investigates value function approximation in the context of zero-sum Markov games, which can be viewed as a generalization of the Markov decision process (MDP) framework to the two-agent case. We generalize error bounds from MDPs…

人工智能 · 计算机科学 2013-01-07 Michail Lagoudakis , Ron Parr

Determining the value of basketball players through analyzing the players' behavior is important for the managers of modern basketball teams. However, conventional methods always utilize isolated statistical data, leading to ineffective and…

社会与信息网络 · 计算机科学 2021-01-01 Xin Du , Weihong Cai , Jianquan Liu , Ding Yu , Kai Xu , Wei Li

Complex interactions between two opposing agents frequently occur in domains of machine learning, game theory, and other application domains. Quantitatively analyzing the strategies involved can provide an objective basis for…

机器学习 · 计算机科学 2023-07-28 Calvin C. K. Yeung , Keisuke Fujii

Planning problems where effects of actions are non-deterministic can be modeled as Markov decision processes. Planning problems are usually goal-directed. This paper proposes several techniques for exploiting the goal-directedness to…

人工智能 · 计算机科学 2013-02-08 Nevin Lianwen Zhang , Weihong Zhang
‹ 上一页 1 2 3 10 下一页 ›