中文
相关论文

相关论文: Entropy Non-increasing Games for the Improvement o…

200 篇论文

Designing agents that are able to achieve different play-styles while maintaining a competitive level of play is a difficult task, especially for games for which the research community has not found super-human performance yet, like…

The analysis of network routing games typically assumes, right at the onset, precise and detailed information about the latency functions. Such information may, however, be unavailable or difficult to obtain. Moreover, one is often…

计算机科学与博弈论 · 计算机科学 2014-08-08 Umang Bhaskar , Katrina Ligett , Leonard J. Schulman , Chaitanya Swamy

Entropy-based objectives are widely used to perform state space exploration in reinforcement learning (RL) and dataset generation for offline RL. Behavioral entropy (BE), a rigorous generalization of classical entropies that incorporates…

机器学习 · 计算机科学 2025-02-07 Wesley A. Suttle , Aamodh Suresh , Carlos Nieto-Granda

In this paper we introduce a novel flow representation for finite games in strategic form. This representation allows us to develop a canonical direct sum decomposition of an arbitrary game into three components, which we refer to as the…

计算机科学与博弈论 · 计算机科学 2015-03-17 Ozan Candogan , Ishai Menache , Asuman Ozdaglar , Pablo A. Parrilo

Traditional game-theoretic research for security applications primarily focuses on the allocation of external protection resources to defend targets. This work puts forward the study of a new class of games centered around strategically…

计算机科学与博弈论 · 计算机科学 2024-10-29 Niclas Boehmer , Minbiao Han , Haifeng Xu , Milind Tambe

The advantages of event-sensing over conventional sensors (e.g., higher dynamic range, lower time latency, and lower power consumption) have spurred research into machine learning for event data. Unsurprisingly, deep learning has emerged as…

机器学习 · 计算机科学 2021-06-11 Fuqiang Gu , Weicong Sng , Xuke Hu , Fangwen Yu

Estimating the unknown reward functions driving agents' behaviors is of central interest in inverse reinforcement learning and game theory. To tackle this problem, we develop a unified framework for reward function recovery in two-player…

机器学习 · 计算机科学 2026-05-20 Junyi Liao , Zihan Zhu , Ethan Fang , Zhuoran Yang , Vahid Tarokh

Applying neural network (NN) methods in games can lead to various new and exciting game dynamics not previously possible. However, they also lead to new challenges such as the lack of large, clean datasets, varying player skill levels, and…

机器学习 · 计算机科学 2021-07-06 Mathias Löwe , Jennifer Villareale , Evan Freed , Aleksanteri Sladek , Jichen Zhu , Sebastian Risi

Online computation is a concept to model uncertainty where not all information on a problem instance is known in advance. An online algorithm receives requests which reveal the instance piecewise and has to respond with irrevocable…

计算复杂性 · 计算机科学 2023-11-28 Janosch Fuchs , Christoph Grüne , Tom Janßen

Petri games have been introduced as a multi-player game model representing causal memory to address the synthesis of distributed systems. For Petri games with one environment player and an arbitrary bounded number of system players,…

计算机科学与博弈论 · 计算机科学 2019-04-12 Manuel Gieseking , Ernst-Rüdiger Olderog

In Multi-Goal Reinforcement Learning, an agent learns to achieve multiple goals with a goal-conditioned policy. During learning, the agent first collects the trajectories into a replay buffer, and later these trajectories are selected…

机器学习 · 计算机科学 2020-05-26 Rui Zhao , Xudong Sun , Volker Tresp

For over a decade now, robotics and the use of artificial agents have become a common thing.Testing the performance of new path finding or search space optimization algorithms has also become a challenge as they require simulation or an…

机器学习 · 计算机科学 2022-07-29 Jerin Paul Selvan , Pravin S. Game

The game industry is moving into an era where old-style game engines are being replaced by re-engineered systems with embedded machine learning technologies for the operation, analysis and understanding of game play. In this paper, we…

计算机与社会 · 计算机科学 2021-01-05 Yilei Zeng , Aayush Shah , Jameson Thai , Michael Zyda

We explore the use of policy approximations to reduce the computational cost of learning Nash equilibria in zero-sum stochastic games. We propose a new Q-learning type algorithm that uses a sequence of entropy-regularized soft policies to…

机器学习 · 计算机科学 2021-06-29 Yue Guan , Qifan Zhang , Panagiotis Tsiotras

Policy makers focus on stable strategies as the ones adopted by rational players. If there are many such solutions an important question is how to select amongst them. We study this question for the Multicommodity Flow Coalition Game, used…

计算机科学与博弈论 · 计算机科学 2020-01-29 Coulter Beeson , Bruce Shepherd

Large Language Models (LLMs) define probability measures on text. By considering the implicit knowledge question of what it means for an LLM to know such a measure and what it entails algorithmically, we are naturally led to formulate a…

人工智能 · 计算机科学 2025-06-24 Clément Hongler , Andrew Emil

Modern vision generators transport a base distribution to data through time-indexed measures, implemented as deterministic flows (ODEs) or stochastic diffusions (SDEs). Despite strong empirical performance, standard flow-matching objectives…

机器学习 · 计算机科学 2026-02-27 Chika Maduabuchi

Advanced Persistent Threats (APTs) are stealthy attacks that threaten the security and privacy of sensitive information. Interactions of APTs with victim system introduce information flows that are recorded in the system logs. Dynamic…

最优化与控制 · 数学 2021-06-29 Dinuka Sahabandu , Shana Moothedath , Joey Allen , Linda Bushnell , Wenke Lee , Radha Poovendran

Collaborative multiple robots for unknown environment exploration have become mainstream due to their remarkable performance and efficiency. However, most existing methods assume perfect robots' communication during exploration, which is…

机器人学 · 计算机科学 2025-05-30 Khattiya Pongsirijinda , Zhiqiang Cao , Billy Pik Lik Lau , Ran Liu , Chau Yuen , U-Xuan Tan

The single- and multi- processor cup games can be used to model natural problems in areas such as processor scheduling, deamortization, and buffer management. At the beginning of the single-processor cup game, $n$ cups are initially empty.…

数据结构与算法 · 计算机科学 2019-04-08 Michael A. Bender , Martin Farach-Colton , William Kuszmaul