中文
相关论文

相关论文: Energy-Based Learning for Cooperative Games, with …

200 篇论文

In mean-payoff games, the objective of the protagonist is to ensure that the limit average of an infinite sequence of numeric weights is nonnegative. In energy games, the objective is to ensure that the running sum of weights is always…

计算机科学中的逻辑 · 计算机科学 2010-10-05 Krishnendu Chatterjee , Laurent Doyen , Thomas A. Henzinger , Jean-Francois Raskin

Measuring individual productivity (or equivalently distributing the overall productivity) in a network structure of workers displaying peer effects has been a subject of ongoing interest in many areas ranging from academia to industry. In…

计算机科学与博弈论 · 计算机科学 2024-02-07 N. Allouch , Luis A. Guardiola , A. Meca

Energy-Based Models (EBMs) have proven to be a highly effective approach for modelling densities on finite-dimensional spaces. Their ability to incorporate domain-specific choices and constraints into the structure of the model through…

机器学习 · 计算机科学 2023-02-24 Jen Ning Lim , Sebastian Vollmer , Lorenz Wolf , Andrew Duncan

The safety alignment of large language models (LLMs) often relies on reinforcement learning from human feedback (RLHF), which requires human annotations to construct preference datasets. Given the challenge of assigning overall quality…

计算与语言 · 计算机科学 2025-11-12 Xiaomin Li , Xupeng Chen , Jingxuan Fan , Eric Hanchen Jiang , Mingye Gao

Shapley values have become one of the go-to methods to explain complex models to end-users. They provide a model agnostic post-hoc explanation with foundations in game theory: what is the worth of a player (in machine learning, a feature…

机器学习 · 计算机科学 2023-06-21 Joran Michiels , Maarten De Vos , Johan Suykens

The learning and evaluation of energy-based latent variable models (EBLVMs) without any structural assumptions are highly challenging, because the true posteriors and the partition functions in such models are generally intractable. This…

机器学习 · 计算机科学 2021-06-08 Fan Bao , Kun Xu , Chongxuan Li , Lanqing Hong , Jun Zhu , Bo Zhang

Federated learning is a distributed learning paradigm where multiple agents, each only with access to local data, jointly learn a global model. There has recently been an explosion of research aiming not only to improve the accuracy rates…

计算机科学与博弈论 · 计算机科学 2021-06-18 Kate Donahue , Jon Kleinberg

Electric storage units constitute a key element in the emerging smart grid system. In this paper, the interactions and energy trading decisions of a number of geographically distributed storage units are studied using a novel framework…

计算机科学与博弈论 · 计算机科学 2013-10-08 Yunpeng Wang , Walid Saad , Zhu Han , H. Vincent Poor , Tamer Başar

Large language models frequently exhibit suboptimal performance on low resource languages, primarily due to inefficient subword segmentation and systemic training data imbalances. In this paper, we propose Variable Entropy Policy…

计算与语言 · 计算机科学 2026-03-20 Chonghan Liu , Yimin Du , Qi An , Xin He , Cunqi Zhai , Fei Tan , Weijia Lin , Xiaochun Gong , Yongchao Deng , Shousheng Jia , Xiangzheng Zhang

This paper studies cooperative games where coalitions are formed online and the value generated by the grand coalition must be irrevocably distributed among the players at each timestep. We investigate the fundamental issue of strategic…

计算机科学与博弈论 · 计算机科学 2025-02-28 Haris Aziz , Yuhang Guo , Zhaohong Sun

Within the task of collaborative filtering two challenges for computing conditional probabilities exist. First, the amount of training data available is typically sparse with respect to the size of the domain. Thus, support for higher-order…

信息检索 · 计算机科学 2012-07-19 Lawrence Zitnick , Takeo Kanade

Measuring the contribution of individual agents is challenging in cooperative multi-agent reinforcement learning (MARL). In cooperative MARL, team performance is typically inferred from a single shared global reward. Arguably, among the…

人工智能 · 计算机科学 2024-01-29 Omayma Mahjoub , Ruan de Kock , Siddarth Singh , Wiem Khlifi , Abidine Vall , Kale-ab Tessera , Arnu Pretorius

We consider an N-player hierarchical game in which the i-th player's objective comprises of an expectation-valued term, parametrized by rival decisions, and a hierarchical term. Such a framework allows for capturing a broad range of…

最优化与控制 · 数学 2024-01-26 Shisheng Cui , Uday V. Shanbhag , Mathias Staudigl

Energy-based models (EBMs) offer a flexible framework for parameterizing probability distributions using neural networks. However, learning EBMs by exact maximum likelihood estimation (MLE) is generally intractable, due to the need to…

机器学习 · 计算机科学 2025-08-20 Michael E. Sander , Vincent Roulet , Tianlin Liu , Mathieu Blondel

The $1-N$ generalized Stackelberg game (single-leader multi-follower game) is intricately intertwined with the interaction between a leader and followers (hierarchical interaction) and the interaction among followers (simultaneous…

计算机科学与博弈论 · 计算机科学 2023-06-12 Jaeyeon Jo , Jihwan Yu , Jinkyoo Park

Recent advances in recommender systems have shown that user-system interaction essentially formulates long-term optimization problems, and online reinforcement learning can be adopted to improve recommendation performance. The general…

信息检索 · 计算机科学 2025-02-04 Xiaobei Wang , Shuchang Liu , Qingpeng Cai , Xiang Li , Lantao Hu , Han li , Guangming Xie

$\omega$-regular energy games, which are weighted two-player turn-based games with the quantitative objective to keep the energy levels non-negative, have been used in the context of verification and synthesis. The logic of modal…

计算机科学中的逻辑 · 计算机科学 2020-10-20 Gal Amram , Shahar Maoz , Or Pistiner , Jan Oliver Ringert

Spatial public goods games model collective dilemmas where individual payoffs depend on population-level strategy configurations. Most existing studies rely on evolutionary update rules or value-based reinforcement learning methods. These…

多智能体系统 · 计算机科学 2025-12-23 Zhaoqilin Yang , Axin Xiang , Kedi Yang , Tianjun Liu , Youliang Tian

An energy community is modeled as a cooperative game, where a veto player is needed beyond the prosumers to manage the community, and the worth of a coalition is its benefit compared to the selfish behaviour of the prosumers. Properties of…

最优化与控制 · 数学 2026-02-27 Giancarlo Bigi , Davide Fioriti , Antonio Frangioni , Mauro Passacantando , Davide Poli

We develop a resource-theoretical approach that allows us to quantify values of two-player, one-round cooperative games with quantum inputs and outputs, as well as values of quantum probabilistic hypergraphs. We analyse the quantum game…

量子物理 · 物理学 2023-10-30 Jason Crann , Rupert H. Levene , Ivan G. Todorov , Lyudmila Turowska