中文
相关论文

相关论文: Reactive Power Compensation Game under Prospect-Th…

200 篇论文

This paper considers an online reinforcement learning algorithm that leverages pre-collected data (passive memory) from the environment for online interaction. We show that using passive memory improves performance and further provide…

机器学习 · 计算机科学 2024-10-21 Anay Pattanaik , Lav R. Varshney

A simple model for cooperation between "selfish" agents, which play an extended version of the Prisoner's Dilemma(PD) game, in which they use arbitrary payoffs, is presented and studied. A continuous variable, representing the probability…

凝聚态物理 · 物理学 2009-11-10 H. Fort

Understanding how individual agents make strategic decisions within collectives is important for advancing fields as diverse as economics, neuroscience, and multi-agent systems. Two complementary approaches can be integrated to this end.…

多智能体系统 · 计算机科学 2025-05-21 Jaime Ruiz-Serra , Patrick Sweeney , Michael S. Harré

Network congestion games are a well-understood model of multi-agent strategic interactions. Despite their ubiquitous applications, it is not clear whether it is possible to design information structures to ameliorate the overall experience…

计算机科学与博弈论 · 计算机科学 2020-02-14 Matteo Castiglioni , Andrea Celli , Alberto Marchesi , Nicola Gatti

The problem of dynamic pricing of electricity in a retail market is considered. A Stackelberg game is used to model interactions between a retailer and its customers; the retailer sets the day-ahead hourly price of electricity and consumers…

最优化与控制 · 数学 2016-03-01 Liyan Jia , Lang Tong

Different from shopping at retail stores, consumers on e-commerce platforms usually cannot touch or try products before purchasing, which means that they have to make decisions when they are uncertain about the outcome (e.g., satisfaction…

信息检索 · 计算机科学 2020-08-20 Zhichao Xu , Yi Han , Yongfeng Zhang , Qingyao Ai

We consider the problem of using logged data to make predictions about what would happen if we changed the `rules of the game' in a multi-agent system. This task is difficult because in many cases we observe actions individuals take but not…

计算机科学与博弈论 · 计算机科学 2019-04-05 Alexander Peysakhovich , Christian Kroer , Adam Lerer

The theory of direct reciprocity explores how individuals cooperate when they interact repeatedly. In repeated interactions, individuals can condition their behaviour on what happened earlier. One prominent example of a conditional strategy…

动力系统 · 数学 2024-11-21 Nataliya A. Balabanova , Manh Hong Duong , Christian Hilbe

Cooperation is usually represented as a Prisoner's Dilemma game. Although individual self-interest may not favour cooperation, cooperation can evolve if, for example, players interact multiple times adjusting their behaviour accordingly to…

物理与社会 · 物理学 2015-04-29 Elton J. S. Júnior , Lucas Wardil , Jafferson K. L. da Silva

In strategic games such as the prisoner's dilemma, allowing players to make binding offers of utility transfers before play has been shown to alter incentives and potentially support cooperative outcomes. These preplay exchange mechanisms…

理论经济学 · 经济学 2026-04-27 Ian Fligler

Demand response is widely employed by today's data centers to reduce energy consumption in response to the increasing of electricity cost. To incentivize users of data centers participate in the demand response programs, i.e., breaking the…

分布式、并行与集群计算 · 计算机科学 2016-04-08 Yong Zhan , Du Xu , Hongfang Yu , Shui Yu

Energy storage and demand-side response will play an increasingly important role in the future electricity system. We extend previous results on a single energy storage unit to the management of two energy storage units cooperating for the…

最优化与控制 · 数学 2020-05-25 Miguel F. Anjos , James R. Cruise , Albert Solà Vilalta

The dynamic pricing of electricity is one of the most crucial demand response (DR) strategies in smart grid, where the utility company typically adjust electricity prices to influence user electricity demand. This paper models the…

最优化与控制 · 数学 2024-07-16 Jiangjiang Cheng , Ge Chen , Zhouming Wu , Yifen Mu

Crowd simulation is important for video-games design, since it enables to populate virtual worlds with autonomous avatars that navigate in a human-like manner. Reinforcement learning has shown great potential in simulating virtual crowds,…

机器学习 · 计算机科学 2023-09-25 Ariel Kwiatkowski , Vicky Kalogeiton , Julien Pettré , Marie-Paule Cani

We consider preference communication in two-player multi-objective normal-form games. In such games, the payoffs resulting from joint actions are vector-valued. Taking a utility-based approach, we assume there exists a utility function for…

计算机科学与博弈论 · 计算机科学 2022-06-13 Willem Röpke , Diederik M. Roijers , Ann Nowé , Roxana Rădulescu

Underlying relationships among Multi-Agent Systems (MAS) in hazardous scenarios can be represented as Game-theoretic models. This paper proposes a new hierarchical network-based model called Game-theoretic Utility Tree (GUT), which…

多智能体系统 · 计算机科学 2023-03-30 Qin Yang , Ramviyas Parasuraman

Cooperation between self-interested individuals is a widespread phenomenon in the natural world, but remains elusive in interactions between artificially intelligent agents. Instead, naive reinforcement learning algorithms typically…

多智能体系统 · 计算机科学 2025-01-16 John L. Zhou , Weizhe Hong , Jonathan C. Kao

This work studies the decentralized and uncoordinated energy source selection problem for smart-grid consumers with heterogeneous energy profiles and risk attitudes: they compete for a limited amount of renewable energy in their local…

系统与控制 · 电气工程与系统科学 2022-03-16 Eleni Stai , Evangelia Kokolaki , Lesia Mitridati , Petros Tatoulis , Ioannis Stavrakakis , Gabriela Hug

Uncoordinated charging of a rapidly growing number of electric vehicles (EVs) and the uncertainty associated with renewable energy resources may constitute a critical issue for the electric mobility (E-Mobility) in the transportation system…

最优化与控制 · 数学 2020-06-30 Hwei-Ming Chung , Sabita Maharjan , Yan Zhang , Frank Eliassen

Strategic decision-making involves interactive reasoning where agents adapt their choices in response to others, yet existing evaluations of large language models (LLMs) often emphasize Nash Equilibrium (NE) approximation, overlooking the…

人工智能 · 计算机科学 2025-11-04 Jingru Jia , Zehua Yuan , Junhao Pan , Paul E. McNamara , Deming Chen
‹ 上一页 1 8 9 10 下一页 ›