中文
相关论文

相关论文: Self-optimization in distributed manufacturing sys…

200 篇论文

In this paper, we propose a hierarchical game approach to model the energy efficiency maximization problem where transmitters individually choose their channel assignment and power control. We conduct a thorough analysis of the existence,…

信息论 · 计算机科学 2016-11-15 Majed Haddad , Piotr Wiecek , Oussama Habachi , Yezekael Hayel

Demand-side management (DSM) enables distribution system operators (DSOs) to steer electricity consumption through dynamic price signals or incentive mechanisms, thereby leveraging end-users' flexibility potential for delivering grid…

最优化与控制 · 数学 2026-05-04 Silvia Cianchi , Reza Rahimi Baghbadorani , Anibal Sanjab , Sergio Grammatico

This work investigates the formal policy synthesis of continuous-state stochastic dynamic systems given high-level specifications in linear temporal logic. To learn an optimal policy that maximizes the satisfaction probability, we take a…

人工智能 · 计算机科学 2023-04-21 Lening Li , Zhentian Qian

Adversarial decision-making in partially observable multi-agent systems requires sophisticated strategies for both deception and counter-deception. This paper presents a sequential hypothesis testing (SHT)-driven framework that captures the…

最优化与控制 · 数学 2026-04-14 Haosheng Zhou , Daniel Ralston , Xu Yang , Ruimeng Hu

We study the problem of computing Stackelberg equilibria Stackelberg games whose underlying structure is in congestion games, focusing on the case where each player can choose a single resource (a.k.a. singleton congestion games) and one of…

计算机科学与博弈论 · 计算机科学 2018-08-31 Matteo Castiglioni , Alberto Marchesi , Nicola Gatti , Stefano Coniglio

In this paper, a new offline actor-critic learning algorithm is introduced: Sampled Policy Gradient (SPG). SPG samples in the action space to calculate an approximated policy gradient by using the critic to evaluate the samples. This…

人工智能 · 计算机科学 2018-09-18 Anton Orell Wiehe , Nil Stolt Ansó , Madalina M. Drugan , Marco A. Wiering

This paper focuses on multi-stage coordination for a population of thermostatically controlled loads (TCL). Each load maximizes the individual utility in response to an energy price, while the coordinator determines the price to maximize…

最优化与控制 · 数学 2016-08-09 Sen Li , Wei Zhang , Jianming Lian , Karanjit Kalsi

Competitive games involving thousands or even millions of players are prevalent in real-world contexts, such as transportation, communications, and computer networks. However, learning in these large-scale multi-agent environments presents…

最优化与控制 · 数学 2025-02-04 Batuhan Yardim , Semih Cayci , Niao He

This study employs gamified experiments to investigate and refine the Schelling Model of Segregation, a framework that demonstrates how individual preferences can lead to systemic segregation. Using a movement selection algorithm derived…

物理与社会 · 物理学 2025-01-15 Aleix Nicolás Olivé , Luce Prignano , Dimitri Marinelli , Emanuele Cozzo

Real-time strategy games have been an important field of game artificial intelligence in recent years. This paper presents a reinforcement learning and curriculum transfer learning method to control multiple units in StarCraft…

人工智能 · 计算机科学 2018-04-04 Kun Shao , Yuanheng Zhu , Dongbin Zhao

This paper studies a nonlinear open-loop mean field Stackelberg stochastic differential game by using the probabilistic method through the FBSDE system and the idea of taking control as the fixed point. We successively construct the…

最优化与控制 · 数学 2026-01-08 Jianhui Huang , Qi Huang

This paper is concerned with a three-level multi-leader-follower incentive Stackelberg game with $H_\infty$ constraint. Based on $H_2/H_\infty$ control theory, we firstly obtain the worst-case disturbance and the team-optimal strategy by…

最优化与控制 · 数学 2024-12-13 Na Xiang , Jingtao Shi

Systems engineering processes coordinate the effort of different individuals to generate a product satisfying certain requirements. As the involved engineers are self-interested agents, the goals at different levels of the systems…

多智能体系统 · 计算机科学 2023-07-19 Salar Safarkhani , Ilias Bilionis , Jitesh Panchal

The rapid progression of sophisticated advance metering infrastructure (AMI), allows us to have a better understanding and data from demand-response (DR) solutions. There are vast amounts of research on the internet of things and its…

信号处理 · 电气工程与系统科学 2019-08-09 Ramin Faraji Fijani , Behrouz Azimian , Ehsan Ghotbi , Xingwu Wang

Multi-robot cooperation requires agents to make decisions that are consistent with the shared goal without disregarding action-specific preferences that might arise from asymmetry in capabilities and individual objectives. To accomplish…

机器人学 · 计算机科学 2021-05-07 Joewie J. Koh , Guohui Ding , Christoffer Heckman , Lijun Chen , Alessandro Roncone

We discuss an open-loop backward Stackelberg differential game involving single leader and single follower. Unlike most Stackelberg game literature, the state to be controlled is characterized by a backward stochastic differential equation…

最优化与控制 · 数学 2021-04-06 Xinwei Feng , Ying Hu , Jianhui Huang

Information uncertainty is one of the major challenges facing applications of game theory. In the context of Stackelberg games, various approaches have been proposed to deal with the leader's incomplete knowledge about the follower's…

计算机科学与博弈论 · 计算机科学 2019-05-21 Jiarui Gan , Haifeng Xu , Qingyu Guo , Long Tran-Thanh , Zinovi Rabinovich , Michael Wooldridge

Governments are motivated to subsidize profit-driven firms that manufacture zero-emission vehicles to ensure they become price-competitive. This paper introduces a dynamic Stackelberg game to determine the government's optimal subsidy…

最优化与控制 · 数学 2025-10-21 Utsav Sadana , Georges Zaccour

We study a two-player Stackelberg game with incomplete information such that the follower's strategy belongs to a known family of parameterized functions with an unknown parameter vector. We design an adaptive learning approach to…

计算机科学与博弈论 · 计算机科学 2021-01-12 Guosong Yang , Radha Poovendran , João P. Hespanha

Stackelberg games originate where there are market leaders and followers, and the actions of leaders influence the behavior of the followers. Mathematical modelling of such games results in what's called a Bilevel Optimization problem.…

计算机科学与博弈论 · 计算机科学 2023-12-07 Pravesh Koirala , Forrest Laine