中文
相关论文

相关论文: Self-optimization in distributed manufacturing sys…

200 篇论文

This paper studies multi-agent reinforcement learning with submodular team utilities for online distributed task allocation. In this setting, each agent selects one action from a local categorical policy, so feasible joint actions form a…

系统与控制 · 电气工程与系统科学 2026-05-14 Jing Liu , Yangyang Yang , Luca Ballotta , Fangfei Li , Yang Tang , Ruggero Carli

We study how to synthesize a robust and safe policy for autonomous systems under signal temporal logic (STL) tasks in adversarial settings against unknown dynamic agents. To ensure the worst-case STL satisfaction, we propose STLGame, a…

机器人学 · 计算机科学 2024-12-03 Shuo Yang , Hongrui Zheng , Cristian-Ioan Vasile , George Pappas , Rahul Mangharam

Power producers can exhibit strategic behavior in electricity markets to maximize their profits. This behavior is more pronounced with the deregulation of distribution markets, which offers an opportunity for profit arbitrage between…

系统与控制 · 电气工程与系统科学 2020-10-19 Hafiz Anwar Ullah Khan , Jip Kim , Yury Dvorkin

As predictive models are deployed into the real world, they must increasingly contend with strategic behavior. A growing body of work on strategic classification treats this problem as a Stackelberg game: the decision-maker "leads" in the…

机器学习 · 计算机科学 2022-02-01 Tijana Zrnic , Eric Mazumdar , S. Shankar Sastry , Michael I. Jordan

Classical game-theoretic approaches for multi-agent systems in both the forward policy design problem and the inverse reward learning problem often make strong rationality assumptions: agents perfectly maximize expected utilities under…

机器学习 · 计算机科学 2021-03-23 Ran Tian , Liting Sun , Masayoshi Tomizuka

One stream of reinforcement learning research is exploring biologically plausible models and algorithms to simulate biological intelligence and fit neuromorphic hardware. Among them, reward-modulated spike-timing-dependent plasticity…

神经与进化计算 · 计算机科学 2022-10-25 Zhile Yang , Shangqi Guo , Ying Fang , Jian K. Liu

A key challenge in multi-agent systems is the design of intelligent agents solving real-world tasks in close interaction with other agents (e.g. humans), thereby being confronted with a variety of behavioral variations and limited knowledge…

多智能体系统 · 计算机科学 2020-07-13 Julian Bernhard , Alois Knoll

In today's uncertain and competitive market, where enterprises are subjected to increasingly shortened product life-cycles and frequent volume changes, reconfigurable manufacturing systems (RMS) applications play a significant role in the…

系统与控制 · 电气工程与系统科学 2023-01-02 Carlos Alberto Barrera-Diaz , Amir Nourmohammdi , Henrik Smedberg , Tehseen Aslam , Amos H. C. Ng

This work adopts the very successful distributional perspective on reinforcement learning and adapts it to the continuous control setting. We combine this within a distributed framework for off-policy learning in order to develop what we…

This paper considers the challenging tasks of Multi-Agent Reinforcement Learning (MARL) under partial observability, where each agent only sees her own individual observations and actions that reveal incomplete information about the…

机器学习 · 计算机科学 2022-10-18 Qinghua Liu , Csaba Szepesvári , Chi Jin

LLM self-play algorithms are notable in that, in principle, nothing bounds their learning: a Conjecturer model creates problems for a Solver, and both improve together. However, in practice, existing LLM self-play methods do not scale well…

机器学习 · 计算机科学 2026-04-23 Luke Bailey , Kaiyue Wen , Kefan Dong , Tatsunori Hashimoto , Tengyu Ma

We study a Stackelberg game to examine how two agents determine to cooperate while competing with each other. Each selects an arrival time to a destination, the earlier one fetching a higher reward. There is, however, an inherent penalty in…

计算机科学与博弈论 · 计算机科学 2024-07-30 Chenlan Wang , Mehrdad Moharrami , Mingyan Liu

Multi-objective optimization aims to solve problems with competing objectives. Evaluating such problems is often slow or expensive, limiting the budget of evaluations. In many applications, historical data from related optimization tasks is…

机器学习 · 计算机科学 2026-05-12 Leonard Papenmeier , Petru Tighineanu

This paper studies a multi-period demand response management problem in the smart grid where multiple utility companies compete among themselves. The user-utility interactions are modeled by a noncooperative game of a Stackelberg type where…

最优化与控制 · 数学 2016-11-17 Khaled Alshehri , Ji Liu , Xudong Chen , Tamer Başar

Optimizing strategic decisions (a.k.a. computing equilibrium) is key to the success of many non-cooperative multi-agent applications. However, in many real-world situations, we may face the exact opposite of this game-theoretic problem --…

计算机科学与博弈论 · 计算机科学 2022-10-05 Jibang Wu , Weiran Shen , Fei Fang , Haifeng Xu

This paper studies the dynamic pricing mechanism for data products in demand-driven markets through a game-theoretic framework. We develop a three-tier Stackelberg game model to capture the hierarchical strategic interactions among key…

最优化与控制 · 数学 2025-12-29 Lijun Bo , Dongfang Yang , Shihua Wang

Decentralised learning enables the training of deep learning algorithms without centralising data sets, resulting in benefits such as improved data privacy, operational efficiency and the fostering of data ownership policies. However,…

机器学习 · 计算机科学 2024-12-23 Sebastian Niehaus , Ingo Roeder , Nico Scherf

Federated learning (FL) rests on the notion of training a global model in a decentralized manner. Under this setting, mobile devices perform computations on their local data before uploading the required updates to improve the global model.…

机器学习 · 计算机科学 2020-05-07 Shashi Raj Pandey , Nguyen H. Tran , Mehdi Bennis , Yan Kyaw Tun , Aunas Manzoor , Choong Seon Hong

This paper is concerned with a Stackelberg game of backward stochastic differential equations (BSDEs), where the coefficients of the backward system and the cost functionals are deterministic, and the control domain is convex. Necessary and…

最优化与控制 · 数学 2019-04-18 Yueyang Zheng , Jingtao Shi

We extend the formalism of Conjectural Variations games to Stackelberg games involving multiple leaders and a single follower. To solve these nonconvex games, a common assumption is that the leaders compute their strategies having perfect…

计算机科学与博弈论 · 计算机科学 2025-07-24 Francesco Morri , Hélène Le Cadre , Luce Brotcorne
‹ 上一页 1 8 9 10 下一页 ›