中文
相关论文

相关论文: Performance Evaluation of a Multi-Agent Risk-Sensi…

200 篇论文

This paper addresses the problem of formation control and tracking a of desired trajectory by an Euler-Lagrange multi-agent systems. It is inspired by recent results by Qingkai et al. and adopts an event-triggered control strategy to reduce…

系统与控制 · 计算机科学 2018-05-31 Christophe Viel , Sylvain Bertrand , Michel Kieffer , Hélène Piet-Lahanier

Reinforcement learning in multi-agent scenarios is important for real-world applications but presents challenges beyond those seen in single-agent settings. We present an actor-critic algorithm that trains decentralized policies in…

机器学习 · 计算机科学 2019-05-29 Shariq Iqbal , Fei Sha

We consider the problem of robust and adaptive model predictive control (MPC) of a linear system, with unknown parameters that are learned along the way (adaptive), in a critical setting where failures must be prevented (robust). This…

机器学习 · 计算机科学 2020-10-22 Edouard Leurent , Denis Efimov , Odalric-Ambrym Maillard

Markov games (MGs) and multi-agent reinforcement learning (MARL) are studied to model decision making in multi-agent systems. Traditionally, the objective in MG and MARL has been risk-neutral, i.e., agents are assumed to optimize a…

计算机科学与博弈论 · 计算机科学 2024-06-11 Hafez Ghaemi , Shirin Jamshidi , Mohammad Mashreghi , Majid Nili Ahmadabadi , Hamed Kebriaei

We propose a simple model that describes the dynamics of efficiencies of competing agents. Agents communicate leading to increase of efficiencies of underachievers, and an efficiency of each agent can increase or decrease irrespectively of…

统计力学 · 物理学 2009-10-31 S. N. Majumdar , P. L. Krapivsky

AI systems have become increasingly capable of dangerous behaviours in many domains. This raises the question: Do models sometimes choose to violate human instructions in order to perform behaviour that is more useful for certain goals? We…

人工智能 · 计算机科学 2026-05-08 Jonas Wiedermann-Möller , Leonard Dung , Maksym Andriushchenko

Gradual advancement of control technology gives rise to the studies of the stability of linear systems. The stability of the linear multiagent system is motivated by increasing utilization of agent dynamics together with the number of…

系统与控制 · 电气工程与系统科学 2023-08-22 Gurmu Meseret Debele

Multi-agent safe systems have become an increasingly important area of study as we can now easily have multiple AI-powered systems operating together. In such settings, we need to ensure the safety of not only each individual agent, but…

人工智能 · 计算机科学 2021-03-08 Zheqing Zhu , Erdem Bıyık , Dorsa Sadigh

In Probabilistic Risk Management, risk is characterized by two quantities: the magnitude (or severity) of the adverse consequences that can potentially result from the given activity or action, and by the likelihood of occurrence of the…

人工智能 · 计算机科学 2009-10-07 Eric Daudé , Pierrick Tranouez , Patrice Langlois

In domain-specific applications, GPT-4, augmented with precise prompts or Retrieval-Augmented Generation (RAG), shows notable potential but faces the critical tri-lemma of performance, cost, and data privacy. High performance requires…

We consider a financial market model which consists of a financial asset and a large number of interacting agents classified into many types. Different types of agents are heterogeneous in their price expectations. Each agent can change its…

概率论 · 数学 2008-12-02 Biao Wu

Learning methods are increasingly used to synthesize controllers from data, yet existing sample-complexity characterizations for continuous control are sharp only in the fully observed setting. This paper studies the partially observed case…

系统与控制 · 电气工程与系统科学 2026-05-19 Bruce D. Lee , Anastasios Tsiamis , Nikolai Matni , Manfred Morari , John Lygeros

A fundamental challenge in multiagent reinforcement learning is to learn beneficial behaviors in a shared environment with other simultaneously learning agents. In particular, each agent perceives the environment as effectively…

We study optimal contract design for large populations of heterogeneous agents whose actions generate network spillovers represented by an interaction function. In a linear-quadratic framework, we solve the finite-agent problem and its…

理论经济学 · 经济学 2026-05-19 Guillermo Alonso Alvarez , Erhan Bayraktar , Ibrahim Ekren

Large Language Models (LLMs) have demonstrated strong capabilities as autonomous agents through tool use, planning, and decision-making abilities, leading to their widespread adoption across diverse tasks. As task complexity grows,…

多智能体系统 · 计算机科学 2025-11-10 Ishan Kavathekar , Hemang Jain , Ameya Rathod , Ponnurangam Kumaraguru , Tanuja Ganu

We study the problem of multi-agent multi-armed bandits with adversarial corruption in a heterogeneous setting, where each agent accesses a subset of arms. The adversary can corrupt the reward observations for all agents. Agents share these…

机器学习 · 计算机科学 2024-11-14 Fatemeh Ghaffari , Xuchuang Wang , Jinhang Zuo , Mohammad Hajiesmaili

Criticality has been proposed as a key principle underlying complex behavior in biological and artificial systems; however, how criticality translates from individual dynamics to collective behavior remains unclear. We study this question…

适应与自组织系统 · 物理学 2026-05-05 Nicolas Bessone , Erwan Plantec

This paper develops a novel approach to the consensus problem of multi-agent systems by minimizing a weighted state error with neighbor agents via linear quadratic (LQ) optimal control theory. Existing consensus control algorithms only…

最优化与控制 · 数学 2024-03-19 Liping Zhang , Juanjuan Xu , Huanshui Zhang , Lihua Xie

Many multi-agent systems have the structure of a single coordinator providing behavioral or financial incentives to a large number of agents. Two challenges faced by the coordinator are a finite budget from which to allocate incentives, and…

最优化与控制 · 数学 2017-02-21 Yonatan Mintz , Anil Aswani , Philip Kaminsky , Elena Flowers , Yoshimi Fukuoka

In this work, we introduce MedAgentSim, an open-source simulated clinical environment with doctor, patient, and measurement agents designed to evaluate and enhance LLM performance in dynamic diagnostic settings. Unlike prior approaches, our…

计算与语言 · 计算机科学 2025-10-02 Mohammad Almansoori , Komal Kumar , Hisham Cholakkal