中文
相关论文

相关论文: Generating Teammates for Training Robust Ad Hoc Te…

200 篇论文

Modern unmanned systems, including aerial, terrestrial, and underwater vehicles, are increasingly utilized in dynamic and unpredictable environments, where the presence of modeling uncertainties necessitates the development of robust and…

系统与控制 · 电气工程与系统科学 2026-03-11 Maryam Norouzi , Mingxi Zhou , Chengzhi Yuan

Social dilemmas have been widely studied to explain how humans are able to cooperate in society. Considerable effort has been invested in designing artificial agents for social dilemmas that incorporate explicit agent motivations that are…

多智能体系统 · 计算机科学 2021-08-30 Nicolas Anastassacos , Stephen Hailes , Mirco Musolesi

Robust Policy Search is the problem of learning policies that do not degrade in performance when subject to unseen environment model parameters. It is particularly relevant for transferring policies learned in a simulation environment to…

机器学习 · 计算机科学 2021-11-23 Sai Kiran Narayanaswami , Nandan Sudarsanam , Balaraman Ravindran

The complexity and scale of IT systems are increasing dramatically, posing many challenges to real-world anomaly detection. Deep learning anomaly detection has emerged, aiming at feature learning and anomaly scoring, which has gained…

机器学习 · 计算机科学 2023-12-05 Xue Yang , Enda Howley , Micheal Schukat

In multi-agent reinforcement learning, centralized training with decentralized execution (CTDE) methods typically assume that agents make decisions based on their local observations independently, which may not lead to a correlated joint…

多智能体系统 · 计算机科学 2024-12-16 Zhiyuan Li , Wenshuai Zhao , Lijun Wu , Joni Pajarinen

We propose an improved algorithm by identifying and encouraging cooperative behavior in multi-agent environments. First, we analyze the shortcomings of existing algorithms in addressing multi-agent reinforcement learning problems. Then,…

多智能体系统 · 计算机科学 2025-08-21 Junjie Qi , Siqi Mao , Tianyi Tan

In dynamic urban logistics, the stochastic emergence of time-sensitive tasks poses a significant optimality challenge for heterogeneous AAVs logistics task allocation. To address this problem, a reinforcement learning enhanced overlapping…

机器人学 · 计算机科学 2026-05-27 Yuze Zhou , Jingliang Sun , Junzhi Li , Jianxin Zhong , Zihan Wang , Teng Long

It is widely known how the human ability to cooperate has influenced the thriving of our species. However, as we move towards a hybrid human-machine future, it is still unclear how the introduction of AI agents in our social interactions…

计算机与社会 · 计算机科学 2022-05-16 Inês Terrucha , Elias Fernández Domingos , Francisco C. Santos , Pieter Simoens , Tom Lenaerts

The purpose of offline multi-task reinforcement learning (MTRL) is to develop a unified policy applicable to diverse tasks without the need for online environmental interaction. Recent advancements approach this through sequence modeling,…

机器学习 · 计算机科学 2024-11-05 Ziqing Fan , Shengchao Hu , Yuhang Zhou , Li Shen , Ya Zhang , Yanfeng Wang , Dacheng Tao

This paper develops the first policy gradient method with global optimality guarantee and complexity analysis for robust reinforcement learning under model mismatch. Robust reinforcement learning is to learn a policy robust to model…

机器学习 · 计算机科学 2022-05-17 Yue Wang , Shaofeng Zou

Many cooperative multi-agent problems require agents to learn individual tasks while contributing to the collective success of the group. This is a challenging task for current state-of-the-art multi-agent reinforcement algorithms that are…

多智能体系统 · 计算机科学 2020-03-25 Hassam Ullah Sheikh , Ladislau Bölöni

Project-based learning improves student engagement and learning outcomes, yet allocating students to appropriately challenging projects while forming cognitively diverse teams remains difficult at scale. Traditional allocation methods…

软件工程 · 计算机科学 2026-05-06 Dhruv Gulwani , Basem Suleiman , Aditya Joshi , Sonit Singh

Decision trees are a popular choice of explainable model, but just like neural networks, they suffer from adversarial examples. Existing algorithms for fitting decision trees robust against adversarial examples are greedy heuristics and…

机器学习 · 计算机科学 2021-09-10 Daniël Vos , Sicco Verwer

As AI technology continues to develop, more and more agents will become capable of long term autonomy alongside people. Thus, a recent line of research has studied the problem of teaching autonomous agents the concept of ethics and human…

计算机与社会 · 计算机科学 2019-05-03 Shani Alkoby , Avilash Rath , Peter Stone

Training autonomous agents with sparse rewards is a long-standing problem in online reinforcement learning (RL), due to low data efficiency. Prior work overcomes this challenge by extracting useful knowledge from offline data, often…

机器学习 · 计算机科学 2024-06-07 Qianlan Yang , Yu-Xiong Wang

Adaptive teaming-the capability of agents to effectively collaborate with unfamiliar teammates without prior coordination-is widely explored in virtual video games but overlooked in real-world multi-robot contexts. Yet, such adaptive…

机器人学 · 计算机科学 2025-05-05 Yang Li , Junfan Chen , Feng Xue , Jiabin Qiu , Wenbin Li , Qingrui Zhang , Ying Wen , Wei Pan

Designing reward functions for efficiently guiding reinforcement learning (RL) agents toward specific behaviors is a complex task. This is challenging since it requires the identification of reward structures that are not sparse and that…

机器学习 · 计算机科学 2023-11-01 Dhawal Gupta , Yash Chandak , Scott M. Jordan , Philip S. Thomas , Bruno Castro da Silva

Multi-Agent Task Assignment and Planning (MATP) has attracted growing attention but remains challenging in terms of scalability, spatial reasoning, and adaptability in obstacle-rich environments. To address these challenges, we propose OATH…

机器人学 · 计算机科学 2026-04-09 Nan Li , Jiming Ren , Haris Miller , Samuel Coogan , Karen M. Feigh , Ye Zhao

Effective teamwork is essential across diverse domains. During the team formation stage, a key challenge is forming teams that effectively balance user preferences with task objectives to enhance overall team satisfaction. In the team…

人机交互 · 计算机科学 2025-06-09 Mohammed Almutairi

We address the problem of maintaining resource availability in a networked multi-robot team performing distributed tracking of unknown number of targets in an environment of interest. Based on our model, robots are equipped with sensing and…

机器人学 · 计算机科学 2020-04-16 Ragesh K. Ramachandran , Nicole Fronda , Gaurav S. Sukhatme