中文
相关论文

相关论文: Distributed Online Life-Long Learning (DOL3) for M…

200 篇论文

Artificial intelligence systems increasingly involve continual learning to enable flexibility in general situations that are not encountered during system training. Human interaction with autonomous systems is broadly studied, but research…

Deep Neural Networks (DNNs) have been widely used to perform real-world tasks in cyber-physical systems such as Autonomous Driving Systems (ADS). Ensuring the correct behavior of such DNN-Enabled Systems (DES) is a crucial topic. Online…

机器学习 · 计算机科学 2022-12-23 Fitash Ul Haq , Donghwan Shin , Lionel Briand

In the rapidly evolving landscape of eCommerce, Artificial Intelligence (AI) based pricing algorithms, particularly those utilizing Reinforcement Learning (RL), are becoming increasingly prevalent. This rise has led to an inextricable…

机器学习 · 计算机科学 2024-06-06 Michael Schlechtinger , Damaris Kosack , Franz Krause , Heiko Paulheim

In this paper, we consider learning dictionary models over a network of agents, where each agent is only in charge of a portion of the dictionary elements. This formulation is relevant in Big Data scenarios where large dictionary models may…

机器学习 · 计算机科学 2015-06-18 Jianshu Chen , Zaid J. Towfic , Ali H. Sayed

In online learning environments, students often lack personalized peer interactions, which are crucial for cognitive development and learning engagement. Although previous studies have employed large language models (LLMs) to simulate…

计算机与社会 · 计算机科学 2026-01-08 Xian Gao , Zongyun Zhang , Ting Liu , Yuzhuo Fu

In this work, we introduce the Resilient Projected Push-Pull (RP3) algorithm designed for distributed optimization in multi-agent cyber-physical systems with directed communication graphs and the presence of malicious agents. Our algorithm…

系统与控制 · 电气工程与系统科学 2024-07-10 Arif Kerem Dayı , Orhan Eren Akgün , Stephanie Gil , Michal Yemini , Angelia Nedić

Autonomous driving has experienced remarkable progress, bolstered by innovations in computational hardware and sophisticated deep learning methodologies. The foundation of these advancements rests on the availability and quality of…

计算机视觉与模式识别 · 计算机科学 2024-06-13 Joshua Tokarsky , Ibrahim Abdulhafiz , Satya Ayyalasomayajula , Mostafa Mohsen , Navya G. Rao , Adam Forbes

Recent advancements in large language models (LLMs) have enabled understanding webpage contexts, product details, and human instructions. Utilizing LLMs as the foundational architecture for either reward models or policies in reinforcement…

机器学习 · 计算机科学 2024-08-30 Shuang Feng , Grace Feng

Maximizing long-term rewards is the primary goal in sequential decision-making problems. The majority of existing methods assume that side information is freely available, enabling the learning agent to observe all features' states before…

机器学习 · 计算机科学 2023-07-19 Saeed Ghoorchian , Evgenii Kortukov , Setareh Maghsudi

Offloading computational tasks from resource-constrained devices to resource-abundant peers constitutes a critical paradigm for collaborative computing. Within this context, accurate trust evaluation of potential collaborating devices is…

人工智能 · 计算机科学 2025-12-24 Botao Zhu , Jeslyn Wang , Dusit Niyato , Xianbin Wang

Cooperative multi-agent reinforcement learning often assumes a fixed execution team, yet many decentralized systems must operate with varying numbers of active agents during deployment. We study this setting under episodic roster variation:…

机器学习 · 计算机科学 2026-05-12 Ahmet Onur Akman , Rafał Kucharski

The significant role of division of labor (DOL) in promoting cooperation is widely recognized in real-world applications.Many cooperative multi-agent reinforcement learning (MARL) methods have incorporated the concept of DOL to improve…

机器学习 · 计算机科学 2025-02-04 Yurui Li , Yuxuan Chen , Li Zhang , Shijian Li , Gang Pan

We study interpersonal trust by means of the all-or-nothing public goods game between agents on a network. The agents are endowed with the simple yet adaptive learning rule, exponential moving average, by which they estimate the behavior of…

计算机科学与博弈论 · 计算机科学 2024-12-31 Benedikt Valentin Meylahn

Federated learning performs distributed model training using local data hosted by agents. It shares only model parameter updates for iterative aggregation at the server. Although it is privacy-preserving by design, federated learning is…

机器学习 · 计算机科学 2019-05-09 Yufei Han , Xiangliang Zhang

Heterogeneous multi-robot sensing systems are able to characterize physical processes more comprehensively than homogeneous systems. Access to multiple modalities of sensory data allow such systems to fuse information between complementary…

机器人学 · 计算机科学 2021-06-30 Andrew McDonald , Lai Wei , Vaibhav Srivastava

By informing the onset of the degradation process, health status evaluation serves as a significant preliminary step for reliable remaining useful life (RUL) estimation of complex equipment. This paper proposes a novel temporal dynamics…

机器学习 · 计算机科学 2024-01-10 Anushiya Arunan , Yan Qin , Xiaoli Li , Chau Yuen

Multi-agent systems have evolved into practical LLM-driven collaborators for many applications, gaining robustness from diversity and cross-checking. However, multi-agent RL (MARL) training is resource-intensive and unstable: co-adapting…

High-speed, low-latency obstacle avoidance that is insensitive to sensor noise is essential for enabling multiple decentralized robots to function reliably in cluttered and dynamic environments. While other distributed multi-agent collision…

人工智能 · 计算机科学 2017-07-07 Pinxin Long , Wenxi Liu , Jia Pan

Most of the prior work on multi-agent reinforcement learning (MARL) achieves optimal collaboration by directly controlling the agents to maximize a common reward. In this paper, we aim to address this from a different angle. In particular,…

人工智能 · 计算机科学 2019-03-08 Tianmin Shu , Yuandong Tian

Increasing a ML model accuracy is not enough, we must also increase its trustworthiness. This is an important step for building resilient AI systems for safety-critical applications such as automotive, finance, and healthcare. For that…

人工智能 · 计算机科学 2022-05-03 Gusseppe Bravo-Rocca , Peini Liu , Jordi Guitart , Ajay Dholakia , David Ellison , Miroslav Hodak