中文
相关论文

相关论文: A New Formalism, Method and Open Issues for Zero-S…

200 篇论文

Multi-agent reinforcement learning (MARL) has attracted much research attention recently. However, unlike its single-agent counterpart, many theoretical and algorithmic aspects of MARL have not been well-understood. In this paper, we study…

机器学习 · 计算机科学 2021-12-08 Siliang Zeng , Tianyi Chen , Alfredo Garcia , Mingyi Hong

Deep subspace clustering (DSC) algorithms face several challenges that hinder their widespread adoption across variois application domains. First, clustering quality is typically assessed using only the encoder's output layer, disregarding…

计算机视觉与模式识别 · 计算机科学 2025-04-28 Lovro Sindicic , Ivica Kopriva

Decentralized air traffic management systems offer a scalable alternative to centralized control, but often assume high levels of cooperation. In practice, such assumptions frequently break down since airspace sectors operate independently…

系统与控制 · 电气工程与系统科学 2025-11-19 Jaehan Im , Daniel Delahaye , David Fridovich-Keil , Ufuk Topcu

We consider a decentralized system with multiple controllers and define substitutability of one controller by another in open-loop strategies. We explore the implications of this property on the optimization of closed-loop strategies. In…

系统与控制 · 计算机科学 2016-01-12 Seyed Mohammad Asghari , Ashutosh Nayyar

This paper investigates repeated win-lose coordination games (WLC-games). We analyse which protocols are optimal for these games, covering both the worst case and average case scenarios, i,e., optimizing the guaranteed and expected…

计算机科学与博弈论 · 计算机科学 2021-09-20 Antti Kuusisto , Raine Rönnholm

Stochastic dynamic teams and games are rich models for decentralized systems and challenging testing grounds for multi-agent learning. Previous work that guaranteed team optimality assumed stateless dynamics, or an explicit coordination…

最优化与控制 · 数学 2024-03-28 Bora Yongacoglu , Gürdal Arslan , Serdar Yüksel

Supervised learning requires a sufficient training dataset which includes all label. However, there are cases that some class is not in the training data. Zero-Shot Learning (ZSL) is the task of predicting class that is not in the training…

机器学习 · 计算机科学 2020-07-02 Toshitaka Hayashi , Hamido Fujita

Ensuring that large language models (LLMs) comply with safety requirements is a central challenge in AI deployment. Existing alignment approaches primarily operate during training, such as through fine-tuning or reinforcement learning from…

机器学习 · 计算机科学 2025-12-03 Tuan Nguyen , Long Tran-Thanh

This paper presents a transferable solution method for optimal control problems with varying objectives using function encoder (FE) policies. Traditional optimization-based approaches must be re-solved whenever objectives change, resulting…

最优化与控制 · 数学 2026-03-12 Xingjian Li , Kelvin Kan , Deepanshu Verma , Krishna Kumar , Stanley Osher , Ján Drgoňa

This paper addresses a Stackelberg stochastic linear-quadratic (LQ) differential game under closed-loop information, a problem inherently time-inconsistent. Existing approaches rely on solving two coupled Hamilton-Jacobi-Bellman (HJB)…

最优化与控制 · 数学 2026-04-27 Qi Lü , Bowen Ma , Hanxiao Wang

This paper analyzes the fundamental limits of strate- gic communication in network settings. Strategic communication differs from the conventional communication paradigms in in- formation theory since it involves different objectives for…

信息论 · 计算机科学 2016-02-24 Emrah Akyol , Cedric Langbort , Tamer Basar

The fundamental goal assignment problem for a multi-robot application aims to assign a unique goal to each robot while ensuring collision-free paths, minimizing the total movement cost. A plausible algorithmic solution to this NP-hard…

多智能体系统 · 计算机科学 2024-02-22 Aakash , Indranil Saha

Humans are capable of abstracting various tasks as different combinations of multiple attributes. This perspective of compositionality is vital for human rapid learning and adaption since previous experiences from related tasks can be…

机器人学 · 计算机科学 2022-10-04 Zheng Wu , Yichen Xie , Wenzhao Lian , Changhao Wang , Yanjiang Guo , Jianyu Chen , Stefan Schaal , Masayoshi Tomizuka

A key goal of ad hoc teamwork is to develop a learning agent that cooperates with unknown teams, without resorting to any pre-coordination protocol. Despite a vast number of ad hoc teamwork algorithms in the literature, most of them cannot…

多智能体系统 · 计算机科学 2022-05-09 Alexandre Neves , Alberto Sardinha

This paper investigates the implementation and performance of a decentralized information transmission mechanism in game with complete or incomplete games. We propose a mechanism that realizes irrational correlated equilibria or irrational…

理论经济学 · 经济学 2025-11-12 Shitong Wang

Multi-agent coordination dilemmas expose a fundamental tension between individual optimization and collective welfare, yet characterizing such coordination requires metrics sensitive to temporal structure and collective dynamics. As a…

多智能体系统 · 计算机科学 2026-03-24 Nikolaos Al. Papadopoulos , Konstantinos Psannis

The decentralized stochastic multi-player multi-armed bandit (MP-MAB) problem, where the collision information is not available to the players, is studied in this paper. Building on the seminal work of Boursier and Perchet (2019), we…

机器学习 · 计算机科学 2020-03-03 Chengshuai Shi , Wei Xiong , Cong Shen , Jing Yang

Safety is a primary concern when applying reinforcement learning to real-world control tasks, especially in the presence of external disturbances. However, existing safe reinforcement learning algorithms rarely account for external…

机器学习 · 计算机科学 2023-10-12 Zeyang Li , Chuxiong Hu , Shengbo Eben Li , Jia Cheng , Yunan Wang

Foundation vision-language models have enabled remarkable zero-shot transferability of the pre-trained representations to a wide range of downstream tasks. However, to solve a new task, zero-shot transfer still necessitates human guidance…

机器学习 · 计算机科学 2024-06-12 Artyom Gadetsky , Yulun Jiang , Maria Brbic

This paper considers the problem of how to allocate power among competing users sharing a frequency-selective interference channel. We model the interaction between selfish users as a non-cooperative game. As opposed to the existing…

计算机科学与博弈论 · 计算机科学 2008-12-16 Yi Su , Mihaela van der Schaar