中文
相关论文

相关论文: End-to-End ILC for Repetitive Untrackable Tasks: A…

200 篇论文

In many coordination problems, independently reasoning humans are able to discover mutually compatible policies. In contrast, independently trained self-play policies are often mutually incompatible. Zero-shot coordination (ZSC) has…

人工智能 · 计算机科学 2023-07-14 Johannes Treutlein , Michael Dennis , Caspar Oesterheld , Jakob Foerster

This paper studies data-driven iterative learning control (ILC) for linear time-invariant (LTI) systems with unknown dynamics, output disturbances and input box-constraints. Our main contributions are: 1) using a non-parametric data-driven…

系统与控制 · 电气工程与系统科学 2023-12-25 Jia Wang , Leander Hemelhof , Ivan Markovsky , Panagiotis Patrinos

Machine learning (ML)-based planners have recently gained significant attention. They offer advantages over traditional optimization-based planning algorithms. These advantages include fewer manually selected parameters and faster…

计算机视觉与模式识别 · 计算机科学 2025-04-22 Hui Zhou , Shaoshuai Shi , Hongsheng Li

Collaborations among various entities, such as companies, research labs, AI agents, and edge devices, have become increasingly crucial for achieving machine learning tasks that cannot be accomplished by a single entity alone. This is likely…

机器学习 · 计算机科学 2023-05-29 Xinran Wang , Qi Le , Ahmad Faraz Khan , Jie Ding , Ali Anwar

Multi-task Inverse Reinforcement Learning (IRL) is the problem of inferring multiple reward functions from expert demonstrations. Prior work, built on Bayesian IRL, is unable to scale to complex environments due to computational…

机器学习 · 计算机科学 2018-07-17 Adam Gleave , Oliver Habryka

This paper considers the leader-follower control problem for a linear multi-agent system with undirected topology and linear coupling subject to integral quadratic constraints (IQCs). A consensus-type control protocol is proposed based on…

系统与控制 · 计算机科学 2013-03-13 Yi Cheng , V. Ugrinovskii

Instruction tuning effectively optimizes Large Language Models (LLMs) for downstream tasks. Due to the changing environment in real-life applications, LLMs necessitate continual task-specific adaptation without catastrophic forgetting.…

计算与语言 · 计算机科学 2024-03-19 Yifan Wang , Yafei Liu , Chufan Shi , Haoling Li , Chen Chen , Haonan Lu , Yujiu Yang

We examine sequential equilibrium in the context of computational games, where agents are charged for computation. In such games, an agent can rationally choose to forget, so issues of imperfect recall arise. In this setting, we consider…

计算机科学与博弈论 · 计算机科学 2014-12-22 Joseph Y. Halpern , Rafael Pass

This abstract aims at presenting an ongoing effort to apply a novel typing mechanism stemming from Implicit Computational Complexity (ICC), that tracks dependencies between variables in three different ways, at different stages of…

计算复杂性 · 计算机科学 2022-05-26 Clément Aubert , Thomas Rubiano , Neea Rusch , Thomas Seiller

Autonomous driving systems need to handle complex scenarios such as lane following, avoiding collisions, taking turns, and responding to traffic signals. In recent years, approaches based on end-to-end behavioral cloning have demonstrated…

机器人学 · 计算机科学 2021-04-23 Keishi Ishihara , Anssi Kanervisto , Jun Miura , Ville Hautamäki

Iterative linear-quadratic (ILQ) methods are widely used in the nonlinear optimal control community. Recent work has applied similar methodology in the setting of multiplayer general-sum differential games. Here, ILQ methods are capable of…

系统与控制 · 电气工程与系统科学 2020-03-20 David Fridovich-Keil , Vicenc Rubies-Royo , Claire J. Tomlin

Two-team zero-sum games are one of the most important paradigms in game theory. In this paper, we focus on finding an unexploitable equilibrium in large team games. An unexploitable equilibrium is a worst-case policy, where members in the…

计算机科学与博弈论 · 计算机科学 2024-03-04 Naming Liu , Mingzhi Wang , Youzhi Zhang , Yaodong Yang , Bo An , Ying Wen

Large language models (LLMs) exhibit remarkable in-context learning (ICL) capabilities. However, the underlying working mechanism of ICL remains poorly understood. Recent research presents two conflicting views on ICL: One emphasizes the…

计算与语言 · 计算机科学 2024-10-10 Anhao Zhao , Fanghua Ye , Jinlan Fu , Xiaoyu Shen

We present Intermittent Control (IC) models as a candidate framework for modelling human input movements in Human--Computer Interaction (HCI). IC differs from continuous control in that users are not assumed to use feedback to adjust their…

人机交互 · 计算机科学 2021-03-16 J. Alberto Álvarez Martín , Henrik Gollee , Jörg Müller , Roderick Murray-Smith

Fast execution of contact-rich manipulation is critical for practical deployment, yet providing fast demonstrations for imitation learning (IL) remains challenging: humans cannot demonstrate at high speed, and naively accelerating…

机器人学 · 计算机科学 2026-04-21 Koki Yamane , Cristian C. Beltran-Hernandez , Steven Oh , Masashi Hamaya , Sho Sakaino

We consider a computing system where a master processor assigns tasks for execution to worker processors through the Internet. We model the workers decision of whether to comply (compute the task) or not (return a bogus result to save the…

分布式、并行与集群计算 · 计算机科学 2015-08-25 Antonio Fernández Anta , Chryssis Georgiou , Miguel A. Mosteiro , Daniel Pareja

In this paper we present a Learning Model Predictive Controller (LMPC) for autonomous racing. We model the autonomous racing problem as a minimum time iterative control task, where an iteration corresponds to a lap. In the proposed approach…

系统与控制 · 电气工程与系统科学 2024-12-20 Ugo Rosolia , Francesco Borrelli

Interaction-aware trajectory planning is crucial for closing the gap between autonomous racing cars and human racing drivers. Prior work has applied game theory as it provides equilibrium concepts for non-cooperative dynamic problems. With…

机器人学 · 计算机科学 2024-02-06 Matthias Rowold , Alexander Langmann , Boris Lohmann , Johannes Betz

This paper explores continuous-time control synthesis for target-driven navigation to satisfy complex high-level tasks expressed as linear temporal logic (LTL). We propose a model-free framework using deep reinforcement learning (DRL) where…

机器人学 · 计算机科学 2023-03-17 Mingyu Cai , Makai Mann , Zachary Serlin , Kevin Leahy , Cristian-Ioan Vasile

We present the problem of inverse constraint learning (ICL), which recovers constraints from demonstrations to autonomously reproduce constrained skills in new scenarios. However, ICL suffers from an ill-posed nature, leading to inaccurate…

机器人学 · 计算机科学 2023-12-11 Jaehwi Jang , Minjae Song , Daehyung Park