中文
相关论文

相关论文: Model-Free versus Model-Based Reinforcement Learni…

200 篇论文

Fixed-wing unmanned aerial vehicles (UAVs) offer significant performance advantages over rotary-wing UAVs in terms of speed, endurance, and efficiency. However, these vehicles have traditionally been severely limited with regards to…

机器人学 · 计算机科学 2020-01-31 Max Basescu , Joseph Moore

We propose a learning-based robust predictive control algorithm that compensates for significant uncertainty in the dynamics for a class of discrete-time systems that are nominally linear with an additive nonlinear component. Such systems…

系统与控制 · 电气工程与系统科学 2022-12-05 Rohan Sinha , James Harrison , Spencer M. Richards , Marco Pavone

Successful aerial manipulation largely depends on how effectively a controller can tackle the coupling dynamic forces between the aerial vehicle and the manipulator. However, this control problem has remained largely unsolved as the…

机器人学 · 计算机科学 2024-10-14 Rishabh Dev Yadav , Swati Dantu , Wei Pan , Sihao Sun , Spandan Roy , Simone Baldi

Fault-tolerant flight control faces challenges, as developing a model-based controller for each unexpected failure is unrealistic, and online learning methods can handle limited system complexity due to their low sample efficiency. In this…

系统与控制 · 电气工程与系统科学 2023-01-02 Killian Dally , Erik-Jan van Kampen

Model-free reinforcement learning has recently been shown to successfully learn navigation policies from raw sensor data. In this work, we address the problem of learning driving policies for an autonomous agent in a high-fidelity…

机器学习 · 计算机科学 2019-02-12 Qadeer Khan , Torsten Schön , Patrick Wenzel

In this paper, we propose an adaptive event-triggered reinforcement learning control for continuous-time nonlinear systems, subject to bounded uncertainties, characterized by complex interactions. Specifically, the proposed method is…

机器学习 · 计算机科学 2024-10-01 Umer Siddique , Abhinav Sinha , Yongcan Cao

Model predictive control (MPC) is an effective method for controlling robotic systems, particularly autonomous aerial vehicles such as quadcopters. However, application of MPC can be computationally demanding, and typically requires…

机器学习 · 计算机科学 2016-02-17 Tianhao Zhang , Gregory Kahn , Sergey Levine , Pieter Abbeel

Model-based reinforcement learning promises to learn an optimal policy from fewer interactions with the environment compared to model-free reinforcement learning by learning an intermediate model of the environment in order to predict…

机器学习 · 计算机科学 2022-06-08 Abhinav Bhatia , Philip S. Thomas , Shlomo Zilberstein

We propose a learning-based robust predictive control algorithm that compensates for significant uncertainty in the dynamics for a class of discrete-time systems that are nominally linear with an additive nonlinear component. Such systems…

系统与控制 · 电气工程与系统科学 2021-10-15 Rohan Sinha , James Harrison , Spencer M. Richards , Marco Pavone

This paper presents a novel model-reference reinforcement learning algorithm for the intelligent tracking control of uncertain autonomous surface vehicles with collision avoidance. The proposed control algorithm combines a conventional…

系统与控制 · 电气工程与系统科学 2020-08-18 Qingrui Zhang , Wei Pan , Vasso Reppa

We present an online model-based reinforcement learning algorithm suitable for controlling complex robotic systems directly in the real world. Unlike prevailing sim-to-real pipelines that rely on extensive offline simulation and model-free…

机器人学 · 计算机科学 2026-05-07 Fang Nan , Hao Ma , Qinghua Guan , Josie Hughes , Michael Muehlebach , Marco Hutter

In this study, we applied reinforcement learning based on the proximal policy optimization algorithm to perform motion planning for an unmanned aerial vehicle (UAV) in an open space with static obstacles. The application of reinforcement…

机器人学 · 计算机科学 2020-12-17 Sanghyun Kim , Jongmin Park , Jae-Kwan Yun , Jiwon Seo

This paper tackles the challenge of learning a generalizable minimum-time flight policy for UAVs, capable of navigating between arbitrary start and goal states while balancing agile flight and stable hovering. Traditional approaches,…

机器人学 · 计算机科学 2025-10-24 Swati Dantu , Robert Pěnička , Martin Saska

There have been numerous attempts in explaining the general learning behaviours using model-based and model-free methods. While the model-based control is flexible yet computationally expensive in planning, the model-free control is quick…

机器学习 · 计算机科学 2019-12-12 Krishn Bera , Yash Mandilwar , Bapi Raju

Learn-to-Fly (L2F) is a new framework that aims to replace the traditional iterative development paradigm for aerial vehicles with a combination of real-time aerodynamic modeling, guidance, and learning control. To ensure safe learning of…

系统与控制 · 电气工程与系统科学 2022-08-08 Steven Snyder , Pan Zhao , Naira Hovakimyan

Attitude control of a novel regional truss-braced wing aircraft with low stability characteristics is addressed in this paper using Reinforcement Learning (RL). In recent years, RL has been increasingly employed in challenging applications,…

系统与控制 · 电气工程与系统科学 2022-10-25 Mohsen Zahmatkesh , Seyyed Ali Emami , Afshin Banazadeh , Paolo Castaldi

Previous results reported in the robotics literature show the relationship between time-delay control (TDC) and proportional-integral-derivative control (PID). In this paper, we show that incremental nonlinear dynamic inversion (INDI) -…

系统与控制 · 计算机科学 2021-12-06 Paul Acquatella B. , Wim van Ekeren , Qi Ping Chu

Controlling turbulent dynamics remains a major challenge because of its chaotic, multi-scale dynamics, which strongly influence the performance of many fluid systems. Here we report REACT (Reinforcement Learning for Environmental Adaptation…

流体动力学 · 物理学 2026-05-18 Junjie Zhang , Chengwei Xia , Xianyang Jiang , Isabella Fumarola , Georgios Rigas

Adapting an interface requires taking into account both the positive and negative effects that changes may have on the user. A carelessly picked adaptation may impose high costs to the user -- for example, due to surprise or relearning…

人机交互 · 计算机科学 2021-03-12 Kashyap Todi , Gilles Bailly , Luis A. Leiva , Antti Oulasvirta

Many robotic systems are underactuated, meaning not all degrees of freedom can be directly controlled due to lack of actuators, input constraints, or state-dependent actuation. This property, compounded by modeling uncertainties and…

系统与控制 · 电气工程与系统科学 2025-10-10 Daniel M. Cherenson , Dimitra Panagou