中文
相关论文

相关论文: End-to-end Reinforcement Learning for Time-Optimal…

200 篇论文

End-to-end deep reinforcement learning (DRL) for quadrotor control promises many benefits -- easy deployment, task generalization and real-time execution capability. Prior end-to-end DRL-based methods have showcased the ability to deploy…

机器人学 · 计算机科学 2024-05-07 Zhehui Huang , Zhaojing Yang , Rahul Krupani , Baskın Şenbaşlar , Sumeet Batra , Gaurav S. Sukhatme

In this work, we consider the complex control problem of making a monopod reach a target with a jump. The monopod can jump in any direction and the terrain underneath its foot can be uneven. This is a template of a much larger class of…

机器人学 · 计算机科学 2024-08-06 Riccardo Bussola , Michele Focchi , Andrea Del Prete , Daniele Fontanelli , Luigi Palopoli

Existing end-to-end autonomous driving (AD) algorithms typically follow the Imitation Learning (IL) paradigm, which faces challenges such as causal confusion and an open-loop gap. In this work, we propose RAD, a 3DGS-based closed-loop…

计算机视觉与模式识别 · 计算机科学 2025-10-22 Hao Gao , Shaoyu Chen , Bo Jiang , Bencheng Liao , Yiang Shi , Xiaoyang Guo , Yuechuan Pu , Haoran Yin , Xiangyu Li , Xinbang Zhang , Ying Zhang , Wenyu Liu , Qian Zhang , Xinggang Wang

Recent advances in deep reinforcement learning (RL) based techniques combined with training in simulation have offered a new approach to developing robust controllers for legged robots. However, the application of such approaches to real…

机器人学 · 计算机科学 2023-08-08 Rohan Pratap Singh , Zhaoming Xie , Pierre Gergondet , Fumio Kanehiro

Autopilot systems are typically composed of an "inner loop" providing stability and control, while an "outer loop" is responsible for mission-level objectives, e.g. way-point navigation. Autopilot systems for UAVs are predominately…

机器人学 · 计算机科学 2018-04-13 William Koch , Renato Mancuso , Richard West , Azer Bestavros

We present research using the latest reinforcement learning algorithm for end-to-end driving without any mediated perception (object recognition, scene understanding). The newly proposed reward and learning strategies lead together to…

计算机视觉与模式识别 · 计算机科学 2018-09-03 Maximilian Jaritz , Raoul de Charette , Marin Toromanoff , Etienne Perot , Fawzi Nashashibi

Applications of Reinforcement Learning (RL) in robotics are often limited by high data demand. On the other hand, approximate models are readily available in many robotics scenarios, making model-based approaches like planning a…

人工智能 · 计算机科学 2021-11-16 Ingmar Schubert , Danny Driess , Ozgur S. Oguz , Marc Toussaint

With the rising popularity of autonomous navigation research, Formula Student (FS) events are introducing a Driverless Vehicle (DV) category to their event list. This paper presents the initial investigation into utilising Deep…

机器人学 · 计算机科学 2023-08-28 Aakaash Salvaji , Harry Taylor , David Valencia , Trevor Gee , Henry Williams

Multi-rotor UAVs suffer from a restricted range and flight duration due to limited battery capacity. Autonomous landing on a 2D moving platform offers the possibility to replenish batteries and offload data, thus increasing the utility of…

机器人学 · 计算机科学 2024-05-17 Pascal Goldschmid , Aamir Ahmad

Although end-to-end (E2E) learning has led to impressive progress on a variety of visual understanding tasks, it is often impeded by hardware constraints (e.g., GPU memory) and is prone to overfitting. When it comes to video captioning, one…

计算机视觉与模式识别 · 计算机科学 2019-01-03 Lijun Li , Boqing Gong

This article introduces a novel sample-efficient curriculum learning (CL) approach for training an end-to-end reinforcement learning (RL) policy for robust stabilization of a Quadrotor. The learning objective is to simultaneously stabilize…

This paper proposes the ProxFly, a residual deep Reinforcement Learning (RL)-based controller for close proximity quadcopter flight. Specifically, we design a residual module on top of a cascaded controller (denoted as basic controller) to…

机器人学 · 计算机科学 2025-05-02 Ruiqi Zhang , Dingqi Zhang , Mark W. Mueller

We explore the reinforcement learning approach to designing controllers by extensively discussing the case of a quadcopter attitude controller. We provide all details allowing to reproduce our approach, starting with a model of the dynamics…

End-to-end autonomous driving policies based on Imitation Learning (IL) often struggle in closed-loop execution due to the misalignment between inadequate open-loop training objectives and real driving requirements. While Reinforcement…

机器人学 · 计算机科学 2026-03-17 Yinfeng Gao , Qichao Zhang , Deqing Liu , Zhongpu Xia , Guang Li , Kun Ma , Guang Chen , Hangjun Ye , Long Chen , Da-Wei Ding , Dongbin Zhao

Designing a driving policy for autonomous vehicles is a difficult task. Recent studies suggested an end-toend (E2E) training of a policy to predict car actuators directly from raw sensory inputs. It is appealing due to the ease of labeled…

机器人学 · 计算机科学 2019-01-07 Yonatan Glassner , Liran Gispan , Ariel Ayash , Tal Furman Shohet

Current control algorithms for aerial robots struggle with robustness in dynamic environments and adverse conditions. Model-based reinforcement learning (RL) has shown strong potential in handling these challenges while remaining…

机器人学 · 计算机科学 2025-11-25 Eashan Vytla , Bhavanishankar Kalavakolanu , Andrew Perrault , Matthew McCrink

End-to-end (E2E) training, optimizing the entire model through error backpropagation, fundamentally supports the advancements of deep learning. Despite its high performance, E2E training faces the problems of memory consumption, parallel…

机器学习 · 计算机科学 2024-06-03 Keitaro Sakamoto , Issei Sato

Quadruped robots are used for primary searches during the early stages of indoor fires. A typical primary search involves quickly and thoroughly looking for victims under hazardous conditions and monitoring flammable materials. However,…

机器人学 · 计算机科学 2026-02-04 Baixiao Huang , Baiyu Huang , Yu Hou

Reinforcement learning (RL) offers transformative potential for robotic control in space. We present the first on-orbit demonstration of RL-based autonomous control of a free-flying robot, the NASA Astrobee, aboard the International Space…

机器人学 · 计算机科学 2026-04-01 Kenneth Stewart , Samantha Chapin , Roxana Leontie , Carl Glen Henshaw

In this paper, we investigate the problem of enabling a drone to fly through a tilted narrow gap, without a traditional planning and control pipeline. To this end, we propose an end-to-end policy network, which imitates from the traditional…

机器人学 · 计算机科学 2019-08-06 Jiarong Lin , Luqi Wang , Fei Gao , Shaojie Shen , Fu Zhang