English
Related papers

Related papers: End-to-end Reinforcement Learning for Time-Optimal…

200 papers

Deep Reinforcement Learning (DRL) for quadrotor flight control typically relies on Domain Randomization (DR) for sim-to-real transfer, resulting in overly conservative policies that struggle with dynamic disturbances. To overcome this, we…

Robotics · Computer Science 2026-05-19 Vishnu Saj , Sushil Vemuri , Dileep Kalathil , Moble Benedict

This chapter addresses the critical challenge of simulation-to-reality (sim-to-real) transfer for deep reinforcement learning (DRL) in bipedal locomotion. After contextualizing the problem within various control architectures, we dissect…

Robotics · Computer Science 2025-11-11 Lingfan Bao , Tianhu Peng , Chengxu Zhou

Deep reinforcement learning (RL) has made it possible to solve complex robotics problems using neural networks as function approximators. However, the policies trained on stationary environments suffer in terms of generalization when…

Robotics · Computer Science 2021-11-09 Aditya M. Deshpande , Ali A. Minai , Manish Kumar

Due to their superior energy efficiency, blimps may replace quadcopters for long-duration aerial tasks. However, designing a controller for blimps to handle complex dynamics, modeling errors, and disturbances remains an unsolved challenge.…

Robotics · Computer Science 2023-03-27 Yang Zuo , Yu Tang Liu , Aamir Ahmad

Existing inverse reinforcement learning methods (e.g. MaxEntIRL, $f$-IRL) search over candidate reward functions and solve a reinforcement learning problem in the inner loop. This creates a rather strange inversion where a harder problem,…

Machine Learning · Computer Science 2024-02-06 David Wu , Sanjiban Choudhury

Deep reinforcement Learning for end-to-end driving is limited by the need of complex reward engineering. Sparse rewards can circumvent this challenge but suffers from long training time and leads to sub-optimal policy. In this work, we…

Robotics · Computer Science 2021-08-03 Pranav Agarwal , Pierre de Beaucorps , Raoul de Charette

This work explores techniques to scale up image-based end-to-end learning for dexterous grasping with an arm + hand system. Unlike state-based RL, vision-based RL is much more memory inefficient, resulting in relatively low batch sizes,…

Robotics · Computer Science 2025-09-23 Ritvik Singh , Karl Van Wyk , Pieter Abbeel , Jitendra Malik , Nathan Ratliff , Ankur Handa

Jumping poses a significant challenge for quadruped robots, despite being crucial for many operational scenarios. While optimisation methods exist for controlling such motions, they are often time-consuming and demand extensive knowledge of…

Robotics · Computer Science 2026-05-19 Riccardo Bussola , Michele Focchi , Giulio Turrisi , Claudio Semini , Luigi Palopoli

Autonomous drone racing (ADR) systems have recently achieved champion-level performance, yet remain highly specific to drone racing. While end-to-end vision-based methods promise broader applicability, no system to date simultaneously…

Robotics · Computer Science 2025-10-17 Aderik Verraest , Stavrow Bahnam , Robin Ferede , Guido de Croon , Christophe De Wagter

Improving sampling efficiency and generalization capability is critical for the successful data-driven control of quadrotor unmanned aerial vehicles (UAVs) that are inherently unstable. While various reinforcement learning (RL) approaches…

Robotics · Computer Science 2025-03-03 Beomyeol Yu , Taeyoung Lee

An oft-ignored challenge of real-world reinforcement learning is that the real world does not pause when agents make learning updates. As standard simulated environments do not address this real-time aspect of learning, most available…

Robotics · Computer Science 2022-04-01 Yufeng Yuan , A. Rupam Mahmood

Learning visuomotor policies for agile quadrotor flight presents significant difficulties, primarily from inefficient policy exploration caused by high-dimensional visual inputs and the need for precise and low-latency control. To address…

Robotics · Computer Science 2024-11-13 Jiaxu Xing , Angel Romero , Leonard Bauersfeld , Davide Scaramuzza

Deep reinforcement learning has emerged as a promising and powerful technique for automatically acquiring control policies that can process raw sensory inputs, such as images, and perform complex behaviors. However, extending deep RL to…

Machine Learning · Computer Science 2017-06-09 Fereshteh Sadeghi , Sergey Levine

Reinforcement Learning (RL) algorithms have found limited success beyond simulated applications, and one main reason is the absence of safety guarantees during the learning process. Real world systems would realistically fail or break…

Machine Learning · Computer Science 2019-03-22 Richard Cheng , Gabor Orosz , Richard M. Murray , Joel W. Burdick

In this paper, we study the whole-body loco-manipulation problem using reinforcement learning (RL). Specifically, we focus on the problem of how to coordinate the floating base and the robotic arm of a wheeled-quadrupedal manipulator robot…

Robotics · Computer Science 2025-08-14 Kaiwen Jiang , Zhen Fu , Junde Guo , Wei Zhang , Hua Chen

In the last decade, data-driven approaches have become popular choices for quadrotor control, thanks to their ability to facilitate the adaptation to unknown or uncertain flight conditions. Among the different data-driven paradigms, Deep…

Robotics · Computer Science 2024-12-30 Alberto Dionigi , Gabriele Costante , Giuseppe Loianno

This paper introduces a flight envelope protection algorithm on a longitudinal axis that leverages reinforcement learning (RL). By considering limits on variables such as angle of attack, load factor, and pitch rate, the algorithm…

Systems and Control · Electrical Eng. & Systems 2024-06-13 Akin Catak , Ege C. Altunkaya , Mustafa Demir , Emre Koyuncu , Ibrahim Ozkol

Learning-based controllers have achieved impressive performance in agile quadrotor flight but typically rely on massive training in simulation, necessitating accurate system identification for effective Sim2Real transfer. However, even with…

Robotics · Computer Science 2026-02-11 Yunfan Ren , Zhiyuan Zhu , Jiaxu Xing , Davide Scaramuzza

We develop provably safe and convergent reinforcement learning (RL) algorithms for control of nonlinear dynamical systems, bridging the gap between the hard safety guarantees of control theory and the convergence guarantees of RL theory.…

Machine Learning · Computer Science 2024-03-08 Wesley A. Suttle , Vipul K. Sharma , Krishna C. Kosaraju , S. Sivaranjani , Ji Liu , Vijay Gupta , Brian M. Sadler

Quadrotors have demonstrated remarkable versatility, yet their full aerobatic potential remains largely untapped due to inherent underactuation and the complexity of aggressive maneuvers. Traditional approaches, separating trajectory…

Robotics · Computer Science 2025-06-02 Zhichao Han , Xijie Huang , Zhuxiu Xu , Jiarui Zhang , Yuze Wu , Mingyang Wang , Tianyue Wu , Fei Gao