中文
相关论文

相关论文: End-to-end Reinforcement Learning for Time-Optimal…

200 篇论文

This paper identifies and addresses the problems with naively combining (reinforcement) learning-based controllers and state estimators for robotic in-hand manipulation. Specifically, we tackle the challenging task of purely tactile,…

机器人学 · 计算机科学 2024-01-09 Lennart Röstel , Johannes Pitz , Leon Sievers , Berthold Bäuml

Due to dynamic variations such as changing payload, aerodynamic disturbances, and varying platforms, a robust solution for quadrotor trajectory tracking remains challenging. To address these challenges, we present a deep reinforcement…

系统与控制 · 电气工程与系统科学 2026-01-06 Varad Vaidya , Jishnu Keshavan

It is well-known that inverse dynamics models can improve tracking performance in robot control. These models need to precisely capture the robot dynamics, which consist of well-understood components, e.g., rigid body dynamics, and effects…

机器人学 · 计算机科学 2022-05-30 Moritz Reuss , Niels van Duijkeren , Robert Krug , Philipp Becker , Vaisakh Shaj , Gerhard Neumann

End-to-end delay is a critical attribute of quality of service (QoS) in application domains such as cloud computing and computer networks. This metric is particularly important in tandem service systems, where the end-to-end service is…

机器学习 · 计算机科学 2021-01-13 Majid Raeis , Ali Tizghadam , Alberto Leon-Garcia

In this paper, we present a deep reinforcement learning method for quadcopter bypassing the obstacle on the flying path. In the past study, the algorithm only controls the forward direction about quadcopter. In this letter, we use two…

人工智能 · 计算机科学 2018-11-13 Tung-Cheng Wu , Shau-Yin Tseng , Chin-Feng Lai , Chia-Yu Ho , Ying-Hsun Lai

Aerial manipulators, which combine robotic arms with multi-rotor drones, face strict constraints on arm weight and mechanical complexity. In this work, we study a lightweight 2-degree-of-freedom (DoF) arm mounted on a quadrotor via a…

机器人学 · 计算机科学 2026-03-12 Shlok Deshmukh , Javier Alonso-Mora , Sihao Sun

This paper presented a deep reinforcement learning method named Double Deep Q-networks to design an end-to-end vision-based adaptive cruise control (ACC) system. A simulation environment of a highway scene was set up in Unity, which is a…

计算机视觉与模式识别 · 计算机科学 2020-01-28 Zhensong Wei , Yu Jiang , Xishun Liao , Xuewei Qi , Ziran Wang , Guoyuan Wu , Peng Hao , Matthew Barth

Learning-based control approaches like reinforcement learning (RL) have recently produced a slew of impressive results for tasks like quadrotor trajectory tracking and drone racing. Naturally, it is common to demonstrate the advantages of…

机器人学 · 计算机科学 2025-06-24 Pratik Kunapuli , Jake Welde , Dinesh Jayaraman , Vijay Kumar

Balancing mutually diverging performance metrics, such as, processing latency, outcome accuracy, and end device energy consumption is a challenging undertaking for deep learning model inference in ad-hoc edge environments. In this paper, we…

分布式、并行与集群计算 · 计算机科学 2024-10-17 Motahare Mounesan , Xiaojie Zhang , Saptarshi Debroy

Reinforcement learning for training end-to-end autonomous driving models in closed-loop simulations is gaining growing attention. However, most simulation environments differ significantly from real-world conditions, creating a substantial…

计算机视觉与模式识别 · 计算机科学 2025-08-22 Chaojun Ni , Guosheng Zhao , Xiaofeng Wang , Zheng Zhu , Wenkang Qin , Xinze Chen , Guanghong Jia , Guan Huang , Wenjun Mei

The focus of this paper is behavior modeling for pilots of unmanned aerial vehicles. The pilot is assumed to make decisions that optimize an unknown cost functional, which is estimated from observed trajectories using a novel inverse…

系统与控制 · 电气工程与系统科学 2023-07-26 Jared Town , Zachary Morrison , Rushikesh Kamalapurkar

Reinforcement learning (RL) algorithms can enable high-maneuverability in unmanned aerial vehicles (MAVs), but transferring them from simulation to real-world use is challenging. Variable-pitch propeller (VPP) MAVs offer greater agility,…

机器人学 · 计算机科学 2025-04-11 Zhikun Wang , Shiyu Zhao

The nonlinear and unstable aerodynamic interference generated by the tandem wings of such biomimetic systems poses substantial challenges for motion control, especially under multiple random operating conditions. To address these…

机器学习 · 计算机科学 2024-12-23 Zhang Minghao , Song Bifeng , Yang Xiaojun , Wang Liang

Deep Reinforcement learning has shown to be a powerful tool for developing policies in environments where an optimal solution is unclear. In this paper, we attempt to apply Twin Delayed Deep Deterministic Policy Gradients to train a neural…

机器人学 · 计算机科学 2024-12-20 Patrick Thomas , Kevin Schroeder , Jonathan Black

We study the inverse reinforcement learning (IRL) problem under a transition dynamics mismatch between the expert and the learner. Specifically, we consider the Maximum Causal Entropy (MCE) IRL learner model and provide a tight upper bound…

机器学习 · 计算机科学 2021-12-01 Luca Viano , Yu-Ting Huang , Parameswaran Kamalaruban , Adrian Weller , Volkan Cevher

Compact quadrupedal robots are proving increasingly suitable for deployment in real-world scenarios. Their smaller size fosters easy integration into human environments. Nevertheless, real-time locomotion on uneven terrains remains…

机器人学 · 计算机科学 2026-02-20 Davide Plozza , Patricia Apostol , Paul Joseph , Simon Schläpfer , Michele Magno

Successful machine learning involves a complete pipeline of data, model, and downstream applications. Instead of treating them separately, there has been a prominent increase of attention within the constrained optimization (CO) and machine…

机器学习 · 计算机科学 2023-12-27 Wangkun Xu , Jianhong Wang , Fei Teng

In this paper, we study the use of robust model independent bounded extremum seeking (ES) feedback control to improve the robustness of deep reinforcement learning (DRL) controllers for a class of nonlinear time-varying systems. DRL has the…

机器学习 · 计算机科学 2026-03-11 Shaifalee Saxena , Alan Williams , Rafael Fierro , Alexander Scheinker

In this work, a novel, end-to-end motion planning method is proposed for quadrotor navigation in cluttered environments. The proposed method circumvents the explicit sensing-reconstructing-planning in contrast to conventional navigation…

机器人学 · 计算机科学 2019-10-08 Efe Camci , Erdal Kayacan

End-to-end approaches to autonomous driving commonly rely on expert demonstrations. Although humans are good drivers, they are not good coaches for end-to-end algorithms that demand dense on-policy supervision. On the contrary, automated…

计算机视觉与模式识别 · 计算机科学 2021-10-06 Zhejun Zhang , Alexander Liniger , Dengxin Dai , Fisher Yu , Luc Van Gool