中文
相关论文

相关论文: Guidance Design for Escape Flight Vehicle Using Ev…

200 篇论文

Safe reinforcement learning (RL) aims to learn policies that satisfy certain constraints before deploying them to safety-critical applications. Previous primal-dual style approaches suffer from instability issues and lack optimality…

机器学习 · 计算机科学 2022-06-20 Zuxin Liu , Zhepeng Cen , Vladislav Isenbaev , Wei Liu , Zhiwei Steven Wu , Bo Li , Ding Zhao

This paper considers the problem of autonomous mobile robot navigation in unknown environments with moving obstacles. We propose a new method to achieve environment-aware safe tracking (EAST) of robot motion plans that integrates an…

机器人学 · 计算机科学 2025-11-17 Zhichao Li , Yinzhuang Yi , Zhuolin Niu , Nikolay Atanasov

Employing unmanned aerial vehicles (UAVs) has attracted growing interests and emerged as the state-of-the-art technology for data collection in Internet-of-Things (IoT) networks. In this paper, with the objective of minimizing the total…

系统与控制 · 电气工程与系统科学 2021-12-02 Botao Zhu , Ebrahim Bedeer , Ha H. Nguyen , Robert Barton , Jerome Henry

Direct Preference Optimization (DPO) improves the alignment of large language models (LLMs) with human values by training directly on human preference datasets, eliminating the need for reward models. However, due to the presence of…

人工智能 · 计算机科学 2024-06-11 Biqing Qi , Pengfei Li , Fangyuan Li , Junqi Gao , Kaiyan Zhang , Bowen Zhou

Reinforcement learning (RL) has emerged as a promising strategy for improving the reasoning capabilities of language models (LMs) in domains such as mathematics and coding. However, most modern RL algorithms were designed to target robotics…

人工智能 · 计算机科学 2025-05-26 Lianghuan Huang , Shuo Li , Sagnik Anupam , Insup Lee , Osbert Bastani

Large language model (LLM)-based agents frequently generate seemingly coherent plans that fail upon execution due to infeasible actions, constraint violations, and compounding errors over extended horizons. PIVOT (Plan-Inspect-eVOlve…

人工智能 · 计算机科学 2026-05-13 Tuo Zhang , Alin-Ionut Popa , Yan Xu , Rui Song , Dimitrios Dimitriadis

To continuously generate trajectories for serial manipulators with high dimensional degrees of freedom (DOF) in the dynamic environment, a real-time optimal trajectory generation method based on machine learning aiming at high dimensional…

机器人学 · 计算机科学 2018-12-19 Shiyu Zhang , Shuling Dai

Recent focus on robustness to adversarial attacks for deep neural networks produced a large variety of algorithms for training robust models. Most of the effective algorithms involve solving the min-max optimization problem for training…

机器学习 · 计算机科学 2021-03-03 Yasaman Esfandiari , Aditya Balu , Keivan Ebrahimi , Umesh Vaidya , Nicola Elia , Soumik Sarkar

Evolution Strategies (ES) is a class of powerful black-box optimisation methods that are highly parallelisable and can handle non-differentiable and noisy objectives. However, na\"ive ES becomes prohibitively expensive at scale on GPUs due…

We present RLSS: a reinforcement learning algorithm for sequential scene generation. This is based on employing the proximal policy optimization (PPO) algorithm for generative problems. In particular, we consider how to effectively reduce…

计算机视觉与模式识别 · 计算机科学 2022-06-07 Azimkhon Ostonov , Peter Wonka , Dominik L. Michels

Deep Reinforcement Learning is quickly becoming a popular method for training autonomous Unmanned Aerial Vehicles (UAVs). Our work analyzes the effects of measurement uncertainty on the performance of Deep Reinforcement Learning (DRL) based…

机器人学 · 计算机科学 2023-03-14 Bhaskar Joshi , Dhruv Kapur , Harikumar Kandath

This paper tackles the challenging task of maintaining formation among multiple unmanned aerial vehicles (UAVs) while avoiding both static and dynamic obstacles during directed flight. The complexity of the task arises from its…

机器人学 · 计算机科学 2025-03-04 Yuqing Xie , Chao Yu , Hongzhi Zang , Feng Gao , Wenhao Tang , Jingyi Huang , Jiayu Chen , Botian Xu , Yi Wu , Yu Wang

We consider a task of surveillance-evading path-planning in a continuous setting. An Evader strives to escape from a 2D domain while minimizing the risk of detection (and immediate capture). The probability of detection is path-dependent…

机器学习 · 计算机科学 2023-02-24 Dongping Qi , David Bindel , Alexander Vladimirsky

Designing continuous trajectories whose time-averaged occupancy provably matches a prescribed spatial density (the \emph{ergodic coverage} problem) is central to UAV-assisted data collection and sensing, robotic exploration, and mobile…

机器学习 · 计算机科学 2026-05-14 Ehsan Aghazadeh , Masoud Malekzadeh , Ahmad Ghasemi , Hossein Pishro-Nik

This paper focuses on the continuous control of the unmanned aerial vehicle (UAV) based on a deep reinforcement learning method for a large-scale 3D complex environment. The purpose is to make the UAV reach any target point from a certain…

机器人学 · 计算机科学 2023-04-13 Xuyang Li , Jianwu Fang , Kai Du , Kuizhi Mei , Jianru Xue

Dynamic flexible assembly flow shop scheduling with multi-product delivery is a critical combinatorial problem, characterized by kitting supply and machine flexibility. Genetic programming is widely used to automatically generate…

神经与进化计算 · 计算机科学 2026-04-01 Junhao Qiu , Haoyang Zhuang , Fei Liu , Jianjun Liu , Qingfu Zhang

Generative diffusion models for end-to-end autonomous driving often suffer from mode collapse, tending to generate conservative and homogeneous behaviors. While DiffusionDrive employs predefined anchors representing different driving…

计算机视觉与模式识别 · 计算机科学 2025-12-09 Jialv Zou , Shaoyu Chen , Bencheng Liao , Zhiyu Zheng , Yuehao Song , Lefei Zhang , Qian Zhang , Wenyu Liu , Xinggang Wang

In this paper, the trajectory planning problem for autonomous rendezvous and docking between a controlled spacecraft and a tumbling target is addressed. The use of a variable planning horizon is proposed in order to construct an appropriate…

系统与控制 · 电气工程与系统科学 2024-09-11 Mirko Leomanni , Renato Quartullo , Gianni Bianchini , Andrea Garulli , Antonio Giannitrapani

The multiple spacecraft guidance problem for proximity flight in libration point orbit is considered. A nonlinear optimal control problem with continuous-time path constraints enforcing minimum separation between each spacecraft is…

最优化与控制 · 数学 2025-09-19 Yuri Shimane , Purnanand Elango , Avishai Weiss

Unmanned aerial vehicle (UAV)-assisted data collection has been emerging as a prominent application due to its flexibility, mobility, and low operational cost. However, under the dynamic and uncertainty of IoT data collection and energy…

网络与互联网体系结构 · 计算机科学 2021-06-22 Nam H. Chu , Dinh Thai Hoang , Diep N. Nguyen , Nguyen Van Huynh , Eryk Dutkiewicz