中文
相关论文

相关论文: Optimizing Mission Planning for Multi-Debris Rende…

200 篇论文

This paper investigates a joint beamforming and resource allocation problem in downlink reconfigurable intelligent surface (RIS)-assisted orthogonal frequency division multiplexing (OFDM) systems to minimize the average delay, where data…

人工智能 · 计算机科学 2025-07-25 Yu Ma , Xiao Li , Chongtao Guo , Le Liang , Michail Matthaiou , Shi Jin

A RL (Reinforcement Learning) algorithm was developed for command automation onboard a 3U CubeSat. This effort focused on the implementation of macro control action RL, a technique in which an onboard agent is provided with compiled…

系统与控制 · 电气工程与系统科学 2025-07-31 Cannon Whitney , Joseph Melville

Multi-UAV pursuit-evasion, where pursuers aim to capture evaders, poses a key challenge for UAV swarm intelligence. Multi-agent reinforcement learning (MARL) has demonstrated potential in modeling cooperative behaviors, but most RL-based…

机器人学 · 计算机科学 2025-07-09 Jiayu Chen , Chao Yu , Guosheng Li , Wenhao Tang , Shilong Ji , Xinyi Yang , Botian Xu , Huazhong Yang , Yu Wang

Efficient load balancing is crucial in cloud computing environments to ensure optimal resource utilization, minimize response times, and prevent server overload. Traditional load balancing algorithms, such as round-robin or least…

分布式、并行与集群计算 · 计算机科学 2024-09-10 Kavish Chawla

Reinforcement learning shows great potential to solve complex contact-rich robot manipulation tasks. However, the safety of using RL in the real world is a crucial problem, since unexpected dangerous collisions might happen when the RL…

机器人学 · 计算机科学 2025-05-27 Xiang Zhu , Shucheng Kang , Jianyu Chen

In this paper, we study a joint detection, mapping and navigation problem for a single unmanned aerial vehicle (UAV) equipped with a low complexity radar and flying in an unknown environment. The goal is to optimize its trajectory with the…

机器人学 · 计算机科学 2020-07-23 Anna Guerra , Francesco Guidi , Davide Dardari , Petar M. Djuric

We study a resource allocation problem for the cooperative aerial-ground vehicle routing application, in which multiple Unmanned Aerial Vehicles (UAVs) with limited battery capacity and multiple Unmanned Ground Vehicles (UGVs) that can also…

This paper presents a Predictive Maneuver Planning with Deep Reinforcement Learning (PMP-DRL) model for maneuver planning. Traditional rule-based maneuver planning approaches often have to improve their abilities to handle the variabilities…

机器人学 · 计算机科学 2023-06-16 Jayabrata Chowdhury , Vishruth Veerendranath , Suresh Sundaram , Narasimhan Sundararajan

Performing autonomous exploration is essential for unmanned aerial vehicles (UAVs) operating in unknown environments. Often, these missions start with building a map for the environment via pure exploration and subsequently using (i.e.…

机器学习 · 计算机科学 2021-05-05 Ashley Peake , Joe McCalmon , Yixin Zhang , Daniel Myers , Sarra Alqahtani , Paul Pauca

Reinforcement learning (RL) has recently been used for solving challenging decision-making problems in the context of automated driving. However, one of the main drawbacks of the presented RL-based policies is the lack of safety guarantees,…

机器人学 · 计算机科学 2021-07-16 Danial Kamran , Yu Ren , Martin Lauer

Multi-access Edge Computing (MEC) addresses computational and battery limitations in devices by allowing them to offload computation tasks. To overcome the difficulties in establishing line-of-sight connections, integrating unmanned aerial…

网络与互联网体系结构 · 计算机科学 2023-12-15 Pyae Sone Aung , Loc X. Nguyen , Yan Kyaw Tun , Zhu Han , Choong Seon Hong

Unmanned aerial vehicles (UAVs) can be utilized as aerial base stations (ABSs) to assist terrestrial infrastructure for keeping wireless connectivity in various emergency scenarios. To maximize the coverage rate of N ground users (GUs) by…

信息论 · 计算机科学 2020-02-06 Jin Qiu , Jiangbin Lyu , Liqun Fu

The objective of this study is to develop a model-free workspace trajectory planner for space manipulators using a Twin Delayed Deep Deterministic Policy Gradient (TD3) agent to enable safe and reliable debris capture. A local control…

机器人学 · 计算机科学 2025-10-09 Vincent Lam , Robin Chhabra

Reinforcement learning (RL) plays a central role in large language model (LLM) post-training. Among existing approaches, Group Relative Policy Optimization (GRPO) is widely used, especially for RL with verifiable rewards (RLVR) fine-tuning.…

Soft real-time applications are becoming increasingly complex, posing significant challenges for scheduling offloaded tasks in edge computing environments while meeting task timing constraints. Moreover, the exponential growth of the search…

机器学习 · 计算机科学 2025-06-11 Amin Avan , Akramul Azim , Qusay Mahmoud

Direct imaging of Earth-like exoplanets is one of the most prominent scientific drivers of the next generation of ground-based telescopes. Typically, Earth-like exoplanets are located at small angular separations from their host stars,…

The proliferation of unmanned aerial vehicles (UAVs) in controlled airspace presents significant risks, including potential collisions, disruptions to air traffic, and security threats. Ensuring the safe and efficient operation of airspace,…

机器人学 · 计算机科学 2025-10-22 Francisco Giral , Ignacio Gómez , Soledad Le Clainche

The significant components of any successful autonomous flight system are task completion and collision avoidance. Most deep learning algorithms successfully execute these aspects under the environment and conditions they are trained.…

Despite the increasing adoption of Deep Reinforcement Learning (DRL) for Autonomous Surface Vehicles (ASVs), there still remain challenges limiting real-world deployment. In this paper, we first integrate buoyancy and hydrodynamics models…

机器人学 · 计算机科学 2024-07-12 Luis F W Batista , Junghwan Ro , Antoine Richard , Pete Schroepfer , Seth Hutchinson , Cedric Pradalier

Unmanned aerial vehicle (UAV)-based base stations offer a promising solution in emergencies where the rapid deployment of cutting-edge networks is crucial for maximizing life-saving potential. Optimizing the strategic positioning of these…

人工智能 · 计算机科学 2025-04-08 Mario Rico Ibanez , Azim Akhtarshenas , David Lopez-Perez , Giovanni Geraci