中文
相关论文

相关论文: Improved Exploration for Safety-Embedded Different…

200 篇论文

In this paper, the problem of the trajectory design for a group of energy-constrained drones operating in dynamic wireless network environments is studied. In the considered model, a team of drone base stations (DBSs) is dispatched to…

机器学习 · 计算机科学 2020-12-08 Ye Hu , Mingzhe Chen , Walid Saad , H. Vincent Poor , Shuguang Cui

Activation steering has emerged as a powerful method for guiding the behavior of generative models towards desired outcomes such as toxicity mitigation. However, most existing methods apply interventions uniformly across all inputs,…

机器学习 · 计算机科学 2025-12-04 Alex Ferrando , Xavier Suau , Jordi Gonzàlez , Pau Rodriguez

This paper investigates secure Directional Modulation (DM) design enhanced by a rotatable active Reconfigurable Intelligent Surface (RIS). In conventional RIS-assisted DM networks, the security performance gain is limited due to the…

信号处理 · 电气工程与系统科学 2026-03-24 Yongqiang Li , Feng Shu , Shaofan Chen , Yuanyuan Wu , Maolin Li , Zhen Chen , Hao Jiang , Jiangzhou Wang

For safe and flexible navigation in multi-robot systems, this paper presents an enhanced and predictive sampling-based trajectory planning approach in complex environments, the Gradient Field-based Dynamic Window Approach (GF-DWA). Building…

机器人学 · 计算机科学 2025-07-09 Ze Zhang , Yifan Xue , Nadia Figueroa , Knut Åkesson

In this paper, we develop a distributionally robust optimal control approach for differentially private dynamical systems, enabling a plant to securely outsource control computation to an untrusted remote server. We consider a plant that…

系统与控制 · 电气工程与系统科学 2026-03-20 Yeongjun Jang , Kaoru Teranishi , Junsoo Kim

Based on the use of different exponential bases to define class-dependent error bounds, a new and highly efficient asymmetric boosting scheme, coined as AdaBoostDB (Double-Base), is proposed. Supported by a fully theoretical derivation…

计算机视觉与模式识别 · 计算机科学 2015-07-09 Iago Landesa-Vázquez , José Luis Alba-Castro

Temporal difference (TD) learning is a widely used method to evaluate policies in reinforcement learning. While many TD learning methods have been developed in recent years, little attention has been paid to preserving privacy and most of…

机器学习 · 计算机科学 2022-01-26 Canzhe Zhao , Yanjie Ze , Jing Dong , Baoxiang Wang , Shuai Li

Maze navigation is a fundamental challenge in robotics, requiring agents to traverse complex environments efficiently. While the Deep Deterministic Policy Gradient (DDPG) algorithm excels in control tasks, its performance in maze navigation…

机器人学 · 计算机科学 2025-08-08 Wenjie Hu , Ye Zhou , Hann Woei Ho

To perform autonomous driving maneuvers, such as parallel or perpendicular parking, a vehicle requires continual speed and steering adjustments to follow a generated path. In consequence, the path's quality is a limiting factor of the…

系统与控制 · 电气工程与系统科学 2025-05-14 Jason Zalev

Soft robots have garnered significant attention due to their promising applications across various domains. A hallmark of these systems is their bilayer structure, where strain mismatch caused by differential expansion between layers…

机器人学 · 计算机科学 2025-02-04 Jiahao Li , Dezhong Tong , Zhuonan Hao , Yinbo Zhu , Hengan Wu , Mingchao Liu , Weicheng Huang

Autonomous mobile agents often operate in hazardous environments, necessitating an awareness of safety. These agents can have non-linear, stochastic dynamics that must be considered during planning to guarantee bounded risk. Most state of…

机器人学 · 计算机科学 2024-04-11 Marlyse Reeves , Brian C. Williams

The imminent integration of autonomous vehicles and mobile robots in urban settings presents a critical safety challenge for future intelligent transportation systems. This paper addresses the complex problem of coordinating heterogeneous…

多智能体系统 · 计算机科学 2026-05-28 Wenzhe Song , Hao Zhang

Reinforcement learning (RL) has achieved promising results on most robotic control tasks. Safety of learning-based controllers is an essential notion of ensuring the effectiveness of the controllers. Current methods adopt whole consistency…

机器人学 · 计算机科学 2023-07-31 Haotian Xu , Shengjie Wang , Zhaolei Wang , Yunzhe Zhang , Qing Zhuo , Yang Gao , Tao Zhang

Many real-world decision-theoretic planning problems can be naturally modeled with discrete and continuous state Markov decision processes (DC-MDPs). While previous work has addressed automated decision-theoretic planning for DCMDPs,…

人工智能 · 计算机科学 2012-02-20 Scott Sanner , Karina Valdivia Delgado , Leliane Nunes de Barros

This paper presents a safe feedback control framework for nonlinear control-affine systems with parametric uncertainty by leveraging adaptive dynamic programming (ADP) with barrier-state augmentation. The developed ADP-based controller…

For combinatorial optimization problems, model-based approaches such as mixed-integer programming (MIP) and constraint programming (CP) aim to decouple modeling and solving a problem: the 'holy grail' of declarative problem solving. We…

人工智能 · 计算机科学 2024-01-26 Ryo Kuroiwa , J. Christopher Beck

Autonomous vehicle path following performance is one of significant consideration. This paper presents discrete time design of robust PD controlled system with disturbance observer (DOB) and communication disturbance observer (CDOB)…

机器人学 · 计算机科学 2023-06-06 Haoan Wang , Levent Guvenc

Most, if not all, robot navigation systems employ a decomposed planning framework that includes global and local planning. To trade-off onboard computation and plan quality, current systems have to limit all robot dynamics considerations…

机器人学 · 计算机科学 2025-10-08 Yuanjie Lu , Tong Xu , Linji Wang , Nick Hawes , Xuesu Xiao

This paper presents a new formulation for model-free robust optimal regulation of continuous-time nonlinear systems. The proposed reinforcement learning based approach, referred to as incremental adaptive dynamic programming (IADP),…

系统与控制 · 电气工程与系统科学 2022-03-25 Cong Li , Yongchao Wang , Fangzhou Liu , Qingchen Liu , Martin Buss

Recently, deep reinforcement learning (DRL) has emerged as a promising approach for robotic control. However, the deployment of DRL in real-world robots is hindered by its sensitivity to environmental perturbations. While existing whitebox…

机器学习 · 计算机科学 2025-03-27 Zongyuan Zhang , Tianyang Duan , Zheng Lin , Dong Huang , Zihan Fang , Zekai Sun , Ling Xiong , Hongbin Liang , Heming Cui , Yong Cui