中文
相关论文

相关论文: A Physics-Informed Machine Learning Framework for …

200 篇论文

Optimal feedback control with implicit Hamiltonians poses a fundamental challenge for learning-based value function methods due to the absence of closed-form optimal control laws. Recent work~\cite{gelphman2025end} introduced an implicit…

最优化与控制 · 数学 2026-04-28 Eric Gelphman , Deepanshu Verma , Nicole Tianjiao Yang , Stanley Osher , Samy Wu Fung

Infinite-time nonlinear optimal regulation control is widely utilized in aerospace engineering as a systematic method for synthesizing stable controllers. However, conventional methods often rely on linearization hypothesis, while recent…

系统与控制 · 电气工程与系统科学 2025-06-13 Han Wang , Di Wu , Lin Cheng , Shengping Gong , Xu Huang

Autonomous vehicles with a self-evolving ability are expected to cope with unknown scenarios in the real-world environment. Take advantage of trial and error mechanism, reinforcement learning is able to self evolve by learning the optimal…

机器人学 · 计算机科学 2024-08-23 Shuo Yang , Liwen Wang , Yanjun Huang , Hong Chen

Learning high-performance control policies that remain consistent with expert behavior is a fundamental challenge in robotics. Reinforcement learning can discover high-performing strategies but often departs from desirable human behavior,…

机器人学 · 计算机科学 2026-04-06 Siwei Ju , Jan Tauberschmidt , Oleg Arenz , Peter van Vliet , Jan Peters

Safety assurance is a critical yet challenging aspect when developing self-driving technologies. Hamilton-Jacobi backward-reachability analysis is a formal verification tool for verifying the safety of dynamic systems in the presence of…

机器人学 · 计算机科学 2021-06-08 Ran Tian , Anjian Li , Masayoshi Tomizuka , Liting Sun

Autonomous vehicles rely on machine learning to solve challenging tasks in perception and motion planning. However, automotive software safety standards have not fully evolved to address the challenges of machine learning safety such as…

机器学习 · 计算机科学 2019-12-23 Sina Mohseni , Mandar Pitale , Vasu Singh , Zhangyang Wang

This paper proposes a novel learning-based framework for autonomous driving based on the concept of maximal safety probability. Efficient learning requires rewards that are informative of desirable/undesirable states, but such rewards are…

机器人学 · 计算机科学 2024-09-06 Hikaru Hoshino , Jiaxing Li , Arnav Menon , John M. Dolan , Yorie Nakahira

Providing safety guarantees for Autonomous Vehicle (AV) systems with machine-learning-based controllers remains a challenging issue. In this work, we propose Simplex-Drive, a framework that can achieve runtime safety assurance for…

机器人学 · 计算机科学 2021-09-29 Shengduo Chen , Yaowei Sun , Dachuan Li , Qiang Wang , Qi Hao , Joseph Sifakis

This paper presents a new approach for guaranteed safety subject to input constraints (e.g., actuator limits) using a composition of multiple control barrier functions (CBFs). First, we present a method for constructing a single CBF from…

系统与控制 · 电气工程与系统科学 2024-09-10 Pedram Rabiee , Jesse B. Hoagg

In the field of autonomous driving, developing safe and trustworthy autonomous driving policies remains a significant challenge. Recently, Reinforcement Learning with Human Feedback (RLHF) has attracted substantial attention due to its…

机器人学 · 计算机科学 2024-09-06 Zilin Huang , Zihao Sheng , Sikai Chen

Co-optimizing safety and performance in large-scale multi-agent systems remains a fundamental challenge. Existing approaches based on multi-agent reinforcement learning (MARL), safety filtering, or Model Predictive Control (MPC) either lack…

机器人学 · 计算机科学 2025-09-30 Manan Tayal , Aditya Singh , Shishir Kolathaya , Somil Bansal

Latent safety filters extend Hamilton-Jacobi (HJ) reachability to operate on latent state representations and dynamics learned directly from high-dimensional observations, enabling safe visuomotor control under hard-to-model constraints.…

机器人学 · 计算机科学 2025-11-25 Kensuke Nakamura , Arun L. Bishop , Steven Man , Aaron M. Johnson , Zachary Manchester , Andrea Bajcsy

Trustworthy Federated Learning (TFL) typically leverages protection mechanisms to guarantee privacy. However, protection mechanisms inevitably introduce utility loss or efficiency reduction while protecting data privacy. Therefore,…

机器学习 · 计算机科学 2024-02-29 Xiaojin Zhang , Yan Kang , Lixin Fan , Kai Chen , Qiang Yang

Safety-critical whole-body robot control demands reactive methods that ensure collision avoidance in real-time. Complementarity constraints and control barrier functions (CBF) have emerged as core tools for ensuring such safety constraints,…

This paper studies the problem of safe control of sampled-data systems under bounded disturbance and measurement errors with piecewise-constant controllers. To achieve this, we first propose the High-Order Doubly Robust Control Barrier…

系统与控制 · 电气工程与系统科学 2023-09-18 Pradeep Sharma Oruganti , Parinaz Naghizadeh , Qadeer Ahmed

High performance, reliability and safety are crucial properties of any Software-Defined-Networking (SDN) system. Although the use of Deep Reinforcement Learning (DRL) algorithms has been widely studied to improve performance, their…

网络与互联网体系结构 · 计算机科学 2024-10-23 Lam Dinh , Pham Tran Anh Quang , Jérémie Leguay

This work explores a collaborative method for ensuring safety in multi-agent formation control problems. We formulate a control barrier function (CBF) based safety filter control law for a generic distributed formation controller and extend…

机器人学 · 计算机科学 2024-10-08 Brooks A. Butler , Chi Ho Leung , Philip E. Paré

This paper presents a two-stage framework for constrained near-optimal feedback control of input-affine nonlinear systems. An approximate value function for the unconstrained control problem is computed offline by solving the…

系统与控制 · 电气工程与系统科学 2026-03-18 Milad Alipour Shahraki , Laurent Lessard

We develop a game-theoretic framework for adversarially robust optimal safe predefined-time stabilization of parameter-dependent nonlinear dynamical systems with nonquadratic cost functionals. Our approach ensures that all system…

In safe offline reinforcement learning (RL), the objective is to develop a policy that maximizes cumulative rewards while strictly adhering to safety constraints, utilizing only offline data. Traditional methods often face difficulties in…

机器学习 · 计算机科学 2026-02-11 Prajwal Koirala , Zhanhong Jiang , Soumik Sarkar , Cody Fleming