中文
相关论文

相关论文: Action-Conditioned Risk Gating for Safety-Critical…

200 篇论文

We propose a method to encourage safety in Model Predictive Control (MPC)-based Reinforcement Learning (RL) via Gaussian Process (GP) regression. This framework consists of 1) a parametric MPC scheme that is employed as model-based…

系统与控制 · 电气工程与系统科学 2024-12-13 Filippo Airaldi , Bart De Schutter , Azita Dabiri

Ensuring safety in autonomous systems with vision-based control remains a critical challenge due to the high dimensionality of image inputs and the fact that the relationship between true system state and its visual manifestation is…

机器人学 · 计算机科学 2025-11-12 Xinhang Ma , Junlin Wu , Hussein Sibai , Yiannis Kantaros , Yevgeniy Vorobeychik

Safety-critical applications require controllers/policies that can guarantee safety with high confidence. The control barrier function is a useful tool to guarantee safety if we have access to the ground-truth system dynamics. In practice,…

机器学习 · 计算机科学 2021-12-30 Athindran Ramesh Kumar , Sulin Liu , Jaime F. Fisac , Ryan P. Adams , Peter J. Ramadge

Active learning of physical systems must commonly respect practical safety constraints, which restricts the exploration of the design space. Gaussian Processes (GPs) and their calibrated uncertainty estimations are widely used for this…

机器学习 · 计算机科学 2024-04-16 Jörn Tebbe , Christoph Zimmer , Ansgar Steland , Markus Lange-Hegermann , Fabian Mies

We present a scheme for sequential decision making with a risk-sensitive objective and constraints in a dynamic environment. A neural network is trained as an approximator of the mapping from parameter space to space of risk and policy with…

人工智能 · 计算机科学 2019-07-10 Shuai Ma , Jia Yuan Yu , Ahmet Satir

Recent advances in learning techniques have garnered attention for their applicability to a diverse range of real-world sequential decision-making problems. Yet, many practical applications have critical constraints for operation in real…

机器学习 · 计算机科学 2024-05-06 Jose A. Ayala-Romero , Andres Garcia-Saavedra , Xavier Costa-Perez

Identifying uncertainty and taking mitigating actions is crucial for safe and trustworthy reinforcement learning agents, especially when deployed in high-risk environments. In this paper, risk sensitivity is promoted in a model-based…

机器学习 · 计算机科学 2021-11-10 Stefan Radic Webster , Peter Flach

Safe decision-making algorithms for control of mobile robots often require the existence of feedback to verify the safety of proposed actions. This feedback is assumed to be directly available during the development or deployment of the…

机器学习 · 计算机科学 2026-05-26 Jeff Pflueger , Michael Everett

To solve multi-step manipulation tasks in the real world, an autonomous robot must take actions to observe its environment and react to unexpected observations. This may require opening a drawer to observe its contents or moving an object…

机器人学 · 计算机科学 2020-03-24 Caelan Reed Garrett , Chris Paxton , Tomás Lozano-Pérez , Leslie Pack Kaelbling , Dieter Fox

Reinforcement learning has been successfully used to solve difficult tasks in complex unknown environments. However, these methods typically do not provide any safety guarantees during the learning process. This is particularly problematic,…

系统与控制 · 电气工程与系统科学 2019-07-02 Torsten Koller , Felix Berkenkamp , Matteo Turchetta , Joschka Boedecker , Andreas Krause

In recent years several learning approaches to point goal navigation in previously unseen environments have been proposed. They vary in the representations of the environments, problem decomposition, and experimental evaluation. In this…

机器人学 · 计算机科学 2022-12-20 Yimeng Li , Arnab Debnath , Gregory J. Stein , Jana Kosecka

When deploying autonomous agents in unstructured environments over sustained periods of time, adaptability and robustness oftentimes outweigh optimality as a primary consideration. In other words, safety and survivability constraints play a…

系统与控制 · 电气工程与系统科学 2021-04-08 Motoya Ohnishi , Gennaro Notomista , Masashi Sugiyama , Magnus Egerstedt

This paper presents an integrated model-learning predictive control scheme for spacecraft orbit-attitude station-keeping in the vicinity of asteroids. The orbiting probe relies on optical and laser navigation while attitude measurements are…

系统与控制 · 电气工程与系统科学 2025-01-23 Julio C. Sanchez , Rafael Vazquez , James D. Biggs , Franco Bernelli-Zazzera

In many practical applications, decision-making processes must balance the costs of acquiring information with the benefits it provides. Traditional control systems often assume full observability, an unrealistic assumption when…

人工智能 · 计算机科学 2025-01-24 Taiyi Wang , Jianheng Liu , Bryan Lee , Zhihao Wu , Yu Wu

We use one-step conditional risk mappings to formulate a risk averse version of a total cost problem on a controlled Markov process in discrete time infinite horizon. The nonnegative one step costs are assumed to be lower semi-continuous…

最优化与控制 · 数学 2018-06-05 Kerem Ugurlu

This paper addresses the problem of risk-aware fixed-time stabilization of a class of uncertain, output-feedback nonlinear systems modeled via stochastic differential equations. First, novel classes of certificate functions, namely…

最优化与控制 · 数学 2024-04-01 Mitchell Black , Georgios Fainekos , Bardh Hoxha , Dimitra Panagou

State of the art reinforcement learning methods sometimes encounter unsafe situations. Identifying when these situations occur is of interest both for post-hoc analysis and during deployment, where it might be advantageous to call out to a…

机器学习 · 计算机科学 2025-05-29 Alexander Grushin , Walt Woods , Alvaro Velasquez , Simon Khan

Occlusion-aware prediction remains a critical challenge in autonomous driving due to the inherent uncertainty of unobserved regions. Existing approaches either overestimate risk based on reachable states or struggle to predict accurate…

机器人学 · 计算机科学 2026-05-22 Jie Jia , Yaofeng Su , Zeyu Bao , Yun Hong , Bingzhao Gao , Zhongxue Gan , Wenchao Ding

We study the problem of system identification and adaptive control in partially observable linear dynamical systems. Adaptive and closed-loop system identification is a challenging problem due to correlations introduced in data collection.…

机器学习 · 计算机科学 2020-06-25 Sahin Lale , Kamyar Azizzadenesheli , Babak Hassibi , Anima Anandkumar

Safety has been recognized as the central obstacle to preventing the use of reinforcement learning (RL) for real-world applications. Different methods have been developed to deal with safety concerns in RL. However, learning reliable…

机器学习 · 计算机科学 2023-02-08 Huiliang Zhang , Di Wu , Benoit Boulet