中文
相关论文

相关论文: On the Equilibrium between Feasible Zone and Uncer…

200 篇论文

In this work we seek for an approach to integrate safety in the learning process that relies on a partly known state-space model of the system and regards the unknown dynamics as an additive bounded disturbance. We introduce a framework for…

机器学习 · 计算机科学 2018-11-12 Stanislav Fedorov , Antonio Candelieri

Safe reinforcement learning is extremely challenging--not only must the agent explore an unknown environment, it must do so while ensuring no safety constraint violations. We formulate this safe reinforcement learning (RL) problem using the…

In this paper, we consider non-convex optimization problems under \textit{unknown} yet safety-critical constraints. Such problems naturally arise in a variety of domains including robotics, manufacturing, and medical procedures, where it is…

机器学习 · 计算机科学 2020-06-25 Mohammad Fereydounian , Zebang Shen , Aryan Mokhtari , Amin Karbasi , Hamed Hassani

Planning safe trajectories under model uncertainty is a fundamental challenge. Robust planning ensures safety by considering worst-case realizations, yet ignores uncertainty reduction and leads to overly conservative behavior. Actively…

机器人学 · 计算机科学 2026-05-26 Kaleb Ben Naveed , Devansh R. Agrawal , Dimitra Panagou

This paper addresses semantic planning problems in unknown environments under perceptual uncertainty. The environment contains multiple unknown semantically labeled regions or objects, and the robot must reach desired locations while…

机器人学 · 计算机科学 2026-02-23 David Smith Sundarsingh , Yifei Li , Tianji Tang , George J. Pappas , Nikolay Atanasov , Yiannis Kantaros

Intelligent agents progress by continually refining their capabilities through actively exploring environments. Yet robot policies often lack sufficient exploration capability due to action mode collapse. Existing methods that encourage…

机器人学 · 计算机科学 2025-09-24 Yang Jin , Jun Lv , Han Xue , Wendi Chen , Chuan Wen , Cewu Lu

An effective approach to exploration in reinforcement learning is to rely on an agent's uncertainty over the optimal policy, which can yield near-optimal exploration strategies in tabular settings. However, in non-tabular settings that…

In Interactive Machine Learning (IML), we iteratively make decisions and obtain noisy observations of an unknown function. While IML methods, e.g., Bayesian optimization and active learning, have been successful in applications, on…

机器学习 · 计算机科学 2019-10-31 Matteo Turchetta , Felix Berkenkamp , Andreas Krause

Robotic tasks involving contact interactions pose significant challenges for trajectory optimization due to discontinuous dynamics. Conventional formulations typically assume deterministic contact events, which limit robustness and…

机器人学 · 计算机科学 2026-02-09 Zhuocheng Zhang , Haizhou Zhao , Xudong Sun , Aaron M. Johnson , Majid Khadiv

We study dynamic relationships in which one party extracts current surplus in ways that degrade the future state, while the counterparty cannot exit but adjusts effort in response. Standard stationary Markov equilibria may sustain collapse…

理论经济学 · 经济学 2026-03-11 Nicholas H. Kirk

Scalable and effective exploration remains a key challenge in reinforcement learning (RL). While there are methods with optimality guarantees in the setting of discrete state and action spaces, these methods cannot be applied in…

机器学习 · 计算机科学 2017-01-30 Rein Houthooft , Xi Chen , Yan Duan , John Schulman , Filip De Turck , Pieter Abbeel

A fundamental challenge in learning to control an unknown dynamical system is to reduce model uncertainty by making measurements while maintaining safety. In this work, we formulate a mathematical definition of what it means to safely learn…

最优化与控制 · 数学 2020-11-25 Amir Ali Ahmadi , Abraar Chaudhry , Vikas Sindhwani , Stephen Tu

Safe navigation in cluttered environments is an important challenge for autonomous systems. Robots navigating through obstacle ridden scenarios need to be able to navigate safely in the presence of obstacles, goals, and ego objects of…

系统与控制 · 电气工程与系统科学 2026-05-05 Omanshu Thapliyal , Malarvizhi Sankaranarayanasamy , Ravigopal Vennelakanti

This paper investigates performance guarantees on coverage-based ergodic exploration methods in environments containing disturbances. Ergodic exploration methods generate trajectories for autonomous robots such that time spent in each area…

机器人学 · 计算机科学 2024-12-11 Henry Berger , Ian Abraham

One of the bottlenecks preventing Deep Reinforcement Learning algorithms (DRL) from real-world applications is how to explore the environment and collect informative transitions efficiently. The present paper describes bounded exploration,…

机器学习 · 计算机科学 2024-12-10 Ting Qiao , Henry Williams , David Valencia , Bruce MacDonald

We study an important yet under-addressed problem of quickly and safely improving policies in online reinforcement learning domains. As its solution, we propose a novel exploration strategy - diverse exploration (DE), which learns and…

机器学习 · 计算机科学 2018-02-26 Andrew Cohen , Lei Yu , Robert Wright

Exploration and mapping of unknown environments is a fundamental task in applications for autonomous robots. In this article, we present a complete framework for deploying MAVs in autonomous exploration missions in unknown subterranean…

Safe offline RL is a promising way to bypass risky online interactions towards safe policy learning. Most existing methods only enforce soft constraints, i.e., constraining safety violations in expectation below thresholds predetermined.…

机器学习 · 计算机科学 2024-01-22 Yinan Zheng , Jianxiong Li , Dongjie Yu , Yujie Yang , Shengbo Eben Li , Xianyuan Zhan , Jingjing Liu

Balancing safety and efficiency when planning in dense traffic is challenging. Interactive behavior planners incorporate prediction uncertainty and interactivity inherent to these traffic situations. Yet, their use of single-objective…

人工智能 · 计算机科学 2021-02-08 Julian Bernhard , Alois Knoll

We study multiplayer quantitative reachability games played on a finite directed graph, where the objective of each player is to reach his target set of vertices as quickly as possible. Instead of the well-known notion of Nash equilibrium…

计算机科学与博弈论 · 计算机科学 2023-06-22 Thomas Brihaye , Véronique Bruyère , Aline Goeminne , Jean-François Raskin , Marie van den Bogaard