中文
相关论文

相关论文: Inverse Optimal Control Adapted to the Noise Chara…

200 篇论文

Inverse optimal control can be used to characterize behavior in sequential decision-making tasks. Most existing work, however, is limited to fully observable or linear systems, or requires the action signals to be known. Here, we introduce…

机器学习 · 计算机科学 2023-10-31 Dominik Straub , Matthias Schultheis , Heinz Koeppl , Constantin A. Rothkopf

In this paper, we consider the inverse optimal control problem for the discrete-time linear quadratic regulator, over finite-time horizons. Given observations of the optimal trajectories, and optimal control inputs, to a linear…

最优化与控制 · 数学 2018-10-31 Han Zhang , Jack Umenberger , Xiaoming Hu

Continuous control and planning remains a major challenge in robotics and machine learning. Neuroscience offers the possibility of learning from animal brains that implement highly successful controllers, but it is unclear how to relate an…

人工智能 · 计算机科学 2019-08-14 Saurabh Daptardar , Paul Schrater , Xaq Pitkow

Stochastic Optimal Control models represent the state-of-the-art in modeling goal-directed human movements. The linear-quadratic sensorimotor (LQS) model based on signal-dependent noise processes in state and output equation is the current…

最优化与控制 · 数学 2023-03-28 Philipp Karg , Simon Stoll , Simon Rothfuß , Sören Hohmann

A fundamental question in neuroscience is how the brain creates an internal model of the world to guide actions using sequences of ambiguous sensory information. This is naturally formulated as a reinforcement learning problem under partial…

机器学习 · 计算机科学 2020-11-02 Minhae Kwon , Saurabh Daptardar , Paul Schrater , Xaq Pitkow

Cost functions have the potential to provide compact and understandable generalizations of motion. The goal of Inverse Optimal Control (IOC) is to analyze an observed behavior which is assumed to be optimal with respect to an unknown cost…

机器人学 · 计算机科学 2021-04-27 John R. Rebula , Stefan Schaal , James Finley , Ludovic Righetti

Given a set of human's decisions that are observed, inverse optimization has been developed and utilized to infer the underlying decision making problem. The majority of existing studies assumes that the decision making problem is with a…

机器学习 · 统计学 2018-08-03 Chaosheng Dong , Bo Zeng

Inverse optimal control, also known as inverse reinforcement learning, is the problem of recovering an unknown reward function in a Markov decision process from expert demonstrations of the optimal policy. We introduce a probabilistic…

机器学习 · 计算机科学 2012-06-22 Sergey Levine , Vladlen Koltun

We consider the problem of estimating the possibly non-convex cost of an agent by observing its interactions with a nonlinear, non-stationary and stochastic environment. For this inverse problem, we give a result that allows to estimate the…

最优化与控制 · 数学 2023-07-24 Émiland Garrabé , Hozefa Jesawada , Carmen Del Vecchio , Giovanni Russo

This work addresses stochastic optimal control problems where the unknown state evolves in continuous time while partial, noisy, and possibly controllable measurements are only available in discrete time. We develop a framework for…

最优化与控制 · 数学 2025-08-19 Christian Bayer , Boualem Djehiche , Eliza Rezvanova , Raul Fidel Tempone

In this paper, we define and solve the Inverse Stochastic Optimal Control (ISOC) problem of the linear-quadratic Gaussian (LQG) and the linear-quadratic sensorimotor (LQS) control model. These Stochastic Optimal Control (SOC) models are…

最优化与控制 · 数学 2022-11-01 Philipp Karg , Simon Stoll , Simon Rothfuß , Sören Hohmann

The problem of continuous inverse optimal control (over finite time horizon) is to learn the unknown cost function over the sequence of continuous control variables from expert demonstrations. In this article, we study this fundamental…

机器学习 · 计算机科学 2022-04-20 Yifei Xu , Jianwen Xie , Tianyang Zhao , Chris Baker , Yibiao Zhao , Ying Nian Wu

The article poses a general model for optimal control subject to information constraints, motivated in part by recent work of Sims and others on information-constrained decision-making by economic agents. In the average-cost optimal control…

最优化与控制 · 数学 2016-02-24 Ehsan Shafieepoorfard , Maxim Raginsky , Sean P. Meyn

Reinforcement learning can acquire complex behaviors from high-level specifications. However, defining a cost function that can be optimized effectively and encodes the correct task is challenging in practice. We explore how inverse optimal…

机器学习 · 计算机科学 2016-05-30 Chelsea Finn , Sergey Levine , Pieter Abbeel

This work studies discrete-time discounted Markov decision processes with continuous state and action spaces and addresses the inverse problem of inferring a cost function from observed optimal behavior. We first consider the case in which…

最优化与控制 · 数学 2024-05-27 Angeliki Kamoutsi , Peter Schmitt-Förster , Tobias Sutter , Volkan Cevher , John Lygeros

In this paper we consider a control problem for a Partially Observable Piecewise Deterministic Markov Process of the following type: After the jump of the process the controller receives a noisy signal about the state and the aim is to…

最优化与控制 · 数学 2021-07-21 Nicole Bäuerle , Dirk Lange

Optimal control of stochastic nonlinear dynamical systems is a major challenge in the domain of robot learning. Given the intractability of the global control problem, state-of-the-art algorithms focus on approximate sequential optimization…

机器学习 · 计算机科学 2020-04-23 Joe Watson , Hany Abdulsamad , Jan Peters

Robust control of complex engineered and biological systems hinges on the integration of feedforward and feedback mechanisms. This is exemplified in neural motor control, where feedforward muscle co-contraction complements sensory-driven…

最优化与控制 · 数学 2026-03-06 Bastien Berret , Frédéric Jean

A new model for controlled sensing for multihypothesis testing is proposed and studied in the sequential setting. This new model, termed {\em controlled Markovian observation} model, exhibits a more complicated memory structure in the…

最优化与控制 · 数学 2014-07-01 Sirin Nitinawarat , Venupogal V. Veeravalli

Inverse Optimal Control (IOC) seeks to recover an unknown cost from expert demonstrations, and it provides a systematic way of modeling experts' decision mechanisms while considering the prior information of the cost functions.…

最优化与控制 · 数学 2025-12-01 Ziliang Wang , Han Zhang , Axel Ringh
‹ 上一页 1 2 3 10 下一页 ›