中文
相关论文

相关论文: Stochastic Trajectory Influence Functions for LQR:…

200 篇论文

When a controller is designed from an identified model, its performance ultimately depends on the trajectories used for identification, but pinpointing which ones help or hurt remains an open problem. We bring influence functions, a data…

系统与控制 · 电气工程与系统科学 2026-03-25 Jiachen Li , Shihao Li , Soovadeep Bakshi , Jiamin Xu , Dongmei Chen

The ability of a brain or a neural network to efficiently learn depends crucially on both the task structure and the learning rule. Previous works have analyzed the dynamical equations describing learning in the relatively simplified…

机器学习 · 计算机科学 2025-02-26 Christian Schmid , James M. Murray

We study the value of stochastic predictions in online optimal control with random disturbances. Prior work provides performance guarantees based on prediction error but ignores the stochastic dependence between predictions and…

最优化与控制 · 数学 2025-06-06 Yiheng Lin , Christopher Yeh , Zaiwei Chen , Adam Wierman

This paper studies the sample complexity of the stochastic Linear Quadratic Regulator when applied to systems with multiplicative noise. We assume that the covariance of the noise is unknown and estimate it using the sample covariance,…

系统与控制 · 电气工程与系统科学 2021-03-05 Peter Coppens , Panagiotis Patrinos

We consider the problem of function estimation in the case where the data distribution may shift between training and test time, and additional information about it may be available at test time. This relates to popular scenarios such as…

机器学习 · 统计学 2013-06-05 Bernhard Schölkopf , Dominik Janzing , Jonas Peters , Kun Zhang

Unlike traditional model-based reinforcement learning approaches that estimate system parameters from data, non-model-based data-driven control learns the optimal policy directly from input-state data without any intermediate model…

最优化与控制 · 数学 2026-05-05 Leilei Cui , Zhong-Ping Jiang , Petter N. Kolm , Grégoire G. Macqueron

The linear quadratic regulator (LQR) problem has reemerged as an important theoretical benchmark for reinforcement learning-based control of complex dynamical systems with continuous state and action spaces. In contrast with nearly all…

机器学习 · 计算机科学 2020-05-04 Benjamin Gravell , Peyman Mohajerin Esfahani , Tyler Summers

Accurate state estimation requires careful consideration of uncertainty surrounding the process and measurement models; these characteristics are usually not well-known and need an experienced designer to select the covariance matrices. An…

机器学习 · 统计学 2025-07-18 Pardha Sai Krishna Ala , Ameya Salvi , Venkat Krovi , Matthias Schmid

We study the problem of system identification for stochastic continuous-time dynamics, based on a single finite-length state trajectory. We present a method for estimating the possibly unstable open-loop matrix by employing properly…

机器学习 · 统计学 2025-09-30 Reza Sadeghi Hafshejani , Mohamad Kazem Shirani Fradonbeh

In this paper we study an imitation and transfer learning setting for Linear Quadratic Gaussian (LQG) control, where (i) the system dynamics, noise statistics and cost function are unknown and expert data is provided (that is, sequences of…

系统与控制 · 电气工程与系统科学 2023-06-23 Taosha Guo , Abed AlRahman Al Makdah , Vishaal Krishnan , Fabio Pasqualetti

This paper presents a one-shot learning approach with performance and robustness guarantees for the linear quadratic regulator (LQR) control of stochastic linear systems. Even though data-based LQR control has been widely considered,…

系统与控制 · 电气工程与系统科学 2024-10-29 Ramin Esmzad , Hamidreza Modares

Influence functions approximate the effect of training samples in test-time predictions and have a wide variety of applications in machine learning interpretability and uncertainty estimation. A commonly-used (first-order) influence…

机器学习 · 计算机科学 2021-02-12 Samyadeep Basu , Philip Pope , Soheil Feizi

The stochastic leverage effect, defined as the standardized covariation between the returns and their related volatility, is analyzed in a stochastic volatility model set-up. A novel estimator of the effect is defined using a pre-estimation…

统计金融 · 定量金融 2021-03-09 Imma Valentina Curato , Simona Sanfelici

Stochastic Optimal Control models represent the state-of-the-art in modeling goal-directed human movements. The linear-quadratic sensorimotor (LQS) model based on signal-dependent noise processes in state and output equation is the current…

最优化与控制 · 数学 2023-03-28 Philipp Karg , Simon Stoll , Simon Rothfuß , Sören Hohmann

Training data attribution (TDA) identifies which training examples most influenced a model's prediction. Influence function methods are a theoretically grounded family of TDA methods and exploit gradients. To overcome the scalability…

机器学习 · 计算机科学 2026-05-15 Shuangqi Li , Hieu Le , Jingyi Xu , Mathieu Salzmann

Stochastic reaction-diffusion models can be analytically studied on complex networks using the linear noise approximation. This is illustrated through the use of a specific stochastic model, which displays traveling waves in its…

统计力学 · 物理学 2015-06-16 Malbor Asllani , Tommaso Biancalani , Duccio Fanelli , Alan J. McKane

The reconstruction and inference of stochastic dynamical systems from data is a fundamental task in inverse problems and statistical learning. While surrogate modeling advances computational methods to approximate these dynamics, standard…

最优化与控制 · 数学 2026-04-14 Nicole Tianjiao Yang

We propose an approach for learning the causal structure in stochastic dynamical systems with a $1$-step functional dependency in the presence of latent variables. We propose an information-theoretic approach that allows us to recover the…

信息论 · 计算机科学 2017-01-25 Saber Salehkaleybar , Jalal Etesami , Negar Kiyavash

How can we explain the predictions of a black-box model? In this paper, we use influence functions -- a classic technique from robust statistics -- to trace a model's prediction through the learning algorithm and back to its training data,…

机器学习 · 统计学 2021-01-01 Pang Wei Koh , Percy Liang

We propose controller synthesis for state regulation problems in which a human operator shares control with an autonomy system, running in parallel. The autonomy system continuously improves over human action, with minimal intervention, and…

系统与控制 · 计算机科学 2019-09-23 Murad Abu-Khalaf , Sertac Karaman , Daniela Rus
‹ 上一页 1 2 3 10 下一页 ›