中文
相关论文

相关论文: Data-Driven Structured Policy Iteration for Homoge…

200 篇论文

Intelligent surgical robots have the potential to revolutionize clinical practice by enabling more precise and automated surgical procedures. However, the automation of such robot for surgical tasks remains under-explored compared to recent…

机器人学 · 计算机科学 2026-03-10 Chonlam Ho , Jianshu Hu , Lei Song , Hesheng Wang , Qi Dou , Yutong Ban

Given a discounted cost, we study deterministic discrete-time systems whose inputs are generated by policy iteration (PI). We provide novel near-optimality and stability properties, while allowing for non stabilizing initial policies. That…

最优化与控制 · 数学 2024-03-29 Jonathan de Brusse , Mathieu Granzotto , Romain Postoyan , Dragan Nešić

In this paper, we propose a distributionally robust control synthesis for an agent with stochastic dynamics that interacts with other agents under uncertainties and constraints expressed by signal temporal logic (STL). We formulate the…

系统与控制 · 电气工程与系统科学 2025-03-14 Arash Bahari Kordabad , Eleftherios E. Vlahakis , Lars Lindemann , Sebastien Gros , Dimos V. Dimarogonas , Sadegh Soudjani

Distributed model predictive control methods for uncertain systems often suffer from considerable conservatism and can tolerate only small uncertainties due to the use of robust formulations that are amenable to distributed design and…

系统与控制 · 电气工程与系统科学 2022-03-03 Simon Muntwiler , Kim P. Wabersich , Lukas Hewing , Melanie N. Zeilinger

Distributed control algorithms are known to reduce overall computation time compared to centralized control algorithms. However, they can result in inconsistent solutions leading to the violation of safety-critical constraints. Inconsistent…

系统与控制 · 电气工程与系统科学 2024-11-26 Julius Beerwerth , Maximilian Kloock , Bassam Alrifaee

We present a structured neural network architecture that is inspired by linear time-varying dynamical systems. The network is designed to mimic the properties of linear dynamical systems which makes analysis and control simple. The…

机器人学 · 计算机科学 2018-08-06 Alexander Broad , Ian Abraham , Todd Murphey , Brenna Argall

In optimal control problem, policy iteration (PI) is a powerful reinforcement learning (RL) tool used for designing optimal controller for the linear systems. However, the need for an initial stabilizing control policy significantly limits…

最优化与控制 · 数学 2024-11-13 Zhen Pang , Shengda Tang , Jun Cheng , Shuping He

Motivated by large-scale but computationally constrained settings, e.g., the Internet of Things, we present a novel data-driven distributed control algorithm that is synthesized directly from trajectory data. Our method, data-driven…

系统与控制 · 电气工程与系统科学 2021-12-24 Carmen Amo Alonso , Fengjun Yang , Nikolai Matni

Robust data-driven controllers typically rely on datasets from previous experiments, which embed information on the variability of the system parameters across past operational conditions. Complementarily, data collected online can…

系统与控制 · 电气工程与系统科学 2025-11-19 Ignacio Sanchez , Filiberto Fele , Daniel Limon

Rather than learning new control policies for each new task, it is possible, when tasks share some structure, to compose a "meta-policy" from previously learned policies. This paper reports results from experiments using Deep Reinforcement…

人工智能 · 计算机科学 2017-11-07 Richard Liaw , Sanjay Krishnan , Animesh Garg , Daniel Crankshaw , Joseph E. Gonzalez , Ken Goldberg

In this paper, we assume that an autonomous exosystem generates a reference output, and we consider the problem of designing a distributed data-driven control law for a family of discrete-time heterogeneous LTI agents, connected through a…

系统与控制 · 电气工程与系统科学 2025-08-08 Giulio Fattore , Maria Elena Valcher

Given the recent surge of interest in data-driven control, this paper proposes a two-step method to study robust data-driven control for a parameter-unknown linear time-invariant (LTI) system that is affected by energy-bounded noises.…

系统与控制 · 电气工程与系统科学 2022-03-15 Jiabao He , Xuan Zhang , Feng Xu , Junbo Tan , Xueqian Wang

This paper proposes efficient policy iteration and value iteration algorithms for the continuous-time linear quadratic regulator problem with unmeasurable states and unknown system dynamics, from the perspective of direct data-driven…

系统与控制 · 电气工程与系统科学 2026-03-17 Jun Xie , Yuan-Hua Ni , Yiqin Yang , Bo Xu

Learning the relationships between various entities from time-series data is essential in many applications. Gaussian graphical models have been studied to infer these relationships. However, existing algorithms process data in a batch at a…

机器学习 · 计算机科学 2021-10-04 Tong Yao , Shreyas Sundaram

This paper proposes two cooperative optimal output tracking (COOT) algorithms based on policy iteration (PI) for discrete-time multi-agent systems with unknown model parameters. First, we establish a stabilizing PI framework that can start…

系统与控制 · 电气工程与系统科学 2026-01-27 Dongdong Li , Jiuxiang Dong

The behaviour of many real-world phenomena can be modelled by nonlinear dynamical systems whereby a latent system state is observed through a filter. We are interested in interacting subsystems of this form, which we model by a set of…

机器学习 · 计算机科学 2017-02-20 Oliver M. Cliff , Mikhail Prokopenko , Robert Fitch

In this paper, we mainly investigate an integrated system operating under a software defined network (SDN) protocol. SDN is a new networking paradigm in which network intelligence is centrally administered and data is communicated via…

最优化与控制 · 数学 2018-12-04 Cheng Tan , Wing Shing Wong , Huanshui Zhang

The fundamental lemma from behavioral systems theory yields a data-driven non-parametric system representation that has shown great potential for the data-efficient control of unknown linear and weakly nonlinear systems, even in the…

系统与控制 · 电气工程与系统科学 2024-09-26 Johannes Teutsch , Sebastian Ellmaier , Sebastian Kerz , Dirk Wollherr , Marion Leibold

We introduce the family of limited model information control design methods, which construct controllers by accessing the plant's model in a constrained way, according to a given design graph. We investigate the closed-loop performance…

最优化与控制 · 数学 2013-01-08 Farhad Farokhi , Cedric Langbort , Karl H. Johansson

The aim of this work is to define a planner that enables robust legged locomotion for complex multi-agent systems consisting of several holonomically constrained quadrupeds. To this end, we employ a methodology based on behavioral systems…

机器人学 · 计算机科学 2022-11-15 Randall T Fawcett , Leila Amanzadeh , Jeeseop Kim , Aaron D Ames , Kaveh Akbari Hamed