中文
相关论文

相关论文: A Moreau Envelope Approach for LQR Meta-Policy Est…

200 篇论文

We consider the problem of conformal prediction under covariate shift. Given labeled data from a source domain and unlabeled data from a covariate shifted target domain, we seek to construct prediction sets with valid marginal coverage in…

机器学习 · 统计学 2025-07-02 Sunay Joshi , Shayan Kiyani , George Pappas , Edgar Dobriban , Hamed Hassani

For large uncertain systems, solving model predictive control problems online can be computationally taxing. Using a shorter prediction horizon can help, but may lead to poor performance and instability without appropriate modifications.…

系统与控制 · 电气工程与系统科学 2025-03-05 E. M. Turan , Z. Mdoe , J. Jäschke

As the benchmark of data-driven control methods, the linear quadratic regulator (LQR) problem has gained significant attention. A growing trend is direct LQR design, which finds the optimal LQR gain directly from raw data and bypassing…

系统与控制 · 电气工程与系统科学 2025-03-06 Feiran Zhao , Alessandro Chiuso , Florian Dörfler

Individual agents in a multi-agent system (MAS) may have decoupled open-loop dynamics, but a cooperative control objective usually results in coupled closed-loop dynamics thereby making the control design computationally expensive. The…

系统与控制 · 电气工程与系统科学 2021-03-09 Gangshan Jing , He Bai , Jemin George , Aranya Chakrabortty

Configuring variational quantum algorithms for combinatorial optimization remains a difficult, expert-driven process requiring coordinated choices over solver family, ansatz, objective, and optimizer. We present AutoQResearch, an LLM-guided…

量子物理 · 物理学 2026-04-28 Monit Sharma , Hoong Chuin Lau

While the topic of mean-field games (MFGs) has a relatively long history, heretofore there has been limited work concerning algorithms for the computation of equilibrium control policies. In this paper, we develop a computable policy…

系统与控制 · 电气工程与系统科学 2020-04-07 Muhammad Aneeq uz Zaman , Kaiqing Zhang , Erik Miehling , Tamer Başar

Linear-Quadratic-Gaussian (LQG) control is a fundamental control paradigm that is studied in various fields such as engineering, computer science, economics, and neuroscience. It involves controlling a system with linear dynamics and…

最优化与控制 · 数学 2023-11-02 Bahar Taşkesen , Dan A. Iancu , Çağıl Koçyiğit , Daniel Kuhn

Multi-agent reinforcement learning has been successfully applied to a number of challenging problems. Despite these empirical successes, theoretical understanding of different algorithms is lacking, primarily due to the curse of…

机器学习 · 计算机科学 2021-12-28 Yuwei Luo , Zhuoran Yang , Zhaoran Wang , Mladen Kolar

Classical linear quadratic (LQ) control centers around linear time-invariant (LTI) systems, where the control-state pairs introduce a quadratic cost with time-invariant parameters. Recent advancement in online optimization and control has…

最优化与控制 · 数学 2020-09-30 Ting-Jui Chang , Shahin Shahrampour

Despite its nonconvexity, policy optimization for the Linear Quadratic Regulator (LQR) admits a favorable structural property known as gradient dominance, which facilitates linear convergence of policy gradient methods to the globally…

最优化与控制 · 数学 2026-02-27 Yuto Watanabe , Yang Zheng

We propose a control design method for linear time-invariant systems that iteratively learns to satisfy unknown polyhedral state constraints. At each iteration of a repetitive task, the method constructs an estimate of the unknown…

系统与控制 · 电气工程与系统科学 2023-06-13 Monimoy Bujarbaruah , Charlott Vallon , Francesco Borrelli

Reinforcement learning methods typically use Deep Neural Networks to approximate the value functions and policies underlying a Markov Decision Process. Unfortunately, DNN-based RL suffers from a lack of explainability of the resulting…

系统与控制 · 电气工程与系统科学 2022-05-19 Shambhuraj Sawant , Sebastien Gros

We propose a new reinforcement learning based approach to designing hierarchical linear quadratic regulator (LQR) controllers for heterogeneous linear multi-agent systems with unknown state-space models and separated control objectives. The…

系统与控制 · 电气工程与系统科学 2020-07-29 He Bai , Jemin George , Aranya Chakrabortty

In this paper, network of agents with identical dynamics is considered. The agents are assumed to be fed by self and neighboring output measurements, while the states are not available for measuring. Viewing distributed estimation as dual…

系统与控制 · 电气工程与系统科学 2020-01-17 Eleftherios Vlahakis , George Halikias

This paper addresses the problem of designing control policies for agents with unknown stochastic dynamics and control objectives specified using Linear Temporal Logic (LTL). Recent Deep Reinforcement Learning (DRL) algorithms have aimed to…

机器人学 · 计算机科学 2025-04-23 Jun Wang , Hosein Hasanbeig , Kaiyuan Tan , Zihe Sun , Yiannis Kantaros

This paper addresses the problem of learning optimal control policies for systems with uncertain dynamics and high-level control objectives specified as Linear Temporal Logic (LTL) formulas. Uncertainty is considered in the workspace…

机器人学 · 计算机科学 2024-10-17 Yiannis Kantaros , Jun Wang

This paper introduces an extension of the LQR-tree algorithm, which is a feedback-motion-planning algorithm for stabilizing a system of ordinary differential equations from a bounded set of initial conditions to a goal. The constructed…

系统与控制 · 电气工程与系统科学 2023-03-02 Jiří Fejlek , Stefan Ratschan

We consider the static output feedback control for Linear Quadratic Regulator problems with structured constraints under the assumption that system parameters are unknown. To solve the problem in the model free setting, we propose the…

最优化与控制 · 数学 2023-03-21 Shokichi Takakura , Kazuhiro Sato

In recent years, $Q$-learning has become indispensable for model-free reinforcement learning (MFRL). However, it suffers from well-known problems such as under- and overestimation bias of the value, which may adversely affect the policy…

机器学习 · 计算机科学 2021-02-09 Youngmin Oh , Jinwoo Shin , Eunho Yang , Sung Ju Hwang

This paper presents a novel model-free and fully data-driven policy iteration scheme for quadratic regulation of linear dynamics with state- and input-multiplicative noise. The implementation is similar to the least-squares temporal…

最优化与控制 · 数学 2022-12-05 Peter Coppens , Panagiotis Patrinos