中文
相关论文

相关论文: Practical Reinforcement Learning For MPC: Learning…

200 篇论文

This paper shows that the optimal policy and value functions of a Markov Decision Process (MDP), either discounted or not, can be captured by a finite-horizon undiscounted Optimal Control Problem (OCP), even if based on an inexact model.…

系统与控制 · 电气工程与系统科学 2023-02-08 Arash Bahari Kordabad , Mario Zanon , Sebastien Gros

Lane change in dense traffic typically requires the recognition of an appropriate opportunity for maneuvers, which remains a challenging problem in self-driving. In this work, we propose a chance-aware lane-change strategy with high-level…

机器人学 · 计算机科学 2024-02-19 Yubin Wang , Yulin Li , Zengqi Peng , Hakim Ghazzai , Jun Ma

In this paper we present a framework for risk-sensitive model predictive control (MPC) of linear systems affected by stochastic multiplicative uncertainty. Our key innovation is to consider a time-consistent, dynamic risk evaluation of the…

最优化与控制 · 数学 2018-04-26 Sumeet Singh , Yin-Lam Chow , Anirudha Majumdar , Marco Pavone

In this article, a model predictive control (MPC) method is proposed for constrained linear systems to track bounded references with arbitrary dynamics. Besides control inputs to be determined, artificial reference is introduced as…

系统与控制 · 电气工程与系统科学 2025-03-27 Shibo Han , Bonan Hou , Yuhao Zhang , Xiaotong Shi , Xingwei Zhao

Model Predictive Control (MPC) offers a versatile framework for constraint handling and multi-objective optimisation, yet practical application faces challenges regarding initial and recursive feasibility, robustness against model…

最优化与控制 · 数学 2026-02-27 Dario Dennstädt

Human demonstration data is often ambiguous and incomplete, motivating imitation learning approaches that also exhibit reliable planning behavior. A common paradigm to perform planning-from-demonstration involves learning a reward function…

Model Predictive Control (MPC) can be applied to safety-critical control problems, providing closed-loop safety and performance guarantees. Implementation of MPC controllers requires solving an optimization problem at every sampling…

系统与控制 · 电气工程与系统科学 2025-03-27 Nicolas Chatzikiriakos , Kim P. Wabersich , Felix Berkel , Patricia Pauli , Andrea Iannelli

We present an online model-based reinforcement learning algorithm suitable for controlling complex robotic systems directly in the real world. Unlike prevailing sim-to-real pipelines that rely on extensive offline simulation and model-free…

机器人学 · 计算机科学 2026-05-07 Fang Nan , Hao Ma , Qinghua Guan , Josie Hughes , Michael Muehlebach , Marco Hutter

Advanced control strategies like Model Predictive Control (MPC) offer significant energy savings for HVAC systems but often require substantial engineering effort, limiting scalability. Reinforcement Learning (RL) promises greater…

系统与控制 · 电气工程与系统科学 2025-10-03 Ozan Baris Mulayim , Elias N. Pergantis , Levi D. Reyes Premer , Bingqing Chen , Guannan Qu , Kevin J. Kircher , Mario Bergés

Model-based reinforcement learning has shown promise for improving sample efficiency and decision-making in complex environments. However, existing methods face challenges in training stability, robustness to noise, and computational…

机器学习 · 计算机科学 2024-10-08 Yutaka Shimizu , Masayoshi Tomizuka

This paper proposes a real-time model predictive control (MPC) scheme to execute multiple tasks using robots over a finite-time horizon. In industrial robotic applications, we must carefully consider multiple constraints for avoiding joint…

机器人学 · 计算机科学 2022-09-27 Jaemin Lee , Mingyo Seo , Andrew Bylard , Robert Sun , Luis Sentis

Recent work in Offline Reinforcement Learning (RL) has shown that a unified Transformer trained under a masked auto-encoding objective can effectively capture the relationships between different modalities (e.g., states, actions, rewards)…

机器学习 · 计算机科学 2025-02-07 Kehan Wen , Yutong Hu , Yao Mu , Lei Ke

This paper studies the optimal control problem for discrete-time nonlinear systems and an approximate dynamic programming-based Model Predictive Control (MPC) scheme is proposed for minimizing a quadratic performance measure. In the…

系统与控制 · 电气工程与系统科学 2023-12-12 Keerthi Chacko , Midhun T. Augustine , S. Janardhanan , Deepak U. Patil , I. N. Kar

Buildings sector is one of the major consumers of energy in the United States. The buildings HVAC (Heating, Ventilation, and Air Conditioning) systems, whose functionality is to maintain thermal comfort and indoor air quality (IAQ), account…

系统与控制 · 电气工程与系统科学 2021-03-24 Chi Zhang , Sanmukh R. Kuppannagari , Rajgopal Kannan , Viktor K. Prasanna

Model predictive control (MPC) is a popular control engineering practice, but requires a sound knowledge of the model. Model-free predictive control (MFPC), a burning issue today, also related to reinforcement learning (RL) in AI, is…

系统与控制 · 电气工程与系统科学 2025-04-23 Cédric Join , Emmanuel Delaleau , Michel Fliess

The reinforcement learning (RL) and model predictive control (MPC) communities have developed vast ecosystems of theoretical approaches and computational tools for solving optimal control problems. Given their conceptual similarities but…

机器学习 · 计算机科学 2025-09-04 Nathan P. Lawrence , Thomas Banker , Ali Mesbah

A model predictive control (MPC) scheme for a permanent-magnet synchronous motor (PMSM) is presented. The torque controller optimizes a quadratic cost consisting of control error and machine losses repeatedly, accounting the voltage and…

系统与控制 · 计算机科学 2013-01-01 Jean-Francois Stumper , Alexander Dötlinger , Ralph Kennel

State-of-the-art model-based Reinforcement Learning (RL) approaches either use gradient-free, population-based methods for planning, learned policy networks, or a combination of policy networks and planning. Hybrid approaches that combine…

机器学习 · 计算机科学 2026-05-25 Jonathan Spieler , Sven Behnke

With a growing interest in data-driven control techniques, Model Predictive Control (MPC) provides an opportunity to exploit the surplus of data reliably, particularly while taking safety and stability into account. In many real-world and…

人工智能 · 计算机科学 2021-06-04 Mayank Mittal , Marco Gallieri , Alessio Quaglino , Seyed Sina Mirrazavi Salehian , Jan Koutník

Achieving safe and coordinated behavior in dynamic, constraint-rich environments remains a major challenge for learning-based control. Pure end-to-end learning often suffers from poor sample efficiency and limited reliability, while…

系统与控制 · 电气工程与系统科学 2025-10-10 Max Studt , Georg Schildbach