中文
相关论文

相关论文: Trustworthiness of Optimality Condition Violation …

200 篇论文

This study is aimed at answering the famous question of how the approximation errors at each iteration of Approximate Dynamic Programming (ADP) affect the quality of the final results considering the fact that errors at each iteration…

系统与控制 · 计算机科学 2015-05-18 Ali Heydari

Autonomous racing creates challenging control problems, but Model Predictive Control (MPC) has made promising steps toward solving both the minimum lap-time problem and head-to-head racing. Yet, accurate models of the system are necessary…

系统与控制 · 电气工程与系统科学 2023-11-06 Tommaso Benciolini , Chen Tang , Marion Leibold , Catherine Weaver , Masayoshi Tomizuka , Wei Zhan

The study of convex optimization has historically been concerned with worst-case convergence rates. The development of the Optimized Gradient Method (OGM), due to \citet{drori2012PerformanceOF,Kim2016optimal}, marked a major milestone in…

最优化与控制 · 数学 2026-04-21 Benjamin Grimmer , Kevin Shu , Alex L. Wang

We introduce a new algorithm to solve constrained nonlinear optimal control problem, with an emphasis on low-thrust trajectory in highly nonlinear dynamics. The algorithm, dubbed Pontryagin-Bellman Differential Dynamic Programming (PDDP),…

最优化与控制 · 数学 2026-05-27 Yanis Sidhoum , Kenshiro Oguri

Here, we explore the problem of error propagation mitigation in modular digital twins as a sequential decision process. Building on a companion study that used a Hidden Markov Model (HMM) to infer latent error regimes from surrogate-physics…

机器学习 · 计算机科学 2026-04-27 Annice Najafi , Shokoufeh Mirzaei

Solving hydrologic inverse problems usually requires repetitive forward simulations. One approach to mitigate the computational cost is to build a surrogate model, i.e., an approximate mapping from model parameters (input) to observable…

最优化与控制 · 数学 2015-06-17 Jiangjiang Zhang , Weixuan Li

We introduce an Implicit Game-Theoretic MPC (IGT-MPC), a decentralized algorithm for two-agent motion planning that uses a learned value function that predicts the game-theoretic interaction outcomes as the terminal cost-to-go function in a…

多智能体系统 · 计算机科学 2025-12-05 Hansung Kim , Edward L. Zhu , Chang Seok Lim , Francesco Borrelli

We present an Imitation Learning approach for the control of dynamical systems with a known model. Our policy search method is guided by solutions from MPC. Typical policy search methods of this kind minimize a distance metric between the…

机器人学 · 计算机科学 2020-02-18 Jan Carius , Farbod Farshidian , Marco Hutter

Collisions are common in many dynamical systems with real applications. They can be formulated as hybrid dynamical systems with discontinuities automatically triggered when states transverse certain manifolds. We present an algorithm for…

最优化与控制 · 数学 2025-01-20 Wei Hu , Jihao Long , Yaohua Zang , Weinan E , Jiequn Han

An insider is defined as a team member who covertly deviates from the team's optimal collaborative control strategy in pursuit of a private objective, while maintaining an outward appearance of cooperation. Such insider threats can severely…

最优化与控制 · 数学 2025-12-04 Gehui Xu , Kaiwen Chen , Thomas Parisini , Andreas A. Malikopoulos

Inverse optimal control can be used to characterize behavior in sequential decision-making tasks. Most existing work, however, is limited to fully observable or linear systems, or requires the action signals to be known. Here, we introduce…

机器学习 · 计算机科学 2023-10-31 Dominik Straub , Matthias Schultheis , Heinz Koeppl , Constantin A. Rothkopf

This paper investigates the simultaneous reconstruction of the running cost function and the internal topological structure within the mean-field games (MFG) system utilizing partial boundary data. The inverse problem is notably challenging…

最优化与控制 · 数学 2024-08-20 Ming-Hui Ding , Hongyu Liu , Guang-Hui Zheng

Most learning algorithms with formal regret guarantees essentially rely on trying all possible behaviors, which is problematic when some errors cannot be recovered from. Instead, we allow the learning agent to ask for help from a mentor and…

机器学习 · 计算机科学 2025-09-17 Benjamin Plaut , Juan Liévano-Karim , Hanlin Zhu , Stuart Russell

Game-theoretic approaches are envisioned to bring human-like reasoning skills and decision-making processes for autonomous vehicles (AVs). However, challenges including game complexity and incomplete information still remain to be addressed…

系统与控制 · 电气工程与系统科学 2023-12-01 Mushuang Liu , H. Eric Tseng , Dimitar Filev , Anouck Girard , Ilya Kolmanovsky

We propose a dynamic information manipulation game (DIMG) to investigate the incentives of an information manipulator (IM) to influence the transition rules of a partially observable Markov decision process (POMDP). DIMG is a hierarchical…

最优化与控制 · 数学 2025-07-15 Shutian Liu , Quanyan Zhu

Dynamic games arise when multiple agents with differing objectives control a dynamic system. They model a wide variety of applications in economics, defense, energy systems and etc. However, compared to single-agent control problems, the…

系统与控制 · 电气工程与系统科学 2020-01-08 Bolei Di , Andrew Lamperski

Mechanism design has found considerable application to the construction of agent-interaction protocols. In the standard setting, the type (e.g., utility function) of an agent is not known by other agents, nor is it known by the mechanism…

计算机科学与博弈论 · 计算机科学 2012-07-19 Nathanael Hyafil , Craig Boutilier

We develop an Integral Transformation Method (ITM) for the study of suitable optimal control and differential game models. This allows for a solution to such dynamic problems to be found through solving a family of optimization problems…

The recent mean field game (MFG) formalism has enabled the application of inverse reinforcement learning (IRL) methods in large-scale multi-agent systems, with the goal of inferring reward signals that can explain demonstrated behaviours of…

机器学习 · 计算机科学 2022-02-15 Yang Chen , Libo Zhang , Jiamou Liu , Shuyue Hu

Computing worst-case robust strategies in pursuit-evasion games (PEGs) is time-consuming, especially when real-world factors like partial observability are considered. While important for general security purposes, real-time applicable…

机器学习 · 计算机科学 2026-05-15 Runyu Lu , Ruochuan Shi , Yuanheng Zhu , Dongbin Zhao