中文
相关论文

相关论文: Efficient Iterative Linear-Quadratic Approximation…

200 篇论文

We formulate and study a class of two-player zero-sum stochastic dynamic games with partial and asymmetric information. Information asymmetry introduces fundamental challenges involving \emph{belief representation} and \emph{theory of mind}…

最优化与控制 · 数学 2026-03-20 Yuxiang Guan , Iman Shames , Tyler Summers

In many settings where multiple agents interact, the optimal choices for each agent depend heavily on the choices of the others. These coupled interactions are well-described by a general-sum differential game, in which players have…

机器人学 · 计算机科学 2020-05-07 Lasse Peters , David Fridovich-Keil , Claire J. Tomlin , Zachary N. Sunberg

We study Stackelberg equilibria in finitely repeated games, where the leader commits to a strategy that picks actions in each round and can be adaptive to the history of play (i.e. they commit to an algorithm). In particular, we study…

计算机科学与博弈论 · 计算机科学 2024-03-08 Natalie Collina , Eshwar Ram Arunachaleswaran , Michael Kearns

We present high order explicit geometric integrators to solve linear-quadratic optimal control problems and $N$-player differential games. These problems are described by a system coupled non-linear differential equations with boundary…

数值分析 · 数学 2013-11-06 Sergio Blanes

Linear-quadratic Gaussian games provide a framework for modeling strategic interactions in multi-agent systems, where agents must estimate system states from noisy observations while also making decisions to optimize a quadratic cost.…

系统与控制 · 电气工程与系统科学 2026-03-19 Tianyu Qiu , Filippos Fotiadis , Xinjie Liu , Christian Ellis , Jesse Milzman , Wesley Suttle , Ufuk Topcu , David Fridovich-Keil

This work presents an algorithmic scheme for solving the infinite-time constrained linear quadratic regulation problem. We employ an accelerated version of a popular proximal gradient scheme, commonly known as the Forward-Backward Splitting…

最优化与控制 · 数学 2015-01-20 Giorgos Stathopoulos , Milan Korda , Colin N. Jones

Trajectory optimization is a popular strategy for planning trajectories for robotic systems. However, many robotic tasks require changing contact conditions, which is difficult due to the hybrid nature of the dynamics. The optimal sequence…

机器人学 · 计算机科学 2021-09-08 Nathan J. Kong , George Council , Aaron M. Johnson

This paper is concerned with a linear-quadratic (LQ) leader-follower differential game with mixed deterministic and stochastic controls. In the game, the follower is a random controller which means that the follower can choose adapted…

最优化与控制 · 数学 2025-09-26 Jingtao Shi , Guangchen Wang

Infinitely repeated games can support cooperative outcomes that are not equilibria in the one-shot game. The idea is to make sure that any gains from deviating will be offset by retaliation in future rounds. However, this model of…

计算机科学与博弈论 · 计算机科学 2024-06-04 Ratip Emin Berker , Vincent Conitzer

Designing the optimal linear quadratic regulator (LQR) for a large-scale multi-agent system (MAS) is time-consuming since it involves solving a large-size matrix Riccati equation. The situation is further exasperated when the design needs…

系统与控制 · 电气工程与系统科学 2021-03-18 Gangshan Jing , He Bai , Jemin George , Aranya Chakrabortty

An open problem in linear quadratic (LQ) games has been characterizing the Nash equilibria. This problem has renewed relevance given the surge of work on understanding the convergence of learning algorithms in dynamic games. This paper…

计算机科学与博弈论 · 计算机科学 2025-04-18 Giulio Salizzoni , Reda Ouhamma , Maryam Kamgarpour

This paper introduces a generalization of the well-known Riccati recursion for solving the discrete-time equality-constrained linear quadratic optimal control problem. The recursion can be used to compute the solutions as well as optimal…

最优化与控制 · 数学 2024-12-31 Lander Vanroye , Joris De Schutter , Wilm Decré

In this paper, we address Linear Quadratic Regulator (LQR) problems through a novel iterative algorithm named EXtremum-seeking Policy iteration LQR (EXP-LQR). The peculiarity of EXP-LQR is that it only needs access to a truncated…

最优化与控制 · 数学 2025-06-13 Guido Carnevale , Nicola Mimmo , Giuseppe Notarstefano

This paper proposes efficient policy iteration and value iteration algorithms for the continuous-time linear quadratic regulator problem with unmeasurable states and unknown system dynamics, from the perspective of direct data-driven…

系统与控制 · 电气工程与系统科学 2026-03-17 Jun Xie , Yuan-Hua Ni , Yiqin Yang , Bo Xu

Robust Reinforcement Learning (RRL) is a promising Reinforcement Learning (RL) paradigm aimed at training robust to uncertainty or disturbances models, making them more efficient for real-world applications. Following this paradigm,…

机器学习 · 计算机科学 2024-05-06 Anton Plaksin , Vitaly Kalev

In this article, we study the repeated routing game problem on a parallel network with affine latency functions on each edge. We cast the game setup in a LQR control theoretic framework, leveraging the Rosenthal potential formulation. We…

最优化与控制 · 数学 2022-01-03 Marsalis Gibson , Yiling You , Alexandre Bayen

The goal of this paper is to investigate new and simple convergence analysis of dynamic programming for linear quadratic regulator problem of discrete-time linear time-invariant systems. In particular, bounds on errors are given in terms of…

最优化与控制 · 数学 2021-06-18 Donghwan Lee

We propose controller synthesis for state regulation problems in which a human operator shares control with an autonomy system, running in parallel. The autonomy system continuously improves over human action, with minimal intervention, and…

系统与控制 · 计算机科学 2019-09-23 Murad Abu-Khalaf , Sertac Karaman , Daniela Rus

There has been significant recent progress in algorithms for approximation of Nash equilibrium in large two-player zero-sum imperfect-information games and exact computation of Nash equilibrium in multiplayer strategic-form games. While…

计算机科学与博弈论 · 计算机科学 2025-10-01 Sam Ganzfried

Quadratic programs arise in robotics, communications, smart grids, and many other applications. As these problems grow in size, finding solutions becomes more computationally demanding, and new algorithms are needed to efficiently solve…

最优化与控制 · 数学 2020-06-17 Matthew Ubl , Matthew T. Hale