中文
相关论文

相关论文: Minimax Iterative Dynamic Game: Application to Non…

200 篇论文

Influence diagrams are widely employed to represent multi-stage decision problems in which each decision is a choice from a discrete set of alternatives, uncertain chance events have discrete outcomes, and prior decisions may influence the…

最优化与控制 · 数学 2022-01-20 Ahti Salo , Juho Andelmin , Fabricio Oliveira

We study continuous action reinforcement learning problems in which it is crucial that the agent interacts with the environment only through safe policies, i.e.,~policies that do not take the agent to undesirable situations. We formulate…

机器学习 · 计算机科学 2019-02-13 Yinlam Chow , Ofir Nachum , Aleksandra Faust , Edgar Duenez-Guzman , Mohammad Ghavamzadeh

We propose policy gradient algorithms for robust infinite-horizon Markov decision processes (MDPs) with non-rectangular uncertainty sets, thereby addressing an open challenge in the robust MDP literature. Indeed, uncertainty sets that…

最优化与控制 · 数学 2025-09-30 Mengmeng Li , Daniel Kuhn , Tobias Sutter

Multi-stage stochastic programming is a well-established framework for sequential decision making under uncertainty by seeking policies that are fully adapted to the uncertainty. Often such flexible policies are not desirable, and the…

最优化与控制 · 数学 2024-08-06 Beste Basciftci , Shabbir Ahmed , Nagi Gebraeel

Iterative learning control (ILC) is a method for reducing system tracking or estimation errors over multiple iterations by using information from past iterations. The disturbance observer (DOB) is used to estimate and mitigate disturbances…

机器人学 · 计算机科学 2024-04-23 Harsh Modi , Zhu Chen , Xiao Liang , Minghui Zheng

This paper considers robust filtering for a nominal Gaussian state-space model, when a relative entropy tolerance is applied to each time increment of a dynamical model. The problem is formulated as a dynamic minimax game where the…

最优化与控制 · 数学 2011-09-26 Bernard C. Levy , Ramine Nikoukhah

This paper deals with a distributed implementation of minimax adaptive control algorithm for networked dynamical systems modeled by a finite set of linear models. To hedge against the uncertainty arising out of finite number of possible…

系统与控制 · 电气工程与系统科学 2022-10-04 Venkatraman Renganathan , Anders Rantzer , Olle Kjellqvist

We propose a dynamic information manipulation game (DIMG) to investigate the incentives of an information manipulator (IM) to influence the transition rules of a partially observable Markov decision process (POMDP). DIMG is a hierarchical…

最优化与控制 · 数学 2025-07-15 Shutian Liu , Quanyan Zhu

Intelligent agents need a physical understanding of the world to predict the impact of their actions in the future. While learning-based models of the environment dynamics have contributed to significant improvements in sample efficiency…

机器学习 · 计算机科学 2020-05-20 Eric Heiden , David Millard , Hejia Zhang , Gaurav S. Sukhatme

Nonzero-sum stochastic differential games with impulse controls offer a realistic and far-reaching modelling framework for applications within finance, energy markets, and other areas, but the difficulty in solving such problems has…

数值分析 · 数学 2020-06-29 Diego Zabaljauregui

This work introduces a novel control strategy called Iterative Linear Quadratic Regulator for Iterative Tasks (i2LQR), which aims to improve closed-loop performance with local trajectory optimization for iterative tasks in a dynamic…

系统与控制 · 电气工程与系统科学 2023-09-08 Yifan Zeng , Suiyi He , Han Hoang Nguyen , Yihan Li , Zhongyu Li , Koushil Sreenath , Jun Zeng

Model-based reinforcement learning (MBRL) agents typically learn world models by minimizing predictive loss. However, powerful RL optimizers inevitably exploit minor model inaccuracies, leading to simulator exploitation and a reality gap…

机器学习 · 计算机科学 2026-05-29 Christoph Dann , Yishay Mansour , Mehryar Mohri

In this work, we analyze the applicability of Inverse Dynamic Game (IDG) methods based on the Minimum Principle (MP). The IDG method determines unknown cost functions in a single- or multi-agent setting from observed system trajectories by…

最优化与控制 · 数学 2024-06-18 Philipp Karg , Adrian Kienzle , Jonas Kaub , Balint Varga , Sören Hohmann

We present Dynamic ReAct, a novel approach for enabling ReAct agents to efficiently operate with extensive Model Control Protocol (MCP) tool sets that exceed the contextual memory limitations of large language models. Our approach addresses…

软件工程 · 计算机科学 2025-09-29 Nishant Gaurav , Adit Akarsh , Ankit Ranjan , Manoj Bajaj

This paper studies the robustness of reinforcement learning algorithms to errors in the learning process. Specifically, we revisit the benchmark problem of discrete-time linear quadratic regulation (LQR) and study the long-standing open…

最优化与控制 · 数学 2021-03-16 Bo Pang , Zhong-Ping Jiang

Multi-stage forceful manipulation tasks, such as twisting a nut on a bolt, require reasoning over interlocking constraints over discrete as well as continuous choices. The robot must choose a sequence of discrete actions, or strategy, such…

机器人学 · 计算机科学 2021-05-11 Rachel Holladay , Tomás Lozano-Pérez , Alberto Rodriguez

We formalize decision-making problems in robotics and automated control using continuous MDPs and actions that take place over continuous time intervals. We then approximate the continuous MDP using finer and finer discretizations. Doing…

机器人学 · 计算机科学 2020-05-22 Nan Rong , Joseph Y. Halpern , Ashutosh Saxena

We consider large-scale Markov decision processes (MDPs) with parameter uncertainty, under the robust MDP paradigm. Previous studies showed that robust MDPs, based on a minimax approach to handle uncertainty, can be solved using dynamic…

机器学习 · 计算机科学 2013-06-27 Aviv Tamar , Huan Xu , Shie Mannor

Purpose of Review: To effectively synthesise and analyse multi-robot behaviour, we require formal task-level models which accurately capture multi-robot execution. In this paper, we review modelling formalisms for multi-robot systems under…

机器人学 · 计算机科学 2023-08-16 Charlie Street , Masoumeh Mansouri , Bruno Lacerda

This paper introduces a novel data-driven hierarchical control scheme for managing a fleet of nonlinear, capacity-constrained autonomous agents in an iterative environment. We propose a control framework consisting of a high-level dynamic…

机器人学 · 计算机科学 2024-04-12 Charlott Vallon , Alessandro Pinto , Bartolomeo Stellato , Francesco Borrelli