中文
相关论文

相关论文: Policy iteration for discrete-time systems with di…

200 篇论文

This paper extends the core results of discrete time infinite horizon dynamic programming to the case of state-dependent discounting. We obtain a condition on the discount factor process under which all of the standard optimality results…

综合经济学 · 经济学 2020-10-15 John Stachurski , Junnan Zhang

A data-based policy for iterative control task is presented. The proposed strategy is model-free and can be applied whenever safe input and state trajectories of a system performing an iterative task are available. These trajectories,…

系统与控制 · 计算机科学 2019-03-22 Ugo Rosolia , Xiaojing Zhang , Francesco Borrelli

Optimal stopping is the problem of determining when to stop a stochastic system in order to maximize reward, which is of practical importance in domains such as finance, operations management and healthcare. Existing methods for…

最优化与控制 · 数学 2022-03-28 Xinyi Guan , Velibor V. Mišić

We present an approach to compute stabilizing controllers for continuous-time linear time-invariant systems directly from an input-output trajectory affected by process and measurement noise. The proposed output-feedback design combines (i)…

系统与控制 · 电气工程与系统科学 2025-11-17 Alessandro Bosso , Marco Borghesi , Andrea Iannelli , Bowen Yi , Giuseppe Notarstefano

Bisimulation metrics define a distance measure between states of a Markov decision process (MDP) based on a comparison of reward sequences. Due to this property they provide theoretical guarantees in value function approximation (VFA). In…

机器学习 · 计算机科学 2022-11-15 Mete Kemertas , Allan Jepson

Providing finite-time probabilistic safety and reach-avoid guarantees is crucial for safety-critical stochastic systems. Existing state-of-the-art barrier methods often rely on a restrictive boundedness assumption for auxiliary functions,…

系统与控制 · 电气工程与系统科学 2026-05-12 Bai Xue , Luke Ong , Dominik Wagner , Peixin Wang

We develop a novel iterative algorithm for locally optimal experimental design under constraints, like budget or performance constraints. It is an adaptive discretization algorithm. In every iteration, a discretized version of the…

最优化与控制 · 数学 2026-04-21 Jochen Schmid , Philipp Seufert , Jan Schwientek , Tobias Seidel , Karl-Heinz Küfer

This paper presents a new safety specification method that is robust against errors in the probability distribution of disturbances. Our proposed distributionally robust safe policy maximizes the probability of a system remaining in a…

最优化与控制 · 数学 2018-10-05 Insoon Yang

We consider the task of estimating a structural model of dynamic decisions by a human agent based upon the observable history of implemented actions and visited states. This problem has an inherent nested structure: in the inner problem, an…

机器学习 · 计算机科学 2024-03-04 Siliang Zeng , Mingyi Hong , Alfredo Garcia

Recently, a framework for controller design of sampled-data nonlinear systems via their approximate discrete-time models has been proposed in the literature. In this paper we develop novel tools that can be used within this framework and…

最优化与控制 · 数学 2007-05-23 Dragan Nesic , Antonio Loria

We study existence and uniqueness of the fixed points solutions of a large class of non-linear variable discounted transfer operators associated to a sequential decision-making process. We establish regularity properties of these solutions,…

动力系统 · 数学 2019-02-20 L. Cioletti , Elismar R. Oliveira

We establish a collection of closed-loop guarantees and propose a scalable optimization algorithm for distributionally robust model predictive control (DRMPC) applied to linear systems, convex constraints, and quadratic costs. Via standard…

最优化与控制 · 数学 2024-11-13 Robert D. McAllister , Peyman Mohajerin Esfahani

Value iteration (VI) is a ubiquitous algorithm for optimal control, planning, and reinforcement learning schemes. Under the right assumptions, VI is a vital tool to generate inputs with desirable properties for the controlled system, like…

最优化与控制 · 数学 2020-11-23 Mathieu Granzotto , Romain Postoyan , Dragan Nešić , Lucian Buşoniu , Jamal Daafouz

We propose a robust data-driven model predictive control (MPC) scheme to control linear time-invariant (LTI) systems. The scheme uses an implicit model description based on behavioral systems theory and past measured trajectories. In…

系统与控制 · 电气工程与系统科学 2021-04-19 Julian Berberich , Johannes Köhler , Matthias A. Müller , Frank Allgöwer

In this paper we develop novel results on self triggering control of nonlinear systems, subject to perturbations and actuation delays. First, considering an unperturbed nonlinear system with bounded actuation delays, we provide conditions…

最优化与控制 · 数学 2011-08-29 M. D. Di Benedetto , S. Di Gennaro , A. D'Innocenzo

Deep Reinforcement Learning (RL) agents often learn policies that achieve the same episodic return yet behave very differently, due to a combination of environmental (random transitions, initial conditions, reward noise) and algorithmic…

机器学习 · 计算机科学 2026-01-06 Dennis Jabs , Aditya Mohan , Marius Lindauer

We study the behaviour of discrete dynamical systems generated by a continuous map $f$ of a compact real interval into itself where at randomly chosen times a function different from $f$ - so called impulse function is applied. We show that…

动力系统 · 数学 2024-10-25 J. Kováč , J. Veselý , K. Janková

In this paper we first study the fixed-time stabilizability of discrete-time switched linear control systems. Using a geometric approach, we derive conditions under which such systems can be stabilized within a prescribed number of steps,…

最优化与控制 · 数学 2026-04-30 Picchiotti Flavio , Thiago Alves Lima , Girard Antoine

We consider synthesis of control policies that maximize the probability of satisfying given temporal logic specifications in unknown, stochastic environments. We model the interaction between the system and its environment as a Markov…

系统与控制 · 计算机科学 2014-05-01 Jie Fu , Ufuk Topcu

We study a sampled-data implementation of linear controllers that depend on the output and its derivatives. First, we consider an LTI system of relative degree $r\ge 2$ that can be stabilized using $r-1$ output derivatives. Then, we…

最优化与控制 · 数学 2019-07-03 Anton Selivanov , Emilia Fridman
‹ 上一页 1 8 9 10 下一页 ›