中文
相关论文

相关论文: Uniform value in Dynamic Programming

200 篇论文

This article studies the solutions of time-dependent differential inclusions which is motivated by their utility in the modeling of certain physical systems. The differential inclusion is described by a time-dependent set-valued mapping…

最优化与控制 · 数学 2021-07-05 Kanat Camlibel , Luigi Iannelli , Aneel Tanwani

The paper is concerned with two-person dynamic zero-sum games. We investigate the limit of value functions of finite horizon games with long run average cost as the time horizon tends to infinity, and the limit of value functions of…

最优化与控制 · 数学 2016-07-21 Dmitry Khlopin

We consider an infinite horizon dynamic mechanism design problem with interdependent valuations. In this setting the type of each agent is assumed to be evolving according to a first order Markov process and is independent of the types of…

计算机科学与博弈论 · 计算机科学 2015-06-26 Swaprava Nath , Onno Zoeter , Y. Narahari , Christopher R. Dance

We obtain the dynamic programming equations and optimality conditions akin to Pontryagin's extremum principle for certain mathematical models of hybrid control systems.

最优化与控制 · 数学 2007-05-23 S. A. Belbas

In many scenarios, it is natural to model a plant's dynamical behavior using a hybrid dynamical system influenced by exogenous continuous-time inputs. While solution concepts and analytical tools for existence and completeness are well…

系统与控制 · 电气工程与系统科学 2026-01-19 W. P. M. H. Heemels , R. Postoyan , P. Bernard , K. J. A. Scheres , R. G. Sanfelice

Markov automata combine non-determinism, probabilistic branching, and exponentially distributed delays. This compositional variant of continuous-time Markov decision processes is used in reliability engineering, performance evaluation and…

计算机科学中的逻辑 · 计算机科学 2017-05-11 Tim Quatmann , Sebastian Junges , Joost-Pieter Katoen

Mechanism design is a well-established game-theoretic paradigm for designing games to achieve desired outcomes. This paper addresses a closely related but distinct concept, equilibrium design. Unlike mechanism design, the designer's…

计算机科学与博弈论 · 计算机科学 2024-08-20 Muhammad Najib , Giuseppe Perelli

In \emph{zero-sum two-player hidden stochastic games}, players observe partial information about the state. We address: $(i)$ the existence of the \emph{uniform value}, i.e., a limiting average payoff that both players can guarantee for…

最优化与控制 · 数学 2026-02-09 Krishnendu Chatterjee , David Lurie , Raimundo Saona , Bruno Ziliotto

We address the stability problem for linear switching systems with mode-dependent restrictions on the switching intervals. Their lengths can be bounded as from below (the guaranteed dwell-time) as from above. The upper bounds make this…

最优化与控制 · 数学 2022-06-01 Vladimir Yu. Protasov , Rinat Kamalov

We study the stability properties of linear time-varying systems in continuous time whose system matrix is Metzler with zero row sums. This class of systems arises naturally in the context of distributed decision problems, coordination and…

最优化与控制 · 数学 2007-05-23 Luc Moreau

This paper studies dynamic stochastic optimization problems parametrized by a random variable. Such problems arise in many applications in operations research and mathematical finance. We give sufficient conditions for the existence of…

最优化与控制 · 数学 2011-05-06 Teemu Pennanen , Ari-Pekka Perkkiö

We consider the problem of approximating the reachability probabilities in Markov decision processes (MDP) with uncountable (continuous) state and action spaces. While there are algorithms that, for special classes of such MDP, provide a…

系统与控制 · 电气工程与系统科学 2022-07-13 Kush Grover , Jan Křetínský , Tobias Meggendorfer , Maximilian Weininger

In this paper we investigate possible approaches to study general time-inconsistent optimization problems without assuming the existence of optimal strategy. This leads immediately to the need to refine the concept of time-consistency as…

最优化与控制 · 数学 2016-04-14 Chandrasekhar Karnam , Jin Ma , Jianfeng Zhang

The parametrization theorem is derived in a flat nD pseudo-complex affine space. The pseudo-complex hyperbolic space accomodates n-number of uncompactified time-like extra dimensions with sugnature (s,r), where s and r are the numbers of…

微分几何 · 数学 2010-03-02 Minh Q. Truong

A general model for zero-sum stochastic games with asymmetric information is considered. In this model, each player's information at each time can be divided into a common information part and a private information part. Under certain…

系统与控制 · 电气工程与系统科学 2019-12-25 Dhruva Kartik , Ashutosh Nayyar

We consider the general model of zero-sum repeated games (or stochastic games with signals), and assume that one of the players is fully informed and controls the transitions of the state variable. We prove the existence of the uniform…

最优化与控制 · 数学 2009-04-20 Jérôme Renault

A very simple example of an algorithmic problem solvable by dynamic programming is to maximize, over sets A in {1,2,...,n}, the objective function |A| - \sum_i \xi_i 1(i \in A,i+1 \in A) for given \xi_i > 0. This problem, with random…

概率论 · 数学 2007-10-04 David J. Aldous , Charles Bordenave , Marc Lelarge

We introduce a quantum extension of dynamic programming, a fundamental computational method that efficiently solves recursive problems using memory. Our innovation lies in showing how to coherently generate recursion step unitaries by using…

量子物理 · 物理学 2025-05-09 Jeongrak Son , Marek Gluza , Ryuji Takagi , Nelly H. Y. Ng

This paper is concerned with finite-level quantum memory systems for retaining initial dynamic variables in the presence of external quantum noise. The system variables have an algebraic structure, similar to that of the Pauli matrices, and…

最优化与控制 · 数学 2026-04-01 Igor G. Vladimirov , Ian R. Petersen , Guodong Shi

We consider a stochastic optimal control problem in a market model with temporary and permanent price impact, which is related to an expected utility maximization problem under finite fuel constraint. We establish the initial condition…

数理金融 · 定量金融 2015-10-13 Mourad Lazgham