中文
相关论文

相关论文: Non-Markovian Impulse Control Under Nonlinear Expe…

200 篇论文

We present FilterDDP, a differential dynamic programming algorithm for solving discrete-time, optimal control problems (OCPs) with nonlinear equality constraints. Unlike prior methods based on merit functions or the augmented Lagrangian…

最优化与控制 · 数学 2026-04-16 Ming Xu , Stephen Gould , Iman Shames

Trajectory optimization is an efficient approach for solving optimal control problems for complex robotic systems. It relies on two key components: first the transcription into a sparse nonlinear program, and second the corresponding solver…

机器人学 · 计算机科学 2022-10-31 Wilson Jallet , Antoine Bambade , Nicolas Mansard , Justin Carpentier

For the classical backward induction algorithm, the input is an arbitrary $n$-person positional game with perfect information modeled by a finite acyclic directed graph (digraph) and the output is a profile $(x_1, \ldots, x_n)$ of pure…

组合数学 · 数学 2017-11-21 Vladimir Gurvich

Stochastic games combine controllable and adversarial non-determinism with stochastic behavior and are a common tool in control, verification and synthesis of reactive systems facing uncertainty. Multi-objective stochastic games are natural…

计算复杂性 · 计算机科学 2022-07-21 Tobias Winkler , Maximilian Weininger

In this paper, we study the existence of equilibrium in a single-leader-multiple-follower game with decision-dependent chance constraints (DDCCs), where decision-dependent uncertainties (DDUs) exist in the constraints of followers. DDUs…

最优化与控制 · 数学 2024-08-06 Jingxiang Wang , Zhaojian Wang , Bo Yang , Feng Liu , Xinping Guan

This paper is concerned with stochastic differential games (SDGs) defined through fully coupled forward-backward stochastic differential equations (FBSDEs) which are governed by Brownian motion and Poisson random measure. For SDGs, the…

最优化与控制 · 数学 2013-02-06 Juan Li , Qingmeng Wei

We consider the problem of optimally controlling stochastic, Markovian systems subject to joint chance constraints over a finite-time horizon. For such problems, standard Dynamic Programming is inapplicable due to the time correlation of…

最优化与控制 · 数学 2024-11-22 Niklas Schmid , Marta Fochesato , Sarah H. Q. Li , Tobias Sutter , John Lygeros

This paper considers a distributed stochastic optimization problem where the goal is to minimize the time average of a cost function subject to a set of constraints on the time averages of a related stochastic processes called penalties. We…

信息论 · 计算机科学 2016-10-06 B. N. Bharath , Vaishali P

We consider a class of zero-sum stopper vs. singular-controller games in which the controller can only act on a subset $d_0<d$ of the $d$ coordinates of a controlled diffusion. Due to the constraint on the control directions these games…

最优化与控制 · 数学 2024-02-02 Andrea Bovo , Tiziano De Angelis , Jan Palczewski

This paper proposes and studies a general form of dynamic $N$-player non-cooperative games called $\alpha$-potential games, where the change of a player's value function upon her unilateral deviation from her strategy is equal to the change…

最优化与控制 · 数学 2025-04-02 Xin Guo , Xinyu Li , Yufei Zhang

State-of-the-art methods for solving 2-player zero-sum imperfect information games rely on linear programming or regret minimization, though not on dynamic programming (DP) or heuristic search (HS), while the latter are often at the core of…

人工智能 · 计算机科学 2022-10-27 Aurélien Delage , Olivier Buffet , Jilles S. Dibangoye , Abdallah Saffidine

We develop a martingale approach for studying continuous-time stochastic differential games of control and stopping, in a non-Markovian framework and with the control affecting only the drift term of the state-process. Under appropriate…

概率论 · 数学 2008-08-28 Ioannis Karatzas , Ingrid-Mona Zamfirescu

For a zero-sum stochastic game which does not satisfy the Isaacs condition, we provide a value function representation for an Isaacs-type equation whose Hamiltonian lies in between the lower and upper Hamiltonians, as a convex combination…

概率论 · 数学 2016-09-30 Daniel Hernández-Hernández , Mihai Sîrbu

Sample average approximation--based stochastic dynamic programming (SDP) and model predictive control (MPC) are two different methods for approaching multistage stochastic optimization. In this paper we investigate the conditions under…

最优化与控制 · 数学 2026-02-10 Dominic S. T. Keehan , Andrew B. Philpott , Edward J. Anderson

We study zero-sum stochastic games for controlled discrete time Markov chains with risk-sensitive average cost criterion with countable state space and Borel action spaces. The payoff function is nonnegative and possibly unbounded. Under a…

最优化与控制 · 数学 2022-01-12 Mrinal K. Ghosh , Subrata Golui , Chandan Pal , Somnath Pradhan

Stochastic games combine controllable and adversarial non-determinism with stochastic behavior and are a common tool in control, verification and synthesis of reactive systems facing uncertainty. Multi-objective stochastic games are natural…

计算机科学与博弈论 · 计算机科学 2021-09-20 Tobias Winkler , Maximilian Weininger

This paper deals with a stochastic recursive optimal control problem, where the diffusion coefficient depends on the control variable and the control domain is not necessarily convex. We focus on the connection between the general maximum…

最优化与控制 · 数学 2016-12-21 Tianyang Nie , Jingtao Shi , Zhen Wu

This paper shows that the optimal policy and value functions of a Markov Decision Process (MDP), either discounted or not, can be captured by a finite-horizon undiscounted Optimal Control Problem (OCP), even if based on an inexact model.…

系统与控制 · 电气工程与系统科学 2023-02-08 Arash Bahari Kordabad , Mario Zanon , Sebastien Gros

In this work, we propose novel offline and online Inverse Differential Game (IDG) methods for nonlinear Differential Games (DG), which identify the cost functions of all players from control and state trajectories constituting a feedback…

最优化与控制 · 数学 2024-11-18 Philipp Karg , Balint Varga , Sören Hohmann

The problem of synthesizing stochastic explicit model predictive control policies is known to be quickly intractable even for systems of modest complexity when using classical control-theoretic methods. To address this challenge, we present…

机器学习 · 计算机科学 2022-05-24 Ján Drgoňa , Sayak Mukherjee , Aaron Tuor , Mahantesh Halappanavar , Draguna Vrabie