English
Related papers

Related papers: Multiple Mean-Payoff Optimization under Local Stab…

200 papers

In a mean field game of controls, a large population of identical players seek to minimize a cost that depends on the joint distribution of the states of the players and their controls. We first consider the classes of mean field games of…

Optimization and Control · Mathematics 2025-12-05 P. Jameson Graber , Kyle Rosengartner

Robust Markov decision processes (RMDPs) extend standard Markov decision processes (MDPs) to account for uncertainty in the transition probabilities. RMDPs have an uncertainty set that defines a set of possible transition functions, each of…

Logic in Computer Science · Computer Science 2026-04-30 Marnix Suilen , Guillermo A. Pérez

This paper is concerned with the development of rigorous approximations to various expectations associated with Markov chains and processes having non-stationary transition probabilities. Such non-stationary models arise naturally in…

Probability · Mathematics 2018-05-07 Zeyu Zheng , Harsha Honnappa , Peter W. Glynn

The parameters of a discrete stationary Markov model are transition probabilities between states. Traditionally, data consist in sequences of observed states for a given number of individuals over the whole observation period. In such a…

Computation · Statistics 2012-04-30 Alberto Pasanisi , Shuai Fu , Nicolas Bousquet

This paper considers a class of reinforcement-based learning (namely, perturbed learning automata) and provides a stochastic-stability analysis in repeatedly-played, positive-utility, finite strategic-form games. Prior work in this class of…

Computer Science and Game Theory · Computer Science 2019-01-29 Georgios C. Chasparis

In this paper, we investigate system theoretic properties of transient average constrained economic model predictive control (MPC) without terminal constraints. We show that the optimal open-loop solution passes by the optimal steady-state…

Systems and Control · Electrical Eng. & Systems 2020-10-21 Mario Rosenfelder , Johannes Köhler , Frank Allgöwer

Two-player games on graphs provide the mathematical foundation for the study of reactive systems. In the quantitative framework, an objective assigns a value to every play, and the goal of player 1 is to minimize the value of the objective.…

Logic in Computer Science · Computer Science 2014-04-30 Yaron Velner

Energy Markov Decision Processes (EMDPs) are finite-state Markov decision processes where each transition is assigned an integer counter update and a rational payoff. An EMDP configuration is a pair s(n), where s is a control state and n is…

Logic in Computer Science · Computer Science 2016-07-05 Tomáš Brázdil , Antonín Kučera , Petr Novotný

We introduce a general framework for Markov decision problems under model uncertainty in a discrete-time infinite horizon setting. By providing a dynamic programming principle we obtain a local-to-global paradigm, namely solving a local,…

Optimization and Control · Mathematics 2023-01-06 Ariel Neufeld , Julian Sester , Mario Šikić

The overall performance or expected excess risk of an iterative machine learning algorithm can be decomposed into training error and generalization error. While the former is controlled by its convergence analysis, the latter can be tightly…

Machine Learning · Statistics 2018-04-06 Yuansi Chen , Chi Jin , Bin Yu

We consider a type of optimal switching problems with non-uniform execution delays and ramping. Such problems frequently occur in the operation of economical and engineering systems. We first provide a solution to the problem by applying a…

Optimization and Control · Mathematics 2017-02-15 Magnus Perninge

The coordinated and efficient distribution of limited resources by individual decisions is a fundamental, unsolved problem. When individuals compete for road capacities, time, space, money, goods, etc., they normally make decisions based on…

Statistical Mechanics · Physics 2009-11-07 Dirk Helbing , Martin Schoenhof , Daniel Kern

It is known that state-dependent, multi-step Lyapunov bounds lead to greatly simplified verification theorems for stability for large classes of Markov chain models. This is one component of the "fluid model" approach to stability of…

Optimization and Control · Mathematics 2012-05-18 Serdar Yüksel , Sean P. Meyn

Model Predictive Control (MPC) is well understood in the deterministic setting, yet rigorous stability and performance guarantees for stochastic MPC remain limited to the consideration of terminal constraints and penalties. In contrast,…

Optimization and Control · Mathematics 2025-10-24 Jonas Schießl , Hannah Selder , Ruchuan Ou , Michael Heinrich Baumann , Timm Faulwasser , Lars Grüne

We consider the problem of optimizing the economic performance of nonlinear constrained systems subject to uncertain time-varying parameters and bounded disturbances. In particular, we propose an adaptive economic model predictive control…

Systems and Control · Electrical Eng. & Systems 2026-01-16 Maximilian Degner , Raffaele Soloperto , Melanie N. Zeilinger , John Lygeros , Johannes Köhler

This paper considers an opportunistic scheduling problem over a renewal system. A controller observes a random event at the beginning of each renewal frame and then chooses an action in response to the event, which affects the duration of…

Optimization and Control · Mathematics 2019-06-10 Xiaohan Wei , Michael J. Neely

We consider concurrent games played on graphs. At every round of the game, each player simultaneously and independently selects a move; the moves jointly determine the transition to a successor state. Two basic objectives are the safety…

Computer Science and Game Theory · Computer Science 2008-12-18 Krishnendu Chatterjee , Luca de Alfaro , Thomas A. Henzinger

Policy optimization methods with function approximation are widely used in multi-agent reinforcement learning. However, it remains elusive how to design such algorithms with statistical guarantees. Leveraging a multi-agent performance…

Machine Learning · Computer Science 2023-05-09 Yulai Zhao , Zhuoran Yang , Zhaoran Wang , Jason D. Lee

In many stochastic games stemming from financial models, the environment evolves with latent factors and there may be common noise across agents' states. Two classic examples are: (i) multi-agent trading on electronic exchanges, and (ii)…

Optimization and Control · Mathematics 2019-07-24 Dena Firoozi , Peter E. Caines , Sebastian Jaimungal

This paper focuses on a class of continuous-time controlled Markov chains with time-inconsistent and distribution-dependent cost functional (in some appropriate sense). A new definition of time-inconsistent distribution-dependent…

Optimization and Control · Mathematics 2019-09-26 Hongwei Mei , George Yin