English
Related papers

Related papers: On Linear Threshold Policies for Continuous-Time D…

200 papers

Model Predictive Control (MPC) is often tuned by trial and error. When a baseline linear controller exists that is already well tuned in the absence of constraints and MPC is introduced to enforce them, one would like to avoid altering the…

Systems and Control · Electrical Eng. & Systems 2021-11-01 Mario Zanon , Alberto Bemporad

Motivated by applications in online marketplaces such as ride-hailing, we study how strategic servers impact the system performance. We consider a discrete-time process in which, heterogeneous types of customers and servers arrive. Each…

Optimization and Control · Mathematics 2021-06-25 Sushil Mahavir Varma , Francisco Castro , Siva Theja Maguluri

We consider a class of infinite-time horizon optimal stopping problems for spectrally negative Levy processes. Focusing on strategies of threshold type, we write explicit expressions for the corresponding expected payoff via the scale…

Optimization and Control · Mathematics 2013-05-03 Masahiko Egami , Kazutoshi Yamazaki

This paper employs a policy iteration reinforcement learning (RL) method to study continuous-time linear-quadratic mean-field control problems in infinite horizon. The drift and diffusion terms in the dynamics involve the states, the…

Optimization and Control · Mathematics 2024-11-05 Na Li , Xun Li , Zuo Quan Xu

We propose a distributed data-based predictive control scheme to stabilize a network system described by linear dynamics. Agents cooperate to predict the future system evolution without knowledge of the dynamics, relying instead on learning…

Optimization and Control · Mathematics 2020-12-02 Ahmed Allibhoy , Jorge Cortés

This paper presents a novel framework which combines a non-iterative solution of Real-Time Nonlinear Receding Horizon Control (NRHC) methodology to achieve consensus within complex network topologies with existing time-delays and in…

Optimization and Control · Mathematics 2019-07-17 Fei Sun , Kamran Turkoglu

We introduce a model of infinite horizon linear dynamic optimization and obtain results concerning existence of solution and satisfaction of the competitive condition and transversality condition being unconditionally sufficient for…

Optimization and Control · Mathematics 2025-06-23 Somdeb Lahiri

This paper presents a data-driven receding horizon control framework for discrete-time linear systems that guarantees robust performance in the presence of bounded disturbances. Unlike the majority of existing data-driven predictive control…

Optimization and Control · Mathematics 2025-10-08 Jian Zheng , Sahand Kiani , Mario Sznaier , Constantino Lagoa

In this paper we address the problem of designing receding horizon control algorithms for linear discrete-time systems with parametric uncertainty. We do not consider presence of stochastic forcing or process noise in the system. It is…

Optimization and Control · Mathematics 2014-02-20 Raktim Bhattacharya , James Fisher

This paper considers the problem of finding a solution to the finite horizon constrained Markov decision processes (CMDP) where the objective as well as constraints are sum of additive and multiplicative utilities. Towards solving this, we…

Optimization and Control · Mathematics 2023-03-16 Uday Kumar M , Sanjay P Bhat , Veeraruna Kavitha , Nandyala Hemachandra

This paper studies convergence properties of optimal values and actions for discounted and average-cost Markov Decision Processes (MDPs) with weakly continuous transition probabilities and applies these properties to the stochastic…

Optimization and Control · Mathematics 2017-03-21 Eugene A. Feinberg , Mark E. Lewis

We consider the problem of service placement at the network edge, in which a decision maker has to choose between $N$ services to host at the edge to satisfy the demands of customers. Our goal is to design adaptive algorithms to minimize…

Networking and Internet Architecture · Computer Science 2021-01-15 Guojun Xiong , Rahul Singh , Jian Li

We explore fixed-horizon temporal difference (TD) methods, reinforcement learning algorithms for a new kind of value function that predicts the sum of rewards over a $\textit{fixed}$ number of future time steps. To learn the value function…

Machine Learning · Computer Science 2020-02-12 Kristopher De Asis , Alan Chan , Silviu Pitis , Richard S. Sutton , Daniel Graves

In this paper we propose an on-line policy iteration (PI) algorithm for finite-state infinite horizon discounted dynamic programming, whereby the policy improvement operation is done on-line, only for the states that are encountered during…

Optimization and Control · Mathematics 2021-06-03 Dimitri Bertsekas

For an investor with constant absolute risk aversion and a long horizon, who trades in a market with constant investment opportunities and small proportional transaction costs, we obtain explicitly the optimal investment policy, its implied…

Portfolio Management · Quantitative Finance 2012-08-01 Paolo Guasoni , Johannes Muhle-Karbe

This paper, based on the compactness-continuity and finite value conditions, establishes the sufficiency of the class of stationary policies out of the general class of history-dependent ones for a constrained continuous-time Markov…

Optimization and Control · Mathematics 2014-10-31 Yi Zhang

The convex analytic method has proved to be a very versatile method for the study of infinite horizon average cost optimal stochastic control problems. In this paper, we revisit the convex analytic method and make three primary…

Optimization and Control · Mathematics 2022-08-04 Ari Arapostathis , Serdar Yüksel

Regime-switching models, in particular Hidden Markov Models (HMMs) where the switching is driven by an unobservable Markov chain, are widely-used in financial applications, due to their tractability and good econometric properties. In this…

Statistical Finance · Quantitative Finance 2016-02-18 Vikram Krishnamurthy , Elisabeth Leoff , Jörn Sass

Markov Decision Processes (MDPs) have been used to formulate many decision-making problems in science and engineering. The objective is to synthesize the best decision (action selection) policies to maximize expected rewards (minimize…

Optimization and Control · Mathematics 2015-07-08 Mahmoud El Chamie , Behcet Acikmese

Parallel server systems in transportation, manufacturing, and computing heavily rely on dynamic routing using connected cyber components for computation and communication. Yet, these components remain vulnerable to random malfunctions and…

Systems and Control · Electrical Eng. & Systems 2023-08-22 Qian Xie , Jiayi Wang , Li Jin