English
Related papers

Related papers: On Linear Threshold Policies for Continuous-Time D…

200 papers

In this paper, we address the optimal control of stochastic matching models on general graphs and single arrivals having fixed arrival rates, as introduced in \cite{MaiMoy16}. On the `N-shaped' graph, by following the dynamic programming…

Optimization and Control · Mathematics 2024-02-06 Loïc Jean , Pascal Moyal

This paper studies a long-term resource allocation problem over multiple periods where each period requires a multi-stage decision-making process. We formulate the problem as an online allocation problem in an episodic finite-horizon…

Data Structures and Algorithms · Computer Science 2023-10-20 Duksang Lee , William Overman , Dabeen Lee

This paper considers receding horizon control of finite deterministic systems, which must satisfy a high level, rich specification expressed as a linear temporal logic formula. Under the assumption that time-varying rewards are associated…

Optimization and Control · Mathematics 2012-03-14 Xuchu Ding , Mircea Lazar , Calin Belta

A multi-class single-server queueing model with finite buffers, in which scheduling and admission of customers are subject to control, is studied in the moderate deviation heavy traffic regime. A risk-sensitive cost set over a finite time…

Probability · Mathematics 2018-05-02 Rami Atar , Asaf Cohen

We propose a formulation of the stochastic cutting stock problem as a discounted infinite-horizon Markov decision process. At each decision epoch, given current inventory of items, an agent chooses in which patterns to cut objects in stock…

Optimization and Control · Mathematics 2022-06-29 Anselmo R. Pitombeira-Neto , Arthur H. Fonseca Murta

This paper introduces a novel stochastic control framework to enhance the capabilities of automated investment managers, or robo-advisors, by accurately inferring clients' investment preferences from past activities. Our approach leverages…

Optimization and Control · Mathematics 2024-06-05 Haoyang Cao , Zhengqi Wu , Renyuan Xu

We explore reinforcement learning methods for finding the optimal policy in the linear quadratic regulator (LQR) problem. In particular, we consider the convergence of policy gradient methods in the setting of known and unknown parameters.…

Machine Learning · Computer Science 2021-06-25 Ben Hambly , Renyuan Xu , Huining Yang

We study an optimal investment and consumption problem over a finite-time horizon, in which an individual invests in a risk-free asset and a risky asset, and evaluate utility using a general utility function that exhibits loss aversion with…

Optimization and Control · Mathematics 2025-07-08 Chonghu Guan , Xinfeng Gu , Wenhao Zhang , Xun Li

We consider a discrete-time bipartite matching model with random arrivals of units of supply and demand that can wait in queues located at the nodes in the network. A control policy determines which are matched at each time. The focus is on…

Discrete Mathematics · Computer Science 2016-06-28 Ana Bušić , Sean Meyn

We consider a finite-horizon linear-quadratic optimal control problem where only a limited number of control messages are allowed for sending from the controller to the actuator. To restrict the number of control actions computed and…

Systems and Control · Computer Science 2017-01-19 Burak Demirel , Euhanna Ghadimi , Daniel E. Quevedo , Mikael Johansson

We study a continuous-time, infinite-horizon dynamic bipartite matching problem. Suppliers arrive according to a Poisson process; while waiting, they may abandon the queue at a uniform rate. Customers on the other hand must be matched upon…

Data Structures and Algorithms · Computer Science 2025-06-03 Alireza AmaniHamedani , Ali Aouad , Amin Saberi

In this paper, we study a stock-rationing queue with two demand classes by means of the sensitivity-based optimization, and develop a complete algebraic solution to the optimal dynamic rationing policy. We show that the optimal dynamic…

Optimization and Control · Mathematics 2021-02-02 Quan-Lin Li , Yi-Meng Li , Jing-Yu Ma , Heng-Li Liu

Markov control algorithms that perform smooth, non-greedy updates of the policy have been shown to be very general and versatile, with policy gradient and Expectation Maximisation algorithms being particularly popular. For these algorithms,…

Systems and Control · Computer Science 2012-02-20 Thomas Furmston , David Barber

In this paper, we investigate an interesting and important stopping problem mixed with stochastic controls and a \textit{nonsmooth} utility over a finite time horizon. The paper aims to develop new methodologies, which are significantly…

Optimization and Control · Mathematics 2015-07-06 Chonghu Guan , Xun Li , Zuoquan Xu , Fahuai Yi

Continuous time financial market models are often motivated as scaling limits of discrete time models. The objective of this paper is to establish such a connection for a robust framework. More specifically, we consider discrete time models…

Probability · Mathematics 2024-10-17 David Criens

Hierarchical Reinforcement Learning (HRL) approaches have shown successful results in solving a large variety of complex, structured, long-horizon problems. Nevertheless, a full theoretical understanding of this empirical evidence is…

Machine Learning · Computer Science 2025-02-05 Gianluca Drappo , Alberto Maria Metelli , Marcello Restelli

Constrained partially observable Markov decision processes (CPOMDPs) have been used to model various real-world phenomena. However, they are notoriously difficult to solve to optimality, and there exist only a few approximation methods for…

Artificial Intelligence · Computer Science 2023-06-27 Robert K. Helmeczi , Can Kavaklioglu , Mucahit Cevik

This article treats both discrete time and continuous time stopping problems for general Markov processes on the real line with general linear costs. Using an auxiliary function of maximum representation type, conditions are given to…

Probability · Mathematics 2020-01-28 Sören Christensen , Tobias Sohr

We introduce a model of infinite horizon linear dynamic optimization with linear constraints and obtain results concerning feasibility of trajectories and optimal solutions necessarily satisfying conditions that resemble the Euler condition…

Optimization and Control · Mathematics 2025-04-02 Somdeb Lahiri

Option-critic learning is a general-purpose reinforcement learning (RL) framework that aims to address the issue of long term credit assignment by leveraging temporal abstractions. However, when dealing with extended timescales, discounting…

Machine Learning · Computer Science 2019-11-21 Akshay Dharmavaram , Matthew Riemer , Shalabh Bhatnagar