English
Related papers

Related papers: Dynamic Control for Random Access in Deadline-Cons…

200 papers

We study the offline data-driven sequential decision making problem in the framework of Markov decision process (MDP). In order to enhance the generalizability and adaptivity of the learned policy, we propose to evaluate each policy by a…

Statistics Theory · Mathematics 2021-11-11 Zhengling Qi , Peng Liao

In this work, we consider the optimal portfolio selection problem under hard constraints on trading volume amounts when the dynamics of the risky asset returns are governed by a discrete-time approximation of the Markov-modulated geometric…

Portfolio Management · Quantitative Finance 2014-10-07 Vladimir Dombrovskii , Tatyana Obyedko

Neural networks have been increasingly employed in Model Predictive Controller (MPC) to control nonlinear dynamic systems. However, MPC still poses a problem that an achievable update rate is insufficient to cope with model uncertainty and…

Robotics · Computer Science 2022-07-15 Taekyung Kim , Hojin Lee , Seongil Hong , Wonsuk Lee

Managing announced task completion times is a fundamental control problem in project management. While extensive research exists on estimating task durations and task scheduling, the problem of when and how to update completion times…

Systems and Control · Electrical Eng. & Systems 2026-03-17 Duncan Eddy , Esen Yel , Emma Passmore , Niles Egan , Grayson Armour , Dylan M. Asmar , Mykel J. Kochenderfer

We propose a real-time signal control framework based on a nonlinear decision rule (NDR), which defines a nonlinear mapping between network states and signal control parameters to actual signal controls based on prevailing traffic…

Optimization and Control · Mathematics 2019-03-20 Junwoo Song , Simon Hu , Ke Han , Chaozhe Jiang

In the optimization of dynamical systems, the variables typically have constraints. Such problems can be modeled as a constrained Markov Decision Process (CMDP). This paper considers a model-free approach to the problem, where the…

Machine Learning · Computer Science 2021-02-02 Qinbo Bai , Vaneet Aggarwal , Ather Gattami

We consider opportunistic communications over multiple channels where the state ("good" or "bad") of each channel evolves as independent and identically distributed Markov processes. A user, with limited sensing and access capability,…

Networking and Internet Architecture · Computer Science 2009-03-11 Sahand H. A. Ahmad , Mingyan Liu , Tara Javidi , Qing Zhao , Bhaskar Krishnamachari

We consider Markov decision processes (MDPs) in which the transition probabilities and rewards belong to an uncertainty set parametrized by a collection of random variables. The probability distributions for these random parameters are…

Logic in Computer Science · Computer Science 2020-02-26 Murat Cubuktepe , Nils Jansen , Sebastian Junges , Joost-Pieter Katoen , Ufuk Topcu

Many real-world decision-making problems face the off-dynamics challenge: the agent learns a policy in a source domain and deploys it in a target domain with different state transitions. The distributionally robust Markov decision process…

Machine Learning · Computer Science 2025-05-26 Zhishuai Liu , Pan Xu

Control applications often feature tasks with similar, but not identical, dynamics. We introduce the Hidden Parameter Markov Decision Process (HiP-MDP), a framework that parametrizes a family of related dynamical systems with a…

Machine Learning · Computer Science 2013-08-19 Finale Doshi-Velez , George Konidaris

In this article, we consider a nonlinear process with delayed dynamics to be controlled over a communication network in the presence of disturbances and study robustness of the resulting closed-loop system with respect to network-induced…

Systems and Control · Computer Science 2016-04-18 Domagoj Tolic , Sandra Hirche

Traffic-responsive signal control is a cost-effective and easy-to-implement network management strategy with high potential in improving performance in congested networks with dynamic characteristics. Max Pressure (MP) distributed…

Systems and Control · Electrical Eng. & Systems 2023-05-03 Dimitrios Tsitsokas , Anastasios Kouvelas , Nikolas Geroliminis

We study a Q learning algorithm for continuous time stochastic control problems. The proposed algorithm uses the sampled state process by discretizing the state and control action spaces under piece-wise constant control processes. We show…

Optimization and Control · Mathematics 2023-03-10 Erhan Bayraktar , Ali Devran Kara

Constrained Markov Decision Processes (CMDPs) formalize sequential decision-making problems whose objective is to minimize a cost function while satisfying constraints on various cost functions. In this paper, we consider the setting of…

Machine Learning · Computer Science 2020-09-25 Krishna C. Kalagarla , Rahul Jain , Pierluigi Nuzzo

We consider online reinforcement learning in episodic Markov decision process (MDP) with unknown transition function and stochastic rewards drawn from some fixed but unknown distribution. The learner aims to learn the optimal policy and…

Machine Learning · Computer Science 2024-03-12 Vincent Leon , S. Rasoul Etesami

In scenarios where high penetration of renewable energy sources (RES) is connected to the grid over long distances, the output of RES exhibits significant fluctuations, making it difficult to accurately characterize. The intermittency and…

Optimization and Control · Mathematics 2025-02-27 Yuhong Wang , Xinyao Wang , Chen Shen , Jianquan Liao , Qianni Cao , Yufei Teng , Huabo Shi , Gang Chen

It is well known that for ergodic channel processes the Generalized Max-Weight Matching (GMWM) scheduling policy stabilizes the network for any supportable arrival rate vector within the network capacity region. This policy, however, often…

Information Theory · Computer Science 2016-11-18 Mahdi Lotfinezhad , Ben Liang , Elvino S. Sousa

Here, we explore the problem of error propagation mitigation in modular digital twins as a sequential decision process. Building on a companion study that used a Hidden Markov Model (HMM) to infer latent error regimes from surrogate-physics…

Machine Learning · Computer Science 2026-04-27 Annice Najafi , Shokoufeh Mirzaei

The standard practice in modeling dynamics and optimal control of a large population, ensemble, multi-agent system represented by it's continuum density, is to model individual decision making using local feedback information. In comparison…

Optimization and Control · Mathematics 2020-09-29 Kaivlaya Bakshi , Evangelos A. Theodorou

Repair mechanisms are important within resilient systems to maintain the system in an operational state after an error occurred. Usually, constraints on the repair mechanisms are imposed, e.g., concerning the time or resources required…

Systems and Control · Computer Science 2017-07-12 Christel Baier , Clemens Dubslaff , Ľuboš Korenčiak , Antonín Kučera Vojtěch Řehák
‹ Prev 1 8 9 10 Next ›