中文
相关论文

相关论文: Capacities, Measurable Selection and Dynamic Progr…

200 篇论文

Temporal abstraction is key to scaling up learning and planning in reinforcement learning. While planning with temporally extended actions is well understood, creating such abstractions autonomously from data has remained challenging. We…

人工智能 · 计算机科学 2016-12-06 Pierre-Luc Bacon , Jean Harb , Doina Precup

We study a catching-up algorithm for a class of differential inclusions driven by maximal monotone operators with continuous perturbations. Using a decomposition of the monotone operator into the closed convex hull of its single-valued part…

最优化与控制 · 数学 2026-04-14 Tan H. Cao , Hassan Saoud

This paper proposes a robust self-triggered distributed model predictive control (DMPC) scheme for a family of Discrete-Time linear systems with local (uncoupled) and global (coupled) constraints. To handle the additive disturbance,…

系统与控制 · 电气工程与系统科学 2020-12-17 Zhengcai Li

We consider the optimization of an uncertain objective over continuous and multi-dimensional decision spaces in problems in which we are only provided with observational data. We propose a novel algorithmic framework that is tractable,…

机器学习 · 统计学 2018-10-30 Dimitris Bertsimas , Christopher McCord

Graded modalities have been proposed in recent work on programming languages as a general framework for refining type systems with intensional properties. In particular, continuous endomaps of the discrete time scale, or time warps, can be…

逻辑 · 数学 2021-08-20 Sam van Gool , Adrien Guatto , George Metcalfe , Simon Santschi

Optimization is an important module of modern machine learning applications. Tremendous efforts have been made to accelerate optimization algorithms. A common formulation is achieving a lower loss at a given time. This enables a…

机器学习 · 计算机科学 2025-05-29 Zhonglin Xie , Yiman Fong , Haoran Yuan , Zaiwen Wen

In this paper, we consider the gradual-impulse control problem of continuous-time Markov decision processes, where the system performance is measured by the expectation of the exponential utility of the total cost. We prove, under very…

最优化与控制 · 数学 2023-11-16 Xin Guo , Aiko Kurushima , Alexey Piunovskiy , Yi Zhang

In this semi-tutorial paper, we first review the information-theoretic approach to account for the computational costs incurred during the search for optimal actions in a sequential decision-making problem. The traditional (MDP) framework…

人工智能 · 计算机科学 2021-02-23 Daniel T. Larsson , Daniel Braun , Panagiotis Tsiotras

We consider both discrete and continuous control problems constrained by a fixed budget of some resource, which may be renewed upon entering a preferred subset of the state space. In the discrete case, we consider both deterministic and…

最优化与控制 · 数学 2014-09-30 Ryo Takei , Weiyan Chen , Zachary Clawson , Slav Kirov , Alexander Vladimirsky

In this paper, we propose a class of discrete-time approximation schemes for stochastic optimal control problems under the $G$-expectation framework. The proposed schemes are constructed recursively based on piecewise constant policy. We…

最优化与控制 · 数学 2021-10-05 Lianzi Jiang

We introduce a framework that represents a dynamic program as a family of operators acting on a partially ordered set. We provide an optimality theory based only on order-theoretic assumptions and show how applications across almost all…

最优化与控制 · 数学 2025-01-07 Thomas J. Sargent , John Stachurski

When modeling an application of practical relevance as an instance of a combinatorial problem X, we are often interested not merely in finding one optimal solution for that instance, but in finding a sufficiently diverse collection of good…

These lecture notes are derived from a graduate-level course in dynamic optimization, offering an introduction to techniques and models extensively used in management science, economics, operations research, engineering, and computer…

最优化与控制 · 数学 2024-10-11 Bar Light

Modern learning systems increasingly interact with data that evolve over time and depend on hidden internal state. We ask a basic question: when is such a dynamical system learnable from observations alone? This paper proposes a research…

机器学习 · 计算机科学 2025-12-23 Elad Hazan , Shai Shalev Shwartz , Nathan Srebro

We present differentiable predictive control (DPC), a method for learning constrained neural control policies for linear systems with probabilistic performance guarantees. We employ automatic differentiation to obtain direct policy…

系统与控制 · 电气工程与系统科学 2022-01-28 Jan Drgona , Aaron Tuor , Draguna Vrabie

Model Predictive Control (MPC) is a powerful method for complex system regulation, but its reliance on an accurate model poses many limitations in real-world applications. Data-driven predictive control (DDPC) aims at overcoming this…

系统与控制 · 电气工程与系统科学 2025-01-08 Alessandro Chiuso , Marco Fabris , Valentina Breschi , Simone Formentin

We prove a general existence result in stochastic optimal control in discrete time where controls take values in conditional metric spaces, and depend on the current state and the information of past decisions through the evolution of a…

最优化与控制 · 数学 2018-12-19 Asgar Jamneshan , Michael Kupper , José Miguel Zapata

This paper provides a compositional scheme based on dissipativity approaches for constructing finite abstractions of continuous-time continuous-space stochastic control systems. The proposed framework enjoys the structure of the…

系统与控制 · 电气工程与系统科学 2020-05-06 Ameneh Nejati , Majid Zamani

A common theme in all the above areas is designing a dynamical system to accomplish desired objectives, possibly in some predefined optimal way. Since control theory advances the idea of suitably modifying the behavior of a dynamical…

最优化与控制 · 数学 2024-07-03 Revati Gunjal , Syed Shadab Nayyer , Sushama Wagh , Navdeep Singh

We study the McKean-Vlasov optimal control problem with common noise in various formulations, namely the strong and weak formulation, as well as the Markovian and non-Markovian formulations, and allowing for the law of the control process…

最优化与控制 · 数学 2020-03-25 Mao Fabrice Djete , Dylan Possamaï , Xiaolu Tan