中文
相关论文

相关论文: A Certainty Equivalence Result in Team-Optimal Con…

200 篇论文

Piecewise constant control approximation provides a practical framework for designing numerical schemes of continuous-time control problems. We analyze the accuracy of such approximations for extended mean field control (MFC) problems,…

最优化与控制 · 数学 2025-09-03 Christoph Reisinger , Wolfgang Stockinger , Maria Olympia Tsianni , Yufei Zhang

This paper considers the decentralized convex optimization problem, which has a wide range of applications in large-scale machine learning, sensor networks, and control theory. We propose novel algorithms that achieve optimal computation…

机器学习 · 计算机科学 2023-10-11 Haishan Ye , Luo Luo , Ziang Zhou , Tong Zhang

Continuous-time Bayesian networks is a natural structured representation language for multicomponent stochastic processes that evolve continuously over time. Despite the compact representation, inference in such models is intractable even…

人工智能 · 计算机科学 2012-05-14 Ido Cohn , Tal El-Hay , Nir Friedman , Raz Kupferman

We study the convergence problem of mean-field control theory in the presence of state constraints and non-degenerate idiosyncratic noise. Our main result is the convergence of the value functions associated to stochastic control problems…

最优化与控制 · 数学 2023-06-02 Samuel Daudin

This paper presents a particle-based optimization method designed for addressing minimization problems with equality constraints, particularly in cases where the loss function exhibits non-differentiability or non-convexity. The proposed…

最优化与控制 · 数学 2026-03-31 José A. Carrillo , Shi Jin , Haoyu Zhang , Yuhua Zhu

This article is concerned with stochastic control problems for backward doubly stochastic differential equations of mean-field type, where the coefficient functions depend on the joint distribution of the state process and the control…

概率论 · 数学 2022-05-26 Jian Song , Meng Wang

We address the problem of distributed uncon- strained convex optimization under separability assumptions, i.e., the framework where each agent of a network is endowed with a local private multidimensional convex cost, is subject to…

最优化与控制 · 数学 2015-11-06 Damiano Varagnolo , Filippo Zanella , Angelo Cenedese , Gianluigi Pillonetto , Luca Schenato

Coordination of distributed agents is required for problems arising in many areas, including multi-robot systems, networking and e-commerce. As a formal framework for such problems, we use the decentralized partially observable Markov…

人工智能 · 计算机科学 2014-01-16 Daniel S. Bernstein , Christopher Amato , Eric A. Hansen , Shlomo Zilberstein

We derive sufficient and necessary optimality conditions in terms of a stochastic maximum principle (SMP) for controls associated with cost functionals of mean-field type, under dynamics driven by a class of Markov chains of mean-field type…

概率论 · 数学 2018-09-07 Salah Eddine Choutri , Hamidou Tembine

We study Markov population processes on large graphs, with the local state transition rates of a single vertex being linear function of its neighborhood. A simple way to approximate such processes is by a system of ODEs called the…

概率论 · 数学 2021-08-30 Dániel Keliger

This paper considers a new approach to using Markov chain Monte Carlo (MCMC) in contexts where one may adopt multilevel (ML) Monte Carlo. The underlying problem is to approximate expectations w.r.t. an underlying probability measure that is…

数值分析 · 数学 2018-06-27 Ajay Jasra , Kody Law , Yaxian Xu

We consider the problem of controlling a Markov decision process (MDP) with a large state space, so as to minimize average cost. Since it is intractable to compete with the optimal policy for large scale problems, we pursue the more modest…

最优化与控制 · 数学 2014-02-28 Yasin Abbasi-Yadkori , Peter L. Bartlett , Alan Malek

In this paper, we consider the gradual-impulse control problem of continuous-time Markov decision processes, where the system performance is measured by the expectation of the exponential utility of the total cost. We prove, under very…

最优化与控制 · 数学 2023-11-16 Xin Guo , Aiko Kurushima , Alexey Piunovskiy , Yi Zhang

We study a family of McKean-Vlasov (mean-field) type ergodic optimal control problems with linear control, and quadratic dependence on control of the cost function. For this class of problems we establish existence and uniqueness of an…

This study considers an optimal reinsurance, investment, and dividend strategy control problem for insurance companies in a regulated Markov regime-switching environment, intending to maximize long-run average reward. Unlike existing single…

最优化与控制 · 数学 2025-12-18 Lingjia Zeng , Manman Li

Suppose there are $n$ Markov chains and we need to pay a per-step \emph{price} to advance them. The "destination" states of the Markov chains contain rewards; however, we can only get rewards for a subset of them that satisfy a…

数据结构与算法 · 计算机科学 2019-02-22 Anupam Gupta , Haotian Jiang , Ziv Scully , Sahil Singla

In the paper, we study a new rate of convergence estimate for homogeneous discrete-time nonlinear Markov chains based on the Markov-Dobrushin condition. This result generalizes the convergence estimates for any positive number of transition…

概率论 · 数学 2021-10-22 Aleksandr A. Shchegolev

This paper shows that the optimal policy and value functions of a Markov Decision Process (MDP), either discounted or not, can be captured by a finite-horizon undiscounted Optimal Control Problem (OCP), even if based on an inexact model.…

系统与控制 · 电气工程与系统科学 2023-02-08 Arash Bahari Kordabad , Mario Zanon , Sebastien Gros

The emergence of the Internet-of-Things and cyber-physical systems necessitates the coordination of access to limited communication resources in an autonomous and distributed fashion. Herein, the optimal design of a wireless sensing system…

系统与控制 · 电气工程与系统科学 2020-05-26 Xu Zhang , Marcos M. Vasconcelos , Wei Cui , Urbashi Mitra

Starting from the Avellaneda-Stoikov framework, we consider a market maker who wants to optimally set bid/ask quotes over a finite time horizon, to maximize her expected utility. The intensities of the orders she receives depend not only on…

交易与市场微观结构 · 定量金融 2020-06-29 Diego Zabaljauregui , Luciano Campi