中文
相关论文

相关论文: The time-dependent expected reward and deviation m…

200 篇论文

We consider Poisson's equation for quasi-birth-and-death processes (QBDs) and we exploit the special transition structure of QBDs to obtain its solutions in two different forms. One is based on a decomposition through first passage times to…

概率论 · 数学 2013-08-13 Sarah Dendievel , Guy Latouche , Yuanyuan Liu

We are interested in the analysis of very large continuous-time Markov chains (CTMCs) with many distinct rates. Such models arise naturally in the context of reliability analysis, e.g., of computer network performability analysis, of power…

计算机科学中的逻辑 · 计算机科学 2015-07-24 Ernst Moritz Hahn , Holger Hermanns , Ralf Wimmer , Bernd Becker

This paper studies the expected value of multiplicative rewards, where rewards obtained in each step are multiplied (instead of the usual addition), in Markov chains (MCs) and Markov decision processes (MDPs). One of the key differences to…

计算机科学中的逻辑 · 计算机科学 2025-06-24 Christel Baier , Krishnendu Chatterjee , Tobias Meggendorfer , Jakob Piribauer

This paper is concerned with the development of rigorous approximations to various expectations associated with Markov chains and processes having non-stationary transition probabilities. Such non-stationary models arise naturally in…

概率论 · 数学 2018-05-07 Zeyu Zheng , Harsha Honnappa , Peter W. Glynn

In this paper, we develop some matrix Poisson's equations satisfied by the mean and variance of the mixing time in an irreducible positive-recurrent discrete-time Markov chain with infinitely-many levels, and provide a computational…

概率论 · 数学 2013-08-21 Quan-Lin Li , Jing Cao

We describe an exact approach for calculating transition probabilities and waiting times in finite-state discrete-time Markov processes. All the states and the rules for transitions between them must be known in advance. We can then…

其他凝聚态物理 · 物理学 2009-11-11 Semen A. Trygubenko , David J. Wales

We study the computational complexity of central analysis problems for One-Counter Markov Decision Processes (OC-MDPs), a class of finitely-presented, countable-state MDPs. OC-MDPs are equivalent to a controlled extension of (discrete-time)…

计算机科学与博弈论 · 计算机科学 2009-09-11 Tomáš Brázdil , Václav Brožek , Kousha Etessami , Antonín Kučera , Dominik Wojtczak

We propose a new approach for estimating the finite dimensional transition matrix of a Markov chain using a large number of independent sample paths observed at random times. The sample paths may be observed as few as two times, and the…

统计方法学 · 统计学 2025-05-20 Daphne Aurouet , Valentin Patilea

The goal of a traditional Markov decision process (MDP) is to maximize expected cumulative reward over a defined horizon (possibly infinite). In many applications, however, a decision maker may be interested in optimizing a specific…

人工智能 · 计算机科学 2025-10-16 Xiaocheng Li , Huaiyang Zhong , Margaret L. Brandeau

In the study of large scale stochastic networks with resource management, differential equations and mean-field limits are two key techniques. Recent research shows that the expected fraction vector (that is, the tailed probability vector)…

概率论 · 数学 2013-05-27 Quan-Lin Li

Consider a system evolving according to an absorbing discrete-time Markov chain with known transition matrix. The state of the system is observed at two points in time, separated by an unknown number of generations. We are interested in…

概率论 · 数学 2015-11-04 Bianca De Sanctis , A. P. Jason de Koning

In this article, we consider a continuous review (s, S) inventory system with failures of demand fulfillment (service) modeled as a Markov-modulated retrial queueing system. The inventory system features a single product that experiences…

概率论 · 数学 2023-07-18 James Cordeiro , Ying-Ju Chen , Andres Larrain-Hubach , Mark Abramson

We consider the problem of estimating the transition rate matrix of a continuous-time Markov chain from a finite-duration realisation of this process. We approach this problem in an imprecise probabilistic framework, using a set of prior…

机器学习 · 统计学 2018-07-12 Thomas Krak , Alexander Erreygers , Jasper De Bock

A discrete-time two-dimensional quasi-birth-and-death process (2d-QBD process), $\{{\boldsymbol{Y}}_n\}=\{(X_{1,n},X_{2,n},J_n)\}$, is a two-dimensional skip-free random walk $\{(X_{1,n},X_{2,n})\}$ on $\mathbb{Z}_+^2$ with a supplemental…

概率论 · 数学 2018-07-23 Toshihisa Ozawa , Masahiro Kobayashi

It is interesting and challenging to study double-ended queues with First-Come-First-Match discipline under customers' impatient behavior and non-Poisson inputs. The system stability can be guaranteed by the customers' impatient behavior,…

概率论 · 数学 2022-04-29 Heng-Li Liu , Quan-Lin Li , Yan-Xia Chang , Chi Zhang

Markov chains are the de facto finite-state model for stochastic dynamical systems, and Markov decision processes (MDPs) extend Markov chains by incorporating non-deterministic behaviors. Given an MDP and rewards on states, a classical…

计算机科学中的逻辑 · 计算机科学 2024-11-13 Krishnendu Chatterjee , Laurent Doyen

Note that the serial structure of blockchain has many essential pitfalls, thus a data network structure and its DAG-based blockchain are introduced to resolve the blockchain pitfalls. From such a network perspective, analysis of the…

性能 · 计算机科学 2022-10-24 Xing-Shuo Song , Quan-Lin Li , Yan-Xia Chang , Chi Zhang

The paper addresses the problem of computing maximal conditional expected accumulated rewards until reaching a target state (briefly called maximal conditional expectations) in finite-state Markov decision processes where the condition is…

计算机科学中的逻辑 · 计算机科学 2023-03-07 Christel Baier , Joachim Klein , Sascha Klüppelholz , Sascha Wunderlich

A Markov decision process can be parameterized by a transition kernel and a reward function. Both play essential roles in the study of reinforcement learning as evidenced by their presence in the Bellman equations. In our inquiry of various…

机器学习 · 计算机科学 2023-09-04 Falcon Z. Dai

In this paper, we present a numerical framework for constructing bounds on stationary performance measures of random walks in the positive orthant using the Markov reward approach. These bounds are established in terms of stationary…

概率论 · 数学 2018-11-22 Xinwei Bai , Jasper Goseling
‹ 上一页 1 2 3 10 下一页 ›