中文
相关论文

相关论文: Reduction of Markov Chains using a Value-of-Inform…

200 篇论文

This paper presents two new approaches to decomposing and solving large Markov decision problems (MDPs), a partial decoupling method and a complete decoupling method. In these approaches, a large, stochastic decision problem is divided into…

人工智能 · 计算机科学 2013-02-01 Ron Parr

We consider the problem of controlling a Markov decision process (MDP) with a large state space, so as to minimize average cost. Since it is intractable to compete with the optimal policy for large scale problems, we pursue the more modest…

最优化与控制 · 数学 2014-02-28 Yasin Abbasi-Yadkori , Peter L. Bartlett , Alan Malek

Mostof the existing literature on supervised machine learning problems focuses on the case when the training data set is drawn from an i.i.d. sample. However, many practical problems are characterized by temporal dependence and strong…

统计理论 · 数学 2023-01-23 Nikola Sandrić , Stjepan Šebek

We propose a bottom-up approach, based on Reinforcement Learning, to the design of a chain achieving efficient excitation-transfer performances. We assume distance-dependent interactions among particles arranged in a chain under…

量子物理 · 物理学 2024-02-27 S. Sgroi , G. Zicari , A. Imparato , M. Paternostro

Parametric Markov chains (pMC) are used to model probabilistic systems with unknown or partially known probabilities. Although (universal) pMC verification for reachability properties is known to be coETR-complete, there have been efforts…

计算机科学中的逻辑 · 计算机科学 2025-04-29 Kasper Engelen , Guillermo A. Pérez , Shrisha Rao

This paper presents a numerical method to calculate the value function for a general discounted impulse control problem for piecewise deterministic Markov processes. Our approach is based on a quantization technique for the underlying…

概率论 · 数学 2011-08-31 Benoîte de Saporta , François Dufour

Parametric Markov chains occur quite naturally in various applications: they can be used for a conservative analysis of probabilistic systems (no matter how the parameter is chosen, the system works to specification); they can be used to…

计算机科学中的逻辑 · 计算机科学 2018-11-05 Paul Gainer , Ernst Moritz Hahn , Sven Schewe

Verification of infinite-state Markov chains is still a challenge despite several fruitful numerical or statistical approaches. For decisive Markov chains, there is a simple numerical algorithm that frames the reachability probability as…

计算机科学中的逻辑 · 计算机科学 2024-09-30 Benoît Barbot , Patricia Bouyer , Serge Haddad

We propose a Markov chain model for credit rating changes. We do not use any distributional assumptions on the asset values of the rated companies but directly model the rating transitions process. The parameters of the model are estimated…

风险管理 · 定量金融 2014-01-21 David Wozabal , Ronald Hochreiter

The problem of estimating an unknown discrete distribution from its samples is a fundamental tenet of statistical learning. Over the past decade, it attracted significant research effort and has been solved for a variety of divergence…

机器学习 · 计算机科学 2018-10-30 Yi Hao , Alon Orlitsky , Venkatadheeraj Pichapati

We propose two different approaches for introducing the information temperature of the binary N-th order Markov chains. The first approach is based on comparing the Markov sequences with the equilibrium Ising chains at given temperatures.…

数据分析、统计与概率 · 物理学 2022-10-05 O. V. Usatenko , S. S. Melnyk , G. M. Pritula , V. A. Yampol'skii

We consider the problem of learning low-dimensional representations for large-scale Markov chains. We formulate the task of representation learning as that of mapping the state space of the model to a low-dimensional state space, called the…

机器学习 · 计算机科学 2020-04-09 Mahsa Ghasemi , Abolfazl Hashemi , Haris Vikalo , Ufuk Topcu

Poyiadjis et al. (2011) show how particle methods can be used to estimate both the score and the observed information matrix for state space models. These methods either suffer from a computational cost that is quadratic in the number of…

统计计算 · 统计学 2015-09-07 Christopher Nemeth , Paul Fearnhead , Lyudmila Mihaylova

Markov chain analysis is a key technique in formal verification. A practical obstacle is that all probabilities in Markov models need to be known. However, system quantities such as failure rates or packet loss ratios, etc. are often not --…

计算机科学中的逻辑 · 计算机科学 2023-11-08 Sebastian Junges , Erika Ábrahám , Christian Hensel , Nils Jansen , Joost-Pieter Katoen , Tim Quatmann , Matthias Volk

An optimal feedback controller for a given Markov decision process (MDP) can in principle be synthesized by value or policy iteration. However, if the system dynamics and the reward function are unknown, a learning agent must discover an…

机器学习 · 计算机科学 2019-07-19 Boris Belousov , Jan Peters

This paper establishes a Markov chain model as a unified framework for understanding information consumption processes in complex networks, with clear implications to the Internet and big-data technologies. In particular, the proposed model…

社会与信息网络 · 计算机科学 2016-02-03 David Shui Wing Hui , Yi-Chao Chen , Gong Zhang , Weijie Wu , Guanrong Chen , John C. S. Lui , Yingtao Li

We introduce a Markov Chain Monte Carlo (MCMC) method that is designed to sample from target distributions with irregular geometry using an adaptive scheme. In cases where targets exhibit non-Gaussian behaviour, we propose that adaption…

统计计算 · 统计学 2023-10-06 Ameer Dharamshi , Vivian Ngo , Jeffrey S. Rosenthal

We introduce the Conditional Mutual Information (CMI) for the estimation of the Markov chain order. For a Markov chain of $K$ symbols, we define CMI of order $m$, $I_c(m)$, as the mutual information of two variables in the chain being $m$…

数据分析、统计与概率 · 物理学 2013-01-03 Maria Papapetrou , Dimitris Kugiumtzis

In this paper, we study a mean-variance optimization problem in an infinite horizon discrete time discounted Markov decision process (MDP). The objective is to minimize the variance of system rewards with the constraint of mean performance.…

最优化与控制 · 数学 2017-08-24 Li Xia

We propose a numerical technique for parameter inference in Markov models of biological processes. Based on time-series data of a process we estimate the kinetic rate constants by maximizing the likelihood of the data. The computation of…

定量方法 · 定量生物学 2011-02-15 Aleksandr Andreychenko , Linar Mikeev , David Spieler , Verena Wolf