English
Related papers

Related papers: A Generalized Fundamental Matrix for Computing Fun…

200 papers

Recently discovered polyhedral structures of the value function for finite state-action discounted Markov decision processes (MDP) shed light on understanding the success of reinforcement learning. We investigate the value function polytope…

Machine Learning · Computer Science 2022-06-27 Yue Wu , Jesús A. De Loera

We introduce a simple approach for testing the reliability of homogeneous generators and the Markov property of the stochastic processes underlying empirical time series of credit ratings. We analyze open access data provided by Moody's and…

Risk Management · Quantitative Finance 2014-10-30 Pedro Lencastre , Frank Raischel , Pedro G. Lind , Tim Rogers

The intermittent on-off switching of feedback control is considered as a major mechanism of postural stabilization during human quiet standing, which can be modeled by switched-type hybrid stochastic delay differential equations with…

Systems and Control · Electrical Eng. & Systems 2023-09-01 Yasuyuki Suzuki , Keigo Togame , Akihiro Nakamura , Taishin Nomura

On a periodic basis, publicly traded companies are required to report fundamentals: financial data such as revenue, operating income, debt, among others. These data points provide some insight into the financial health of a company.…

Machine Learning · Statistics 2018-04-27 John Alberg , Zachary C. Lipton

Modeling the dynamics of non-stationary stochastic systems requires balancing the representational power of deep learning with the mathematical transparency of classical models. While classical Markov transition operators provide explicit,…

Machine Learning · Computer Science 2026-05-07 Jan Rovirosa , Jesse Schmolze

Biological systems need to react to stimuli over a broad spectrum of timescales. If and how this ability can emerge without external fine-tuning is a puzzle. We consider here this problem in discrete Markovian systems, where we can leverage…

Disordered Systems and Neural Networks · Physics 2021-08-11 Faheem Mosam , Diego Vidaurre , Eric De Giuli

A sequence of real numbers (x_n) is Benford if the significands, i.e. the fraction parts in the floating-point representation of (x_n) are distributed logarithmically. Similarly, a discrete-time irreducible and aperiodic finite-state Markov…

Probability · Mathematics 2010-03-05 Bahar Kaynar , Arno Berger , Theodore P. Hill , Ad Ridder

Markov decision processes (MDPs) are a standard model for sequential decision-making problems and are widely used across many scientific areas, including formal methods and artificial intelligence (AI). MDPs do, however, come with the…

Artificial Intelligence · Computer Science 2024-12-11 Marnix Suilen , Thom Badings , Eline M. Bovy , David Parker , Nils Jansen

We consider the estimation of the transition matrix of a hidden Markovian process by using information geometry with respect to transition matrices. In this paper, only the histogram of $k$-memory data is used for the estimation. To…

Statistics Theory · Mathematics 2024-09-10 Masahito Hayashi

We consider a robust approach to address uncertainty in model parameters in Markov Decision Processes (MDPs), which are widely used to model dynamic optimization in many applications. Most prior works consider the case where the uncertainty…

Optimization and Control · Mathematics 2021-09-02 Vineet Goyal , Julien Grand-Clément

Given noisy, partial observations of a time-homogeneous, finite-statespace Markov chain, conceptually simple, direct statistical inference is available, in theory, via its rate matrix, or infinitesimal generator, $\mathsf{Q}$, since $\exp…

Methodology · Statistics 2020-03-23 Chris Sherlock

Phase-type distribution has been an important probabilistic tool in the analysis of complex stochastic system evolution. It was introduced by Neuts \cite{Neuts1975} in 1975. The model describes the lifetime distribution of a finite-state…

Methodology · Statistics 2016-11-14 B. A. Surya

In the first part of this study (Paper I), we introduced the systematic improvement probability (SIP) as a tool to assess the level of improvement on absolute errors to be expected when switching between two computational chemistry methods.…

Chemical Physics · Physics 2020-04-30 Pascal Pernot , Andreas Savin

This paper deals with control of partially observable discrete-time stochastic systems. It introduces and studies Markov Decision Processes with Incomplete Information and with semi-uniform Feller transition probabilities. The important…

Optimization and Control · Mathematics 2022-08-30 Eugene A. Feinberg , Pavlo O. Kasyanov , Michael Z. Zgurovsky

Structural results impose sufficient conditions on the model parameters of a Markov decision process (MDP) so that the optimal policy is an increasing function of the underlying state. The classical assumptions for MDP structural results…

Systems and Control · Electrical Eng. & Systems 2023-03-07 Vikram Krishnamurthy

Robust Markov decision processes (r-MDPs) extend MDPs by explicitly modelling epistemic uncertainty about transition dynamics. Learning r-MDPs from interactions with an unknown environment enables the synthesis of robust policies with…

Machine Learning · Computer Science 2025-11-21 Yannik Schnitzer , Alessandro Abate , David Parker

Policy Iteration (PI) is a widely used family of algorithms to compute optimal policies for Markov Decision Problems (MDPs). We derive upper bounds on the running time of PI on Deterministic MDPs (DMDPs): the class of MDPs in which every…

Discrete Mathematics · Computer Science 2023-10-10 Ritesh Goenka , Eashan Gupta , Sushil Khyalia , Pratyush Agarwal , Mulinti Shaik Wajid , Shivaram Kalyanakrishnan

Risk averse decision making under uncertainty in partially observable domains is a fundamental problem in AI and essential for reliable autonomous agents. In our case, the problem is modeled using partially observable Markov decision…

Artificial Intelligence · Computer Science 2024-06-11 Yaacov Pariente , Vadim Indelman

The Markov assumption (MA) is fundamental to the empirical validity of reinforcement learning. In this paper, we propose a novel Forward-Backward Learning procedure to test MA in sequential decision making. The proposed test does not assume…

Machine Learning · Statistics 2020-02-06 Chengchun Shi , Runzhe Wan , Rui Song , Wenbin Lu , Ling Leng

Standard Markov decision process (MDP) and reinforcement learning algorithms optimize the policy with respect to the expected gain. We propose an algorithm which enables to optimize an alternative objective: the probability that the gain is…

Machine Learning · Computer Science 2023-03-06 Vincent Corlay , Jean-Christophe Sibel
‹ Prev 1 4 5 6 7 8 10 Next ›