English
Related papers

Related papers: Linear Algebraic Truncation Algorithm with A Poste…

200 papers

At the working heart of policy iteration algorithms commonly used and studied in the discounted setting of reinforcement learning, the policy evaluation step estimates the value of states with samples from a Markov reward process induced by…

Machine Learning · Computer Science 2021-03-04 Falcon Z. Dai , Matthew R. Walter

We study the error introduced by entropy regularization in infinite-horizon discrete discounted Markov decision processes. We show that this error decreases exponentially in the inverse regularization strength, both in a weighted…

Optimization and Control · Mathematics 2025-12-16 Johannes Müller , Semih Cayci

The theory of imprecise Markov chains has achieved significant progress in recent years. Its applicability, however, is still very much limited, due in large part to the lack of efficient computational methods for calculating…

Optimization and Control · Mathematics 2022-03-30 Damjan Škulj

A delay Lyapunov matrix corresponding to an exponentially stable system of linear time-invariant delay differential equations can be characterized as the solution of a boundary value problem involving a matrix valued delay differential…

Numerical Analysis · Mathematics 2018-08-28 Wim Michiels , Bin Zhou

An important class of physical systems that are of interest in practice are input-output open quantum systems that can be described by quantum stochastic differential equations and defined on an infinite-dimensional underlying Hilbert…

Quantum Physics · Physics 2016-09-26 O. Techakesari , H. I. Nurdin

We analyze structure-preserving model order reduction methods for Ornstein-Uhlenbeck processes and linear S(P)DEs with multiplicative noise based on balanced truncation. For the first time, we include in this study the analysis of non-zero…

Optimization and Control · Mathematics 2022-03-18 Simon Becker , Carsten Hartmann , Martin Redmann , Lorenz Richter

Model order reduction is a technique that is used to construct low-order approximations of large-scale dynamical systems. In this paper, we investigate a balancing based model order reduction method for dynamical systems with a linear…

Optimization and Control · Mathematics 2019-09-11 Peter Benner , Pawan Goyal , Igor Pontes Duff

We give computable bounds on the rate of convergence of the transition probabilities to the stationary distribution for a certain class of geometrically ergodic Markov chains. Our results are different from earlier estimates of Meyn and…

Probability · Mathematics 2007-05-23 Peter H. Baxendale

In this paper we obtain several informative error bounds on function approximation for the policy evaluation algorithm proposed by Basu et al. when the aim is to find the risk-sensitive cost represented using exponential utility. The main…

Machine Learning · Computer Science 2019-10-23 Prasenjit Karmakar , Shalabh Bhatnagar

Inference in hidden Markov model has been challenging in terms of scalability due to dependencies in the observation data. In this paper, we utilize the inherent memory decay in hidden Markov models, such that the forward and backward…

Machine Learning · Statistics 2025-01-14 Felix X. -F. Ye , Yi-an Ma , Hong Qian

We derive an approximation error bound that holds simultaneously for a function and all its derivatives up to any prescribed order. The bounds apply to elementary functions, including multivariate polynomials, the exponential function, and…

Machine Learning · Computer Science 2025-12-29 Konstantin Yakovlev , Nikita Puchkin

The problems of optimal recovery of unbounded operators are studied. Optimality means the highest possible accuracy and the minimal amount of discrete information involved. It is established that the truncation method, when certain…

Numerical Analysis · Mathematics 2025-05-13 Oleg Davydov , Sergei Solodky

We consider the Bayesian approach to the linear Gaussian inference problem of inferring the initial condition of a linear dynamical system from noisy output measurements taken after the initial time. In practical applications, the large…

Systems and Control · Electrical Eng. & Systems 2021-11-29 Elizabeth Qian , Jemima M. Tabeart , Christopher Beattie , Serkan Gugercin , Jiahua Jiang , Peter R. Kramer , Akil Narayan

We present a method to find an optimal policy with respect to a reward function for a discounted Markov decision process under general linear temporal logic (LTL) specifications. Previous work has either focused on maximizing a cumulative…

Systems and Control · Electrical Eng. & Systems 2021-03-24 Krishna C. Kalagarla , Rahul Jain , Pierluigi Nuzzo

We study continuous-time Markov chains on the non-negative integers under mild regularity conditions (in particular, the set of jump vectors is finite and both forward and backward jumps are possible). Based on the so-called flux balance…

Probability · Mathematics 2024-11-26 Mads Chr Hansen , Carsten Wiuf , Chuang Xu

Iterative gradient-based optimization algorithms are widely used to solve difficult or large-scale optimization problems. There are many algorithms to choose from, such as gradient descent and its accelerated variants such as Polyak's Heavy…

Optimization and Control · Mathematics 2023-09-21 Bryan Van Scoy , Laurent Lessard

Markov chains are the de facto finite-state model for stochastic dynamical systems, and Markov decision processes (MDPs) extend Markov chains by incorporating non-deterministic behaviors. Given an MDP and rewards on states, a classical…

Logic in Computer Science · Computer Science 2024-11-13 Krishnendu Chatterjee , Laurent Doyen

This paper presents a novel model order reduction framework tailored for fully nonlinear stochastic dynamics without lifting them to quadratic systems and without using linearization techniques. By directly leveraging structural properties…

Probability · Mathematics 2025-08-05 Martin Redmann

We develop a pivot-shifted Carleman linearization framework for quantum algorithms solving quadratic nonlinear ordinary differential equations. By shifting the dynamics by a pivot state prior to Carleman lifting, and combining this with a…

Quantum Physics · Physics 2026-05-20 Ke Wang , Zikang Jia , Shravan Veerapaneni , Zhiyan Ding

Estimation of the degree of stability and the bounds of solutions to non-autonomous nonlinear systems present major concerns in numerous applied problems. Yet, current techniques are frequently yield overconservative conditions which are…

Dynamical Systems · Mathematics 2020-12-29 Mark A. Pinsky
‹ Prev 1 3 4 5 6 7 10 Next ›