English
Related papers

Related papers: Computing Quantiles in Markov Reward Models

200 papers

Specifying informative and dense reward functions remains a pivotal challenge in Reinforcement Learning, as it directly affects the efficiency of agent training. In this work, we harness the expressive power of quantitative Linear Temporal…

Machine Learning · Computer Science 2025-12-30 Omar Adalat , Francesco Belardinelli

Networked applications have software components that reside on different computers. Email, for example, has database, processing, and user interface components that can be distributed across a network and shared by users in different…

Methodology · Statistics 2007-08-03 John M. Chambers , David A. James , Diane Lambert , Scott Vander Wiel

We study observation-based strategies for partially-observable Markov decision processes (POMDPs) with omega-regular objectives. An observation-based strategy relies on partial information about the history of a play, namely, on the past…

Logic in Computer Science · Computer Science 2015-05-14 Krishnendu Chatterjee , Laurent Doyen , Thomas A. Henzinger

In many spatial and spatial-temporal models, and more generally in models with complex dependencies, it may be too difficult to carry out full maximum likelihood (ML) analysis. Remedies include the use of pseudo-likelihood (PL) and…

Methodology · Statistics 2026-04-24 Nils Lid Hjort , Cristiano Varin

The development of large databases of material properties, together with the availability of powerful computers, has allowed machine learning (ML) modeling to become a widely used tool for predicting material performances. While confidence…

Materials Science · Physics 2023-10-23 Francesca Tavazza , Kamal Choudhary , Brian DeCost

In using multiple regression methods for prediction, one often considers the linear combination of explanatory variables as an index. Seeking a single such index when here are multiple responses is rather more complicated. One classical…

Methodology · Statistics 2020-11-19 Stephen Portnoy , Joseph Haimberg

Quantization of a probability measure means representing it with a finite set of Dirac masses that approximates the input distribution well enough (in some metric space of probability measures). Various methods exists to do so, but the…

Machine Learning · Statistics 2024-02-12 Gabriel Turinici

We discuss the efficient computation of performance, reliability, and availability measures for Markov chains; these metrics, and the ones obtained by combining them, are often called performability measures. We show that this computational…

Numerical Analysis · Mathematics 2019-10-11 Giulio Masetti , Leonardo Robol

Motivated by applications in computing and telecommunication systems, we investigate the problem of estimating p-quantile of steady-state sojourn times in a single-server multi-class queueing system with non-preemptive priorities for p…

Probability · Mathematics 2022-07-11 Jin Guang , Guiyu Hong , Xinyun Chen , Xi Peng , Li Chen , Bo Bai , Gong Zhang

While model checking PCTL for Markov chains is decidable in polynomial-time, the decidability of PCTL satisfiability, as well as its finite model property, are long standing open problems. While general satisfiability is an intriguing…

Logic in Computer Science · Computer Science 2015-03-20 Nathalie Bertrand , John Fearnley , Sven Schewe

The study of Markov models is central to control theory and machine learning. A quantum analogue of partially observable Markov decision process was studied in (Barry, Barry, and Aaronson, Phys. Rev. A, 90, 2014). It was proved that…

Quantum Physics · Physics 2019-11-06 Christino Tamon , Weichen Xie

We investigate the statistical complexity of estimating the parameters of a discrete-state Markov chain kernel from a single long sequence of state observations. In the finite case, we characterize (modulo logarithmic factors) the minimax…

Machine Learning · Statistics 2020-08-14 Geoffrey Wolfer , Aryeh Kontorovich

We present an algorithm that can efficiently compute a broad class of inferences for discrete-time imprecise Markov chains, a generalised type of Markov chains that allows one to take into account partially specified probabilities and other…

Probability · Mathematics 2019-07-02 Natan T'Joens , Thomas Krak , Jasper De Bock , Gert de Cooman

Quantile regression permits describing how quantiles of a scalar response variable depend on a set of predictors. Because a unique definition of multivariate quantiles is lacking, extending quantile regression to multivariate responses is…

Methodology · Statistics 2021-04-22 Silvia Columbu , Paolo Frumento , Matteo Bottai

Among the many ways of quantifying uncertainty in a regression setting, specifying the full quantile function is attractive, as quantiles are amenable to interpretation and evaluation. A model that predicts the true conditional quantiles…

Machine Learning · Computer Science 2021-12-10 Youngseog Chung , Willie Neiswanger , Ian Char , Jeff Schneider

We present a novel algorithm to solve a non-linear system of equations, whose solution can be interpreted as a tight lower bound on the vector of expected hitting times of a Markov chain whose transition probabilities are only partially…

Probability · Mathematics 2022-03-30 Thomas Krak

Probabilistic Computation Tree Logic (PCTL) is frequently used to formally specify control objectives such as probabilistic reachability and safety. In this work, we focus on model checking PCTL specifications statistically on Markov…

Machine Learning · Computer Science 2020-04-23 Yu Wang , Nima Roohi , Matthew West , Mahesh Viswanathan , Geir E. Dullerud

Quantile regression is a powerful tool capable of offering a richer view of the data as compared to least-squares regression. Quantile regression is typically performed individually on a few quantiles or a grid of quantiles without…

Methodology · Statistics 2026-03-26 Ta-Hsin Li , Nimrod Megiddo

We consider a class of restless bandit problems that finds a broad application area in reinforcement learning and stochastic optimization. We consider $N$ independent discrete-time Markov processes, each of which had two possible states: 1…

Machine Learning · Computer Science 2024-05-14 Keqin Liu , Richard Weber , Chengzhong Zhang

Quantum trajectories are Markov chains modeling quantum systems subjected to repeated indirect measurements. Their stationary regime depends on what observables are measured on the probes used to indirectly measure the system. In this…

Mathematical Physics · Physics 2026-03-31 Tristan Benoist , Sascha Lill , Cornelia Vogel