English
Related papers

Related papers: Comparing the Efficiency of General State Space Re…

200 papers

We establish that an optimistic variant of Q-learning applied to a fixed-horizon episodic Markov decision process with an aggregated state representation incurs regret $\tilde{\mathcal{O}}(\sqrt{H^5 M K} + \epsilon HK)$, where $H$ is the…

Machine Learning · Statistics 2020-02-20 Shi Dong , Benjamin Van Roy , Zhengyuan Zhou

We study the stability properties of nonlinear multi-task regression in reproducing Hilbert spaces with operator-valued kernels. Such kernels, a.k.a. multi-task kernels, are appropriate for learning prob- lems with nonscalar outputs like…

Machine Learning · Computer Science 2013-06-18 Julien Audiffren , Hachem Kadri

The dynamics of open quantum systems and manipulation of quantum resources are both of fundamental interest in quantum physics. Here, we investigate the relation between quantum Markovianity and coherence, providing an effective way for…

Quantum Physics · Physics 2020-06-23 Kang-Da Wu , Zhibo Hou , Guo-Yong Xiang , Chuan-Feng Li , Guang-Can Guo , Daoyi Dong , Franco Nori

We show how the essential spectral radius of a bounded positive kernel, acting on bounded functions, is linked to its lower approximation by certain absolutely continuous kernels. The standart Doeblin's condition can be interpreted in this…

Probability · Mathematics 2007-05-23 Hubert Hennion

Temporal-difference learning is a popular algorithm for policy evaluation. In this paper, we study the convergence of the regularized non-parametric TD(0) algorithm, in both the independent and Markovian observation settings. In particular,…

Optimization and Control · Mathematics 2022-05-25 Eloïse Berthier , Ziad Kobeissi , Francis Bach

We present a Markov-chain analysis of blockwise-stochastic algorithms for solving partially block-separable optimization problems. Our main contributions to the extensive literature on these methods are statements about the Markov operators…

Optimization and Control · Mathematics 2023-11-01 D. Russell Luke

The preparation of the stationary distribution of irreducible, time-reversible Markov chains is a fundamental building block in many heuristic approaches to algorithmically hard problems. It has been conjectured that quantum analogs of…

Quantum Physics · Physics 2015-02-20 Vedran Dunjko , Hans J. Briegel

We provide new asymptotic theory for kernel density estimators, when these are applied to autoregressive processes exhibiting moderate deviations from a unit root. This fills a gap in the existing literature, which has to date considered…

Statistics Theory · Mathematics 2019-08-19 James A. Duffy

We provide abstract, general and highly uniform rates of asymptotic regularity for a generalized stochastic Halpern-style iteration, which incorporates a second mapping in the style of a Krasnoselskii-Mann iteration. This iteration is…

Optimization and Control · Mathematics 2025-12-19 Nicholas Pischke , Thomas Powell

We consider continuous-time Markov chain on a finite state space X. We assume X can be clustered into several subsets such that the intra-transition rates within these subsets are of order $\mathcal{O}(\frac{1}{\epsilon})$ comparing to the…

Probability · Mathematics 2016-01-28 Wei Zhang

A major factor in the success of deep neural networks is the use of sophisticated architectures rather than the classical multilayer perceptron (MLP). Residual networks (ResNets) stand out among these powerful modern architectures. Previous…

Machine Learning · Computer Science 2021-05-25 Tom Tirer , Joan Bruna , Raja Giryes

We study reinforcement learning with function approximation for large-scale Partially Observable Markov Decision Processes (POMDPs) where the state space and observation space are large or even continuous. Particularly, we consider Hilbert…

Machine Learning · Computer Science 2022-06-27 Masatoshi Uehara , Ayush Sekhari , Jason D. Lee , Nathan Kallus , Wen Sun

This article investigates nonparametric estimation of variance functions for functional data when the mean function is unknown. We obtain asymptotic results for the kernel estimator based on squared residuals. Similar to the finite…

Methodology · Statistics 2008-12-16 Heng Lian

In this paper, we investigate a nonparametric approach to provide a recursive estimator of the transition density of a non-stationary piecewise-deterministic Markov process, from only one observation of the path within a long time. In this…

Statistics Theory · Mathematics 2013-05-07 Romain Azaïs

The Importance Markov chain is a novel algorithm bridging the gap between rejection sampling and importance sampling, moving from one to the other through a tuning parameter. Based on a modified sample of an instrumental Markov chain…

Computation · Statistics 2024-02-27 Charly Andral , Randal Douc , Hugo Marival , Christian P. Robert

The basic motivation and primary goal of this paper is a qualitative evaluation of the performance of a new weighted statistic for a nonparametric test for stochastic dominance based on two samples, which was introduced in Ledwina and…

Statistics Theory · Mathematics 2018-06-07 Inglot Tadeusz , Ledwina Teresa , Ćmiel Bogdan

Do phenomenological master equations with memory kernel always describe a non-Markovian quantum dynamics characterized by reverse flow of information? Is the integration over the past states of the system an unmistakable signature of…

Quantum Physics · Physics 2015-05-18 L. Mazzola , E. -M. Laine , H. -P. Breuer , S. Maniscalco , J. Piilo

In this paper, we analyze the asymptotic behavior of the main characteristics of the mean-variance efficient frontier employing random matrix theory. Our particular interest covers the case when the dimension $p$ and the sample size $n$…

Statistical Finance · Quantitative Finance 2024-09-24 Taras Bodnar , Nikolaus Hautsch , Yarema Okhrin , Nestor Parolya

The paper studies machine learning problems where each example is described using a set of Boolean features and where hypotheses are represented by linear threshold elements. One method of increasing the expressiveness of learned hypotheses…

Machine Learning · Computer Science 2011-09-13 R. Khardon , D. Roth , R. A. Servedio

For a Markov transition kernel $P$ and a probability distribution $ \mu$ on nonnegative integers, a time-sampled Markov chain evolves according to the transition kernel $P_{\mu} = \sum_k \mu(k)P^k.$ In this note we obtain CLT conditions for…

Probability · Mathematics 2011-06-07 Krzysztof Latuszynski , Gareth O. Roberts