中文
相关论文

相关论文: Comparing the Efficiency of General State Space Re…

200 篇论文

We provide a theoretical foundation for non-parametric estimation of functions of random variables using kernel mean embeddings. We show that for any continuous function $f$, consistent estimators of the mean embedding of a random variable…

机器学习 · 统计学 2018-06-04 Carl-Johann Simon-Gabriel , Adam Ścibior , Ilya Tolstikhin , Bernhard Schölkopf

Particle Marginal Metropolis-Hastings (PMMH) is a general approach to Bayesian inference when the likelihood is intractable, but can be estimated unbiasedly. Our article develops an efficient PMMH method that scales up better to higher…

统计计算 · 统计学 2023-05-10 David Gunawan , Pratiti Chatterjee , Robert Kohn

Let $P$ be a Markov kernel on a measurable space $\X$ and let $V:\X\r[1,+\infty)$. This paper provides explicit connections between the $V$-geometric ergodicity of $P$ and that of finite-rank nonnegative sub-Markov kernels $\Pc_k$…

概率论 · 数学 2014-01-24 Loïc Hervé , James Ledoux

We develop a stochastic approximation framework for learning nonlinear operators between infinite-dimensional spaces utilizing general Mercer operator-valued kernels. Our framework encompasses two key classes: (i) compact kernels, which…

机器学习 · 统计学 2026-01-13 Jia-Qi Yang , Lei Shi

Reinforcement learning algorithms often require finiteness of state and action spaces in Markov decision processes (MDPs) (also called controlled Markov chains) and various efforts have been made in the literature towards the applicability…

机器学习 · 计算机科学 2023-09-08 Ali Devran Kara , Naci Saldi , Serdar Yüksel

An efficient Krylov subspace algorithm for computing actions of the $\varphi$ matrix function for large matrices is proposed. This matrix function is widely used in exponential time integration, Markov chains and network analysis and many…

数值分析 · 数学 2020-10-20 Mike A. Botchev , Leonid A. Knizhnerman , Eugene E. Tyrtyshnikov

Deploying quantum machine learning on NISQ devices requires architectures where training overhead does not negate computational advantages. We systematically compare two quantum approaches for chaotic time-series prediction on the Lorenz…

量子物理 · 物理学 2026-04-28 Tushar Pandey

The classical Metropolis-Hastings (MH) algorithm can be extended to generate non-reversible Markov chains. This is achieved by means of a modification of the acceptance probability, using the notion of vorticity matrix. The resulting Markov…

概率论 · 数学 2020-09-29 Joris Bierkens

We introduce QPU micro-kernels: shallow quantum circuits that perform a stencil node update and return a Monte Carlo estimate from repeated measurements. We show how to use them to solve Partial Differential Equations (PDEs) explicitly…

新兴技术 · 计算机科学 2025-11-18 Stefano Markidis , Luca Pennati , Marco Pasquale , Gilbert Netzer , Ivy Peng

This work contributes to the programme of studying effective versions of "almost everywhere" theorems in analysis and ergodic theory via algorithmic randomness. We determine the level of randomness needed for a point in a Cantor space $…

逻辑 · 数学 2016-05-10 Rodney G. Downey , Satyadev Nandakumar , Andre Nies

This paper investigates the optimization problem of an infinite stage discrete time Markov decision process (MDP) with a long-run average metric considering both mean and variance of rewards together. Such performance metric is important…

最优化与控制 · 数学 2020-08-11 Li Xia

Reversible Markov chains play a central role in stochastic modelling and in algorithms such as Markov chain Monte Carlo (MCMC). Motivated by the fundamental importance of reversibility in classical settings, this paper develops a…

概率论 · 数学 2025-10-28 Damjan Škulj

Via operator theoretic methods, we formalize the concentration phenomenon for a given observable `$r$' of a discrete time Markov chain with `$\mu_{\pi}$' as invariant ergodic measure, possibly having support on an unbounded state space. The…

机器学习 · 计算机科学 2023-06-01 Muhammad Abdullah Naeem , Miroslav Pajic

A Peskun ordering between two samplers, implying a dominance of one over the other, is known among the Markov chain Monte Carlo community for being a remarkably strong result. It is however also known for being a result that is notably…

统计计算 · 统计学 2024-05-20 Philippe Gagnon , Florian Maire

We consider the problem of estimating the asymptotic variance of a function defined on a Markov chain, an important step for statistical inference of the stationary mean. We design a novel recursive estimator that requires $O(1)$…

统计理论 · 数学 2024-09-24 Shubhada Agrawal , Prashanth L. A. , Siva Theja Maguluri

We prove a complete class theorem that characterizes \emph{all} stationary time reversible Markov processes whose finite dimensional marginal distributions (of all orders) are infinitely divisible. Aside from two degenerate cases (iid and…

概率论 · 数学 2021-06-01 Robert L Wolpert , Lawrence D. Brown

We study the generalized eigenvalue problem on the whole space for a class of integro-differential elliptic operators. The nonlocal operator is over a finite measure, but this has no particular structure. Some of our results even hold for…

偏微分方程分析 · 数学 2022-11-24 Ari Arapostathis , Anup Biswas , Prasun Roychowdhury

We study model-free reinforcement learning (RL) algorithms in episodic non-stationary constrained Markov Decision Processes (CMDPs), in which an agent aims to maximize the expected cumulative reward subject to a cumulative constraint on the…

机器学习 · 计算机科学 2023-03-13 Honghao Wei , Arnob Ghosh , Ness Shroff , Lei Ying , Xingyu Zhou

Quantum processors may enhance machine learning by mapping high-dimensional data onto quantum systems for processing. Conventional feature maps, for encoding data onto a quantum circuit are currently impractical, as the number of entangling…

量子物理 · 物理学 2026-03-27 Utkarsh Singh , Jean-Frédéric Laprade , Aaron Z. Goldberg , Khabat Heshami

We consider online learning for minimizing regret in unknown, episodic Markov decision processes (MDPs) with continuous states and actions. We develop variants of the UCRL and posterior sampling algorithms that employ nonparametric Gaussian…

机器学习 · 计算机科学 2019-01-04 Sayak Ray Chowdhury , Aditya Gopalan