中文
相关论文

相关论文: Computing the minimal rebinding effect for non-rev…

200 篇论文

Recurrent neural networks (RNN) are simple dynamical systems whose computational power has been attributed to their short-term memory. Short-term memory of RNNs has been previously studied analytically only for the case of orthogonal…

神经与进化计算 · 计算机科学 2016-04-26 Alireza Goudarzi , Sarah Marzen , Peter Banda , Guy Feldman , Christof Teuscher , Darko Stefanovic

We introduce the notion of order of magnitude reversibility (OM-reversibility) in Markov chains that are parametrized by a positive parameter $\ep$. OM-reversibility is a weaker condition than reversibility, and requires only the knowledge…

概率论 · 数学 2011-10-26 Badal Joshi

Beyond the conventional quantum regression theorem, a general formula for non-Markovian correlation functions of arbitrary system operators both in the time- and frequency-domain is given. We approach the problem by transforming the…

量子物理 · 物理学 2016-09-21 Jinshuang Jin , Christian Karlewski , Michael Marthaler

Birth and death Markov processes can model stochastic physical systems from percolation to disease spread and, in particular, wildfires. We introduce and analyze a birth-death-suppression Markov process as a model of controlled culling of…

适应与自组织系统 · 物理学 2023-10-11 George Hulsey , David L. Alderson , Jean Carlson

We define the spectral gap of a Markov chain on a finite state space as the second-smallest singular value of the generator of the chain, generalizing the usual definition of spectral gap for reversible chains. We then define the relaxation…

概率论 · 数学 2025-01-07 Sourav Chatterjee

This work studies discrete-time discounted Markov decision processes with continuous state and action spaces and addresses the inverse problem of inferring a cost function from observed optimal behavior. We first consider the case in which…

最优化与控制 · 数学 2024-05-27 Angeliki Kamoutsi , Peter Schmitt-Förster , Tobias Sutter , Volkan Cevher , John Lygeros

We develop a method for computing policies in Markov decision processes with risk-sensitive measures subject to temporal logic constraints. Specifically, we use a particular risk-sensitive measure from cumulative prospect theory, which has…

人工智能 · 计算机科学 2020-04-21 Murat Cubuktepe , Ufuk Topcu

This paper studies the statistical theory of batch data reinforcement learning with function approximation. Consider the off-policy evaluation problem, which is to estimate the cumulative value of a new target policy from logged history…

机器学习 · 计算机科学 2020-02-25 Yaqi Duan , Mengdi Wang

Meta reinforcement learning sets a distribution over a set of tasks on which the agent can train at will, then is asked to learn an optimal policy for any test task efficiently. In this paper, we consider a finite set of tasks modeled…

机器学习 · 计算机科学 2024-06-05 Mirco Mutti , Aviv Tamar

Real-world sequential decision making problems commonly involve partial observability, which requires the agent to maintain a memory of history in order to infer the latent states, plan and make good decisions. Coping with partial…

机器学习 · 计算机科学 2022-02-09 Yonathan Efroni , Chi Jin , Akshay Krishnamurthy , Sobhan Miryoosefi

We study the problem of deinterleaving a set of finite-memory (Markov) processes over disjoint finite alphabets, which have been randomly interleaved by a finite-memory switch. The deinterleaver has access to a sample of the resulting…

信息论 · 计算机科学 2011-08-29 Gadiel Seroussi , Wojciech Szpankowski , Marcelo J. Weinberger

Understanding the stability and long-time behavior of generative models is a fundamental problem in modern machine learning. This paper provides quantitative bounds on the sampling error of score-based generative models by leveraging…

We define the concept of an `open' Markov process, a continuous-time Markov chain equipped with specified boundary states through which probability can flow in and out of the system. External couplings which fix the probabilities of…

数学物理 · 物理学 2017-10-03 Blake S. Pollard

In this paper we study limit behavior for a Markov-modulated (MM) binomial counting process, also called a binomial counting process under regime switching. Such a process naturally appears in the context of credit risk when multiple…

概率论 · 数学 2020-03-25 Peter Spreij , Jaap Storm

Irreversibility is commonly quantified by entropy production. An external observer can estimate it through measuring an observable that is antisymmetric under time-reversal like a current. We introduce a general framework that, inter alia,…

统计力学 · 物理学 2023-07-05 Jann van der Meer , Julius Degünther , Udo Seifert

Recently two approximate Newton methods were proposed for the optimisation of Markov Decision Processes. While these methods were shown to have desirable properties, such as a guarantee that the preconditioner is negative-semidefinite when…

最优化与控制 · 数学 2015-08-05 Thomas Furmston , Guy Lever

We study quasi-stationary distributions and quasi-limiting behavior of Markov chains in general reducible state spaces with absorption. We propose a set of assumptions dealing with particular situations where the state space can be…

概率论 · 数学 2026-01-14 Nicolas Champagnat , Denis Villemonais

In two phase materials, each phase having a non-local response in time, it has been found that for some driving fields the response somehow untangles at specific times, and allows one to directly infer useful information about the geometry…

数学物理 · 物理学 2021-01-06 Ornella Mattei , Graeme W. Milton , Mihai Putinar

The well-known Mori-Zwanzig theory tells us that model reduction leads to memory effect. For a long time, modeling the memory effect accurately and efficiently has been an important but nearly impossible task in developing a good reduced…

机器学习 · 计算机科学 2018-08-14 Chao Ma , Jianchun Wang , Weinan E

The duration, strength and structure of memory effects are crucial properties of physical evolution. Due to the invasive nature of quantum measurement, such properties must be defined with respect to the probing instruments employed. Here,…

量子物理 · 物理学 2021-06-18 Yu Guo , Philip Taranto , Bi-Heng Liu , Xiao-Min Hu , Yun-Feng Huang , Chuan-Feng Li , Guang-Can Guo