English
Related papers

Related papers: Entropy-Rate Selection for Partially Observed Proc…

200 papers

In this paper, we consider the problem of controlling a partially observed Markov decision process (POMDP) in order to actively estimate its state trajectory over a fixed horizon with minimal uncertainty. We pose a novel active smoothing…

Systems and Control · Electrical Eng. & Systems 2021-04-06 Timothy L. Molloy , Girish N. Nair

Efficient exploration is a central problem in reinforcement learning and is often formalized as maximizing the entropy of the state-action occupancy measure. While unconstrained maximum-entropy exploration is relatively well understood,…

Machine Learning · Computer Science 2026-05-01 Florian Wolf , Ilyas Fatkhullin , Niao He

We study finite particle systems on the one-dimensional integer lattice, where each particle performs a continuous-time nearest-neighbour random walk, with jump rates intrinsic to each particle, subject to an exclusion interaction which…

Probability · Mathematics 2024-05-07 Vadim Malyshev , Mikhail Menshikov , Serguei Popov , Andrew Wade

In a recent paper, the authors proposed a general methodology for probabilistic learning on manifolds. The method was used to generate numerical samples that are statistically consistent with an existing dataset construed as a realization…

Probability · Mathematics 2018-03-30 C. Soizea , R. Ghanem , C. Safta , X. Huan , Z. P. Vane , J. Oefelein , G. Lacaz , H. N. Najm , Q. Tang , X. Chen

Many natural and engineered systems can be modeled as discrete state Markov processes. Often, only a subset of states are directly observable. Inferring the conditional probability that a system occupies a particular hidden state, given the…

Signal Processing · Electrical Eng. & Systems 2023-01-04 Daniel Chen , Alexander G. Strang , Andrew W. Eckford , Peter J. Thomas

Depending on context, the term entropy is used for a thermodynamic quantity, a~measure of available choice, a quantity to measure information, or, in the context of statistical inference, a maximum configuration predictor. For systems in…

Statistical Mechanics · Physics 2018-11-14 Rudolf Hanel , Stefan Thurner

We investigate entropy minimization problems for quantum states subject to convex block-separable constraints. Our principal result is a quantitative stability theorem: under a natural confining (fixed-support) hypothesis, if a state has…

Quantum Physics · Physics 2026-01-21 Hassan Nasreddine

Finding observing path creating its observer is important problem in physics and information science. In observing processes, each observation is act changing the observing process that generates interactive observation. Each interaction is…

Adaptation and Self-Organizing Systems · Physics 2020-07-09 Vladimir S. Lerner

We propose and study a general framework for regularized Markov decision processes (MDPs) where the goal is to find an optimal policy that maximizes the expected discounted total reward plus a policy regularization term. The extant…

Machine Learning · Statistics 2019-10-22 Xiang Li , Wenhao Yang , Zhihua Zhang

We extend the notion of estimation entropy of autonomous dynamical systems proposed by Liberzon and Mitra [1] to nonlinear dynamical systems with uncertain inputs with bounded variation. We call this new notion the {$\epsilon$}-estimation…

Systems and Control · Electrical Eng. & Systems 2023-11-14 Hussein Sibai , Sayan Mitra

A randomized algorithm for finding sparse cuts is given which is based on constructing a dual markov chain called multiscale rings process(MRP) and a new concept of entropy. It is shown how the time to absorption of the dual process…

Probability · Mathematics 2022-03-16 Farshad Noravesh

In this paper, we explore a scenario where a sender provides an information policy and a receiver, upon observing a realization of this policy, decides whether to take a particular action, such as making a purchase. The sender's objective…

Numerical Analysis · Mathematics 2024-12-13 Jorge Justiniano , Andreas Kleiner , Benny Moldovanu , Martin Rumpf , Philipp Strack

We consider the problem of finding the best memoryless stochastic policy for an infinite-horizon partially observable Markov decision process (POMDP) with finite state and action spaces with respect to either the discounted or mean reward…

Optimization and Control · Mathematics 2022-05-02 Johannes Müller , Guido Montúfar

We show that the naive application of the maximum entropy principle can yield answers which depend on the level of description, i.e. the result is not invariant under coarse-graining. We demonstrate that the correct approach, even for…

Statistical Mechanics · Physics 2007-05-23 Jayanth Banavar , Amos Maritan

The maximum entropy principle from statistical mechanics states that a closed system attains an equilibrium distribution that maximizes its entropy. We first show that for graphs with fixed number of edges one can define a stochastic edge…

Disordered Systems and Neural Networks · Physics 2007-05-23 Jesse S. A. Bridgewater , P. Oscar Boykin , Vwani P. Roychowdhury

Information theory on a time-discrete setting in the framework of time series analysis is generalized to the time-continuous case. Considerations of the Roessler and Lorenz dynamics as well as the Ornstein-Uhlenbeck process yield for…

Chaotic Dynamics · Physics 2008-06-04 Detlef Holstein

In the theory of Partially Observed Markov Decision Processes (POMDPs), existence of optimal policies have in general been established via converting the original partially observed stochastic control problem to a fully observed one on the…

Optimization and Control · Mathematics 2022-01-11 Ali Devran Kara , Serdar Yuksel

We investigate the memory properties of discrete sequences built upon a finite number of states. We find that the block entropy can reliably determine the memory for systems modeled as Markov chains of arbitrary finite order. Further, we…

Statistical Mechanics · Physics 2022-11-21 Juan De Gregorio , David Sanchez , Raul Toral

Starting from the Avellaneda-Stoikov framework, we consider a market maker who wants to optimally set bid/ask quotes over a finite time horizon, to maximize her expected utility. The intensities of the orders she receives depend not only on…

Trading and Market Microstructure · Quantitative Finance 2020-06-29 Diego Zabaljauregui , Luciano Campi

In this paper, we use entropy functions to characterise the set of rate-capacity tuples achievable with either zero decoding error, or vanishing decoding error, for general network coding problems. We show that when sources are colocated,…

Information Theory · Computer Science 2015-03-19 Terence H. Chan , Alex Grant
‹ Prev 1 3 4 5 6 7 10 Next ›