English
Related papers

Related papers: When are Kalman-filter restless bandits indexable?

200 papers

We consider the neural contextual bandit problem. In contrast to the existing work which primarily focuses on ReLU neural nets, we consider a general set of smooth activation functions. Under this more general setting, (i) we derive…

Machine Learning · Statistics 2022-06-02 Sudeep Salgia , Sattar Vakili , Qing Zhao

This article is concerned with the inverse problem on determining the temporal component of the source term in a coupled system of time-fractional diffusion equations by single point observation. Under a non-degeneracy condition on the…

Analysis of PDEs · Mathematics 2026-03-13 Mohamed BenSalah , Yikan Liu

Given a linear dynamical system affected by stochastic noise, we consider the problem of selecting an optimal set of sensors (at design-time) to minimize the trace of the steady state a priori or a posteriori error covariance of the Kalman…

Optimization and Control · Mathematics 2020-07-13 Lintao Ye , Nathaniel Woodford , Sandip Roy , Shreyas Sundaram

The classical state-space approach to optimal estimation of stochastic processes is efficient when the driving noises are generated by martingales. In particular, the weight function of the optimal linear filter, which solves a complicated…

Probability · Mathematics 2022-06-13 D. Afterman , P. Chigansky , M. Kleptsyna , D. Marushkevych

The paper studies the asymptotic behavior of Random Algebraic Riccati Equations (RARE) arising in Kalman filtering when the arrival of the observations is described by a Bernoulli i.i.d. process. We model the RARE as an order-preserving,…

Information Theory · Computer Science 2010-05-31 Soummya Kar , Bruno Sinopoli , Jose M. F. Moura

The dynamic allocation problem, also known as the `multi-armed bandit' problem, simulates a situation in which an agent is faced with a tradeoff between actions that yield an immediate reward and actions whose benefits can only be perceived…

Probability · Mathematics 2026-02-03 Christopher Wang

Reinforcement learning is an attractive approach to learn good resource allocation and scheduling policies based on data when the system model is unknown. However, the cumulative regret of most RL algorithms scales as $\tilde O(\mathsf{S}…

Machine Learning · Computer Science 2023-04-28 Nima Akbarzadeh , Aditya Mahajan

Ill-posed inverse problems are ubiquitous in applications. Under- standing of algorithms for their solution has been greatly enhanced by a deep understanding of the linear inverse problem. In the applied communities ensemble-based filtering…

Statistics Theory · Mathematics 2015-12-08 Marco A. Iglesias , Kui Lin , Shuai Lu , Andrew M. Stuart

The ensemble Kalman inversion is widely used in practice to estimate unknown parameters from noisy measurement data. Its low computational costs, straightforward implementation, and non-intrusive nature makes the method appealing in various…

Numerical Analysis · Mathematics 2019-09-04 Dirk Blömker , Claudia Schillings , Philipp Wacker , Simon Weissmann

We extend the relative index theorem on non-compact manifolds to encompass a wide variety of hypoelliptic differential operators of arbitrary order, demonstrating that the change in index when changing a differential operator locally can be…

K-Theory and Homology · Mathematics 2025-11-11 Magnus Fries

This paper extends the ensemble Kalman filter (EnKF) for inverse problems to identify trending model coefficients. This is done by repeatedly inflating the ensemble while maintaining the mean of the particles. As a benchmark serves a…

Optimization and Control · Mathematics 2020-01-30 M. Schwenzer , G. Visconti , M. Ay , T. Bergs , M. Herty , D. Abel

The use of Kalman filtering, as well as its nonlinear extensions, for the estimation of system variables and parameters has played a pivotal role in many fields of scientific inquiry where observations of the system are restricted to a…

Dynamical Systems · Mathematics 2017-02-15 Joseph Arthur , Adam Attarian , Franz Hamilton , Hien Tran

A Schmidt filter is a modification of the Kalman filter that allows to append system parameters as states and considers their uncertainty effect in the filtering process without attempting to estimate such parameters. The states that are…

Systems and Control · Electrical Eng. & Systems 2022-08-29 J Humberto Ramos

We consider Calderon -- Zygmund singular integral in the discrete half-space $h{\bf Z}^m_{+}$, where ${\bf Z}^m$ is entire lattice ($h>0$) in ${\bf R}^m$, and prove that the discrete singular integral operator is invertible in $L_2(h{\bf…

Analysis of PDEs · Mathematics 2014-10-07 Alexander V. Vasilyev , Vladimir B. Vasilyev

In this paper, we analyze the convergence of a risk sensitive like filter where the risk sensitivity parameter is time varying. Such filter has a Kalman like structure and its gain matrix is updated according to a Riccati like iteration. We…

Optimization and Control · Mathematics 2015-09-29 Mattia Zorzi , Bernard C. Levy

We provide a rigorous derivation of the Ensemble Kalman-Bucy Filter as well as the Ensemble Transform Kalman-Bucy Filter in case of nonlinear, unbounded model and observation operators. We identify them as the continuous time limit of the…

Probability · Mathematics 2021-11-29 Theresa Lange

We consider the restless Markov bandit problem, in which the state of each arm evolves according to a Markov process independently of the learner's actions. We suggest an algorithm that after $T$ steps achieves $\tilde{O}(\sqrt{T})$ regret…

Machine Learning · Computer Science 2012-10-23 Ronald Ortner , Daniil Ryabko , Peter Auer , Rémi Munos

Directional estimation is a common problem in many tracking applications. Traditional filters such as the Kalman filter perform poorly because they fail to take the periodic nature of the problem into account. We present a recursive filter…

Systems and Control · Computer Science 2013-05-01 Gerhard Kurz , Igor Gilitschenski , Simon Julier , Uwe D. Hanebeck

The problem of rested and restless multi-armed bandits with constrained availability of arms is considered. The states of arms evolve in Markovian manner and the exact states are hidden from the decision maker. First, some structural…

Systems and Control · Computer Science 2017-10-20 Varun Mehta , Rahul Meshram , Kesav Kaza , S. N. Merchant

Time-scale theory, due to its ability to unify the continuous and discrete cases, allows handling intractable non-uniform measurements, such as intermittent received signals. In this work, we address the state estimation problem of a…

Systems and Control · Electrical Eng. & Systems 2022-03-31 Wenqi Cai , Bacem Ben Nasser , Mohamed Djemai , Taous Meriem Laleg-Kirati