English
Related papers

Related papers: Reinforcement Learning, Optimal Control, and Bayes…

200 papers

Reinforcement learning with verifiable rewards (RLVR) has become a standard approach for improving reasoning in language models, yet models trained with RLVR often suffer from diversity collapse: while single-sample accuracy improves,…

Machine Learning · Computer Science 2026-05-05 Marc Dymetman

Online reinforcement learning with verifiable rewards (RLVR) turns checkable outcomes into a scalable training signal, but it keeps rollout generation, verifier scoring, and reference-policy evaluations on the optimization path. Static…

Machine Learning · Computer Science 2026-05-05 Yao Shu , Chenxing Wei , Hongbin Lin , Shuang Qiu , Hui Xiong

We derive a novel, provably robust, and closed-form Bayesian update rule for online filtering in state-space models in the presence of outliers and misspecified measurement models. Our method combines generalised Bayesian inference with…

Data assimilation (DA) aims to optimally combine model forecasts and observations that are both partial and noisy. Multi-model DA generalizes the variational or Bayesian formulation of the Kalman filter, and we prove that it is also the…

Methodology · Statistics 2023-01-23 Eviatar Bach , Michael Ghil

Backpropagation dominates modern machine learning, yet it is not the only principled method for optimizing dynamical systems. We propose Kalman World Models (KWM), a class of learned state-space models trained via recursive Bayesian…

Machine Learning · Computer Science 2026-03-17 Andrew Kiruluta

This paper aims at the algorithmic/theoretical core of reinforcement learning (RL) by introducing the novel class of proximal Bellman mappings. These mappings are defined in reproducing kernel Hilbert spaces (RKHSs), to benefit from the…

Signal Processing · Electrical Eng. & Systems 2023-09-15 Yuki Akiyama , Konstantinos Slavakis

While Contrastive Learning (CL) has revolutionized self-supervised representation learning, its latent representations remain highly entangled and opaque, limiting their interpretability in safety-critical applications. We identify that a…

Computer Vision and Pattern Recognition · Computer Science 2026-05-28 Peng Cui , Jiahao Zhang , Lijie Hu

We use Markov categories to generalize the basic theory of Markov chains and hidden Markov models to an abstract setting. This comprises characterizations of hidden Markov models in terms of conditional independences and algorithms for…

Statistics Theory · Mathematics 2025-08-26 Tobias Fritz , Andreas Klingler , Drew McNeely , Areeb Shah-Mohammed , Yuwen Wang

A hybrid data assimilation algorithm is developed for complex dynamical systems with partial observations. The method starts with applying a spectral decomposition to the entire spatiotemporal fields, followed by creating a machine learning…

Computational Physics · Physics 2022-12-27 Changhong Mou , Leslie M. Smith , Nan Chen

Data assimilation algorithms are used to estimate the states of a dynamical system using partial and noisy observations. The ensemble Kalman filter has become a popular data assimilation scheme due to its simplicity and robustness for a…

Numerical Analysis · Mathematics 2021-06-23 Gottfried Hastermann , Maria Reinhardt , Rupert Klein , Sebastian Reich

This paper presents an adaptive Kalman filter for a linear dynamic system perturbed by an additive disturbance. The objective is to estimate both of the state and the unknown disturbance concurrently, while learning the disturbance as a…

Optimization and Control · Mathematics 2019-10-23 Taeyoung Lee

This paper analyzes a popular computational framework to solve infinite-dimensional Bayesian inverse problems, discretizing the prior and the forward model in a finite-dimensional weighted inner product space. We demonstrate the benefit of…

Numerical Analysis · Mathematics 2024-02-22 Daniel Sanz-Alonso , Nathan Waniorek

We propose a reinforcement learning (RL)-based algorithm to jointly train (1) a trajectory planner and (2) a tracking controller in a layered control architecture. Our algorithm arises naturally from a rewrite of the underlying optimal…

Systems and Control · Electrical Eng. & Systems 2024-12-18 Fengjun Yang , Nikolai Matni

Bayesian filtering deals with computing the posterior distribution of the state of a stochastic dynamic system given noisy observations. In this paper, motivated by applications in counter-adversarial systems, we consider the following…

Systems and Control · Electrical Eng. & Systems 2020-10-28 Robert Mattila , Cristian R. Rojas , Vikram Krishnamurthy , Bo Wahlberg

We consider particle filters with weakly informative observations (or `potentials') relative to the latent state dynamics. The particular focus of this work is on particle filters to approximate time-discretisations of continuous-time…

Computation · Statistics 2022-07-12 Nicolas Chopin , Sumeetpal S. Singh , Tomás Soto , Matti Vihola

The theory of Bayesian learning incorporates the use of Student-t Processes to model heavy-tailed distributions and datasets with outliers. However, despite Student-t Processes having a similar computational complexity as Gaussian…

Machine Learning · Computer Science 2025-08-12 Jian Xu , Delu Zeng

We study the problem of optimal estimation and control of linear systems using quantized measurements, with a focus on applications over sensor networks. We show that the state conditioned on a causal quantization of the measurements can be…

Information Theory · Computer Science 2015-03-13 Ravi Teja Sukhavasi , Babak Hassibi

Gradient-based methods have been widely used for system design and optimization in diverse application domains. Recently, there has been a renewed interest in studying theoretical properties of these methods in the context of control and…

Optimization and Control · Mathematics 2022-10-11 Bin Hu , Kaiqing Zhang , Na Li , Mehran Mesbahi , Maryam Fazel , Tamer Başar

The Kalman filter is an algorithm for the estimation of hidden variables in dynamical systems under linear Gauss-Markov assumptions with widespread applications across different fields. Recently, its Bayesian interpretation has received a…

Neurons and Cognition · Quantitative Biology 2021-11-23 Manuel Baltieri , Takuya Isomura

We develop a purely categorical theory of action filtrations and their associated growth invariants. When specialized to categories of geometric interest, such as the wrapped Fukaya category of a Weinstein manifold, and the bounded derived…

Symplectic Geometry · Mathematics 2023-05-22 Laurent Côté , Yusuf Barış Kartal
‹ Prev 1 4 5 6 7 8 10 Next ›