English
Related papers

Related papers: Distributed Time-Varying Gaussian Regression via K…

200 papers

We devise a distributional variant of gradient temporal-difference (TD) learning. Distributional reinforcement learning has been demonstrated to outperform the regular one in the recent study \citep{bellemare2017distributional}. In the…

Machine Learning · Computer Science 2019-04-04 Chao Qu , Shie Mannor , Huan Xu

This paper considers the distributed filtering problem for a class of stochastic uncertain systems under quantized data flowing over switching sensor networks. Employing the biased noisy observations of the local sensor and…

Signal Processing · Electrical Eng. & Systems 2019-10-08 Xingkang He , Wenchao Xue , Xiaocheng Zhang , Haitao Fang

We propose DistGP: a multi-robot learning method for collaborative learning of a global function using only local experience and computation. We utilise a sparse Gaussian process (GP) model with a factorisation that mirrors the multi-robot…

Robotics · Computer Science 2026-03-10 Seth Nabarro , Mark van der Wilk , Andrew J. Davison

Distributional reinforcement learning improves performance by capturing environmental stochasticity, but a comprehensive theoretical understanding of its effectiveness remains elusive. In addition, the intractable element of the infinite…

Machine Learning · Computer Science 2025-05-14 Taehyun Cho , Seungyub Han , Seokhun Ju , Dohyeong Kim , Kyungjae Lee , Jungwoo Lee

Many safety-critical real-world problems, such as autonomous driving and collaborative robots, are of a distributed multi-agent nature. To optimize the performance of these systems while ensuring safety, we can cast them as distributed…

Systems and Control · Electrical Eng. & Systems 2025-08-20 Abdullah Tokmak , Thomas B. Schön , Dominik Baumann

Large-scale distributed systems such as sensor networks, often need to achieve filtering and consensus on an estimated parameter from high-dimensional measurements. Running a Kalman filter on every node in such a network is computationally…

Optimization and Control · Mathematics 2017-04-12 Mathias Hudoba de Badyn , Mehran Mesbahi

This paper examines learning the optimal filtering policy, known as the Kalman gain, for a linear system with unknown noise covariance matrices using noisy output data. The learning problem is formulated as a stochastic policy optimization…

Systems and Control · Electrical Eng. & Systems 2023-10-27 Shahriar Talebi , Amirhossein Taghvaei , Mehran Mesbahi

We introduce a variational Bayesian neural network where the parameters are governed via a probability distribution on random matrices. Specifically, we employ a matrix variate Gaussian \cite{gupta1999matrix} parameter posterior…

Machine Learning · Statistics 2016-06-24 Christos Louizos , Max Welling

We study the policy evaluation problem in multi-agent reinforcement learning, modeled by a Markov decision process. In this problem, the agents operate in a common environment under a fixed control policy, working together to discover the…

Optimization and Control · Mathematics 2020-01-13 Thinh T. Doan , Siva Theja Maguluri , Justin Romberg

In this work, we propose a method for training distributed GAN with sequential temporary discriminators. Our proposed method tackles the challenge of training GAN in the federated learning manner: How to update the generator with a flow of…

Computer Vision and Pattern Recognition · Computer Science 2020-07-21 Hui Qu , Yikai Zhang , Qi Chang , Zhennan Yan , Chao Chen , Dimitris Metaxas

Sequential fine-tuning of transformers is useful when new data arrive sequentially, especially with shifting distributions. Unlike batch learning, sequential learning demands that training be stabilized despite a small amount of data by…

Machine Learning · Computer Science 2025-09-16 Haoming Jing , Oren Wright , José M. F. Moura , Yorie Nakahira

We study a distributed Kalman filtering problem in which a number of nodes cooperate without central coordination to estimate a common state based on local measurements and data received from neighbors. This is typically done by running a…

Systems and Control · Electrical Eng. & Systems 2021-02-18 Damián Marelli , Tianju Sui , Minyue Fu

This paper is concerned with developing a novel distributed Kalman filtering algorithm over wireless sensor networks based on randomized consensus strategy. Compared with the centralized algorithm, distributed filtering techniques require…

Systems and Control · Computer Science 2018-10-08 Jiahu Qin , Jie Wang , Ling Shi , Yu Kang

This paper studies distributed continuous-time optimization for time-varying quadratic cost functions with uncertain parameters. We first propose a centralized adaptive optimization algorithm using partial information of the cost function.…

Systems and Control · Electrical Eng. & Systems 2024-07-30 Liangze Jiang , Zheng-Guang Wu , Lei Wang

In this work, we consider the problem of steering the first two moments of the uncertain state of an unknown discrete-time stochastic nonlinear system to a given terminal distribution in finite time. Toward that goal, first, a…

Optimization and Control · Mathematics 2021-04-05 Alexandros Tsolovikos , Efstathios Bakolas

Conventional Bayesian estimation requires an accurate stochastic model of a system. However, this requirement is not always met in many practical cases where the system is not completely known or may differ from the assumed model. For such…

Signal Processing · Electrical Eng. & Systems 2023-04-05 Ranjeet Kumar Tiwari , Shovan Bhaumik

Distributed state estimation strongly depends on collaborative signal processing, which often requires excessive communication and computation to be executed on resource-constrained sensor nodes. To address this problem, we propose an…

Systems and Control · Computer Science 2020-02-19 Amr Alanwar , Hazem Said , Ankur Mehta , Matthias Althoff

The goal of this paper is to study a distributed version of the gradient temporal-difference (GTD) learning algorithm for a class of multi-agent Markov decision processes (MDPs). The temporal-difference (TD) learning is a reinforcement…

Optimization and Control · Mathematics 2020-04-29 Donghwan Lee , Jianghai Hu

Latent variable time-series models are among the most heavily used tools from machine learning and applied statistics. These models have the advantage of learning latent structure both from noisy observations and from the temporal ordering…

Machine Learning · Statistics 2015-11-24 Evan Archer , Il Memming Park , Lars Buesing , John Cunningham , Liam Paninski

Motivated by the maneuvering target tracking with sensors such as radar and sonar, this paper considers the joint and recursive estimation of the dynamic state and the time-varying process noise covariance in nonlinear state space models.…

Systems and Control · Electrical Eng. & Systems 2023-05-09 Hua Lan , Jinjie Hu , Zengfu Wang , Qiang Cheng