中文
相关论文

相关论文: Reinforcement Solver for H-infinity Filter with Bo…

200 篇论文

In this paper, the reinforcement learning (RL)-based optimal control problem is studied for multiplicative-noise systems, where input delay is involved and partial system dynamics is unknown. To solve a variant of Riccati-ZXL equations,…

最优化与控制 · 数学 2023-01-10 Hongxia Wang , Fuyu Zhao , Zhaorong Zhang , Juanjuan Xu , Xun Li

In this paper, we propose an improved numerical algorithm for solving minimax problems based on nonsmooth optimization, quadratic programming and iterative process. We also provide a rigorous proof of convergence for our algorithm under…

人工智能 · 计算机科学 2025-07-02 Qing Xu , Xiaohua Xuan

We consider an affine process $X$ which is only observed up to an additive white noise, and we ask for its law, for some time $t > 0 $, conditional on all observations up to this time $ t $. This is a general, possibly high dimensional…

概率论 · 数学 2018-01-25 Lukas Gonon , Josef Teichmann

According to the fundamental laws of quantum optics, noise is necessarily added to the system when one tries to clone or amplify a quantum state. However, it has recently been shown that the quantum noise related to the operation of a…

量子物理 · 物理学 2013-12-18 Mikko Partanen , Teppo Häyrynen , Jani Oksanen , Jukka Tulkki

Deconvolution is a widely used strategy to mitigate the blurring and noisy degradation of hyperspectral images~(HSI) generated by the acquisition devices. This issue is usually addressed by solving an ill-posed inverse problem. While…

图像与视频处理 · 电气工程与系统科学 2023-05-03 Xiuheng Wang , Jie Chen , Cédric Richard

When measurements from dynamical systems are noisy, it is useful to have estimation algorithms that have low sensitivity to measurement noises and outliers. In the first set of results described in this paper we obtain optimal estimators…

系统与控制 · 电气工程与系统科学 2022-09-20 Krishan Mohan Nagpal

Nonholonomic mechanical systems have been attracting more interest in recent years because of their rich geometric properties and their applications in Engineering. In all generality, we discuss the reduction of a Hamilton-Jacobi theory for…

数学物理 · 物理学 2019-10-23 Oğul Esen , Manuel de León , Víctor Manuel Jiménez Morales , Cristina Sardón

We consider reinforcement learning (RL) methods for finding optimal policies in linear quadratic (LQ) mean field control (MFC) problems over an infinite horizon in continuous time, with common noise and entropy regularization. We study…

最优化与控制 · 数学 2024-08-06 Noufel Frikha , Huyên Pham , Xuanye Song

We consider a convex optimization problem with many linear inequality constraints. To deal with a large number of constraints, we provide a penalty reformulation of the problem, where the penalty is a variant of the one-sided Huber loss…

最优化与控制 · 数学 2023-11-03 Angelia Nedich , Tatiana Tatarenko

Commonly in reinforcement learning (RL), rewards are discounted over time using an exponential function to model time preference, thereby bounding the expected long-term reward. In contrast, in economics and psychology, it has been shown…

机器学习 · 计算机科学 2022-12-08 Matthias Schultheis , Constantin A. Rothkopf , Heinz Koeppl

In this paper, a synthesis method for distributed estimation is presented, which is suitable for dealing with large-scale interconnected linear systems with disturbance. The main feature of the proposed method is that local estimators only…

系统与控制 · 计算机科学 2015-12-08 Jingbo Wu , Valery Ugrinovskii , Frank Allgöwer

The problem of reinforcement learning is considered where the environment or the model undergoes a change. An algorithm is proposed that an agent can apply in such a problem to achieve the optimal long-time discounted reward. The algorithm…

系统与控制 · 电气工程与系统科学 2023-04-25 Wuxia Chen , Taposh Banerjee , Jemin George , Carl Busart

The aim of this paper is to design a band-limited optimal input with power constraints for identifying a linear multi-input multi-output system. It is assumed that the nominal system parameters are specified. The key idea is to use the…

系统与控制 · 计算机科学 2017-06-14 Shravan Mohan , Mithun Im , Bharath Bhikkaji

This paper characterizes the solution to a finite horizon min-max optimal control problem where the system is linear and discrete-time with control and state constraints, and the cost quadratic; the disturbance is negatively costed, as in…

最优化与控制 · 数学 2017-10-13 D. Q. Mayne , S. V. Rakovic , R. B. Vinter , E. C. Kerrigan

This paper studies the continuous-time reinforcement learning for stochastic singular control with the application to an infinite-horizon irreversible reinsurance problems. The singular control is equivalently characterized as a pair of…

最优化与控制 · 数学 2025-12-03 Zongxia Liang , Xiaodong Luo , Xiang Yu

Gradient-descent based iterative algorithms pervade a variety of problems in estimation, prediction, learning, control, and optimization. Recently iterative algorithms based on higher-order information have been explored in an attempt to…

机器学习 · 计算机科学 2021-03-25 Spencer McDonald , Yingnan Cui , Joseph E. Gaudio , Anuradha M. Annaswamy

We present for the first time an asymptotic convergence analysis of two time-scale stochastic approximation driven by "controlled" Markov noise. In particular, the faster and slower recursions have non-additive controlled Markov noise…

机器学习 · 计算机科学 2020-12-03 Prasenjit Karmakar

We address the problem of blind gain and phase calibration of a sensor array from ambient noise. The key motivation is to ease the calibration process by avoiding a complex procedure setup. We show that computing the sample covariance…

仪器与探测器 · 物理学 2023-03-22 Charles Vanwynsberghe , Simon Bouley , Jérôme Antoni

We consider entanglement-assisted frequency estimation by Ramsey interferometry, in the presence of dephasing noise from spatiotemporally correlated environments.By working in the widely employed local estimation regime, we show that even…

量子物理 · 物理学 2023-09-20 Francisco Riberi , Gerardo Paz-Silva , Lorenza Viola

In this paper, we develop stochastic variance reduced algorithms for solving a class of finite-sum hemivariational inequality (HVI) problem. In this HVI problem, the associated function is assumed to be differentiable, and both the vector…

最优化与控制 · 数学 2025-09-12 Kevin Huang , Nuozhou Wang , Shuzhong Zhang