中文
相关论文

相关论文: Reinforcement Solver for H-infinity Filter with Bo…

200 篇论文

This paper investigates the fuzzy $H_{\infty}$ filter design issue for nonlinear systems with time-varying delay. In order to obtain less conservative fuzzy $H_{\infty}$ filter design method, a novel integral inequality is employed to…

系统与控制 · 电气工程与系统科学 2022-09-14 Qianqian Ma , Li Li , Junhui Shen , Haowei Guan , Guangcheng Ma , Hongwei Xia

This paper mainly discusses the $H_{\infty}$ filtering of general nonlinear discrete time-varying stochastic systems. A nonlinear discrete-time stochastic bounded real lemma (SBRL) is firstly obtained by means of the smoothness of the…

最优化与控制 · 数学 2018-12-21 Tianliang Zhang , Feiqi Deng , Weihai Zhang

Inverse optimization refers to the inference of unknown parameters of an optimization problem based on knowledge of its optimal solutions. This paper considers inverse optimization in the setting where measurements of the optimal solutions…

最优化与控制 · 数学 2017-12-27 Anil Aswani , Zuo-Jun Max Shen , Auyon Siddiq

The paper introduces an interactive machine learning mechanism to process the measurements of an uncertain, nonlinear dynamic process and hence advise an actuation strategy in real-time. For concept demonstration, a trajectory-following…

系统与控制 · 电气工程与系统科学 2023-03-16 Mohammed Abouheaf , Derek Boase , Wail Gueaieb , Davide Spinello , Salah Al-Sharhan

This paper is concerned with the problem of Model Predictive Control and Rolling Horizon Control of discrete-time systems subject to possibly unbounded random noise inputs, while satisfying hard bounds on the control inputs. We use a…

最优化与控制 · 数学 2010-09-08 Peter Hokayem , Debasish Chatterjee , John Lygeros

We study momentum-based first-order optimization algorithms in which the iterations utilize information from the two previous steps and are subject to an additive white noise. This setup uses noise to account for uncertainty in either…

最优化与控制 · 数学 2024-06-21 Hesameddin Mohammadi , Meisam Razaviyayn , Mihailo R. Jovanović

Reinforcement Learning (RL) has emerged as a powerful framework for sequential decision-making in dynamic environments, particularly when system parameters are unknown. This paper investigates RL-based control for entropy-regularized…

系统与控制 · 电气工程与系统科学 2025-12-02 Gabriel Diaz , Lucky Li , Wenhao Zhang

Recent literature has proposed approaches that learn control policies with high performance while maintaining safety guarantees. Synthesizing Hamilton-Jacobi (HJ) reachable sets has become an effective tool for verifying safety and…

系统与控制 · 电气工程与系统科学 2024-08-23 Milan Ganai , Sicun Gao , Sylvia Herbert

In this paper, we address the inverse problem for linear-quadratic differential non-cooperative games with output-feedback. Given players' stabilizing feedback laws, the goal is to find cost function parameters that lead to a game for which…

最优化与控制 · 数学 2024-10-27 Emin Martirosyan , Ming Cao

In this paper, we give a causal solution to the problem of spline interpolation using H-infinity optimal approximation. Generally speaking, spline interpolation requires filtering the whole sampled data, the past and the future, to…

信息论 · 计算机科学 2013-08-14 Masaaki Nagahara , Yutaka Yamamoto

Deep learning-based hearing loss compensation (HLC) seeks to enhance speech intelligibility and quality for hearing impaired listeners using neural networks. One major challenge of HLC is the lack of a ground-truth target. Recent works have…

音频与语音处理 · 电气工程与系统科学 2025-11-04 Philippe Gonzalez , Torsten Dau , Tobias May

In this paper, we consider the problem of estimating parameters of a linear regression model. Using a hybrid systems framework, a hybrid algorithm is proposed allowing the estimate to converge to the exact value of the unknown parameters in…

系统与控制 · 电气工程与系统科学 2026-03-04 Adnane Saoud , Ryan S. Johnson , Ricardo G. Sanfelice

We propose a mesh-free policy iteration framework that combines classical dynamic programming with physics-informed neural networks (PINNs) to solve high-dimensional, nonconvex Hamilton--Jacobi--Isaacs (HJI) equations arising in stochastic…

数值分析 · 数学 2025-07-24 Hee Jun Yang , Minjung Gim , Yeoneung Kim

In most machine learning applications, classification accuracy is not the primary metric of interest. Binary classifiers which face class imbalance are often evaluated by the $F_\beta$ score, area under the precision-recall curve, Precision…

机器学习 · 计算机科学 2018-03-02 Alan Mackey , Xiyang Luo , Elad Eban

Policy-gradient methods are widely used in reinforcement learning, yet training often becomes unstable or slows down as learning progresses. We study this phenomenon through the noise-to-signal ratio (NSR) of a policy-gradient estimator,…

最优化与控制 · 数学 2026-02-10 Haoyu Han , Heng Yang

Quadratic programming is a workhorse of modern nonlinear optimization, control, and data science. Although regularized methods offer convergence guarantees under minimal assumptions on the problem data, they can exhibit the slow…

最优化与控制 · 数学 2026-05-18 Jeremy Bertoncini , Alberto De Marchi , Matthias Gerdts , Simon Gottschalk

We explore the use of policy gradient methods in reinforcement learning for quantum control via energy landscape shaping of XX-Heisenberg spin chains in a model agnostic fashion. Their performance is compared to finding controllers using…

量子物理 · 物理学 2022-07-19 I. Khalid , C. A. Weidner , E. A. Jonckheere , S. G. Schirmer , F. C. Langbein

This paper introduces a novel approach to design of functional H_\infty filters for a class of nonlinear descriptor systems subjected to disturbances. Departing from conventional assumptions regarding system regularity, we adopt a more…

最优化与控制 · 数学 2024-09-10 Rishabh Sharma , Mahendra Kumar Gupta , Nutan Kumar Tomar

This paper studies linear quadratic Gaussian robust mean field social control problems in the presence of multiplicative noise. We aim to compute asymptotic decentralized strategies without requiring full prior knowledge of agents'…

系统与控制 · 电气工程与系统科学 2025-09-16 Zhenhui Xu , Jiayu Chen , Bing-Chang Wang , Yuhu Wu , Tielong Shen

The linear functional strategy for the regularization of inverse problems is considered. For selecting the regularization parameter therein, we propose the heuristic quasi-optimality principle and some modifications including the smoothness…

数值分析 · 数学 2018-05-23 Stefan Kindermann , Sergiy Pereverzyev , Andrey Pilipenko