中文
相关论文

相关论文: Sharp Spectral Thresholds for Logit Fixed Points

200 篇论文

Robust reinforcement learning (RL) is to find a policy that optimizes the worst-case performance over an uncertainty set of MDPs. In this paper, we focus on model-free robust RL, where the uncertainty set is defined to be centering at a…

机器学习 · 计算机科学 2021-10-29 Yue Wang , Shaofeng Zou

Symbiotic control synergistically integrates fixed-gain control and adaptive learning architectures to mitigate system uncertainties more predictably than adaptive learning alone and without requiring prior knowledge of uncertainty bounds…

系统与控制 · 电气工程与系统科学 2024-11-18 Emre Yildirim , Tansel Yucelen , John T. Hrynuk

Standard negative log-likelihood (NLL) for Supervised Fine-Tuning (SFT) applies uniform token-level weighting. This rigidity creates a two-fold failure mode: (i) overemphasizing low-probability targets can amplify gradients on noisy…

计算与语言 · 计算机科学 2026-02-13 Zecheng Wang , Deyuan Liu , Chunshan Li , Yupeng Zhang , Zhengyun Zhao , Dianhui Chu , Bingning Wang , Dianbo Sui

This note studies the robust output feedback stabilization problem of a class of multi-input multi-output invertible nonlinear systems, for which an "ideal" state feedback based on feedback linearization can be designed under certain mild…

系统与控制 · 电气工程与系统科学 2021-01-07 Lei Wang , Christopher M. Kellett

The primary objective of this paper is to demonstrate that problems related to stability and robust control in the harmonic context can be effectively addressed by formulating them as semidefinite optimization problems, invoking the concept…

系统与控制 · 电气工程与系统科学 2023-11-13 Flora Vernerey , Pierre Riedinger , Jamal Daafouz

The softmax content-based attention mechanism has proven to be very beneficial in many applications of recurrent neural networks. Nevertheless it suffers from two major computational limitations. First, its computations for an attention…

机器学习 · 计算机科学 2016-09-20 Alexandre de Brébisson , Pascal Vincent

We extend the definition of $n$-dimensional difference equations to complex order $\alpha\in \mathbb{C} $. We investigate the stability of linear systems defined by an $n$-dimensional matrix $A$ and derive conditions for the stability of…

动力系统 · 数学 2022-08-29 Sachin Bhalekar , Prashant M. Gade , Divya Joshi

For a class of linear switched systems in continuous time a controllability condition implies that state feedbacks allow to achieve almost sure stabilization with arbitrary exponential decay rates. This is based on the Multiplicative…

动力系统 · 数学 2019-01-11 Fritz Colonius , Guilherme Mazanti

This article is concerned with stability analysis and stabilization of randomly switched nonlinear systems. These systems may be regarded as piecewise deterministic stochastic systems: the discrete switches are triggered by a stochastic…

最优化与控制 · 数学 2010-09-08 Debasish Chatterjee , Daniel Liberzon

Linear attention has attracted interest as a computationally efficient approximation to softmax attention, especially for long sequences. Recent studies have explored distilling softmax attention in pre-trained Transformers into linear…

机器学习 · 计算机科学 2025-07-08 Naoki Nishikawa , Rei Higuchi , Taiji Suzuki

Stabilizing feedback operators are presented which depend only on the orthogonal projection of the state onto the finite-dimensional control space. A class of monotone feedback operators mapping the finite-dimensional control space into…

最优化与控制 · 数学 2025-03-10 Karl Kunisch , Sérgio S. Rodrigues , Daniel Walter

This paper deals with designing a robust fixed-order dynamic output feedback controller for uncertain fractional order linear time invariant (FO-LTI) systems by means of linear matrix inequalities (LMIs). Our purpose is to design a low…

最优化与控制 · 数学 2018-02-22 Pouya Badri , Mahdi Sojoodi

This study investigates the absolute stability criteria based on the framework of integral quadratic constraint (IQC) for feedback systems with slope-restricted nonlinearities. In existing works, well-known absolute stability certificates…

Many recent works on stabilization of nonlinear systems target the case of locally stabilizing an unstable steady state solutions against small perturbation. In this work we explicitly address the goal of driving a system into a…

动力系统 · 数学 2020-03-11 Peter Benner , Jan Heiland

This paper proposes a simulation-based reinforcement learning algorithm for controlling systems with uncertain and varying system parameters. While simulators are useful for safely learning control policies, the reality gap remains a major…

系统与控制 · 电气工程与系统科学 2026-05-14 Junya Ikemoto

We construct a patchy feedback for a general control system on $\R^n$ which realizes practical stabilization to a target set $\Sigma$, when the dynamics is constrained to a given set of states $S$. The main result is that $S$--constrained…

最优化与控制 · 数学 2014-08-07 Fabio S. Priuli

Time-delayed feedback control, attributed to Pyragas (1992 Physics Letters 170(6) 421-428), is a method known to stabilise periodic orbits in low dimensional chaotic dynamical systems. A system of the form…

流体动力学 · 物理学 2022-01-21 Dan Lucas , Tatsuya Yasuda

The current series of papers is concerned with stochastic stability of monotone dynamical systems by identifying the basic dynamical units that can survive in the presence of noise interference. In the first of the series, for the…

动力系统 · 数学 2025-11-18 Jifa Jiang , Xi Sheng , Yi Wang

We study finite-time horizon continuous-time linear-convex reinforcement learning problems in an episodic setting. In this problem, the unknown linear jump-diffusion process is controlled subject to nonsmooth convex costs. We show that the…

最优化与控制 · 数学 2022-03-03 Xin Guo , Anran Hu , Yufei Zhang

In adaptive dynamical networks, the dynamics of the nodes and the edges influence each other. We show that we can treat such systems as a closed feedback loop between edge and node dynamics. Using recent advances on the stability of…

适应与自组织系统 · 物理学 2024-11-25 Nina Kastendiek , Jakob Niehues , Robin Delabays , Thilo Gross , Frank Hellmann