中文
相关论文

相关论文: Sharp Spectral Thresholds for Logit Fixed Points

200 篇论文

One of the main criticisms to game theory concerns the assumption of full rationality. Logit dynamics is a decentralized algorithm in which a level of irrationality (a.k.a. "noise") is introduced in players' behavior. In this context, the…

计算机科学与博弈论 · 计算机科学 2014-11-05 Diodato Ferraioli , Carmine Ventre

For a general class of dynamical systems (of which the canonical continuous and uniform discrete versions are but special cases), we prove that there is a state feedback gain such that the resulting closed-loop system is uniformly…

最优化与控制 · 数学 2009-10-19 Billy J. Jackson , John M. Davis , Ian A. Gravagne , Robert J. Marks

Soft robotic manipulators offer operational advantage due to their compliant and deformable structures. However, their inherently nonlinear dynamics presents substantial challenges. Traditional analytical methods often depend on simplifying…

机器人学 · 计算机科学 2024-10-28 Uljad Berdica , Matthew Jackson , Niccolò Enrico Veronese , Jakob Foerster , Perla Maiolino

Policy gradient methods are known to be highly sensitive to the choice of policy parameterization. In particular, the widely used softmax parameterization can induce ill-conditioned optimization landscapes and lead to exponentially slow…

机器学习 · 计算机科学 2026-04-02 Safwan Labbi , Daniil Tiapkin , Paul Mangold , Eric Moulines

In this paper, the feedback stabilization of a linear time-invariant (LTI) multiple-input multiple-output (MIMO) system cascaded by a linear stochastic system is studied in the mean-square sense. Here, the linear stochastic system can model…

系统与控制 · 电气工程与系统科学 2024-05-06 Junhui Li , Jieying Lu , Weizhou Su

This article presents input-output stability analysis of nonlinear feedback systems based on the notion of soft and hard scaled relative graphs (SRGs). The soft and hard SRGs acknowledge the distinction between incremental positivity and…

系统与控制 · 电气工程与系统科学 2026-05-12 Chao Chen , Sei Zhen Khong , Rodolphe Sepulchre

Deep reinforcement learning policies achieve strong performance in complex continuous control environments with nonlinear contact forces. However, these policies often produce chaotic state dynamics, with trivially small changes to the…

机器学习 · 计算机科学 2026-04-28 Rory Young , Nicolas Pugeault

Attention is a core component of transformer architecture, whether encoder-only, decoder-only, or encoder-decoder model. However, the standard softmax attention often produces noisy probability distribution, which can impair effective…

计算与语言 · 计算机科学 2025-11-11 Dhananjay Ram , Wei Xia , Stefano Soatto

The softmax policy gradient (PG) method, which performs gradient ascent under softmax policy parameterization, is arguably one of the de facto implementations of policy optimization in modern reinforcement learning. For $\gamma$-discounted…

机器学习 · 计算机科学 2022-12-19 Gen Li , Yuting Wei , Yuejie Chi , Yuxin Chen

Despite their central role in the success of foundational models and large-scale language modeling, the theoretical foundations governing the operation of Transformers remain only partially understood. Contemporary research has largely…

机器学习 · 计算机科学 2025-06-02 Sagar Ghosh , Kushal Bose , Swagatam Das

This paper addresses the problem of stabilization for infinite-dimensional systems. In particular, we design nonlinear stabilizers for both linear and nonlinear abstract systems. We focus on two classes of systems: the first class comprises…

系统与控制 · 电气工程与系统科学 2025-09-19 Kamal Fenza , Moussa Labbadi , Mohamed Ouzahra

There exist many ways to stabilize an infinite-dimensional linear autonomous control systems when it is possible. Anyway, finding an exponentially stabilizing feedback control that is as simple as possible may be a challenge. The Riccati…

最优化与控制 · 数学 2019-06-26 Emmanuel Trélat , Gengsheng Wang , Yashan Xu

In this paper, we analyze gradient-free methods with one-point feedback for stochastic saddle point problems $\min_{x}\max_{y} \varphi(x, y)$. For non-smooth and smooth cases, we present analysis in a general geometric setup with arbitrary…

最优化与控制 · 数学 2022-09-12 Aleksandr Beznosikov , Vasilii Novitskii , Alexander Gasnikov

We consider control systems of the type $\dot x = A x +\alpha(t)bu$, where $u\in\R$, $(A,b)$ is a controllable pair and $\alpha$ is an unknown time-varying signal with values in $[0,1]$ satisfying a persistent excitation condition i.e.,…

最优化与控制 · 数学 2009-05-18 Yacine Chitour , Mario Sigalotti

This paper investigates the global stability and the global asymptotic stability independent of the sizes of the delays of linear time-varying Caputo fractional dynamic systems of real fractional order possessing internal point delays. The…

动力系统 · 数学 2010-10-18 M. De La Sen

Soft robotics has advanced rapidly, yet its control methods remain fragmented: different morphologies and actuation schemes still require task-specific controllers, hindering theoretical integration and large-scale deployment. A generic…

Designing a stabilizing controller for nonlinear systems is a challenging task, especially for high-dimensional problems with unknown dynamics. Traditional reinforcement learning algorithms applied to stabilization tasks tend to drive the…

系统与控制 · 电气工程与系统科学 2024-09-16 Thanin Quartz , Ruikun Zhou , Hans De Sterck , Jun Liu

This work addresses the design of static output feedback control of discrete-time nonlinear systems satisfying a local Lipschitz continuity condition with time-varying uncertainties. The controller has also a guaranteed disturbance…

系统与控制 · 计算机科学 2016-06-28 Masoud Abbaszadeh , Horacio J. Marquez

Output stabilizability of a class of infinite dimensional linear systems is studied in this paper. A criterion for the system to be output stabilizable by a linear bounded feedback $u=Fx$, $F\in L(Z,\mathbb{R}^{^{p}})$ will be given.

最优化与控制 · 数学 2014-08-08 Faouzi Haddouchi

Scaled graphs (SGs) offer a geometric framework for feedback stability analysis. This paper develops containment conditions for SGs within multiplier-defined regions, addressing both circular and conic geometries. For circular regions, we…

最优化与控制 · 数学 2026-04-08 Eder Baron-Prada , Julius P. J. Krebbekx , Adolfo Anta , Florian Dörfler