中文
相关论文

相关论文: The Infinite-Dimensional Standard and Strict Bound…

200 篇论文

The paper deals with the controllability of finite-dimensional linear difference delay equations, i.e., dynamics for which the state at a given time $t$ is obtained as a linear combination of the control evaluated at time $t$ and of the…

最优化与控制 · 数学 2025-07-16 Yacine Chitour , Sébastien Fueyo , Guilherme Mazanti , Mario Sigalotti

Policy-based methods currently dominate reinforcement learning (RL) pipelines for large language model (LLM) reasoning, leaving value-based approaches largely unexplored. We revisit the classical paradigm of Bellman Residual Minimization…

机器学习 · 计算机科学 2025-11-13 Yurun Yuan , Fan Chen , Zeyu Jia , Alexander Rakhlin , Tengyang Xie

We consider recognizability for Infinite Time Blum-Shub-Smale machines, a model of infinitary computability introduced in Koepke and Seyfferth [KS]. In particular, we show that the lost melody theorem (originally proved for ITTMs in Hamkins…

逻辑 · 数学 2026-05-19 Merlin Carl

We consider an abstract class of infinite-dimensional dynamical systems with inputs. For this class, the significance of noncoercive Lyapunov functions is analyzed. It is shown that the existence of such Lyapunov functions implies…

最优化与控制 · 数学 2022-11-21 B. Jacob , A. Mironchenko , J. R. Partington , F. Wirth

This paper is concerned with conditionally structure-preserving, low regularity time integration methods for a class of semilinear parabolic equations of Allen-Cahn type. Important properties of such equations include maximum bound…

数值分析 · 数学 2022-11-09 Cao-Kha Doan , Thi-Thao-Phuong Hoang , Lili Ju , Katharina Schratz

A delay Lyapunov matrix corresponding to an exponentially stable system of linear time-invariant delay differential equations can be characterized as the solution of a boundary value problem involving a matrix valued delay differential…

数值分析 · 数学 2018-08-28 Wim Michiels , Bin Zhou

This paper discusses a general and useful stability principle which, roughly speaking, says that given a uniformly continuous function defined on an arbitrary metric space, if the function is bounded on the constraint set and we slightly…

最优化与控制 · 数学 2020-09-04 Daniel Reem , Simeon Reich , Alvaro De Pierro

The exponential growth of data-intensive applications has placed unprecedented demands on modern storage systems, necessitating dynamic and efficient optimization strategies. Traditional heuristics employed for storage performance…

操作系统 · 计算机科学 2025-08-25 Chiyu Cheng , Chang Zhou , Yang Zhao

Reinforcement Learning (RL) has gained substantial attention across diverse application domains and theoretical investigations. Existing literature on RL theory largely focuses on risk-neutral settings where the decision-maker learns to…

机器学习 · 计算机科学 2024-12-24 Zhengqi Wu , Renyuan Xu

Reinforcement Learning (RL) has shown promise in various robotics applications, yet its deployment on real systems is still limited due to safety and operational constraints. The safe RL field has gained considerable attention in recent…

机器人学 · 计算机科学 2026-03-19 Sadık Bera Yüksel , Ali Tevfik Buyukkocak , Derya Aksaray

Recent progress in generative modeling has highlighted the importance of Reinforcement Learning (RL) for fine-tuning, with KL-regularized methods in particular proving to be highly effective for both autoregressive and diffusion models.…

机器学习 · 计算机科学 2025-09-03 Tristan Deleu , Padideh Nouri , Yoshua Bengio , Doina Precup

One reason for the well known fact that the Complex Langevin (CL) method sometimes fails to converge or converges to the wrong limit has been identified long ago: it is insufficient decay of the probability density either near infinity or…

高能物理 - 格点 · 物理学 2020-01-15 M. Scherzer , E. Seiler , D. Sexty , I. -O. Stamatescu

The Krylov subspace projection approach is a well-established tool for the reduced order modeling of dynamical systems in the time domain. In this paper, we address the main issues obstructing the application of this powerful approach to…

数学物理 · 物理学 2012-04-16 Vladimir Druskin , Rob Remis

This paper studies data-driven approaches to the continuous-time linear quadratic regulator (LQR) problem based on two existing parameterizations, namely a closed-loop (CL) parameterization from behavioral system theory and an integral…

最优化与控制 · 数学 2026-05-01 Armin Gießler , Felix Thömmes , Sören Hohmann

The gloabal objective of inverse Reinforcement Learning (IRL) is to estimate the unknown cost function of some MDP base on observed trajectories generated by (approximate) optimal policies. The classical approach consists in tuning this…

机器学习 · 计算机科学 2021-05-26 Firas Jarboui , Vianney Perchet

This study proposes a novel method for developing discretization-consistent closure schemes for implicitly filtered Large Eddy Simulation (LES). Here, the induced filter kernel, and thus the closure terms, are determined by the properties…

流体动力学 · 物理学 2023-12-14 Andrea Beck , Marius Kurz

Offline reinforcement learning (RL) aims to learn decision policies from a fixed batch of logged transitions, without additional environment interaction. Despite remarkable empirical progress, offline RL remains fragile under distribution…

统计方法学 · 统计学 2026-03-16 Debashis Chatterjee

We suggest a new relativity principle, which asserts the impossibility to distinguish the state of rest and the state of motion at the constant velocity of a system, if no work is done to the system in question during its motion. We suggest…

综合物理 · 物理学 2014-07-25 Alexander Kholmetskii , Tolga Yarman , Oleg Missevitch

Deep reinforcement learning (RL) has been recognized as a promising tool to address the challenges in real-time control of power systems. However, its deployment in real-world power systems has been hindered by a lack of formal stability…

系统与控制 · 电气工程与系统科学 2021-10-01 Yuanyuan Shi , Guannan Qu , Steven Low , Anima Anandkumar , Adam Wierman

We extend the Malitsky-Tam forward-reflected-backward (FRB) splitting method for inclusion problems of monotone operators to nonconvex minimization problems. By assuming the generalized concave Kurdyka-{\L}ojasiewicz (KL) property of a…

最优化与控制 · 数学 2021-11-18 Xianfu Wang , Ziyuan Wang
‹ 上一页 1 8 9 10 下一页 ›