中文
相关论文

相关论文: The Infinite-Dimensional Standard and Strict Bound…

200 篇论文

We study reinforcement learning (RL) in the setting of continuous time and space, for an infinite horizon with a discounted objective and the underlying dynamics driven by a stochastic differential equation. Built upon recent advances in…

机器学习 · 计算机科学 2023-10-19 Hanyang Zhao , Wenpin Tang , David D. Yao

This paper introduces the Progressive Barrier Lyapunov Function (p-BLF) for output- and full-state-constrained nonlinear control systems. Unlike traditional BLF methods, where control effort continuously increases as the state approaches…

系统与控制 · 电气工程与系统科学 2025-07-03 Hamed Rahimi Nohooji , Holger Voos

Statistical Relational Learning (SRL) methods for anomaly detection are introduced via a security-related application. Operational requirements for online learning stability are outlined and compared to mathematical definitions as applied…

机器学习 · 计算机科学 2017-05-19 Magnus Jändel , Pontus Svenson , Niclas Wadströmer

We study the possibility of testing local realistic theory (LRT), envisioned implicitly by Einstein, Podolsky and Rosen in 1935, based on the Bell inequality for the correlations in the decay modes of entangled K or B-mesons. It is shown…

量子物理 · 物理学 2008-12-18 Tsubasa Ichikawa , Satoshi Tamura , Izumi Tsutsui

We propose a continuous-time formulation of persistent contrastive divergence (PCD) for maximum likelihood estimation (MLE) of unnormalised densities. Our approach expresses PCD as a coupled, multiscale system of stochastic differential…

机器学习 · 统计学 2025-10-03 Paul Felix Valsecchi Oliva , O. Deniz Akyildiz , Andrew Duncan

Limit theorems for the time average of some observation functions in an infinite measure dynamical system are studied. It is known that intermittent phenomena, such as the Rayleigh-Benard convection and Belousov-Zhabotinsky reaction, are…

统计力学 · 物理学 2010-05-14 Takuma Akimoto

In this paper, we delve into the intricacies of boundary stabilization for the linearized KP-II equation within the constraints of a bounded domain, a phenomenon known as ``critical length." Our primary aim is to design a feedback law that…

偏微分方程分析 · 数学 2025-04-09 F. A. Gallego , J. R. Muñoz

Inspired by the widespread concept of Lyapunov-Krasovskii functionals of complete type, this article proposes an alternative class of functionals, termed Lyapunov-Krasovskii functionals of robust type. Their construction aims at improving…

系统与控制 · 电气工程与系统科学 2025-11-12 Tessina H. Scholl

We study reinforcement learning in infinite-horizon discounted Markov decision processes with continuous state spaces, where data are generated online from a single trajectory under a Markovian behavior policy. To avoid maintaining an…

机器学习 · 计算机科学 2026-03-05 Shengbo Wang

Reinforcement Learning (RL) is a computational approach to reward-driven learning in sequential decision problems. It implements the discovery of optimal actions by learning from an agent interacting with an environment rather than from…

统计方法学 · 统计学 2022-10-06 Mauricio Tec , Yunshan Duan , Peter Müller

This paper examines reinforcement learning (RL) in infinite-horizon decision processes with almost-sure safety constraints, crucial for applications like autonomous systems, finance, and resource management. We propose a doubly-regularized…

机器学习 · 计算机科学 2025-09-17 Pekka Malo , Lauri Viitasaari , Antti Suominen , Eeva Vilkkumaa , Olli Tahvonen

The ability to direct a Probabilistic Boolean Network (PBN) to a desired state is important to applications such as targeted therapeutics in cancer biology. Reinforcement Learning (RL) has been proposed as a framework that solves a…

机器学习 · 计算机科学 2022-10-26 Sotiris Moschoyiannis , Evangelos Chatzaroulas , Vytenis Sliogeris , Yuhu Wu

Deep reinforcement learning (RL) uses model-free techniques to optimize task-specific control policies. Despite having emerged as a promising approach for complex problems, RL is still hard to use reliably for real-world applications. Apart…

机器人学 · 计算机科学 2020-02-25 Siddhant Gangapurwala , Alexander Mitchell , Ioannis Havoutis

Interest in reinforcement learning (RL) for large-scale systems, comprising extensive populations of intelligent agents interacting with heterogeneous environments, has surged significantly across diverse scientific domains in recent years.…

系统与控制 · 电气工程与系统科学 2025-09-16 Wei Zhang , Jr-Shin Li

Reinforcement learning algorithms are typically designed for discrete-time dynamics, even though the underlying real-world control systems are often continuous in time. In this paper, we study the problem of continuous-time reinforcement…

机器学习 · 计算机科学 2026-03-03 Klemens Iten , Lenart Treven , Bhavya Sukhija , Florian Dörfler , Andreas Krause

Reinforcement learning (RL) is a key paradigm for post-training large language models (LLMs), but the widely used Group Relative Policy Optimization (GRPO) often suffers from entropy collapse: exploration quickly disappears, policies…

机器学习 · 计算机科学 2026-05-19 Chen Wang , Zhaochun Li , Jionghao Bai , Hexuan Deng , Ge Lan , Yue Wang

LCRL is a software tool that implements model-free Reinforcement Learning (RL) algorithms over unknown Markov Decision Processes (MDPs), synthesising policies that satisfy a given linear temporal specification with maximal probability. LCRL…

机器学习 · 计算机科学 2022-09-22 Hosein Hasanbeig , Daniel Kroening , Alessandro Abate

We present a fundamentally new proof of the dimensionless Lp boundedness of the Bakry Riesz vector on manifolds with bounded geometry. Our proof has the significant advantage that it allows for a much stronger conclusion than previous…

概率论 · 数学 2023-03-30 Komla Domelevo , Stefanie Petermichl , Kristina Ana Škreb

Infinite-time nonlinear optimal regulation control is widely utilized in aerospace engineering as a systematic method for synthesizing stable controllers. However, conventional methods often rely on linearization hypothesis, while recent…

系统与控制 · 电气工程与系统科学 2025-06-13 Han Wang , Di Wu , Lin Cheng , Shengping Gong , Xu Huang

In this paper, we study the relative controllability of linear difference equations with multiple delays in the state by using a suitable formula for the solutions of such systems in terms of their initial conditions, their control inputs,…

最优化与控制 · 数学 2017-10-27 Guilherme Mazanti