中文
相关论文

相关论文: Near Optimal Interpolation based Time-Limited Mode…

200 篇论文

A method is given for solving an optimal H2 approximation problem for SISO linear time-invariant stable systems. The method, based on constructive algebra, guarantees that the global optimum is found; it does not involve any gradient-based…

最优化与控制 · 数学 2007-06-14 Bernard Hanzon , Jan M. Maciejowski , Chun Tung Chou

This paper presents a quasi time optimal receding horizon control algorithm. The proposed algorithm generates near time optimal control when the state of the system is far from the target. When the state attains a certain neighbourhood of…

最优化与控制 · 数学 2007-05-23 Piotr Bania

During the past decade, Model Order Reduction (MOR) has become key enabler for the efficient simulation of large circuit models. MOR techniques based on moment-matching are well established due to their simplicity and computational…

Policy optimization methods are one of the most widely used classes of Reinforcement Learning (RL) algorithms. However, theoretical understanding of these methods remains insufficient. Even in the episodic (time-inhomogeneous) tabular…

机器学习 · 计算机科学 2022-12-06 Tianhao Wu , Yunchang Yang , Han Zhong , Liwei Wang , Simon S. Du , Jiantao Jiao

This work presents a novel algorithm for impulsive optimal control of linear time-varying systems with the inclusion of input magnitude constraints. Impulsive optimal control problems, where the optimal input solution is a sum of delta…

最优化与控制 · 数学 2026-03-17 Ethan Foss , Simone D'Amico

In this paper, we study the use of state-of-the-art nonlinear system identification techniques for the optimal control of nonlinear systems. We show that the nonlinear systems identification problem is equivalent to estimating the…

最优化与控制 · 数学 2023-10-23 Aayushman Sharma , Suman Chakravorty

In current model-free reinforcement learning (RL) algorithms, stability criteria based on sampling methods are commonly utilized to guide policy optimization. However, these criteria only guarantee the infinite-time convergence of the…

机器人学 · 计算机科学 2023-10-16 Shengjie Wang , Fengbo Lan , Xiang Zheng , Yuxue Cao , Oluwatosin Oseni , Haotian Xu , Tao Zhang , Yang Gao

The Krylov subspace projection approach is a well-established tool for the reduced order modeling of dynamical systems in the time domain. In this paper, we address the main issues obstructing the application of this powerful approach to…

数学物理 · 物理学 2012-04-16 Vladimir Druskin , Rob Remis

In this paper, we propose to combine imitation and reinforcement learning via the idea of reward shaping using an oracle. We study the effectiveness of the near-optimal cost-to-go oracle on the planning horizon and demonstrate that the…

机器学习 · 计算机科学 2018-05-30 Wen Sun , J. Andrew Bagnell , Byron Boots

Vision-Language Models (VLMs) have become essential backbones of modern multimodal intelligence, yet their outputs remain prone to hallucination-plausible text misaligned with visual inputs. Existing alignment approaches often rely on…

计算机视觉与模式识别 · 计算机科学 2025-10-28 Kejia Chen , Jiawen Zhang , Jiacong Hu , Kewei Gao , Jian Lou , Zunlei Feng , Mingli Song

We consider low-order controller design for large-scale linear time-invariant dynamical systems with inputs and outputs. Model order reduction is a popular technique, but controllers designed for reduced-order models may result in unstable…

最优化与控制 · 数学 2018-03-20 Peter Benner , Tim Mitchell , Michael L. Overton

In this paper we present a novel extended Krylov subspace reduced-order modeling technique to efficiently simulate time- and frequency-domain wavefields in open complex structures. To simulate the extension to infinity, we use an optimal…

数学物理 · 物理学 2015-06-18 Vladimir Druskin , Rob Remis , Mikhail Zaslavsky

Currently, existing tensor recovery methods fail to recognize the impact of tensor scale variations on their structural characteristics. Furthermore, existing studies face prohibitive computational costs when dealing with large-scale…

机器学习 · 计算机科学 2025-07-09 Wenjin Qin , Hailin Wang , Jingyao Hou , Jianjun Wang

This paper is devoted to the question, whether there is an order barrier $p\leq2$ for time integration in computational elasto-plasticity. In the analysis we use an implicit Runge-Kutta (RK) method of order $p=3$ for integrating the…

数值分析 · 数学 2015-12-22 Bernhard Eidel , Charlotte Kuhn

This paper proposes Proximal Policy Optimization with Linear Temporal Logic Constraints (PPO-LTL), a framework that integrates safety constraints written in LTL into PPO for safe reinforcement learning. LTL constraints offer rigorous…

机器学习 · 计算机科学 2026-03-03 Maifang Zhang , Hang Yu , Qian Zuo , Cheng Wang , Vaishak Belle , Fengxiang He

This paper presents a novel space-time topology optimisation framework for time-dependent thermal conduction problems, aiming to significantly reduce the time-to-solution. By treating time as an additional spatial dimension, we discretise…

计算工程、金融与科学 · 计算机科学 2025-08-14 Joe Alexandersen , Magnus Appel

We introduce a computationally efficient and accurate reduced order modelling approach for the optimization of spatiotemporally chaotic systems. The proposed method combines quantized local reduced order modelling with adjoint-based…

混沌动力学 · 物理学 2026-04-10 Defne E. Ozan , Antonio Colanera , Luca Magri

We propose, analyze, and test new iterative solvers for large-scale systems of linear algebraic equations arising from the finite element discretization of reduced optimality systems defining the finite element approximations to the…

数值分析 · 数学 2023-12-20 Ulrich Langer , Richard Löscher , Olaf Steinbach , Huidong Yang

Krylov quantum diagonalization methods have emerged as a promising use case for quantum computers. However, many existing implementations rely on controlled operations, which pose challenges to near-term quantum hardware. We introduce a…

量子物理 · 物理学 2025-10-15 Nicola Mariella , Enrique Rico , Adam Byrne , Sergiy Zhuk

Recently, Large Language Models (LLMs) have rapidly evolved, approaching Artificial General Intelligence (AGI) while benefiting from large-scale reinforcement learning to enhance Human Alignment (HA) and Reasoning. Recent reward-based…

机器学习 · 计算机科学 2025-06-19 Xuerui Su , Shufang Xie , Guoqing Liu , Yingce Xia , Renqian Luo , Peiran Jin , Zhiming Ma , Yue Wang , Zun Wang , Yuting Liu