中文
相关论文

相关论文: Dynamical Low-Rank Approximation Strategies for No…

200 篇论文

Efficient and accurate numerical approximation of the full Boltzmann equation has been a longstanding challenging problem in kinetic theory. This is mainly due to the high dimensionality of the problem and the complicated collision…

数值分析 · 数学 2021-12-07 Jingwei Hu , Yubo Wang

Reinforcement learning (RL) is a promising, upcoming topic in automatic control applications. Where classical control approaches require a priori system knowledge, data-driven control approaches like RL allow a model-free controller design…

系统与控制 · 电气工程与系统科学 2022-02-01 Daniel Weber , Maximilian Schenke , Oliver Wallscheid

Reliable simultaneous localization and mapping (SLAM) algorithms are necessary for safety-critical autonomous navigation. In the communication-constrained multi-agent setting, navigation systems increasingly use point-to-point range sensors…

机器人学 · 计算机科学 2025-05-15 Alexander Thoms , Alan Papalia , Jared Velasquez , David M. Rosen , Sriram Narasimhan

Computing optimal feedback controls for nonlinear systems generally requires solving Hamilton-Jacobi-Bellman (HJB) equations, which are notoriously difficult when the state dimension is large. Existing strategies for high-dimensional…

最优化与控制 · 数学 2021-04-09 Tenavi Nakamura-Zimmerer , Qi Gong , Wei Kang

Optimal control deals with optimization problems in which variables steer a dynamical system, and its outcome contributes to the objective function. Two classical approaches to solving these problems are Dynamic Programming and the…

最优化与控制 · 数学 2023-12-18 Alessandro Betti , Michele Casoni , Marco Gori , Simone Marullo , Stefano Melacci , Matteo Tiezzi

Inspired by REINFORCE, we introduce a novel receding-horizon algorithm for the Linear Quadratic Regulator (LQR) problem with unknown dynamics. Unlike prior methods, our algorithm avoids reliance on two-point gradient estimates while…

最优化与控制 · 数学 2025-10-07 Amirreza Neshaei Moghaddam , Alex Olshevsky , Bahman Gharesifard

This paper focuses on the discrete-time backward stochastic linear quadratic (BSLQ) optimal control problem with nonhomogeneous system terms and cost function cross terms. The terminal constraint of such systems distinguishes it from…

最优化与控制 · 数学 2026-04-14 Hu Ligui , Meng Qingxin , Tang Maoning

We present a dynamic subspace approach for efficiently approximating large-scale systems by learning time-continuous trajectories on the Grassmannian manifold. By parameterizing a low-dimensional basis as a geodesic path, the method allows…

数值分析 · 数学 2026-05-26 Jack DeChant , Rudy Geelen , Shane A. McQuarrie , Johann Guilleminot

In this work, the Parareal algorithm is applied to evolution problems that admit good low-rank approximations and for which the dynamical low-rank approximation (DLRA) can be used as time stepper. Many discrete integrators for DLRA have…

数值分析 · 数学 2022-09-14 Benjamin Carrel , Martin J. Gander , Bart Vandereycken

We study the problem of estimating Dynamic Discrete Choice (DDC) models, also known as offline Maximum Entropy-Regularized Inverse Reinforcement Learning (offline MaxEnt-IRL) in machine learning. The objective is to recover reward or $Q^*$…

机器学习 · 计算机科学 2026-05-06 Enoch H. Kang , Hema Yoganarasimhan , Lalit Jain

This paper develops and analyzes feedback-based online optimization methods to regulate the output of a linear time-invariant (LTI) dynamical system to the optimal solution of a time-varying convex optimization problem. The design of the…

最优化与控制 · 数学 2018-05-31 Marcello Colombino , Emiliano Dall'Anese , Andrey Bernstein

Deep Actor-Critic algorithms, which combine Actor-Critic with deep neural network (DNN), have been among the most prevalent reinforcement learning algorithms for decision-making problems in simulated environments. However, the existing deep…

机器学习 · 计算机科学 2024-09-19 Kexuan Wang , An Liu , Baishuo Lin

With the ever-growing size of pretrained models (PMs), fine-tuning them has become more expensive and resource-hungry. As a remedy, low-rank adapters (LoRA) keep the main pretrained weights of the model frozen and just introduce some…

计算与语言 · 计算机科学 2023-04-20 Mojtaba Valipour , Mehdi Rezagholizadeh , Ivan Kobyzev , Ali Ghodsi

Feedback control problems involving autonomous quadratic systems are prevalent, yet there are only a limited number of software tools available for approximating their solution due to the complexity of the problem. This paper represents a…

最优化与控制 · 数学 2019-10-09 Jeff Borggaard , Lizette Zietsman

This paper develops three high-order accurate discontinuous Galerkin (DG) methods for the one-dimensional (1D) and two-dimensional (2D) nonlinear Dirac (NLD) equations with a general scalar self-interaction. They are the Runge-Kutta DG…

数值分析 · 数学 2020-11-03 Shu-Cun Li , Huazhong Tang

This paper studies linear quadratic Gaussian robust mean field social control problems in the presence of multiplicative noise. We aim to compute asymptotic decentralized strategies without requiring full prior knowledge of agents'…

系统与控制 · 电气工程与系统科学 2025-09-16 Zhenhui Xu , Jiayu Chen , Bing-Chang Wang , Yuhu Wu , Tielong Shen

In robotics, contemporary strategies are learning-based, characterized by a complex black-box nature and a lack of interpretability, which may pose challenges in ensuring stability and safety. To address these issues, we propose integrating…

机器人学 · 计算机科学 2024-08-23 Mehdi Heydari Shahna , Seyed Adel Alizadeh Kolagar , Jouni Mattila

In this paper we propose a new computational method for designing optimal regulators for high-dimensional nonlinear systems. The proposed approach leverages physics-informed machine learning to solve high-dimensional Hamilton-Jacobi-Bellman…

最优化与控制 · 数学 2021-04-09 Tenavi Nakamura-Zimmerer , Qi Gong , Wei Kang

We introduce a novel data-driven method to mitigate the risk of cascading failures in delayed discrete-time Linear Time-Invariant (LTI) systems. Our approach involves formulating a distributionally robust finite-horizon optimal control…

最优化与控制 · 数学 2023-10-19 Guangyi Liu , Arash Amini , Vivek Pandey , Nader Motee

Low-Rank Adaptation (LoRA) enables efficient Continual Learning but often suffers from catastrophic forgetting due to destructive interference between tasks. Our analysis reveals that this degradation is primarily driven by antagonistic…

机器学习 · 计算机科学 2025-12-11 Yueer Zhou , Yichen Wu , Ying Wei