中文
相关论文

相关论文: Robust Performance Analysis of Source-Seeking Dyna…

200 篇论文

The nonlinear response coefficient, $\chi_{4,22}$, is a crucial observable for probing the dynamical properties of the quark-gluon plasma (QGP). While traditionally understood as a signature of medium response, recent studies suggest that…

核理论 · 物理学 2026-04-23 Zhi-Jie Yang , Hao-jie Xu , Jie Zhao , Hanlin Li

This paper introduces a novel data-driven approach to design a linear quadratic regulator (LQR) using a reinforcement learning (RL) algorithm that does not require a system model. The key contribution is to perform policy iteration (PI) by…

系统与控制 · 电气工程与系统科学 2023-11-20 Soroush Asri , Luis Rodrigues

We study the global convergence of generative adversarial imitation learning for linear quadratic regulators, which is posed as minimax optimization. To address the challenges arising from non-convex-concave geometry, we analyze the…

机器学习 · 计算机科学 2019-01-15 Qi Cai , Mingyi Hong , Yongxin Chen , Zhaoran Wang

The Quark--Meson--Coupling (QMC) model self-consistently relates the dynamics of the internal quark structure of a hadron to the relativistic mean fields arising in nuclear matter. It offers a natural explanation to some open questions in…

The sample inefficiency of reinforcement learning (RL) remains a significant challenge in robotics. RL requires large-scale simulation and can still cause long training times, slowing research and innovation. This issue is particularly…

机器人学 · 计算机科学 2026-01-16 Johannes Heeg , Yunlong Song , Davide Scaramuzza

Under dynamic traffic, service function chain (SFC) migration is considered as an effective way to improve resource utilization. However, the lack of future network information leads to non-optimal solutions, which motivates us to study…

网络与互联网体系结构 · 计算机科学 2019-11-14 Ruoyun Chen , Hancheng Lu , Yujiao Lu , Jinxue Liu

Learning in multi-agent systems is highly challenging due to several factors including the non-stationarity introduced by agents' interactions and the combinatorial nature of their state and action spaces. In particular, we consider the…

机器学习 · 统计学 2023-05-10 Barna Pásztor , Ilija Bogunovic , Andreas Krause

A scalable and resource-efficient quantum reinforcement learning framework is presented that eliminates the linear qubit-scaling barrier in multi-step quantum Markov decision processes (QMDPs). The proposed framework integrates a QMDP…

量子物理 · 物理学 2026-04-23 Thet Htar Su , Shaswot Shresthamali , Masaaki Kondo

We introduce an extended nonlinear Lugiato-Lefever equation (LLE) with the pseudo-stimulated-Raman-scattering (pseudo-SRS) cubic term, linear damping/gain, and spatial inhomogeneous (weakly or strongly localized) pump. The LLE is derived,…

斑图形成与孤子 · 物理学 2025-12-02 Evgeny M. Gromov , Boris A. Malomed

Traffic assignment methods are some of the key approaches used to model flow patterns that arise in transportation networks. Since static traffic assignment does not have a notion of time, it is not designed to represent temporal dynamics…

分布式、并行与集群计算 · 计算机科学 2021-05-20 Cy Chan , Anu Kuncheria , Bingyu Zhao , Theophile Cabannes , Alexander Keimer , Bin Wang , Alexandre Bayen , Jane Macfarlane

Reinforcement learning (RL) has shown great effectiveness in quadrotor control, enabling specialized policies to develop even human-champion-level performance in single-task scenarios. However, these specialized policies often struggle with…

机器人学 · 计算机科学 2024-12-18 Jiaxu Xing , Ismail Geles , Yunlong Song , Elie Aljalbout , Davide Scaramuzza

We propose a method for designing policies for convex stochastic control problems characterized by random linear dynamics and convex stage cost. We consider policies that employ quadratic approximate value functions as a substitute for the…

最优化与控制 · 数学 2023-11-10 Alan Yang , Stephen Boyd

Deep reinforcement learning has shown promise in various engineering applications, including vehicular traffic control. The non-stationary nature of traffic, especially in the lane-free environment with more degrees of freedom in vehicle…

机器人学 · 计算机科学 2024-06-24 Mehran Berahman , Majid Rostami-Shahrbabaki , Klaus Bogenberger

A data-driven framework is proposed for online estimation of quadrotor motor efficiency via residual minimization. The problem is formulated as a constrained nonlinear optimization that minimizes trajectory residuals between measured flight…

系统与控制 · 电气工程与系统科学 2026-03-09 Sheng-Wen Cheng , Teng-Hu Cheng

Traffic congestion, primarily driven by intersection queuing, significantly impacts urban living standards, safety, environmental quality, and economic efficiency. While Traffic Signal Control (TSC) systems hold potential for congestion…

机器学习 · 计算机科学 2026-01-14 Qiang Li , Jin Niu , Lina Yu

This paper presents a formulation of Lagrangian dynamics of constrained mechanical systems in terms of reduced quasi-velocities and quasi-forces that can be used for simulation, analysis, and control purposes. In this formulation, Cholesky…

计算物理 · 物理学 2021-08-17 Farhad Aghili

An iterative optimization approach that simultaneously minimizes the energy and optimizes the Lagrange multipliers enforcing desired constraints is presented. The method is tested on previously established benchmark systems and it is proved…

计算物理 · 物理学 2018-08-15 D. Kidd , A. S. Umar , K. Varga

In this paper, we present a computationally efficient trajectory optimizer that can exploit GPUs to jointly compute trajectories of tens of agents in under a second. At the heart of our optimizer is a novel reformulation of the non-convex…

机器人学 · 计算机科学 2020-11-10 Fatemeh Rastgar , Houman Masnavi , Jatan Shrestha , Karl Kruusamae , Alvo Aabloo , Arun Kumar Singh

We study the long-time dynamics of a bulk-surface convective Cahn--Hilliard system describing phase separation processes with bulk-surface interaction. The presence of convection terms leads to a non-autonomous dynamical system and prevents…

偏微分方程分析 · 数学 2026-03-12 Patrik Knopf , Andrea Poiatti , Jonas Stange , Sema Yayla

This article develops a strengthened convex quadratic convex (QC) relaxation of the AC Optimal Power Flow (AC-OPF) problem and presents an optimization-based bound-tightening (OBBT) algorithm to compute tight, feasible bounds on the voltage…

最优化与控制 · 数学 2019-01-30 Kaarthik Sundar , Harsha Nagarajan , Sidhant Misra , Mowen Lu , Carleton Coffrin , Russell Bent