中文
相关论文

相关论文: Preview Reference Governors: A Constraint Manageme…

200 篇论文

Reinforcement learning (RL) has become a promising paradigm for optimizing Retrieval-Augmented Generation (RAG) in complex reasoning tasks. However, traditional outcome-based RL approaches often suffer from reward sparsity and inefficient…

人工智能 · 计算机科学 2026-01-30 Zhao Wang , Ziliang Zhao , Zhicheng Dou

Boom cranes are among the most used cranes to lift heavy loads. Although fairly simple mechanically, from the control viewpoint this kind of crane is a nonlinear underactuated system which presents several challenges, especially when…

系统与控制 · 电气工程与系统科学 2021-03-04 Michele Ambrosino , Arnaud Dawans , Emanuele Garone

Explicit reference governor (ERG) is an add-on unit that provides constraint handling capability to pre-stabilized systems. The main idea behind ERG is to manipulate the derivative of the applied reference in continuous time such that the…

系统与控制 · 电气工程与系统科学 2024-06-03 Mu'taz A. Momani , Mehdi Hosseinzadeh

In this paper, we consider the problem of learning safe policies for probabilistic-constrained reinforcement learning (RL). Specifically, a safe policy or controller is one that, with high probability, maintains the trajectory of the agent…

机器学习 · 计算机科学 2024-03-14 Weiqin Chen , Dharmashankar Subramanian , Santiago Paternain

This paper focuses on a passivity-based distributed reference governor (RG) applied to a pre-stabilized mobile robotic network. The novelty of this paper lies in the method used to solve the RG problem, where a passivity-based distributed…

多智能体系统 · 计算机科学 2017-03-21 Tam Nguyen , Takeshi Hatanaka , Mamoru Doi , Emanuele Garone , Masayuki Fujita

Safe navigation around obstacles is a fundamental challenge for highly dynamic robots. The state-of-the-art approach for adapting simple reference path planners to complex robot dynamics using trajectory optimization and tracking control is…

机器人学 · 计算机科学 2022-02-28 Aykut İşleyen , Nathan van de Wouw , Ömür Arslan

Reinforcement learning with human feedback for aligning large language models (LLMs) trains a reward model typically using ranking loss with comparison pairs.However, the training procedure suffers from an inherent problem: the uncontrolled…

计算与语言 · 计算机科学 2024-09-19 Hang Zhou , Chenglong Wang , Yimin Hu , Tong Xiao , Chunliang Zhang , Jingbo Zhu

Continual learning aims to avoid catastrophic forgetting and effectively leverage learned experiences to master new knowledge. Existing gradient projection approaches impose hard constraints on the optimization space for new tasks to…

机器学习 · 计算机科学 2023-01-31 Zeyuan Yang , Zonghan Yang , Peng Li , Yang Liu

This paper introduces an explicit reference governor-based control scheme tailored for addressing the velocity-free spacecraft attitude maneuver problem. This problem is subject to specific constraints, namely the pointing constraint,…

地球与行星天体物理 · 物理学 2024-06-25 Qingqing Dang , Wenbo Libo , Haichao Gui

The paper addresses the problem of vehicle rollover avoidance using reference governors applied to modify the driver steering input in vehicles with an active steering system. Several reference governor designs are presented and tested with…

系统与控制 · 计算机科学 2016-08-09 Ricardo Bencatel , Anouck Girard , Ilya Kolmanovsky

Reinforcement learning is essential for neural architecture search and hyperparameter optimization, but the conventional approaches impede widespread use due to prohibitive time and computational costs. Inspired by DeepSeek-V3 multi-token…

机器学习 · 计算机科学 2025-06-19 Zheng Li , Jerry Cheng , Huanying Helen Gu

Generative Recommenders (GRs), exemplified by the Hierarchical Sequential Transduction Unit (HSTU), have emerged as a powerful paradigm for modeling long user interaction sequences. However, we observe that their "flat-sequence" assumption…

信息检索 · 计算机科学 2026-03-03 Zerui Chen , Heng Chang , Tianying Liu , Chuantian Zhou , Yi Cao , Jiandong Ding , Ming Liu , Bing Qin

In this work, we propose a control scheme for linear systems subject to pointwise in time state and input constraints that aims to minimize time-varying and a priori unknown cost functions. The proposed controller is based on online convex…

系统与控制 · 电气工程与系统科学 2024-12-02 Marko Nonhoff , Johannes Köhler , Matthias A. Müller

Mechanical search (MS) in cluttered environments remains a significant challenge for autonomous manipulators, requiring long-horizon planning and robust state estimation under occlusions and partial observability. In this work, we introduce…

机器人学 · 计算机科学 2025-06-17 Yiting Zhang , Shichen Li , Elena Shrestha

Reinforcement Learning (RL) has emerged as a powerful framework for sequential decision-making in dynamic environments, particularly when system parameters are unknown. This paper investigates RL-based control for entropy-regularized…

系统与控制 · 电气工程与系统科学 2025-12-02 Gabriel Diaz , Lucky Li , Wenhao Zhang

A predictive triggering (PT) framework for the distributed control of resource constrained multi-agent systems is proposed. By predicting future communication demands and deriving a probabilistic priority measure, the PT framework is able…

系统与控制 · 电气工程与系统科学 2019-07-30 José Mario Mastrangelo , Dominik Baumann , Sebastian Trimpe

Hierarchical reinforcement learning (HRL) decomposes the policy into a manager and a worker, enabling long-horizon planning but introducing a performance gap on tasks requiring agility. We identify a root cause: in subgoal-based HRL, the…

人工智能 · 计算机科学 2026-04-22 Shashank Sharma , Janina Hoffmann , Vinay Namboodiri

Specifying rewards for reinforcement learned (RL) agents is challenging. Preference-based RL (PbRL) mitigates these challenges by inferring a reward from feedback over sets of trajectories. However, the effectiveness of PbRL is limited by…

机器学习 · 计算机科学 2022-10-20 Mudit Verma , Katherine Metcalf

Recent generative models based on score matching and flow matching have significantly advanced generation tasks, but their potential in discriminative tasks remains underexplored. Previous approaches, such as generative classifiers, have…

计算机视觉与模式识别 · 计算机科学 2025-08-14 Rongkun Xue , Jinouwen Zhang , Yazhe Niu , Dazhong Shen , Bingqi Ma , Yu Liu , Jing Yang

The goal of a recommendation system is to model the relevance between each user and each item through the user-item interaction history, so that maximize the positive samples score and minimize negative samples. Currently, two popular loss…

信息检索 · 计算机科学 2022-07-08 Chun Yang , Shicai Fan