English
Related papers

Related papers: Feedback Linearization Control for Systems with Mi…

200 papers

We study a new paradigm for sequential decision making, called offline policy learning from observations (PLfO). Offline PLfO aims to learn policies using datasets with substandard qualities: 1) only a subset of trajectories is labeled with…

Machine Learning · Computer Science 2023-08-08 Anqi Li , Byron Boots , Ching-An Cheng

Flow matching has recently emerged as a powerful alternative to diffusion models, providing a continuous-time formulation for generative modeling and representation learning. Yet, we show that this framework suffers from a fundamental…

Machine Learning · Computer Science 2025-09-26 Weili Zeng , Yichao Yan

Learning from Observations (LfO) is a practical reinforcement learning scenario from which many applications can benefit through the reuse of incomplete resources. Compared to conventional imitation learning (IL), LfO is more challenging…

Machine Learning · Computer Science 2021-03-01 Zhuangdi Zhu , Kaixiang Lin , Bo Dai , Jiayu Zhou

For linear control systems, the usual state feedback stabilizability has two components: one is a continuous observation mode (i.e., to observe solutions continuously in time), and the other is a class of feedback laws (which is usually the…

Optimization and Control · Mathematics 2022-08-29 Hanbing Liu , Gengsheng Wang , Huaiqiang Yu

This work investigates how disturbance-aware, robustness-embedded reference trajectories translate into driving performance when executed by professional drivers in a dynamic simulator. Three planned reference trajectories are compared…

This paper develops learning-enabled safe controllers for linear systems subject to system uncertainties and bounded disturbances. Given the disturbance zonotope, the databased closed-loop dynamics (CLDs) are first characterized using a…

Systems and Control · Electrical Eng. & Systems 2025-10-22 Amir Modares , Niyousha Ghiasi , Bahare Kiumarsi , Hamidreza Modares

In this paper, we first study the leader-following output synchronization problem for a class of uncertain nonlinear multi-agent systems over jointly connected switching networks. Our approach integrates the output-based adaptive…

Optimization and Control · Mathematics 2024-11-04 Jie Huang

Hallucination occurs when large language models exhibit behavior that deviates from the boundaries of their knowledge during response generation. To address this critical issue, previous learning-based methods attempt to finetune models but…

Computation and Language · Computer Science 2025-05-27 Xueru Wen , Jie Lou , Xinyu Lu , Ji Yuqiu , Xinyan Guan , Yaojie Lu , Hongyu Lin , Ben He , Xianpei Han , Debing Zhang , Le Sun

This paper addresses the stabilization of linear systems with multiple time-varying input delays. In scenarios where neither the exact delays information nor their bound is known, we propose a class of linear time-varying state feedback…

Dynamical Systems · Mathematics 2025-05-01 Bin Zhou , Kai Zhang

With the rapid advances in Large Language Models (LLMs), aligning LLMs with human preferences become increasingly important. Although Reinforcement Learning with Human Feedback (RLHF) proves effective, it is complicated and highly…

Computation and Language · Computer Science 2024-10-31 Shiqi Wang , Zhengze Zhang , Rui Zhao , Fei Tan , Cam Tu Nguyen

Finding a control Lyapunov function (CLF) in a dynamical system with a controller is an effective way to guarantee stability, which is a crucial issue in safety-concerned applications. Recently, deep learning models representing CLFs have…

Machine Learning · Computer Science 2025-11-04 Yupu Lu , Shijie Lin , Hao Xu , Zeqing Zhang , Jia Pan

We consider the problem of controlling a possibly unknown linear dynamical system with adversarial perturbations, adversarially chosen convex loss functions, and partially observed states, known as non-stochastic control. We introduce a…

Machine Learning · Computer Science 2020-06-26 Max Simchowitz , Karan Singh , Elad Hazan

Nonlinear friction has long been, and continues to be, one of the major challenges for precision motion control systems. A linear asymptotic observer of the motion state variables with nonlinear friction uses a dedicated state-space…

Systems and Control · Electrical Eng. & Systems 2025-12-22 Michael Ruderman

We propose an approach to design a Model Predictive Controller (MPC) for constrained Linear Time Invariant systems performing an iterative task. The system is subject to an additive disturbance, and the goal is to learn to satisfy state and…

Systems and Control · Electrical Eng. & Systems 2023-06-13 Monimoy Bujarbaruah , Akhil Shetty , Kameshwar Poolla , Francesco Borrelli

To improve detection robustness in adverse conditions (e.g., haze and low light), image restoration is commonly applied as a pre-processing step to enhance image quality for the detector. However, the functional mismatch between restoration…

Computer Vision and Pattern Recognition · Computer Science 2026-03-10 Qing Zhao , Weijian Deng , Pengxu Wei , ZiYi Dong , Hannan Lu , Xiangyang Ji , Liang Lin

This paper develops systematically the output feedback exponential stabilization for a one-dimensional unstable/anti-stable wave equation where the control boundary suffers from both internal nonlinear uncertainty and external disturbance.…

Optimization and Control · Mathematics 2017-06-08 Hua-Cheng Zhou , George Weiss

This paper considers the fixed-time control problem of a multi-agent system composed of a class of Euler-Lagrange dynamics with parametric uncertainty and a dynamic leader under a directed communication network. A distributed fixed-time…

Optimization and Control · Mathematics 2022-02-17 Yi Dong , Zhiyong Chen

This project develops a self correcting framework for large language models (LLMs) that detects and mitigates hallucinations during multi-step reasoning. Rather than relying solely on final answer correctness, our approach leverages fine…

Artificial Intelligence · Computer Science 2025-11-21 Chelsea Zou , Yiheng Yao , Basant Khalil

Group-Relative Policy Optimization (GRPO) is a key technique for training large reasoning models, yet it suffers from a critical vulnerability: the \emph{Think-Answer Mismatch}, where noisy reward signals corrupt the learning process. This…

Machine Learning · Computer Science 2025-08-11 Si Shen , Peijun Shen , Wenhua Zhao , Danhao Zhu

Temporal logic inference is the process of extracting formal descriptions of system behaviors from data in the form of temporal logic formulas. The existing temporal logic inference methods mostly neglect uncertainties in the data, which…

Artificial Intelligence · Computer Science 2021-06-01 Nasim Baharisangari , Jean-Raphaël Gaglione , Daniel Neider , Ufuk Topcu , Zhe Xu