English
Related papers

Related papers: Learning-based Attitude Estimation with Noisy Meas…

200 papers

Training Large Language Models (LLMs) for reasoning tasks is increasingly driven by Reinforcement Learning with Verifiable Rewards (RLVR), where Proximal Policy Optimization (PPO) provides a principled framework for stable policy updates.…

Machine Learning · Computer Science 2026-01-13 Xue Gong , Qi Yi , Ziyuan Nan , Guanhua Huang , Kejiao Li , Yuhao Jiang , Ruibin Xiong , Zenan Xu , Jiaming Guo , Shaohui Peng , Bo Zhou

This paper investigates the reinforcement learning (RL) based disturbance rejection control for uncertain nonlinear systems having non-simple nominal models. An extended state observer (ESO) is first designed to estimate the system state…

Dynamical Systems · Mathematics 2020-11-25 Maopeng Ran , Juncheng Li , Lihua Xie

This work presents a novel adaptive framework for simultaneously estimating spacecraft attitude and sensor misalignment. Uncorrected star tracker misalignment can introduce significant pointing errors that compromise mission objectives in…

Systems and Control · Electrical Eng. & Systems 2025-07-29 Ridma Ganganath , Simone Servadio , David Lee

Recency bias is a useful inductive prior for sequential modeling: it emphasizes nearby observations and can still allow longer-range dependencies. Standard Transformer attention lacks this property, relying on all-to-all interactions that…

Machine Learning · Computer Science 2026-04-23 Kareem Hegazy , Michael W. Mahoney , N. Benjamin Erichson

Robotic manipulation in unstructured environments requires the generation of robust and long-horizon trajectory-level policy with conditions of perceptual observations and benefits from the advantages of SE(3)-equivariant diffusion models…

Robotics · Computer Science 2025-09-30 Zhitao Wang , Yanke Wang , Jiangtao Wen , Roberto Horowitz , Yuxing Han

In recent times, Reinforcement learning (RL) has been widely applied to many challenging tasks. However, in order to perform well, it requires access to a good reward function which is often sparse or manually engineered with scope for…

Machine Learning · Computer Science 2024-09-25 Yuxuan Li , Srijita Das , Matthew E. Taylor

We use Reinforcement Meta-Learning to optimize an adaptive integrated guidance, navigation, and control system suitable for exoatmospheric interception of a maneuvering target. The system maps observations consisting of strapdown seeker…

Systems and Control · Electrical Eng. & Systems 2021-12-14 Brian Gaudet , Roberto Furfaro , Richard Linares , Andrea Scorsoglio

When deploying machine learning estimators in science and engineering (SAE) domains, it is critical to avoid failed estimations that can have disastrous consequences, e.g., in aero engine design. This work focuses on detecting and…

Machine Learning · Computer Science 2023-10-31 Ruiyuan Kang , Tingting Mu , Panos Liatsis , Dimitrios C. Kyritsis

Risk-averse Constrained Reinforcement Learning (RaCRL) aims to learn policies that minimise the likelihood of rare and catastrophic constraint violations caused by an environment's inherent randomness. In general, risk-aversion leads to…

Machine Learning · Computer Science 2025-08-28 James McCarthy , Radu Marinescu , Elizabeth Daly , Ivana Dusparic

Powered by deep representation learning, reinforcement learning (RL) provides an end-to-end learning framework capable of solving self-driving (SD) tasks without manual designs. However, time-varying nonstationary environments cause…

Robotics · Computer Science 2023-03-09 Tao Li , Haozhe Lei , Quanyan Zhu

Intuitive human-machine interfaces may be developed using pattern classification to estimate executed human motions from electromyogram (EMG) signals generated during muscle contraction. The continual use of EMG-based interfaces gradually…

Signal Processing · Electrical Eng. & Systems 2023-10-03 Seitaro Yoneda , Akira Furui

In this note an intrinsic version of the Cram\'er-Rao bound on estimation accuracy is established on the Special Orthogonal group $SO(3)$. It is intrinsic in the sense that it does not rely on a specific choice of coordinates on $SO(3)$:…

Optimization and Control · Mathematics 2015-10-13 Silvère Bonnabel , Axel Barrau

This paper introduces two novel nonlinear stochastic attitude estimators developed on the Special Orthogonal Group \mathbb{SO}\left(3\right) with the tracking error of the normalized Euclidean distance meeting predefined transient and…

Systems and Control · Electrical Eng. & Systems 2020-06-23 Hashim A. Hashim

Algorithmic recourse recommends a cost-efficient action to a subject to reverse an unfavorable machine learning classification decision. Most existing methods in the literature generate recourse under the assumption of complete knowledge…

Machine Learning · Computer Science 2024-02-26 Duy Nguyen , Bao Nguyen , Viet Anh Nguyen

We propose CARE (Collision Avoidance via Repulsive Estimation) to improve the robustness of learning-based visual navigation methods. Recently, visual navigation models, particularly foundation models, have demonstrated promising…

Robotics · Computer Science 2025-08-11 Joonkyung Kim , Joonyeol Sim , Woojun Kim , Katia Sycara , Changjoo Nam

Reinforcement Learning (RL) has the promise of providing data-driven support for decision-making in a wide range of problems in healthcare, education, business, and other domains. Classical RL methods focus on the mean of the total return…

Machine Learning · Computer Science 2022-02-02 Elynn Y. Chen , Rui Song , Michael I. Jordan

Recent advancements in computer vision have accelerated the development of autonomous driving. Despite these advancements, training machines to drive in a way that aligns with human expectations remains a significant challenge. Human…

Computer Vision and Pattern Recognition · Computer Science 2026-04-01 Zhuoli Zhuang , Yu-Cheng Chang , Yu-Kai Wang , Thomas Do , Chin-Teng Lin

We demonstrate a new deep learning autoencoder network, trained by a nonnegativity constraint algorithm (NCAE), that learns features which show part-based representation of data. The learning algorithm is based on constraining negative…

Machine Learning · Computer Science 2016-01-13 Ehsan Hosseini-Asl , Jacek M. Zurada , Olfa Nasraoui

High fidelity behavior prediction of intelligent agents is critical in many applications. However, the prediction model trained on the training set may not generalize to the testing set due to domain shift and time variance. The challenge…

Machine Learning · Computer Science 2020-04-29 Abulikemu Abuduweili , Changliu Liu

In this work, we propose a novel generative method to identify the causal impact and apply it to prediction tasks. We conduct causal impact analysis using interventional and counterfactual perspectives. First, applying interventions, we…

Machine Learning · Computer Science 2025-09-03 Soma Bandyopadhyay , Sudeshna Sarkar
‹ Prev 1 4 5 6 7 8 10 Next ›