中文
相关论文

相关论文: Line of Sight Curvature for Missile Guidance using…

200 篇论文

We present a novel guidance law that uses observations consisting solely of seeker line of sight angle measurements and their rate of change. The policy is optimized using reinforcement meta-learning and demonstrated in a simulated terminal…

系统与控制 · 电气工程与系统科学 2024-09-23 Brian Gaudet , Roberto Furfaro , Richard Linares

We apply a reinforcement meta-learning framework to optimize an integrated and adaptive guidance and flight control system for an air-to-air missile. The system is implemented as a policy that maps navigation system outputs directly to…

系统与控制 · 电气工程与系统科学 2022-05-05 Brian Gaudet , Roberto Furfaro

We use Reinforcement Meta-Learning to optimize an adaptive integrated guidance, navigation, and control system suitable for exoatmospheric interception of a maneuvering target. The system maps observations consisting of strapdown seeker…

系统与控制 · 电气工程与系统科学 2021-12-14 Brian Gaudet , Roberto Furfaro , Richard Linares , Andrea Scorsoglio

Current practice for asteroid close proximity maneuvers requires extremely accurate characterization of the environmental dynamics and precise spacecraft positioning prior to the maneuver. This creates a delay of several months between the…

系统与控制 · 电气工程与系统科学 2020-09-16 Brian Gaudet , Richard Linares , Roberto Furfaro

An adaptive guidance system suitable for the terminal phase trajectory of a hypersonic strike weapon is optimized using reinforcement meta learning. The guidance system maps observations directly to commanded bank angle, angle of attack,…

系统与控制 · 电气工程与系统科学 2021-10-19 Brian Gaudet , Roberto Furfaro

We use Reinforcement Meta Learning to optimize an adaptive guidance system suitable for the approach phase of a gliding hypersonic vehicle. Adaptability is achieved by optimizing over a range of off-nominal flight conditions including…

机器人学 · 计算机科学 2021-08-02 Brian Gaudet , Kris Drozd , Ryan Meltzer , Roberto Furfaro

This paper proposes a novel adaptive guidance system developed using reinforcement meta-learning with a recurrent policy and value function approximator. The use of recurrent network layers allows the deployed policy to adapt real time to…

系统与控制 · 计算机科学 2020-02-19 Brian Gaudet , Richard Linares , Roberto Furfaro

This paper proposes a novel adaptive guidance system developed using reinforcement meta-learning with a recurrent policy and value function approximator. The use of recurrent network layers allows the deployed policy to adapt real time to…

系统与控制 · 电气工程与系统科学 2024-12-20 Brian Gaudet , Richard Linares

Autonomy is a key challenge for future space exploration endeavours. Deep Reinforcement Learning holds the promises for developing agents able to learn complex behaviours simply by interacting with their environment. This paper investigates…

机器人学 · 计算机科学 2025-05-02 Matteo El Hariry , Andrea Cini , Giacomo Mellone , Alessandro Balossino

Guided policy search algorithms can be used to optimize complex nonlinear policies, such as deep neural networks, without directly computing policy gradients in the high-dimensional parameter space. Instead, these methods use supervised…

机器学习 · 计算机科学 2016-07-18 William Montgomery , Sergey Levine

We propose meta-curvature (MC), a framework to learn curvature information for better generalization and fast model adaptation. MC expands on the model-agnostic meta-learner (MAML) by learning to transform the gradients in the inner…

机器学习 · 计算机科学 2020-01-10 Eunbyung Park , Junier B. Oliva

Handling orientations of robots and objects is a crucial aspect of many applications. Yet, ever so often, there is a lack of mathematical correctness when dealing with orientations, especially in learning pipelines involving, for example,…

机器人学 · 计算机科学 2025-10-13 Martin Schuck , Jan Brüdigam , Sandra Hirche , Angela Schoellig

This paper addresses a boosting method for mapping functionality of neural networks in visual recognition such as image classification and face recognition. We present reversible learning for generating and learning latent features using…

机器学习 · 计算机科学 2019-10-22 Jongmin Yu

This paper considers meta-cognitive radars in an adversarial setting. A cognitive radar optimally adapts its waveform (response) in response to maneuvers (probes) of a possibly adversarial moving target. A meta-cognitive radar is aware of…

信号处理 · 电气工程与系统科学 2022-05-05 Kunal Pattanayak , Vikram Krishnamurthy , Christopher Berry

In this paper, we present a novel guidance scheme based on model-based deep reinforcement learning (RL) technique. With model-based deep RL method, a deep neural network is trained as a predictive model of guidance dynamics which is…

机器人学 · 计算机科学 2019-04-16 Chen Liang , Weihong Wang , Zhenghua Liu , Chao Lai , Benchun Zhou

Imitation learning (IL) can train computationally-efficient sensorimotor policies from a resource-intensive Model Predictive Controller (MPC), but it often requires many samples, leading to long training times or limited robustness. To…

机器人学 · 计算机科学 2024-02-27 Andrea Tagliabue , Jonathan P. How

This paper investigates the problem of impact-time-control and proposes a learning-based computational guidance algorithm to solve this problem. The proposed guidance algorithm is developed based on a general prediction-correction concept:…

机器学习 · 计算机科学 2021-05-31 Zichao Liu , Jiang Wang , Shaoming He , Hyo-Sang Shin , Antonios Tsourdos

Reinforcement learning has been applied to human movement through physiologically-based biomechanical models to add insights into the neural control of these movements; it is also useful in the design of prosthetics and robotics. In this…

机器学习 · 计算机科学 2020-08-13 Julie Iskander , Mohammed Hossny

We propose to meta-learn an a self-supervised patient trajectory forecast learning rule by meta-training on a meta-objective that directly optimizes the utility of the patient representation over the subsequent clinical outcome prediction.…

机器学习 · 计算机科学 2024-07-30 Yuan Xue , Nan Du , Anne Mottram , Martin Seneviratne , Andrew M. Dai

Reinforcement learning has proven its power on various occasions. However, its performance is not always guaranteed when system dynamics change. Instead, it largely relies on users' empirical experience. For reinforcement learning…

机器学习 · 计算机科学 2026-05-05 Jingyi Liu , Jian Guo , Eberhard Gill
‹ 上一页 1 2 3 10 下一页 ›