English

Residual Reinforcement Learning for Robot Teleoperation under Stochastic Delays

Robotics 2026-05-18 v1 Artificial Intelligence

Abstract

Stochastic communication delays in teleoperation introduce signal discontinuities that undermine control stability and degrade control performance. Consequently, the conventional reinforcement learning (RL) methods struggle with the delayed observations due to the delay-induced observations, leading to high-frequency chattering. To address this, we propose a hybrid control framework, delay-resilient RL, integrating a state estimator utilizing Long Short-Term Memory (LSTM) with a residual RL policy, which is resilient to stochastic delays. The LSTM reconstructs smooth, continuous state estimates from delayed observations, enabling the RL agent to learn a residual torque compensation policy that balances tracking accuracy with velocity smoothness. Experimental validation on Franka Panda robots demonstrates that our approach significantly outperforms the state-of-the-art baselines, ensuring robust and stable teleoperation even under high-variance stochastic delays.

Keywords

Cite

@article{arxiv.2605.15480,
  title  = {Residual Reinforcement Learning for Robot Teleoperation under Stochastic Delays},
  author = {Kaize Deng and Zewen Yang},
  journal= {arXiv preprint arXiv:2605.15480},
  year   = {2026}
}

Comments

Accepted at 23rd IFAC World Congress 2026

R2 v1 2026-07-22T07:13:28.668Z