English

WarmPrior: Straightening Flow-Matching Policies with Temporal Priors

Machine Learning 2026-05-15 v1 Artificial Intelligence Robotics

Abstract

Generative policies based on diffusion and flow matching have become a dominant paradigm for visuomotor robotic control. We show that replacing the standard Gaussian source distribution with WarmPrior, a simple temporally grounded prior constructed from readily available recent action history, consistently improves success rates on robotic manipulation tasks. We trace this gain to markedly straighter probability paths, echoing the effect of optimal-transport couplings in Rectified Flow. Beyond standard behavior cloning, WarmPrior also reshapes the exploration distribution in prior-space reinforcement learning, improving both sample efficiency and final performance. Collectively, these results identify the source distribution as an important and underexplored design axis in generative robot control.

Keywords

Cite

@article{arxiv.2605.13959,
  title  = {WarmPrior: Straightening Flow-Matching Policies with Temporal Priors},
  author = {Sinjae Kang and Chanyoung Kim and Kaixin Wang and Li Zhao and Kimin Lee},
  journal= {arXiv preprint arXiv:2605.13959},
  year   = {2026}
}