English

STB-VMM: Swin Transformer Based Video Motion Magnification

Computer Vision and Pattern Recognition 2023-03-29 v2 Artificial Intelligence

Abstract

The goal of video motion magnification techniques is to magnify small motions in a video to reveal previously invisible or unseen movement. Its uses extend from bio-medical applications and deepfake detection to structural modal analysis and predictive maintenance. However, discerning small motion from noise is a complex task, especially when attempting to magnify very subtle, often sub-pixel movement. As a result, motion magnification techniques generally suffer from noisy and blurry outputs. This work presents a new state-of-the-art model based on the Swin Transformer, which offers better tolerance to noisy inputs as well as higher-quality outputs that exhibit less noise, blurriness, and artifacts than prior-art. Improvements in output image quality will enable more precise measurements for any application reliant on magnified video sequences, and may enable further development of video motion magnification techniques in new technical fields.

Keywords

Cite

@article{arxiv.2302.10001,
  title  = {STB-VMM: Swin Transformer Based Video Motion Magnification},
  author = {Ricard Lado-Roigé and Marco A. Pérez},
  journal= {arXiv preprint arXiv:2302.10001},
  year   = {2023}
}

Comments

Code available at: https://github.com/RLado/STB-VMM

R2 v1 2026-06-28T08:44:34.484Z