English
Related papers

Related papers: Learning Event-guided Exposure-agnostic Video Fram…

200 papers

Recent works have shown the ability of Implicit Neural Representations (INR) to carry meaningful representations of signal derivatives. In this work, we leverage this property to perform Video Frame Interpolation (VFI) by explicitly…

Computer Vision and Pattern Recognition · Computer Science 2022-06-23 Weihao Zhuang , Tristan Hascoet , Ryoichi Takashima , Tetsuya Takiguchi

Video deblurring is a challenging task due to the spatially variant blur caused by camera shake, object motions, and depth variations, etc. Existing methods usually estimate optical flow in the blurry video to align consecutive frames or…

Computer Vision and Pattern Recognition · Computer Science 2019-08-02 Shangchen Zhou , Jiawei Zhang , Jinshan Pan , Haozhe Xie , Wangmeng Zuo , Jimmy Ren

Event cameras are biologically-inspired sensors that gather the temporal evolution of the scene. They capture pixel-wise brightness variations and output a corresponding stream of asynchronous events. Despite having multiple advantages with…

Computer Vision and Pattern Recognition · Computer Science 2019-12-11 Stefano Pini , Guido Borghi , Roberto Vezzani

In autonomous driving, relying solely on frame-based cameras can lead to inaccuracies caused by factors like long exposure times, high-speed motion, and challenging lighting conditions. To address these issues, we introduce a bio-inspired…

Computer Vision and Pattern Recognition · Computer Science 2026-03-31 Hu Cao , Jiong Liu , Xingzhuo Yan , Rui Song , Yan Xia , Walter Zimmer , Guang Chen , Alois Knoll

Motion modeling is critical in flow-based Video Frame Interpolation (VFI). Existing paradigms either consider linear combinations of bidirectional flows or directly predict bilateral flows for given timestamps without exploring favorable…

Computer Vision and Pattern Recognition · Computer Science 2025-02-11 Zujin Guo , Wei Li , Chen Change Loy

Audio-visual emotion recognition (AVER) methods typically fuse utterance-level features, and even frame-level attention models seldom address the frame-rate mismatch across modalities. In this paper, we propose a Transformer-based framework…

Multimedia · Computer Science 2026-03-13 Inyong Koo , yeeun Seong , Minseok Son , Jaehyuk Jang , Changick Kim

Clear imaging under hazy conditions is a critical task. Prior-based and neural methods have improved results. However, they operate on RGB frames, which suffer from limited dynamic range. Therefore, dehazing remains ill-posed and can erase…

Computer Vision and Pattern Recognition · Computer Science 2025-11-18 Ling Wang , Yunfan Lu , Wenzong Ma , Huizai Yao , Pengteng Li , Hui Xiong

Under extreme low-light conditions, frame-based cameras suffer from severe detail loss due to limited dynamic range. Recent studies have introduced event cameras for event-guided low-light image enhancement. However, existing approaches…

Computer Vision and Pattern Recognition · Computer Science 2025-11-11 Zhanwen Liu , Huanna Song , Yang Wang , Nan Yang , Weiping Ding , Yisheng An

Traditional visual place recognition (VPR), usually using standard cameras, is easy to fail due to glare or high-speed motion. By contrast, event cameras have the advantages of low latency, high temporal resolution, and high dynamic range,…

Computer Vision and Pattern Recognition · Computer Science 2022-11-24 Kuanxu Hou , Delei Kong , Junjie Jiang , Hao Zhuang , Xinjie Huang , Zheng Fang

Event cameras have shown promise in vision applications like optical flow estimation and stereo matching, with many specialized architectures leveraging the asynchronous and sparse nature of event data. However, existing works only focus…

Computer Vision and Pattern Recognition · Computer Science 2024-11-25 Pengjie Zhang , Lin Zhu , Xiao Wang , Lizhi Wang , Wanxuan Lu , Hua Huang

Text-conditioned diffusion models have emerged as powerful tools for high-quality video generation. However, enabling Interactive Video Generation (IVG), where users control motion elements such as object trajectory, remains challenging.…

Computer Vision and Pattern Recognition · Computer Science 2025-06-02 Ishaan Rawal , Suryansh Kumar

Turbulence mitigation (TM) aims to remove the stochastic distortions and blurs introduced by atmospheric turbulence into frame cameras. Existing state-of-the-art deep-learning TM methods extract turbulence cues from multiple degraded frames…

Computer Vision and Pattern Recognition · Computer Science 2025-09-05 Huanan Li , Rui Fan , Juntao Guan , Weidong Hao , Lai Rui , Tong Wu , Yikai Wang , Lin Gu

Video prediction is an extrapolation task that predicts future frames given past frames, and video frame interpolation is an interpolation task that estimates intermediate frames between two frames. We have witnessed the tremendous…

Computer Vision and Pattern Recognition · Computer Science 2022-06-28 Yue Wu , Qiang Wen , Qifeng Chen

This paper proposes a novel deep learning-based video object matting method that can achieve temporally coherent matting results. Its key component is an attention-based temporal aggregation module that maximizes image matting networks'…

Computer Vision and Pattern Recognition · Computer Science 2021-07-30 Yunke Zhang , Chi Wang , Miaomiao Cui , Peiran Ren , Xuansong Xie , Xian-sheng Hua , Hujun Bao , Qixing Huang , Weiwei Xu

Super-resolution (SR) has been widely used to convert low-resolution legacy videos to high-resolution (HR) ones, to suit the increasing resolution of displays (e.g. UHD TVs). However, it becomes easier for humans to notice motion artifacts…

Computer Vision and Pattern Recognition · Computer Science 2022-02-08 Soo Ye Kim , Jihyong Oh , Munchurl Kim

Event cameras offer significant advantages, including a wide dynamic range, high temporal resolution, and immunity to motion blur, making them highly promising for addressing challenging visual conditions. Extracting and utilizing effective…

Computer Vision and Pattern Recognition · Computer Science 2025-08-05 Xin Dong , Yiwei Zhang , Yangjie Cui , Jinwu Xiang , Daochun Li , Zhan Tu

Low-light image enhancement (LLIE) aims to improve the visibility of images captured in poorly lit environments. Prevalent event-based solutions primarily utilize events triggered by motion, i.e., ''motion events'' to strengthen only the…

Computer Vision and Pattern Recognition · Computer Science 2025-04-15 Lei Sun , Yuhan Bao , Jiajun Zhai , Jingyun Liang , Yulun Zhang , Kaiwei Wang , Danda Pani Paudel , Luc Van Gool

Generating intermediate video content of varying lengths based on given first and last frames, along with text prompt information, offers significant research and application potential. However, traditional frame interpolation tasks…

Computer Vision and Pattern Recognition · Computer Science 2025-07-08 Yijia Hong , Jiangning Zhang , Ran Yi , Yuji Wang , Weijian Cao , Xiaobin Hu , Zhucun Xue , Yabiao Wang , Chengjie Wang , Lizhuang Ma

Event cameras generate asynchronous signals in response to pixel-level brightness changes, offering a sensing paradigm with theoretically microsecond-scale latency that can significantly enhance the performance of multi-sensor systems.…

Robotics · Computer Science 2025-08-19 Jiayao Mai , Xiuyuan Lu , Kuan Dai , Shaojie Shen , Yi Zhou

Recording fast motion in a high FPS (frame-per-second) requires expensive high-speed cameras. As an alternative, interpolating low-FPS videos from commodity cameras has attracted significant attention. If only low-FPS videos are available,…

Computer Vision and Pattern Recognition · Computer Science 2022-03-29 Weihua He , Kaichao You , Zhendong Qiao , Xu Jia , Ziyang Zhang , Wenhui Wang , Huchuan Lu , Yaoyuan Wang , Jianxing Liao