中文
相关论文

相关论文: Robust Online Matrix Factorization for Dynamic Bac…

200 篇论文

We introduce a robust, real-time, high-resolution human video matting method that achieves new state-of-the-art performance. Our method is much lighter than previous approaches and can process 4K at 76 FPS and HD at 104 FPS on an Nvidia GTX…

计算机视觉与模式识别 · 计算机科学 2021-08-27 Shanchuan Lin , Linjie Yang , Imran Saleemi , Soumyadip Sengupta

We present a novel camera path optimization framework for the task of online video stabilization. Typically, a stabilization pipeline consists of three steps: motion estimating, path smoothing, and novel view rendering. Most previous…

计算机视觉与模式识别 · 计算机科学 2023-08-16 Zhuofan Zhang , Zhen Liu , Ping Tan , Bing Zeng , Shuaicheng Liu

3D Gaussian Splatting (3DGS) has emerged as a powerful representation due to its efficiency and high-fidelity rendering. 3DGS training requires a known camera pose for each input view, typically obtained by Structure-from-Motion (SfM)…

计算机视觉与模式识别 · 计算机科学 2025-05-27 Zhen-Hui Dong , Sheng Ye , Yu-Hui Wen , Nannan Li , Yong-Jin Liu

Extracting accurate foreground objects from a scene is an essential step for many video applications. Traditional background subtraction algorithms can generate coarse estimates, but generating high quality masks requires professional…

计算机视觉与模式识别 · 计算机科学 2020-09-29 Xiran Wang , Jason Juang , Stanley H. Chan

Video object segmentation, aiming to segment the foreground objects given the annotation of the first frame, has been attracting increasing attentions. Many state-of-the-art approaches have achieved great performance by relying on online…

计算机视觉与模式识别 · 计算机科学 2021-07-12 Siyue Yu , Jimin Xiao , BingFeng Zhang , Eng Gee Lim

We present MoBGS, a novel motion deblurring 3D Gaussian Splatting (3DGS) framework capable of reconstructing sharp and high-quality novel spatio-temporal views from blurry monocular videos in an end-to-end manner. Existing dynamic novel…

计算机视觉与模式识别 · 计算机科学 2025-12-04 Minh-Quan Viet Bui , Jongmin Park , Juan Luis Gonzalez Bello , Jaeho Moon , Jihyong Oh , Munchurl Kim

This paper presents a unified framework that allows high-quality dynamic Gaussian Splatting from both defocused and motion-blurred monocular videos. Due to the significant difference between the formation processes of defocus blur and…

计算机视觉与模式识别 · 计算机科学 2025-11-03 Xuankai Zhang , Junjin Xiao , Qing Zhang

The data storage has been one of the bottlenecks in surveillance systems. The conventional video compression algorithms such as H.264 and H.265 do not fully utilize the low information density characteristic of the surveillance video. In…

计算机视觉与模式识别 · 计算机科学 2020-09-29 Lirong Wu , Kejie Huang , Haibin Shen , Lianli Gao

The segmentation of video sequences into foreground and background regions is a low-level process commonly used in video content analysis and smart surveillance applications. Using a multispectral camera setup can improve this process by…

计算机视觉与模式识别 · 计算机科学 2018-12-27 Pierre-Luc St-Charles , Guillaume-Alexandre Bilodeau , Robert Bergevin

Recent advancements in human video synthesis have enabled the generation of high-quality videos through the application of stable diffusion models. However, existing methods predominantly concentrate on animating solely the human element…

计算机视觉与模式识别 · 计算机科学 2024-05-29 Jinlin Liu , Kai Yu , Mengyang Feng , Xiefan Guo , Miaomiao Cui

We in this paper solve the problem of high-quality automatic real-time background cut for 720p portrait videos. We first handle the background ambiguity issue in semantic segmentation by proposing a global background attenuation model. A…

计算机视觉与模式识别 · 计算机科学 2017-05-01 Xiaoyong Shen , Ruixing Wang , Hengshuang Zhao , Jiaya Jia

This paper aims to tackle the problem of modeling dynamic urban streets for autonomous driving scenes. Recent methods extend NeRF by incorporating tracked vehicle poses to animate vehicles, enabling photo-realistic view synthesis of dynamic…

计算机视觉与模式识别 · 计算机科学 2024-08-20 Yunzhi Yan , Haotong Lin , Chenxu Zhou , Weijie Wang , Haiyang Sun , Kun Zhan , Xianpeng Lang , Xiaowei Zhou , Sida Peng

Frame sampling is a fundamental problem in video action recognition due to the essential redundancy in time and limited computation resources. The existing sampling strategy often employs a fixed frame selection and lacks the flexibility to…

计算机视觉与模式识别 · 计算机科学 2021-08-23 Yuan Zhi , Zhan Tong , Limin Wang , Gangshan Wu

Cross-modal video retrieval aims to retrieve the semantically relevant videos given a text as a query, and is one of the fundamental tasks in Multimedia. Most of top-performing methods primarily leverage Visual Transformer (ViT) to extract…

计算机视觉与模式识别 · 计算机科学 2022-10-18 Ning Han , Xun Yang , Ee-Peng Lim , Hao Chen , Qianru Sun

We design a real-time portrait matting pipeline for everyday use, particularly for "virtual backgrounds" in video conferences. Existing segmentation and matting methods prioritize accuracy and quality over throughput and efficiency, and our…

计算机视觉与模式识别 · 计算机科学 2020-06-12 Jo Chuang , Qian Dong

Video diffusion models achieve strong frame-level fidelity but still struggle with motion coherence, dynamics and realism, often producing jitter, ghosting, or implausible dynamics. A key limitation is that the standard denoising MSE…

计算机视觉与模式识别 · 计算机科学 2025-11-27 Haotian Xue , Qi Chen , Zhonghao Wang , Xun Huang , Eli Shechtman , Jinrong Xie , Yongxin Chen

Video rain/snow removal from surveillance videos is an important task in the computer vision community since rain/snow existed in videos can severely degenerate the performance of many surveillance system. Various methods have been…

计算机视觉与模式识别 · 计算机科学 2019-09-16 Minghan Li , Xiangyong Cao , Qian Zhao , Lei Zhang , Chenqiang Gao , Deyu Meng

We have recently seen great progress in 3D scene reconstruction through explicit point-based 3D Gaussian Splatting (3DGS), notable for its high quality and fast rendering speed. However, reconstructing dynamic scenes such as complex human…

计算机视觉与模式识别 · 计算机科学 2025-03-10 Chao Zhang , Yifeng Zhou , Shuheng Wang , Wenfa Li , Degang Wang , Yi Xu , Shaohui Jiao

Graph Neural Networks (GNNs) have demonstrated superior performance across various graph learning tasks but face significant computational challenges when applied to large-scale graphs. One effective approach to mitigate these challenges is…

机器学习 · 计算机科学 2024-10-04 Guibin Zhang , Xiangguo Sun , Yanwei Yue , Chonghe Jiang , Kun Wang , Tianlong Chen , Shirui Pan

Due to the difficulty of solving the matting problem, lots of methods use some kinds of assistance to acquire high quality alpha matte. Green screen matting methods rely on physical equipment. Trimap-based methods take manual interactions…

计算机视觉与模式识别 · 计算机科学 2022-03-15 Jinlin Liu