English
Related papers

Related papers: Self-Supervised Video Desmoking for Laparoscopic S…

200 papers

Semantic tool segmentation in surgical videos is important for surgical scene understanding and computer-assisted interventions as well as for the development of robotic automation. The problem is challenging because different illumination…

Computer Vision and Pattern Recognition · Computer Science 2020-07-28 Emanuele Colleoni , Philip Edwards , Danail Stoyanov

Modern consumer cameras usually employ the rolling shutter (RS) mechanism, where images are captured by scanning scenes row-by-row, yielding RS distortions for dynamic scenes. To correct RS distortions, existing methods adopt a fully…

Computer Vision and Pattern Recognition · Computer Science 2023-09-15 Wei Shang , Dongwei Ren , Chaoyu Feng , Xiaotao Wang , Lei Lei , Wangmeng Zuo

In endoscopic surgery, a clear and high-quality visual field is critical for surgeons to make accurate intraoperative decisions. However, persistent visual degradation, including smoke generated by energy devices, lens fogging from thermal…

Computer Vision and Pattern Recognition · Computer Science 2026-03-19 Jialun Pei , Diandian Guo , Donghui Yang , Zhixi Li , Yuxin Feng , Long Ma , Bo Du , Pheng-Ann Heng

Ultrasound (US) is widely used for its advantages of real-time imaging, radiation-free and portability. In clinical practice, analysis and diagnosis often rely on US sequences rather than a single image to obtain dynamic anatomical…

Computer Vision and Pattern Recognition · Computer Science 2022-07-04 Jiamin Liang , Xin Yang , Yuhao Huang , Kai Liu , Xinrui Zhou , Xindi Hu , Zehui Lin , Huanjia Luo , Yuanji Zhang , Yi Xiong , Dong Ni

Depth cameras are frequently used in robotic manipulation, e.g. for visual servoing. The quality of small and compact depth cameras is though often not sufficient for depth reconstruction, which is required for precise tracking in and…

Machine Learning · Computer Science 2023-05-11 Claudius Kienle , David Petri

Pre-training video transformers generally requires a large amount of data, presenting significant challenges in terms of data collection costs and concerns related to privacy, licensing, and inherent biases. Synthesizing data is one of the…

Computer Vision and Pattern Recognition · Computer Science 2024-09-11 Yuchi Ishikawa , Masayoshi Kondo , Yoshimitsu Aoki

Supervised training has led to state-of-the-art results in image and video denoising. However, its application to real data is limited since it requires large datasets of noisy-clean pairs that are difficult to obtain. For this reason,…

Image and Video Processing · Electrical Eng. & Systems 2022-04-26 Valéry Dewil , Aranud Barral , Gabriele Facciolo , Pablo Arias

Self-supervised learning has transformed 2D computer vision by enabling models trained on large, unannotated datasets to provide versatile off-the-shelf features that perform similarly to models trained with labels. However, in 3D scene…

Computer Vision and Pattern Recognition · Computer Science 2025-04-10 Pedro Hermosilla , Christian Stippel , Leon Sick

To facilitate video denoising research, we construct a compelling dataset, namely, "Practical Video Denoising Dataset" (PVDD), containing 200 noisy-clean dynamic video pairs in both sRGB and RAW format. Compared with existing datasets…

Computer Vision and Pattern Recognition · Computer Science 2022-12-02 Xiaogang Xu , Yitong Yu , Nianjuan Jiang , Jiangbo Lu , Bei Yu , Jiaya Jia

Owing to recent advances in machine learning and the ability to harvest large amounts of data during robotic-assisted surgeries, surgical data science is ripe for foundational work. We present a large dataset of surgical videos and their…

Computer Vision and Pattern Recognition · Computer Science 2026-05-11 Aneeq Zia , Max Berniker , Rogerio Nespolo , Xiaorui Zhang , Conor Perreault , Ziheng Wang , Benjamin Mueller , Ryan Schmidt , Kiran Bhattacharyya , Xi Liu , Anthony Jarc

In this paper, we propose self-supervised training for video transformers using unlabeled video data. From a given video, we create local and global spatiotemporal views with varying spatial sizes and frame rates. Our self-supervised…

Computer Vision and Pattern Recognition · Computer Science 2022-03-22 Kanchana Ranasinghe , Muzammal Naseer , Salman Khan , Fahad Shahbaz Khan , Michael Ryoo

Shadow removal is an important computer vision task aiming at the detection and successful removal of the shadow produced by an occluded light source and a photo-realistic restoration of the image contents. Decades of re-search produced a…

Computer Vision and Pattern Recognition · Computer Science 2020-10-23 Florin-Alexandru Vasluianu , Andres Romero , Luc Van Gool , Radu Timofte

This paper tackles the challenge of automatically performing realistic surgical simulations from readily available surgical videos. Recent efforts have successfully integrated physically grounded dynamics within 3D Gaussians to perform…

Computer Vision and Pattern Recognition · Computer Science 2024-12-04 Kailing Wang , Chen Yang , Keyang Zhao , Xiaokang Yang , Wei Shen

Surgical data science (SDS) is a field that analyzes patient data before, during, and after surgery to improve surgical outcomes and skills. However, surgical data is scarce, heterogeneous, and complex, which limits the applicability of…

Computer Vision and Pattern Recognition · Computer Science 2024-10-31 Yousef Yeganeh , Rachmadio Lazuardi , Amir Shamseddin , Emine Dari , Yash Thirani , Nassir Navab , Azade Farshad

In clinical practice, 2D magnetic resonance (MR) sequences are widely adopted. While individual 2D slices can be stacked to form a 3D volume, the relatively large slice spacing can pose challenges for both image visualization and subsequent…

Image and Video Processing · Electrical Eng. & Systems 2024-06-11 Xin Wang , Zhiyun Song , Yitao Zhu , Sheng Wang , Lichi Zhang , Dinggang Shen , Qian Wang

Existing deep learning-based video super-resolution (SR) methods usually depend on the supervised learning approach, where the training data is usually generated by the blurring operation with known or predefined kernels (e.g., Bicubic…

Computer Vision and Pattern Recognition · Computer Science 2022-01-20 Haoran Bai , Jinshan Pan

Recent single-image super-resolution (SISR) networks, which can adapt their network parameters to specific input images, have shown promising results by exploiting the information available within the input data as well as large external…

Computer Vision and Pattern Recognition · Computer Science 2021-03-19 Jinsu Yoo , Tae Hyun Kim

This paper proposes a vision-based fire and smoke segmentation system which use spatial, temporal and motion information to extract the desired regions from the video frames. The fusion of information is done using multiple features such as…

Computer Vision and Pattern Recognition · Computer Science 2019-10-01 Meenu Ajith , Manel Martínez-Ramón

The performance of medical image analysis systems is constrained by the quantity of high-quality image annotations. Such systems require data to be annotated by experts with years of training, especially when diagnostic decisions are…

Computer Vision and Pattern Recognition · Computer Science 2019-02-08 Siqi Liu , Eli Gibson , Sasa Grbic , Zhoubing Xu , Arnaud Arindra Adiyoso Setio , Jie Yang , Bogdan Georgescu , Dorin Comaniciu

Decomposing a video into a layer-based representation is crucial for easy video editing for the creative industries, as it enables independent editing of specific layers. Existing video-layer decomposition models rely on implicit neural…

Computer Vision and Pattern Recognition · Computer Science 2025-03-24 Maria Pilligua , Danna Xue , Javier Vazquez-Corral