English
Related papers

Related papers: SVDC: Consistent Direct Time-of-Flight Video Depth…

200 papers

Consecutive frames in a video contain redundancy, but they may also contain relevant complementary information for the detection task. The objective of our work is to leverage this complementary information to improve detection. Therefore,…

Computer Vision and Pattern Recognition · Computer Science 2024-02-19 Noreen Anwar , Guillaume-Alexandre Bilodeau , Wassim Bouachir

Voice spoofing attacks pose a significant threat to automated speaker verification systems. Existing anti-spoofing methods often simulate specific attack types, such as synthetic or replay attacks. However, in real-world scenarios, the…

Sound · Computer Science 2023-09-19 Awais Khan , Khalid Mahmood Malik , Shah Nawaz

3D depth sensors using single-photon avalanche diodes (SPADs) are becoming increasingly common in applications such as autonomous navigation and object detection. Recent designs implement on-chip histogramming time-to-digital converters…

Image and Video Processing · Electrical Eng. & Systems 2024-02-13 Jack MacLean , Brian Stewart , Istvan Gyongy

Long-distance depth imaging holds great promise for applications such as autonomous driving and robotics. Direct time-of-flight (dToF) imaging offers high-precision, long-distance depth sensing, yet demands ultra-short pulse light sources…

Computer Vision and Pattern Recognition · Computer Science 2025-05-29 Manchao Bao , Shengjiang Fang , Tao Yue , Xuemei Hu

We propose SparseDC, a model for Depth Completion of Sparse and non-uniform depth inputs. Unlike previous methods focusing on completing fixed distributions on benchmark datasets (e.g., NYU with 500 points, KITTI with 64 lines), SparseDC is…

Computer Vision and Pattern Recognition · Computer Science 2023-12-04 Chen Long , Wenxiao Zhang , Zhe Chen , Haiping Wang , Yuan Liu , Zhen Cao , Zhen Dong , Bisheng Yang

3D Time-of-Flight (ToF) image sensors are used widely in applications such as self-driving cars, Augmented Reality (AR) and robotics. When implemented with Single-Photon Avalanche Diodes (SPADs), compact, array format sensors can be made…

Image and Video Processing · Electrical Eng. & Systems 2023-03-22 Germán Mora Martín , Stirling Scholes , Alice Ruget , Robert K. Henderson , Jonathan Leach , Istvan Gyongy

Indirect Time-of-Flight (I-ToF) imaging is a widespread way of depth estimation for mobile devices due to its small size and affordable price. Previous works have mainly focused on quality improvement for I-ToF imaging especially curing the…

Computer Vision and Pattern Recognition · Computer Science 2022-06-17 HyunJun Jung , Nikolas Brasch , Ales Leonardis , Nassir Navab , Benjamin Busam

Multi-focus is a technique of focusing on different aspects of a particular object or scene. Wireless Visual Sensor Networks (WVSN) use multi-focus image fusion, which combines two or more images to create a more accurate output image that…

Computer Vision and Pattern Recognition · Computer Science 2023-05-22 Krishnendu K. S.

As one of the automotive sensors that have emerged in recent years, 4D millimeter-wave radar has a higher resolution than conventional 3D radar and provides precise elevation measurements. But its point clouds are still sparse and noisy,…

Computer Vision and Pattern Recognition · Computer Science 2026-01-14 Hongsi Liu , Jun Liu , Guangfeng Jiang , Xin Jin

We present MobiFuse, a high-precision depth perception system on mobile devices that combines dual RGB and Time-of-Flight (ToF) cameras. To achieve this, we leverage physical principles from various environmental factors to propose the…

Computer Vision and Pattern Recognition · Computer Science 2024-12-19 Jinrui Zhang , Deyu Zhang , Tingting Long , Wenxin Chen , Ju Ren , Yunxin Liu , Yudong Zhao , Yaoxue Zhang , Youngki Lee

Indirect Time-of-Flight (iToF) cameras are a widespread type of 3D sensor, which perform multiple captures to obtain depth values of the captured scene. While recent approaches to correct iToF depths achieve high performance when removing…

Computer Vision and Pattern Recognition · Computer Science 2022-10-20 Michael Schelling , Pedro Hermosilla , Timo Ropinski

Applying salient object detection (SOD) to RGB-D videos is an emerging task called RGB-D VSOD and has recently gained increasing interest, due to considerable performance gains of incorporating motion and depth and that RGB-D videos can be…

Computer Vision and Pattern Recognition · Computer Science 2025-07-30 Jiahao He , Daerji Suolang , Keren Fu , Qijun Zhao

Semantic Scene Completion (SSC) aims to jointly infer semantics and occupancies of 3D scenes. Truncated Signed Distance Function (TSDF), a 3D encoding of depth, has been a common input for SSC. Furthermore, RGB-TSDF fusion, seems promising…

Computer Vision and Pattern Recognition · Computer Science 2024-11-26 Laiyan Ding , Panwen Hu , Jie Li , Rui Huang

Spatially and temporally highly resolved depth information enables numerous applications including human-machine interaction in gaming or safety functions in the automotive industry. In this paper, we address this issue using Time-of-flight…

Numerical Analysis · Mathematics 2018-12-27 Stephan Antholzer , Christoph Wolf , Michael Sandbichler , Markus Dielacher , Markus Haltmeier

Time-of-Flight (ToF) depth sensing camera is able to obtain depth maps at a high frame rate. However, its low resolution and sensitivity to the noise are always a concern. A popular solution is upsampling the obtained noisy low resolution…

Computer Vision and Pattern Recognition · Computer Science 2015-06-18 Wei Liu , Yijun Li , Xiaogang Chen , Jie Yang , Qiang Wu , Jingyi Yu

Depth cameras are emerging as a cornerstone modality with diverse applications that directly or indirectly rely on measured depth, including personal devices, robotics, and self-driving vehicles. Although time-of-flight (ToF) methods have…

Computer Vision and Pattern Recognition · Computer Science 2021-05-26 Seung-Hwan Baek , Noah Walsh , Ilya Chugunov , Zheng Shi , Felix Heide

Video semantic segmentation aims to generate accurate semantic maps for each video frame. To this end, many works dedicate to integrate diverse information from consecutive frames to enhance the features for prediction, where a feature…

Computer Vision and Pattern Recognition · Computer Science 2023-01-11 Jiafan Zhuang , Zilei Wang , Junjie Li

Depth completion in dynamic scenes poses significant challenges due to rapid ego-motion and object motion, which can severely degrade the quality of input modalities such as RGB images and LiDAR measurements. Conventional RGB-D sensors…

Computer Vision and Pattern Recognition · Computer Science 2025-05-21 Zhiqiang Yan , Jianhao Jiao , Zhengxue Wang , Gim Hee Lee

Audio-visual saliency prediction aims to mimic human visual attention by identifying salient regions in videos through the integration of both visual and auditory information. Although visual-only approaches have significantly advanced,…

Computer Vision and Pattern Recognition · Computer Science 2025-04-17 Kiana Hooshanfar , Alireza Hosseini , Ahmad Kalhor , Babak Nadjar Araabi

State-of-the-art LiDAR-camera 3D object detectors usually focus on feature fusion. However, they neglect the factor of depth while designing the fusion strategy. In this work, we are the first to observe that different modalities play…

Computer Vision and Pattern Recognition · Computer Science 2025-05-13 Mingqian Ji , Jian Yang , Shanshan Zhang