中文
相关论文

相关论文: SBEVNet: End-to-End Deep Stereo Layout Estimation

200 篇论文

In this paper, we introduce the problem of estimating the real world depth of elements in a scene captured by two cameras with different field of views, where the first field of view (FOV) is a Wide FOV (WFOV) captured by a wide angle lens,…

计算机视觉与模式识别 · 计算机科学 2020-10-20 Mostafa El-Khamy , Haoyu Ren , Xianzhi Du , Jungwon Lee

At present, supervised stereo methods based on deep neural network have achieved impressive results. However, in some scenarios, accurate three-dimensional labels are inaccessible for supervised training. In this paper, a self-supervised…

计算机视觉与模式识别 · 计算机科学 2019-07-02 Xiaoyu Chen , Qixin Wang , Jinzhou Ge , Yi Zhang , Jing Han

Automating sleep staging is vital to scale up sleep assessment and diagnosis to serve millions experiencing sleep deprivation and disorders and enable longitudinal sleep monitoring in home environments. Learning from raw polysomnography…

信号处理 · 电气工程与系统科学 2021-04-06 Huy Phan , Oliver Y. Chén , Minh C. Tran , Philipp Koch , Alfred Mertins , Maarten De Vos

Stereo matching for inland waterways is one of the key technologies for the autonomous navigation of Unmanned Surface Vehicles (USVs), which involves dividing the stereo images into reference images and target images for pixel-level…

计算机视觉与模式识别 · 计算机科学 2024-10-11 Jing Su , Yiqing Zhou , Yu Zhang , Chao Wang , Yi Wei

We present a novel bird's-eye-view (BEV) detector with perspective supervision, which converges faster and better suits modern image backbones. Existing state-of-the-art BEV detectors are often tied to certain depth pre-trained backbones…

计算机视觉与模式识别 · 计算机科学 2022-11-21 Chenyu Yang , Yuntao Chen , Hao Tian , Chenxin Tao , Xizhou Zhu , Zhaoxiang Zhang , Gao Huang , Hongyang Li , Yu Qiao , Lewei Lu , Jie Zhou , Jifeng Dai

Accurate localization in challenging garage environments -- marked by poor lighting, sparse textures, repetitive structures, dynamic scenes, and the absence of GPS -- is crucial for automated valet parking (AVP) tasks. Addressing these…

机器人学 · 计算机科学 2024-07-02 Ye Li , Wenchao Yang , Dekun Lin , Qianlei Wang , Zhe Cui , Xiaolin Qin

Applying pseudo labeling techniques has been found to be advantageous in semi-supervised 3D object detection (SSOD) in Bird's-Eye-View (BEV) for autonomous driving, particularly where labeled data is limited. In the literature, Exponential…

计算机视觉与模式识别 · 计算机科学 2024-12-06 Saheli Hazra , Sudip Das , Rohit Choudhary , Arindam Das , Ganesh Sistu , Ciaran Eising , Ujjwal Bhattacharya

Efficient reasoning about the semantic, spatial, and temporal structure of a scene is a crucial prerequisite for autonomous driving. We present NEural ATtention fields (NEAT), a novel representation that enables such reasoning for…

计算机视觉与模式识别 · 计算机科学 2021-09-10 Kashyap Chitta , Aditya Prakash , Andreas Geiger

In this paper we present ActiveStereoNet, the first deep learning solution for active stereo systems. Due to the lack of ground truth, our method is fully self-supervised, yet it produces precise depth with a subpixel precision of $1/30th$…

Motivated by the need to identify erroneous disparity assignments, various approaches for uncertainty and confidence estimation of dense stereo matching have been presented in recent years. As in many other fields, especially deep learning…

计算机视觉与模式识别 · 计算机科学 2020-02-11 Max Mehltretter

End-to-end autonomous driving offers a streamlined alternative to the traditional modular pipeline, integrating perception, prediction, and planning within a single framework. While Deep Reinforcement Learning (DRL) has recently gained…

人工智能 · 计算机科学 2024-09-27 Siyi Lu , Lei He , Shengbo Eben Li , Yugong Luo , Jianqiang Wang , Keqiang Li

As a cornerstone technique for autonomous driving, Bird's Eye View (BEV) segmentation has recently achieved remarkable progress with pinhole cameras. However, it is non-trivial to extend the existing methods to fisheye cameras with severe…

计算机视觉与模式识别 · 计算机科学 2025-09-18 Hang Li , Dianmo Sheng , Qiankun Dong , Zichun Wang , Zhiwei Xu , Tao Li

Most automated driving systems comprise a diverse sensor set, including several cameras, Radars, and LiDARs, ensuring a complete 360\deg coverage in near and far regions. Unlike Radar and LiDAR, which measure directly in 3D, cameras capture…

机器人学 · 计算机科学 2023-09-20 David Unger , Nikhil Gosala , Varun Ravi Kumar , Shubhankar Borse , Abhinav Valada , Senthil Yogamani

Place recognition is a challenging but crucial task in robotics. Current description-based methods may be limited by representation capabilities, while pairwise similarity-based methods require exhaustive searches, which is time-consuming.…

计算机视觉与模式识别 · 计算机科学 2024-07-24 Chencan Fu , Lin Li , Jianbiao Mei , Yukai Ma , Linpeng Peng , Xiangrui Zhao , Yong Liu

Stereo matching is a fundamental task for 3D scene reconstruction. Recently, deep learning based methods have proven effective on some benchmark datasets, such as KITTI and Scene Flow. UAVs (Unmanned Aerial Vehicles) are commonly utilized…

计算机视觉与模式识别 · 计算机科学 2023-02-21 Zhang Xiaoyi , Cao Xuefeng , Yu Anzhu , Yu Wenshuai , Li Zhenqi , Quan Yujun

Monocular visual-inertial odometry (VIO) is a critical problem in robotics and autonomous driving. Traditional methods solve this problem based on filtering or optimization. While being fully interpretable, they rely on manual interference…

机器人学 · 计算机科学 2022-09-20 Zexi Chen , Haozhe Du , Xuecheng Xu , Rong Xiong , Yiyi Liao , Yue Wang

Pairwise point cloud registration is a critical task for many applications, which heavily depends on finding correct correspondences from the two point clouds. However, the low overlap between input point clouds causes the registration to…

计算机视觉与模式识别 · 计算机科学 2023-03-15 Lin Li , Wendong Ding , Yongkun Wen , Yufei Liang , Yong Liu , Guowei Wan

Disparity/depth estimation from sequences of stereo images is an important element in 3D vision. Owing to occlusions, imperfect settings and homogeneous luminance, accurate estimate of depth remains a challenging problem. Targetting view…

图像与视频处理 · 电气工程与系统科学 2020-03-17 Nantheera Anantrasirichai , Majid Geravand , David Braendler , David R. Bull

End-to-end deep learning methods have advanced stereo vision in recent years and obtained excellent results when the training and test data are similar. However, large datasets of diverse real-world scenes with dense ground truth are…

计算机视觉与模式识别 · 计算机科学 2020-08-26 Jialiang Wang , Varun Jampani , Deqing Sun , Charles Loop , Stan Birchfield , Jan Kautz

LiDAR-based place recognition (LPR) is one of the most crucial components of autonomous vehicles to identify previously visited places in GPS-denied environments. Most existing LPR methods use mundane representations of the input point…

计算机视觉与模式识别 · 计算机科学 2023-10-09 Junyi Ma , Guangming Xiong , Jingyi Xu , Xieyuanli Chen