中文
相关论文

相关论文: Does it work outside this benchmark? Introducing t…

200 篇论文

Deploying autonomous robots in crowded indoor environments usually requires them to have accurate dynamic obstacle perception. Although plenty of previous works in the autonomous driving field have investigated the 3D object detection…

机器人学 · 计算机科学 2024-02-28 Zhefan Xu , Xiaoyang Zhan , Yumeng Xiu , Christopher Suzuki , Kenji Shimada

We propose a novel framework for creating large-scale photorealistic datasets of indoor scenes, with ground truth geometry, material, lighting and semantics. Our goal is to make the dataset creation process widely accessible, transforming…

Fisheye cameras are increasingly adopted in robotics for near-field manipulation, navigation, and immersive perception, yet indoor depth benchmarks with accurate ground truth are still missing. To address this, we introduce WideDepth - the…

计算机视觉与模式识别 · 计算机科学 2026-05-26 Ilia Indyk , Ignat Penshin , Ivan Sosin , Maxim Monastyrny , Aleksei Valenkov , Ilya Makarov

Understanding visually-rich business documents to extract structured data and automate business workflows has been receiving attention both in academia and industry. Although recent multi-modal language models have achieved impressive…

计算与语言 · 计算机科学 2023-09-19 Zilong Wang , Yichao Zhou , Wei Wei , Chen-Yu Lee , Sandeep Tata

With the increasing availability of large databases of 3D CAD models, depth-based recognition methods can be trained on an uncountable number of synthetically rendered images. However, discrepancies with the real data acquired from various…

计算机视觉与模式识别 · 计算机科学 2018-05-25 Sergey Zakharov , Benjamin Planche , Ziyan Wu , Andreas Hutter , Harald Kosch , Slobodan Ilic

In this paper, we propose a new technique that applies automated image analysis in the area of structural corrosion monitoring and demonstrate improved efficacy compared to existing approaches. Structural corrosion monitoring is the initial…

计算机视觉与模式识别 · 计算机科学 2022-01-12 Abdur Rahim Mohammad Forkan , Yong-Bin Kang , Prem Prakash Jayaraman , Kewen Liao , Rohit Kaul , Graham Morgan , Rajiv Ranjan , Samir Sinha

Most existing algorithms for depth estimation from single monocular images need large quantities of metric groundtruth depths for supervised learning. We show that relative depth can be an informative cue for metric depth estimation and can…

计算机视觉与模式识别 · 计算机科学 2019-07-12 Yuanzhouhan Cao , Tianqi Zhao , Ke Xian , Chunhua Shen , Zhiguo Cao , Shugong Xu

Deep reinforcement learning (DRL) demonstrates its potential in learning a model-free navigation policy for robot visual navigation. However, the data-demanding algorithm relies on a large number of navigation trajectories in training.…

机器人学 · 计算机科学 2018-02-27 Kaichun Mo , Haoxiang Li , Zhe Lin , Joon-Young Lee

Accurate three-dimensional perception is essential for modern industrial robotic systems that perform manipulation, inspection, and navigation tasks. RGB-D and stereo vision sensors are widely used for this purpose, but the depth maps they…

计算机视觉与模式识别 · 计算机科学 2025-12-10 Tony Salloom , Dandi Zhou , Xinhai Sun

We consider the problem of reconstructing a dynamic scene observed from a stereo camera. Most existing methods for depth from stereo treat different stereo frames independently, leading to temporally inconsistent depth predictions. Temporal…

计算机视觉与模式识别 · 计算机科学 2023-05-04 Nikita Karaev , Ignacio Rocco , Benjamin Graham , Natalia Neverova , Andrea Vedaldi , Christian Rupprecht

In recent years, self-supervised methods for monocular depth estimation has rapidly become an significant branch of depth estimation task, especially for autonomous driving applications. Despite the high overall precision achieved, current…

计算机视觉与模式识别 · 计算机科学 2020-09-10 Feng Xue , Guirong Zhuo , Ziyuan Huang , Wufei Fu , Zhuoyue Wu , Marcelo H. Ang

As an essential procedure of data fusion, LiDAR-camera calibration is critical for autonomous vehicles and robot navigation. Most calibration methods rely on hand-crafted features and require significant amounts of extracted features or…

机器人学 · 计算机科学 2021-04-27 Xudong Lv , Boya Wang , Ziwen Dou , Dong Ye , Shuo Wang

Accurate depth estimation plays a critical role in the navigation of endoscopic surgical robots, forming the foundation for 3D reconstruction and safe instrument guidance. Fine-tuning pretrained models heavily relies on endoscopic surgical…

计算机视觉与模式识别 · 计算机科学 2026-02-27 Yinheng Lin , Yiming Huang , Beilei Cui , Long Bai , Huxin Gao , Hongliang Ren , Jiewen Lai

Learned confidence measures gain increasing importance for outlier removal and quality improvement in stereo vision. However, acquiring the necessary training data is typically a tedious and time consuming task that involves manual…

计算机视觉与模式识别 · 计算机科学 2016-04-19 Christian Mostegel , Markus Rumpler , Friedrich Fraundorfer , Horst Bischof

This work presents the Large Depth Completion Model (LDCM), a simple, effective, and robust framework for single-view metric depth estimation with sparse observations. Without relying on complex architectural designs, LDCM generates…

计算机视觉与模式识别 · 计算机科学 2026-05-29 Zhu Yu , Zhengyi Zhao , Runmin Zhang , Lingteng Qiu , Kejie Qiu , Yisheng He , Siyu Zhu , Zilong Dong , Si-Yuan Cao , Hui-Liang Shen

Data depth is a concept in multivariate statistics that measures the centrality of a point in a given data cloud in $\IR^d$. If the depth of a point can be represented as the minimum of the depths with respect to all one-dimensional…

统计计算 · 统计学 2020-07-17 Rainer Dyckerhoff , Pavlo Mozharovskyi , Stanislav Nagy

Software development in the aerospace domain requires adhering to strict, high-quality standards. While there exist regulatory guidelines for commercial software in this domain (e.g., ARP-4754 and DO-178), these do not apply to software…

软件工程 · 计算机科学 2024-08-06 Guy Katz , Natan Levy , Idan Refaeli , Raz Yerushalmi

Monocular depth estimation has drawn widespread attention from the vision community due to its broad applications. In this paper, we propose a novel physics (geometry)-driven deep learning framework for monocular depth estimation by…

计算机视觉与模式识别 · 计算机科学 2023-09-26 Shuwei Shao , Zhongcai Pei , Weihai Chen , Xingming Wu , Zhengguo Li

Underwater imaging is fundamentally challenging due to wavelength-dependent light attenuation, strong scattering from suspended particles, turbidity-induced blur, and non-uniform illumination. These effects impair standard cameras and make…

计算机视觉与模式识别 · 计算机科学 2026-01-16 Nick Truong , Pritam P. Karmokar , William J. Beksi

This paper presents a simulation workflow for generating synthetic LiDAR datasets to support autonomous vehicle perception, robotics research, and sensor security analysis. Leveraging the CoppeliaSim simulation environment and its Python…

机器人学 · 计算机科学 2025-06-24 Abhishek Phadke , Shakib Mahmud Dipto , Pratip Rana