中文
相关论文

相关论文: Depth Anything with Any Prior

200 篇论文

Self-supervised monocular depth estimation (DE) is an approach to learning depth without costly depth ground truths. However, it often struggles with moving objects that violate the static scene assumption during training. To address this…

计算机视觉与模式识别 · 计算机科学 2023-12-19 Jaeho Moon , Juan Luis Gonzalez Bello , Byeongjun Kwon , Munchurl Kim

Depth map estimation from images is an important task in robotic systems. Existing methods can be categorized into two groups including multi-view stereo and monocular depth estimation. The former requires cameras to have large overlapping…

计算机视觉与模式识别 · 计算机科学 2022-10-06 Jialei Xu , Xianming Liu , Yuanchao Bai , Junjun Jiang , Kaixuan Wang , Xiaozhi Chen , Xiangyang Ji

This work presents EndoStreamDepth, a monocular depth estimation framework for endoscopic video streams. It provides accurate depth maps with sharp anatomical boundaries for each frame, temporally consistent predictions across frames, and…

计算机视觉与模式识别 · 计算机科学 2026-01-05 Hao Li , Daiwei Lu , Jiacheng Wang , Robert J. Webster , Ipek Oguz

This paper tackles the problem of depth estimation from a single image. Existing work either focuses on generalization performance disregarding metric scale, i.e. relative depth estimation, or state-of-the-art results on specific datasets,…

计算机视觉与模式识别 · 计算机科学 2023-02-27 Shariq Farooq Bhat , Reiner Birkl , Diana Wofk , Peter Wonka , Matthias Müller

Monocular depth estimation (MDE) is inherently ambiguous, as a given image may result from many different 3D scenes and vice versa. To resolve this ambiguity, an MDE system must make assumptions about the most likely 3D scenes for a given…

计算机视觉与模式识别 · 计算机科学 2024-03-26 Dylan Auty , Krystian Mikolajczyk

This paper introduces Stereo Any Video, a powerful framework for video stereo matching. It can estimate spatially accurate and temporally consistent disparities without relying on auxiliary information such as camera poses or optical flow.…

计算机视觉与模式识别 · 计算机科学 2025-07-22 Junpeng Jing , Weixun Luo , Ye Mao , Krystian Mikolajczyk

Generalizable depth completion enables the acquisition of dense metric depth maps for unseen environments, offering robust perception capabilities for various downstream tasks. However, training such models typically requires large-scale…

计算机视觉与模式识别 · 计算机科学 2025-07-11 Haotian Wang , Aoran Xiao , Xiaoqin Zhang , Meng Yang , Shijian Lu

Current self-supervised monocular depth estimation (MDE) approaches encounter performance limitations due to insufficient semantic-spatial knowledge extraction. To address this challenge, we propose Hybrid-depth, a novel framework that…

计算机视觉与模式识别 · 计算机科学 2025-10-13 Wenyao Zhang , Hongsi Liu , Bohan Li , Jiawei He , Zekun Qi , Yunnan Wang , Shengyang Zhao , Xinqiang Yu , Wenjun Zeng , Xin Jin

Multi-Camera arrays are increasingly employed in both consumer and industrial applications, and various passive techniques are documented to estimate depth from such camera arrays. Current depth estimation methods provide useful estimations…

计算机视觉与模式识别 · 计算机科学 2018-06-22 Hossein Javidnia , Peter Corcoran

Estimating depth from a single image is a challenging visual task. Compared to relative depth estimation, metric depth estimation attracts more attention due to its practical physical significance and critical applications in real-life…

计算机视觉与模式识别 · 计算机科学 2026-02-25 Ruijie Zhu , Chuxin Wang , Ziyang Song , Li Liu , Tianzhu Zhang , Yongdong Zhang

Depth estimation is one of the essential tasks to be addressed when creating mobile autonomous systems. While monocular depth estimation methods have improved in recent times, depth completion provides more accurate and reliable depth maps…

计算机视觉与模式识别 · 计算机科学 2023-08-17 Wolfgang Boettcher , Lukas Hoyer , Ozan Unal , Ke Li , Dengxin Dai

Accurate Monocular Depth Estimation (MDE) is critical for autonomous robotic surgery. However, existing self-supervised methods often exhibit a severe "ex-vivo to in-vivo gap": they achieve high accuracy on public datasets but struggle in…

计算机视觉与模式识别 · 计算机科学 2026-04-21 Ankan Aich , Emma D. Ryan , Kris Moe , Isaac Schmale , Li-Xing Man , Yangming Lee

This work presents the Large Depth Completion Model (LDCM), a simple, effective, and robust framework for single-view metric depth estimation with sparse observations. Without relying on complex architectural designs, LDCM generates…

计算机视觉与模式识别 · 计算机科学 2026-05-29 Zhu Yu , Zhengyi Zhao , Runmin Zhang , Lingteng Qiu , Kejie Qiu , Yisheng He , Siyu Zhu , Zilong Dong , Si-Yuan Cao , Hui-Liang Shen

Monocular relative and metric depth estimation has seen a tremendous boost in the last few years due to the sharp advancements in foundation models and in particular transformer based networks. As we start to see applications to the domain…

Early accident anticipation from dashcam videos is a highly desirable yet challenging task for improving the safety of intelligent vehicles. Existing advanced accident anticipation approaches commonly model the interaction among traffic…

计算机视觉与模式识别 · 计算机科学 2025-02-27 Hongpu Huang , Wei Zhou , Chen Wang

There has been a recent surge of interest in learning to perceive depth from monocular videos in an unsupervised fashion. A key challenge in this field is achieving robust and accurate depth estimation in challenging scenarios, particularly…

计算机视觉与模式识别 · 计算机科学 2025-01-22 Mengtan Zhang , Yi Feng , Qijun Chen , Rui Fan

One of the most popular recent areas of machine learning predicates the use of neural networks augmented by information about the underlying process in the form of Partial Differential Equations (PDEs). These physics-informed neural…

流体动力学 · 物理学 2025-06-17 Luca Menicali , David H. Richter , Stefano Castruccio

Model generalizability to unseen datasets, concerned with in-the-wild robustness, is less studied for indoor single-image depth prediction. We leverage gradient-based meta-learning for higher generalizability on zero-shot cross-dataset…

计算机视觉与模式识别 · 计算机科学 2024-01-31 Cho-Ying Wu , Yiqi Zhong , Junying Wang , Ulrich Neumann

Recently, large models (Segment Anything model) came on the scene to provide a new baseline for polyp segmentation tasks. This demonstrates that large models with a sufficient image level prior can achieve promising performance on a given…

计算机视觉与模式识别 · 计算机科学 2024-02-06 Zhuoran Zheng , Chen Wu , Wei Wang , Yeying Jin , Xiuyi Jia

We present an algorithm for reconstructing dense, geometrically consistent depth for all pixels in a monocular video. We leverage a conventional structure-from-motion reconstruction to establish geometric constraints on pixels in the video.…

计算机视觉与模式识别 · 计算机科学 2020-08-28 Xuan Luo , Jia-Bin Huang , Richard Szeliski , Kevin Matzen , Johannes Kopf