中文
相关论文

相关论文: CaLDiff: Camera Localization in NeRF via Pose Diff…

200 篇论文

We introduce Corr2Distrib, the first correspondence-based method which estimates a 6D camera pose distribution from an RGB image, explaining the observations. Indeed, symmetries and occlusions introduce visual ambiguities, leading to…

计算机视觉与模式识别 · 计算机科学 2025-05-06 Asma Brazi , Boris Meden , Fabrice Mayran de Chamisso , Steve Bourgeois , Vincent Lepetit

Numerous works have recently integrated 3D camera control into foundational text-to-video models, but the resulting camera control is often imprecise, and video generation quality suffers. In this work, we analyze camera motion from a first…

We present VERF, a collection of two methods (VERF-PnP and VERF-Light) for providing runtime assurance on the correctness of a camera pose estimate of a monocular camera without relying on direct depth measurements. We leverage the ability…

机器人学 · 计算机科学 2023-08-14 Dominic Maggio , Courtney Mario , Luca Carlone

In this work, we aim to detect the changes caused by object variations in a scene represented by the neural radiance fields (NeRFs). Given an arbitrary view and two sets of scene images captured at different timestamps, we can predict the…

计算机视觉与模式识别 · 计算机科学 2023-12-27 Rui Huang , Binbin Jiang , Qingyi Zhao , William Wang , Yuxiang Zhang , Qing Guo

Image diffusion has recently shown remarkable performance in image synthesis and implicitly as an image prior. Such a prior has been used with conditioning to solve the inpainting problem, but only supporting binary user-based conditioning.…

计算机视觉与模式识别 · 计算机科学 2023-07-04 Majed El Helou

Neural Radiance Field (NeRF) has recently emerged as a powerful representation to synthesize photorealistic novel views. While showing impressive performance, it relies on the availability of dense input views with highly accurate camera…

计算机视觉与模式识别 · 计算机科学 2023-06-14 Prune Truong , Marie-Julie Rakotosaona , Fabian Manhardt , Federico Tombari

Monocular depth predictors are typically trained on large-scale training sets which are naturally biased w.r.t the distribution of camera poses. As a result, trained predictors fail to make reliable depth predictions for testing examples…

计算机视觉与模式识别 · 计算机科学 2021-03-30 Yunhan Zhao , Shu Kong , Charless Fowlkes

Feature point matching for camera localization suffers from scalability problems. Even when feature descriptors associated with 3D scene points are locally unique, as coverage grows, similar or repeated features become increasingly common.…

计算机视觉与模式识别 · 计算机科学 2017-05-23 Raúl Díaz , Charless C. Fowlkes

Multi-resolution hash encoding has recently been proposed to reduce the computational cost of neural renderings, such as NeRF. This method requires accurate camera poses for the neural renderings of given scenes. However, contrary to…

计算机视觉与模式识别 · 计算机科学 2023-02-06 Hwan Heo , Taekyung Kim , Jiyoung Lee , Jaewon Lee , Soohyun Kim , Hyunwoo J. Kim , Jin-Hwa Kim

We present an approach to generate a 360-degree view of a person with a consistent, high-resolution appearance from a single input image. NeRF and its variants typically require videos or images from different viewpoints. Most existing…

计算机视觉与模式识别 · 计算机科学 2023-11-16 Badour AlBahar , Shunsuke Saito , Hung-Yu Tseng , Changil Kim , Johannes Kopf , Jia-Bin Huang

Animating stylized avatars with dynamic poses and expressions has attracted increasing attention for its broad range of applications. Previous research has made significant progress by training controllable generative models to synthesize…

Image retouching aims to enhance the visual quality of photos. Considering the different aesthetic preferences of users, the target of retouching is subjective. However, current retouching methods mostly adopt deterministic models, which…

计算机视觉与模式识别 · 计算机科学 2024-07-08 Zheng-Peng Duan , Jiawei zhang , Zheng Lin , Xin Jin , Dongqing Zou , Chunle Guo , Chongyi Li

Leveraging multi-modal fusion, especially between camera and LiDAR, has become essential for building accurate and robust 3D object detection systems for autonomous vehicles. Until recently, point decorating approaches, in which point…

计算机视觉与模式识别 · 计算机科学 2023-04-28 Philip Jacobson , Yiyang Zhou , Wei Zhan , Masayoshi Tomizuka , Ming C. Wu

Recently, denoising diffusion models have achieved promising results in 2D image generation and editing. Instruct-NeRF2NeRF (IN2N) introduces the success of diffusion into 3D scene editing through an "Iterative dataset update" (IDU)…

计算机视觉与模式识别 · 计算机科学 2025-02-25 Yuxuan Xiong , Yue Shi , Yishun Dou , Bingbing Ni

Image-based localization is a core component of many augmented/mixed reality (AR/MR) and autonomous robotic systems. Current localization systems rely on the persistent storage of 3D point clouds of the scene to enable camera pose…

计算机视觉与模式识别 · 计算机科学 2019-03-14 Pablo Speciale , Johannes L. Schönberger , Sing Bing Kang , Sudipta N. Sinha , Marc Pollefeys

Accurate visual localization is crucial for autonomous driving, yet existing methods face a fundamental dilemma: While high-definition (HD) maps provide high-precision localization references, their costly construction and maintenance…

计算机视觉与模式识别 · 计算机科学 2026-02-04 Li Gao , Hongyang Sun , Liu Liu , Yunhao Li , Yang Cai

Addressing pose ambiguity in 6D object pose estimation from single RGB images presents a significant challenge, particularly due to object symmetries or occlusions. In response, we introduce a novel score-based diffusion method applied to…

计算机视觉与模式识别 · 计算机科学 2024-04-09 Tsu-Ching Hsiao , Hao-Wei Chen , Hsuan-Kung Yang , Chun-Yi Lee

Neural Radiance Fields (NeRFs) have emerged as a powerful neural 3D representation for objects and scenes derived from 2D data. Generating NeRFs, however, remains difficult in many scenarios. For instance, training a NeRF with only a small…

计算机视觉与模式识别 · 计算机科学 2023-05-01 Guandao Yang , Abhijit Kundu , Leonidas J. Guibas , Jonathan T. Barron , Ben Poole

The ability to create high-quality 3D faces from a single image has become increasingly important with wide applications in video conferencing, AR/VR, and advanced video editing in movie industries. In this paper, we propose Face Diffusion…

计算机视觉与模式识别 · 计算机科学 2023-12-05 Hao Zhang , Yanbo Xu , Tianyuan Dai , Yu-Wing Tai , Chi-Keung Tang

Cascaded regression method is a fast and accurate method on finding 2D pose of objects in RGB images. It is able to find the accurate pose of objects in an image by a great number of corrections on the good initial guess of the pose of…

计算机视觉与模式识别 · 计算机科学 2017-09-26 Wenye He