中文
相关论文

相关论文: Photoreal Scene Reconstruction from an Egocentric …

200 篇论文

Spatial scene-understanding, including dense depth and ego-motion estimation, is an important problem in computer vision for autonomous vehicles and advanced driver assistance systems. Thus, it is beneficial to design perception modules…

计算机视觉与模式识别 · 计算机科学 2023-02-03 Hemang Chawla , Matti Jukola , Shabbir Marzban , Elahe Arani , Bahram Zonooz

With the development of eXtended Reality (XR), photo capturing and display technology based on head-mounted displays (HMDs) have experienced significant advancements and gained considerable attention. Egocentric spatial images and videos…

计算机视觉与模式识别 · 计算机科学 2025-02-24 Xilei Zhu , Liu Yang , Huiyu Duan , Xiongkuo Min , Guangtao Zhai , Patrick Le Callet

External effects such as shocks and temperature variations affect the calibration of visual-inertial sensor systems and thus they cannot fully rely on factory calibrations. Re-calibrations performed on short user-collected datasets might…

机器人学 · 计算机科学 2019-01-23 Thomas Schneider , Mingyang Li , Cesar Cadena , Juan Nieto , Roland Siegwart

The fusion of sensor data from heterogeneous sensors is crucial for robust perception in various robotics applications that involve moving platforms, for instance, autonomous vehicle navigation. In particular, combining camera and lidar…

机器人学 · 计算机科学 2020-03-10 Mao Shan , Julie Stephany Berrio , Stewart Worrall , Eduardo Nebot

We present EgoRenderer, a system for rendering full-body neural avatars of a person captured by a wearable, egocentric fisheye camera that is mounted on a cap or a VR headset. Our system renders photorealistic novel views of the actor and…

计算机视觉与模式识别 · 计算机科学 2021-11-25 Tao Hu , Kripasindhu Sarkar , Lingjie Liu , Matthias Zwicker , Christian Theobalt

We present a new solution to egocentric 3D body pose estimation from monocular images captured from a downward looking fish-eye camera installed on the rim of a head mounted virtual reality device. This unusual viewpoint, just 2 cm. away…

计算机视觉与模式识别 · 计算机科学 2019-07-24 Denis Tome , Patrick Peluse , Lourdes Agapito , Hernan Badino

Visual-inertial systems have been widely studied and applied in the last two decades (from the early 2000s to the present), mainly due to their low cost and power consumption, small footprint, and high availability. Such a trend…

机器人学 · 计算机科学 2024-11-01 Shuolong Chen , Xingxing Li , Shengyu Li , Yuxuan Zhou

Incrementally recovering real-sized 3D geometry from a pose-free RGB stream is a challenging task in 3D reconstruction, requiring minimal assumptions on input data. Existing methods can be broadly categorized into end-to-end and visual…

计算机视觉与模式识别 · 计算机科学 2025-08-07 Linqing Zhao , Xiuwei Xu , Yirui Wang , Hao Wang , Wenzhao Zheng , Yansong Tang , Haibin Yan , Jiwen Lu

High-dynamic scene reconstruction aims to represent static background with rigid spatial features and dynamic objects with deformed continuous spatiotemporal features. Typically, existing methods adopt unified representation model (e.g.,…

计算机视觉与模式识别 · 计算机科学 2025-07-01 Hanyu Zhou , Haonan Wang , Haoyue Liu , Yuxing Duan , Luxin Yan , Gim Hee Lee

Gaussian Splatting (GS) is a popular approach for 3D reconstruction, mostly due to its ability to converge reasonably fast, faithfully represent the scene and render (novel) views in a fast fashion. However, it suffers from large storage…

计算机视觉与模式识别 · 计算机科学 2025-04-10 Anil Armagan , Albert Saà-Garriga , Bruno Manganelli , Kyuwon Kim , M. Kerim Yucel

Deformable Gaussian Splatting (GS) accomplishes photorealistic dynamic 3-D reconstruction from dense multi-view video (MVV) by learning to deform a canonical GS representation. However, in filmmaking, tight budgets can result in sparse…

计算机视觉与模式识别 · 计算机科学 2026-04-21 Adrian Azzarelli , Nantheera Anantrasirichai , David R Bull

Photo-realistic novel view synthesis from multi-view images, such as neural radiance field (NeRF) and 3D Gaussian Splatting (3DGS), has gained significant attention for its superior performance. However, most existing methods rely on low…

计算机视觉与模式识别 · 计算机科学 2025-08-19 Shucheng Gong , Lingzhe Zhao , Wenpu Li , Hong Xie , Yin Zhang , Shiyu Zhao , Peidong Liu

Precise 6-DoF simultaneous localization and mapping (SLAM) from onboard sensors is critical for wearable devices capturing egocentric data, which exhibits specific challenges, such as a wider diversity of motions and viewpoints, prevalent…

计算机视觉与模式识别 · 计算机科学 2025-10-01 Anusha Krishnan , Shaohui Liu , Paul-Edouard Sarlin , Oscar Gentilhomme , David Caruso , Maurizio Monge , Richard Newcombe , Jakob Engel , Marc Pollefeys

We present a framework that enables fast reconstruction and real-time rendering of urban-scale scenes while maintaining robustness against appearance variations across multi-view captures. Our approach begins with scene partitioning for…

计算机视觉与模式识别 · 计算机科学 2025-08-01 Zhensheng Yuan , Haozhi Huang , Zhen Xiong , Di Wang , Guanghua Yang

Recent approaches have successfully focused on the segmentation of static reconstructions, thereby equipping downstream applications with semantic 3D understanding. However, the world in which we live is dynamic, characterized by numerous…

机器人学 · 计算机科学 2025-03-12 Tjark Behrens , René Zurbrügg , Marc Pollefeys , Zuria Bauer , Hermann Blum

Implicit neural representation and explicit 3D Gaussian Splatting (3D-GS) for novel view synthesis have achieved remarkable progress with frame-based camera (e.g. RGB and RGB-D cameras) recently. Compared to frame-based camera, a novel type…

计算机视觉与模式识别 · 计算机科学 2025-03-26 Jian Huang , Chengrui Dong , Xuanhua Chen , Peidong Liu

Dynamic scenes that contain both object motion and egomotion are a challenge for monocular visual odometry (VO). Another issue with monocular VO is the scale ambiguity, i.e. these methods cannot estimate scene depth and camera motion in…

计算机视觉与模式识别 · 计算机科学 2020-08-31 Hirak J Kashyap , Charless Fowlkes , Jeffrey L Krichmar

Novel view synthesis techniques predominantly utilize RGB cameras, inheriting their limitations such as the need for sufficient lighting, susceptibility to motion blur, and restricted dynamic range. In contrast, event cameras are…

计算机视觉与模式识别 · 计算机科学 2025-02-18 Sohaib Zahid , Viktor Rudnev , Eddy Ilg , Vladislav Golyanik

Egocentric videos provide valuable insights into human interactions with the physical world, which has sparked growing interest in the computer vision and robotics communities. A critical challenge in fully understanding the geometry and…

计算机视觉与模式识别 · 计算机科学 2025-07-15 Chengbo Yuan , Geng Chen , Li Yi , Yang Gao

We present FunRec, a method for reconstructing functional 3D digital twins of indoor scenes directly from egocentric RGB-D interaction videos. Unlike existing methods on articulated reconstruction, which rely on controlled setups,…