中文
相关论文

相关论文: Visual Vibration Tomography: Estimating Interior M…

200 篇论文

We obtain quantitative measurements of the oscillation amplitude of vibrating objects by using sideband digital holography. The frequency sidebands on the light scattered by the object, shifted by n times the vibration frequency, are…

光学 · 物理学 2009-03-13 Fadwa Joud , Franck Laloë , Michael Atlan , Jean Hare , Michel Gross

Learning to predict scene depth from RGB inputs is a challenging task both for indoor and outdoor robot navigation. In this work we address unsupervised learning of scene depth and robot ego-motion where supervision is provided by monocular…

计算机视觉与模式识别 · 计算机科学 2018-11-16 Vincent Casser , Soeren Pirk , Reza Mahjourian , Anelia Angelova

When interacting in a three dimensional world, humans must estimate 3D structure from visual inputs projected down to two dimensional retinal images. It has been shown that humans use the persistence of object shape over motion-induced…

神经元与认知 · 定量生物学 2023-04-03 Marissa Connor , Bruno Olshausen , Christopher Rozell

Video object insertion is a critical task for dynamically inserting new objects into existing environments. Previous video generation methods focus primarily on synthesizing entire scenes while struggling with ensuring consistent object…

计算机视觉与模式识别 · 计算机科学 2026-04-17 Xia Qi , Peishan Cong , Yichen Yao , Ziyi Wang , Yaoqin Ye , Yuexin Ma

Inferring rigid-body physical states and properties from monocular videos is a fundamental step toward physics-based perception and simulation. Existing approaches assume specific underlying physical systems, object types, and camera poses,…

计算机视觉与模式识别 · 计算机科学 2026-05-21 Chia-Hsiang Kao , Cong Phuoc Huynh , Chien-Yi Wang , Noranart Vesdapunt , Stefan Stojanov , Bharath Hariharan , Oleksandr Obiednikov , Ning Zhou

Traditional 3D morphable face models (3DMMs) provide fine-grained control over expression but cannot easily capture geometric and appearance details. Neural volumetric representations approach photorealism but are hard to animate and do not…

计算机视觉与模式识别 · 计算机科学 2022-11-07 Yufeng Zheng , Victoria Fernández Abrevaya , Marcel C. Bühler , Xu Chen , Michael J. Black , Otmar Hilliges

Video inpainting is the task of filling a region in a video in a visually convincing manner. It is very challenging due to the high dimensionality of the data and the temporal consistency required for obtaining convincing results. Recently,…

计算机视觉与模式识别 · 计算机科学 2025-04-29 Nicolas Cherel , Andrés Almansa , Yann Gousseau , Alasdair Newson

Optical vibration sensing enables recovering the scene sound directly from the surface vibration of nearby objects, turning everyday objects into ``visual microphones''. However, most prior methods had focused on capturing the vibrations of…

计算机视觉与模式识别 · 计算机科学 2026-04-30 Shai Bagon , Matan Kichler , Mark Sheinin

We present a method for jointly training the estimation of depth, ego-motion, and a dense 3D translation field of objects relative to the scene, with monocular photometric consistency being the sole source of supervision. We show that this…

计算机视觉与模式识别 · 计算机科学 2020-11-10 Hanhan Li , Ariel Gordon , Hang Zhao , Vincent Casser , Anelia Angelova

In developing organisms, internal cellular processes generate mechanical stresses at the tissue scale. The resulting deformations depend on the material properties of the tissue, which can exhibit long-ranged orientational order and…

软凝聚态物质 · 物理学 2021-01-20 Carles Blanch-Mercader , Pau Guillamat , Aurélien Roux , Karsten Kruse

Estimating accurate camera poses, 3D scene geometry, and object motion from in-the-wild videos is a long-standing challenge for classical structure from motion pipelines due to the presence of dynamic objects. Recent learning-based methods…

计算机视觉与模式识别 · 计算机科学 2025-12-08 Zhuoyuan Wu , Xurui Yang , Jiahui Huang , Yue Wang , Jun Gao

Three-dimensional (3D) reconstruction from a single image is an ill-posed problem with inherent ambiguities, i.e. scale. Predicting a 3D scene from text description(s) is similarly ill-posed, i.e. spatial arrangements of objects described.…

计算机视觉与模式识别 · 计算机科学 2024-06-04 Ziyao Zeng , Daniel Wang , Fengyu Yang , Hyoungseob Park , Yangchao Wu , Stefano Soatto , Byung-Woo Hong , Dong Lao , Alex Wong

During VR demos we have performed over last few years, many participants (in the absence of any haptic feedback) have commented on their perceived ability to 'feel' differences between simulated molecular objects. The mechanisms for such…

Motion blur is one of the major challenges remaining for visual odometry methods. In low-light conditions where longer exposure times are necessary, motion blur can appear even for relatively slow camera motions. In this paper we present a…

计算机视觉与模式识别 · 计算机科学 2021-03-26 Peidong Liu , Xingxing Zuo , Viktor Larsson , Marc Pollefeys

We demonstrate that, under orthographic projection and with a camera fixated on a point located on a rigid body, the rotation of that body can be analytically obtained by tracking only one other feature in the image. With some exceptions,…

计算机视觉与模式识别 · 计算机科学 2025-07-08 Daniel Raviv , Juan D. Yepes , Eiki M. Martinson

The field of indoor monocular 3D object detection is gaining significant attention, fueled by the increasing demand in VR/AR and robotic applications. However, its advancement is impeded by the limited availability and diversity of 3D…

计算机视觉与模式识别 · 计算机科学 2024-12-17 Jin-Cheng Jhang , Tao Tu , Fu-En Wang , Ke Zhang , Min Sun , Cheng-Hao Kuo

This paper proposes a new objective metric of exceptional motion in VR video contents for VR sickness assessment. In VR environment, VR sickness can be caused by several factors which are mismatched motion, field of view, motion parallax,…

计算机视觉与模式识别 · 计算机科学 2018-04-12 Hak Gu Kim , Wissam J. Baddar , Heoun-taek Lim , Hyunwook Jeong , Yong Man Ro

In this paper, we propose a scale-aware method for inserting virtual objects with proper sizes into monocular videos. To tackle the scale ambiguity problem of geometry recovery from monocular videos, we estimate the global scale objects in…

计算机视觉与模式识别 · 计算机科学 2020-12-07 Songhai Zhang , Xiangli Li , Yingtian Liu , Hongbo Fu

Computer vision seeks to infer a wide range of information about objects and events. However, vision systems based on conventional imaging are limited to extracting information only from the visible surfaces of scene objects. For instance,…

计算机视觉与模式识别 · 计算机科学 2025-07-29 Matan Kichler , Shai Bagon , Mark Sheinin

Monocular Depth Estimation (MDE) is performed to produce 3D information that can be used in downstream tasks such as those related to on-board perception for Autonomous Vehicles (AVs) or driver assistance. Therefore, a relevant arising…

计算机视觉与模式识别 · 计算机科学 2023-02-21 Akhil Gurram , Antonio M. Lopez