中文
相关论文

相关论文: Ray-ONet: Efficient 3D Reconstruction From A Singl…

200 篇论文

The focus of our work is speeding up evaluation of deep neural networks in retrieval scenarios, where conventional architectures may spend too much time on negative examples. We propose to replace a monolithic network with our novel cascade…

计算机视觉与模式识别 · 计算机科学 2016-08-10 Martin Simonovsky , Nikos Komodakis

We present Mobile Video Networks (MoViNets), a family of computation and memory efficient video networks that can operate on streaming video for online inference. 3D convolutional neural networks (CNNs) are accurate at video recognition but…

计算机视觉与模式识别 · 计算机科学 2021-04-20 Dan Kondratyuk , Liangzhe Yuan , Yandong Li , Li Zhang , Mingxing Tan , Matthew Brown , Boqing Gong

Computational imaging enables compact infrared systems, but deep-learning pipelines that combine image reconstruction and object detection often introduce substantial inference latency. Most existing acceleration strategies compress the…

计算机视觉与模式识别 · 计算机科学 2026-05-22 Xuquan Wang , Guishuo Yang , Dapeng Yan , Yujie Xing , Xuanyu Qian , Kai Zhang , Xiong Dun , Jiande Sun , Zhanshan Wang , Xinbin Cheng

Real-time 3D reconstruction enables fast dense mapping of the environment which benefits numerous applications, such as navigation or live evaluation of an emergency. In contrast to most real-time capable approaches, our approach does not…

计算机视觉与模式识别 · 计算机科学 2021-04-22 Max Hermann , Boitumelo Ruf , Martin Weinmann

Human driver can easily describe the complex traffic scene by visual system. Such an ability of precise perception is essential for driver's planning. To achieve this, a geometry-aware representation that quantizes the physical 3D scene…

计算机视觉与模式识别 · 计算机科学 2023-06-27 Chonghao Sima , Wenwen Tong , Tai Wang , Li Chen , Silei Wu , Hanming Deng , Yi Gu , Lewei Lu , Ping Luo , Dahua Lin , Hongyang Li

We present a learning-based model to infer the personalized 3D shape of people from a few frames (1-8) of a monocular video in which the person is moving, in less than 10 seconds with a reconstruction accuracy of 5mm. Our model learns to…

计算机视觉与模式识别 · 计算机科学 2019-04-09 Thiemo Alldieck , Marcus Magnor , Bharat Lal Bhatnagar , Christian Theobalt , Gerard Pons-Moll

We introduce a View-Volume convolutional neural network (VVNet) for inferring the occupancy and semantic labels of a volumetric 3D scene from a single depth image. The VVNet concatenates a 2D view CNN and a 3D volume CNN with a…

计算机视觉与模式识别 · 计算机科学 2018-06-15 Yu-Xiao Guo , Xin Tong

Current geometry-based monocular 3D object detection models can efficiently detect objects by leveraging perspective geometry, but their performance is limited due to the absence of accurate depth information. Though this issue can be…

计算机视觉与模式识别 · 计算机科学 2021-07-29 Chenhang He , Jianqiang Huang , Xian-Sheng Hua , Lei Zhang

We are witnessing an explosion of neural implicit representations in computer vision and graphics. Their applicability has recently expanded beyond tasks such as shape generation and image-based rendering to the fundamental problem of…

计算机视觉与模式识别 · 计算机科学 2022-05-26 Jiaming Sun , Xi Chen , Qianqian Wang , Zhengqi Li , Hadar Averbuch-Elor , Xiaowei Zhou , Noah Snavely

Monocular 3D object detection is well-known to be a challenging vision task due to the loss of depth information; attempts to recover depth using separate image-only approaches lead to unstable and noisy depth estimates, harming 3D…

计算机视觉与模式识别 · 计算机科学 2019-05-15 Ivan Barabanau , Alexey Artemov , Evgeny Burnaev , Vyacheslav Murashkin

Custom and natural lighting conditions can be emulated in images of the scene during post-editing. Extraordinary capabilities of the deep learning framework can be utilized for such purpose. Deep image relighting allows automatic photo…

计算机视觉与模式识别 · 计算机科学 2021-06-17 Sourya Dipta Das , Nisarg A. Shah , Saikat Dutta , Himanshu Kumar

In the field of monocular 3D detection, it is common practice to utilize scene geometric clues to enhance the detector's performance. However, many existing works adopt these clues explicitly such as estimating a depth map and…

计算机视觉与模式识别 · 计算机科学 2023-09-27 Junkai Xu , Liang Peng , Haoran Cheng , Hao Li , Wei Qian , Ke Li , Wenxiao Wang , Deng Cai

3D reconstruction from a single 2D image was extensively covered in the literature but relies on depth supervision at training time, which limits its applicability. To relax the dependence to depth we propose SceneRF, a self-supervised…

计算机视觉与模式识别 · 计算机科学 2023-08-28 Anh-Quan Cao , Raoul de Charette

Traditionally, 3D indoor scene reconstruction from posed images happens in two phases: per-image depth estimation, followed by depth merging and surface reconstruction. Recently, a family of methods have emerged that perform reconstruction…

计算机视觉与模式识别 · 计算机科学 2022-09-01 Mohamed Sayed , John Gibson , Jamie Watson , Victor Prisacariu , Michael Firman , Clément Godard

Computer vision tasks are often expected to be executed on compressed images. Classical image compression standards like JPEG 2000 are widely used. However, they do not account for the specific end-task at hand. Motivated by works on…

图像与视频处理 · 电气工程与系统科学 2021-07-27 Alex Golts , Yoav Y. Schechner

We introduce UprightNet, a learning-based approach for estimating 2DoF camera orientation from a single RGB image of an indoor scene. Unlike recent methods that leverage deep learning to perform black-box regression from image to…

计算机视觉与模式识别 · 计算机科学 2019-08-21 Wenqi Xian , Zhengqi Li , Matthew Fisher , Jonathan Eisenmann , Eli Shechtman , Noah Snavely

People live in a 3D world. However, existing works on person re-identification (re-id) mostly consider the semantic representation learning in a 2D space, intrinsically limiting the understanding of people. In this work, we address this…

计算机视觉与模式识别 · 计算机科学 2022-10-26 Zhedong Zheng , Nenggan Zheng , Yi Yang

In this paper, we propose a novel approach, 3D-RecGAN++, which reconstructs the complete 3D structure of a given object from a single arbitrary depth view using generative adversarial networks. Unlike existing work which typically requires…

计算机视觉与模式识别 · 计算机科学 2018-09-11 Bo Yang , Stefano Rosa , Andrew Markham , Niki Trigoni , Hongkai Wen

Robust face reconstruction from monocular image in general lighting conditions is challenging. Methods combining deep neural network encoders with differentiable rendering have opened up the path for very fast monocular reconstruction of…

计算机视觉与模式识别 · 计算机科学 2021-11-23 Abdallah Dib , Cedric Thebault , Junghyun Ahn , Philippe-Henri Gosselin , Christian Theobalt , Louis Chevallier

In this paper, we present a new algorithm for fast, online 3D reconstruction of dynamic scenes using times of arrival of photons recorded by single-photon detector arrays. One of the main challenges in 3D imaging using single-photon lidar…