English
Related papers

Related papers: 1st Place Solution for ICCV 2023 OmniObject3D Chal…

200 papers

In this paper, we show our solution to the Google Landmark Recognition 2021 Competition. Firstly, embeddings of images are extracted via various architectures (i.e. CNN-, Transformer- and hybrid-based), which are optimized by ArcFace loss.…

Computer Vision and Pattern Recognition · Computer Science 2021-10-08 Cheng Xu , Weimin Wang , Shuai Liu , Yong Wang , Yuxiang Tang , Tianling Bian , Yanyu Yan , Qi She , Cheng Yang

Reconstructing 3D objects from extremely sparse views is a long-standing and challenging problem. While recent techniques employ image diffusion models for generating plausible images at novel viewpoints or for distilling pre-trained…

Computer Vision and Pattern Recognition · Computer Science 2023-12-21 Zi-Xin Zou , Weihao Cheng , Yan-Pei Cao , Shi-Sheng Huang , Ying Shan , Song-Hai Zhang

While recent deep neural networks have achieved promising results for 3D reconstruction from a single-view image, these rely on the availability of RGB textures in images and extra information as supervision. In this work, we propose novel…

Computer Vision and Pattern Recognition · Computer Science 2017-01-18 Xinhan Di , Pengqian Yu

In this paper, we introduce a data-efficient instance segmentation method we used in the 2021 VIPriors Instance Segmentation Challenge. Our solution is a modified version of Swin Transformer, based on the mmdetection which is a powerful…

Computer Vision and Pattern Recognition · Computer Science 2022-11-08 Pengyu Chen , Wanhua Li

Accurately predicting the 3D shape of any arbitrary object in any pose from a single image is a key goal of computer vision research. This is challenging as it requires a model to learn a representation that can infer both the visible and…

Computer Vision and Pattern Recognition · Computer Science 2021-09-03 Anh Thai , Stefan Stojanov , Vijay Upadhya , James M. Rehg

Weakly supervised localization aims at finding target object regions using only image-level supervision. However, localization maps extracted from classification networks are often not accurate due to the lack of fine pixel-level…

Computer Vision and Pattern Recognition · Computer Science 2020-08-13 Xiaolin Zhang , Yunchao Wei , Yi Yang

Neural Radiance Fields (NeRF) have demonstrated impressive potential in synthesizing novel views from dense input, however, their effectiveness is challenged when dealing with sparse input. Existing approaches that incorporate additional…

Computer Vision and Pattern Recognition · Computer Science 2023-12-18 Zhangkai Ni , Peiqi Yang , Wenhan Yang , Hanli Wang , Lin Ma , Sam Kwong

In the present work, we propose a Self-supervised COordinate Projection nEtwork (SCOPE) to reconstruct the artifacts-free CT image from a single SV sinogram by solving the inverse tomography imaging problem. Compared with recent related…

Image and Video Processing · Electrical Eng. & Systems 2023-08-14 Qing Wu , Ruimin Feng , Hongjiang Wei , Jingyi Yu , Yuyao Zhang

Recent works in hand-object reconstruction mainly focus on the single-view and dense multi-view settings. On the one hand, single-view methods can leverage learned shape priors to generalise to unseen objects but are prone to inaccuracies…

Computer Vision and Pattern Recognition · Computer Science 2024-05-03 Yik Lung Pang , Changjae Oh , Andrea Cavallaro

We present Rewis3d, a framework that leverages recent advances in feed-forward 3D reconstruction to significantly improve weakly supervised semantic segmentation on 2D images. Obtaining dense, pixel-level annotations remains a costly…

Computer Vision and Pattern Recognition · Computer Science 2026-03-09 Jonas Ernst , Wolfgang Boettcher , Lukas Hoyer , Jan Eric Lenssen , Bernt Schiele

We present a new framework to reconstruct holistic 3D indoor scenes including both room background and indoor objects from single-view images. Existing methods can only produce 3D shapes of indoor objects with limited geometry quality…

Computer Vision and Pattern Recognition · Computer Science 2022-08-09 Haolin Liu , Yujian Zheng , Guanying Chen , Shuguang Cui , Xiaoguang Han

We present an approach for recognizing all objects in a scene and estimating their full pose from an accurate 3D instance-aware semantic reconstruction using an RGB-D camera. Our framework couples convolutional neural networks (CNNs) and a…

Robotics · Computer Science 2019-10-01 Dinh-Cuong Hoang , Todor Stoyanov , Achim J. Lilienthal

This paper studies the challenging two-view 3D reconstruction in a rigorous sparse-view configuration, which is suffering from insufficient correspondences in the input image pairs for camera pose estimation. We present a novel Neural…

Computer Vision and Pattern Recognition · Computer Science 2023-09-14 Bin Tan , Nan Xue , Tianfu Wu , Gui-Song Xia

3D object reconstruction is a fundamental task of many robotics and AI problems. With the aid of deep convolutional neural networks (CNNs), 3D object reconstruction has witnessed a significant progress in recent years. However, possibly due…

Computer Vision and Pattern Recognition · Computer Science 2018-09-11 Hanqing Wang , Jiaolong Yang , Wei Liang , Xin Tong

Monocular 3D object parsing is highly desirable in various scenarios including occlusion reasoning and holistic scene interpretation. We present a deep convolutional neural network (CNN) architecture to localize semantic parts in 2D image…

Computer Vision and Pattern Recognition · Computer Science 2017-04-24 Chi Li , M. Zeeshan Zia , Quoc-Huy Tran , Xiang Yu , Gregory D. Hager , Manmohan Chandraker

Approaches for single-view reconstruction typically rely on viewpoint annotations, silhouettes, the absence of background, multiple views of the same instance, a template shape, or symmetry. We avoid all such supervision and assumptions by…

Computer Vision and Pattern Recognition · Computer Science 2022-07-26 Tom Monnier , Matthew Fisher , Alexei A. Efros , Mathieu Aubry

Scene and object reconstruction is an important problem in robotics, in particular in planning collision-free trajectories or in object manipulation. This paper compares two strategies for the reconstruction of nonvisible parts of the…

Robotics · Computer Science 2025-01-28 Rafał Staszak , Piotr Michałek , Jakub Chudziński , Marek Kopicki , Dominik Belter

Visual relocalization is a key technique to autonomous driving, robotics, and virtual/augmented reality. After decades of explorations, absolute pose regression (APR), scene coordinate regression (SCR), and hierarchical methods (HMs) have…

Computer Vision and Pattern Recognition · Computer Science 2024-04-16 Fei Xue , Ignas Budvytis , Daniel Olmeda Reino , Roberto Cipolla

In this paper, we introduce 3rd place solution for PVUW2023 VSS track. Semantic segmentation is a fundamental task in computer vision with numerous real-world applications. We have explored various image-level visual backbones and…

Computer Vision and Pattern Recognition · Computer Science 2023-06-07 Shijie Chang , Zeqi Hao , Ben Kang , Xiaoqi Zhao , Jiawen Zhu , Zhenyu Chen , Lihe Zhang , Lu Zhang , Huchuan Lu

We propose Filtering Inversion (FINV), a learning framework and optimization process that predicts a renderable 3D object representation from one or few partial views. FINV addresses the challenge of synthesizing novel views of objects from…

Computer Vision and Pattern Recognition · Computer Science 2024-08-20 Fan-Yun Sun , Jonathan Tremblay , Valts Blukis , Kevin Lin , Danfei Xu , Boris Ivanovic , Peter Karkus , Stan Birchfield , Dieter Fox , Ruohan Zhang , Yunzhu Li , Jiajun Wu , Marco Pavone , Nick Haber