English
Related papers

Related papers: OASIS: A Large-Scale Dataset for Single Image 3D i…

200 papers

3D face reconstruction is a challenging problem but also an important task in the field of computer vision and graphics. Recently, many researchers put attention to the problem and a large number of articles have been published. Single…

Computer Vision and Pattern Recognition · Computer Science 2021-11-09 Hanxin Wang

Image alignment and image restoration are classical computer vision tasks. However, there is still a lack of datasets that provide enough data to train and evaluate end-to-end deep learning models. Obtaining ground-truth data for image…

Computer Vision and Pattern Recognition · Computer Science 2024-03-05 Monika Kwiatkowski , Simon Matern , Olaf Hellwich

We introduce LAESI, a Synthetic Leaf Dataset of 100,000 synthetic leaf images on millimeter paper, each with semantic masks and surface area labels. This dataset provides a resource for leaf morphology analysis primarily aimed at beech and…

The task of estimating 3D occupancy from surrounding-view images is an exciting development in the field of autonomous driving, following the success of Bird's Eye View (BEV) perception. This task provides crucial 3D attributes of the…

Computer Vision and Pattern Recognition · Computer Science 2023-11-20 Wanshui Gan , Ningkai Mo , Hongbin Xu , Naoto Yokoya

We present the All-Seeing (AS) project: a large-scale data and model for recognizing and understanding everything in the open world. Using a scalable data engine that incorporates human feedback and efficient models in the loop, we create a…

Computer Vision and Pattern Recognition · Computer Science 2023-08-04 Weiyun Wang , Min Shi , Qingyun Li , Wenhai Wang , Zhenhang Huang , Linjie Xing , Zhe Chen , Hao Li , Xizhou Zhu , Zhiguo Cao , Yushi Chen , Tong Lu , Jifeng Dai , Yu Qiao

In continual instruction tuning (CIT) scenarios, where new instruction tuning data continuously arrive in an online streaming manner, training delays from large-scale data significantly hinder real-time adaptation. Data selection can…

Computer Vision and Pattern Recognition · Computer Science 2025-10-10 Minjae Lee , Minhyuk Seo , Tingyu Qu , Tinne Tuytelaars , Jonghyun Choi

Visual navigation and three-dimensional (3D) scene reconstruction are essential for robotics to interact with the surrounding environment. Large-scale scenes and critical camera motions are great challenges facing the research community to…

Computer Vision and Pattern Recognition · Computer Science 2021-09-21 Qi Cai , Lilian Zhang , Yuanxin Wu , Wenxian Yu , Dewen Hu

As 3D Gaussian Splatting (3DGS) gains attention in immersive media and digital content creation, assessing the aesthetics of 3D scenes becomes important in helping creators build more visually compelling 3D content. However, existing…

Computer Vision and Pattern Recognition · Computer Science 2026-05-29 Chuanzhi Xu , Boyu Wei , Haoxian Zhou , Xuanhua Yin , Zihan Deng , Haodong Chen , Qiang Qu , Weidong Cai

Omnidirectional (or 360-degree) images are increasingly being used for 3D applications since they allow the rendering of an entire scene with a single image. Existing works based on neural radiance fields demonstrate successful 3D…

Computer Vision and Pattern Recognition · Computer Science 2024-10-29 Suyoung Lee , Jaeyoung Chung , Jaeyoo Huh , Kyoung Mu Lee

3D reconstruction from 2D inputs, especially for non-rigid objects like humans, presents unique challenges due to the significant range of possible deformations. Traditional methods often struggle with non-rigid shapes, which require…

Computer Vision and Pattern Recognition · Computer Science 2025-05-26 Fahd Alhamazani , Yu-Kun Lai , Paul L. Rosin

The development of generalizable Novel View Synthesis (NVS) models is critically limited by the scarcity of large-scale training data featuring diverse and precise camera trajectories. While real-world captures are photorealistic, they are…

Computer Vision and Pattern Recognition · Computer Science 2026-04-14 Chenhan Jiang , Yu Chen , Qingwen Zhang , Jifei Song , Songcen Xu , Dit-Yan Yeung , Jiankang Deng

We present a dataset of large-scale indoor spaces that provides a variety of mutually registered modalities from 2D, 2.5D and 3D domains, with instance-level semantic and geometric annotations. The dataset covers over 6,000m2 and contains…

Computer Vision and Pattern Recognition · Computer Science 2017-04-07 Iro Armeni , Sasha Sax , Amir R. Zamir , Silvio Savarese

Instance shape reconstruction from a 3D scene involves recovering the full geometries of multiple objects at the semantic instance level. Many methods leverage data-driven learning due to the intricacies of scene complexity and significant…

Computer Vision and Pattern Recognition · Computer Science 2023-12-20 Haolin Liu , Chongjie Ye , Yinyu Nie , Yingfan He , Xiaoguang Han

Reconstruction of a 3D shape from a single 2D image is a classical computer vision problem, whose difficulty stems from the inherent ambiguity of recovering occluded or only partially observed surfaces. Recent methods address this challenge…

Computer Vision and Pattern Recognition · Computer Science 2020-02-04 Yuan Yao , Nico Schertler , Enrique Rosales , Helge Rhodin , Leonid Sigal , Alla Sheffer

Estimating the 6D pose of arbitrary unseen objects from a single reference image is critical for robotics operating in the long-tail of real-world instances. However, this setting is notoriously challenging: 3D models are rarely available,…

Computer Vision and Pattern Recognition · Computer Science 2025-09-10 Zheng Geng , Nan Wang , Shaocong Xu , Chongjie Ye , Bohan Li , Zhaoxi Chen , Sida Peng , Hao Zhao

Reconstructing three-dimensional (3D) structures from two-dimensional (2D) X-ray images is a valuable and efficient technique in medical applications that requires less radiation exposure than computed tomography scans. Recent approaches…

Image and Video Processing · Electrical Eng. & Systems 2025-03-11 Chengrui Zhu , Ryoichi Ishikawa , Masataka Kagesawa , Tomohisa Yuzawa , Toru Watsuji , Takeshi Oishi

Single-view place recognition, that we can define as finding an image that corresponds to the same place as a given query image, is a key capability for autonomous navigation and mapping. Although there has been a considerable amount of…

Computer Vision and Pattern Recognition · Computer Science 2018-08-21 Daniel Olid , José M. Fácil , Javier Civera

This report demonstrates our solution for the Open Images 2018 Challenge. Based on our detailed analysis on the Open Images Datasets (OID), it is found that there are four typical features: large-scale, hierarchical tag system, severe…

Computer Vision and Pattern Recognition · Computer Science 2018-10-16 Yuan Gao , Xingyuan Bu , Yang Hu , Hui Shen , Ti Bai , Xubin Li , Shilei Wen

Learning 3D scene representation from a single-view image is a long-standing fundamental problem in computer vision, with the inherent ambiguity in predicting contents unseen from the input view. Built on the recently proposed 3D Gaussian…

Computer Vision and Pattern Recognition · Computer Science 2024-06-04 Jianghao Shen , Nan Xue , Tianfu Wu

We propose SiM3D, the first benchmark considering the integration of multiview and multimodal information for comprehensive 3D anomaly detection and segmentation (ADS), where the task is to produce a voxel-based Anomaly Volume. Moreover,…

Computer Vision and Pattern Recognition · Computer Science 2025-08-05 Alex Costanzino , Pierluigi Zama Ramirez , Luigi Lella , Matteo Ragaglia , Alessandro Oliva , Giuseppe Lisanti , Luigi Di Stefano
‹ Prev 1 4 5 6 7 8 10 Next ›