English
Related papers

Related papers: TO-Scene: A Large-scale Dataset for Understanding …

200 papers

Deep neural network models have achieved remarkable progress in 3D scene understanding while trained in the closed-set setting and with full labels. However, the major bottleneck is that these models do not have the capacity to recognize…

Computer Vision and Pattern Recognition · Computer Science 2025-02-20 Kangcheng Liu , Yong-Jin Liu , Baoquan Chen

We present a dataset of large-scale indoor spaces that provides a variety of mutually registered modalities from 2D, 2.5D and 3D domains, with instance-level semantic and geometric annotations. The dataset covers over 6,000m2 and contains…

Computer Vision and Pattern Recognition · Computer Science 2017-04-07 Iro Armeni , Sasha Sax , Amir R. Zamir , Silvio Savarese

Autonomous trucking is a promising technology that can greatly impact modern logistics and the environment. Ensuring its safety on public roads is one of the main duties that requires an accurate perception of the environment. To achieve…

Computer Vision and Pattern Recognition · Computer Science 2024-11-12 Felix Fent , Fabian Kuttenreich , Florian Ruch , Farija Rizwin , Stefan Juergens , Lorenz Lechermann , Christian Nissler , Andrea Perl , Ulrich Voll , Min Yan , Markus Lienkamp

Panorama images have a much larger field-of-view thus naturally encode enriched scene context information compared to standard perspective images, which however is not well exploited in the previous scene understanding methods. In this…

Computer Vision and Pattern Recognition · Computer Science 2021-08-25 Cheng Zhang , Zhaopeng Cui , Cai Chen , Shuaicheng Liu , Bing Zeng , Hujun Bao , Yinda Zhang

The integration of language and 3D perception is critical for embodied AI and robotic systems to perceive, understand, and interact with the physical world. Spatial reasoning, a key capability for understanding spatial relationships between…

Computer Vision and Pattern Recognition · Computer Science 2025-07-11 Jiaxin Huang , Ziwen Li , Hanlve Zhang , Runnan Chen , Xiao He , Yandong Guo , Wenping Wang , Tongliang Liu , Mingming Gong

Understanding scene contexts is crucial for machines to perform tasks and adapt prior knowledge in unseen or noisy 3D environments. As data-driven learning is intractable to comprehensively encapsulate diverse ranges of layouts and open…

Computer Vision and Pattern Recognition · Computer Science 2025-08-01 Junho Kim , Gwangtak Bae , Eun Sun Lee , Young Min Kim

Prior to the deep learning era, shape was commonly used to describe the objects. Nowadays, state-of-the-art (SOTA) algorithms in medical imaging are predominantly diverging from computer vision, where voxel grids, meshes, point clouds, and…

Computer Vision and Pattern Recognition · Computer Science 2025-06-06 Jianning Li , Zongwei Zhou , Jiancheng Yang , Antonio Pepe , Christina Gsaxner , Gijs Luijten , Chongyu Qu , Tiezheng Zhang , Xiaoxi Chen , Wenxuan Li , Marek Wodzinski , Paul Friedrich , Kangxian Xie , Yuan Jin , Narmada Ambigapathy , Enrico Nasca , Naida Solak , Gian Marco Melito , Viet Duc Vu , Afaque R. Memon , Christopher Schlachta , Sandrine De Ribaupierre , Rajnikant Patel , Roy Eagleson , Xiaojun Chen , Heinrich Mächler , Jan Stefan Kirschke , Ezequiel de la Rosa , Patrick Ferdinand Christ , Hongwei Bran Li , David G. Ellis , Michele R. Aizenberg , Sergios Gatidis , Thomas Küstner , Nadya Shusharina , Nicholas Heller , Vincent Andrearczyk , Adrien Depeursinge , Mathieu Hatt , Anjany Sekuboyina , Maximilian Löffler , Hans Liebl , Reuben Dorent , Tom Vercauteren , Jonathan Shapey , Aaron Kujawa , Stefan Cornelissen , Patrick Langenhuizen , Achraf Ben-Hamadou , Ahmed Rekik , Sergi Pujades , Edmond Boyer , Federico Bolelli , Costantino Grana , Luca Lumetti , Hamidreza Salehi , Jun Ma , Yao Zhang , Ramtin Gharleghi , Susann Beier , Arcot Sowmya , Eduardo A. Garza-Villarreal , Thania Balducci , Diego Angeles-Valdez , Roberto Souza , Leticia Rittner , Richard Frayne , Yuanfeng Ji , Vincenzo Ferrari , Soumick Chatterjee , Florian Dubost , Stefanie Schreiber , Hendrik Mattern , Oliver Speck , Daniel Haehn , Christoph John , Andreas Nürnberger , João Pedrosa , Carlos Ferreira , Guilherme Aresta , António Cunha , Aurélio Campilho , Yannick Suter , Jose Garcia , Alain Lalande , Vicky Vandenbossche , Aline Van Oevelen , Kate Duquesne , Hamza Mekhzoum , Jef Vandemeulebroucke , Emmanuel Audenaert , Claudia Krebs , Timo van Leeuwen , Evie Vereecke , Hauke Heidemeyer , Rainer Röhrig , Frank Hölzle , Vahid Badeli , Kathrin Krieger , Matthias Gunzer , Jianxu Chen , Timo van Meegdenburg , Amin Dada , Miriam Balzer , Jana Fragemann , Frederic Jonske , Moritz Rempe , Stanislav Malorodov , Fin H. Bahnsen , Constantin Seibold , Alexander Jaus , Zdravko Marinov , Paul F. Jaeger , Rainer Stiefelhagen , Ana Sofia Santos , Mariana Lindo , André Ferreira , Victor Alves , Michael Kamp , Amr Abourayya , Felix Nensa , Fabian Hörst , Alexander Brehmer , Lukas Heine , Yannik Hanusrichter , Martin Weßling , Marcel Dudda , Lars E. Podleska , Matthias A. Fink , Julius Keyl , Konstantinos Tserpes , Moon-Sung Kim , Shireen Elhabian , Hans Lamecker , Dženan Zukić , Beatriz Paniagua , Christian Wachinger , Martin Urschler , Luc Duong , Jakob Wasserthal , Peter F. Hoyer , Oliver Basu , Thomas Maal , Max J. H. Witjes , Gregor Schiele , Ti-chiun Chang , Seyed-Ahmad Ahmadi , Ping Luo , Bjoern Menze , Mauricio Reyes , Thomas M. Deserno , Christos Davatzikos , Behrus Puladi , Pascal Fua , Alan L. Yuille , Jens Kleesiek , Jan Egger

We present a novel object detection pipeline for localization and recognition in three dimensional environments. Our approach makes use of an RGB-D sensor and combines state-of-the-art techniques from the robotics and computer vision…

Robotics · Computer Science 2017-03-16 Alexander Broad , Brenna Argall

Discovering 3D arrangements of objects from single indoor images is important given its many applications including interior design, content creation, etc. Although heavily researched in the recent years, existing approaches break down…

Computer Vision and Pattern Recognition · Computer Science 2017-12-05 Moos Hueting , Pradyumna Reddy , Vladimir Kim , Ersin Yumer , Nathan Carr , Niloy Mitra

In this work we study indoor scene object placement. Given a 3D indoor scene and an object, the task is to predict placement locations within the scene. Empirical observations of data-driven approaches to the problem show their tendency to…

Graphics · Computer Science 2026-05-05 Adrian Chang , Kai Wang , Yuanbo Li , Manolis Savva , Angel X. Chang , Daniel Ritchie

Traditionally, 3d indoor datasets have generally prioritized scale over ground-truth accuracy in order to obtain improved generalization. However, using these datasets to evaluate dense geometry tasks, such as depth rendering, can be…

Computer Vision and Pattern Recognition · Computer Science 2025-01-07 HyunJun Jung , Weihang Li , Shun-Cheng Wu , William Bittner , Nikolas Brasch , Jifei Song , Eduardo Pérez-Pellitero , Zhensong Zhang , Arthur Moreau , Nassir Navab , Benjamin Busam

Object grasping is critical for many applications, which is also a challenging computer vision problem. However, for the clustered scene, current researches suffer from the problems of insufficient training data and the lacking of…

Computer Vision and Pattern Recognition · Computer Science 2020-01-03 Hao-Shu Fang , Chenxi Wang , Minghao Gou , Cewu Lu

Continual learning refers to the ability of humans and animals to incrementally learn over time in a given environment. Trying to simulate this learning process in machines is a challenging task, also due to the inherent difficulty in…

Computer Vision and Pattern Recognition · Computer Science 2021-09-17 Enrico Meloni , Alessandro Betti , Lapo Faggi , Simone Marullo , Matteo Tiezzi , Stefano Melacci

A proper scene representation is central to the pursuit of spatial intelligence where agents can robustly reconstruct and efficiently understand 3D scenes. A scene representation is either metric, such as landmark maps in 3D reconstruction,…

Computer Vision and Pattern Recognition · Computer Science 2024-11-21 Juexiao Zhang , Gao Zhu , Sihang Li , Xinhao Liu , Haorui Song , Xinran Tang , Chen Feng

We address the problem of discovering 3D parts for objects in unseen categories. Being able to learn the geometry prior of parts and transfer this prior to unseen categories pose fundamental challenges on data-driven shape segmentation…

Computer Vision and Pattern Recognition · Computer Science 2021-09-21 Tiange Luo , Kaichun Mo , Zhiao Huang , Jiarui Xu , Siyu Hu , Liwei Wang , Hao Su

To endow machines with the ability to perceive the real-world in a three dimensional representation as we do as humans is a fundamental and long-standing topic in Artificial Intelligence. Given different types of visual inputs such as…

Computer Vision and Pattern Recognition · Computer Science 2020-10-20 Bo Yang

Automatic scene generation is an essential area of research with applications in robotics, recreation, visual representation, training and simulation, education, and more. This survey provides a comprehensive review of the current…

Computer Vision and Pattern Recognition · Computer Science 2024-10-04 Awal Ahmed Fime , Saifuddin Mahmud , Arpita Das , Md. Sunzidul Islam , Hong-Hoon Kim

Recent perception-generalist approaches based on language models have achieved state-of-the-art results across diverse tasks, including 3D scene layout estimation and 3D object detection, via unified architecture and interface. However,…

Computer Vision and Pattern Recognition · Computer Science 2026-04-01 Ruihong Yin , Xuepeng Shi , Oleksandr Bailo , Marco Manfredi , Theo Gevers

Generating 3D scenes from human motion sequences supports numerous applications, including virtual reality and architectural design. However, previous auto-regression-based human-aware 3D scene generation methods have struggled to…

Computer Vision and Pattern Recognition · Computer Science 2024-08-21 Xiaolin Hong , Hongwei Yi , Fazhi He , Qiong Cao

Spatial reasoning in large-scale 3D environments remains challenging for current vision-language models, which are typically constrained to room-scale scenarios. We introduce H$^2$U3D (Holistic House Understanding in 3D), a 3D visual…

Computer Vision and Pattern Recognition · Computer Science 2025-12-04 Hongpei Zheng , Shijie Li , Yanran Li , Hujun Yin