English
Related papers

Related papers: CRAG: Can 3D Generative Models Help 3D Assembly?

200 papers

In multi-view human body capture systems, the recovered 3D geometry or even the acquired imagery data can be heavily corrupted due to occlusions, noise, limited field of- view, etc. Direct estimation of 3D pose, body shape or motion on…

Computer Vision and Pattern Recognition · Computer Science 2018-02-02 Zhong Li , Yu Ji , Wei Yang , Jinwei Ye , Jingyi Yu

Self-assembly plays an essential role in many natural processes, involving the formation and evolution of living or non-living structures, and shows potential applications in many emerging domains. In existing research and practice, there…

Multiagent Systems · Computer Science 2021-02-24 Wenjie Chu , Wei Zhang , Haiyan Zhao , Zhi Jin , Hong Mei

Buildings are primary components of cities, often featuring repeated elements such as windows and doors. Traditional 3D building asset creation is labor-intensive and requires specialized skills to develop design rules. Recent generative…

Computer Vision and Pattern Recognition · Computer Science 2024-12-11 Yixuan Li , Xingjian Ran , Linning Xu , Tao Lu , Mulin Yu , Zhenzhi Wang , Yuanbo Xiangli , Dahua Lin , Bo Dai

Imagine a robot that can assemble a functional product from the individual parts presented in any configuration to the robot. Designing such a robotic system is a complex problem which presents several open challenges. To bypass these…

Accurate 6D pose estimation is key for robotic manipulation, enabling precise object localization for tasks like grasping. We present RAG-6DPose, a retrieval-augmented approach that leverages 3D CAD models as a knowledge base by integrating…

Computer Vision and Pattern Recognition · Computer Science 2025-06-24 Kuanning Wang , Yuqian Fu , Tianyu Wang , Yanwei Fu , Longfei Liang , Yu-Gang Jiang , Xiangyang Xue

We introduce 3inGAN, an unconditional 3D generative model trained from 2D images of a single self-similar 3D scene. Such a model can be used to produce 3D "remixes" of a given scene, by mapping spatial latent codes into a 3D volumetric…

Computer Vision and Pattern Recognition · Computer Science 2022-11-29 Animesh Karnewar , Oliver Wang , Tobias Ritschel , Niloy Mitra

3D modeling has long been an important area in computer vision and computer graphics. Recently, thanks to the breakthroughs in neural representations and generative models, we witnessed a rapid development of 3D modeling. 3D human modeling,…

Computer Vision and Pattern Recognition · Computer Science 2024-06-07 Ruihe Wang , Yukang Cao , Kai Han , Kwan-Yee K. Wong

Generating 3D models lies at the core of computer graphics and has been the focus of decades of research. With the emergence of advanced neural representations and generative models, the field of 3D content generation is developing rapidly,…

Computer Vision and Pattern Recognition · Computer Science 2024-02-01 Xiaoyu Li , Qi Zhang , Di Kang , Weihao Cheng , Yiming Gao , Jingbo Zhang , Zhihao Liang , Jing Liao , Yan-Pei Cao , Ying Shan

Deep generative models seek to recover the process with which the observed data was generated. They may be used to synthesize new samples or to subsequently extract representations. Successful approaches in the domain of images are driven…

Computer Vision and Pattern Recognition · Computer Science 2020-07-27 Sjoerd van Steenkiste , Karol Kurach , Jürgen Schmidhuber , Sylvain Gelly

Text- or image-to-3D generators and 3D scanners can now produce 3D assets with high-quality shapes and textures. These assets typically consist of a single, fused representation, like an implicit neural field, a Gaussian mixture, or a mesh,…

Computer Vision and Pattern Recognition · Computer Science 2024-12-31 Minghao Chen , Roman Shapovalov , Iro Laina , Tom Monnier , Jianyuan Wang , David Novotny , Andrea Vedaldi

Large Language Models are increasingly capable of interpreting multimodal inputs to generate complex 3D shapes, yet robust methods to evaluate geometric and structural fidelity remain underdeveloped. This paper introduces a human in the…

Computer Vision and Pattern Recognition · Computer Science 2025-09-10 Ahmed R. Sadik , Mariusz Bujny

Retrieval-Augmented Generation (RAG) systems leverage Large Language Models (LLMs) to generate accurate and reliable responses that are grounded in retrieved context. However, LLMs often generate inconsistent outputs for semantically…

Computation and Language · Computer Science 2025-10-17 Xujun Peng , Anoop Kumar , Jingyu Wu , Parker Glenn , Daben Liu

Retrieval-Augmented Generation (RAG) helps large language models (LLMs) answer knowledge-intensive and time-sensitive questions by conditioning generation on external evidence. However, most RAG systems still retrieve unstructured chunks…

Computation and Language · Computer Science 2026-03-11 Jiashuo Sun , Yixuan Xie , Jimeng Shi , Shaowen Wang , Jiawei Han

Reconstructing complete 3D shapes from incomplete or noisy observations is a fundamentally ill-posed problem that requires balancing measurement consistency with shape plausibility. Existing methods for shape reconstruction can achieve…

Computer Vision and Pattern Recognition · Computer Science 2026-03-31 Linus Härenstam-Nielsen , Dmitrii Pozdeev , Thomas Dagès , Nikita Araslanov , Daniel Cremers

Numerous methods have been proposed for probabilistic generative modelling of 3D objects. However, none of these is able to produce textured objects, which renders them of limited use for practical tasks. In this work, we present the first…

Computer Vision and Pattern Recognition · Computer Science 2020-04-10 Paul Henderson , Vagia Tsiminaki , Christoph H. Lampert

We investigate the problem of estimating the 3D shape of an object defined by a set of 3D landmarks, given their 2D correspondences in a single image. A successful approach to alleviating the reconstruction ambiguity is the 3D deformable…

Computer Vision and Pattern Recognition · Computer Science 2017-01-12 Xiaowei Zhou , Menglong Zhu , Spyridon Leonardos , Kostas Daniilidis

Three-dimensional scene generation holds significant potential in gaming, film, and virtual reality. However, most existing methods adopt a single-step generation process, making it difficult to balance scene complexity with minimal user…

Computer Vision and Pattern Recognition · Computer Science 2025-11-03 Jiacheng Hong , Kunzhen Wu , Mingrui Yu , Yichao Gu , Shengze Xue , Shuangjiu Xiao , Deli Dong

Realistic 3D indoor scene generation is crucial for virtual reality, interior design, embodied intelligence, and scene understanding. While existing methods have made progress in coarse-scale furniture arrangement, they struggle to capture…

Computer Vision and Pattern Recognition · Computer Science 2025-09-05 Xiping Wang , Yuxi Wang , Mengqi Zhou , Junsong Fan , Zhaoxiang Zhang

We investigate the problem of training generative models on a very sparse collection of 3D models. We use geometrically motivated energies to augment and thus boost a sparse collection of example (training) models. We analyze the Hessian of…

Computer Vision and Pattern Recognition · Computer Science 2022-05-02 Sanjeev Muralikrishnan , Siddhartha Chaudhuri , Noam Aigerman , Vladimir Kim , Matthew Fisher , Niloy Mitra

To represent people in mixed reality applications for collaboration and communication, we need to generate realistic and faithful avatar poses. However, the signal streams that can be applied for this task from head-mounted devices (HMDs)…

Computer Vision and Pattern Recognition · Computer Science 2022-03-14 Sadegh Aliakbarian , Pashmina Cameron , Federica Bogo , Andrew Fitzgibbon , Thomas J. Cashman