English
Related papers

Related papers: BrickAnything: Geometry-Conditioned Buildable Bric…

200 papers

Animatable 3D assets, defined as geometry equipped with an articulated skeleton and skinning weights, are fundamental to interactive graphics, embodied agents, and animation production. While recent 3D generative models can synthesize…

3D generative shape modeling is a fundamental research area in computer vision and interactive computer graphics, with many real-world applications. This paper investigates the novel problem of generating 3D shape point cloud geometry from…

Computer Vision and Pattern Recognition · Computer Science 2020-07-17 Kaichun Mo , He Wang , Xinchen Yan , Leonidas J. Guibas

We present a framework for generating physically realizable assembly instructions from natural language descriptions. Unlike unconstrained text-to-3D approaches, our method operates within a discrete parts vocabulary, enforcing geometric…

Artificial Intelligence · Computer Science 2025-12-19 David Noever

Vision-language models (VLMs) commonly formulate visual grounding and detection as a coordinate-token generation problem, serializing each 2D box into multiple 1D tokens that are learned and decoded largely independently. This…

Computer Vision and Pattern Recognition · Computer Science 2026-05-28 Shihao Wang , Shilong Liu , Yuanguo Kuang , Xinyu Wei , Yangzhou Liu , Zhiqi Li , Yunze Man , Guo Chen , Andrew Tao , Guilin Liu , Jan Kautz , Lei Zhang , Zhiding Yu

Aggregating base elements into rigid objects such as furniture or sculptures is a great way for designers to convey a specific look and feel. Unfortunately, there is no existing solution to help model structurally sound aggregates. The…

Graphics · Computer Science 2018-11-08 Jérémie Dumas , Jonàs Martínez , Sylvain Lefebvre , Li-Yi Wei

We propose a method to detect and reconstruct multiple 3D objects from a single RGB image. The key idea is to optimize for detection, alignment and shape jointly over all objects in the RGB image, while focusing on realistic and physically…

Computer Vision and Pattern Recognition · Computer Science 2021-06-23 Francis Engelmann , Konstantinos Rematas , Bastian Leibe , Vittorio Ferrari

While recent 3D generative models can produce high-quality texture images, they often fail to capture human preferences or meet task-specific requirements. Moreover, a core challenge in the 3D texture generation domain is that most existing…

Computer Vision and Pattern Recognition · Computer Science 2025-12-10 AmirHossein Zamani , Tianhao Xie , Amir G. Aghdam , Tiberiu Popa , Eugene Belilovsky

Data-driven machine learning methods have the potential to dramatically accelerate the rate of materials design over conventional human-guided approaches. These methods would help identify or, in the case of generative models, even create…

Materials Science · Physics 2022-07-28 Victor Fung , Shuyi Jia , Jiaxin Zhang , Sirui Bi , Junqi Yin , P. Ganesh

Inferring step-wise actions to assemble 3D objects with primitive bricks from images is a challenging task due to complex constraints and the vast number of possible combinations. Recent studies have demonstrated promising results on…

Computer Vision and Pattern Recognition · Computer Science 2024-07-23 Mengqi Guo , Chen Li , Yuyang Zhao , Gim Hee Lee

Shape-based virtual screening is widely employed in ligand-based drug design to search chemical libraries for molecules with similar 3D shapes yet novel 2D chemical structures compared to known ligands. 3D deep generative models have the…

Chemical Physics · Physics 2022-10-12 Keir Adams , Connor W. Coley

Accurate object geometry estimation is essential for many downstream tasks, including robotic manipulation and physical interaction. Although vision is the dominant modality for shape perception, it becomes unreliable under occlusions or…

Computer Vision and Pattern Recognition · Computer Science 2026-04-13 Langzhe Gu , Hung-Jui Huang , Mohamad Qadri , Michael Kaess , Wenzhen Yuan

Generative models have achieved success in producing semantically plausible 2D images, but it remains challenging in 3D generation due to the absence of spatial geometry constraints. Typically, existing methods utilize geometric features as…

Computer Vision and Pattern Recognition · Computer Science 2026-03-09 Haonan Wang , Hanyu Zhou , Haoyue Liu , Tao Gu , Luxin Yan

Protein structure tokenization converts 3D structures into discrete or vectorized representations, enabling the integration of structural and sequence data. Despite many recent works on structure tokenization, the properties of the…

Machine Learning · Computer Science 2025-11-14 Zijing Liu , Bin Feng , He Cao , Yu Li

Generating complete 3D objects under partial occlusions (i.e., amodal scenarios) is a practically important yet challenging problem, as large portions of object geometry are unobserved in real-world scenarios. Existing approaches either…

Computer Vision and Pattern Recognition · Computer Science 2026-03-17 Junwei Zhou , Yu-Wing Tai

This paper presents a novel decoder-based approach for generating manufacturable 3D structures optimized for additive manufacturing. We introduce a deep learning framework that decodes latent representations into geometrically valid,…

Computer Vision and Pattern Recognition · Computer Science 2026-01-14 Abhishek Kumar

Shape primitive abstraction, which decomposes complex 3D shapes into simple geometric elements, plays a crucial role in human visual cognition and has broad applications in computer vision and graphics. While recent advances in 3D content…

Graphics · Computer Science 2025-05-08 Jingwen Ye , Yuze He , Yanning Zhou , Yiqin Zhu , Kaiwen Xiao , Yong-Jin Liu , Wei Yang , Xiao Han

We introduce a new approach to high-fidelity 3D scene reconstruction from multi-view RGB images that tightly couples reconstruction with a strong generative 3D prior. We cast scene reconstruction as conditional 3D generation over a set of…

Computer Vision and Pattern Recognition · Computer Science 2026-05-25 Katharina Schmid , Nicolas von Lützow , Jozef Hladký , Angela Dai , Matthias Nießner

The automated generation of interactive 3D cities is a critical challenge with broad applications in autonomous driving, virtual reality, and embodied intelligence. While recent advances in generative models and procedural techniques have…

Computer Vision and Pattern Recognition · Computer Science 2026-03-02 Zishan Liu , Zecong Tang , RuoCheng Wu , Xinzhe Zheng , Jingyu Hu , Ka-Hei Hui , Haoran Xie , Bo Dai , Zhengzhe Liu

Previous representation and generation approaches for the B-rep relied on graph-based representations that disentangle geometric and topological features through decoupled computational pipelines, thereby precluding the application of…

Computer Vision and Pattern Recognition · Computer Science 2026-03-31 Jiahao Li , Yunpeng Bai , Yongkang Dai , Hao Guo , Hongping Gan , Yilei Shi

Although LEGO sets have entertained generations of children and adults, the challenge of designing customized builds matching the complexity of real-world or imagined scenes remains too great for the average enthusiast. In order to make…

Computer Vision and Pattern Recognition · Computer Science 2021-08-20 Kyle Lennon , Katharina Fransen , Alexander O'Brien , Yumeng Cao , Matthew Beveridge , Yamin Arefeen , Nikhil Singh , Iddo Drori