English
Related papers

Related papers: Guiding Diffusion-Based Articulated Object Generat…

200 papers

Lidar point cloud synthesis based on generative models offers a promising solution to augment deep learning pipelines, particularly when real-world data is scarce or lacks diversity. By enabling flexible object manipulation, this synthesis…

Computer Vision and Pattern Recognition · Computer Science 2025-07-01 Zhengkang Xiang , Zizhao Li , Amir Khodabandeh , Kourosh Khoshelham

Curating datasets for object segmentation is a difficult task. With the advent of large-scale pre-trained generative models, conditional image generation has been given a significant boost in result quality and ease of use. In this paper,…

Computer Vision and Pattern Recognition · Computer Science 2023-09-06 Mischa Dombrowski , Hadrien Reynaud , Matthew Baugh , Bernhard Kainz

Virtual try-on can significantly improve the garment shopping experiences in both online and in-store scenarios, attracting broad interest in computer vision. However, to achieve high-fidelity try-on performance, most state-of-the-art…

Computer Vision and Pattern Recognition · Computer Science 2024-02-06 Yunfang Niu , Dong Yi , Lingxiang Wu , Zhiwei Liu , Pengxiang Cai , Jinqiao Wang

Generating natural hand-object interactions in 3D is challenging as the resulting hand and object motions are expected to be physically plausible and semantically meaningful. Furthermore, generalization to unseen objects is hindered by the…

Computer Vision and Pattern Recognition · Computer Science 2024-12-24 Sammy Christen , Shreyas Hampali , Fadime Sener , Edoardo Remelli , Tomas Hodan , Eric Sauser , Shugao Ma , Bugra Tekin

This paper focuses on motion prediction for point cloud sequences in the challenging case of deformable 3D objects, such as human body motion. First, we investigate the challenges caused by deformable shapes and complex motions present in…

Computer Vision and Pattern Recognition · Computer Science 2023-07-20 Pedro Gomes , Silvia Rossi , Laura Toni

Recent advances in generative modeling with diffusion processes (DPs) enabled breakthroughs in image synthesis. Despite impressive image quality, these models have various prompt compliance problems, including low recall in generating…

Computer Vision and Pattern Recognition · Computer Science 2024-10-30 Deepak Sridhar , Abhishek Peri , Rohith Rachala , Nuno Vasconcelos

Text-to-image generation models have revolutionized content creation, but diffusion-based vision-language models still face challenges in precisely controlling the shape, appearance, and positional placement of objects in generated images…

Computer Vision and Pattern Recognition · Computer Science 2024-12-19 Shan Yang

Understanding and generating the fine-grained structure of objects -- such as birds with species-specific beaks, wings, and tails -- is a long-standing challenge in computer vision. We propose Chirpy3D, a part-aware multi-view diffusion…

Computer Vision and Pattern Recognition · Computer Science 2026-05-28 Kam Woh Ng , Jing Yang , Jia Wei Sii , Chee Seng Chan , Jiankang Deng , Yi-Zhe Song , Tao Xiang , Xiatian Zhu

Predicting and generating human hand grasp over objects is critical for animation and robotic tasks. In this work, we focus on generating both the hand and objects in a grasp by a single diffusion model. Our proposed Joint Hand-Object…

Computer Vision and Pattern Recognition · Computer Science 2026-01-28 Jinkun Cao , Jingyuan Liu , Kris Kitani , Yi Zhou

We introduce the Quartet of Diffusions, a structure-aware point cloud generation framework that explicitly models part composition and symmetry. Unlike prior methods that treat shape generation as a holistic process or only support part…

Computer Vision and Pattern Recognition · Computer Science 2026-01-29 Chenliang Zhou , Fangcheng Zhong , Weihao Xia , Albert Miao , Canberk Baykal , Cengiz Oztireli

3D human generation is an important problem with a wide range of applications in computer vision and graphics. Despite recent progress in generative AI such as diffusion models or rendering methods like Neural Radiance Fields or Gaussian…

Computer Vision and Pattern Recognition · Computer Science 2025-06-06 Maksym Ivashechkin , Oscar Mendez , Richard Bowden

Generating physically realistic 3D molecular structures remains a core challenge in molecular generative modeling. While diffusion models equipped with equivariant neural networks have made progress in capturing molecular geometries, they…

Machine Learning · Computer Science 2025-08-25 Zhijian Zhou , Junyi An , Zongkai Liu , Yunfei Shi , Xuan Zhang , Fenglei Cao , Chao Qu , Yuan Qi

The problem of text-guided image generation is a complex task in Computer Vision, with various applications, including creating visually appealing artwork and realistic product images. One popular solution widely used for this task is the…

Computer Vision and Pattern Recognition · Computer Science 2023-04-04 Halil Faruk Karagoz , Gulcin Baykal , Irem Arikan Eksi , Gozde Unal

In this paper, we present a novel shape reconstruction method leveraging diffusion model to generate 3D sparse point cloud for the object captured in a single RGB image. Recent methods typically leverage global embedding or local…

Computer Vision and Pattern Recognition · Computer Science 2023-08-16 Yan Di , Chenyangguang Zhang , Pengyuan Wang , Guangyao Zhai , Ruida Zhang , Fabian Manhardt , Benjamin Busam , Xiangyang Ji , Federico Tombari

3D object detection is essential for understanding 3D scenes. Contemporary techniques often require extensive annotated training data, yet obtaining point-wise annotations for point clouds is time-consuming and laborious. Recent…

Computer Vision and Pattern Recognition · Computer Science 2024-08-02 Jiacheng Deng , Jiahao Lu , Tianzhu Zhang

In this paper, we study the problem of task-oriented grasp synthesis from partial point cloud data using an eye-in-hand camera configuration. In task-oriented grasp synthesis, a grasp has to be selected so that the object is not lost during…

Robotics · Computer Science 2023-09-22 Aditya Patankar , Khiem Phi , Dasharadhan Mahalingam , Nilanjan Chakraborty , IV Ramakrishnan

LiDAR is an important method for autonomous driving systems to sense the environment. The point clouds obtained by LiDAR typically exhibit sparse and irregular distribution, thus posing great challenges to the detection of 3D objects,…

Computer Vision and Pattern Recognition · Computer Science 2020-10-28 Tai Wang , Xinge Zhu , Dahua Lin

Denoising diffusion probabilistic models (DDPMs) have achieved impressive performance on various image generation tasks, including image super-resolution. By learning to reverse the process of gradually diffusing the data distribution into…

Image and Video Processing · Electrical Eng. & Systems 2023-07-25 Kai Zhao , Alex Ling Yu Hung , Kaifeng Pang , Haoxin Zheng , Kyunghyun Sung

Articulated objects are common in the real world, yet modeling their structure and motion remains a challenging task for 3D reconstruction methods. In this work, we introduce Part$^{2}$GS, a novel framework for modeling articulated digital…

Computer Vision and Pattern Recognition · Computer Science 2026-04-10 Tianjiao Yu , Vedant Shah , Muntasir Wahed , Ying Shen , Kiet A. Nguyen , Ismini Lourentzou

Single-image point cloud reconstruction must infer complete 3D geometry, including occluded parts, from a single RGB image. While diffusion-based reconstructors achieve high accuracy, they typically require many denoising iterations,…

Computer Vision and Pattern Recognition · Computer Science 2026-04-28 Yuta Baba , Keiji Yanai
‹ Prev 1 8 9 10 Next ›