中文
相关论文

相关论文: TexVerse: A Universe of 3D Objects with High-Resol…

200 篇论文

Learning to generate textures for a novel 3D mesh given a collection of 3D meshes and real-world 2D images is an important problem with applications in various domains such as 3D simulation, augmented and virtual reality, gaming,…

计算机视觉与模式识别 · 计算机科学 2024-03-08 Dharma KC , Clayton T. Morrison

Recent advances in modeling 3D objects mostly rely on synthetic datasets due to the lack of large-scale realscanned 3D databases. To facilitate the development of 3D perception, reconstruction, and generation in the real world, we propose…

计算机视觉与模式识别 · 计算机科学 2023-04-12 Tong Wu , Jiarui Zhang , Xiao Fu , Yuxin Wang , Jiawei Ren , Liang Pan , Wayne Wu , Lei Yang , Jiaqi Wang , Chen Qian , Dahua Lin , Ziwei Liu

The quality of the video dataset (image quality, resolution, and fine-grained caption) greatly influences the performance of the video generation model. The growing demand for video applications sets higher requirements for high-quality…

计算机视觉与模式识别 · 计算机科学 2025-06-17 Zhucun Xue , Jiangning Zhang , Teng Hu , Haoyang He , Yinan Chen , Yuxuan Cai , Yabiao Wang , Chengjie Wang , Yong Liu , Xiangtai Li , Dacheng Tao

Motion synthesis for diverse object categories holds great potential for 3D content creation but remains underexplored due to two key challenges: (1) the lack of comprehensive motion datasets that include a wide range of high-quality…

计算机视觉与模式识别 · 计算机科学 2025-07-01 Wonkwang Lee , Jongwon Jeong , Taehong Moon , Hyeon-Jong Kim , Jaehyeon Kim , Gunhee Kim , Byeong-Uk Lee

We present PartNet: a consistent, large-scale dataset of 3D objects annotated with fine-grained, instance-level, and hierarchical 3D part information. Our dataset consists of 573,585 part instances over 26,671 3D models covering 24 object…

计算机视觉与模式识别 · 计算机科学 2018-12-07 Kaichun Mo , Shilin Zhu , Angel X. Chang , Li Yi , Subarna Tripathi , Leonidas J. Guibas , Hao Su

3D scene understanding has been transformed by open-vocabulary language models that enable interaction via natural language. However, at present the evaluation of these representations is limited to datasets with closed-set semantics that…

计算机视觉与模式识别 · 计算机科学 2025-10-15 Christina Kassab , Sacha Morin , Martin Büchner , Matías Mattamala , Kumaraditya Gupta , Abhinav Valada , Liam Paull , Maurice Fallon

There has been increasing interest in smart factories powered by robotics systems to tackle repetitive, laborious tasks. One impactful yet challenging task in robotics-powered smart factory applications is robotic grasping: using robotic…

计算机视觉与模式识别 · 计算机科学 2022-08-31 Yuhao Chen , E. Zhixuan Zeng , Maximilian Gilles , Alexander Wong

Despite the latest remarkable advances in generative modeling, efficient generation of high-quality 3D assets from textual prompts remains a difficult task. A key challenge lies in data scarcity: the most extensive 3D datasets encompass…

计算机视觉与模式识别 · 计算机科学 2024-01-17 Antoine Mercier , Ramin Nakhli , Mahesh Reddy , Rajeev Yasarla , Hong Cai , Fatih Porikli , Guillaume Berger

Transparent objects are ubiquitous in household settings and pose distinct challenges for visual sensing and perception systems. The optical properties of transparent objects leave conventional 3D sensors alone unreliable for object depth…

计算机视觉与模式识别 · 计算机科学 2022-07-22 Xiaotong Chen , Huijie Zhang , Zeren Yu , Anthony Opipari , Odest Chadwicke Jenkins

We present a new object representation, called Dense RepPoints, that utilizes a large set of points to describe an object at multiple levels, including both box level and pixel level. Techniques are proposed to efficiently process these…

计算机视觉与模式识别 · 计算机科学 2020-05-19 Ze Yang , Yinghao Xu , Han Xue , Zheng Zhang , Raquel Urtasun , Liwei Wang , Stephen Lin , Han Hu

Texturing 3D humans with semantic UV maps remains a challenge due to the difficulty of acquiring reasonably unfolded UV. Despite recent text-to-3D advancements in supervising multi-view renderings using large text-to-image (T2I) models,…

计算机视觉与模式识别 · 计算机科学 2024-03-20 Yufei Liu , Junwei Zhu , Junshu Tang , Shijie Zhang , Jiangning Zhang , Weijian Cao , Chengjie Wang , Yunsheng Wu , Dongjin Huang

Synthesizing high-fidelity head avatars is a central problem for computer vision and graphics. While head avatar synthesis algorithms have advanced rapidly, the best ones still face great obstacles in real-world scenarios. One of the vital…

计算机视觉与模式识别 · 计算机科学 2023-05-24 Dongwei Pan , Long Zhuo , Jingtan Piao , Huiwen Luo , Wei Cheng , Yuxin Wang , Siming Fan , Shengqi Liu , Lei Yang , Bo Dai , Ziwei Liu , Chen Change Loy , Chen Qian , Wayne Wu , Dahua Lin , Kwan-Yee Lin

We present AircraftVerse, a publicly available aerial vehicle design dataset. Aircraft design encompasses different physics domains and, hence, multiple modalities of representation. The evaluation of these cyber-physical system (CPS)…

Urban embodied AI agents, ranging from delivery robots to quadrupeds, are increasingly populating our cities, navigating chaotic streets to provide last-mile connectivity. Training such agents requires diverse, high-fidelity urban…

计算机视觉与模式识别 · 计算机科学 2026-03-03 Mingxuan Liu , Honglin He , Elisa Ricci , Wayne Wu , Bolei Zhou

Image-text interleaved data, consisting of multiple images and texts arranged in a natural document format, aligns with the presentation paradigm of internet data and closely resembles human reading habits. Recent studies have shown that…

Neural implicit functions have brought impressive advances to the state-of-the-art of clothed human digitization from multiple or even single images. However, despite the progress, current arts still have difficulty generalizing to unseen…

计算机视觉与模式识别 · 计算机科学 2024-11-06 Zhongjin Luo , Haolin Liu , Chenghong Li , Wanghao Du , Zirong Jin , Wanhu Sun , Yinyu Nie , Weikai Chen , Xiaoguang Han

We consider the problem of scaling deep generative shape models to high-resolution. Drawing motivation from the canonical view representation of objects, we introduce a novel method for the fast up-sampling of 3D objects in voxel space…

计算机视觉与模式识别 · 计算机科学 2018-11-16 Edward Smith , Scott Fujimoto , David Meger

Understanding objects through multiple sensory modalities is fundamental to human perception, enabling cross-sensory integration and richer comprehension. For AI and robotic systems to replicate this ability, access to diverse, high-quality…

计算机视觉与模式识别 · 计算机科学 2025-04-04 Samuel Clarke , Suzannah Wistreich , Yanjie Ze , Jiajun Wu

Recent advances in deep learning, such as neural radiance fields and implicit neural representations, have significantly advanced 3D reconstruction. However, accurately reconstructing objects with complex optical properties, such as metals,…

计算机视觉与模式识别 · 计算机科学 2025-11-04 Zheng Dang , Jialu Huang , Fei Wang , Mathieu Salzmann

In this paper, we propose NeoVerse, a versatile 4D world model that is capable of 4D reconstruction, novel-trajectory video generation, and rich downstream applications. We first identify a common limitation of scalability in current 4D…

计算机视觉与模式识别 · 计算机科学 2026-03-27 Yuxue Yang , Lue Fan , Ziqi Shi , Junran Peng , Feng Wang , Zhaoxiang Zhang