English
Related papers

Related papers: Artiverse: A Diverse and Physically Grounded Datas…

200 papers

Generating grasps for a dexterous hand often requires numerous grasping annotations. However, annotating high DoF dexterous hand poses is quite challenging. Especially for functional grasps, requiring the hand to grasp the object in a…

Robotics · Computer Science 2024-10-25 Rina Wu , Tianqiang Zhu , Xiangbo Lin , Yi Sun

Representing articulated objects remains a difficult problem within the field of robotics. Objects such as pliers, clamps, or cabinets require representations that capture not only geometry and color information, but also part seperation,…

Robotics · Computer Science 2025-06-17 Stanley Lewis , Vishal Chandra , Tom Gao , Odest Chadwicke Jenkins

Autonomous bin picking poses significant challenges to vision-driven robotic systems given the complexity of the problem, ranging from various sensor modalities, to highly entangled object layouts, to diverse item properties and gripper…

Computer Vision and Pattern Recognition · Computer Science 2022-08-09 Maximilian Gilles , Yuhao Chen , Tim Robin Winter , E. Zhixuan Zeng , Alexander Wong

This paper focuses on the challenging problem of 3D pose estimation of a diverse spectrum of articulated objects from single depth images. A novel structured prediction approach is considered, where 3D poses are represented as skeletal…

Computer Vision and Pattern Recognition · Computer Science 2016-12-05 Yu Zhang , Chi Xu , Li Cheng

Generalizable articulated object manipulation is essential for home-assistant robots. Recent efforts focus on imitation learning from demonstrations or reinforcement learning in simulation, however, due to the prohibitive costs of…

Robotics · Computer Science 2024-02-22 Wenke Xia , Dong Wang , Xincheng Pang , Zhigang Wang , Bin Zhao , Di Hu , Xuelong Li

Multimodal object recognition is still an emerging field. Thus, publicly available datasets are still rare and of small size. This dataset was developed to help fill this void and presents multimodal data for 63 objects with some visual and…

Current 3D visual grounding tasks only process sentence level detection or segmentation, which critically fails to leverage the rich compositional contextual reasonings within natural language expressions. To address this challenge, we…

Computer Vision and Pattern Recognition · Computer Science 2026-03-04 Qi Chen , Changli Wu , Jiayi Ji , Yiwei Ma , Liujuan Cao

Reconstructing 3D visuals from functional Magnetic Resonance Imaging (fMRI) data, introduced as Recon3DMind, is of significant interest to both cognitive neuroscience and computer vision. To advance this task, we present the fMRI-3D…

Computer Vision and Pattern Recognition · Computer Science 2025-08-15 Jianxiong Gao , Yanwei Fu , Yuqian Fu , Yun Wang , Xuelin Qian , Jianfeng Feng

Accurate and efficient object detection is crucial for safe and efficient operation of earth-moving equipment in mining. Traditional 2D image-based methods face limitations in dynamic and complex mine environments. To overcome these…

Computer Vision and Pattern Recognition · Computer Science 2023-12-12 Mehala Balamurali , Ehsan Mihankhah

Recently, there has been growing interest in developing learning-based methods to detect and utilize salient semi-global or global structures, such as junctions, lines, planes, cuboids, smooth surfaces, and all types of symmetries, for 3D…

Computer Vision and Pattern Recognition · Computer Science 2020-07-20 Jia Zheng , Junfei Zhang , Jing Li , Rui Tang , Shenghua Gao , Zihan Zhou

In the era of foundation models, achieving a unified understanding of different dynamic objects through a single network has the potential to empower stronger spatial intelligence. Moreover, accurate estimation of animal pose and shape…

Computer Vision and Pattern Recognition · Computer Science 2025-11-18 Liang An , Jin Lyu , Li Lin , Pujin Cheng , Yebin Liu , Xiaoying Tang

The increasing production of waste, driven by population growth, has created challenges in managing and recycling materials effectively. Manual waste sorting is a common practice; however, it remains inefficient for handling large-scale…

Computer Vision and Pattern Recognition · Computer Science 2026-01-08 Sara Inácio , Hugo Proença , João C. Neves

Endowing robots with the ability to rearrange various large and heavy objects, such as furniture, can substantially alleviate human workload. However, this task is extremely challenging due to the need to interact with diverse objects and…

Robotics · Computer Science 2026-02-05 Zhihai Bi , Yushan Zhang , Kai Chen , Guoyang Zhao , Yulin Li , Jun Ma

We introduce D3D-HOI: a dataset of monocular videos with ground truth annotations of 3D object pose, shape and part motion during human-object interactions. Our dataset consists of several common articulated objects captured from diverse…

Computer Vision and Pattern Recognition · Computer Science 2021-08-20 Xiang Xu , Hanbyul Joo , Greg Mori , Manolis Savva

Being heavily reliant on animals, it is our ethical obligation to improve their well-being by understanding their needs. Several studies show that animal needs are often expressed through their faces. Though remarkable progress has been…

Computer Vision and Pattern Recognition · Computer Science 2019-09-12 Muhammad Haris Khan , John McDonagh , Salman Khan , Muhammad Shahabuddin , Aditya Arora , Fahad Shahbaz Khan , Ling Shao , Georgios Tzimiropoulos

Rigging and skinning are essential steps to create realistic 3D animations, often requiring significant expertise and manual effort. Traditional attempts at automating these processes rely heavily on geometric heuristics and often struggle…

Graphics · Computer Science 2025-07-08 Yufan Deng , Yuhao Zhang , Chen Geng , Shangzhe Wu , Jiajun Wu

Metaverse platforms are rapidly evolving to provide immersive spaces for user interaction and content creation. However, the generation of dynamic and interactive 3D objects remains challenging due to the need for advanced 3D modeling and…

Human-Computer Interaction · Computer Science 2025-05-01 Ryutaro Kurai , Takefumi Hiraki , Yuichi Hiroi , Yutaro Hirao , Monica Perusquía-Hernández , Hideaki Uchiyama , Kiyoshi Kiyokawa

Parametric Computer-Aided Design (CAD) of articulated assemblies is essential for product development, yet generating these multi-part, movable models from high-level descriptions remains unexplored. To address this, we propose ArtiCAD, the…

Computer Vision and Pattern Recognition · Computer Science 2026-04-15 Yuan Shui , Yandong Guan , Zhanwei Zhang , Juncheng Hu , Jing Zhang , Dong Xu , Qian Yu

The advancement of Embodied AI heavily relies on large-scale, simulatable 3D scene datasets characterized by scene diversity and realistic layouts. However, existing datasets typically suffer from limitations in data scale or diversity,…

Computer Vision and Pattern Recognition · Computer Science 2026-04-29 Weipeng Zhong , Peizhou Cao , Yichen Jin , Li Luo , Wenzhe Cai , Jingli Lin , Hanqing Wang , Zhaoyang Lyu , Tai Wang , Bo Dai , Xudong Xu , Jiangmiao Pang

The studies of human clothing for digital avatars have predominantly relied on synthetic datasets. While easy to collect, synthetic data often fall short in realism and fail to capture authentic clothing dynamics. Addressing this gap, we…

Computer Vision and Pattern Recognition · Computer Science 2024-04-30 Wenbo Wang , Hsuan-I Ho , Chen Guo , Boxiang Rong , Artur Grigorev , Jie Song , Juan Jose Zarate , Otmar Hilliges