中文
相关论文

相关论文: HumanRig: Learning Automatic Rigging for Humanoid …

200 篇论文

Musculoskeletal humanoids are robots that closely mimic the human musculoskeletal system, offering various advantages such as variable stiffness control, redundancy, and flexibility. However, their body structure is complex, and muscle…

机器人学 · 计算机科学 2025-07-08 Kento Kawaharazuka , Takahiro Hattori , Keita Yoneda , Kei Okada

To effectively engage in human society, the ability to adapt, filter information, and make informed decisions in ever-changing situations is critical. As robots and intelligent agents become more integrated into human life, there is a…

StyleGAN generates photorealistic portrait images of faces with eyes, teeth, hair and context (neck, shoulders, background), but lacks a rig-like control over semantic face parameters that are interpretable in 3D, such as face pose,…

计算机视觉与模式识别 · 计算机科学 2020-06-16 Ayush Tewari , Mohamed Elgharib , Gaurav Bharaj , Florian Bernard , Hans-Peter Seidel , Patrick Pérez , Michael Zollhöfer , Christian Theobalt

3D human pose estimation and mesh recovery have attracted widespread research interest in many areas, such as computer vision, autonomous driving, and robotics. Deep learning on 3D human pose estimation and mesh recovery has recently…

计算机视觉与模式识别 · 计算机科学 2024-07-04 Yang Liu , Changzhen Qiu , Zhiyong Zhang

This paper introduces an enormous dataset, HaGRID (HAnd Gesture Recognition Image Dataset), to build a hand gesture recognition (HGR) system concentrating on interaction with devices to manage them. That is why all 18 chosen gestures are…

计算机视觉与模式识别 · 计算机科学 2024-10-10 Alexander Kapitanov , Karina Kvanchiani , Alexander Nagaev , Roman Kraynov , Andrei Makhliarchuk

Magnetic Resonance Fingerprinting (MRF) enables fast quantitative imaging, yet reconstructing high-resolution 3D data remains computationally demanding. Non-Cartesian reconstructions require repeated non-uniform FFTs, and the commonly used…

图像与视频处理 · 电气工程与系统科学 2026-02-13 Yonatan Urman , Mark Nishimura , Daniel Abraham , Xiaozhi Cao , Kawin Setsompop

In the current deep learning paradigm, the amount and quality of training data are as critical as the network architecture and its training details. However, collecting, processing, and annotating real data at scale is difficult, expensive,…

计算机视觉与模式识别 · 计算机科学 2023-09-21 Zheng Dang , Mathieu Salzmann

Accurately estimating the 3D pose and shape is an essential step towards understanding animal behavior, and can potentially benefit many downstream applications, such as wildlife conservation. However, research in this area is held back by…

We present SpineTrack, the first comprehensive dataset for 2D spine pose estimation in unconstrained settings, addressing a crucial need in sports analytics, healthcare, and realistic animation. Existing pose datasets often simplify the…

计算机视觉与模式识别 · 计算机科学 2025-04-14 Muhammad Saif Ullah Khan , Stephan Krauß , Didier Stricker

Teeth localization, segmentation, and labeling in 2D images have great potential in modern dentistry to enhance dental diagnostics, treatment planning, and population-based studies on oral health. However, general instance segmentation…

计算机视觉与模式识别 · 计算机科学 2024-04-02 Bo Zou , Shaofeng Wang , Hao Liu , Gaoyue Sun , Yajie Wang , FeiFei Zuo , Chengbin Quan , Youjian Zhao

Full-body Human inverse rendering based on physically-based rendering aims to acquire high-quality materials, which helps achieve photo-realistic rendering under arbitrary illuminations. This task requires estimating multiple material maps…

计算机视觉与模式识别 · 计算机科学 2025-07-25 Yu Jiang , Jiahao Xia , Jiongming Qin , Yusen Wang , Tuo Cao , Chunxia Xiao

A key challenge in the task of human pose and shape estimation is occlusion, including self-occlusions, object-human occlusions, and inter-person occlusions. The lack of diverse and accurate pose and shape training data becomes a major…

计算机视觉与模式识别 · 计算机科学 2022-03-02 Kaibing Yang , Renshu Gu , Maoyu Wang , Masahiro Toyoura , Gang Xu

We present Human3R, a unified, feed-forward framework for online 4D human-scene reconstruction, in the world frame, from casually captured monocular videos. Unlike previous approaches that rely on multi-stage pipelines, iterative…

计算机视觉与模式识别 · 计算机科学 2026-03-04 Yue Chen , Xingyu Chen , Yuxuan Xue , Anpei Chen , Yuliang Xiu , Gerard Pons-Moll

The facial expression generation capability of humanoid social robots is critical for achieving natural and human-like interactions, playing a vital role in enhancing the fluidity of human-robot interactions and the accuracy of emotional…

机器人学 · 计算机科学 2025-10-28 Yongtong Zhu , Lei Li , Iggy Qian , WenBin Zhou , Ye Yuan , Qingdu Li , Na Liu , Jianwei Zhang

Animating realistic character interactions with the surrounding environment is important for autonomous agents in gaming, AR/VR, and robotics. However, current methods for human motion reconstruction struggle with accurately placing humans…

计算机视觉与模式识别 · 计算机科学 2025-10-20 Joshua Li , Brendan Chharawala , Chang Shu , Xue Bin Peng , Pengcheng Xi

Humans exhibit adaptive, context-sensitive responses to egocentric visual input. However, faithfully modeling such reactions from egocentric video remains challenging due to the dual requirements of strictly causal generation and precise 3D…

计算机视觉与模式识别 · 计算机科学 2026-01-06 Libo Zhang , Zekun Li , Tianyu Li , Zeyu Cao , Rui Xu , Xiaoxiao Long , Wenjia Wang , Jingbo Wang , Yuan Liu , Wenping Wang , Daquan Zhou , Taku Komura , Zhiyang Dou

This paper proposes an end-to-end framework for generating 3D human pose datasets using Neural Radiance Fields (NeRF). Public datasets generally have limited diversity in terms of human poses and camera viewpoints, largely due to the…

计算机视觉与模式识别 · 计算机科学 2023-12-25 Mohsen Gholami , Rabab Ward , Z. Jane Wang

Deformable image registration plays a critical role in various tasks of medical image analysis. A successful registration algorithm, either derived from conventional energy optimization or deep networks requires tremendous efforts from…

计算机视觉与模式识别 · 计算机科学 2023-08-15 Xin Fan , Zi Li , Ziyang Li , Xiaolin Wang , Risheng Liu , Zhongxuan Luo , Hao Huang

Our paper aims to generate diverse and realistic animal motion sequences from textual descriptions, without a large-scale animal text-motion dataset. While the task of text-driven human motion synthesis is already extensively studied and…

计算机视觉与模式识别 · 计算机科学 2023-12-01 Zhangsihao Yang , Mingyuan Zhou , Mengyi Shan , Bingbing Wen , Ziwei Xuan , Mitch Hill , Junjie Bai , Guo-Jun Qi , Yalin Wang

Creating a 360{\deg} parametric model of a human head is a very challenging task. While recent advancements have demonstrated the efficacy of leveraging synthetic data for building such parametric head models, their performance remains…

计算机视觉与模式识别 · 计算机科学 2024-08-02 Yuxiao He , Yiyu Zhuang , Yanwen Wang , Yao Yao , Siyu Zhu , Xiaoyu Li , Qi Zhang , Xun Cao , Hao Zhu