中文
相关论文

相关论文: DogMo: A Large-Scale Multi-View RGB-D Dataset for …

200 篇论文

Success in generative modeling across language, image, and video demonstrates that large, well-curated datasets are the key driver for building capable models. 3D Human motion, however, has lagged behind, constrained by an unsatisfying…

Large-scale high-quality 3D motion datasets with multi-person interactions are crucial for data-driven models in autonomous driving to achieve fine-grained pedestrian interaction understanding in dynamic urban environments. However,…

计算机视觉与模式识别 · 计算机科学 2025-08-14 Guangxun Zhu , Shiyu Fan , Hang Dai , Edmond S. L. Ho

Modern canine applications span medical and service roles, while robotic legged dogs serve as autonomous platforms for high-risk industrial inspection, disaster response, and search and rescue operations. For both, accurate positioning…

机器人学 · 计算机科学 2026-03-10 Gal Versano , Itai Savin , Itzik Klein

Computer vision for animals holds great promise for wildlife research but often depends on large-scale data, while existing collection methods rely on controlled capture setups. Recent data-driven approaches show the potential of…

计算机视觉与模式识别 · 计算机科学 2025-11-04 Brian Nlong Zhao , Jiajun Wu , Shangzhe Wu

We introduce a novel musculoskeletal model of a dog, procedurally generated from accurate 3D muscle meshes. Accompanying this model is a motion capture-based locomotion task compatible with a variety of control algorithms, as well as an…

机器人学 · 计算机科学 2025-07-01 Vittorio La Barbera , Steven Bohez , Leonard Hasenclever , Yuval Tassa , John R. Hutchinson

Estimating the pose of animals can facilitate the understanding of animal motion which is fundamental in disciplines such as biomechanics, neuroscience, ethology, robotics and the entertainment industry. Human pose estimation models have…

计算机视觉与模式识别 · 计算机科学 2021-08-03 Moira Shooter , Charles Malleson , Adrian Hilton

The automatic extraction of animal \reb{3D} pose from images without markers is of interest in a range of scientific fields. Most work to date predicts animal pose from RGB images, based on 2D labelling of joint positions. However, due to…

计算机视觉与模式识别 · 计算机科学 2020-04-17 Sinead Kearney , Wenbin Li , Martin Parsons , Kwang In Kim , Darren Cosker

Monocular 3D animal reconstruction is challenging due to complex articulation, self-occlusion, and fine-scale details such as fur. Existing methods often produce distorted geometry and inconsistent textures due to the lack of articulated 3D…

计算机视觉与模式识别 · 计算机科学 2026-03-10 Shufan Sun , Chenchen Wang , Zongfu Yu

While automatic monitoring and coaching of exercises are showing encouraging results in non-medical applications, they still have limitations such as errors and limited use contexts. To allow the development and assessment of physical…

机器学习 · 计算机科学 2025-01-14 Sao Mai Nguyen , Maxime Devanne , Olivier Remy-Neris , Mathieu Lempereur , André Thepaut

In this paper, we present a novel framework designed to reconstruct long-sequence 3D human motion in the world coordinates from in-the-wild videos with multiple shot transitions. Such long-sequence in-the-wild motions are highly valuable to…

计算机视觉与模式识别 · 计算机科学 2025-03-11 Yuhong Zhang , Guanlin Wu , Ling-Hao Chen , Zhuokai Zhao , Jing Lin , Xiaoke Jiang , Jiamin Wu , Zhuoheng Li , Hao Frank Yang , Haoqian Wang , Lei Zhang

There has been extensive progress in the reconstruction and generation of 4D scenes from monocular casually-captured video. While these tasks rely heavily on known camera poses, the problem of finding such poses using structure-from-motion…

计算机视觉与模式识别 · 计算机科学 2024-12-02 Lily Goli , Sara Sabour , Mark Matthews , Marcus Brubaker , Dmitry Lagun , Alec Jacobson , David J. Fleet , Saurabh Saxena , Andrea Tagliasacchi

The recognition of dynamic and social behavior in animals is fundamental for advancing ethology, ecology, medicine and neuroscience. Recent progress in deep learning has enabled automated behavior recognition from video, yet an accurate…

计算机视觉与模式识别 · 计算机科学 2026-02-24 Lucas Martini , Alexander Lappe , Anna Bognár , Rufin Vogels , Martin A. Giese

This paper studies full-body 3D human motion recovery from head-mounted device signals. Existing diffusion-based methods often rely on global distribution matching, leading to local joint reconstruction errors. We propose MotionGRPO, a…

计算机视觉与模式识别 · 计算机科学 2026-05-13 Nanjie Yao , Junlong Ren , Wenhao Shen , Hao Wang

In this paper, we introduce RoleMotion, a large-scale human motion dataset that encompasses a wealth of role-playing and functional motion data tailored to fit various specific scenes. Existing text datasets are mainly constructed…

计算机视觉与模式识别 · 计算机科学 2025-12-02 Junran Peng , Yiheng Huang , Silei Shen , Zeji Wei , Jingwei Yang , Baojie Wang , Yonghao He , Chuanchen Luo , Man Zhang , Xucheng Yin , Wei Sui

Recent advances in image-based human pose estimation make it possible to capture 3D human motion from a single RGB video. However, the inherent depth ambiguity and self-occlusion in a single view prohibit the recovery of as high-quality…

计算机视觉与模式识别 · 计算机科学 2020-08-20 Junting Dong , Qing Shuai , Yuanqing Zhang , Xian Liu , Xiaowei Zhou , Hujun Bao

Studying animal locomotion improves our understanding of motor control and aids in the treatment of motor impairment. Mice are a premier model of human disease and are the model system of choice for much of basic neuroscience. High frame…

计算机视觉与模式识别 · 计算机科学 2018-02-08 Omid Haji Maghsoudi , Mahdi Alizadeh

We introduce a new benchmark analysis focusing on 3D canine pose estimation from monocular in-the-wild images. A multi-modal dataset 3DDogs-Lab was captured indoors, featuring various dog breeds trotting on a walkway. It includes data from…

计算机视觉与模式识别 · 计算机科学 2024-06-21 Moira Shooter , Charles Malleson , Adrian Hilton

Human Activity Recognition in RGB-D videos has been an active research topic during the last decade. However, no efforts have been found in the literature, for recognizing human activity in RGB-D videos where several performers are…

计算机视觉与模式识别 · 计算机科学 2018-07-10 Snehasis Mukherjee , Leburu Anvitha , T. Mohana Lahari

Recent approaches in depth-based human activity analysis achieved outstanding performance and proved the effectiveness of 3D representation for classification of action classes. Currently available depth-based and RGB+D-based action…

计算机视觉与模式识别 · 计算机科学 2016-04-12 Amir Shahroudy , Jun Liu , Tian-Tsong Ng , Gang Wang

We present DuoMo, a generative method that recovers human motion in world-space coordinates from unconstrained videos with noisy or incomplete observations. Reconstructing such motion requires solving a fundamental trade-off: generalizing…

计算机视觉与模式识别 · 计算机科学 2026-03-04 Yufu Wang , Evonne Ng , Soyong Shin , Rawal Khirodkar , Yuan Dong , Zhaoen Su , Jinhyung Park , Kris Kitani , Alexander Richard , Fabian Prada , Michael Zollhofer
‹ 上一页 1 2 3 10 下一页 ›