中文
相关论文

相关论文: MoVi: A Large Multipurpose Motion and Video Datase…

200 篇论文

When we physically interact with our environment using our hands, we touch objects and force them to move: contact and motion are defining properties of manipulation. In this paper, we present an active, bottom-up method for the detection…

计算机视觉与模式识别 · 计算机科学 2019-02-05 Konstantinos Zampogiannis , Kanishka Ganguly , Cornelia Fermuller , Yiannis Aloimonos

The lack of large-scale, labeled data sets impedes progress in developing robust and generalized predictive models for on-body sensor-based human activity recognition (HAR). Labeled data in human activity recognition is scarce and hard to…

计算机视觉与模式识别 · 计算机科学 2020-08-05 Hyeokhyen Kwon , Catherine Tong , Harish Haresamudram , Yan Gao , Gregory D. Abowd , Nicholas D. Lane , Thomas Ploetz

We construct the first markerless deformable interaction dataset recording interactive motions of the hands and deformable objects, called HMDO (Hand Manipulation with Deformable Objects). With our built multi-view capture system, it…

计算机视觉与模式识别 · 计算机科学 2023-01-19 Wei Xie , Zhipeng Yu , Zimeng Zhao , Binghui Zuo , Yangang Wang

We present the first event-based learning approach for motion segmentation in indoor scenes and the first event-based dataset - EV-IMO - which includes accurate pixel-wise motion masks, egomotion and ground truth depth. Our approach is…

计算机视觉与模式识别 · 计算机科学 2020-01-14 Anton Mitrokhin , Chengxi Ye , Cornelia Fermuller , Yiannis Aloimonos , Tobi Delbruck

This work focuses on tracking and understanding human motion using consumer wearable devices, such as VR/AR headsets, smart glasses, cellphones, and smartwatches. These devices provide diverse, multi-modal sensor inputs, including…

计算机视觉与模式识别 · 计算机科学 2025-04-14 Jian Wang , Rishabh Dabral , Diogo Luvizon , Zhe Cao , Lingjie Liu , Thabo Beeler , Christian Theobalt

The volumetric representation of human interactions is one of the fundamental domains in the development of immersive media productions and telecommunication applications. Particularly in the context of the rapid advancement of Extended…

计算机视觉与模式识别 · 计算机科学 2024-02-15 Fatemeh Ghorbani Lohesara , Davi Rabbouni Freitas , Christine Guillemot , Karen Eguiazarian , Sebastian Knorr

Generating accurate descriptions of human actions in videos remains a challenging task for video captioning models. Existing approaches often struggle to capture fine-grained motion details, resulting in vague or semantically inconsistent…

计算机视觉与模式识别 · 计算机科学 2025-10-30 Guorui Song , Guocun Wang , Zhe Huang , Jing Lin , Xuefei Zhe , Jian Li , Haoqian Wang

Real-time human motion reconstruction from a sparse set of (e.g. six) wearable IMUs provides a non-intrusive and economic approach to motion capture. Without the ability to acquire position information directly from IMUs, recent works took…

计算机视觉与模式识别 · 计算机科学 2022-12-12 Yifeng Jiang , Yuting Ye , Deepak Gopinath , Jungdam Won , Alexander W. Winkler , C. Karen Liu

A new event camera dataset, EVIMO2, is introduced that improves on the popular EVIMO dataset by providing more data, from better cameras, in more complex scenarios. As with its predecessor, EVIMO2 provides labels in the form of per-pixel…

计算机视觉与模式识别 · 计算机科学 2022-05-10 Levi Burner , Anton Mitrokhin , Cornelia Fermüller , Yiannis Aloimonos

Neuromorphic sensors, specifically event cameras, revolutionize visual data acquisition by capturing pixel intensity changes with exceptional dynamic range, minimal latency, and energy efficiency, setting them apart from conventional…

计算机视觉与模式识别 · 计算机科学 2024-07-16 Qi Wang , Zhou Xu , Yuming Lin , Jingtao Ye , Hongsheng Li , Guangming Zhu , Syed Afaq Ali Shah , Mohammed Bennamoun , Liang Zhang

Wearable cameras allow to collect images and videos of humans interacting with the world. While human-object interactions have been thoroughly investigated in third person vision, the problem has been understudied in egocentric settings and…

计算机视觉与模式识别 · 计算机科学 2020-10-13 Francesco Ragusa , Antonino Furnari , Salvatore Livatino , Giovanni Maria Farinella

We present a unified perspective on tackling various human-centric video tasks by learning human motion representations from large-scale and heterogeneous data resources. Specifically, we propose a pretraining stage in which a motion…

计算机视觉与模式识别 · 计算机科学 2023-08-15 Wentao Zhu , Xiaoxuan Ma , Zhaoyang Liu , Libin Liu , Wayne Wu , Yizhou Wang

Video object segmentation (VOS) aims to segment specified target objects throughout a video. Although state-of-the-art methods have achieved impressive performance (e.g., 90+% J&F) on benchmarks such as DAVIS and YouTube-VOS, these datasets…

计算机视觉与模式识别 · 计算机科学 2025-09-23 Henghui Ding , Kaining Ying , Chang Liu , Shuting He , Xudong Jiang , Yu-Gang Jiang , Philip H. S. Torr , Song Bai

Fine-grained capturing of 3D HOI boosts human activity understanding and facilitates downstream visual tasks, including action recognition, holistic scene reconstruction, and human motion synthesis. Despite its significance, existing works…

计算机视觉与模式识别 · 计算机科学 2023-12-19 Nan Jiang , Tengyu Liu , Zhexuan Cao , Jieming Cui , Zhiyuan zhang , Yixin Chen , He Wang , Yixin Zhu , Siyuan Huang

Beyond possessing large enough size to feed data hungry machines (eg, transformers), what attributes measure the quality of a dataset? Assuming that the definitions of such attributes do exist, how do we quantify among their relative…

计算机视觉与模式识别 · 计算机科学 2022-04-19 Rajat Modi , Aayush Jung Rana , Akash Kumar , Praveen Tirupattur , Shruti Vyas , Yogesh Singh Rawat , Mubarak Shah

Identifying human behaviors is a challenging research problem due to the complexity and variation of appearances and postures, the variation of camera settings, and view angles. In this paper, we try to address the problem of human behavior…

计算机视觉与模式识别 · 计算机科学 2019-03-08 Eissa Jaber Alreshidi , Mohammad Bilal

Recent advances in world models have demonstrated strong capabilities in simulating physical reality, making them an increasingly important foundation for embodied intelligence. For UAV agents in particular, accurate prediction of complex…

计算机视觉与模式识别 · 计算机科学 2026-04-10 Zile Guo , Zhan Chen , Enze Zhu , Kan Wei , Yongkang Zou , Xiaoxuan Liu , Lei Wang

Human pose estimation faces hurdles in real-world applications due to factors like lighting changes, occlusions, and cluttered environments. We introduce a unique RGB-Thermal Nearly Paired and Annotated 2D Pose Dataset, comprising over…

计算机视觉与模式识别 · 计算机科学 2024-04-17 Avinash Upadhyay , Bhipanshu Dhupar , Manoj Sharma , Ankit Shukla , Ajith Abraham

The creation of large, diverse, high-quality robot manipulation datasets is an important stepping stone on the path toward more capable and robust robotic manipulation policies. However, creating such datasets is challenging: collecting…

机器人学 · 计算机科学 2025-04-23 Alexander Khazatsky , Karl Pertsch , Suraj Nair , Ashwin Balakrishna , Sudeep Dasari , Siddharth Karamcheti , Soroush Nasiriany , Mohan Kumar Srirama , Lawrence Yunliang Chen , Kirsty Ellis , Peter David Fagan , Joey Hejna , Masha Itkina , Marion Lepert , Yecheng Jason Ma , Patrick Tree Miller , Jimmy Wu , Suneel Belkhale , Shivin Dass , Huy Ha , Arhan Jain , Abraham Lee , Youngwoon Lee , Marius Memmel , Sungjae Park , Ilija Radosavovic , Kaiyuan Wang , Albert Zhan , Kevin Black , Cheng Chi , Kyle Beltran Hatch , Shan Lin , Jingpei Lu , Jean Mercat , Abdul Rehman , Pannag R Sanketi , Archit Sharma , Cody Simpson , Quan Vuong , Homer Rich Walke , Blake Wulfe , Ted Xiao , Jonathan Heewon Yang , Arefeh Yavary , Tony Z. Zhao , Christopher Agia , Rohan Baijal , Mateo Guaman Castro , Daphne Chen , Qiuyu Chen , Trinity Chung , Jaimyn Drake , Ethan Paul Foster , Jensen Gao , Vitor Guizilini , David Antonio Herrera , Minho Heo , Kyle Hsu , Jiaheng Hu , Muhammad Zubair Irshad , Donovon Jackson , Charlotte Le , Yunshuang Li , Kevin Lin , Roy Lin , Zehan Ma , Abhiram Maddukuri , Suvir Mirchandani , Daniel Morton , Tony Nguyen , Abigail O'Neill , Rosario Scalise , Derick Seale , Victor Son , Stephen Tian , Emi Tran , Andrew E. Wang , Yilin Wu , Annie Xie , Jingyun Yang , Patrick Yin , Yunchu Zhang , Osbert Bastani , Glen Berseth , Jeannette Bohg , Ken Goldberg , Abhinav Gupta , Abhishek Gupta , Dinesh Jayaraman , Joseph J Lim , Jitendra Malik , Roberto Martín-Martín , Subramanian Ramamoorthy , Dorsa Sadigh , Shuran Song , Jiajun Wu , Michael C. Yip , Yuke Zhu , Thomas Kollar , Sergey Levine , Chelsea Finn

As research on neural volumetric video reconstruction and compression flourishes, there is a need for diverse and realistic datasets, which can be used to develop and validate reconstruction and compression models. However, existing…

计算机视觉与模式识别 · 计算机科学 2026-04-23 Adrian Azzarelli , Ge Gao , Ho Man Kwan , Fan Zhang , Nantheera Anantrasirichai , Ollie Moolan-Feroze , David Bull
‹ 上一页 1 8 9 10 下一页 ›