中文
相关论文

相关论文: PointOdyssey: A Large-Scale Synthetic Dataset for …

200 篇论文

Most monocular and physics-based human pose tracking methods, while achieving state-of-the-art results, suffer from artifacts when the scene does not have a strictly flat ground plane or when the camera is moving. Moreover, these methods…

计算机视觉与模式识别 · 计算机科学 2025-07-24 Ayce Idil Aytekin , Chuqiao Li , Diogo Luvizon , Rishabh Dabral , Martin Oswald , Marc Habermann , Christian Theobalt

The ultimate goal of Dataset Distillation is to synthesize a small synthetic dataset such that a model trained on this synthetic set will perform equally well as a model trained on the full, real dataset. Until now, no method of Dataset…

计算机视觉与模式识别 · 计算机科学 2024-03-19 Ziyao Guo , Kai Wang , George Cazenavette , Hui Li , Kaipeng Zhang , Yang You

Tracking pixels in videos is typically studied as an optical flow estimation problem, where every pixel is described with a displacement vector that locates it in the next frame. Even though wider temporal context is freely available, prior…

计算机视觉与模式识别 · 计算机科学 2022-07-26 Adam W. Harley , Zhaoyuan Fang , Katerina Fragkiadaki

World engines aim to synthesize long, 3D-consistent videos that support interactive exploration of a scene under user-controlled camera motion. However, existing systems struggle under aggressive 6-DoF trajectories and complex outdoor…

计算机视觉与模式识别 · 计算机科学 2026-03-25 Yu-Cheng Chou , Xingrui Wang , Yitong Li , Jiahao Wang , Hanting Liu , Cihang Xie , Alan Yuille , Junfei Xiao

Video synchronization-aligning multiple video streams capturing the same event from different angles-is crucial for applications such as reality TV show production, sports analysis, surveillance, and autonomous systems. Prior work has…

计算机视觉与模式识别 · 计算机科学 2025-06-23 Yosub Shin , Igor Molybog

Large-scale pre-trained video diffusion models have exhibited remarkable capabilities in diverse video generation. However, existing solutions face several challenges in generating long videos with rich human-scene interactions (HSI),…

计算机视觉与模式识别 · 计算机科学 2026-04-20 Zekun Li , Rui Zhou , Rahul Sajnani , Xiaoyan Cong , Daniel Ritchie , Srinath Sridhar

We study the problem of synthesizing a long-term dynamic video from only a single image. This is challenging since it requires consistent visual content movements given large camera motions. Existing methods either hallucinate inconsistent…

计算机视觉与模式识别 · 计算机科学 2023-08-22 Liao Shen , Xingyi Li , Huiqiang Sun , Juewen Peng , Ke Xian , Zhiguo Cao , Guosheng Lin

Acquiring labeled datasets for 3D human mesh estimation is challenging due to depth ambiguities and the inherent difficulty of annotating 3D geometry from monocular images. Existing datasets are either real, with manually annotated 3D…

计算机视觉与模式识别 · 计算机科学 2026-03-31 Lorenza Prospero , Orest Kupyn , Ostap Viniavskyi , João F. Henriques , Christian Rupprecht

With the emergence of deep learning in the last years, new opportunities arose in Earth observation research. Nevertheless, they also brought with them new challenges. The data-hungry training processes of deep learning models demand large,…

计算机视觉与模式识别 · 计算机科学 2022-04-21 Thorsten Hoeser , Claudia Kuenzer

In today's world, the amount of data produced in every field has increased at an unexpected level. In the face of increasing data, the importance of data processing has increased remarkably. Our resource topic is on the processing of video…

计算机视觉与模式识别 · 计算机科学 2021-06-24 Talha Dilber , Mehmet Serdar Guzel , Erkan Bostanci

Image- and video-based 3D human recovery (i.e., pose and shape estimation) have achieved substantial progress. However, due to the prohibitive cost of motion capture, existing datasets are often limited in scale and diversity. In this work,…

计算机视觉与模式识别 · 计算机科学 2024-09-11 Zhongang Cai , Mingyuan Zhang , Jiawei Ren , Chen Wei , Daxuan Ren , Zhengyu Lin , Haiyu Zhao , Lei Yang , Chen Change Loy , Ziwei Liu

Numerous methods for human activity recognition have been proposed in the past two decades. Many of these methods are based on sparse representation, which describes the whole video content by a set of local features. Trajectories, being…

计算机视觉与模式识别 · 计算机科学 2019-05-15 Pejman Habashi , Boubakeur Boufama , Imran Shafiq Ahmad

Trampoline gymnastics involves extreme human poses and uncommon viewpoints, on which state-of-the art pose estimation models tend to under-perform. We demonstrate that this problem can be addressed by fine-tuning a pose estimation model on…

计算机视觉与模式识别 · 计算机科学 2026-04-03 Léa Drolet-Roy , Victor Nogues , Sylvain Gaudet , Eve Charbonneau , Mickaël Begon , Lama Séoud

Continual learning refers to the ability of humans and animals to incrementally learn over time in a given environment. Trying to simulate this learning process in machines is a challenging task, also due to the inherent difficulty in…

计算机视觉与模式识别 · 计算机科学 2021-09-17 Enrico Meloni , Alessandro Betti , Lapo Faggi , Simone Marullo , Matteo Tiezzi , Stefano Melacci

In this paper, we introduce RoleMotion, a large-scale human motion dataset that encompasses a wealth of role-playing and functional motion data tailored to fit various specific scenes. Existing text datasets are mainly constructed…

计算机视觉与模式识别 · 计算机科学 2025-12-02 Junran Peng , Yiheng Huang , Silei Shen , Zeji Wei , Jingwei Yang , Baojie Wang , Yonghao He , Chuanchen Luo , Man Zhang , Xucheng Yin , Wei Sui

Developing robotic systems capable of robustly executing long-horizon manipulation tasks with human-level dexterity is challenging, as such tasks require both physical dexterity and seamless sequencing of manipulation skills while robustly…

机器人学 · 计算机科学 2025-08-26 Weikang Wan , Jiawei Fu , Xiaodi Yuan , Yifeng Zhu , Hao Su

In this paper, we propose 3DBodyTex.Pose, a dataset that addresses the task of 3D human pose estimation in-the-wild. Generalization to in-the-wild images remains limited due to the lack of adequate datasets. Existent ones are usually…

计算机视觉与模式识别 · 计算机科学 2020-04-22 Renato Baptista , Alexandre Saint , Kassem Al Ismaeil , Djamila Aouada

Video synopsis is an efficient method for condensing surveillance videos. This technique begins with the detection and tracking of objects, followed by the creation of object tubes. These tubes consist of sequences, each containing…

计算机视觉与模式识别 · 计算机科学 2024-09-10 Ramtin Malekpour , M. Mehrdad Morsali , Hoda Mohammadzade

Video generative models have made remarkable progress, yet they often yield visual artifacts that violate grounding in physical dynamics. Recent works such as PhysGen3D tackle single image-to-3D physics through mesh reconstruction and…

计算机视觉与模式识别 · 计算机科学 2026-05-19 Hwidong Kim , Yunho Kim , Tae-Kyun Kim

Video grounding is a fundamental problem in multimodal content understanding, aiming to localize specific natural language queries in an untrimmed video. However, current video grounding datasets merely focus on simple events and are either…

计算机视觉与模式识别 · 计算机科学 2024-08-20 Chaolei Tan , Zihang Lin , Junfu Pu , Zhongang Qi , Wei-Yi Pei , Zhi Qu , Yexin Wang , Ying Shan , Wei-Shi Zheng , Jian-Fang Hu
‹ 上一页 1 8 9 10 下一页 ›