中文
相关论文

相关论文: SyncUp: Vision-based Practice Support for Synchron…

200 篇论文

We present AlignNet, a model that synchronizes videos with reference audios under non-uniform and irregular misalignments. AlignNet learns the end-to-end dense correspondence between each frame of a video and an audio. Our method is…

计算机视觉与模式识别 · 计算机科学 2020-02-13 Jianren Wang , Zhaoyuan Fang , Hang Zhao

Human pose detection systems based on state-of-the-art DNNs are on the go to be extended, adapted and re-trained to fit the application domain of specific sports. Therefore, plenty of noisy pose data will soon be available from videos…

计算机视觉与模式识别 · 计算机科学 2020-04-22 Rainer Lienhart , Moritz Einfalt , Dan Zecha

Generative models for audio-conditioned dance motion synthesis map music features to dance movements. Models are trained to associate motion patterns to audio patterns, usually without an explicit knowledge of the human body. This approach…

计算机视觉与模式识别 · 计算机科学 2022-07-25 Davide Moltisanti , Jinyi Wu , Bo Dai , Chen Change Loy

This letter presents a physical human-robot interaction scenario in which a robot guides and performs the role of a teacher within a defined dance training framework. A combined cognitive and physical feedback of performance is proposed for…

机器人学 · 计算机科学 2022-12-15 Diego Felipe Paez Granados , Breno A. Yamamoto , Hiroko Kamide , Jun Kinugawa , Kazuhiro Kosuge

We present SyncLight, a method to enable consistent, parametric control over light sources across multiple uncalibrated views of a static scene conditioned on a single view. While single-view relighting has advanced significantly, existing…

计算机视觉与模式识别 · 计算机科学 2026-05-15 David Serrano-Lozano , Anand Bhattad , Luis Herranz , Jean-François Lalonde , Javier Vazquez-Corral

Controllable video generation (CVG) has advanced rapidly, yet current systems falter when more than one actor must move, interact, and exchange positions under noisy control signals. We address this gap with DanceTogether, the first…

计算机视觉与模式识别 · 计算机科学 2025-05-26 Junhao Chen , Mingjin Chen , Jianjin Xu , Xiang Li , Junting Dong , Mingze Sun , Puhua Jiang , Hongxiang Li , Yuhang Yang , Hao Zhao , Xiaoxiao Long , Ruqi Huang

3D human body shape and pose estimation from RGB images is a challenging problem with potential applications in augmented/virtual reality, healthcare and fitness technology and virtual retail. Recent solutions have focused on three types of…

计算机视觉与模式识别 · 计算机科学 2024-01-31 Darshan Venkatrayappa , Alain Tremeau , Damien Muselet , Philippe Colantoni

In this paper we address the issue of photo galleries synchronization, where pictures related to the same event are collected by different users. Existing solutions to address the problem are usually based on unrealistic assumptions, like…

多媒体 · 计算机科学 2017-01-17 E. Sansone , K. Apostolidis , N. Conci , G. Boato , V. Mezaris , F. G. B. De Natale

Text-to-video and image-to-video generation have made rapid progress in visual quality, but they remain limited in controlling the precise timing of motion. In contrast, audio provides temporal cues aligned with video motion, making it a…

计算机视觉与模式识别 · 计算机科学 2025-09-29 Jibin Song , Mingi Kwon , Jaeseok Jeong , Youngjung Uh

Understanding the underlying semantics of performing arts like dance is a challenging task. Dance is multimedia in nature and spans over time as well as space. Capturing and analyzing the multimedia content of the dance is useful for the…

计算机视觉与模式识别 · 计算机科学 2019-09-25 Tanwi Mallick , Partha Pratim Das , Arun Kumar Majumdar

Multi-camera surveillance has been an active research topic for understanding and modeling scenes. Compared to a single camera, multi-cameras provide larger field-of-view and more object cues, and the related applications are multi-view…

计算机视觉与模式识别 · 计算机科学 2022-05-03 Qi Zhang , Antoni B. Chan

We present DanceAnyWay, a generative learning method to synthesize beat-guided dances of 3D human characters synchronized with music. Our method learns to disentangle the dance movements at the beat frames from the dance movements at all…

声音 · 计算机科学 2024-11-26 Aneesh Bhattacharya , Manas Paranjape , Uttaran Bhattacharya , Aniket Bera

Good posture and form are essential for safe and productive exercising. Even in gym settings, trainers may not be readily available for feedback. Rehabilitation therapies and fitness workouts can thus benefit from recommender systems that…

人工智能 · 计算机科学 2023-10-12 Abhishek Jaiswal , Gautam Chauhan , Nisheeth Srivastava

Humans naturally integrate vision and haptics for robust object perception during manipulation. The loss of either modality significantly degrades performance. Inspired by this multisensory integration, prior object pose estimation research…

机器人学 · 计算机科学 2025-09-12 Hongyu Li , Mingxi Jia , Tuluhan Akbulut , Yu Xiang , George Konidaris , Srinath Sridhar

Video stabilization is a longstanding computer vision problem, particularly pixel-level synthesis solutions for video stabilization which synthesize full frames add to the complexity of this task. These techniques aim to stabilize videos by…

计算机视觉与模式识别 · 计算机科学 2024-04-10 Muhammad Kashif Ali , Eun Woo Im , Dongjin Kim , Tae Hyun Kim

Online dance tutorials have gained widespread popularity. However, many novices encounter difficulties when dance motion complexity exceeds their skill level, potentially leading to discouragement. This study explores dance motion…

人机交互 · 计算机科学 2026-04-14 Hyunyoung Han , Murad Eynizada , Son Xuan Nghiem , Sang Ho Yoon

In order to be effective teammates, robots need to be able to understand high-level human behavior to recognize, anticipate, and adapt to human motion. We have designed a new approach to enable robots to perceive human group motion in…

机器人学 · 计算机科学 2016-11-15 Tariq Iqbal , Samantha Rack , Laurel D. Riek

Editing portrait videos is a challenging task that requires flexible yet precise control over a wide range of modifications, such as appearance changes, expression edits, or the addition of objects. The key difficulty lies in preserving the…

计算机视觉与模式识别 · 计算机科学 2025-12-03 Sagi Polaczek , Or Patashnik , Ali Mahdavi-Amiri , Daniel Cohen-Or

We present a framework for generating music-synchronized, choreography aware animal dance videos. Our framework introduces choreography patterns -- structured sequences of motion beats that define the long-range structure of a dance -- as a…

计算机视觉与模式识别 · 计算机科学 2025-11-26 Xiaojuan Wang , Aleksander Holynski , Brian Curless , Ira Kemelmacher , Steve Seitz

Surveillance cameras are widely applied for indoor occupancy measurement and human movement perception, which benefit for building energy management and social security. To address the challenges of limited view angle of single camera as…

计算机视觉与模式识别 · 计算机科学 2021-04-13 Ping Zhang , Zhenxiang Tao , Wenjie Yang , Minze Chen , Shan Ding , Xiaodong Liu , Rui Yang , Hui Zhang