中文
相关论文

相关论文: PKU-DyMVHumans: A Multi-View Video Benchmark for H…

200 篇论文

Meeting online is becoming the new normal. Creating an immersive experience for online meetings is a necessity towards more diverse and seamless environments. Efficient photorealistic rendering of human 3D dynamics is the core of immersive…

计算机视觉与模式识别 · 计算机科学 2023-06-30 Chuanyue Shen , Letian Zhang , Zhangsihao Yang , Masood Mortazavi , Xiyun Song , Liang Peng , Heather Yu

Tracking a crowd in 3D using multiple RGB cameras is a challenging task. Most previous multi-camera tracking algorithms are designed for offline setting and have high computational complexity. Robust real-time multi-camera 3D tracking is…

计算机视觉与模式识别 · 计算机科学 2020-03-27 Quanzeng You , Hao Jiang

Neural Radiance Fields (NeRF) has achieved impressive results in single object scene reconstruction and novel view synthesis, which have been demonstrated on many single modality and single object focused indoor scene datasets like DTU,…

计算机视觉与模式识别 · 计算机科学 2023-01-18 Chongshan Lu , Fukun Yin , Xin Chen , Tao Chen , Gang YU , Jiayuan Fan

Synthesizing high-fidelity videos from real-world multi-view input is challenging because of the complexities of real-world environments and highly dynamic motions. Previous works based on neural radiance fields have demonstrated…

计算机视觉与模式识别 · 计算机科学 2023-10-10 Feng Wang , Sinan Tan , Xinghang Li , Zeyue Tian , Yafei Song , Huaping Liu

Novel view synthesis (NVS) of multi-human scenes imposes challenges due to the complex inter-human occlusions. Layered representations handle the complexities by dividing the scene into multi-layered radiance fields, however, they are…

计算机视觉与模式识别 · 计算机科学 2023-09-22 Youssef Abdelkareem , Shady Shehata , Fakhri Karray

The analysis of the ubiquitous human-human interactions is pivotal for understanding humans as social beings. Existing human-human interaction datasets typically suffer from inaccurate body motions, lack of hand gestures and fine-grained…

计算机视觉与模式识别 · 计算机科学 2023-12-27 Liang Xu , Xintao Lv , Yichao Yan , Xin Jin , Shuwen Wu , Congsheng Xu , Yifan Liu , Yizhou Zhou , Fengyun Rao , Xingdong Sheng , Yunhui Liu , Wenjun Zeng , Xiaokang Yang

This study seeks to automate camera movement control for filming existing subjects into attractive videos, contrasting with the creation of non-existent content by directly generating the pixels. We select drone videos as our test case due…

计算机视觉与模式识别 · 计算机科学 2024-12-13 Yunzhong Hou , Liang Zheng , Philip Torr

Current benchmarks for facial expression recognition (FER) mainly focus on static images, while there are limited datasets for FER in videos. It is still ambiguous to evaluate whether performances of existing methods remain satisfactory in…

计算机视觉与模式识别 · 计算机科学 2022-03-22 Yan Wang , Yixuan Sun , Yiwen Huang , Zhongying Liu , Shuyong Gao , Wei Zhang , Weifeng Ge , Wenqiang Zhang

Existing 4D human datasets fall short for fashion-specific research, lacking either realistic garment dynamics or task-specific annotations. Synthetic datasets suffer from a realism gap, whereas real-world captures lack the detailed…

计算机视觉与模式识别 · 计算机科学 2026-03-10 Hunor Laczkó , Libang Jia , Loc-Phat Truong , Diego Hernández , Sergio Escalera , Jordi Gonzalez , Meysam Madadi

We introduce AG-VPReID, a new large-scale dataset for aerial-ground video-based person re-identification (ReID) that comprises 6,632 subjects, 32,321 tracklets and over 9.6 million frames captured by drones (altitudes ranging from 15-120m),…

计算机视觉与模式识别 · 计算机科学 2025-03-18 Huy Nguyen , Kien Nguyen , Akila Pemasiri , Feng Liu , Sridha Sridharan , Clinton Fookes

Person search has recently been a challenging task in the computer vision domain, which aims to search specific pedestrians from real cameras.Nevertheless, most surveillance videos comprise only a handful of images of each pedestrian, which…

计算机视觉与模式识别 · 计算机科学 2023-08-09 Huibing Wang , Tianxiang Cui , Mingze Yao , Huijuan Pang , Yushan Du

The prevalence of violence in daily life poses significant threats to individuals' physical and mental well-being. Using surveillance cameras in public spaces has proven effective in proactively deterring and preventing such incidents.…

计算机视觉与模式识别 · 计算机科学 2023-10-24 Yiting Dong , Yang Li , Dongcheng Zhao , Guobin Shen , Yi Zeng

Image recapture seriously breaks the fairness of artificial intelligent (AI) systems, which deceives the system by recapturing others' images. Most of the existing recapture models can only address a single pattern of recapture (e.g.,…

计算机视觉与模式识别 · 计算机科学 2022-06-14 Shuyu Miao , Lin Zheng , Hong Jin

We have built a custom mobile multi-camera large-space dense light field capture system, which provides a series of high-quality and sufficiently dense light field images for various scenarios. Our aim is to contribute to the development of…

计算机视觉与模式识别 · 计算机科学 2024-03-18 Xiaohang Yu , Zhengxian Yang , Shi Pan , Yuqi Han , Haoxiang Wang , Jun Zhang , Shi Yan , Borong Lin , Lei Yang , Tao Yu , Lu Fang

In this paper, we introduce a new benchmark dataset named IPN Hand with sufficient size, variety, and real-world elements able to train and evaluate deep neural networks. This dataset contains more than 4,000 gesture samples and 800,000 RGB…

计算机视觉与模式识别 · 计算机科学 2020-10-21 Gibran Benitez-Garcia , Jesus Olivares-Mercado , Gabriel Sanchez-Perez , Keiji Yanai

High-quality 3D human body reconstruction requires high-fidelity and large-scale training data and appropriate network design that effectively exploits the high-resolution input images. To tackle these problems, we propose a simple yet…

计算机视觉与模式识别 · 计算机科学 2023-03-28 Sang-Hun Han , Min-Gyu Park , Ju Hong Yoon , Ju-Mi Kang , Young-Jae Park , Hae-Gon Jeon

Multimodal human action recognition (HAR) leverages complementary sensors for activity classification. Beyond recognition, recent advances in large language models (LLMs) enable detailed descriptions and causal reasoning, motivating new…

计算机视觉与模式识别 · 计算机科学 2025-12-09 Siyang Jiang , Mu Yuan , Xiang Ji , Bufang Yang , Zeyu Liu , Lilin Xu , Yang Li , Yuting He , Liran Dong , Wenrui Lu , Zhenyu Yan , Xiaofan Jiang , Wei Gao , Hongkai Chen , Guoliang Xing

The volumetric representation of human interactions is one of the fundamental domains in the development of immersive media productions and telecommunication applications. Particularly in the context of the rapid advancement of Extended…

计算机视觉与模式识别 · 计算机科学 2024-02-15 Fatemeh Ghorbani Lohesara , Davi Rabbouni Freitas , Christine Guillemot , Karen Eguiazarian , Sebastian Knorr

Neuromorphic vision sensors, such as the dynamic vision sensor (DVS) and spike camera, have gained increasing attention in recent years. The spike camera can detect fine textures by mimicking the fovea in the human visual system, and output…

计算机视觉与模式识别 · 计算机科学 2024-12-17 Wei Zhang , Weiquan Yan , Yun Zhao , Wenxiang Cheng , Gang Chen , Huihui Zhou , Yonghong Tian

Current vision-language multimodal models are well-adapted for general visual understanding tasks. However, they perform inadequately when handling complex visual tasks related to human poses and actions due to the lack of specialized…

计算机视觉与模式识别 · 计算机科学 2025-06-03 Dewen Zhang , Wangpeng An , Hayaru Shouno