中文
相关论文

相关论文: Closing the Gap in Human Behavior Analysis: A Pipe…

200 篇论文

This paper proposes a method to track human figures in physical spaces and then utilizes this data to generate several data points such as footfall distribution, demographic analysis,heat maps as well as gender distribution. The proposed…

计算机视觉与模式识别 · 计算机科学 2017-11-23 Namit Juneja , Rajesh Kumar Muthu

Compared to RGB images, raw sensor data provides a richer representation of information, which is crucial for accurate recognition, particularly under challenging conditions such as low-light environments. The traditional Image Signal…

计算机视觉与模式识别 · 计算机科学 2026-05-11 Hanxi Li , Yao Cheng , Bo Zhang , Li Zeng

This study uses domain randomization to generate a synthetic RGB-D dataset for training multimodal instance segmentation models, aiming to achieve colour-agnostic hand localization in cluttered industrial environments. Domain randomization…

人机交互 · 计算机科学 2026-02-23 Stefan Grushko , Aleš Vysocký , Jakub Chlebek , Petr Prokop

Synthesizing multi-character interactions is a challenging task due to the complex and varied interactions between the characters. In particular, precise spatiotemporal alignment between characters is required in generating close…

图形学 · 计算机科学 2022-08-05 Aman Goel , Qianhui Men , Edmond S. L. Ho

The real-world is inherently multi-modal at its core. Our tools observe and take snapshots of it, in digital form, such as videos or sounds, however much of it is lost. Similarly for actions and information passing between humans, languages…

计算机视觉与模式识别 · 计算机科学 2026-05-26 Mihai-Cristian Pîrvu , Marius Leordeanu

A crucial component of an autonomous vehicle (AV) is the artificial intelligence (AI) is able to drive towards a desired destination. Today, there are different paradigms addressing the development of AI drivers. On the one hand, we find…

计算机视觉与模式识别 · 计算机科学 2020-10-27 Yi Xiao , Felipe Codevilla , Akhil Gurram , Onay Urfalioglu , Antonio M. López

We address the problem of text-guided video temporal grounding, which aims to identify the time interval of a certain event based on a natural language description. Different from most existing methods that only consider RGB images as…

计算机视觉与模式识别 · 计算机科学 2021-11-01 Yi-Wen Chen , Yi-Hsuan Tsai , Ming-Hsuan Yang

Monitoring the movement and actions of humans in video in real-time is an important task. We present a deep learning based algorithm for human action recognition for both RGB and thermal cameras. It is able to detect and track humans and…

计算机视觉与模式识别 · 计算机科学 2023-04-05 Hannes Fassold , Karlheinz Gutjahr , Anna Weber , Roland Perko

The focal point of egocentric video understanding is modelling hand-object interactions. Standard models -- CNNs, Vision Transformers, etc. -- which receive RGB frames as input perform well, however, their performance improves further by…

计算机视觉与模式识别 · 计算机科学 2022-10-11 Gorjan Radevski , Dusan Grujicic , Matthew Blaschko , Marie-Francine Moens , Tinne Tuytelaars

This paper demonstrates human synthesis based on the Radio Frequency (RF) signals, which leverages the fact that RF signals can record human movements with the signal reflections off the human body. Different from existing RF sensing works…

多媒体 · 计算机科学 2021-12-08 Cong Yu , Zhi Wu , Dongheng Zhang , Zhi Lu , Yang Hu , Yan Chen

Skeleton-based human action recognition is a powerful approach for understanding human behaviour from pose data, but collecting large-scale, diverse, and well-annotated 3D skeleton datasets is both expensive and labor-intensive. To address…

计算机视觉与模式识别 · 计算机科学 2026-04-17 Xu Dong , Wanqing Li , Anthony Adeyemi-Ejeye , Andrew Gilbert

RGB and thermal source data suffer from both shared and specific challenges, and how to explore and exploit them plays a critical role to represent the target appearance in RGBT tracking. In this paper, we propose a novel challenge-aware…

计算机视觉与模式识别 · 计算机科学 2020-07-28 Chenglong Li , Lei Liu , Andong Lu , Qing Ji , Jin Tang

Methods for generating synthetic data have become of increasing importance to build large datasets required for Convolution Neural Networks (CNN) based deep learning techniques for a wide range of computer vision applications. In this work,…

计算机视觉与模式识别 · 计算机科学 2020-09-02 Muhammad Ali Farooq , Peter Corcoran

We propose a method for human activity recognition from RGB data that does not rely on any pose information during test time and does not explicitly calculate pose information internally. Instead, a visual attention module learns to predict…

计算机视觉与模式识别 · 计算机科学 2018-08-22 Fabien Baradel , Christian Wolf , Julien Mille , Graham W. Taylor

Multi-modality of color and depth, i.e., RGB-D, is of great importance in recent research of indoor scene recognition. In this kind of data representation, depth map is able to describe the 3D structure of scenes and geometric relations…

计算机视觉与模式识别 · 计算机科学 2026-04-02 Qiong Liu , Ruofei Xiong , Xingzhen Chen , Muyao Peng , You Yang

Computational models have emerged as powerful tools for multi-scale energy modeling research at the building and urban scale, supporting data-driven analysis across building and urban energy systems. However, these models require large…

人工智能 · 计算机科学 2026-04-09 Jackson Eshbaugh , Chetan Tiwari , Jorge Silveyra

Human emotion recognition is an important direction in the field of biometric and information forensics. However, most existing human emotion research are based on the single RGB view. In this paper, we introduce a RGBD video-emotion…

计算机视觉与模式识别 · 计算机科学 2018-11-09 Shenglan Liu , Shuai Guo , Hong Qiao , Yang Wang , Bin Wang , Wenbo Luo , Mingming Zhang , Keye Zhang , Bixuan Du

This paper strives for action recognition and detection in video modalities like RGB, depth maps or 3D-skeleton sequences when only limited modality-specific labeled examples are available. For the RGB, and derived optical-flow, modality…

计算机视觉与模式识别 · 计算机科学 2021-08-10 Fida Mohammad Thoker , Cees G. M. Snoek

Paired RGB-thermal data has shown significant utility across a range of applications, including image fusion, object tracking, and anomaly detection; however, its broader adoption is constrained by the limited availability of aligned…

计算机视觉与模式识别 · 计算机科学 2026-05-19 Tseten Sherpa , Sikandar Ali , Shubham Parab , Haoyun Feng , Matthew Dennis , Keenan Gibbons , Verrah Otiende , Geoffrey H. Siwo

Image and video synthesis has become a blooming topic in computer vision and machine learning communities along with the developments of deep generative models, due to its great academic and application value. Many researchers have been…

计算机视觉与模式识别 · 计算机科学 2024-05-27 Zhen Jia , Zhang Zhang , Liang Wang , Tieniu Tan