中文
相关论文

相关论文: DeSPITE: Exploring Contrastive Deep Skeleton-Point…

200 篇论文

Combining different sensing modalities with multiple positions helps form a unified perception and understanding of complex situations such as human behavior. Hence, human activity recognition (HAR) benefits from combining redundant and…

机器学习 · 计算机科学 2024-04-26 Hymalai Bello

Human Activity Recognition (HAR) systems aim to understand human behaviour and assign a label to each action, attracting significant attention in computer vision due to their wide range of applications. HAR can leverage various data…

计算机视觉与模式识别 · 计算机科学 2024-09-17 Jungpil Shin , Najmul Hassan , Abu Saleh Musa Miah1 , Satoshi Nishimura

Human Action Recognition (HAR) aims to understand human behavior and assign a label to each action. It has a wide range of applications, and therefore has been attracting increasing attention in the field of computer vision. Human actions…

计算机视觉与模式识别 · 计算机科学 2022-06-22 Zehua Sun , Qiuhong Ke , Hossein Rahmani , Mohammed Bennamoun , Gang Wang , Jun Liu

One of the primary challenges in the field of human activity recognition (HAR) is the lack of large labeled datasets. This hinders the development of robust and generalizable models. Recently, cross modality transfer approaches have been…

计算机视觉与模式识别 · 计算机科学 2024-02-05 Zikang Leng , Amitrajit Bhattacharjee , Hrudhai Rajasekhar , Lizhe Zhang , Elizabeth Bruda , Hyeokhyen Kwon , Thomas Plötz

We present Lepard, a Learning based approach for partial point cloud matching in rigid and deformable scenes. The key characteristics are the following techniques that exploit 3D positional knowledge for point cloud matching: 1) An…

计算机视觉与模式识别 · 计算机科学 2022-03-08 Yang Li , Tatsuya Harada

Visible-Infrared Person Re-Identification (VI-ReID) is a challenging retrieval task due to the substantial modality gap between visible and infrared images. While existing methods attempt to bridge this gap by learning modality-invariant…

计算机视觉与模式识别 · 计算机科学 2026-03-17 Haoxuan Xu , Guanglin Niu

3D object detection from LiDAR point cloud is of critical importance for autonomous driving and robotics. While sequential point cloud has the potential to enhance 3D perception through temporal information, utilizing these temporal…

计算机视觉与模式识别 · 计算机科学 2023-07-06 Zheyuan Zhou , Jiachen Lu , Yihan Zeng , Hang Xu , Li Zhang

Generating realistic intermediate shapes between non-rigidly deformed shapes is a challenging task in computer vision, especially with unstructured data (e.g., point clouds) where temporal consistency across frames is lacking, and…

计算机视觉与模式识别 · 计算机科学 2025-02-28 Lu Sang , Zehranaz Canfes , Dongliang Cao , Riccardo Marin , Florian Bernard , Daniel Cremers

Human Activity Recognition (HAR) with wearable sensors is essential for applications in healthcare, fitness, and human-computer interaction. Bio-impedance sensing offers unique advantages for fine-grained motion capture but remains…

计算机视觉与模式识别 · 计算机科学 2025-07-10 Lala Shakti Swarup Ray , Mengxi Liu , Deepika Gurung , Bo Zhou , Sungho Suh , Paul Lukowicz

The paper introduces the Decouple Re-identificatiOn and human Parsing (DROP) method for occluded person re-identification (ReID). Unlike mainstream approaches using global features for simultaneous multi-task learning of ReID and human…

计算机视觉与模式识别 · 计算机科学 2024-02-01 Shuguang Dou , Xiangyang Jiang , Yuanpeng Tu , Junyao Gao , Zefan Qu , Qingsong Zhao , Cairong Zhao

Intrinsic image decomposition (IID) is the task that decomposes a natural image into albedo and shade. While IID is typically solved through supervised learning methods, it is not ideal due to the difficulty in observing ground truth albedo…

计算机视觉与模式识别 · 计算机科学 2023-03-29 Shogo Sato , Yasuhiro Yao , Taiga Yoshida , Takuhiro Kaneko , Shingo Ando , Jun Shimamura

We present IMU2CLIP, a novel pre-training approach to align Inertial Measurement Unit (IMU) motion sensor recordings with video and text, by projecting them into the joint representation space of Contrastive Language-Image Pre-training…

计算机视觉与模式识别 · 计算机科学 2022-10-27 Seungwhan Moon , Andrea Madotto , Zhaojiang Lin , Alireza Dirafzoon , Aparajita Saraf , Amy Bearman , Babak Damavandi

Inertial Measurement Unit (IMU) sensors are present in everyday devices such as smartphones and fitness watches. As a result, the array of health-related research and applications that tap onto this data has been growing, but little…

机器学习 · 计算机科学 2021-08-23 Davi Pedrosa de Aguiar , Fabricio Murai

Video-based gait recognition has achieved impressive results in constrained scenarios. However, visual cameras neglect human 3D structure information, which limits the feasibility of gait recognition in the 3D wild world. Instead of…

计算机视觉与模式识别 · 计算机科学 2023-03-31 Chuanfu Shen , Chao Fan , Wei Wu , Rui Wang , George Q. Huang , Shiqi Yu

Recent advancements in vision-language pre-training (e.g. CLIP) have shown that vision models can benefit from language supervision. While many models using language modality have achieved great success on 2D vision tasks, the joint…

计算机视觉与模式识别 · 计算机科学 2023-01-19 Rui Huang , Xuran Pan , Henry Zheng , Haojun Jiang , Zhifeng Xie , Shiji Song , Gao Huang

We introduce HiSC4D, a novel Human-centered interaction and 4D Scene Capture method, aimed at accurately and efficiently creating a dynamic digital world, containing large-scale indoor-outdoor scenes, diverse human motions, rich human-human…

计算机视觉与模式识别 · 计算机科学 2024-09-17 Yudi Dai , Zhiyong Wang , Xiping Lin , Chenglu Wen , Lan Xu , Siqi Shen , Yuexin Ma , Cheng Wang

The Visible-Infrared Person Re-identification (VI ReID) aims to match visible and infrared images of the same pedestrians across non-overlapped camera views. These two input modalities contain both invariant information, such as shape, and…

计算机视觉与模式识别 · 计算机科学 2024-06-19 Ruiqi Wu , Bingliang Jiao , Wenxuan Wang , Meng Liu , Peng Wang

Human activity recognition (HAR) using inertial measurement units (IMUs) increasingly leverages large language models (LLMs), yet existing approaches focus on coarse activities like walking or running. Our preliminary study indicates that…

计算机视觉与模式识别 · 计算机科学 2025-04-07 Lilin Xu , Kaiyuan Hou , Xiaofan Jiang

Existing contrastive language-image pre-training aims to learn a joint representation by matching abundant image-text pairs. However, the number of image-text pairs in medical datasets is usually orders of magnitude smaller than that in…

计算机视觉与模式识别 · 计算机科学 2024-01-04 Jiarun Liu , Hong-Yu Zhou , Cheng Li , Weijian Huang , Hao Yang , Yong Liang , Shanshan Wang

Millimeter-wave (mmWave) radar offers robust sensing capabilities in diverse environments, making it a highly promising solution for human body reconstruction due to its privacy-friendly and non-intrusive nature. However, the significant…

计算机视觉与模式识别 · 计算机科学 2025-03-05 Jiarui Yang , Songpengcheng Xia , Zengyuan Lai , Lan Sun , Qi Wu , Wenxian Yu , Ling Pei