中文
相关论文

相关论文: Modeling Cross-view Interaction Consistency for Pa…

200 篇论文

Capturing complex temporal relationships between video and audio modalities is vital for Audio-Visual Emotion Recognition (AVER). However, existing methods lack attention to local details, such as facial state changes between video frames,…

计算机视觉与模式识别 · 计算机科学 2024-05-28 Tong Shi , Xuri Ge , Joemon M. Jose , Nicolas Pugeault , Paul Henderson

As artificial intelligence (AI) systems become increasingly embedded in our daily life, the ability to recognize and adapt to human emotions is essential for effective human-computer interaction. Facial expression recognition (FER) provides…

计算机视觉与模式识别 · 计算机科学 2025-12-16 Thibault Geoffroy , Myriam Maumy , Lionel Prevost

Mirror neurons have been observed in the primary motor cortex of primate species, in particular in humans and monkeys. A mirror neuron fires when a person performs a certain action, and also when he observes the same action being performed…

计算机视觉与模式识别 · 计算机科学 2016-12-20 Shervin Ardeshir , Krishna Regmi , Ali Borji

Human Object Interaction (HOI) detection is a challenging task that requires to distinguish the interaction between a human-object pair. Attention based relation parsing is a popular and effective strategy utilized in HOI. However, current…

计算机视觉与模式识别 · 计算机科学 2022-07-19 Jingjia Huang , Baixiang Yang

As Augmented Reality (AR) devices become more prevalent and commercially viable, the need for quick, efficient, and secure schemes for pairing these devices has become more pressing. Current methods to securely exchange holograms require…

密码学与安全 · 计算机科学 2023-03-15 Matthew Corbett , Jiacheng Shang , Bo Ji

Thanks to the availability and increasing popularity of Egocentric cameras such as GoPro cameras, glasses, and etc. we have been provided with a plethora of videos captured from the first person perspective. Surveillance cameras and…

计算机视觉与模式识别 · 计算机科学 2016-09-15 Shervin Ardeshir , Ali Borji

Entity Resolution (ER) is a constitutional part for integrating different knowledge graphs in order to identify entities referring to the same real-world object. A promising approach is the use of graph embeddings for ER in order to…

机器学习 · 计算机科学 2021-01-18 Daniel Obraczka , Jonathan Schuchart , Erhard Rahm

Autonomous driving requires efficient reasoning about the location and appearance of the different agents in the scene, which aids in downstream tasks such as object detection, object tracking, and path planning. The past few years have…

计算机视觉与模式识别 · 计算机科学 2022-11-10 Sarthak Sharma , Unnikrishnan R. Nair , Udit Singh Parihar , Midhun Menon S , Srikanth Vidapanakal

Given a dataset of individuals each described by a covariate vector, a treatment, and an observed outcome on the treatment, the goal of the individual treatment effect (ITE) estimation task is to predict outcome changes resulting from a…

机器学习 · 计算机科学 2024-06-07 Lokesh Nagalapatti , Pranava Singhal , Avishek Ghosh , Sunita Sarawagi

In this paper, we explore the dynamic grasping of moving objects through active pose tracking and reinforcement learning for hand-eye coordination systems. Most existing vision-based robotic grasping methods implicitly assume target objects…

机器人学 · 计算机科学 2023-10-11 Baichuan Huang , Jingjin Yu , Siddarth Jain

When humans work together to complete a joint task, each person builds an internal model of the situation and how it will evolve. Efficient collaboration is dependent on how these individual models overlap to form a shared mental model…

机器人学 · 计算机科学 2022-08-26 Wesley P. Chan , Morgan Crouch , Khoa Hoang , Charlie Chen , Nicole Robinson , Elizabeth Croft

Unsupervised graph representation learning aims to distill various graph information into a downstream task-agnostic dense vector embedding. However, existing graph representation learning approaches are designed mainly under the node…

机器学习 · 计算机科学 2022-03-04 You Li , Bei Lin , Binli Luo , Ning Gui

The body pose of a person wearing a camera is of great interest for applications in augmented reality, healthcare, and robotics, yet much of the person's body is out of view for a typical wearable camera. We propose a learning-based…

计算机视觉与模式识别 · 计算机科学 2020-03-31 Evonne Ng , Donglai Xiang , Hanbyul Joo , Kristen Grauman

Image retrieval plays a pivotal role in applications from wildlife conservation to healthcare, for finding individual animals or relevant images to aid diagnosis. Although deep learning techniques for image retrieval have advanced…

计算机视觉与模式识别 · 计算机科学 2024-07-15 Vaibhav Balloli , Sara Beery , Elizabeth Bondi-Kelly

Computer vision has achieved great success in interpreting semantic meanings from images, yet estimating underlying (non-visual) physical properties of an object is often limited to their bulk values rather than reconstructing a dense map.…

计算机视觉与模式识别 · 计算机科学 2022-01-31 Shuangjun Liu , Sarah Ostadabbas

Unified video and action prediction models hold great potential for robotic manipulation, as future observations offer contextual cues for planning, while actions reveal how interactions shape the environment. However, most existing…

机器人学 · 计算机科学 2025-12-08 Yijie Zhu , Rui Shao , Ziyang Liu , Jie He , Jizhihui Liu , Jiuru Wang , Zitong Yu

The massive availability of cameras results in a wide variability of imaging conditions, producing large intra-class variations and a significant performance drop if heterogeneous images are compared for person recognition. However, as…

计算机视觉与模式识别 · 计算机科学 2022-03-31 Fernando Alonso-Fernandez , Kiran B. Raja , R. Raghavendra , Cristoph Busch , Josef Bigun , Ruben Vera-Rodriguez , Julian Fierrez

We propose PARSE, a novel semi-supervised architecture for learning strong EEG representations for emotion recognition. To reduce the potential distribution mismatch between the large amounts of unlabeled data and the limited amount of…

机器学习 · 计算机科学 2022-09-28 Guangyi Zhang , Vandad Davoodnia , Ali Etemad

NIR-to-VIS face recognition is identifying faces of two different domains by extracting domain-invariant features. However, this is a challenging problem due to the two different domain characteristics, and the lack of NIR face dataset. In…

计算机视觉与模式识别 · 计算机科学 2022-08-05 MyeongAh Cho , Tae-young Chun , g Taeoh Kim , Sangyoun Lee

Event cameras action recognition (EAR) offers compelling privacy-protecting and efficiency advantages, where temporal motion dynamics is of great importance. Existing spatiotemporal multi-view representation learning (SMVRL) methods for…

计算机视觉与模式识别 · 计算机科学 2026-01-27 Rui Fan , Weidong Hao