中文
相关论文

相关论文: RecruitView: A Multimodal Dataset for Predicting P…

200 篇论文

Gait recognition has emerged as a compelling biometric modality for surveillance and security applications, offering inherent advantages such as non-intrusiveness, resistance to disguise, and long-range identification capability. However,…

计算机视觉与模式识别 · 计算机科学 2026-05-01 Yabo Luo , Xiaoyun Wang , Cunrong Li

The rapid progress of Multimodal Large Language Models(MLLMs) has transformed the AI landscape. These models combine pre-trained LLMs with various modality encoders. This integration requires a systematic understanding of how different…

计算与语言 · 计算机科学 2025-06-06 Jisu An , Junseok Lee , Jeoungeun Lee , Yongseok Son

Automatic facial expression classification (FER) from videos is a critical problem for the development of intelligent human-computer interaction systems. Still, it is a challenging problem that involves capturing high-dimensional…

计算机视觉与模式识别 · 计算机科学 2016-07-22 Arnaud Dapogny , Kévin Bailly , Séverine Dubuisson

Synthesizing high-fidelity head avatars is a central problem for computer vision and graphics. While head avatar synthesis algorithms have advanced rapidly, the best ones still face great obstacles in real-world scenarios. One of the vital…

计算机视觉与模式识别 · 计算机科学 2023-05-24 Dongwei Pan , Long Zhuo , Jingtan Piao , Huiwen Luo , Wei Cheng , Yuxin Wang , Siming Fan , Shengqi Liu , Lei Yang , Bo Dai , Ziwei Liu , Chen Change Loy , Chen Qian , Wayne Wu , Dahua Lin , Kwan-Yee Lin

Inter-modal interaction plays an indispensable role in multimodal sentiment analysis. Due to different modalities sequences are usually non-alignment, how to integrate relevant information of each modality to learn fusion representations…

计算与语言 · 计算机科学 2022-12-23 Kaicheng Yang , Ruxuan Zhang , Hua Xu , Kai Gao

3D face reconstruction (3DFR) algorithms are based on specific assumptions tailored to the limits and characteristics of the different application scenarios. In this study, we investigate how multiple state-of-the-art 3DFR algorithms can be…

计算机视觉与模式识别 · 计算机科学 2025-04-29 Simone Maurizio La Cava , Roberto Casula , Sara Concas , Giulia Orrù , Ruben Tolosana , Martin Drahansky , Julian Fierrez , Gian Luca Marcialis

Point cloud registration has seen significant advancements with the application of deep learning techniques. However, existing approaches often overlook the potential of integrating radiometric information from RGB images. This limitation…

计算机视觉与模式识别 · 计算机科学 2025-06-09 Zhaoyi Wang , Shengyu Huang , Jemil Avers Butt , Yuanzhou Cai , Matej Varga , Andreas Wieser

Information ecosystems increasingly shape how people internalize exposure to adverse digital experiences, raising concerns about the long-term consequences for information health. In modern search and recommendation systems, ranking and…

计算机与社会 · 计算机科学 2026-02-18 Victor De Lima , Jiqun Liu , Grace Hui Yang

Audio-visual embodied navigation aims to enable an agent to autonomously localize and reach a sound source in unseen 3D environments by leveraging auditory cues. The key challenge of this task lies in effectively modeling the interaction…

计算机视觉与模式识别 · 计算机科学 2026-01-15 Yi Wang , Yinfeng Yu , Bin Ren

Recent progress in face detection (including keypoint detection), and recognition is mainly being driven by (i) deeper convolutional neural network architectures, and (ii) larger datasets. However, most of the large datasets are maintained…

计算机视觉与模式识别 · 计算机科学 2017-05-23 Ankan Bansal , Anirudh Nanduri , Carlos Castillo , Rajeev Ranjan , Rama Chellappa

Graph based molecular representation learning is essential for accurately predicting molecular properties in drug discovery and materials science; however, it faces significant challenges due to the intricate relationships among molecules…

计算工程、金融与科学 · 计算机科学 2025-05-28 Zhengyang Zhou , Yunrui Li , Pengyu Hong , Hao Xu

AI-enhanced personality assessments are increasingly shaping hiring decisions, using affective computing to predict traits from the Big Five (OCEAN) model. However, integrating AI into these assessments raises ethical concerns, especially…

人机交互 · 计算机科学 2025-11-24 Dena F. Mujtaba , Nihar R. Mahapatra

Effectively describing features for cross-modal remote sensing image matching remains a challenging task due to the significant geometric and radiometric differences between multimodal images. Existing methods primarily extract features at…

计算机视觉与模式识别 · 计算机科学 2025-10-03 Abu Sadat Mohammad Salehin Amit , Xiaoli Zhang , Md Masum Billa Shagar , Zhaojun Liu , Xiongfei Li , Fanlong Meng

The increasing global prevalence of mental disorders, such as depression and PTSD, requires objective and scalable diagnostic tools. Traditional clinical assessments often face limitations in accessibility, objectivity, and consistency.…

音频与语音处理 · 电气工程与系统科学 2025-04-03 Abdelrahaman A. Hassan , Abdelrahman A. Ali , Aya E. Fouda , Radwa J. Hanafy , Mohammed E. Fouda

Robust face clustering is a vital step in enabling computational understanding of visual character portrayal in media. Face clustering for long-form content is challenging because of variations in appearance and lack of supporting…

计算机视觉与模式识别 · 计算机科学 2022-03-01 Krishna Somandepalli , Rajat Hebbar , Shrikanth Narayanan

Due to its ability to accurately predict emotional state using multimodal features, audiovisual emotion recognition has recently gained more interest from researchers. This paper proposes two methods to predict emotional attributes from…

音频与语音处理 · 电气工程与系统科学 2022-07-22 Bagus Tris Atmaja , Masato Akagi

In human-centered environments such as restaurants, homes, and warehouses, robots often face challenges in accurately recognizing 3D objects. These challenges stem from the complexity and variability of these environments, including diverse…

计算机视觉与模式识别 · 计算机科学 2025-08-14 Songsong Xiong , Hamidreza Kasaei

We present KeyMorph, a deep learning-based image registration framework that relies on automatically detecting corresponding keypoints. State-of-the-art deep learning methods for registration often are not robust to large misalignments, are…

计算机视觉与模式识别 · 计算机科学 2023-09-04 Alan Q. Wang , Evan M. Yu , Adrian V. Dalca , Mert R. Sabuncu

In recent years, Deep Learning has been successfully applied to multimodal learning problems, with the aim of learning useful joint representations in data fusion applications. When the available modalities consist of time series data such…

计算机视觉与模式识别 · 计算机科学 2017-04-12 Xitong Yang , Palghat Ramesh , Radha Chitta , Sriganesh Madhvanath , Edgar A. Bernal , Jiebo Luo

The video-based person re-identification is to recognize a person under different cameras, which is a crucial task applied in visual surveillance system. Most previous methods mainly focused on the feature of full body in the frame. In this…

计算机视觉与模式识别 · 计算机科学 2018-07-04 Jie Liu , Cheng Sun , Xiang Xu , Baomin Xu , Shuangyuan Yu
‹ 上一页 1 8 9 10 下一页 ›