中文
相关论文

相关论文: AuthFormer: Adaptive Multimodal biometric authenti…

200 篇论文

Driver action recognition, aiming to accurately identify drivers' behaviours, is crucial for enhancing driver-vehicle interactions and ensuring driving safety. Unlike general action recognition, drivers' environments are often challenging,…

计算机视觉与模式识别 · 计算机科学 2024-08-20 Ruoyu Wang , Wenqian Wang , Jianjun Gao , Dan Lin , Kim-Hui Yap , Bingbing Li

Healthcare data now span EHRs, medical imaging, genomics, and wearable sensors, but most diagnostic models still process these modalities in isolation. This limits their ability to capture early, cross-modal disease signatures. This paper…

机器学习 · 计算机科学 2025-12-18 Md Talha Mohsin , Ismail Abdulrashid

We introduce SurgFormer, a multiresolution gated transformer for data driven soft tissue simulation on volumetric meshes. High fidelity biomechanical solvers are often too costly for interactive use, so we train SurgFormer on solver…

Face recognition is a crucial task in various multimedia applications such as security check, credential access and motion sensing games. However, the task is challenging when an input face is noisy (e.g. poor-condition RGB image) or lacks…

计算机视觉与模式识别 · 计算机科学 2021-12-09 Wenbin Teng , Chongyang Bai

The prediction of mild cognitive impairment (MCI) conversion to Alzheimer's disease (AD) is important for early treatment to prevent or slow the progression of AD. To accurately predict the MCI conversion to stable MCI or progressive MCI,…

计算机视觉与模式识别 · 计算机科学 2023-07-17 Linfeng Liu , Junyan Lyu , Siyu Liu , Xiaoying Tang , Shekhar S. Chandra , Fatima A. Nasrallah

With the development of technology, the usage areas and importance of biometric systems have started to increase. Since the characteristics of each person are different from each other, a single model biometric system can yield successful…

机器学习 · 计算机科学 2019-03-20 Cihan Akın , Umit Kacar , Murvet Kirci

Image editing techniques have rapidly advanced, facilitating both innovative use cases and malicious manipulation of digital images. Deep learning-based methods have recently achieved high accuracy in pixel-level forgery localization, yet…

计算机视觉与模式识别 · 计算机科学 2025-06-27 Ju-Hyeon Nam , Dong-Hyun Moon , Sang-Chul Lee

Authentication plays a significant part in dealing with security in public and private sectors such as healthcare systems, banking system, transportation system and law and security. Biometric technology has grown quickly recently,…

密码学与安全 · 计算机科学 2023-01-03 Konark Modi , Lakshmipathi Devaraj

Biometric authentication systems are crucial for security, but developing them involves various complexities, including privacy, security, and achieving high accuracy without directly storing pure biometric data in storage. We introduce an…

计算机视觉与模式识别 · 计算机科学 2026-04-13 Dmytro Zakharov , Oleksandr Kuznetsov , Emanuele Frontoni , Natalia Kryvinska

In this work, we propose a novel two-stage framework, called FaceShifter, for high fidelity and occlusion aware face swapping. Unlike many existing face swapping works that leverage only limited information from the target image when…

计算机视觉与模式识别 · 计算机科学 2020-09-16 Lingzhi Li , Jianmin Bao , Hao Yang , Dong Chen , Fang Wen

The multimodal datasets can be leveraged to pre-train large-scale vision-language models by providing cross-modal semantics. Current endeavors for determining the usage of datasets mainly focus on single-modal dataset ownership verification…

信息检索 · 计算机科学 2025-04-18 Wenyi Zhang , Ju Jia , Xiaojun Jia , Yihao Huang , Xinfeng Li , Cong Wu , Lina Wang

In the field of cross-modal retrieval, single encoder models tend to perform better than dual encoder models, but they suffer from high latency and low throughput. In this paper, we present a dual encoder model called BagFormer that…

信息检索 · 计算机科学 2023-01-02 Haowen Hou , Xiaopeng Yan , Yigeng Zhang , Fengzong Lian , Zhanhui Kang

Biometric systems suffer from some drawbacks: a biometric system can provide in general good performances except with some individuals as its performance depends highly on the quality of the capture. One solution to solve some of these…

神经与进化计算 · 计算机科学 2012-05-16 Romain Giot , Christophe Rosenberger

To ensure safe clinical integration, deep learning models must provide more than just high accuracy; they require dependable uncertainty quantification. While current Medical Vision Transformers perform well, they frequently struggle with…

图像与视频处理 · 电气工程与系统科学 2026-04-13 Mohammed Maaz Sibhai , Abedalrhman Alkhateeb , Saad B. Ahmed

Unmanned aerial vehicle (UAV) detection and aerial object recognition are critical for modern surveillance and security, prompting a need for robust systems that overcome limitations of single-modality approaches. This research addresses…

计算机视觉与模式识别 · 计算机科学 2025-11-20 Mauro Larrat , Claudomiro Sales

Multimodal recommendation systems utilize various types of information, including images and text, to enhance the effectiveness of recommendations. The key challenge is predicting user purchasing behavior from the available data. Current…

信息检索 · 计算机科学 2025-11-04 Ke Shi , Yan Zhang , Miao Zhang , Lifan Chen , Jiali Yi , Kui Xiao , Xiaoju Hou , Zhifei Li

Various types of sensors have been considered to develop human action recognition (HAR) models. Robust HAR performance can be achieved by fusing multimodal data acquired by different sensors. In this paper, we introduce a new multimodal…

计算机视觉与模式识别 · 计算机科学 2023-09-12 Kyoung Ok Yang , Junho Koh , Jun Won Choi

Recently, transformers have demonstrated great potential for modeling long-term dependencies from skeleton sequences and thereby gained ever-increasing attention in skeleton action recognition. However, the existing transformer-based…

计算机视觉与模式识别 · 计算机科学 2024-07-31 Wenhan Wu , Ce Zheng , Zihao Yang , Chen Chen , Srijan Das , Aidong Lu

Biometric authentication systems are increasingly being deployed in critical applications, but they remain susceptible to spoofing. Since most of the research efforts focus on modality-specific anti-spoofing techniques, building a unified,…

计算机视觉与模式识别 · 计算机科学 2025-06-10 Nidheesh Gorthi , Kartik Thakral , Rishabh Ranjan , Richa Singh , Mayank Vatsa

Multimodal surface material classification plays a critical role in advancing tactile perception for robotic manipulation and interaction. In this paper, we present Surformer v2, an enhanced multi-modal classification architecture designed…

机器人学 · 计算机科学 2025-09-08 Manish Kansana , Sindhuja Penchala , Shahram Rahimi , Noorbakhsh Amiri Golilarz