中文
相关论文

相关论文: MobileFaceSwap: A Lightweight Framework for Video …

200 篇论文

Video-language modeling has attracted much attention with the rapid growth of web videos. Most existing methods assume that the video frames and text description are semantically correlated, and focus on video-language modeling at video…

计算机视觉与模式识别 · 计算机科学 2022-12-06 Haoyu Lu , Mingyu Ding , Nanyi Fei , Yuqi Huo , Zhiwu Lu

The existing action recognition methods are mainly based on clip-level classifiers such as two-stream CNNs or 3D CNNs, which are trained from the randomly selected clips and applied to densely sampled clips during testing. However, this…

计算机视觉与模式识别 · 计算机科学 2020-08-26 Yin-Dong Zheng , Zhaoyang Liu , Tong Lu , Limin Wang

Face aging has become a crucial task in computer vision, with applications ranging from entertainment to healthcare. However, existing methods struggle with achieving a realistic and seamless transformation across the entire lifespan,…

计算机视觉与模式识别 · 计算机科学 2025-10-21 Tao Liu , Dafeng Zhang , Gengchen Li , Shizhuo Liu , Yongqi Song , Senmao Li , Shiqi Yang , Boqian Li , Kai Wang , Yaxing Wang

Autonomous vehicles (AVs) are more vulnerable to network attacks due to the high connectivity and diverse communication modes between vehicles and external networks. Deep learning-based Intrusion detection, an effective method for detecting…

密码学与安全 · 计算机科学 2023-09-27 Pengzhou Cheng , Lei Hua , Haobin Jiang , Gongshen Liu

Existing methods for driver facial expression recognition (DFER) are often computationally intensive, rendering them unsuitable for real-time applications. In this work, we introduce a novel transfer learning-based dual architecture, named…

计算机视觉与模式识别 · 计算机科学 2024-09-06 Ibtissam Saadi , Douglas W. Cunningham , Taleb-ahmed Abdelmalik , Abdenour Hadid , Yassin El Hillali

Temporal modeling still remains challenging for action recognition in videos. To mitigate this issue, this paper presents a new video architecture, termed as Temporal Difference Network (TDN), with a focus on capturing multi-scale temporal…

计算机视觉与模式识别 · 计算机科学 2021-04-02 Limin Wang , Zhan Tong , Bin Ji , Gangshan Wu

Active Appearance Model (AAM) is a commonly used method for facial image analysis with applications in face identification and facial expression recognition. This paper proposes a new approach based on image alignment for AAM fitting called…

计算机视觉与模式识别 · 计算机科学 2015-11-23 Ali Mollahosseini , Mohammad H. Mahoor

Conventional voice conversion modifies voice characteristics from a source speaker to a target speaker, relying on audio input from both sides. However, this process becomes infeasible when clean audio is unavailable, such as in silent…

声音 · 计算机科学 2025-08-05 Yifan Liu , Yu Fang , Zhouhan Lin

We present IMU2Face, a gesture-driven facial reenactment system. To this end, we combine recent advances in facial motion capture and inertial measurement units (IMUs) to control the facial expressions of a person in a target video based on…

计算机视觉与模式识别 · 计算机科学 2018-01-08 Justus Thies , Michael Zollhöfer , Matthias Nießner

Numerous activities in our daily life require us to verify who we are by showing our ID documents containing face images, such as passports and driver licenses, to human operators. However, this process is slow, labor intensive and…

计算机视觉与模式识别 · 计算机科学 2018-09-19 Yichun Shi , Anil K. Jain

With the advancement of large pre-trained vision-language models, effectively transferring the knowledge embedded within these foundational models to downstream tasks has become a pivotal topic, particularly in data-scarce environments.…

计算机视觉与模式识别 · 计算机科学 2024-09-12 Tianxiang Hao , Mengyao Lyu , Hui Chen , Sicheng Zhao , Xiaohan Ding , Jungong Han , Guiguang Ding

Most of the current top-down multi-person pose estimation lightweight methods are based on multi-branch parallel pure CNN network architecture, which often struggle to capture the global context required for detecting semantically complex…

计算机视觉与模式识别 · 计算机科学 2025-11-14 Biao Guo , Cong Zhou , Fangmin Guo , Xiaonan Luo , Guibo Luo , Feng Zhang

Recent advances in text-to-image generation have driven interest in generating personalized human images that depict specific identities from reference images. Although existing methods achieve high-fidelity identity preservation, they are…

计算机视觉与模式识别 · 计算机科学 2025-07-22 Xirui Hu , Jiahao Wang , Hao Chen , Weizhan Zhang , Benqi Wang , Yikun Li , Haishun Nan

Designing Deep Neural Networks (DNNs) running on edge hardware remains a challenge. Standard designs have been adopted by the community to facilitate the deployment of Neural Network models. However, not much emphasis is put on adapting the…

计算机视觉与模式识别 · 计算机科学 2022-08-24 Simon Narduzzi , Engin Türetken , Jean-Philippe Thiran , L. Andrea Dunbar

The increasing demand for large-scale visual data, coupled with strict privacy regulations, has driven research into anonymization methods that hide personal identities without seriously degrading data quality. In this paper, we explore the…

计算机视觉与模式识别 · 计算机科学 2025-05-28 Mustafa İzzet Muştu , Hazım Kemal Ekenel

Facial appearance editing is crucial for digital avatars, AR/VR, and personalized content creation, driving realistic user experiences. However, preserving identity with generative models is challenging, especially in scenarios with limited…

计算机视觉与模式识别 · 计算机科学 2025-03-10 MD Wahiduzzaman Khan , Mingshan Jia , Xiaolin Zhang , En Yu , Caifeng Shan , Kaska Musial-Gabrys

Flexible intelligent metasurfaces (FIMs) show great potential for improving the wireless network capacity in an energy-efficient manner. An FIM is a soft array consisting of several low-cost radiating elements. Each element can…

信息论 · 计算机科学 2025-03-12 Jiancheng An , Zhu Han , Dusit Niyato , Mérouane Debbah , Chau Yuen , Lajos Hanzo

Facial expression recognition is an essential task for various applications, including emotion detection, mental health analysis, and human-machine interactions. In this paper, we propose a multi-modal facial expression recognition method…

计算机视觉与模式识别 · 计算机科学 2023-03-21 Jun-Hwa Kim , Namho Kim , Chee Sun Won

Training of deep learning models for computer vision requires large image or video datasets from real world. Often, in collecting such datasets, we need to protect the privacy of the people captured in the images or videos, while still…

计算机视觉与模式识别 · 计算机科学 2019-02-13 Yuezun Li , Siwei Lyu

Recent works on convolutional neural networks (CNNs) for facial alignment have demonstrated unprecedented accuracy on a variety of large, publicly available datasets. However, the developed models are often both cumbersome and…

计算机视觉与模式识别 · 计算机科学 2019-06-12 TianXing Li , Zhi Yu , Edmund Phung , Brendan Duke , Irina Kezele , Parham Aarabi