中文
相关论文

相关论文: MIDV-2020: A Comprehensive Benchmark Dataset for I…

200 篇论文

Steel surface defect analysis is critical for industrial quality control, yet existing benchmarks rely primarily on label-only annotations, limiting fine-grained semantic understanding and systematic evaluation of vision-language models. To…

计算机视觉与模式识别 · 计算机科学 2026-05-11 Shuxian Zhao , Jie Gui , Baosheng Yu , Dacheng Tao

The rapid advancement of generative AI has raised concerns about the authenticity of digital images, as highly realistic fake images can now be generated at low cost, potentially increasing societal risks. In response, several datasets have…

计算机视觉与模式识别 · 计算机科学 2026-02-12 Hanzhe Yu , Yun Ye , Jintao Rong , Qi Xuan , Chen Ma

The misuse of advanced generative AI models has resulted in the widespread proliferation of falsified data, particularly forged human-centric audiovisual content, which poses substantial societal risks (e.g., financial fraud and social…

Advances in radiance fields have enabled photorealistic novel view synthesis. In several domains, large-scale real-world datasets have been developed to support comprehensive benchmarking and to facilitate progress beyond scene-specific…

计算机视觉与模式识别 · 计算机科学 2026-04-16 Cheng-You Lu , Yi-Shan Hung , Wei-Ling Chi , Hao-Ping Wang , Charlie Li-Ting Tsai , Yu-Cheng Chang , Yu-Lun Liu , Thomas Do , Chin-Teng Lin

Human head detection, keypoint estimation, and 3D head model fitting are essential tasks with many applications. However, traditional real-world datasets often suffer from bias, privacy, and ethical concerns, and they have been recorded in…

计算机视觉与模式识别 · 计算机科学 2024-12-06 Orest Kupyn , Eugene Khvedchenia , Christian Rupprecht

Being data-driven is one of the most iconic properties of deep learning algorithms. The birth of ImageNet drives a remarkable trend of "learning from large-scale data" in computer vision. Pretraining on ImageNet to obtain rich universal…

Amid the proliferation of forged images, notably the tsunami of deepfake content, extensive research has been conducted on using artificial intelligence (AI) to identify forged content in the face of continuing advancements in…

计算机视觉与模式识别 · 计算机科学 2024-07-29 Shuhan Cui , Huy H. Nguyen , Trung-Nghia Le , Chun-Shien Lu , Isao Echizen

Humans are arguably one of the most important subjects in video streams, many real-world applications such as video summarization or video editing workflows often require the automatic search and retrieval of a person of interest. Despite…

计算机视觉与模式识别 · 计算机科学 2021-06-04 Juan Leon Alcazar , Long Mai , Federico Perazzi , Joon-Young Lee , Pablo Arbelaez , Bernard Ghanem , Fabian Caba Heilbron

Identity authentication is the process of verifying one's identity. There are several identity authentication methods, among which biometric authentication is of utmost importance. Facial recognition is a sort of biometric authentication…

计算机视觉与模式识别 · 计算机科学 2022-12-15 Matineh Pooshideh

Within the field of image and video recognition, the traditional approach is a dataset split into fixed training and test partitions. However, the labelling of the training set is time-consuming, especially as datasets grow in size and…

计算机视觉与模式识别 · 计算机科学 2016-12-08 Andrew Gilbert , Richard Bowden

Because of the explosive growth of face photos as well as their widespread dissemination and easy accessibility in social media, the security and privacy of personal identity information becomes an unprecedented challenge. Meanwhile, the…

计算机视觉与模式识别 · 计算机科学 2021-03-03 Yunqian Wen , Li Song , Bo Liu , Ming Ding , Rong Xie

The increasing use of digital technologies and mobile-based registration procedures highlights the vital role of personal identity documents (IDs) in verifying users and safeguarding sensitive information. However, the rise in counterfeit…

计算机视觉与模式识别 · 计算机科学 2024-07-02 Musab Al-Ghadi , Joris Voerman , Souhail Bakkali , Mickaël Coustaty , Nicolas Sidere , Xavier St-Georges

The rapid advancement of AI-generated multimodal video-audio content has raised significant concerns regarding information security and content authenticity. Existing synthetic video datasets predominantly focus on the visual modality…

计算机视觉与模式识别 · 计算机科学 2026-04-21 Mengxue Hu , Yunfeng Diao , Changtao Miao , Zhiqing Guo , Jianshu Li , Zhe Li , Joey Tianyi Zhou

The wearing of the face masks appears as a solution for limiting the spread of COVID-19. In this context, efficient recognition systems are expected for checking that people faces are masked in regulated areas. To perform this task, a large…

计算机视觉与模式识别 · 计算机科学 2020-12-07 Adnane Cabani , Karim Hammoudi , Halim Benhabiles , Mahmoud Melkemi

Video Camouflaged Object Detection (VCOD) is a challenging task which aims to identify objects that seamlessly concealed within the background in videos. The dynamic properties of video enable detection of camouflaged objects through motion…

计算机视觉与模式识别 · 计算机科学 2025-06-02 Shuyong Gao , Yu'ang Feng , Qishan Wang , Lingyi Hong , Xinyu Zhou , Liu Fei , Yan Wang , Wenqiang Zhang

Face morphing attack detection (MAD) algorithms have become essential to overcome the vulnerability of face recognition systems. To solve the lack of large-scale and public-available datasets due to privacy concerns and restrictions, in…

计算机视觉与模式识别 · 计算机科学 2025-03-25 Haoyu Zhang , Raghavendra Ramachandra , Kiran Raja , Christoph Busch

Face detection is a long-standing challenge in the field of computer vision, with the ultimate goal being to accurately localize human faces in an unconstrained environment. There are significant technical hurdles in making these systems…

计算机视觉与模式识别 · 计算机科学 2021-11-03 Necdet Gurkan , Jordan W. Suchow

In this paper, we present our approach to the DataCV ICCV Challenge, which centers on building a high-quality face dataset to train a face recognition model. The constructed dataset must not contain identities overlapping with any existing…

计算机视觉与模式识别 · 计算机科学 2025-08-19 Feiran Li , Qianqian Xu , Shilong Bao , Boyu Han , Zhiyong Yang , Qingming Huang

Anomaly detection (AD) aims to identify defects using normal-only training data. Existing anomaly detection benchmarks (e.g., MVTec-AD with 15 categories) cover only a narrow range of categories, limiting the evaluation of cross-context…

计算机视觉与模式识别 · 计算机科学 2025-11-26 Hai Ling , Jia Guo , Zhulin Tao , Yunkang Cao , Donglin Di , Hongyan Xu , Xiu Su , Yang Song , Lei Fan

Visual persuasion, which uses visual elements to influence cognition and behaviors, is crucial in fields such as advertising and political communication. With recent advancements in artificial intelligence, there is growing potential to…

计算与语言 · 计算机科学 2025-10-29 Junseo Kim , Jongwook Han , Dongmin Choi , Jongwook Yoon , Eun-Ju Lee , Yohan Jo