中文
相关论文

相关论文: Exploiting Alpha Transparency In Language And Visi…

200 篇论文

The use of deep learning for human identification and object detection is becoming ever more prevalent in the surveillance industry. These systems have been trained to identify human body's or faces with a high degree of accuracy. However,…

计算机视觉与模式识别 · 计算机科学 2020-11-26 Morgan Frearson , Kien Nguyen

Deep hiding, embedding images into another using deep neural networks, has shown its great power in increasing the message capacity and robustness. In this paper, we conduct an in-depth study of state-of-the-art deep hiding schemes and…

密码学与安全 · 计算机科学 2021-06-08 Tao Xiang , Hangcheng Liu , Shangwei Guo , Tianwei Zhang

Deepfakes pose severe threats of visual misinformation to our society. One representative deepfake application is face manipulation that modifies a victim's facial attributes in an image, e.g., changing her age or hair color. The…

密码学与安全 · 计算机科学 2022-10-04 Zheng Li , Ning Yu , Ahmed Salem , Michael Backes , Mario Fritz , Yang Zhang

AI-powered generative models have significantly expanded the possibilities for editing, manipulating, and creating high-quality images. Particularly, images that falsely appear to originate from trusted sources pose a serious threat,…

密码学与安全 · 计算机科学 2026-04-28 Mathias Graf , Marco Willi , Melanie Mathys , Michael Aerni , Christian Schwarzer , Martin Melchior , Michael H. Graber

Explainable Artificial Intelligence (XAI) strategies play a crucial part in increasing the understanding and trustworthiness of neural networks. Nonetheless, these techniques could potentially generate misleading explanations. Blinding…

机器学习 · 计算机科学 2024-03-26 Md Abdul Kadir , GowthamKrishna Addluri , Daniel Sonntag

Recent advancements in AI-based multimedia generation have enabled the creation of hyper-realistic images and videos, raising concerns about their potential use in spreading misinformation. The widespread accessibility of generative…

计算机视觉与模式识别 · 计算机科学 2025-04-30 Joy Battocchio , Stefano Dell'Anna , Andrea Montibeller , Giulia Boato

The deployment of large language models (LLMs) in production environments has created an urgent need for observability systems that span the full stack -- from model internals to GPU kernels. Yet existing monitoring approaches address…

软件工程 · 计算机科学 2026-04-30 Twinkll Sisodia

Artificial Intelligence (AI) is rapidly integrating into various aspects of our daily lives, influencing decision-making processes in areas such as targeted advertising and matchmaking algorithms. As AI systems become increasingly…

人工智能 · 计算机科学 2025-03-11 Md. Tanzib Hosain , Mehedi Hasan Anik , Sadman Rafi , Rana Tabassum , Khaleque Insia , Md. Mehrab Siddiky

Multimodal AI systems have achieved remarkable performance across a broad range of real-world tasks, yet the mechanisms underlying visual-language reasoning remain surprisingly poorly understood. We report three findings that challenge…

The rapid advancement of generative AI has enabled the mass production of photorealistic synthetic images, blurring the boundary between authentic and fabricated visual content. This challenge is particularly evident in deepfake scenarios…

计算机视觉与模式识别 · 计算机科学 2025-09-30 Minsun Jeon , Simon S. Woo

We introduce the first method, to the best of our knowledge, for adapting image-to-video models to layer-aware text (glyph) animation, a capability critical for practical dynamic visual design. Existing approaches predominantly handle the…

计算机视觉与模式识别 · 计算机科学 2026-03-20 Fei Zhang , Zijian Zhou , Bohao Tang , Sen He , Hang Li , Zhe Wang , Soubhik Sanyal , Pengfei Liu , Viktar Atliha , Tao Xiang , Frost Xu , Semih Gunel

Vision-language models (VLMs) have achieved strong performance across diverse multimodal tasks, but their adversarial robustness in visible-infrared (VIS-IR) scenarios remains underexplored. This gap is critical because VIS-IR sensing is…

计算机视觉与模式识别 · 计算机科学 2026-05-22 Xiang Chen , Yuxian Dong , Chao Li , Chengyin Hu , Jiaju Han , Fengyu Zhang , Yiwei Wei , Jiahuan Long , Jiujiang Guo

We assess the vulnerabilities of deep face recognition systems for images that falsify/spoof multiple identities simultaneously. We demonstrate that, by manipulating the deep feature representation extracted from a face image via…

计算机视觉与模式识别 · 计算机科学 2021-10-05 Takuma Amada , Seng Pei Liew , Kazuya Kakizaki , Toshinori Araki

The existence of adversarial attacks on convolutional neural networks (CNN) questions the fitness of such models for serious applications. The attacks manipulate an input image such that misclassification is evoked while still looking…

计算机视觉与模式识别 · 计算机科学 2022-08-25 Mohammadreza Amirian , Friedhelm Schwenker , Thilo Stadelmann

The vulnerability of deep neural networks to adversarial attacks has been widely demonstrated (e.g., adversarial example attacks). Traditional attacks perform unstructured pixel-wise perturbation to fool the classifier. An alternative…

机器学习 · 计算机科学 2022-05-23 Shuo Wang , Surya Nepal , Carsten Rudolph , Marthie Grobler , Shangyu Chen , Tianle Chen

Face morphing attacks threaten biometric verification, yet most morphing attack detection (MAD) systems require task-specific training and generalize poorly to unseen attack types. Meanwhile, open-source multimodal large language models…

计算机视觉与模式识别 · 计算机科学 2026-02-18 Marija Ivanovska , Vitomir Štruc

Physical adversarial attacks against deep neural networks (DNNs) have recently gained increasing attention. The current mainstream physical attacks use printed adversarial patches or camouflage to alter the appearance of the target object.…

计算机视觉与模式识别 · 计算机科学 2023-07-18 Donghua Wang , Wen Yao , Tingsong Jiang , Chao Li , Xiaoqian Chen

With the significant advances in deep generative models for image and video synthesis, Deepfakes and manipulated media have raised severe societal concerns. Conventional machine learning classifiers for deepfake detection often fail to cope…

计算机视觉与模式识别 · 计算机科学 2024-10-14 Aakash Varma Nadimpalli , Ajita Rattani

The rapid advancement of text-to-image generation systems, exemplified by models like Stable Diffusion, Midjourney, Imagen, and DALL-E, has heightened concerns about their potential misuse. In response, companies like Meta and Google have…

计算机视觉与模式识别 · 计算机科学 2024-08-21 Niyar R Barman , Krish Sharma , Ashhar Aziz , Shashwat Bajpai , Shwetangshu Biswas , Vasu Sharma , Vinija Jain , Aman Chadha , Amit Sheth , Amitava Das

Generative large vision-language models (LVLMs) have recently achieved impressive performance gains, and their user base is growing rapidly. However, the security of LVLMs, in particular in a long-context multi-turn setting, is largely…

计算机视觉与模式识别 · 计算机科学 2026-02-19 Christian Schlarmann , Matthias Hein