中文
相关论文

相关论文: The Visual Counter Turing Test (VCT2): A Benchmark…

200 篇论文

Contemporary Text-to-Image (T2I) models frequently depend on qualitative human evaluations to assess the consistency between synthesized images and the text prompts. There is a demand for quantitative and automatic evaluation tools, given…

计算机视觉与模式识别 · 计算机科学 2024-12-10 Ziyuan Qin , Dongjie Cheng , Haoyu Wang , Huahui Yi , Yuting Shao , Zhiyuan Fan , Kang Li , Qicheng Lao

The spreading of AI-generated images (AIGI), driven by advances in generative AI, poses a significant threat to information security and public trust. Existing AIGI detectors, while effective against images in clean laboratory settings,…

计算机视觉与模式识别 · 计算机科学 2025-08-20 Cheng Xia , Manxi Lin , Jiexiang Tan , Xiaoxiong Du , Yang Qiu , Junjun Zheng , Xiangheng Kong , Yuning Jiang , Bo Zheng

Traditional supervised methods for detecting AI-generated images depend on large, curated datasets for training and fail to generalize to novel, out-of-domain image generators. As an alternative, we explore pre-trained Vision-Language…

机器学习 · 计算机科学 2026-01-27 Zoher Kachwala , Danishjeet Singh , Danielle Yang , Filippo Menczer

Recent progress in Text-to-Image (T2I) generative models has enabled high-quality image generation. As performance and accessibility increase, these models are gaining significant attraction and popularity: ensuring their fairness and…

计算机视觉与模式识别 · 计算机科学 2024-08-30 Moreno D'Incà , Elia Peruzzo , Massimiliano Mancini , Xingqian Xu , Humphrey Shi , Nicu Sebe

With the rapid advancements in AI-Generated Content (AIGC), AI-Generated Images (AIGIs) have been widely applied in entertainment, education, and social media. However, due to the significant variance in quality among different AIGIs, there…

计算机视觉与模式识别 · 计算机科学 2024-04-05 Chunyi Li , Tengchuan Kou , Yixuan Gao , Yuqin Cao , Wei Sun , Zicheng Zhang , Yingjie Zhou , Zhichao Zhang , Weixia Zhang , Haoning Wu , Xiaohong Liu , Xiongkuo Min , Guangtao Zhai

The rapid proliferation of highly realistic AI-generated images poses serious security threats such as misinformation and identity fraud. Detecting generated images in open-world settings is particularly challenging when they originate from…

密码学与安全 · 计算机科学 2026-01-19 Li Wang , Wenyu Chen , Xiangtao Meng , Zheng Li , Shanqing Guo

We provide a new multi-task benchmark for evaluating text-to-image models. We perform a human evaluation comparing the most common open-source (Stable Diffusion) and commercial (DALL-E 2) models. Twenty computer science AI graduate students…

While recent research suggests Large Language Models match human creative performance in divergent thinking tasks, visual creativity remains underexplored. This study compared image generation in human participants (Visual Artists and Non…

Popular text-to-image (T2I) systems are trained on web-scraped data, which is heavily Amero and Euro-centric, underrepresenting the cultures of the Global South. To analyze these biases, we introduce CuRe, a novel and scalable benchmarking…

计算机视觉与模式识别 · 计算机科学 2025-06-11 Aniket Rege , Zinnia Nie , Mahesh Ramesh , Unmesh Raskar , Zhuoran Yu , Aditya Kusupati , Yong Jae Lee , Ramya Korlakai Vinayak

Image encoders, a fundamental component of vision-language models (VLMs), are typically pretrained independently before being aligned with a language model. This standard paradigm results in encoders that process images agnostically,…

计算机视觉与模式识别 · 计算机科学 2025-11-27 Raghuveer Thirukovalluru , Xiaochuang Han , Bhuwan Dhingra , Emily Dinan , Maha Elbayad

We study human AI-detection behaviour at scale using a year of activity from r/RealOrAI, a Reddit community where users collaboratively assess whether visual media is real or AI-generated. The community is moderated by a bot that solicits…

社会与信息网络 · 计算机科学 2026-05-26 Tuğrulcan Elmas

New advancements for the detection of synthetic images are critical for fighting disinformation, as the capabilities of generative AI models continuously evolve and can lead to hyper-realistic synthetic imagery at unprecedented scale and…

计算机视觉与模式识别 · 计算机科学 2023-04-25 Pantelis Dogoulis , Giorgos Kordopatis-Zilos , Ioannis Kompatsiaris , Symeon Papadopoulos

The unprecedented photorealistic results achieved by recent text-to-image generative systems and their increasing use as plug-and-play content creation solutions make it crucial to understand their potential biases. In this work, we…

计算机视觉与模式识别 · 计算机科学 2024-03-19 Melissa Hall , Candace Ross , Adina Williams , Nicolas Carion , Michal Drozdzal , Adriana Romero Soriano

This paper reports on the NTIRE 2024 Quality Assessment of AI-Generated Content Challenge, which will be held in conjunction with the New Trends in Image Restoration and Enhancement Workshop (NTIRE) at CVPR 2024. This challenge is to…

计算机视觉与模式识别 · 计算机科学 2024-05-08 Xiaohong Liu , Xiongkuo Min , Guangtao Zhai , Chunyi Li , Tengchuan Kou , Wei Sun , Haoning Wu , Yixuan Gao , Yuqin Cao , Zicheng Zhang , Xiele Wu , Radu Timofte , Fei Peng , Huiyuan Fu , Anlong Ming , Chuanming Wang , Huadong Ma , Shuai He , Zifei Dou , Shu Chen , Huacong Zhang , Haiyi Xie , Chengwei Wang , Baoying Chen , Jishen Zeng , Jianquan Yang , Weigang Wang , Xi Fang , Xiaoxin Lv , Jun Yan , Tianwu Zhi , Yabin Zhang , Yaohui Li , Yang Li , Jingwen Xu , Jianzhao Liu , Yiting Liao , Junlin Li , Zihao Yu , Yiting Lu , Xin Li , Hossein Motamednia , S. Farhad Hosseini-Benvidi , Fengbin Guan , Ahmad Mahmoudi-Aznaveh , Azadeh Mansouri , Ganzorig Gankhuyag , Kihwan Yoon , Yifang Xu , Haotian Fan , Fangyuan Kong , Shiling Zhao , Weifeng Dong , Haibing Yin , Li Zhu , Zhiling Wang , Bingchen Huang , Avinab Saha , Sandeep Mishra , Shashank Gupta , Rajesh Sureddi , Oindrila Saha , Luigi Celona , Simone Bianco , Paolo Napoletano , Raimondo Schettini , Junfeng Yang , Jing Fu , Wei Zhang , Wenzhi Cao , Limei Liu , Han Peng , Weijun Yuan , Zhan Li , Yihang Cheng , Yifan Deng , Haohui Li , Bowen Qu , Yao Li , Shuqing Luo , Shunzhou Wang , Wei Gao , Zihao Lu , Marcos V. Conde , Xinrui Wang , Zhibo Chen , Ruling Liao , Yan Ye , Qiulin Wang , Bing Li , Zhaokun Zhou , Miao Geng , Rui Chen , Xin Tao , Xiaoyu Liang , Shangkun Sun , Xingyuan Ma , Jiaze Li , Mengduo Yang , Haoran Xu , Jie Zhou , Shiding Zhu , Bohan Yu , Pengfei Chen , Xinrui Xu , Jiabin Shen , Zhichao Duan , Erfan Asadi , Jiahe Liu , Qi Yan , Youran Qu , Xiaohui Zeng , Lele Wang , Renjie Liao

\underline{AI} \underline{G}enerated \underline{C}ontent (\textbf{AIGC}) has gained widespread attention with the increasing efficiency of deep learning in content creation. AIGC, created with the assistance of artificial intelligence…

计算机视觉与模式识别 · 计算机科学 2023-03-23 Zicheng Zhang , Chunyi Li , Wei Sun , Xiaohong Liu , Xiongkuo Min , Guangtao Zhai

In today's day and age, we face a challenge in detecting deepfake images because of the fast evolution of modern generative models and the poor generalization capability of existing methods. In this paper, we use an ensemble of fine-tuned…

计算机视觉与模式识别 · 计算机科学 2026-04-21 Kaliki V Srinanda , M Manvith Prabhu , Hemanth K Mogilipalem , Jayavarapu S Abhinai , Vaibhav Santhosh , Aryan Herur , Deepu Vijayasenan

Visual counting is a fundamental yet challenging task, especially when users need to count objects of a specific type in complex scenes. While recent models, including class-agnostic counting models and large vision-language models (VLMs),…

计算机视觉与模式识别 · 计算机科学 2025-09-18 Gia Khanh Nguyen , Yifeng Huang , Minh Hoai

We present VB, a benchmark that tests whether vision-language models can determine what is and is not visible in a photograph, and abstain when a human viewer cannot reliably answer. Each item pairs a single photo with a short yes/no…

计算机视觉与模式识别 · 计算机科学 2026-03-10 Neil Tripathi

In this paper, we design and train a Generative Image-to-text Transformer, GIT, to unify vision-language tasks such as image/video captioning and question answering. While generative models provide a consistent network architecture between…

计算机视觉与模式识别 · 计算机科学 2022-12-19 Jianfeng Wang , Zhengyuan Yang , Xiaowei Hu , Linjie Li , Kevin Lin , Zhe Gan , Zicheng Liu , Ce Liu , Lijuan Wang

With the rapid advancement of generative AI, it is now possible to synthesize high-quality images in a few seconds. Despite the power of these technologies, they raise significant concerns regarding misuse. Current efforts to distinguish…

计算机视觉与模式识别 · 计算机科学 2025-03-25 Siyuan Cheng , Lingjuan Lyu , Zhenting Wang , Xiangyu Zhang , Vikash Sehwag
‹ 上一页 1 8 9 10 下一页 ›