English
Related papers

Related papers: OASIS Uncovers: High-Quality T2I Models, Same Old …

200 papers

Text-to-image models, such as Stable Diffusion (SD), undergo iterative updates to improve image quality and address concerns such as safety. Improvements in image quality are straightforward to assess. However, how model updates resolve…

Cryptography and Security · Computer Science 2024-09-02 Yixin Wu , Yun Shen , Michael Backes , Yang Zhang

Evaluating the quality of synthesized images remains a significant challenge in the development of text-to-image (T2I) generation. Most existing studies in this area primarily focus on evaluating text-image alignment, image quality, and…

Computer Vision and Pattern Recognition · Computer Science 2026-04-20 Ziwei Huang , Wanggui He , Quanyu Long , Yandi Wang , Haoyuan Li , Zhelun Yu , Fangxun Shu , Long Chan , Hao Jiang , Fei Wu , Leilei Gan

Text-to-image (T2I) diffusion models have gained widespread application across various domains, demonstrating remarkable creative potential. However, the strong generalization capabilities of these models can inadvertently led they to…

Computer Vision and Pattern Recognition · Computer Science 2025-02-19 Die Chen , Zhiwen Li , Cen Chen , Xiaodan Li , Jinyan Ye

State-of-the-art generative text-to-image models are known to exhibit social biases and over-represent certain groups like people of perceived lighter skin tones and men in their outcomes. In this work, we propose a method to mitigate such…

Computer Vision and Pattern Recognition · Computer Science 2023-10-12 Piero Esposito , Parmida Atighehchian , Anastasis Germanidis , Deepti Ghadiyaram

How the general public perceives scientists has been of interest to science educators for decades. While there can be many factors of it, the impact of recent generative artificial intelligence (AI) models is noteworthy, as these are…

Physics Education · Physics 2025-04-29 Gyeonggeon Lee

This work presents a novel strategy to measure bias in text-to-image models. Using paired prompts that specify gender and vaguely reference an object (e.g. "a man/woman holding an item") we can examine whether certain objects are associated…

Computer Vision and Pattern Recognition · Computer Science 2023-07-18 Harvey Mannering

In image classification, a lot of development has happened in detecting out-of-distribution (OoD) data. However, most OoD detection methods are evaluated on a standard set of datasets, arbitrarily different from training data. There is no…

Computer Vision and Pattern Recognition · Computer Science 2022-09-27 Jishnu Mukhoti , Tsung-Yu Lin , Bor-Chun Chen , Ashish Shah , Philip H. S. Torr , Puneet K. Dokania , Ser-Nam Lim

Image Quality Assessment (IQA) measures and predicts perceived image quality by human observers. Although recent studies have highlighted the critical influence that variations in the scale of an image have on its perceived quality, this…

Computer Vision and Pattern Recognition · Computer Science 2025-08-14 Vlad Hosu , Lorenzo Agnolucci , Daisuke Iso , Dietmar Saupe

Recent text-to-image generative models can generate high-fidelity images from text inputs, but the quality of these generated images cannot be accurately evaluated by existing evaluation metrics. To address this issue, we introduce Human…

Computer Vision and Pattern Recognition · Computer Science 2023-09-26 Xiaoshi Wu , Yiming Hao , Keqiang Sun , Yixiong Chen , Feng Zhu , Rui Zhao , Hongsheng Li

There are not one but two dimensions of bias that can be revealed through the study of large AI models: not only bias in training data or the products of an AI, but also bias in society, such as disparity in employment or health outcomes…

Computers and Society · Computer Science 2025-04-02 Marinus Ferreira

Evaluation of Text to Speech (TTS) systems is challenging and resource-intensive. Subjective metrics such as Mean Opinion Score (MOS) are not easily comparable between works. Objective metrics are frequently used, but rarely validated…

Sound · Computer Science 2026-03-03 Christoph Minixhofer , Ondrej Klejch , Peter Bell

Large Language Models (LLM) have made significant advances in the recent past becoming more mainstream in Artificial Intelligence (AI) enabled human-facing applications. However, LLMs often generate stereotypical output inherited from…

Computation and Language · Computer Science 2023-11-27 Wu Zekun , Sahan Bulathwela , Adriano Soares Koshiyama

Text to image generation methods (T2I) are widely popular in generating art and other creative artifacts. While visual hallucinations can be a positive factor in scenarios where creativity is appreciated, such artifacts are poorly suited…

Computer Vision and Pattern Recognition · Computer Science 2023-05-30 Rodrigo Valerio , Joao Bordalo , Michal Yarom , Yonatan Bitton , Idan Szpektor , Joao Magalhaes

This paper examines the limitations of advanced text-to-image models in accurately rendering unconventional concepts which are scarcely represented or absent in their training datasets. We identify how these limitations not only confine the…

Computer Vision and Pattern Recognition · Computer Science 2024-05-31 Jiyoon Myung , Jihyeon Park

With the advancement of generative models, the assessment of generated images becomes more and more important. Previous methods measure distances between features of reference and generated images from trained vision models. In this paper,…

Computer Vision and Pattern Recognition · Computer Science 2024-04-10 Jaehui Hwang , Junghyuk Lee , Jong-Seok Lee

Despite the tremendous success of neural networks, benign images can be corrupted by adversarial perturbations to deceive these models. Intriguingly, images differ in their attackability. Specifically, given an attack configuration, some…

Computer Vision and Pattern Recognition · Computer Science 2026-04-28 Jiaming Liang , Haowei Liu , Chi-Man Pun

Social media has exacerbated the promotion of Western beauty norms, leading to negative self-image, particularly in women and girls, and causing harm such as body dysmorphia. Increasingly content on the internet has been artificially…

Computer Vision and Pattern Recognition · Computer Science 2025-11-06 Tanvi Dinkar , Aiqi Jiang , Gavin Abercrombie , Ioannis Konstas

Text-to-image (T2I) diffusion models have revolutionized generative modeling by producing high-fidelity, diverse, and visually realistic images from textual prompts. Despite these advances, existing models struggle with complex prompts…

Computer Vision and Pattern Recognition · Computer Science 2024-11-27 Eric Hanchen Jiang , Yasi Zhang , Zhi Zhang , Yixin Wan , Andrew Lizarraga , Shufan Li , Ying Nian Wu

Bias and stereotypes in language models can cause harm, especially in sensitive areas like content moderation and decision-making. This paper addresses bias and stereotype detection by exploring how jointly learning these tasks enhances…

Computation and Language · Computer Science 2025-07-03 Aditya Tomar , Rudra Murthy , Pushpak Bhattacharyya

Following the initial excitement, Text-to-Image (TTI) models are now being examined more critically. While much of the discourse has focused on biases and stereotypes embedded in large-scale training datasets, the sociotechnical dynamics of…

Human-Computer Interaction · Computer Science 2025-04-22 Maria-Teresa De Rosa Palmini , Eva Cetinic