English
Related papers

Related papers: Seeing Twice: How Side-by-Side T2I Comparison Chan…

200 papers

Focusing on text-to-image (T2I) generation, we propose Text and Image Mutual-Translation Adversarial Networks (TIME), a lightweight but effective model that jointly learns a T2I generator G and an image captioning discriminator D under the…

Computer Vision and Pattern Recognition · Computer Science 2020-12-24 Bingchen Liu , Kunpeng Song , Yizhe Zhu , Gerard de Melo , Ahmed Elgammal

Accessibility forums and, more recently, generative AI tools have become vital resources for blind users seeking solutions to computer-interaction issues and learning about new assistive technologies, screen reader features, tutorials, and…

Human-Computer Interaction · Computer Science 2026-02-24 Satwik Ram Kodandaram , Jiawei Zhou , Xiaojun Bi , IV Ramakrishnan , Vikas Ashok

Although generative AI is being deployed into classrooms with promises of aiding teachers, educators caution that these tools can have unintended pedagogical repercussions, including cultural misrepresentation and bias. These concerns are…

Despite advances in multimodal AI, current vision-based assistants often remain inefficient in collaborative tasks. We identify two key gulfs: a communication gulf, where users must translate rich parallel intentions into verbal commands…

Human-Computer Interaction · Computer Science 2026-03-16 Zhuyu Teng , Pei Chen , Yichen Cai , Ruoqing Lu , Zhaoqu Jiang , Jiayang Li , Weitao You , Lingyun Sun

Multimodal sarcasm detection has attracted growing interest due to the rise of multimedia posts on social media. Understanding sarcastic image-text posts often requires external contextual knowledge, such as cultural references or…

Computation and Language · Computer Science 2025-10-30 Soumyadeep Jana , Abhrajyoti Kundu , Sanasam Ranbir Singh

Recent advancements in generative models have significantly enhanced their capacity for image generation, enabling a wide range of applications such as image editing, completion and video editing. A specialized area within generative…

Computer Vision and Pattern Recognition · Computer Science 2025-01-14 Jiaxin Cheng , Zixu Zhao , Tong He , Tianjun Xiao , Yicong Zhou , Zheng Zhang

Text-to-image (T2I) models have rapidly advanced, enabling the generation of high-quality images from text prompts across various domains. However, these models present notable safety concerns, including the risk of generating harmful,…

Computation and Language · Computer Science 2025-07-28 Lijun Li , Zhelun Shi , Xuhao Hu , Bowen Dong , Yiran Qin , Xihui Liu , Lu Sheng , Jing Shao

Generative AI (GenAI) systems offer unprecedented opportunities for transforming professional and personal work, yet present challenges around prompting, evaluating and relying on outputs, and optimizing workflows. We argue that…

Human-Computer Interaction · Computer Science 2024-07-08 Lev Tankelevitch , Viktor Kewenig , Auste Simkute , Ava Elizabeth Scott , Advait Sarkar , Abigail Sellen , Sean Rintel

Generative AI tools are used to create art-like outputs and sometimes aid in the creative process. These tools have potential benefits for artists, but they also have the potential to harm the art workforce and infringe upon artistic and…

Computers and Society · Computer Science 2024-10-02 Juniper Lovato , Julia Zimmerman , Isabelle Smith , Peter Dodds , Jennifer Karson

Capturing professionals' decision-making in creative workflows (e.g., UI/UX) is essential for reflection, collaboration, and knowledge sharing, yet existing methods often leave rationales incomplete and implicit decisions hidden. To address…

Human-Computer Interaction · Computer Science 2026-02-26 Kihoon Son , DaEun Choi , Tae Soo Kim , Young-Ho Kim , Sangdoo Yun , Juho Kim

How can we use generative AI to design tools that augment rather than replace human cognition? In this position paper, we review our own research on AI-assisted decision-making for lessons to learn. We observe that in both AI-assisted…

Human-Computer Interaction · Computer Science 2025-04-07 Zelun Tony Zhang , Leon Reicherts

As generative AI (GenAI) increasingly permeates design workflows, its impact on design outcomes and designers' creative capabilities warrants investigation. We conducted a within-subjects experiment where we asked participants to design…

Human-Computer Interaction · Computer Science 2024-11-04 Yue Fu , Han Bin , Tony Zhou , Marx Wang , Yixin Chen , Zelia Gomes Da Costa Lai , Jacob O. Wobbrock , Alexis Hiniker

As generative AI (GenAI) systems become increasingly proficient at simulating human-like and well-reasoned text, users may attribute authority to AI outputs, shaping how they engage with writing and reasoning tasks. While prior work has…

Human-Computer Interaction · Computer Science 2026-05-18 Vitor H. A. Welzel , Nicholas Vincent

This paper introduces LACE, a co-creative system enabling professional artists to leverage generative AI through controlled prompting and iterative refinement within Photoshop. Addressing challenges in precision, iterative coherence, and…

Human-Computer Interaction · Computer Science 2025-04-22 YenKai Huang , Zheng Ning , Ming Cheng

Generative AI (GenAI) has revolutionized content generation, offering transformative capabilities for improving language coherence, readability, and overall quality. This manuscript explores the application of qualitative, quantitative, and…

Computation and Language · Computer Science 2024-11-28 Saman Sarraf

Generative AI's novel capacities raise questions about the future role of human expertise: does AI level the playing field between professional artists and laypeople, or does expertise enhance AI use? Do the cognitive skills experts make…

Human-Computer Interaction · Computer Science 2026-05-29 Thomas F. Eisenmann , Andres Karjus , Mar Canet Sola , Levin Brinkmann , Bramantyo Ibrahim Supriyatno , Iyad Rahwan

Unified multimodal generation architectures that jointly produce text and images have recently emerged as a promising direction for text-to-image (T2I) synthesis. However, many existing systems rely on explicit modality switching,…

Generative AI tools have lowered barriers to producing branded social media images and captions, yet small-business owners (SBOs) still struggle to create on-brand posts without access to professional designers or marketing consultants.…

Human-Computer Interaction · Computer Science 2026-04-14 Taehyun Yang , Eunhye Kim , Zhongzheng Xu , Fumeng Yang

Large Language Models (LLMs) have shown remarkable capabilities in environmental perception, reasoning-based decision-making, and simulating complex human behaviors, particularly in interactive role-playing contexts. This paper introduces…

Computation and Language · Computer Science 2026-01-21 Yin Cai , Zhouhong Gu , Zhaohan Du , Zheyu Ye , Shaosheng Cao , Yiqian Xu , Hongwei Feng , Ping Chen

Since NFTs and large generative models (such as DALLE2 and Stable Diffusion) have been publicly available, artists have seen their jobs threatened and stolen. While artists depend on sharing their art on online platforms such as Deviantart,…

Computers and Society · Computer Science 2024-06-14 Diego Porres , Alex Gomez-Villa
‹ Prev 1 8 9 10 Next ›