English
Related papers

Related papers: BlendScape: Enabling End-User Customization of Vid…

200 papers

People with visual impairments often struggle to create content that relies heavily on visual elements, particularly when conveying spatial and structural information. Existing accessible drawing tools, which construct images line by line,…

Human-Computer Interaction · Computer Science 2025-04-30 Seonghee Lee , Maho Kohga , Steve Landau , Sile O'Modhrain , Hari Subramonyam

While generative artificial intelligence (GenAI) is finding increased adoption in workplaces, current tools are primarily designed for individual use. Prior work established the potential for these tools to enhance personal creativity and…

Human-Computer Interaction · Computer Science 2026-02-25 Janet G. Johnson , Macarena Peralta , Mansanjam Kaur , Ruijie Sophia Huang , Sheng Zhao , Ruijia Guan , Shwetha Rajaram , Michael Nebeling

Despite their impressive visual fidelity, existing personalized image generators lack interactive control over spatial composition and scale poorly to multiple humans. To address these limitations, we present LayerComposer, an interactive…

Designers rely on visual search to explore and develop ideas in early design stages. However, designers can struggle to identify suitable text queries to initiate a search or to discover images for similarity-based search that can…

Human-Computer Interaction · Computer Science 2024-03-05 Kihoon Son , DaEun Choi , Tae Soo Kim , Young-Ho Kim , Juho Kim

Traditional visualisation designers often start with sketches before implementation. With generative AI, these sketches can be turned into AI-generated visualisations using specific prompts. However, guiding AI to create compelling visuals…

Human-Computer Interaction · Computer Science 2024-09-04 Aron E. Owen , Jonathan C. Roberts

Humans can intuitively compose and arrange scenes in the 3D space for photography. However, can advanced AI image generators plan scenes with similar 3D spatial awareness when creating images from text or image prompts? We present GenSpace,…

Computer Vision and Pattern Recognition · Computer Science 2025-06-09 Zehan Wang , Jiayang Xu , Ziang Zhang , Tianyu Pang , Chao Du , Hengshuang Zhao , Zhou Zhao

Automated live visual descriptions can aid blind people in understanding their surroundings with autonomy and independence. However, providing descriptions that are rich, contextual, and just-in-time has been a long-standing challenge in…

Human-Computer Interaction · Computer Science 2024-08-14 Ruei-Che Chang , Yuxuan Liu , Anhong Guo

Automatic presentation slide generation can greatly streamline content creation. However, since preferences of each user may vary, existing under-specified formulations often lead to suboptimal results that fail to align with individual…

Computation and Language · Computer Science 2025-12-24 Wenzheng Zeng , Mingyu Ouyang , Langyuan Cui , Hwee Tou Ng

3D content creation has long been a complex and time-consuming process, often requiring specialized skills and resources. While recent advancements have allowed for text-guided 3D object and scene generation, they still fall short of…

Computer Vision and Pattern Recognition · Computer Science 2024-08-06 Xingyi Li , Yizheng Wu , Jun Cen , Juewen Peng , Kewei Wang , Ke Xian , Zhe Wang , Zhiguo Cao , Guosheng Lin

Customizing pre-trained text-to-image generation model has attracted massive research interest recently, due to its huge potential in real-world applications. Although existing methods are able to generate creative content for a novel…

Computer Vision and Pattern Recognition · Computer Science 2024-05-10 Yufan Zhou , Ruiyi Zhang , Jiuxiang Gu , Tong Sun

As generative AI technologies are pressed into service in workplace settings, current approaches to account for the contexts in which such technologies are used fall short of users' expectations and needs. This paper empirically…

Computers and Society · Computer Science 2026-04-08 Emanuel Moss , Elizabeth Watkins , Christopher Persaud , Dawn Nafus , Passant Karunaratne , Mona Sloane

Recent advances in text-to-3D scene generation have demonstrated significant potential to transform content creation across multiple industries. Although the research community has made impressive progress in addressing the challenges of…

Customized text-to-image generation renders user-specified concepts into novel contexts based on textual prompts. Scaling the number of concepts in customized generation meets a broader demand for user creation, whereas existing methods…

Computer Vision and Pattern Recognition · Computer Science 2025-03-11 Jian Jin , Zhenbo Yu , Yang Shen , Zhenyong Fu , Jian Yang

Programming has become an essential component of K-12 education and serves as a pathway for developing computational thinking skills. Given the complexity of programming and the advanced skills it requires, previous research has introduced…

Human-Computer Interaction · Computer Science 2024-12-13 Yunnong Chen , Shuhong Xiao , Yaxuan Song , Zejian Li , Lingyun Sun , Liuqing Chen

Immersive environments have gradually become standard for visualizing and analyzing large or complex datasets that would otherwise be cumbersome, if not impossible, to explore through smaller scale computing devices. However, this type of…

Human-Computer Interaction · Computer Science 2019-03-12 Marco Cavallo , Mishal Dholakia , Matous Havlena , Kenneth Ocheltree , Mark Podlaseck

Chinese literati gatherings (Wenren Yaji), as a situated form of Chinese traditional culture, remain underexplored in depth. Although generative AI supports powerful multimodal generation, current cultural applications largely emphasize…

Human-Computer Interaction · Computer Science 2026-02-16 You Zhou , Bingyuan Wang , Hongcheng Guo , Rui Cao , Zeyu Wang

Video-based world models hold significant potential for generating high-quality embodied manipulation data. However, current video generation methods struggle to achieve stable long-horizon generation: classical diffusion-based approaches…

Computer Vision and Pattern Recognition · Computer Science 2025-09-29 Yu Shang , Lei Jin , Yiding Ma , Xin Zhang , Chen Gao , Wei Wu , Yong Li

Creating high-fidelity 3D models of indoor environments is essential for applications in design, virtual reality, and robotics. However, manual 3D modeling remains time-consuming and labor-intensive. While recent advances in generative AI…

Computer Vision and Pattern Recognition · Computer Science 2026-01-16 Chuan Fang , Heng Li , Yixun Liang , Jia Zheng , Yongsen Mao , Yuan Liu , Rui Tang , Zihan Zhou , Ping Tan

Generative AI has demonstrated significant potential in creative design, enabling the rapid generation of visual content and imaginative concepts. Although deep AI models achieve effective featurization in the latent space, navigating the…

Human-Computer Interaction · Computer Science 2026-04-23 Mingwei Li , Suyang Li , Daisuke Sakurai , Bei Wang , Remco Chang

Meeting online is becoming the new normal. Creating an immersive experience for online meetings is a necessity towards more diverse and seamless environments. Efficient photorealistic rendering of human 3D dynamics is the core of immersive…

Computer Vision and Pattern Recognition · Computer Science 2023-06-30 Chuanyue Shen , Letian Zhang , Zhangsihao Yang , Masood Mortazavi , Xiyun Song , Liang Peng , Heather Yu