English
Related papers

Related papers: PosterIQ: A Design Perspective Benchmark for Poste…

200 papers

Generative AI is widely used to create commercial posters. However, rapid advances in generation have outpaced automated quality assessment. Existing models emphasize generic esthetics or low level distortions and lack the functional…

Computer Vision and Pattern Recognition · Computer Science 2026-02-26 Meiqi Sun , Mingyu Li , Junxiong Zhu

Computational Design approaches facilitate the generation of typographic design, but evaluating these designs remains a challenging task. In this paper, we propose a set of heuristic metrics for typographic design evaluation, focusing on…

Multimedia · Computer Science 2024-02-13 Sérgio M. Rebelo , J. J. Merelo , João Bicker , Penousal Machado

Recent advancements in the text-rendering capabilities of image generation models have made the end-to-end creation of graphic design content, such as posters, increasingly feasible. However, existing reward models fall short of accurately…

Content-aware visual-textual presentation layout aims at arranging spatial space on the given canvas for pre-defined elements, including text, logo, and underlay, which is a key to automatic template-free creative graphic design. In…

Computer Vision and Pattern Recognition · Computer Science 2023-03-29 HsiaoYuan Hsu , Xiangteng He , Yuxin Peng , Hao Kong , Qing Zhang

Visual layout plays a critical role in graphic design fields such as advertising, posters, and web UI design. The recent trend towards content-aware layout generation through generative models has shown promise, yet it often overlooks the…

Computer Vision and Pattern Recognition · Computer Science 2024-07-30 Jaejung Seol , Seojun Kim , Jaejun Yoo

Academic poster generation is a crucial yet challenging task in scientific communication, requiring the compression of long-context interleaved documents into a single, visually coherent page. To address this challenge, we introduce the…

Computer Vision and Pattern Recognition · Computer Science 2025-10-31 Wei Pang , Kevin Qinghong Lin , Xiangru Jian , Xi He , Philip Torr

Generating high-quality images without prompt engineering expertise remains a challenge for text-to-image (T2I) models, which often misinterpret poorly structured prompts, leading to distortions and misalignments. While humans easily…

Computer Vision and Pattern Recognition · Computer Science 2025-05-13 Nisan Chhetri , Arpan Sainju

Posters play a crucial role in marketing and advertising by enhancing visual communication and brand visibility, making significant contributions to industrial design. With the latest advancements in controllable T2I diffusion models,…

Computer Vision and Pattern Recognition · Computer Science 2025-02-13 Jian Ma , Yonglin Deng , Chen Chen , Nanyang Du , Haonan Lu , Zhenyu Yang

Text design is one of the most critical procedures in poster design, as it relies heavily on the creativity and expertise of humans to design text images considering the visual harmony and text-semantic. This study introduces TextPainter, a…

Computer Vision and Pattern Recognition · Computer Science 2023-08-15 Yifan Gao , Jinpeng Lin , Min Zhou , Chuanbin Liu , Hongtao Xie , Tiezheng Ge , Yuning Jiang

Product posters, which integrate subject, scene, and text, are crucial promotional tools for attracting customers. Creating such posters using modern image generation methods is valuable, while the main challenge lies in accurately…

Computer Vision and Pattern Recognition · Computer Science 2025-04-10 Yifan Gao , Zihang Lin , Chuanbin Liu , Min Zhou , Tiezheng Ge , Bo Zheng , Hongtao Xie

Generating aesthetic posters is more challenging than simple design images: it requires not only precise text rendering but also the seamless integration of abstract artistic content, striking layouts, and overall stylistic harmony. To…

Computer Vision and Pattern Recognition · Computer Science 2025-06-13 SiXiang Chen , Jianyu Lai , Jialin Gao , Tian Ye , Haoyu Chen , Hengyu Shi , Shitong Shao , Yunlong Lin , Song Fei , Zhaohu Xing , Yeying Jin , Junfeng Luo , Xiaoming Wei , Lei Zhu

Multi-agent systems built upon large language models (LLMs) have demonstrated remarkable capabilities in tackling complex compositional tasks. In this work, we apply this paradigm to the paper-to-poster generation problem, a practical yet…

Artificial Intelligence · Computer Science 2026-04-14 Zhilin Zhang , Xiang Zhang , Jiaqi Wei , Yiwei Xu , Chenyu You

Poster design is a critical medium for visual communication. Prior work has explored automatic poster design using deep learning techniques, but these approaches lack text accuracy, user customization, and aesthetic appeal, limiting their…

Graphics · Computer Science 2025-03-20 Haoyu Chen , Xiaojie Xu , Wenbo Li , Jingjing Ren , Tian Ye , Songhua Liu , Ying-Cong Chen , Lei Zhu , Xinchao Wang

Unified multimodal models integrate the reasoning capacity of large language models with both image understanding and generation, showing great promise for advanced multimodal intelligence. However, the community still lacks a rigorous…

Computer Vision and Pattern Recognition · Computer Science 2026-03-24 Hongxiang Li , Yaowei Li , Bin Lin , Yuwei Niu , Yuhang Yang , Xiaoshuang Huang , Jiayin Cai , Xiaolong Jiang , Yao Hu , Long Chen

Image-to-poster generation is a high-demand task requiring not only local adjustments but also high-level design understanding. Models must generate text, layout, style, and visual elements while preserving semantic fidelity and aesthetic…

Computer Vision and Pattern Recognition · Computer Science 2026-02-13 Sixiang Chen , Jianyu Lai , Jialin Gao , Hengyu Shi , Zhongying Liu , Tian Ye , Junfeng Luo , Xiaoming Wei , Lei Zhu

Recent progress in Vision Language Models (VLMs) has raised the question of whether they can reliably perform nonverbal reasoning. To this end, we introduce VRIQ (Visual Reasoning IQ), a novel benchmark designed to assess and analyze the…

Computer Vision and Pattern Recognition · Computer Science 2026-02-06 Tina Khezresmaeilzadeh , Jike Zhong , Konstantinos Psounis

Automated academic poster generation aims to distill lengthy research papers into concise, visually coherent presentations. Existing Multimodal Large Language Models (MLLMs) based approaches, however, suffer from three critical limitations:…

Computer Vision and Pattern Recognition · Computer Science 2026-03-03 Wenxin Tang , Jingyu Xiao , Yanpei Gong , Fengyuan Ran , Tongchuan Xia , Junliang Liu , Man Ho Lam , Wenxuan Wang , Michael R. Lyu

Text-to-image (T2I) generation aims to synthesize images from textual prompts, which jointly specify what must be shown and imply what can be inferred, which thus correspond to two core capabilities: \textbf{\textit{composition}} and…

Computer Vision and Pattern Recognition · Computer Science 2026-03-03 Ouxiang Li , Yuan Wang , Xinting Hu , Huijuan Huang , Rui Chen , Jiarong Ou , Xin Tao , Pengfei Wan , Xiaojuan Qi , Fuli Feng

Compositionality is a critical capability in Text-to-Image (T2I) models, as it reflects their ability to understand and combine multiple concepts from text descriptions. Existing evaluations of compositional capability rely heavily on…

Computer Vision and Pattern Recognition · Computer Science 2024-08-27 Xindi Wu , Dingli Yu , Yangsibo Huang , Olga Russakovsky , Sanjeev Arora

A molecule's properties are fundamentally determined by its composition and structure encoded in its molecular graph. Thus, reasoning about molecular properties requires the ability to parse and understand the molecular graph. Large…

Machine Learning · Computer Science 2026-01-22 Christoph Bartmann , Johannes Schimunek , Mykyta Ielanskyi , Philipp Seidl , Günter Klambauer , Sohvi Luukkonen
‹ Prev 1 2 3 10 Next ›