中文
相关论文

相关论文: PosterIQ: A Design Perspective Benchmark for Poste…

200 篇论文

Generative AI is widely used to create commercial posters. However, rapid advances in generation have outpaced automated quality assessment. Existing models emphasize generic esthetics or low level distortions and lack the functional…

计算机视觉与模式识别 · 计算机科学 2026-02-26 Meiqi Sun , Mingyu Li , Junxiong Zhu

Computational Design approaches facilitate the generation of typographic design, but evaluating these designs remains a challenging task. In this paper, we propose a set of heuristic metrics for typographic design evaluation, focusing on…

多媒体 · 计算机科学 2024-02-13 Sérgio M. Rebelo , J. J. Merelo , João Bicker , Penousal Machado

Recent advancements in the text-rendering capabilities of image generation models have made the end-to-end creation of graphic design content, such as posters, increasingly feasible. However, existing reward models fall short of accurately…

Content-aware visual-textual presentation layout aims at arranging spatial space on the given canvas for pre-defined elements, including text, logo, and underlay, which is a key to automatic template-free creative graphic design. In…

计算机视觉与模式识别 · 计算机科学 2023-03-29 HsiaoYuan Hsu , Xiangteng He , Yuxin Peng , Hao Kong , Qing Zhang

Visual layout plays a critical role in graphic design fields such as advertising, posters, and web UI design. The recent trend towards content-aware layout generation through generative models has shown promise, yet it often overlooks the…

计算机视觉与模式识别 · 计算机科学 2024-07-30 Jaejung Seol , Seojun Kim , Jaejun Yoo

Academic poster generation is a crucial yet challenging task in scientific communication, requiring the compression of long-context interleaved documents into a single, visually coherent page. To address this challenge, we introduce the…

计算机视觉与模式识别 · 计算机科学 2025-10-31 Wei Pang , Kevin Qinghong Lin , Xiangru Jian , Xi He , Philip Torr

Generating high-quality images without prompt engineering expertise remains a challenge for text-to-image (T2I) models, which often misinterpret poorly structured prompts, leading to distortions and misalignments. While humans easily…

计算机视觉与模式识别 · 计算机科学 2025-05-13 Nisan Chhetri , Arpan Sainju

Posters play a crucial role in marketing and advertising by enhancing visual communication and brand visibility, making significant contributions to industrial design. With the latest advancements in controllable T2I diffusion models,…

计算机视觉与模式识别 · 计算机科学 2025-02-13 Jian Ma , Yonglin Deng , Chen Chen , Nanyang Du , Haonan Lu , Zhenyu Yang

Text design is one of the most critical procedures in poster design, as it relies heavily on the creativity and expertise of humans to design text images considering the visual harmony and text-semantic. This study introduces TextPainter, a…

计算机视觉与模式识别 · 计算机科学 2023-08-15 Yifan Gao , Jinpeng Lin , Min Zhou , Chuanbin Liu , Hongtao Xie , Tiezheng Ge , Yuning Jiang

Product posters, which integrate subject, scene, and text, are crucial promotional tools for attracting customers. Creating such posters using modern image generation methods is valuable, while the main challenge lies in accurately…

计算机视觉与模式识别 · 计算机科学 2025-04-10 Yifan Gao , Zihang Lin , Chuanbin Liu , Min Zhou , Tiezheng Ge , Bo Zheng , Hongtao Xie

Generating aesthetic posters is more challenging than simple design images: it requires not only precise text rendering but also the seamless integration of abstract artistic content, striking layouts, and overall stylistic harmony. To…

计算机视觉与模式识别 · 计算机科学 2025-06-13 SiXiang Chen , Jianyu Lai , Jialin Gao , Tian Ye , Haoyu Chen , Hengyu Shi , Shitong Shao , Yunlong Lin , Song Fei , Zhaohu Xing , Yeying Jin , Junfeng Luo , Xiaoming Wei , Lei Zhu

Multi-agent systems built upon large language models (LLMs) have demonstrated remarkable capabilities in tackling complex compositional tasks. In this work, we apply this paradigm to the paper-to-poster generation problem, a practical yet…

人工智能 · 计算机科学 2026-04-14 Zhilin Zhang , Xiang Zhang , Jiaqi Wei , Yiwei Xu , Chenyu You

Poster design is a critical medium for visual communication. Prior work has explored automatic poster design using deep learning techniques, but these approaches lack text accuracy, user customization, and aesthetic appeal, limiting their…

图形学 · 计算机科学 2025-03-20 Haoyu Chen , Xiaojie Xu , Wenbo Li , Jingjing Ren , Tian Ye , Songhua Liu , Ying-Cong Chen , Lei Zhu , Xinchao Wang

Unified multimodal models integrate the reasoning capacity of large language models with both image understanding and generation, showing great promise for advanced multimodal intelligence. However, the community still lacks a rigorous…

计算机视觉与模式识别 · 计算机科学 2026-03-24 Hongxiang Li , Yaowei Li , Bin Lin , Yuwei Niu , Yuhang Yang , Xiaoshuang Huang , Jiayin Cai , Xiaolong Jiang , Yao Hu , Long Chen

Image-to-poster generation is a high-demand task requiring not only local adjustments but also high-level design understanding. Models must generate text, layout, style, and visual elements while preserving semantic fidelity and aesthetic…

计算机视觉与模式识别 · 计算机科学 2026-02-13 Sixiang Chen , Jianyu Lai , Jialin Gao , Hengyu Shi , Zhongying Liu , Tian Ye , Junfeng Luo , Xiaoming Wei , Lei Zhu

Recent progress in Vision Language Models (VLMs) has raised the question of whether they can reliably perform nonverbal reasoning. To this end, we introduce VRIQ (Visual Reasoning IQ), a novel benchmark designed to assess and analyze the…

计算机视觉与模式识别 · 计算机科学 2026-02-06 Tina Khezresmaeilzadeh , Jike Zhong , Konstantinos Psounis

Automated academic poster generation aims to distill lengthy research papers into concise, visually coherent presentations. Existing Multimodal Large Language Models (MLLMs) based approaches, however, suffer from three critical limitations:…

计算机视觉与模式识别 · 计算机科学 2026-03-03 Wenxin Tang , Jingyu Xiao , Yanpei Gong , Fengyuan Ran , Tongchuan Xia , Junliang Liu , Man Ho Lam , Wenxuan Wang , Michael R. Lyu

Text-to-image (T2I) generation aims to synthesize images from textual prompts, which jointly specify what must be shown and imply what can be inferred, which thus correspond to two core capabilities: \textbf{\textit{composition}} and…

计算机视觉与模式识别 · 计算机科学 2026-03-03 Ouxiang Li , Yuan Wang , Xinting Hu , Huijuan Huang , Rui Chen , Jiarong Ou , Xin Tao , Pengfei Wan , Xiaojuan Qi , Fuli Feng

Compositionality is a critical capability in Text-to-Image (T2I) models, as it reflects their ability to understand and combine multiple concepts from text descriptions. Existing evaluations of compositional capability rely heavily on…

计算机视觉与模式识别 · 计算机科学 2024-08-27 Xindi Wu , Dingli Yu , Yangsibo Huang , Olga Russakovsky , Sanjeev Arora

A molecule's properties are fundamentally determined by its composition and structure encoded in its molecular graph. Thus, reasoning about molecular properties requires the ability to parse and understand the molecular graph. Large…

‹ 上一页 1 2 3 10 下一页 ›