English

Brickify: Enabling Expressive Design Intent Specification through Direct Manipulation on Design Tokens

Human-Computer Interaction 2025-03-03 v1

Abstract

Expressing design intent using natural language prompts requires designers to verbalize the ambiguous visual details concisely, which can be challenging or even impossible. To address this, we introduce Brickify, a visual-centric interaction paradigm -- expressing design intent through direct manipulation on design tokens. Brickify extracts visual elements (e.g., subject, style, and color) from reference images and converts them into interactive and reusable design tokens that can be directly manipulated (e.g., resize, group, link, etc.) to form the visual lexicon. The lexicon reflects users' intent for both what visual elements are desired and how to construct them into a whole. We developed Brickify to demonstrate how AI models can interpret and execute the visual lexicon through an end-to-end pipeline. In a user study, experienced designers found Brickify more efficient and intuitive than text-based prompts, allowing them to describe visual details, explore alternatives, and refine complex designs with greater ease and control.

Keywords

Cite

@article{arxiv.2502.21219,
  title  = {Brickify: Enabling Expressive Design Intent Specification through Direct Manipulation on Design Tokens},
  author = {Xinyu Shi and Yinghou Wang and Ryan Rossi and Jian Zhao},
  journal= {arXiv preprint arXiv:2502.21219},
  year   = {2025}
}

Comments

CHI 2025

R2 v1 2026-06-28T22:02:09.078Z