English

Design a Delicious Lunchbox in Style

Computer Vision and Pattern Recognition 2023-05-25 v1

Abstract

We propose a cyclic generative adversarial network with spatial-wise and channel-wise attention modules for text-to-image synthesis. To accurately depict and design scenes with multiple occluded objects, we design a pre-trained ordering recovery model and a generative adversarial network to predict layout and composite novel box lunch presentations. In the experiments, we devise the Bento800 dataset to evaluate the performance of the text-to-image synthesis model and the layout generation & image composition model. This paper is the continuation of our previous paper works. We also present additional experiments and qualitative performance comparisons to verify the effectiveness of our proposed method. Bento800 dataset is available at https://github.com/Yutong-Zhou-cv/Bento800_Dataset

Keywords

Cite

@article{arxiv.2305.14522,
  title  = {Design a Delicious Lunchbox in Style},
  author = {Yutong Zhou},
  journal= {arXiv preprint arXiv:2305.14522},
  year   = {2023}
}

Comments

Accepted by WiCV @CVPR2023 (In Progress). Dataset: https://github.com/Yutong-Zhou-cv/Bento800_Dataset

R2 v1 2026-06-28T10:43:41.269Z