错误:搜索内容不能为空,请输入英文关键词
错误:关键词超出字数限制,请精简
高级检索

An Improved AttnGAN Model for Text-to-Image Synthesis

  • Remya Gopalakrishnan,
  • Naveen Sambagni,
  • P. V. Sudeep

摘要

Text-to-image generation models generate photo-realistic images from textual descriptions, typically using GANs and BiLSTM networks. However, as input text sequence length increases, these models suffer from a loss of information, leading to missed keywords and unsatisfactory results. To address this, we propose an attentional GAN (AttnGAN) model with a text attention mechanism. We evaluate AttnGAN variants on the MS-COCO dataset qualitatively and quantitatively. For the image quality analysis, we utilize performance measures such as FID score, R-precision, and IS score. Our results show that the proposed model outperforms existing approaches, producing more realistic images by preserving vital information in the input sequence.