Accès ouvert

Arabic Text to Image Generation Tasks

Article scientifique 2022 Anglais

Résumé

Abstract Current AI systems have shown impressive results in the Automatic synthesis realistic images from texts descriptions tasks. In fact, Generative Adversarial Networks (GANs) are used mostly in this tasks. The Generator generates realistic images given noise and sentence vectors, and the discriminator produce a probability of how the synthetic images are reals. In this paper, in order to generate images from Arabic text, we fuse DF-GAN as a sample and efficient text-to-image generation framework and AraBERT architecture. To achieve this purpose, firstly, we re-create new datasets matching the Arabic text-to-image generation task by applying DeepL-Translator from English to Arabic on texts descriptions of original datasets. Secondly, we leverage the power of AraBERT which is trained on billions of Arabic words to produce a strong sentence embedding, and we reduce that vector's dimension to match with DF-GAN shape. Thirdly, we inject the reduced sentence embedding into the UPBlocks sections of DF-GAN and we train the proposed architecture on two challenging datasets. Following the previous works, we use CUB and Oxford-102 flowers as original datasets. Further, we measure our framework with FID and IS. Our framework is the first that achieve much success in generating high-resolution realistic and text matching images conditioned with Arabic text.

Citer ce document

Bahani, M. (2022). Arabic Text to Image Generation Tasks. https://doi.org/10.21203/rs.3.rs-2163664/v1

Accès au document

Voir sur le dépôt source

Ce document est hébergé sur son dépôt institutionnel d'origine.

Auteur(s)

Statistiques

Consultations : 1

Téléchargements : 0