Abstract:The emergence and rapid development of neural networks have been pivotal in advancing text-to-image generative models, with particular emphasis on generative adversarial networks (GANs), variational autoencoders (VAEs), and augmented reality (AR). These models have greatly enriched the field, offering diverse avenues for image generation. Critical support has been provided by databases such as MS COCO, Flickr30K, Visual Genome, and Conceptual Captions, along with essential evaluation metrics, including Inception Score (IS), Frchet Inception Distance (FID), precision, and recall. In this comprehensive review, we delve into the mechanisms and significance of each model and technique, ensuring a holistic examination of their contributions. Both GANs and VAEs stand out as significant models within image generative frameworks, each excelling in distinct aspects. Therefore, it is imperative to discuss both models in this review, as they offer complementary strengths. Additionally, we include noteworthy models such as augmented reality to provide a well-rounded assessment of the current advancements in the field. In terms of datasets, MS COCO offers a diverse and extensive collection of images, serving as a cornerstone for model training. Other datasets like Flickr 30k, Visual Genome, and Conceptual Captions contribute valuable labeled examples, further enriching the learning process for these models. The incorporation of widely recognized metrics and methodologies in the field allows for effective evaluation and comparison of their relative significance. In conclusion, the field's recent achievements owe much to the integration of its various components. VAEs and GANs, with their unique strengths, complement each other, while metrics and datasets play complementary roles in advancing the capabilities of generative models in the context of text-to-image synthesis. This survey underscores the collaborative synergy between models, metrics, and datasets, propelling the field toward new horizons.

Investigation related to application of Generative Adversarial Networks in text-to-image synthesis

A survey of generative adversarial networks and their application in text-to-image synthesis

A Survey and Taxonomy of Adversarial Neural Networks for Text-to-Image Synthesis

A Survey on Adversarial Image Synthesis

Image Synthesis with Adversarial Networks: a Comprehensive Survey and Case Studies

Text-to-Image Synthesis With Generative Models: Methods, Datasets, Performance Metrics, Challenges, and Future Direction

A Survey of Image Synthesis and Editing with Generative Adversarial Networks

Generative Adversarial Networks for Image and Video Synthesis: Algorithms and Applications

Text-To-Image with Generative Adversarial Networks

A Comparative Study of Generative Adversarial Networks for Text-to-Image Synthesis

Survey on Generative Adversarial Behavior in Artificial Neural Tasks

Recent Advances of Generative Adversarial Networks in Computer Vision

Review of Generative Adversarial Networks in Image Generation

Generative adversarial networks (GANs): Introduction, Taxonomy, Variants, Limitations, and Applications

A survey of generative models used in text-to-image

Handwritten Digits Image Generation with help of Generative Adversarial Network: Machine Learning Approach

Revisiting the Evaluation of Image Synthesis with GANs

A brief study of generative adversarial networks and their applications in image synthesis

Narrative-guided synthesis: Revolutionizing text-to-image translation based on Generative Adversarial Networks

A Review: Generative Adversarial Networks

Generative Adversarial Networks and Other Generative Models