Abstract:Recent advances in neural-based generative modeling have reignited the hopes of having computer systems capable of conversing with humans and able to understand natural language. The employment of deep neural architectures has been largely explored in a multitude of context and tasks to fulfill various user needs. On one hand, producing textual content that meets specific requirements is of priority for a model to seamlessly conduct conversations with different groups of people. On the other hand, latent variable models (LVM) such as variational auto-encoders (VAEs) as one of the most popular genres of generative models are designed to characterize the distributional pattern of textual data. Thus they are inherently capable of learning the integral textual features that are worth exploring for controllable pursuits. \noindent This overview gives an introduction to existing generation schemes, problems associated with text variational auto-encoders, and a review of several applications about the controllable generation that are instantiations of these general formulations,\footnote{A detailed paper list is available at \url{<a class="link-external link-https" href="https://github.com/ImKeTT/CTG-latentAEs" rel="external noopener nofollow">this https URL</a>}} as well as related datasets, metrics and discussions for future researches. Hopefully, this overview will provide an overview of living questions, popular methodologies and raw thoughts for controllable language generation under the scope of variational auto-encoder.

FET-LM: Flow-Enhanced Variational Autoencoder for Topic-Guided Language Modeling

Multimodal Latent Language Modeling with Next-Token Diffusion

LlaMaVAE: Guiding Large Language Model Generation via Continuous Latent Sentence Spaces

AdaVAE: Exploring Adaptive GPT-2s in Variational Auto-Encoders for Language Modeling

Flow-Based Variational Sequence Autoencoder

A neural topic model with word vectors and entity vectors for short texts

Topic-VQ-VAE: Leveraging Latent Codebooks for Flexible Topic-Guided Document Generation

f-VAEs: Improve VAEs with Conditional Flows

RegaVAE: A Retrieval-Augmented Gaussian Mixture Variational Auto-Encoder for Language Modeling

A Bayesian Nonparametric Topic Model with Variational Auto-Encoders

Topic-word-constrained sentence generation with variational autoencoder

Dispersed Exponential Family Mixture VAEs for Interpretable Text Generation

Improve Variational Autoencoder for Text Generationwith Discrete Latent Bottleneck

Dispersed EM-VAEs for Interpretable Text Generation

Optimus: Organizing Sentences via Pre-trained Modeling of a Latent Space

Conditional Flow Variational Autoencoders for Structured Sequence Prediction

TG-LLaVA: Text Guided LLaVA via Learnable Latent Embeddings

An Overview on Controllable Text Generation via Variational Auto-Encoders

Topic Modeling with Wasserstein Autoencoders

Text Adversarial Generation Based on Latent Variable Models

LLM2FEA: Discover Novel Designs with Generative Evolutionary Multitasking