Abstract:Recent advances in neural-based generative modeling have reignited the hopes of having computer systems capable of conversing with humans and able to understand natural language. The employment of deep neural architectures has been largely explored in a multitude of context and tasks to fulfill various user needs. On one hand, producing textual content that meets specific requirements is of priority for a model to seamlessly conduct conversations with different groups of people. On the other hand, latent variable models (LVM) such as variational auto-encoders (VAEs) as one of the most popular genres of generative models are designed to characterize the distributional pattern of textual data. Thus they are inherently capable of learning the integral textual features that are worth exploring for controllable pursuits. \noindent This overview gives an introduction to existing generation schemes, problems associated with text variational auto-encoders, and a review of several applications about the controllable generation that are instantiations of these general formulations,\footnote{A detailed paper list is available at \url{<a class="link-external link-https" href="https://github.com/ImKeTT/CTG-latentAEs" rel="external noopener nofollow">this https URL</a>}} as well as related datasets, metrics and discussions for future researches. Hopefully, this overview will provide an overview of living questions, popular methodologies and raw thoughts for controllable language generation under the scope of variational auto-encoder.

Topic-VQ-VAE: Leveraging Latent Codebooks for Flexible Topic-Guided Document Generation

Vector-Quantization-Based Topic Modeling

Topic-Guided Variational Auto-Encoder for Text Generation.

Nested Variational Autoencoder for Topic Modeling on Microtexts with Word Vectors

Gaussian Mixture Variational Autoencoder For Semi-Supervised Topic Modeling

Topic Modeling Using Distributed Word Embeddings

Topic-word-constrained sentence generation with variational autoencoder

SenGen: Sentence Generating Neural Variational Topic Model.

Variational Gaussian Topic Model with Invertible Neural Projections

A Bayesian Nonparametric Topic Model with Variational Auto-Encoders

Vector Quantized Time Series Generation with a Bidirectional Prior Model

FET-LM: Flow-Enhanced Variational Autoencoder for Topic-Guided Language Modeling

Gaussian Mixture Vector Quantization with Aggregated Categorical Posterior

Vector Quantized Wasserstein Auto-Encoder

Long Text Generation with Topic-aware Discrete Latent Variable Model.

Topic2Vec: Learning distributed representations of topics

LG-VQ: Language-Guided Codebook Learning

Extracting nonlinear neural topics with neural variational bayes

An Overview on Controllable Text Generation via Variational Auto-Encoders

Improve Variational Autoencoder for Text Generationwith Discrete Latent Bottleneck