Abstract:Text summarization (TS) is considered one of the most difficult tasks in natural language processing (NLP). It is one of the most important challenges that stand against the modern computer system's capabilities with all its new improvement. Many papers and research studies address this task in literature but are being carried out in extractive summarization, and few of them are being carried out in abstractive summarization, especially in the Arabic language due to its complexity. In this paper, an abstractive Arabic text summarization system is proposed, based on a sequence-to-sequence model. This model works through two components, encoder and decoder. Our aim is to develop the sequence-to-sequence model using several deep artificial neural networks to investigate which of them achieves the best performance. Different layers of Gated Recurrent Units (GRU), Long Short-Term Memory (LSTM), and Bidirectional Long Short-Term Memory (BiLSTM) have been used to develop the encoder and the decoder. In addition, the global attention mechanism has been used because it provides better results than the local attention mechanism. Furthermore, AraBERT preprocess has been applied in the data preprocessing stage that helps the model to understand the Arabic words and achieves state-of-the-art results. Moreover, a comparison between the skip-gram and the continuous bag of words (CBOW) word2Vec word embedding models has been made. We have built these models using the Keras library and run-on Google Colab Jupiter notebook to run seamlessly. Finally, the proposed system is evaluated through ROUGE-1, ROUGE-2, ROUGE-L, and BLEU evaluation metrics. The experimental results show that three layers of BiLSTM hidden states at the encoder achieve the best performance. In addition, our proposed system outperforms the other latest research studies. Also, the results show that abstractive summarization models that use the skip-gram word2Vec model outperform the models that use the CBOW word2Vec model.

End to End Urdu Abstractive Text Summarization With Dataset and Improvement in Evaluation Metric

Abstractive Text Summarization for the Urdu Language: Data and Methods

Abstractive Summary Generation for the Urdu Language

Low Resource Summarization using Pre-trained Language Models

Abstractive Arabic Text Summarization Based on Deep Learning

Abstractive text summarization: State of the art, challenges, and improvements

Improving the readability and saliency of abstractive text summarization using combination of deep neural networks equipped with auxiliary attention mechanism

Assessment of Transformer-Based Encoder-Decoder Model for Human-Like Summarization

Document vector embedding based extractive text summarization system for Hindi and English text

Abstractive Summarization Using Attentive Neural Techniques

Improving ROUGE‐1 by 6%: A novel multilingual transformer for abstractive news summarization

Automatic Extractive Text Summarization using Multiple Linguistic Features

Abstractive Text Summarization using Pre-Trained Language Model "Text-to-Text Transfer Transformer (T5)"

UTSA: Urdu Text Sentiment Analysis Using Deep Learning Methods

On the State of German (Abstractive) Text Summarization

Topic-Guided Abstractive Text Summarization: a Joint Learning Approach

Enhancing extractive text summarization using natural language processing with an optimal deep learning model

Efficient Urdu Caption Generation using Attention based LSTM

Contextual embedded text summarizer system: A hybrid approach

Abstractive Summarization Improved by WordNet-based Extractive Sentences