Abstract:Software Log anomaly event detection with masked event prediction has various technical approaches with countless configurations and parameters. Our objective is to provide a baseline of settings for similar studies in the future. The models we use are the N-Gram model, which is a classic approach in the field of natural language processing (NLP), and two deep learning (DL) models long short-term memory (LSTM) and convolutional neural network (CNN). For datasets we used four datasets Profilence, BlueGene/L (BGL), Hadoop Distributed File System (HDFS) and Hadoop. Other settings are the size of the sliding window which determines how many surrounding events we are using to predict a given event, mask position (the position within the window we are predicting), the usage of only unique sequences, and the portion of data that is used for training. The results show clear indications of settings that can be generalized across datasets. The performance of the DL models does not deteriorate as the window size increases while the N-Gram model shows worse performance with large window sizes on the BGL and Profilence datasets. Despite the popularity of Next Event Prediction, the results show that in this context it is better not to predict events at the edges of the subsequence, i.e., first or last event, with the best result coming from predicting the fourth event when the window size is five. Regarding the amount of data used for training, the results show differences across datasets and models. For example, the N-Gram model appears to be more sensitive toward the lack of data than the DL models. Overall, for similar experimental setups we suggest the following general baseline: Window size 10, mask position second to last, do not filter out non-unique sequences, and use a half of the total data for training.

Multi-Parameter Log Anomaly Detection with an Unsupervised Learning Approach

Learning Discrimination from Contaminated Data: Multi-Instance Learning for Unsupervised Anomaly Detection

Natural Language Processing-based Model for Log Anomaly Detection

LAnoBERT: System Log Anomaly Detection based on BERT Masked Language Model

An Anomaly Detection Approach of Part-of-Speech Log Sequence Via Population Based Training

TPLogAD: Unsupervised Log Anomaly Detection Based on Event Templates and Key Parameters

LogBERT: Log Anomaly Detection via BERT

LLMeLog: an Approach for Anomaly Detection Based on LLM-enriched Log Events

Log Anomaly Detection method based on BERT model optimization

SemLog: A Semantics-based Approach for Anomaly Detection in Big Data System Logs

Angel or Devil: Discriminating Hard Samples and Anomaly Contaminations for Unsupervised Time Series Anomaly Detection

LogAnomaly: Unsupervised Detection of Sequential and Quantitative Anomalies in Unstructured Logs

How to Configure Masked Event Anomaly Detection on Software Logs?

HitAnomaly: Hierarchical Transformers for Anomaly Detection in System Log

Load Balancing Based on Process Migration for MPI

End-to-End AutoML for Unsupervised Log Anomaly Detection

Research on Log Anomaly Detection Based on Sentence-BERT

OneLog: Towards End-to-End Training in Software Log Anomaly Detection

FastLogAD: Log Anomaly Detection with Mask-Guided Pseudo Anomaly Generation and Discrimination

LogPS: A Robust Log Sequential Anomaly Detection Approach Based on Natural Language Processing

Detecting Log Anomalies with Multi-Head Attention (LAMA)