Discovering Useful Compact Sets of Sequential Rules in a Long Sequence

Erwan Bourrand,Luis Galárraga,Esther Galbrun,Elisa Fromont,Alexandre Termier
DOI: https://doi.org/10.48550/arXiv.2109.07519
2022-12-30
Abstract:We are interested in understanding the underlying generation process for long sequences of symbolic events. To do so, we propose COSSU, an algorithm to mine small and meaningful sets of sequential rules. The rules are selected using an MDL-inspired criterion that favors compactness and relies on a novel rule-based encoding scheme for sequences. Our evaluation shows that COSSU can successfully retrieve relevant sets of closed sequential rules from a long sequence. Such rules constitute an interpretable model that exhibits competitive accuracy for the tasks of next-element prediction and classification.
Machine Learning,Artificial Intelligence
What problem does this paper attempt to address?