Abstract:Multi-label text classification tasks face challenges such as sample diversity, complexity, and the need for effective utilization of label correlations. In this paper, we propose a model that integrates multi-granularity fusion of text sequence features and label semantic correlation information. Our model leverages graph convolutional networks to extract label semantic correlation, which enhances classification performance for samples with similar labels and addresses label omission issues. Additionally, text convolutional neural networks are employed to extract multi-granularity sense group features from text sequences, calculate their similarity with semantic correlation label distributions, and dynamically adjust the similarity between text context and label information. This approach tackles the limitations of feature extraction in short texts and label confusion. We replace the original multi-hot label encoding in model training with a label distribution that fuses text multi-granularity sense group features and label correlation information, using a more precise encoding method for soft alignment based on label probability distributions. This enhances the model's resilience to noisy data, avoiding the issue of assigning high-confidence probabilities to incorrect categories due to hard-coded supervision. Our model's performance improvement on noisy datasets significantly surpasses that achieved by label smoothing. Extensive experiments on three legal text datasets and two generalized multi-label datasets demonstrate the model's excellent performance. Our approach is applicable in various real-world scenarios, such as legal judgment prediction, news categorization, and recommendation systems, where accurate multi-label classification is crucial. Ablation and experiments on noisy datasets validate the model's effectiveness and robustness.

Label-Specific Document Representation for Multi-Label Text Classification.

Label-aware Document Representation via Hybrid Attention for Extreme Multi-Label Text Classification

A Label-Specific Attention-Based Network with Regularized Loss for Multi-label Classification

Label-Specific Feature Augmentation for Long-Tailed Multi-Label Text Classification

Research of multi-label text classification based on label attention and correlation networks

A Label Information Aware Model for Multi-label Text Classification

Learning Label-Adaptive Representation for Large-Scale Multi-Label Text Classification

Label-Attentive Hierarchical Attention Network for Text Classification

Label Attention Network for sequential multi-label classification: you were looking at a wrong self-attention

MLAN: Multi-Level Attention Network

Deeply Supervised Layer Selective Attention Network: Towards Label-Efficient Learning for Medical Image Classification

Label Attention Network for Structured Prediction

KeNet:Knowledge-enhanced Doc-Label Attention Network for Multi-label text classification

LA-HCN: Label-based Attention for Hierarchical Multi-label TextClassification Neural Network

MLGN:A Multi-Label Guided Network for Improving Text Classification

Label-Guided Cross-Modal Attention Network for Multi-Label Aerial Image Classification

Semi-Supervised Dual Relation Learning for Multi-Label Classification

Jointly Trained Sequential Labeling and Classification by Sparse Attention Neural Networks.

MFLSCI: Multi-granularity fusion and label semantic correlation information for multi-label legal text classification

A Hybrid Model Based on Convolutional Neural Network and Long Short-Term Memory for Multi-label Text Classification

Multi-label text classification based on semantic-sensitive graph convolutional network