Abstract:Multi-label text classification tasks face challenges such as sample diversity, complexity, and the need for effective utilization of label correlations. In this paper, we propose a model that integrates multi-granularity fusion of text sequence features and label semantic correlation information. Our model leverages graph convolutional networks to extract label semantic correlation, which enhances classification performance for samples with similar labels and addresses label omission issues. Additionally, text convolutional neural networks are employed to extract multi-granularity sense group features from text sequences, calculate their similarity with semantic correlation label distributions, and dynamically adjust the similarity between text context and label information. This approach tackles the limitations of feature extraction in short texts and label confusion. We replace the original multi-hot label encoding in model training with a label distribution that fuses text multi-granularity sense group features and label correlation information, using a more precise encoding method for soft alignment based on label probability distributions. This enhances the model's resilience to noisy data, avoiding the issue of assigning high-confidence probabilities to incorrect categories due to hard-coded supervision. Our model's performance improvement on noisy datasets significantly surpasses that achieved by label smoothing. Extensive experiments on three legal text datasets and two generalized multi-label datasets demonstrate the model's excellent performance. Our approach is applicable in various real-world scenarios, such as legal judgment prediction, news categorization, and recommendation systems, where accurate multi-label classification is crucial. Ablation and experiments on noisy datasets validate the model's effectiveness and robustness.

MFMGC: A Multi-modal Data Fusion Model for Movie Genre Classification

A multimodal approach for multi-label movie genre classification

Gated Multimodal Units for Information Fusion

Multimodal Sentiment Analysis Based on Cross-Modal Attention and Gated Cyclic Hierarchical Fusion Networks

MMGA: Multimodal Learning with Graph Alignment

Label distribution for multimodal machine learning

Rethinking movie genre classification with fine-grained semantic clustering

Unleashing the Power of Imbalanced Modality Information for Multi-modal Knowledge Graph Completion

Multi-Channel Attentive Graph Convolutional Network with Sentiment Fusion for Multimodal Sentiment Analysis

Multimodal Remote Sensing Data Classification Based on Gaussian Mixture Variational Dynamic Fusion Network

Multi-graph Fusion Graph Convolutional Networks with pseudo-label supervision

Multilevel profiling of situation and dialogue-based deep networks for movie genre classification using movie trailers

MFLSCI: Multi-granularity fusion and label semantic correlation information for multi-label legal text classification

Multimodal Attentive Representation Learning for Micro-video Multi-label Classification

A Multimodal Sentiment Analysis Method Integrating Multi-Layer Attention Interaction and Multi-Feature Enhancement

MGMNet: Mutual-Guidance Mechanism for Joint Classification of Multisource Remote Sensing Data

Deep Multimodal Data Fusion

Towards Coarse and Fine-grained Multi-Graph Multi-Label Learning

MFSC: A Multimodal Aspect-Level Sentiment Classification Framework with Multi-Image Gate and Fusion Networks

MERGE: A Modal Equilibrium Relational Graph Framework for Multi-Modal Knowledge Graph Completion

Mutual information maximization and feature space separation and bi-bimodal mo-dality fusion for multimodal sentiment analysis