Abstract:Text classification is one of the fundamental tasks in natural language processing, which requires an agent to determine the most appropriate category for input sentences. Recently, deep neural networks have achieved impressive performance in this area, especially pretrained language models (PLMs). Usually, these methods concentrate on input sentences and corresponding semantic embedding generation. However, for another essential component: labels, most existing works either treat them as meaningless one-hot vectors or use vanilla embedding methods to learn label representations along with model training, underestimating the semantic information and guidance that these labels reveal. To alleviate this problem and better exploit label information, in this article, we employ self-supervised learning (SSL) in model learning process and design a novel self-supervised relation of relation (R <sup>2</sup> ) classification task for label utilization from a one-hot manner perspective. Then, we propose a novel () for text classification, in which text classification and R <sup>2</sup> classification are treated as optimization targets. Meanwhile, triplet loss is employed to enhance the analysis of differences and connections among labels. Moreover, considering that one-hot usage is still short of exploiting label information, we incorporate external knowledge from WordNet to obtain multiaspect descriptions for label semantic learning and extend to a novel () from a label embedding perspective. One step further, since these fine-grained descriptions may introduce unexpected noise, we develop a mutual interaction module to select appropriate parts from input sentences and labels simultaneously based on contrastive learning (CL) for noise mitigation. Extensive experiments on different text classification tasks reveal that can effectively improve the classification performance and can make better use of label information and further improve the performance. As a byproduct, we have released the codes to facilitate other research.

Text Clustering as Classification with LLMs

CLDA: Feature Selection for Text Categorization Based on Constrained LDA

ClusterLLM: Large Language Models as a Guide for Text Clustering

Text Clustering with Large Language Model Embeddings

TeC: A Novel Method for Text Clustering with Large Language Models Guidance and Weakly-Supervised Contrastive Learning

LLMEmbed: Rethinking Lightweight LLM's Genuine Function in Text Classification

Joint unsupervised contrastive learning and robust GMM for text clustering

Adaptable and Reliable Text Classification using Large Language Models

CEIL: A General Classification-Enhanced Iterative Learning Framework for Text Clustering

Clustering Algorithms and RAG Enhancing Semi-Supervised Text Classification with Large LLMs

Large Language Models Enable Few-Shot Clustering

Context-Aware Clustering using Large Language Models

A Lda-Based Algorithm For Length-Aware Text Clustering

Mitigating Boundary Ambiguity and Inherent Bias for Text Classification in the Era of Large Language Models

Dial-In LLM: Human-Aligned Dialogue Intent Clustering with LLM-in-the-loop

ZeroDL: Zero-shot Distribution Learning for Text Clustering via Large Language Models

Description-Enhanced Label Embedding Contrastive Learning for Text Classification

An end-to-end Neural Network Framework for Text Clustering

An Unsupervised Learning Short Text Clustering Method

Research on Multilingual News Clustering Based on Cross-Language Word Embeddings