Enhancing Label Correlation Feedback in Multi-Label Text Classification via Multi-Task Learning
In multi-label text classification (MLTC), each given document is associated with a set of correlated labels. To capture label correlations, previous classifier-chain and sequence-to-sequence models transform MLTC to a sequence prediction task. However, they tend to suffer from label order dependency, label combination over-fitting and error propagation problems. To address these problems, we introduce a novel approach with multi-task learning to enhance label correlation feedback. We first utilize a joint embedding (JE) mechanism to obtain the text and label representation simultaneously. In MLTC task, a document-label cross attention (CA) mechanism is adopted to generate a more discriminative document representation. Furthermore, we propose two auxiliary label co-occurrence prediction tasks to enhance label correlation learning: 1) Pairwise Label Co-occurrence Prediction (PLCP), and 2) Conditional Label Co-occurrence Prediction (CLCP). Experimental results on AAPD and RCV1-V2 datasets show that our method outperforms competitive baselines by a large margin. We analyze low-frequency label performance, label dependency, label combination diversity and coverage speed to show the effectiveness of our proposed method on label correlation learning.
Code (1)
Tasks
DiversityMulti Label Text ClassificationMulti-Label Text ClassificationMulti-Task LearningPredictiontext-classificationText ClassificationSimilar Papers 제목 키워드 기반
A Deep Reinforced Sequence-to-Set Model for Multi-Label Classification
Multi-label classification (MLC) aims to predict a set of labels for a given instance. Based on a pre-defined label order, the sequence-to-sequence (Seq2Seq) model trained via maximum likelihood estimation method has bee…
General ClassificationMulti-Label ClassificationMUlTI-LABEL-ClASSIFICATIONReinforcement Learning+1Deconfounded and Explainable Interactive Vision-Language Retrieval of Complex Scenes
In vision-language retrieval systems, users provide natural language feedback to find target images. Vision-language explanations in the systems can better guide users to provide feedback and thus improve the retrieval. …
Explainable ModelsLanguage ModellingRetrievalImproved Knowledge Transfer for Semi-Supervised Domain Adaptation via Trico Training Strategy
The motivation of the semi-supervised domain adaptation (SSDA) is to train a model by leveraging knowledge acquired from the plentiful labeled source combined with extremely scarce labeled target data to achieve the …
Domain AdaptationSemi-supervised Domain AdaptationTransfer LearningThe Semantic Architect: How FEAML Bridges Structured Data and LLMs for Multi-Label Tasks
Existing feature engineering methods based on large language models (LLMs) have not yet been applied to multi-label learning tasks. They lack the ability to model complex label dependencies and are not specifically adapt…
Multi-Label ClassificationMulti-Label LearningFeature EngineeringCode GenerationImproving Classification Performance With Human Feedback: Label a few, we label the rest
In the realm of artificial intelligence, where a vast majority of data is unstructured, obtaining substantial amounts of labeled data to train supervised machine learning models poses a significant challenge. To address …
Active Learningtext-classificationText Classification