Journal:Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics (ACL), CCF-A
Abstract:Open-Set Semi-Supervised Text Classification (OSTC) aims to train a classification model on a limited set of labeled texts, alongside plenty of unlabeled texts that include both in-distribution and out-of-distribution examples. In this paper, we revisit the main challenge in OSTC, i.e., outlier detection, from a measurement disagreement perspective and innovatively propose to improve OSTC performance by directly maximizing the measurement disagreements. Based on the properties of in-measurement and crossmeasurements, we design an Adversarial Disagreement Maximization (ADM) model that synergeticly optimizes the measurement disagreements. In addition, we develop an abnormal example detection and measurement calibration approach to guarantee the effectiveness of ADM training. Experiment results and comprehensive analysis of three benchmarks demonstrate the effectiveness of our model.
Co-author:Junfan Chen,Richong Zhang, Junchi Chen,Chunming Hu
Indexed by:国际学术会议
Page Number:2170-2180
Translation or Not:no
Date of Publication:2024-01-01
