Abstract
Extreme Multi-label text Classification (XMC) is a fundamental text mining task, which aims to assign multiple labels related to the given text from a large-scale label set. Various models and many data augmentation methods are proposed to improve classification performance. However, the classification performance is limited due to the long tail distribution of labels, which is an essential characteristic of XMC. To address this problem, we propose a novel data augmentation method named Attentional Data Augmentation Method (ADAM) for long tail labels. Specifically, we split each sentence into several segments of equal length and use an attention-based neural network to explore the core segments of long tail labels. The unimportant segments of each instance from the dataset are considered to be replaced by those segments related to the long tail labels. Extensive experiments show that ADAM has an improvement based on the XMC method, especially on the prediction of long tail labels.
Access this chapter
Tax calculation will be finalised at checkout
Purchases are for personal use only
Similar content being viewed by others
Notes
- 1.
If there is more than one long tail label in an instance, segments that have a low attention score with head labels belong to the long tail label that has the highest attention score with them.
- 2.
- 3.
References
Babbar, R., Schölkopf, B.: Data scarcity, robustness and extreme multi-label classification. Mach. Learn. 108(8–9), 1329–1351 (2019)
Chang, W., Yu, H., Zhong, K., Yang, Y., Dhillon, I.S.: Taming pretrained transformers for extreme multi-label text classification. In: KDD 2020. ACM (2020)
Devlin, J., Chang, M., Lee, K., Toutanova, K.: BERT: pre-training of deep bidirectional transformers for language understanding. In: NAACL-HLT 2019 (2019)
Edunov, S., Ott, M., Auli, M., Grangier, D.: Understanding back-translation at scale. In: EMNLP 2018. Association for Computational Linguistics (2018)
Jasinska, K., Dembczynski, K., Busa-Fekete, R., et al.: Extreme f-measure maximization using sparse probability estimates. In: ICML 2016 (2016)
Jiang, T., Wang, D., Sun, L., et al.: LightXML: transformer with dynamic negative sampling for high-performance extreme multi-label text classification. CoRR (2021)
Liu, J., Chang, W., Wu, Y., Yang, Y.: Deep learning for extreme multi-label text classification. In: Proceedings of the 40th International ACM SIGIR (2017)
Liu, Y., et al.: Roberta: a robustly optimized BERT pretraining approach
Mencía, E.L., Fürnkranz, J.: Efficient pairwise multilabel classification for large-scale problems in the legal domain. In: ECML/PKDD 2008 (2008)
Prabhu, Y., Kag, A., Harsola, S., et al.: Parabel: partitioned label trees for extreme classification with application to dynamic search advertising. In: WWW (2018)
Sennrich, R., et al.: Improving neural machine translation models with monolingual data. In: ACL 2016 (2016)
Vaswani, A., et al.: Attention is all you need. In: NeurIPS 2017 (2017)
Wei, J.W., Zou, K.: EDA: easy data augmentation techniques for boosting performance on text classification tasks. In: EMNLP-IJCNLP 2019 (2019)
Xie, Q., et al.: Unsupervised data augmentation for consistency training
Yang, P., Sun, X., Li, W., Ma, S., Wu, W., Wang, H.: SGM: sequence generation model for multi-label classification. In: COLING 2018 (2018)
Yang, Z., Dai, Z., Yang, Y., et al.: XLNet: generalized autoregressive pretraining for language understanding. In: NeurIPS 2019 (2019)
You, R., Zhang, Z., Wang, Z., et al.: AttentionXML: label tree-based attention-aware deep model for high-performance extreme multi-label text classification
Zubiaga, A.: Enhancing navigation on Wikipedia with social tags. CoRR (2012)
Acknowledgements
This research is supported by the National Natural Science Foundation of China under the grant No. 61976119 and the Natural Science Foundation of Tianjin under the grant No. 18ZXZNGX00310.
Author information
Authors and Affiliations
Corresponding author
Editor information
Editors and Affiliations
Rights and permissions
Copyright information
© 2022 The Author(s), under exclusive license to Springer Nature Switzerland AG
About this paper
Cite this paper
Zhang, J., Liu, J., Chen, S., Lin, S., Wang, B., Wang, S. (2022). ADAM: An Attentional Data Augmentation Method for Extreme Multi-label Text Classification. In: Gama, J., Li, T., Yu, Y., Chen, E., Zheng, Y., Teng, F. (eds) Advances in Knowledge Discovery and Data Mining. PAKDD 2022. Lecture Notes in Computer Science(), vol 13280. Springer, Cham. https://doi.org/10.1007/978-3-031-05933-9_11
Download citation
DOI: https://doi.org/10.1007/978-3-031-05933-9_11
Published:
Publisher Name: Springer, Cham
Print ISBN: 978-3-031-05932-2
Online ISBN: 978-3-031-05933-9
eBook Packages: Computer ScienceComputer Science (R0)Springer Nature Proceedings Computer Science

