MELM: Data Augmentation with Masked Entity Language Modeling for Low-Resource NER

Ran Zhou; Xin Li; Ruidan He; Lidong Bing; Erik Cambria; Luo Si; Chunyan Miao

Conference ProceedingsOPEN ACCESS

MELM: Data Augmentation with Masked Entity Language Modeling for Low-Resource NER

Proceedings of the Annual Meeting of the Association for Computational Linguistics (2022) 1 2251-2262

DOI: 10.18653/v1/2022.acl-long.160

70Citations

87Readers

Get full text

Abstract

Data augmentation is an effective solution to data scarcity in low-resource scenarios. However, when applied to token-level tasks such as NER, data augmentation methods often suffer from token-label misalignment, which leads to unsatsifactory performance. In this work, we propose Masked Entity Language Modeling (MELM) as a novel data augmentation framework for low-resource NER. To alleviate the token-label misalignment issue, we explicitly inject NER labels into sentence context, and thus the fine-tuned MELM is able to predict masked entity tokens by explicitly conditioning on their labels. Thereby, MELM generates high-quality augmented data with novel entities, which provides rich entity regularity knowledge and boosts NER performance. When training data from multiple languages are available, we also integrate MELM with code-mixing for further improvement. We demonstrate the effectiveness of MELM on monolingual, cross-lingual and multilingual NER across various low-resource levels. Experimental results show that our MELM presents substantial improvement over the baseline methods.

References Powered by Scopus

View more at Scopus

Cited by Powered by Scopus

View more at Scopus

Cite

CITATION STYLE

APA

Zhou, R., Li, X., He, R., Bing, L., Cambria, E., Si, L., & Miao, C. (2022). MELM: Data Augmentation with Masked Entity Language Modeling for Low-Resource NER. In Proceedings of the Annual Meeting of the Association for Computational Linguistics (Vol. 1, pp. 2251–2262). Association for Computational Linguistics (ACL). https://doi.org/10.18653/v1/2022.acl-long.160

Readers' Seniority

PhD / Post grad / Masters / Doc 21

75%

Researcher 5

18%

Professor / Associate Prof. 1

Lecturer / Post doc 1

Readers' Discipline

Computer Science 30

86%

Linguistics 3

Neuroscience 1

Engineering 1

MELM: Data Augmentation with Masked Entity Language Modeling for Low-Resource NER

Abstract

References Powered by Scopus

Neural architectures for named entity recognition

Improving neural machine translation models with monolingual data

Multilingual denoising pre-training for neural machine translation

Cited by Powered by Scopus

Revisiting DocRED - Addressing the False Negative Problem in Relation Extraction

PromptNER: Prompt Locating and Typing for Named Entity Recognition

Few-shot biomedical named entity recognition via knowledge-guided instance generation and prompt contrastive learning

Register to see more suggestions

Cite

Readers' Seniority

Readers' Discipline