Exploration of noise strategies in semi-supervised named entity classification

Pooja Lakshmi Narayan; Ajay Nagesh; Mihai Surdeanu

Conference ProceedingsOPEN ACCESS

Exploration of noise strategies in semi-supervised named entity classification

*SEM@NAACL-HLT 2019 - 8th Joint Conference on Lexical and Computational Semantics (2019) 186-191

DOI: 10.18653/v1/S19-1020

13Citations

77Readers

Abstract

Noise is inherent in real world datasets and modeling noise is critical during training as it is effective in regularization. Recently, novel semi-supervised deep learning techniques have demonstrated tremendous potential when learning with very limited labeled training data in image processing tasks. A critical aspect of these semi-supervised learning techniques is augmenting the input or the network with noise to be able to learn robust models. While modeling noise is relatively straightforward in continuous domains such as image classification, it is not immediately apparent how noise can be modeled in discrete domains such as language. Our work aims to address this gap by exploring different noise strategies for the semi-supervised named entity classification task, including statistical methods such as adding Gaussian noise to input embeddings, and linguistically-inspired ones such as dropping words and replacing words with their synonyms. We compare their performance on two benchmark datasets (OntoNotes and CoNLL) for named entity classification. Our results indicate that noise strategies that are linguistically informed perform at least as well as statistical approaches, while being simpler and requiring minimal tuning.

Cite

CITATION STYLE

APA

Narayan, P. L., Nagesh, A., & Surdeanu, M. (2019). Exploration of noise strategies in semi-supervised named entity classification. In *SEM@NAACL-HLT 2019 - 8th Joint Conference on Lexical and Computational Semantics (pp. 186–191). Association for Computational Linguistics (ACL). https://doi.org/10.18653/v1/S19-1020

Exploration of noise strategies in semi-supervised named entity classification

Abstract

Cite

Register to see more suggestions