TextBrewer: An open-source knowledge distillation toolkit for natural language processing

21Citations
Citations of this article
143Readers
Mendeley users who have this article in their library.

Abstract

In this paper, we introduce TextBrewer, an open-source knowledge distillation toolkit designed for natural language processing. It works with different neural network models and supports various kinds of supervised learning tasks, such as text classification, reading comprehension, sequence labeling. TextBrewer provides a simple and uniform workflow that enables quick setting up of distillation experiments with highly flexible configurations. It offers a set of predefined distillation methods and can be extended with custom code. As a case study, we use TextBrewer to distill BERT on several typical NLP tasks. With simple configurations, we achieve results that are comparable with or even higher than the public distilled BERT models with similar numbers of parameters.

Cite

CITATION STYLE

APA

Yang, Z., Cui, Y., Chen, Z., Che, W., Liu, T., Wang, S., & Hu, G. (2020). TextBrewer: An open-source knowledge distillation toolkit for natural language processing. In Proceedings of the Annual Meeting of the Association for Computational Linguistics (pp. 9–16). Association for Computational Linguistics (ACL). https://doi.org/10.18653/v1/2020.acl-demos.2

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free