Prompt Tuning for Discriminative Pre-trained Language Models

15Citations
Citations of this article
68Readers
Mendeley users who have this article in their library.

Abstract

Recent works have shown promising results of prompt tuning in stimulating pre-trained language models (PLMs) for natural language processing (NLP) tasks. However, to the best of our knowledge, existing works focus on prompt-tuning generative PLMs that are pre-trained to generate target tokens, such as BERT (Devlin et al., 2019). It is still unknown whether and how discriminative PLMs, e.g., ELECTRA (Clark et al., 2020), can be effectively prompt-tuned. In this work, we present DPT, the first prompt tuning framework for discriminative PLMs, which reformulates NLP tasks into a discriminative language modeling problem. Comprehensive experiments on text classification and question answering show that, compared with vanilla fine-tuning, DPT achieves significantly higher performance, and also prevents the unstable problem in tuning large PLMs in both full-set and low-resource settings. The source code and experiment details of this paper can be obtained from https://github.com/thunlp/DPT.

Cite

CITATION STYLE

APA

Yao, Y., Dong, B., Zhang, A., Zhang, Z., Xie, R., Liu, Z., … Wang, J. (2022). Prompt Tuning for Discriminative Pre-trained Language Models. In Proceedings of the Annual Meeting of the Association for Computational Linguistics (pp. 3468–3473). Association for Computational Linguistics (ACL). https://doi.org/10.18653/v1/2022.findings-acl.273

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free