AdaPrompt: Adaptive Model Training for Prompt-based NLP

11Citations
Citations of this article
74Readers
Mendeley users who have this article in their library.

Abstract

Prompt-based learning, with its capability to tackle zero-shot and few-shot NLP tasks, has gained much attention in community. The main idea is to bridge the gap between NLP downstream tasks and language modeling (LM), by mapping these tasks into natural language prompts, which are then filled by pretrained language models (PLMs). However, for prompt learning, there are still two salient gaps between NLP tasks and pretraining. First, prompt information is not necessarily sufficiently present during LM pretraining. Second, task-specific data are not necessarily well represented during pretraining. We address these two issues by proposing AdaPrompt, adaptively retrieving external data for continual pretraining of PLMs by making use of both task and prompt characteristics. In addition, we make use of knowledge in Natural Language Inference models for deriving adaptive verbalizers. Experimental results on five NLP benchmarks show that AdaPrompt can improve over standard PLMs in few-shot settings. In addition, in zero-shot settings, our method outperforms standard prompt-based methods by up to 26.35% relative error reduction.

Cite

CITATION STYLE

APA

Chen, Y., Liu, Y., Dong, L., Wang, S., Zhu, C., Zeng, M., & Zhang, Y. (2022). AdaPrompt: Adaptive Model Training for Prompt-based NLP. In Findings of the Association for Computational Linguistics: EMNLP 2022 (pp. 6086–6097). Association for Computational Linguistics (ACL). https://doi.org/10.18653/v1/2022.findings-emnlp.448

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free