Question Answering Infused Pre-training of General-Purpose Contextualized Representations

14Citations
Citations of this article
86Readers
Mendeley users who have this article in their library.

Abstract

We propose a pre-training objective based on question answering (QA) for learning general-purpose contextual representations, motivated by the intuition that the representation of a phrase in a passage should encode all questions that the phrase can answer in context. To this end, we train a bi-encoder QA model, which independently encodes passages and questions, to match the predictions of a more accurate cross-encoder model on 80 million synthesized QA pairs. By encoding QA-relevant information, the bi-encoder's token-level representations are useful for non-QA downstream tasks without extensive (or in some cases, any) fine-tuning. We show large improvements over both RoBERTa-large and previous state-of-the-art results on zero-shot and few-shot paraphrase detection on four datasets, few-shot named entity recognition on two datasets, and zero-shot sentiment analysis on three datasets.

Cite

CITATION STYLE

APA

Jia, R., Lewis, M., & Zettlemoyer, L. (2022). Question Answering Infused Pre-training of General-Purpose Contextualized Representations. In Proceedings of the Annual Meeting of the Association for Computational Linguistics (pp. 711–728). Association for Computational Linguistics (ACL). https://doi.org/10.18653/v1/2022.findings-acl.59

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free