TextGAIL: Generative Adversarial Imitation Learning for Text Generation

N/ACitations
Citations of this article
73Readers
Mendeley users who have this article in their library.

Abstract

Generative Adversarial Networks (GANs) for text generation have recently received many criticisms, as they perform worse than their MLE counterparts (Caccia et al. 2020; Tevet et al. 2019; Semeniuta, Severyn, and Gelly 2018). We suspect previous text GANs' inferior performance is due to the lack of a reliable guiding signal in their discriminators. To address this problem, we propose a generative adversarial imitation learning framework for text generation that uses large pre-trained language models to provide more reliable reward guidance. As previous text GANs suffer from high variance of gradients, we apply contrastive discriminator, and proximal policy optimization (PPO) to stabilize and improve text generation performance. For evaluation, we conduct experiments on a diverse set of unconditional and conditional text generation tasks. Experimental results show that TextGAIL achieves better performance in terms of both quality and diversity than the MLE baseline. We also validate our intuition that TextGAIL's discriminator demonstrates the capability of providing reasonable rewards with an additional task.

Cite

CITATION STYLE

APA

Wu, Q., Li, L., & Yu, Z. (2021). TextGAIL: Generative Adversarial Imitation Learning for Text Generation. In 35th AAAI Conference on Artificial Intelligence, AAAI 2021 (Vol. 16, pp. 14067–14075). Association for the Advancement of Artificial Intelligence. https://doi.org/10.1609/aaai.v35i16.17656

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free