Abstract
Recent transformer-based approaches to NLG like GPT-2 can generate syntactically coherent original texts. However, these generated texts have serious flaws: global discourse incoherence and meaninglessness of sentences in terms of entity values. We address both of these flaws: they are independent but can be combined to generate original texts that will be both consistent and truthful. This paper presents an approach to estimate the quality of discourse structure. Empirical results confirm that the discourse structure of currently generated texts is inaccurate. We propose the research directions to correct it using discourse features during the fine-tuning procedure. The suggested approach is universal and can be applied to different languages. Apart from that, we suggest a method to correct wrong entity values based on Web Mining and text alignment.
Cite
CITATION STYLE
Chernyavskiy, A., Ilvovsky, D., & Galitsky, B. (2021). Correcting Texts Generated by Transformers using Discourse Features and Web Mining. In International Conference Recent Advances in Natural Language Processing, RANLP (Vol. 2021-September, pp. 36–43). Incoma Ltd. https://doi.org/10.26615/issn.2603-2821.2021_006
Register to see more suggestions
Mendeley helps you to discover research relevant for your work.