Learning and Evaluating Character Representations in Novels

13Citations
Citations of this article
40Readers
Mendeley users who have this article in their library.

Abstract

We address the problem of learning fixed-length vector representations of characters in novels. Recent advances in word embeddings have proven successful in learning entity representations from short texts, but fall short on longer documents because they do not capture full book-level information. To overcome the weakness of such text-based embeddings, we propose two novel methods for representing characters: (i) graph neural network-based embeddings from a full corpus-based character network; and (ii) low-dimensional embeddings constructed from the occurrence pattern of characters in each novel. We test the quality of these character embeddings using a new benchmark suite to evaluate character representations, encompassing 12 different tasks. We show that our representation techniques combined with text-based embeddings lead to the best character representations, outperforming text-based embeddings in four tasks. Our dataset is made publicly available to stimulate additional work in this area.

Cite

CITATION STYLE

APA

Inoue, N., Pethe, C., Kim, A., & Skiena, S. (2022). Learning and Evaluating Character Representations in Novels. In Proceedings of the Annual Meeting of the Association for Computational Linguistics (pp. 1008–1019). Association for Computational Linguistics (ACL). https://doi.org/10.18653/v1/2022.findings-acl.81

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free