Revisiting the evaluation of theory of mind through question answering

Matthew Le; Y. Lan Boureau; Maximilian Nickel

Conference ProceedingsOPEN ACCESS

Revisiting the evaluation of theory of mind through question answering

EMNLP-IJCNLP 2019 - 2019 Conference on Empirical Methods in Natural Language Processing and 9th International Joint Conference on Natural Language Processing, Proceedings of the Conference (2019) 5872-5877

DOI: 10.18653/v1/D19-1598

87Citations

108Readers

Abstract

Theory of mind, i.e., the ability to reason about intents and beliefs of agents is an important task in artificial intelligence and central to resolving ambiguous references in natural language dialogue. In this work, we revisit the evaluation of theory of mind through question answering. We show that current evaluation methods are flawed and that existing benchmark tasks can be solved without theory of mind due to dataset biases. Based on prior work, we propose an improved evaluation protocol and dataset in which we explicitly control for data regularities via a careful examination of the answer space. We show that state-of-the-art methods which are successful on existing benchmarks fail to solve theory-of-mind tasks in our proposed approach.

Cite

CITATION STYLE

APA

Le, M., Boureau, Y. L., & Nickel, M. (2019). Revisiting the evaluation of theory of mind through question answering. In EMNLP-IJCNLP 2019 - 2019 Conference on Empirical Methods in Natural Language Processing and 9th International Joint Conference on Natural Language Processing, Proceedings of the Conference (pp. 5872–5877). Association for Computational Linguistics. https://doi.org/10.18653/v1/D19-1598

Revisiting the evaluation of theory of mind through question answering

Abstract

Cite

Register to see more suggestions