A semi-supervised learning approach to why-question answering

46Citations
Citations of this article
45Readers
Mendeley users who have this article in their library.

Abstract

We propose a semi-supervised learning method for improving why-question answering (why-QA). The key of our method is to generate training data (question-Answer pairs) from causal relations in texts such as "[Tsunamis are generated]effect because [the ocean's water mass is displaced by an earthquake]cause." A naive method for the generation would be to make a question-Answer pair by simply converting the effect part of the causal relations into a why-question, like "Why are tsunamis generated?" from the above example, and using the source text of the causal relations as an answer. However, in our preliminary experiments, this naive method actually failed to improve the why-QA performance. The main reason was that the machine-generated questions were often incomprehensible like "Why does (it) happen?", and that the system suffered from overfitting to the results of our automatic causality recognizer. Hence, we developed a novel method that effectively filters out incomprehensible questions and retrieves from texts answers that are likely to be paraphrases of a given causal relation. Through a series of experiments, we showed that our approach significantly improved the precision of the top answer by 8% over the current state-of-The-Art system for Japanese why-QA.

Cite

CITATION STYLE

APA

Oh, J. H., Torisawa, K., Hashimoto, C., Iida, R., Tanaka, M., & Kloetzer, J. (2016). A semi-supervised learning approach to why-question answering. In 30th AAAI Conference on Artificial Intelligence, AAAI 2016 (pp. 3022–3029). AAAI press. https://doi.org/10.1609/aaai.v30i1.10388

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free