Uncertainty Modeling for Machine Comprehension Systems using Efficient Bayesian Neural Networks

0Citations
Citations of this article
68Readers
Mendeley users who have this article in their library.

Abstract

While neural approaches have achieved significant improvement in machine comprehension tasks, models often work as a black-box, resulting in lower interpretability, which requires special attention in domains such as healthcare or education. Quantifying uncertainty helps pave the way towards more interpretable neural networks. In classification and regression tasks, Bayesian neural networks have been effective in estimating model uncertainty. However, inference time increases linearly due to the required sampling process in Bayesian neural networks. Thus speed becomes a bottleneck in tasks with high system complexity such as question-answering or dialogue generation. In this work, we propose a hybrid neural architecture to quantify model uncertainty using Bayesian weight approximation but boosts up the inference speed by 80% relative at test time, and apply it for a clinical dialogue comprehension task. The proposed approach is also used to enable active learning so that an updated model can be trained more optimally with new incoming data by selecting samples that are not well-represented in the current training scheme.

Cite

CITATION STYLE

APA

Liu, Z., Krishnaswamy, P., Aw, A. T., & Chen, N. F. (2020). Uncertainty Modeling for Machine Comprehension Systems using Efficient Bayesian Neural Networks. In COLING 2020 - 28th International Conference on Computational Linguistics, Proceedings of the Industry Track (pp. 228–235). Association for Computational Linguistics (ACL). https://doi.org/10.18653/v1/2020.coling-industry.21

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free