Consistency-Preserving Visual Question Answering in Medical Imaging

Sergio Tascon-Morales; Pablo Márquez-Neila; Raphael Sznitman

Conference Proceedings

Consistency-Preserving Visual Question Answering in Medical Imaging

Lecture Notes in Computer Science (including subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics) (2022) 13438 LNCS 386-395

DOI: 10.1007/978-3-031-16452-1_37

4Citations

12Readers

Get full text

Abstract

Visual Question Answering (VQA) models take an image and a natural-language question as input and infer the answer to the question. Recently, VQA systems in medical imaging have gained popularity thanks to potential advantages such as patient engagement and second opinions for clinicians. While most research efforts have been focused on improving architectures and overcoming data-related limitations, answer consistency has been overlooked even though it plays a critical role in establishing trustworthy models. In this work, we propose a novel loss function and corresponding training procedure that allows the inclusion of relations between questions into the training process. Specifically, we consider the case where implications between perception and reasoning questions are known a-priori. To show the benefits of our approach, we evaluate it on the clinically relevant task of Diabetic Macular Edema (DME) staging from fundus imaging. Our experiments show that our method outperforms state-of-the-art baselines, not only by improving model consistency, but also in terms of overall model accuracy. Our code and data are available at https://github.com/sergiotasconmorales/consistency_vqa.

Author supplied keywords

Cite

CITATION STYLE

APA

Tascon-Morales, S., Márquez-Neila, P., & Sznitman, R. (2022). Consistency-Preserving Visual Question Answering in Medical Imaging. In Lecture Notes in Computer Science (including subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics) (Vol. 13438 LNCS, pp. 386–395). Springer Science and Business Media Deutschland GmbH. https://doi.org/10.1007/978-3-031-16452-1_37

Consistency-Preserving Visual Question Answering in Medical Imaging

Abstract

Author supplied keywords

Cite

Register to see more suggestions