Enabling Multi-modal Conversational Interface for Clinical Imaging

10Citations
Citations of this article
11Readers
Mendeley users who have this article in their library.
Get full text

Abstract

Human-computer interaction research has to play a vital role in increasing the adoption of deep learning models in clinical settings, as their adoption is low despite models surpassing/matching the clinician's performance on many medical imaging tasks. Conversational AI has been successful as an interface for general information; however, there is a research gap for multi-modal conversational interface design for safety-critical clinical imaging systems. Our research points to the important role of multi-modal chat in improving usability and explainability through textual and visual explanations. Our main contributions include design principles for conversational interfaces in clinical imaging systems, the importance of multi-modal responses, and an understanding of the usefulness of mimicking clinician/radiologist interactions to improve usability. We show that diagnosis descriptions and visual responses improve the multi-modal conversational interface. The multi-modal conversational interface can help improve the adoption of deep learning systems in clinical settings, improving clinicians' efficiency and patient outcomes.

Cite

CITATION STYLE

APA

Dayanandan, K., & Lall, B. (2024). Enabling Multi-modal Conversational Interface for Clinical Imaging. In Conference on Human Factors in Computing Systems - Proceedings. Association for Computing Machinery. https://doi.org/10.1145/3613905.3650805

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free