Speech, voice, text, and meaning: A multidisciplinary approach to interview data through the use of digital tools

1Citations
Citations of this article
7Readers
Mendeley users who have this article in their library.
Get full text

Abstract

Interview data is multimodal data: it consists of speech sound, facial expression and gestures, captured in a particular situation, and containing textual information and emotion. This workshop shows how a multidisciplinary approach may exploit the full potential of interview data. The workshop first gives a systematic overview of the research fields working with interview data. It then presents the speech technology currently available to support transcribing and annotating interview data, such as automatic speech recognition, speaker diarization, and emotion detection. Finally, scholars who work with interview data and tools may present their work and discover how to make use of existing technology.

Cite

CITATION STYLE

APA

Van Hessen, A., Calamai, S., Van Den Heuvel, H., Scagliola, S., Karrouche, N., Beeken, J., … Draxler, C. (2020). Speech, voice, text, and meaning: A multidisciplinary approach to interview data through the use of digital tools. In ICMI 2020 Companion - Companion Publication of the 2020 International Conference on Multimodal Interaction (pp. 454–455). Association for Computing Machinery, Inc. https://doi.org/10.1145/3395035.3425657

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free