Semantics-Based Automatic Literal Reconstruction of Dictations

  • Jancsary J
  • Klein A
  • Matiasek J
  • et al.
N/ACitations
Citations of this article
1Readers
Mendeley users who have this article in their library.

Abstract

This paper describes a method for the automatic literal reconstruction of dictations in the domain of medical reports. The raw output of an automatic speech recognition system and the final report edited by a professional medical transcriptionist serve as input to the reconstruction algorithm. Reconstruction is based on automatic alignment between the speech recognition result and the edited report. Based on an ontology (i.e. UMLS) and lexical resources (i.e. WordNet and an inventory of spoken variants for each concept), semantic representations are assigned to terms and phrases. Alignment takes into account semantic similarity scores, based on the similarity between semantic representations of the two sources, and phonetic similarity scores. This paper explains how the speech recognition output is compared and aligned to the edited written documents and how the two different input sources are complementary for the task of reconstructing a literal transcript.

Cite

CITATION STYLE

APA

Jancsary, J., Klein, A., Matiasek, J., & Trost, H. (2007). Semantics-Based Automatic Literal Reconstruction of Dictations. In CAEPIA 2007 Workshop on Semantic Representation of Spoken Language. Salamanca, Spain.

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free