Learning diphone-based segmentation

68Citations
Citations of this article
58Readers
Mendeley users who have this article in their library.

This article is free to access.

Abstract

This paper reconsiders the diphone-based word segmentation model of Cairns, Shillcock, Chater, and Levy (1997) and Hockema (2006), previously thought to be unlearnable. A statistically principled learning model is developed using Bayes' theorem and reasonable assumptions about infants' implicit knowledge. The ability to recover phrase-medial word boundaries is tested using phonetic corpora derived from spontaneous interactions with children and adults. The (unsupervised and semi-supervised) learning models are shown to exhibit several crucial properties. First, only a small amount of language exposure is required to achieve the model's ceiling performance, equivalent to between 1day and 1month of caregiver input. Second, the models are robust to variation, both in the free parameter and the input representation. Finally, both the learning and baseline models exhibit undersegmentation, argued to have significant ramifications for speech processing as a whole. Copyright © 2010 Cognitive Science Society, Inc.

Cite

CITATION STYLE

APA

Daland, R., & Pierrehumbert, J. B. (2011). Learning diphone-based segmentation. Cognitive Science, 35(1), 119–155. https://doi.org/10.1111/j.1551-6709.2010.01160.x

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free