Distributed training strategies for the structured perceptron

240Citations
Citations of this article
288Readers
Mendeley users who have this article in their library.

Abstract

Perceptron training is widely applied in the natural language processing community for learning complex structured models. Like all structured prediction learning frameworks, the structured perceptron can be costly to train as training complexity is proportional to inference, which is frequently non-linear in example sequence length. In this paper we investigate distributed training strategies for the structured perceptron as a means to reduce training times when computing clusters are available. We look at two strategies and provide convergence bounds for a particular mode of distributed structured perceptron training based on iterative parameter mixing (or averaging). We present experiments on two structured prediction problems - named-entity recognition and dependency parsing - to highlight the efficiency of this method. © 2010 Association for Computational Linguistics.

Cite

CITATION STYLE

APA

McDonald, R., Hall, K., & Mann, G. (2010). Distributed training strategies for the structured perceptron. In NAACL HLT 2010 - Human Language Technologies: The 2010 Annual Conference of the North American Chapter of the Association for Computational Linguistics, Proceedings of the Main Conference (pp. 456–464).

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free