TnT - A Statistical Part-of-Speech Tagger

  • Brants T
ArXiv: cs/0003055
N/ACitations
Citations of this article
440Readers
Mendeley users who have this article in their library.

Abstract

Trigrams'n'Tags (TnT) is an efficient statistical part-of-speech tagger. Contrary to claims found elsewhere in the literature, we argue that a tagger based on Markov models performs at least as well as other current approaches, including the Maximum Entropy framework. A recent comparison has even shown that TnT performs significantly better for the tested corpora. We describe the basic model of TnT, the techniques used for smoothing and for handling unknown words. Furthermore, we present evaluations on two corpora.

Cite

CITATION STYLE

APA

Brants, T. (2000). TnT - A Statistical Part-of-Speech Tagger. Retrieved from http://arxiv.org/abs/cs/0003055

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free