Abstract
The paper describes the process of building the electronic corpus of 17th- and 18th-century Polish texts, a relatively large, balanced, structurally and morphologically annotated resource of the Middle Polish language, available for searching at https://www.korba.edu.pl. The corpus consists of samples extracted from over seven hundred texts written and published between 1601 and 1772, summing up to a total size of 13.5 million tokens which makes it one of the largest historical corpora for a Slavic language.
Author supplied keywords
Cite
CITATION STYLE
Gruszczyński, W., Adamiec, D., Bronikowska, R., Kieraś, W., Modrzejewski, E., Wieczorek, A., & Woliński, M. (2022). The Electronic Corpus of 17th- and 18th-century Polish Texts. Language Resources and Evaluation, 56(1), 309–332. https://doi.org/10.1007/s10579-021-09549-1
Register to see more suggestions
Mendeley helps you to discover research relevant for your work.