Evaluation of two systems on multi-class multi-label document classification

Xiao Luo; A. Nur Zincir-Heywood

Conference Proceedings

Evaluation of two systems on multi-class multi-label document classification

Lecture Notes in Computer Science (including subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics) (2005) 3488 LNAI 161-169

DOI: 10.1007/11425274_17

37Citations

37Readers

Get full text

Abstract

In the world of text document classification, the most general case is that in which a document can be classified into more than one category, the multi-label problem. This paper investigates the performance of two document classification systems applied to the task of multi-class multi-label document classification. Both systems consider the pattern of co-occurrences in documents of multiple categories. One system is based on a novel sequential data representation combined with a kNN classifier designed to make use of sequence information. The other is based on the "Latent Semantic Indexing" analysis combined with the traditional kNN classifier. The experimental results show that the first system performs better than the second on multi-labeled documents, while the second performs better on uni-labeled documents. Performance therefore depends on the dataset applied and the objective of the application. © Springer-Verlag Berlin Heidelberg 2005.

Cite

CITATION STYLE

APA

Luo, X., & Zincir-Heywood, A. N. (2005). Evaluation of two systems on multi-class multi-label document classification. In Lecture Notes in Computer Science (including subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics) (Vol. 3488 LNAI, pp. 161–169). Springer Verlag. https://doi.org/10.1007/11425274_17

Evaluation of two systems on multi-class multi-label document classification

Abstract

Cite

Register to see more suggestions