Estimate unlabeled-data-distribution for semi-supervised PU learning

Haoji Hu; Chaofeng Sha; Xiaoling Wang; Aoying Zhou

Conference Proceedings

Estimate unlabeled-data-distribution for semi-supervised PU learning

Lecture Notes in Computer Science (including subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics) (2012) 7235 LNCS 22-33

DOI: 10.1007/978-3-642-29253-8_3

0Citations

6Readers

Get full text

Abstract

Traditional supervised classifiers use only labeled data (features/label pairs) as the training set, while the unlabeled data is used as the testing set. In practice, it is often the case that the labeled data is hard to obtain and the unlabeled data contains the instances that belong to the predefined class beyond the labeled data categories. This problem has been widely studied in recent years and the semi-supervised learning is an efficient solution to learn from positive and unlabeled examples(or PU learning). Among all the semi-supervised PU learning methods, it's hard to choose just one approach to fit all unlabeled data distribution. This paper proposes an automatic KL-divergence based semi-supervised learning method by using unlabeled data distribution knowledge. Meanwhile, a new framework is designed to integrate different semi-supervised PU learning algorithms in order to take advantage of the former methods. The experimental results show that (1)data distribution information is very helpful for the semi-supervised PU learning method; (2)the proposed framework can achieve higher precision when compared with the-state-of-the-art method. © 2012 Springer-Verlag Berlin Heidelberg.

Cite

CITATION STYLE

APA

Hu, H., Sha, C., Wang, X., & Zhou, A. (2012). Estimate unlabeled-data-distribution for semi-supervised PU learning. In Lecture Notes in Computer Science (including subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics) (Vol. 7235 LNCS, pp. 22–33). https://doi.org/10.1007/978-3-642-29253-8_3

Estimate unlabeled-data-distribution for semi-supervised PU learning

Abstract

Cite

Register to see more suggestions