Positive and Unlabeled Examples Help Learning

Abstract : In many learning problems, labeled examples are rare or expensive while numerous unlabeled and positive examples are available. However, most learning algorithms only use labeled examples. Thus we address the problem of learning with the help of positive and unlabeled data given a small number of labeled examples. We present both theoretical and empirical arguments showing that learning algorithms can be improved by the use of both unlabeled and positive data. As an illustrating problem, we consider the learning algorithm from statistics for monotone conjunctions in the presence of classification noise and give empirical evidence of our assumptions. We give theoretical results for the improvement of Statistical Query learning algorithms from positive and unlabeled data. Lastly, we apply these ideas to tree induction algorithms. We modify the code of C4.5 to get an algorithm which takes as input a set LAB of labeled examples, a set POS of positive examples and a set UNL of unlabeled data and which uses these three sets to construct the decision tree. We provide experimental results based on data taken from UCI repository which confirm the relevance of this approach.
Type de document :
Communication dans un congrès
Proceedings of the Tenth International Conference on Algorithmic Learning Theory, ALT'99, 1999, Tokyo, Japan. Springer Verlag, pp.219--230, 1999, Lecture Notes in Artificial Intelligence
Liste complète des métadonnées

https://hal.inria.fr/inria-00538885
Contributeur : Rémi Gilleron <>
Soumis le : mardi 23 novembre 2010 - 14:48:40
Dernière modification le : jeudi 11 janvier 2018 - 06:21:19

Identifiants

  • HAL Id : inria-00538885, version 1

Collections

Citation

Françesco De Comité, François Denis, Rémi Gilleron, Fabien Letouzey. Positive and Unlabeled Examples Help Learning. Proceedings of the Tenth International Conference on Algorithmic Learning Theory, ALT'99, 1999, Tokyo, Japan. Springer Verlag, pp.219--230, 1999, Lecture Notes in Artificial Intelligence. 〈inria-00538885〉

Partager

Métriques

Consultations de la notice

101