belgium netherlands conference on artificial intelligence(2007)
被引用1|浏览0
摘要
There are a number of supervised machine learning methods such as classiffers pretrained using restricted
Boltzmann machines and convolutional networks that work very well for handwritten character recognition.
However, they require a large amount of labeled training data to achieve good performance which
unlike unlabeled data is often expensive to obtain. In this paper a number of novel semi-supervised learning
methods for handwritten character recognition are presented based on the previous algorithms. These
methods are oriented towards learning from as little labeled data as possible and for this goal they use
unlabeled data and active learning. The proposed techniques are of varying complexity and involve simple
K-means clustering, feature mapping with self organizing maps, dimensionality reduction with deep
auto-encoders, and sub-sampling techniques. The presented algorithms outperform both a generic semisupervised
active learning algorithm and two well known supervised algorithms.