pith. machine review for the scientific record.
sign in

arxiv: 1003.0024 · v1 · submitted 2010-02-26 · 💻 cs.LG

Asymptotic Analysis of Generative Semi-Supervised Learning

classification 💻 cs.LG
keywords learningaccuracyanalysisasymptoticframeworkgenerativelabelingsemi-supervised
0
0 comments X
read the original abstract

Semisupervised learning has emerged as a popular framework for improving modeling accuracy while controlling labeling cost. Based on an extension of stochastic composite likelihood we quantify the asymptotic accuracy of generative semi-supervised learning. In doing so, we complement distribution-free analysis by providing an alternative framework to measure the value associated with different labeling policies and resolve the fundamental question of how much data to label and in what manner. We demonstrate our approach with both simulation studies and real world experiments using naive Bayes for text classification and MRFs and CRFs for structured prediction in NLP.

This paper has not been read by Pith yet.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.