Pith. sign in

REVIEW 1 cited by

Why the pseudo label based semi-supervised learning algorithm is effective?

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2211.10039 v2 pith:7EXL2R2Q submitted 2022-11-18 cs.LG cs.AI

classification cs.LGcs.AI
keywords pseudodatalearningmodelsemi-supervisedlabelalgorithmbound
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Recently, pseudo label based semi-supervised learning has achieved great success in many fields. The core idea of the pseudo label based semi-supervised learning algorithm is to use the model trained on the labeled data to generate pseudo labels on the unlabeled data, and then train a model to fit the previously generated pseudo labels. We give a theory analysis for why pseudo label based semi-supervised learning is effective in this paper. We mainly compare the generalization error of the model trained under two settings: (1) There are N labeled data. (2) There are N unlabeled data and a suitable initial model. Our analysis shows that, firstly, when the amount of unlabeled data tends to infinity, the pseudo label based semi-supervised learning algorithm can obtain model which have the same generalization error upper bound as model obtained by normally training in the condition of the amount of labeled data tends to infinity. More importantly, we prove that when the amount of unlabeled data is large enough, the generalization error upper bound of the model obtained by pseudo label based semi-supervised learning algorithm can converge to the optimal upper bound with linear convergence rate. We also give the lower bound on sampling complexity to achieve linear convergence rate. Our analysis contributes to understanding the empirical successes of pseudo label-based semi-supervised learning.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Label-Free Concept Drift Assessment for Reliable AI in Emerging Wireless Applications

    cs.NI 2025-07 reject novelty 5.0 of 10

    CFPT and TabAutoDrift detect concept drift by comparing macro-F1 scores from pseudo-label or transfer-learning retraining, reaching F1 of 0.94 in fingerprinting and 1.00 in link anomalies without post-deployment labels.

Pith tools