Pith. sign in

REVIEW 3 cited by

MixMatch: A Holistic Approach to Semi-Supervised Learning

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1905.02249 v2 pith:MRLTCUKI submitted 2019-05-06 cs.LG cs.AIcs.CVstat.ML

classification cs.LGcs.AIcs.CVstat.ML
keywords mixmatchdatalabeledlearningsemi-supervisedunlabeleddatasetsfactor
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Semi-supervised learning has proven to be a powerful paradigm for leveraging unlabeled data to mitigate the reliance on large labeled datasets. In this work, we unify the current dominant approaches for semi-supervised learning to produce a new algorithm, MixMatch, that works by guessing low-entropy labels for data-augmented unlabeled examples and mixing labeled and unlabeled data using MixUp. We show that MixMatch obtains state-of-the-art results by a large margin across many datasets and labeled data amounts. For example, on CIFAR-10 with 250 labels, we reduce error rate by a factor of 4 (from 38% to 11%) and by a factor of 2 on STL-10. We also demonstrate how MixMatch can help achieve a dramatically better accuracy-privacy trade-off for differential privacy. Finally, we perform an ablation study to tease apart which components of MixMatch are most important for its success.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 3 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. OpenAlex reports about 604 citations worldwide. Full citation record

  1. Scaling Up Audio-Synchronized Visual Animation: An Efficient Training Paradigm

    cs.CV 2025-08 conditional novelty 6.0 of 10

    An audio-conditioned video animation model is pretrained on noisy auto-curated videos and fine-tuned on a few clean examples, achieving top synchronization scores on a new 48-class benchmark with only 1.9% additional ...

  2. Automated Solar Radio Burst Detection Using Deep Learning on Augmented e-Callisto Data

    astro-ph.SR 2026-07 conditional novelty 5.0 of 10

    FlareSense, a ResNet detector trained on 304,750 e-Callisto spectrograms with SpecAugment and TimeWarp, reaches 93% precision and 73.15% recall, outperforming routine expert cataloging at matched precision.

  3. Energy Landscapes Enable Reliable Abstention in Retrieval-Augmented Large Language Models for Healthcare

    cs.CL 2025-08 conditional novelty 5.0 of 10

    An energy-based scoring head trained on dense embeddings improves abstention decisions for medical RAG systems on semantically hard out-of-distribution queries compared to softmax and kNN baselines.

Pith tools