training techniqueIntroduced by Billion-scale semi-supervised · 2019
Teacher–student self-training
A teacher labels a huge unlabelled pool; a student learns from those pseudo-labels.
Drafted by AI · not yet reviewed
How this idea evolved
Suggested from citations. No curator has recorded what this idea builds on yet. These ideas from the same theme come from papers that Billion-scale semi-supervised cites, directly or one step removed. Citation is a fact; the connection between the ideas is not verified.
Pre-train on billions of social-media images labelled only by their hashtags.
Performance keeps rising (logarithmically) with 100× more labelled images.
Use huge amounts of cheap, noisy labels instead of clean annotations.
Papers using this
- 2017TagLM
- 2018Instance discrimination
- 2019Noisy Student
- 2020SimCLR
- 2020BYOL
- 2020ViT
- 2021ALIGN