objectiveIntroduced by VirTex · 2020
Captioning as visual pre-training
Learn visual backbones by generating captions, which is data-efficient compared to classification.
Drafted by AI · not yet reviewed
How this idea evolved
No earlier ideas recorded for this concept yet.
Papers using this
- 2021CLIP