evaluationIntroduced by Scaling ViTs (ViT-G) · 2021
Scaling laws for vision Transformers
How ViT error falls with model size, data and compute, up to billions of parameters.
Drafted by AI · not yet reviewed
How this idea evolved
Suggested from citations. No curator has recorded what this idea builds on yet. These ideas from the same theme come from papers that Scaling ViTs (ViT-G) cites, directly or one step removed. Citation is a fact; the connection between the ideas is not verified.
Split an image into patches and feed them to a plain Transformer as tokens.