architectureIntroduced by Bottom-Up Top-Down attention · 2017
Object-detector region features
Represent an image by features of detected object regions instead of a pixel grid.
Drafted by AI · not yet reviewed
How this idea evolved
No earlier ideas recorded for this concept yet.
Papers using this
- 2018Bilinear Attention Networks
- 2019ViLBERT
- 2019LXMERT
- 2020Grid features for VQA
- 2020Oscar
- 2021VinVL
- 2021ViLT
- 2021VLMo
- 2022VL-BEiT