objectiveIntroduced by Oscar · 2020
Object tags as alignment anchors
Feed detected object labels as text so they anchor image–text alignment.
Drafted by AI · not yet reviewed
How this idea evolved
Each step is the idea’s introducing paper. The line between steps is the citation link between those papers.
- 2017
Represent an image by features of detected object regions instead of a pixel grid.
No direct link between these papers in this dataset - 2020
Feed detected object labels as text so they anchor image–text alignment.