Paper Lineage
Esc
DatasetMar 2015arXiv 1503.01817cs.MM

YFCC100M: The New Data in Multimedia Research

Bart Thomee, David A. Shamma, Gerald Friedland and 5 others

We present the Yahoo Flickr Creative Commons 100 Million Dataset (YFCC100M), the largest public multimedia collection that has ever been released. 8 million are videos, all of which carry a Creative Commons license.

From the abstract

Built on

0 papers · 0 verifiedSee as graph

YFCC100M has no earlier papers in this dataset.

Led to

  • “This type of data is available in great abundance via photo-sharing websites: specifically, we use a publicly available dataset of 100 million Flickr images and captions thomee15 (see Figure 1 for six randomly picked Fli…”
    From Weakly supervised visual features (Jouli · §Introduction
  • Visual Genome2016 · cited 4×
    “YFCC100M Thomee et al., 2016 is another large database of 100100 million images that is still largely unexplored.”
    From Visual Genome · §Related Work
  • Grid features for VQA2020 · cited 2×
    “For classification, we include a model trained on YFCC thomee2016yfcc100m, which has 92M images with image tags.”
    From Grid features for VQA · §Why do Our Grid Features Work?
  • DALL·E2021 · cited 1×, 1 in Method
    “This dataset does not include MS-COCO, but does include Conceptual Captions and a filtered subset of YFCC100M (Thomee et al. 2016).”
    From DALL·E · §Method
  • CLIP2021 · cited 2×, 1 in Method
    “Existing work has mainly used three datasets, MS-COCO (Lin et al. 2014), Visual Genome (Krishna et al. 2017), and YFCC100M (Thomee et al. 2016).”
    From CLIP · §Approach
  • LiT2021 · cited 2×
    “Furthermore, we explore publicly available datasets such as YFCC100m yfcc100m and CC12M cc12m.”
    From LiT · §Introduction
Abstract

We present the Yahoo Flickr Creative Commons 100 Million Dataset (YFCC100M), the largest public multimedia collection that has ever been released. The dataset contains a total of 100 million media objects, of which approximately 99.2 million are photos and 0.8 million are videos, all of which carry a Creative Commons license. Each media object in the dataset is represented by several pieces of metadata, e.g. Flickr identifier, owner name, camera, title, tags, geo, media source. The collection provides a comprehensive snapshot of how photos and videos were taken, described, and shared over the years, from the inception of Flickr in 2004 until early 2014. In this article we explain the rationale behind its creation, as well as the implications the dataset has for science, research, engineering, and development. We further present several new challenges in multimedia research that can now be expanded upon with our dataset.