Datasets › MIR-FLICKR25K

MIR-FLICKR25K

archive 2025-07-28

Click to add a brief description of the dataset (Markdown and LaTeX enabled).

Provide:

Data Documentation

This dataset is a classic multi-label collection from Flickr featuring 25,000 images with text descriptions and tag annotations, widely used for image classification, multi-label classification, and cross-modal retrieval. The package contains two main files: mirflickr25k (with 25,000 images named im*.jpg) and mirflickr25k_annotations_v080 (containing 24 category labels). The text descriptions are preprocessed and include 1386 tags that appear in at least 20 images. Most research filters out zero vectors, resulting in 20,015 usable samples. Traditional approaches extract features using pre-trained models like VGG19 for images (4096-D) and BOW or TextCNN for text (1386-D or 300-D), while modern methods often employ Transformer architectures for downstream tasks. * a high-level explanation of the dataset characteristics * explain motivations and summary of its content * potential use cases of the dataset

Benchmarks archive 2025-07-28

No leaderboard in the archive resolves to this dataset.

Papers archive 2025-07-28

No paper in the archive has a leaderboard row on this dataset; the archive counts 1 paper for it but never published that list.

Dataset loaders archive 2025-07-28

No loader listed in the archive.

Tasks archive 2025-07-28

No task tagged in the archive.

License archive 2025-07-28

No licence recorded in the archive. Absence here is not a statement about the dataset's terms.

Modalities archive 2025-07-28

No modality tagged.

Languages archive 2025-07-28

No language tagged.

Variants archive 2025-07-28

  • MIR-FLICKR25K

1 variant name, as the archive lists them.

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections