Datasets › CIRCO

CIRCO (Composed Image Retrieval on Common Objects in context)

Introduced by Alberto Baldrati et al. in Zero-Shot Composed Image Retrieval with Textual Inversion27 Mar 2023 archive 2025-07-28

CIRCO (Composed Image Retrieval on Common Objects in context) is an open-domain benchmarking dataset for Composed Image Retrieval (CIR) based on real-world images from COCO 2017 unlabeled set. It is the first CIR dataset with multiple ground truths and aims to address the problem of false negatives in existing datasets. CIRCO comprises a total of 1020 queries, randomly divided into 220 and 800 for the validation and test set, respectively, with an average of 4.53 ground truths per query.

Source: Zero-Shot Composed Image Retrieval with Textual Inversion

Image Source: Zero-Shot Composed Image Retrieval with Textual Inversion

Benchmarks archive 2025-07-28

All 1 leaderboard whose dataset resolves to this page shown (sort by any header). "First row" is the archive's own first row at snapshot, in the archive's row order; nothing here re-ranks and metric direction is not asserted.

First row (archive order)PaperCode
Zero-Shot Composed Image Retrieval (ZS-CIR) CIRCO MMRet-MLLM mAP@10 43.4 MegaPairs: Massive Data Synthesis For Universal... VectorSpaceLab/MegaPairs 43 Compare

Papers archive 2025-07-28

20 shown of 20 papers with a leaderboard row on this dataset's benchmarks, newest first. The archive's own "papers using this dataset" list was never published, so this is the benchmark-backed subset; the archive's count for this dataset is 35. The Syntology column is from Syntology's graph (read 2026-09-24), stated per sample; it is not part of any archive number.

DateSamples run Syntology
CoLLM: A Large Language Model for Composed Image Retrieval 1 2 25 Mar 2025 not harvested
ImageScope: Unifying Language-Guided Image Retrieval via Large Multimodal Model Collective Reasoning 1 1 13 Mar 2025 not harvested
SCOT: Self-Supervised Contrastive Pretraining For Zero-Shot Compositional Retrieval 0 1 12 Jan 2025 not harvested
MegaPairs: Massive Data Synthesis For Universal Multimodal Retrieval 1 3 19 Dec 2024 not harvested
Reason-before-Retrieve: One-Stage Reflective Chain-of-Thoughts for Training-Free Zero-Shot Composed Image Retrieval 1 3 15 Dec 2024 not harvested
Imagine and Seek: Improving Composed Image Retrieval with an Imagined Proxy 0 2 24 Nov 2024 not harvested
Semantic Editing Increment Benefits Zero-Shot Composed Image Retrieval 2 3 28 Oct 2024 not harvested
LDRE: LLM-based Divergent Reasoning and Ensemble for Zero-Shot Composed Image Retrieval 2 3 11 Jul 2024 not harvested
An Efficient Post-hoc Framework for Reducing Task Discrepancy of Text Encoders for Composed Image Retrieval 1 2 13 Jun 2024 not harvested
iSEARLE: Improving Textual Inversion for Zero-Shot Composed Image Retrieval 2 4 5 May 2024 ran 3 of 4 samples (1 unverified; 4 pointer-only for licence)
MagicLens: Self-Supervised Image Retrieval with Open-Ended Instructions 1 4 28 Mar 2024 ran 1 of 1 samples (0 unverified)
Language-only Efficient Training of Zero-shot Composed Image Retrieval 1 2 4 Dec 2023 ran 3 of 8 samples (5 unverified; 8 pointer-only for licence)
Pretrain like Your Inference: Masked Tuning Improves Zero-Shot Composed Image Retrieval 1 2 13 Nov 2023 ran 3 of 4 samples (1 unverified; 4 pointer-only for licence)
Vision-by-Language for Training-Free Compositional Image Retrieval 1 3 13 Oct 2023 ran 4 of 8 samples (4 unverified)
Context-I2W: Mapping Images to Context-dependent Words for Accurate Zero-Shot Composed Image Retrieval 1 1 28 Sep 2023 ran 1 of 1 samples (0 unverified)
CoVR-2: Automatic Data Construction for Composed Video Retrieval 1 1 28 Aug 2023 ran 1 of 1 samples (0 unverified)
Zero-Shot Composed Image Retrieval with Textual Inversion 2 2 27 Mar 2023 not harvested
CompoDiff: Versatile Composed Image Retrieval With Latent Diffusion 1 2 21 Mar 2023 ran 3 of 6 samples (3 unverified)
Pic2Word: Mapping Pictures to Words for Zero-shot Composed Image Retrieval 1 1 6 Feb 2023 ran 1 of 1 samples (0 unverified)
"This is my unicorn, Fluffy": Personalizing frozen vision-language representations 2 1 4 Apr 2022 ran 1 of 1 samples (0 unverified; 1 pointer-only for licence)

Dataset loaders archive 2025-07-28

1 loader as listed in the archive; links are outbound and not re-checked here.

Tasks archive 2025-07-28

License archive 2025-07-28

Creative Commons BY-NC 4.0

Modalities archive 2025-07-28

Languages archive 2025-07-28

Variants archive 2025-07-28

  • CIRCO

1 variant name, as the archive lists them.

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections