Datasets › DiDeMo › Papers where code ran, page 1
DiDeMo (Distinct Describable Moments)
Papers archive 2025-07-28
papers with a benchmark row: 46 · with a code link: 40 · where Syntology ran a sample: 23 (19 with a run with no instrument failure, 4 where every run was a failure of Syntology's instrument) Syntology
Show:
all papers with a benchmark row only where code ran (23 of 46 with a benchmark row: 19 with a run with no instrument failure, 4 where every run was a failure of Syntology's instrument)
Syntology We ran code from the paper's repository; we did not run it on this dataset or check it against this dataset's benchmarks.
Page 1 of 1: papers 1 to 23 of the 23 papers with a benchmark row here where Syntology ran at least one harvested sample (19 with a run with no instrument failure, 4 where every run was a failure of Syntology's instrument), newest first by the archive's date (ties by arXiv id). This is a filter on Syntology's record ordered by date only, not a ranking; a run is not a correctness claim. A paper missing from this list is not a recorded non-run: it may have no arXiv id, no harvested code, or only samples that have not run yet.
A run counts when it is a run with no instrument failure any run, instrument failures included The first hides, in your browser, the papers where every run was a failure of Syntology's instrument; the second shows every paper on this page.
The archive's own "papers using this dataset" list was never published, so this is the benchmark-backed subset, not that list; the archive's count for this dataset is 216. The Syntology column is from Syntology's graph, stated per sample; it is not part of any archive number. A line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified” (C is a part of N, never taken away from it); the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. When the archive marks a repository official for the paper, the cell starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from. Syntology's record for this page has not changed since 2026-09-28 , the first build that kept a record date for it; when this build read Syntology's graph is in the build record .
Paper Code Results Date Samples run Syntology
Gramian Multimodal Representation Learning and Alignment
2
2
16 Dec 2024
official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 1 honoured, 1 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (1 pointer-only for licence )
vid-TLDR: Training Free Token merging for Light-weight Video Transformer
1
2
20 Mar 2024
official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified
TESTA: Temporal-Spatial Token Aggregation for Long-form Video-Language Understanding
1
1
29 Oct 2023
official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 1 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (5 pointer-only for licence )
LanguageBind: Extending Video-Language Pretraining to N-modality by Language-based Semantic Alignment
6
2
3 Oct 2023
official (archive's flag): 5 ran · 10 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 2 honoured, 0 violated, 7 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (3 pointer-only for licence )
Prototype-based Aleatoric Uncertainty Quantification for Cross-modal Retrieval
1
1
29 Sep 2023
official (archive's flag): 17 ran · 17 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 2 honoured, 1 violated, 7 with no contract checked; 7 where Syntology's instrument failed) · 2 unverified (8 pointer-only for licence )
BT-Adapter: Video Conversation is Feasible Without Video Instruction Tuning
1
1
27 Sep 2023
official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (2 pointer-only for licence )
VAST: A Vision-Audio-Subtitle-Text Omni-Modality Foundation Model and Dataset
2
2
29 May 2023
official (archive's flag): 12 ran · 35 ran (of which 4 constructed an object rather than computing a result; 29 with no instrument failure: 2 honoured, 1 violated, 26 with no contract checked; 6 where Syntology's instrument failed) · 7 unverified (8 pointer-only for licence )
Unmasked Teacher: Towards Training-Efficient Video Foundation Models
1
2
28 Mar 2023
official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 1 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (3 pointer-only for licence )
Video-Text as Game Players: Hierarchical Banzhaf Interaction for Cross-Modal Representation Learning
4
1
25 Mar 2023
official (archive's flag): 1 ran · 12 ran (of which 7 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified
DiffusionRet: Generative Text-Video Retrieval with Diffusion Model
4
2
17 Mar 2023
official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified
mPLUG-2: A Modularized Multi-modal Foundation Model Across Text, Image and Video
4
2
1 Feb 2023
community repositories only · 17 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 8 where Syntology's instrument failed) · 2 unverified (2 pointer-only for licence )
Cap4Video: What Can Auxiliary Captions Do for Text-Video Retrieval?
4
1
31 Dec 2022
official (archive's flag): 14 ran · 20 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 2 honoured, 1 violated, 9 with no contract checked; 8 where Syntology's instrument failed) · 5 unverified (11 pointer-only for licence )
InternVideo: General Video Foundation Models via Generative and Discriminative Learning
2
2
6 Dec 2022
official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified
CLIP-ViP: Adapting Pre-trained Image-Text Model to Video-Language Representation Alignment
1
1
14 Sep 2022
official (archive's flag): 4 ran · 4 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (4 pointer-only for licence )
X-CLIP: End-to-End Multi-grained Contrastive Learning for Video-Text Retrieval
3
1
15 Jul 2022
official: harvested, nothing ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (1 pointer-only for licence )
Revealing Single Frame Bias for Video-and-Language Learning
2
3
7 Jun 2022
community repositories only · 11 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (2 pointer-only for licence )
Disentangled Representation Learning for Text-Video Retrieval
2
1
14 Mar 2022
official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 2 honoured, 1 violated, 3 with no contract checked; 4 where Syntology's instrument failed) · 2 unverified (4 pointer-only for licence )
Bridging Video-text Retrieval with Multiple Choice Questions
2
1
13 Jan 2022
official (archive's flag): 2 ran · 13 ran (of which 8 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 4 where Syntology's instrument failed) · 11 unverified (6 pointer-only for licence )
Cross Modal Retrieval with Querybank Normalisation
1
1
23 Dec 2021
official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified
CLIP4Clip: An Empirical Study of CLIP for End to End Video Clip Retrieval
5
1
18 Apr 2021
official: harvested, nothing ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (3 pointer-only for licence )
Frozen in Time: A Joint Video and Image Encoder for End-to-End Retrieval
5
3
1 Apr 2021
official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (3 pointer-only for licence )
VLG-Net: Video-Language Graph Matching Network for Video Grounding
1
1
19 Nov 2020
1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified
Use What You Have: Video Retrieval Using Representations From Collaborative Experts
3
1
31 Jul 2019
3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (1 pointer-only for licence )