Browse State-of-the-Art › Cross-Modal Retrieval › Papers, page 3
Cross-Modal Retrieval
Papers archive 2025-07-28
archive papers tagged: 522 · with a code link: 244 · where Syntology ran a sample: 70 (56 with a run with no instrument failure, 14 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (70 of 522 tagged: 56 with a run with no instrument failure, 14 where every run was a failure of Syntology's instrument)
Page 3 of 6: papers 201 to 300 of 522, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
4 Feb 2021 1 repository listed
-
5 Jan 2021 1 repository listed Syntology official (archive's flag): 10 ran · 10 ran (of which 6 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
1 Jan 2021 1 repository listed
-
1 Jan 2021 1 repository listed
-
8 Dec 2020 1 repository listed
-
1 Dec 2020 1 repository listed
-
1 Nov 2020 1 repository listed Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
30 Oct 2020 1 repository listed
-
22 Oct 2020 1 repository listed
-
19 Oct 2020 1 repository listed
-
12 Aug 2020 1 repository listed Syntology official (archive's flag): 14 ran · 14 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 0 honoured, 1 violated, 12 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 16 harvested samples) · 2 pointer-only (licence)
-
1 Aug 2020 1 repository listed
-
9 Jun 2020 1 repository listed Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 2 violated, 9 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 16 harvested samples) · 3 pointer-only (licence)
-
1 Jun 2020 1 repository listed
-
7 May 2020 1 repository listed Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples)
-
1 Apr 2020 1 repository listed
-
8 Mar 2020 1 repository listed Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
24 Feb 2020 1 repository listed
-
24 Feb 2020 1 repository listed Syntology 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
15 Oct 2019 1 repository listed
-
8 Oct 2019 1 repository listed
-
1 Oct 2019 1 repository listed
-
1 Oct 2019 1 repository listed
-
12 Sep 2019 1 repository listed Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples) · 1 pointer-only (licence)
-
14 Aug 2019 1 repository listed
-
25 Jul 2019 1 repository listed
-
1 Jun 2019 1 repository listed
-
1 Jun 2019 1 repository listed
-
30 Apr 2019 1 repository listed
-
11 Apr 2019 1 repository listed
-
10 Apr 2019 1 repository listed
-
9 Apr 2019 1 repository listed
-
14 Mar 2019 1 repository listed
-
1 Sep 2018 1 repository listed
-
4 May 2018 1 repository listed
-
2 May 2018 1 repository listed Syntology 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
30 Apr 2018 1 repository listed
-
5 Apr 2018 1 repository listed
-
4 Apr 2018 1 repository listed
-
17 Aug 2017 1 repository listed
-
16 Aug 2017 1 repository listed
-
3 Jun 2017 1 repository listed
-
1 Dec 2015 1 repository listed
-
An analysis of vision-language models for fabric retrieval7 Jul 2025 0 repositories listed
-
Mask-aware Text-to-Image Retrieval: Referring Expression Segmentation Meets Cross-modal Retrieval28 Jun 2025 0 repositories listed
-
Maximal Matching Matters: Preventing Representation Collapse for Robust Cross-Modal Retrieval26 Jun 2025 0 repositories listed
-
Multimodal Medical Image Binding via Shared Text Embeddings22 Jun 2025 0 repositories listed
-
FedNano: Toward Lightweight Federated Tuning for Pretrained Multimodal Large Language Models12 Jun 2025 0 repositories listed
-
SA-Person: Text-Based Person Retrieval with Scene-aware Re-ranking30 May 2025 0 repositories listed
-
EmotionRankCLAP: Bridging Natural Language Speaking Styles and Ordinal Speech Emotion via Rank-N-Contrast29 May 2025 0 repositories listed
-
FOLIAGE: Towards Physical Intelligence World Models Via Unbounded Surface Evolution29 May 2025 0 repositories listed
-
DocMMIR: A Framework for Document Multi-modal Information Retrieval25 May 2025 0 repositories listed
-
GMM-Based Comprehensive Feature Extraction and Relative Distance Preservation For Few-Shot Cross-Modal Retrieval19 May 2025 0 repositories listed
-
Sat2Sound: A Unified Framework for Zero-Shot Soundscape Mapping19 May 2025 0 repositories listed
-
Towards Cross-modal Retrieval in Chinese Cultural Heritage Documents: Dataset and Solution16 May 2025 0 repositories listed
-
CellCLIP -- Learning Perturbation Effects in Cell Painting via Text-Guided Contrastive Learning16 May 2025 0 repositories listed
-
10 May 2025 0 repositories listed Syntology 4 ran (of which 4 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; every one of the 4 samples that ran constructed an object rather than computing a result (of 5 harvested samples) · 5 pointer-only (licence)
-
Improving Sound Source Localization with Joint Slot Attention on Image and Audio21 Apr 2025 0 repositories listed
-
The 1st EReL@MIR Workshop on Efficient Representation Learning for Multimodal Information Retrieval21 Apr 2025 0 repositories listed
-
SemCORE: A Semantic-Enhanced Generative Cross-Modal Retrieval Framework with MLLMs17 Apr 2025 0 repositories listed
-
PATFinger: Prompt-Adapted Transferable Fingerprinting against Unauthorized Multimodal Dataset Usage15 Apr 2025 0 repositories listed
-
Learning Sparse Disentangled Representations for Multimodal Exclusion Retrieval4 Apr 2025 0 repositories listed
-
FineLIP: Extending CLIP's Reach via Fine-Grained Alignment with Longer Text Inputs2 Apr 2025 0 repositories listed
-
Seeing Speech and Sound: Distinguishing and Locating Audios in Visual Scenes24 Mar 2025 0 repositories listed
-
PromptHash: Affinity-Prompted Collaborative Cross-Modal Learning for Adaptive Hashing Retrieval20 Mar 2025 0 repositories listed
-
Adaptive Inner Speech-Text Alignment for LLM-based Speech Translation13 Mar 2025 0 repositories listed
-
Astrea: A MOE-based Visual Understanding Model with Progressive Alignment12 Mar 2025 0 repositories listed
-
A Recipe for Improving Remote Sensing VLM Zero Shot Generalization10 Mar 2025 0 repositories listed
-
X2CT-CLIP: Enable Multi-Abnormality Detection in Computed Tomography from Chest Radiography via Tri-Modal Contrastive Learning4 Mar 2025 0 repositories listed
-
Lightweight Contrastive Distilled Hashing for Online Cross-modal Retrieval27 Feb 2025 0 repositories listed
-
On the Importance of Text Preprocessing for Multimodal Representation Learning and Pathology Report Generation26 Feb 2025 0 repositories listed
-
Pathology Report Generation and Multimodal Representation Learning for Cutaneous Melanocytic Lesions26 Feb 2025 0 repositories listed
-
17 Feb 2025 0 repositories listed
-
Zero-Shot Interactive Text-to-Image Retrieval via Diffusion-Augmented Representations26 Jan 2025 0 repositories listed
-
TSVC:Tripartite Learning with Semantic Variation Consistency for Robust Image-Text Retrieval19 Jan 2025 0 repositories listed
-
Cross-Modal 3D Representation with Multi-View Images and Point Clouds1 Jan 2025 0 repositories listed
-
Incorporating Dense Knowledge Alignment into Unified Multimodal Representation Models1 Jan 2025 0 repositories listed
-
Seeing Speech and Sound: Distinguishing and Locating Audio Sources in Visual Scenes1 Jan 2025 0 repositories listed
-
Maybe you are looking for CroQS: Cross-modal Query Suggestion for Text-to-Image Retrieval18 Dec 2024 0 repositories listed
-
Rebalanced Vision-Language Retrieval Considering Structure-Aware Distillation14 Dec 2024 0 repositories listed
-
CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance5 Dec 2024 0 repositories listed
-
Explainable and Interpretable Multimodal Large Language Models: A Comprehensive Survey3 Dec 2024 0 repositories listed
-
Fusing Physics-Driven Strategies and Cross-Modal Adversarial Learning: Toward Multi-Domain Applications30 Nov 2024 0 repositories listed
-
FLEX-CLIP: Feature-Level GEneration Network Enhanced CLIP for X-shot Cross-modal Retrieval26 Nov 2024 0 repositories listed
-
CLIPS: An Enhanced CLIP Framework for Learning with Synthetic Captions25 Nov 2024 0 repositories listed
-
FINECAPTION: Compositional Image Captioning Focusing on Wherever You Want at Any Granularity23 Nov 2024 0 repositories listed
-
Improving Factuality of 3D Brain MRI Report Generation with Paired Image-domain Retrieval and Text-domain Augmentation23 Nov 2024 0 repositories listed
-
Everything is a Video: Unifying Modalities through Next-Frame Prediction15 Nov 2024 0 repositories listed
-
4 Nov 2024 0 repositories listed
-
MM-Embed: Universal Multimodal Retrieval with Multimodal LLMs4 Nov 2024 0 repositories listed
-
Test-time Adaptation for Cross-modal Retrieval with Query Shift21 Oct 2024 0 repositories listed
-
GleanVec: Accelerating vector search with minimalist nonlinear dimensionality reduction14 Oct 2024 0 repositories listed
-
13 Oct 2024 0 repositories listed
-
CSA: Data-efficient Mapping of Unimodal Features to Multimodal Features10 Oct 2024 0 repositories listed
-
Snap and Diagnose: An Advanced Multimodal Retrieval System for Identifying Plant Diseases in the Wild27 Aug 2024 0 repositories listed
-
Leveraging Chemistry Foundation Models to Facilitate Structure Focused Retrieval Augmented Generation in Multi-Agent Workflows for Catalyst and Materials Design21 Aug 2024 0 repositories listed
-
Limitations in Employing Natural Language Supervision for Sensor-Based Human Activity Recognition -- And Ways to Overcome Them21 Aug 2024 0 repositories listed
-
Bridging Information Asymmetry in Text-video Retrieval: A Data-centric Approach14 Aug 2024 0 repositories listed
-
Contrastive masked auto-encoders based self-supervised hashing for 2D image and 3D point cloud cross-modal retrieval11 Aug 2024 0 repositories listed
Syntology lines on 11 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.