Browse State-of-the-Art › Image-text Retrieval › Papers, page 2
Image-text Retrieval
Papers archive 2025-07-28
archive papers tagged: 248 · with a code link: 131 · where Syntology ran a sample: 49 (44 with a run with no instrument failure, 5 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (49 of 248 tagged: 44 with a run with no instrument failure, 5 where every run was a failure of Syntology's instrument)
Page 2 of 3: papers 101 to 200 of 248, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
11 Oct 2022 1 repository listed
-
28 Sep 2022 1 repository listed
-
12 Sep 2022 1 repository listed
-
8 Sep 2022 1 repository listed
-
29 Aug 2022 1 repository listed Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
11 Jul 2022 1 repository listed
-
16 Jun 2022 1 repository listed Syntology official: harvested, nothing ran · 0 ran · 2 unverified (of 2 harvested samples)
-
15 Jun 2022 1 repository listed Syntology official: harvested, nothing ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 2 pointer-only (licence)
-
1 Jun 2022 1 repository listed
-
1 Jun 2022 1 repository listed Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 3 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
8 May 2022 1 repository listed Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples)
-
27 Mar 2022 1 repository listed
-
8 Mar 2022 1 repository listed
-
26 Feb 2022 1 repository listed
-
21 Feb 2022 1 repository listed Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
9 Nov 2021 1 repository listed
-
5 Nov 2021 1 repository listed
-
22 Jul 2021 1 repository listed Syntology official (archive's flag): 3 ran · 3 ran (of which 3 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified; every one of the 3 samples that ran constructed an object rather than computing a result (of 9 harvested samples) · 9 pointer-only (licence)
-
11 Jul 2021 1 repository listed
-
19 Jun 2021 1 repository listed
-
4 Jun 2021 1 repository listed
-
28 May 2021 1 repository listed Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
9 Nov 2020 1 repository listed
-
3 Sep 2020 1 repository listed
-
26 Jun 2020 1 repository listed Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 1 honoured, 2 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 10 harvested samples) · 3 pointer-only (licence)
-
2 Apr 2020 1 repository listed
-
8 Mar 2020 1 repository listed Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
11 Jan 2020 1 repository listed
-
11 Oct 2019 1 repository listed Syntology 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
11 Jun 2018 1 repository listed
-
Maximal Matching Matters: Preventing Representation Collapse for Robust Cross-Modal Retrieval26 Jun 2025 0 repositories listed
-
Distill CLIP (DCLIP): Enhancing Image-Text Retrieval via Cross-Modal Transformer Distillation25 May 2025 0 repositories listed
-
EvdCLIP: Improving Vision-Language Retrieval with Entity Visual Descriptions from Large Language Models24 May 2025 0 repositories listed
-
Representation Discrepancy Bridging Method for Remote Sensing Image-Text Retrieval22 May 2025 0 repositories listed
-
Breaking Language Barriers or Reinforcing Bias? A Study of Gender and Racial Disparities in Multilingual Contrastive Vision Language Models20 May 2025 0 repositories listed
-
AGATE: Stealthy Black-box Watermarking for Multimodal Model Copyright Protection28 Apr 2025 0 repositories listed
-
Breaking the Modality Barrier: Universal Embedding Learning with Multimodal LLMs24 Apr 2025 0 repositories listed
-
FocalLens: Instruction Tuning Enables Zero-Shot Conditional Image Representations11 Apr 2025 0 repositories listed
-
SeLIP: Similarity Enhanced Contrastive Language Image Pretraining for Multi-modal Head MRI25 Mar 2025 0 repositories listed
-
Anatomy-Aware Conditional Image-Text Retrieval10 Mar 2025 0 repositories listed
-
Variance-Aware Loss Scheduling for Multimodal Alignment in Low-Data Settings5 Mar 2025 0 repositories listed
-
LLaVE: Large Language and Vision Embedding Models with Hardness-Weighted Contrastive Learning4 Mar 2025 0 repositories listed
-
MedUnifier: Unifying Vision-and-Language Pre-training on Medical Data with Vision Generation Task using Discrete Visual Representations2 Mar 2025 0 repositories listed
-
Progressive Local Alignment for Medical Multimodal Pre-training25 Feb 2025 0 repositories listed
-
Fine-tuning Multimodal Transformers on Edge: A Parallel Split Learning Approach10 Feb 2025 0 repositories listed
-
DCFormer: Efficient 3D Vision-Language Modeling with Decomposed Convolutions7 Feb 2025 0 repositories listed
-
MASS: Overcoming Language Bias in Image-Text Matching20 Jan 2025 0 repositories listed
-
TSVC:Tripartite Learning with Semantic Variation Consistency for Robust Image-Text Retrieval19 Jan 2025 0 repositories listed
-
Advancing Myopia To Holism: Fully Contrastive Language-Image Pre-training1 Jan 2025 0 repositories listed
-
Barking Up The Syntactic Tree: Enhancing VLM Training with Syntactic Losses11 Dec 2024 0 repositories listed
-
Explaining and Mitigating the Modality Gap in Contrastive Multimodal Learning10 Dec 2024 0 repositories listed
-
VladVA: Discriminative Fine-tuning of LVLMs5 Dec 2024 0 repositories listed
-
Approximate Fiber Product: A Preliminary Algebraic-Geometric Perspective on Multimodal Embedding Alignment30 Nov 2024 0 repositories listed
-
Knowledge Transfer Across Modalities with Natural Language Supervision23 Nov 2024 0 repositories listed
-
Uni-Mlip: Unified Self-supervision for Medical Vision Language Pre-training20 Nov 2024 0 repositories listed
-
CtrlSynth: Controllable Image Text Synthesis for Data-Efficient Multimodal Learning15 Oct 2024 0 repositories listed
-
AnyAttack: Towards Large-scale Self-supervised Adversarial Attacks on Vision-language Models7 Oct 2024 0 repositories listed
-
NEVLP: Noise-Robust Framework for Efficient Vision-Language Pre-training15 Sep 2024 0 repositories listed
-
Pushing the Limits of Vision-Language Models in Remote Sensing without Human Annotations11 Sep 2024 0 repositories listed
-
Toward Automatic Relevance Judgment using Vision--Language Models for Image--Text Retrieval Evaluation2 Aug 2024 0 repositories listed
-
Assessing Brittleness of Image-Text Retrieval Benchmarks from Vision-Language Models Perspective21 Jul 2024 0 repositories listed
-
CosmoCLIP: Generalizing Large Vision-Language Models for Astronomical Imaging10 Jul 2024 0 repositories listed
-
How to Make Cross Encoder a Good Teacher for Efficient Image-Text Retrieval?10 Jul 2024 0 repositories listed
-
Beat: Bi-directional One-to-Many Embedding Alignment for Text-based Person Retrieval9 Jun 2024 0 repositories listed
-
Knowledge-grounded Adaptation Strategy for Vision-language Models: Building Unique Case-set for Screening Mammograms for Residents Training30 May 2024 0 repositories listed
-
Multimodal Adversarial Defense for Vision-Language Models by Leveraging One-To-Many Relationships29 May 2024 0 repositories listed
-
Active Learning for Finely-Categorized Image-Text Retrieval by Selecting Hard Negative Unpaired Samples25 May 2024 0 repositories listed
-
14 May 2024 0 repositories listed
-
UrbanCross: Enhancing Satellite Image-Text Retrieval with Cross-Domain Adaptation22 Apr 2024 0 repositories listed
-
Self-Training Large Language Models for Improved Visual Program Synthesis With Visual Reinforcement6 Apr 2024 0 repositories listed
-
Improving Adversarial Transferability of Vision-Language Pre-training Models through Collaborative Multimodal Interaction16 Mar 2024 0 repositories listed
-
LuoJiaHOG: A Hierarchy Oriented Geo-aware Image Caption Dataset for Remote Sensing Image-Text Retrival16 Mar 2024 0 repositories listed
-
Enhancing Conceptual Understanding in Multimodal Contrastive Learning through Hard Negative Samples5 Mar 2024 0 repositories listed
-
4 Jan 2024 0 repositories listed
-
Filter & Align: Leveraging Human Knowledge to Curate Image-Text Data11 Dec 2023 0 repositories listed
-
LightCLIP: Learning Multi-Level Interaction for Lightweight Vision-Language Models1 Dec 2023 0 repositories listed
-
IG Captioner: Information Gain Captioners are Strong Zero-shot Classifiers27 Nov 2023 0 repositories listed
-
A New Fine-grained Alignment Method for Image-text Matching3 Nov 2023 0 repositories listed
-
MCAD: Multi-teacher Cross-modal Alignment Distillation for efficient image-text retrieval30 Oct 2023 0 repositories listed
-
12 Oct 2023 0 repositories listed
-
Ziya-Visual: Bilingual Large Vision-Language Model via Multi-Task Instruction Tuning12 Oct 2023 0 repositories listed
-
Constructing Image-Text Pair Dataset from Books3 Oct 2023 0 repositories listed
-
Dual Relation Alignment for Composed Image Retrieval5 Sep 2023 0 repositories listed
-
2 Sep 2023 0 repositories listed
-
DLIP: Distilling Language-Image Pre-training24 Aug 2023 0 repositories listed
-
EVE: Efficient Vision-Language Pre-training with Masked Prediction and Modality-Aware MoE23 Aug 2023 0 repositories listed
-
Free-ATM: Exploring Unsupervised Learning on Diffusion-Generated Images with Free Attention Masks13 Aug 2023 0 repositories listed
-
Distilling Knowledge from Text-to-Image Generative Models Improves Visio-Linguistic Reasoning in CLIP18 Jul 2023 0 repositories listed
-
Switch-BERT: Learning to Model Multimodal Interactions by Switching Attention and Input25 Jun 2023 0 repositories listed
-
Hypernymization of named entity-rich captions for grounding-based multi-modal pretraining25 Apr 2023 0 repositories listed
-
RECLIP: Resource-efficient CLIP by Training with Small Images12 Apr 2023 0 repositories listed
-
Exposing and Mitigating Spurious Correlations for Cross-Modal Retrieval6 Apr 2023 0 repositories listed
-
Scene Graph Based Fusion Network For Image-Text Retrieval20 Mar 2023 0 repositories listed
-
Efficient Image-Text Retrieval via Keyword-Guided Pre-Screening14 Mar 2023 0 repositories listed
-
Understanding and Constructing Latent Modality Structures in Multi-modal Representation Learning10 Mar 2023 0 repositories listed
-
The style transformer with common knowledge optimization for image-text retrieval1 Mar 2023 0 repositories listed
-
GAFNet: A Global Fourier Self Attention Based Novel Network for multi-modal downstream tasks1 Jan 2023 0 repositories listed
-
Multilateral Semantic Relations Modeling for Image Text Retrieval1 Jan 2023 0 repositories listed
-
1 Jan 2023 0 repositories listed
Syntology lines on 12 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.