Browse State-of-the-Art › Image-text Retrieval › Papers, page 3
Image-text Retrieval
Papers archive 2025-07-28
archive papers tagged: 248 · with a code link: 131 · where Syntology ran a sample: 49 (44 with a run with no instrument failure, 5 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (49 of 248 tagged: 44 with a run with no instrument failure, 5 where every run was a failure of Syntology's instrument)
Page 3 of 3: papers 201 to 248 of 248, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
VL-Match: Enhancing Vision-Language Pretraining with Token-Level and Instance-Level Matching1 Jan 2023 0 repositories listed
-
Efficient Image Captioning for Edge Devices18 Dec 2022 0 repositories listed
-
HGAN: Hierarchical Graph Alignment Network for Image-Text Retrieval16 Dec 2022 0 repositories listed
-
NLIP: Noise-robust Language-Image Pre-training14 Dec 2022 0 repositories listed
-
Scale-Semantic Joint Decoupling Network for Image-text Retrieval in Remote Sensing12 Dec 2022 0 repositories listed
-
2 Dec 2022 0 repositories listed
-
Generative Negative Text Replay for Continual Vision-Language Pretraining31 Oct 2022 0 repositories listed
-
Image-Text Retrieval with Binary and Continuous Label Supervision20 Oct 2022 0 repositories listed
-
CPL: Counterfactual Prompt Learning for Vision and Language Models19 Oct 2022 0 repositories listed
-
MAMO: Masked Multimodal Modeling for Fine-Grained Vision-Language Representation Learning9 Oct 2022 0 repositories listed
-
Learning to embed semantic similarity for joint image-text retrieval7 Oct 2022 0 repositories listed
-
Efficient Multilingual Multi-modal Pre-training through Triple Contrastive Loss1 Oct 2022 0 repositories listed
-
29 Sep 2022 0 repositories listed
-
Revising Image-Text Retrieval via Multi-Modal Entailment22 Aug 2022 0 repositories listed
-
CODER: Coupled Diversity-Sensitive Momentum Contrastive Learning for Image-Text Retrieval21 Aug 2022 0 repositories listed
-
VLMAE: Vision-Language Masked Autoencoder19 Aug 2022 0 repositories listed
-
Dynamic Contrastive Distillation for Image-Text Retrieval4 Jul 2022 0 repositories listed
-
VL-BEiT: Generative Vision-Language Pretraining2 Jun 2022 0 repositories listed
-
Prompt-based Learning for Unpaired Image Captioning26 May 2022 0 repositories listed
-
25 May 2022 0 repositories listed
-
HiVLP: Hierarchical Vision-Language Pre-Training for Fast Image-Text Retrieval24 May 2022 0 repositories listed
-
Progressive Learning for Image Retrieval with Hybrid-Modality Queries24 Apr 2022 0 repositories listed
-
15 Apr 2022 0 repositories listed
-
10 Apr 2022 0 repositories listed
-
Image-text Retrieval: A Survey on Recent Research and Development28 Mar 2022 0 repositories listed
-
LoopITR: Combining Dual and Cross Encoder Architectures for Image-Text Retrieval10 Mar 2022 0 repositories listed
-
CommerceMM: Large-Scale Commerce MultiModal Representation Learning with Omni Retrieval15 Feb 2022 0 repositories listed
-
Negative Sample is Negative in Its Own Way: Tailoring Negative Sentences for Image-Text Retrieval17 Dec 2021 0 repositories listed
-
Unified Multimodal Pre-training and Prompt-based Tuning for Vision-Language Understanding and Generation10 Dec 2021 0 repositories listed
-
UFO: A UniFied TransfOrmer for Vision-Language Representation Learning19 Nov 2021 0 repositories listed
-
Constructing Phrase-level Semantic Labels to Form Multi-GrainedSupervision for Image-Text Retrieval16 Nov 2021 0 repositories listed
-
SwAMP: Swapped Assignment of Multi-Modal Pairs for Cross-Modal Retrieval10 Nov 2021 0 repositories listed
-
Constructing Phrase-level Semantic Labels to Form Multi-Grained Supervision for Image-Text Retrieval12 Sep 2021 0 repositories listed
-
Probing Inter-modality: Visual Parsing with Self-Attention for Vision-Language Pre-training25 Jun 2021 0 repositories listed
-
Survey of Visual-Semantic Embedding Methods for Zero-Shot Image Retrieval16 May 2021 0 repositories listed
-
Playing Lottery Tickets with Vision and Language23 Apr 2021 0 repositories listed
-
Continual learning in cross-modal retrieval14 Apr 2021 0 repositories listed
-
UC2: Universal Cross-lingual Cross-modal Vision-and-Language Pre-training1 Apr 2021 0 repositories listed
-
Learning Multi-Modal Nonlinear Embeddings: Performance Bounds and an Algorithm3 Jun 2020 0 repositories listed
-
Context-Aware Attention Network for Image-Text Retrieval1 Jun 2020 0 repositories listed
-
3 Mar 2020 0 repositories listed
-
UNITER: Learning UNiversal Image-TExt Representations25 Sep 2019 0 repositories listed
-
16 Aug 2019 0 repositories listed
-
Deep Semantic Multimodal Hashing Network for Scalable Image-Text and Video-Text Retrievals9 Jan 2019 0 repositories listed
-
Webly Supervised Joint Embedding for Cross-Modal lmage-Text Retrieval1 Oct 2018 0 repositories listed
-
Webly Supervised Joint Embedding for Cross-Modal Image-Text Retrieval23 Aug 2018 0 repositories listed
-
Look, Imagine and Match: Improving Textual-Visual Cross-Modal Retrieval with Generative Models17 Nov 2017 0 repositories listed
-
Asymmetrically Weighted CCA And Hierarchical Kernel Sentence Embedding For Image & Text Retrieval19 Nov 2015 0 repositories listed