Browse State-of-the-Art › Text Retrieval › Papers, page 5
Text Retrieval
Papers archive 2025-07-28
archive papers tagged: 671 · with a code link: 335 · where Syntology ran a sample: 117 (99 with a run with no instrument failure, 18 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (117 of 671 tagged: 99 with a run with no instrument failure, 18 where every run was a failure of Syntology's instrument)
Page 5 of 7: papers 401 to 500 of 671, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Robotic State Recognition with Image-to-Text Retrieval Task of Pre-Trained Vision-Language Model and Black-Box Optimization30 Oct 2024 0 repositories listed
-
Do Audio-Language Models Understand Linguistic Variations?21 Oct 2024 0 repositories listed
-
Improving General Text Embedding Model: Tackling Task Conflict and Data Imbalance through Model Merging19 Oct 2024 0 repositories listed
-
Beyond Coarse-Grained Matching in Video-Text Retrieval16 Oct 2024 0 repositories listed
-
CtrlSynth: Controllable Image Text Synthesis for Data-Efficient Multimodal Learning15 Oct 2024 0 repositories listed
-
LaMP: Language-Motion Pretraining for Motion Generation, Retrieval, and Captioning9 Oct 2024 0 repositories listed
-
AnyAttack: Towards Large-scale Self-supervised Adversarial Attacks on Vision-language Models7 Oct 2024 0 repositories listed
-
CoLLAP: Contrastive Long-form Language-Audio Pretraining with Musical Temporal Structure Augmentation3 Oct 2024 0 repositories listed
-
Robotic Environmental State Recognition with Pre-Trained Vision-Language Models and Black-Box Optimization26 Sep 2024 0 repositories listed
-
DiffATR: Diffusion-based Generative Modeling for Audio-Text Retrieval16 Sep 2024 0 repositories listed
-
NEVLP: Noise-Robust Framework for Efficient Vision-Language Pre-training15 Sep 2024 0 repositories listed
-
Enhancing Q&A Text Retrieval with Ranking Models: Benchmarking, fine-tuning and deploying Rerankers for RAG12 Sep 2024 0 repositories listed
-
Pushing the Limits of Vision-Language Models in Remote Sensing without Human Annotations11 Sep 2024 0 repositories listed
-
Benchmarking and Building Zero-Shot Hindi Retrieval Model with Hindi-BEIR and NLLB-E59 Sep 2024 0 repositories listed
-
Improving embedding with contrastive fine-tuning on small datasets with expert-augmented scores19 Aug 2024 0 repositories listed
-
NAVERO: Unlocking Fine-Grained Semantics for Video-Language Compositionality18 Aug 2024 0 repositories listed
-
Mamba Retriever: Utilizing Mamba for Effective and Efficient Dense Retrieval15 Aug 2024 0 repositories listed
-
Pairing Clustered Inverted Indexes with kNN Graphs for Fast Approximate Retrieval over Learned Sparse Representations8 Aug 2024 0 repositories listed
-
Toward Automatic Relevance Judgment using Vision--Language Models for Image--Text Retrieval Evaluation2 Aug 2024 0 repositories listed
-
29 Jul 2024 0 repositories listed Syntology 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Assessing Brittleness of Image-Text Retrieval Benchmarks from Vision-Language Models Perspective21 Jul 2024 0 repositories listed
-
Multimodal Misinformation Detection using Large Vision-Language Models19 Jul 2024 0 repositories listed
-
CosmoCLIP: Generalizing Large Vision-Language Models for Astronomical Imaging10 Jul 2024 0 repositories listed
-
EA-VTR: Event-Aware Video-Text Retrieval10 Jul 2024 0 repositories listed
-
How to Make Cross Encoder a Good Teacher for Efficient Image-Text Retrieval?10 Jul 2024 0 repositories listed
-
CEIA: CLIP-Based Event-Image Alignment for Open-World Event-Based Understanding9 Jul 2024 0 repositories listed
-
Memory³: Language Modeling with Explicit Memory1 Jul 2024 0 repositories listed
-
PathAlign: A vision-language model for whole slide images in histopathology27 Jun 2024 0 repositories listed
-
ACE: A Generative Cross-Modal Retrieval Framework with Coarse-To-Fine Semantic Modeling25 Jun 2024 0 repositories listed
-
Evaluating D-MERIT of Partial-annotation on Information Retrieval23 Jun 2024 0 repositories listed
-
Multi-Scale Temporal Difference Transformer for Video-Text Retrieval23 Jun 2024 0 repositories listed
-
RE-AdaptIR: Improving Information Retrieval through Reverse Engineered Adaptation20 Jun 2024 0 repositories listed
-
Unifying Multimodal Retrieval via Document Screenshot Embedding17 Jun 2024 0 repositories listed
-
Enhancing Knowledge Retrieval with In-Context Learning and Semantic Search through Generative AI13 Jun 2024 0 repositories listed
-
Beat: Bi-directional One-to-Many Embedding Alignment for Text-based Person Retrieval9 Jun 2024 0 repositories listed
-
HENASY: Learning to Assemble Scene-Entities for Egocentric Video-Language Model1 Jun 2024 0 repositories listed
-
Jina CLIP: Your CLIP Model Is Also Your Text Retriever30 May 2024 0 repositories listed
-
Knowledge-grounded Adaptation Strategy for Vision-language Models: Building Unique Case-set for Screening Mammograms for Residents Training30 May 2024 0 repositories listed
-
Uncertainty-aware sign language video retrieval with probability distribution modeling30 May 2024 0 repositories listed
-
Multimodal Adversarial Defense for Vision-Language Models by Leveraging One-To-Many Relationships29 May 2024 0 repositories listed
-
Multilingual Diversity Improves Vision-Language Representations27 May 2024 0 repositories listed
-
Understanding the Effect of using Semantically Meaningful Tokens for Visual Representation Learning26 May 2024 0 repositories listed
-
Active Learning for Finely-Categorized Image-Text Retrieval by Selecting Hard Negative Unpaired Samples25 May 2024 0 repositories listed
-
An Empirical Study of Excitation and Aggregation Design Adaptions in CLIP4Clip for Video-Text Retrieval25 May 2024 0 repositories listed
-
14 May 2024 0 repositories listed
-
RETTA: Retrieval-Enhanced Test-Time Adaptation for Zero-Shot Video Captioning11 May 2024 0 repositories listed
-
Distance Sampling-based Paraphraser Leveraging ChatGPT for Text Data Manipulation1 May 2024 0 repositories listed
-
VISLA Benchmark: Evaluating Embedding Sensitivity to Semantic and Lexical Alterations25 Apr 2024 0 repositories listed
-
UrbanCross: Enhancing Satellite Image-Text Retrieval with Cross-Domain Adaptation22 Apr 2024 0 repositories listed
-
MindTuner: Cross-Subject Visual Decoding with Visual Fingerprint and Semantic Correction19 Apr 2024 0 repositories listed
-
FecTek: Enhancing Term Weight in Lexicon-Based Retrieval with Feature Context and Term-level Knowledge18 Apr 2024 0 repositories listed
-
TEXT2TASTE: A Versatile Egocentric Vision System for Intelligent Reading Assistance Using Large Language Model14 Apr 2024 0 repositories listed
-
13 Apr 2024 0 repositories listed
-
HaVTR: Improving Video-Text Retrieval Through Augmentation Using Large Foundation Models7 Apr 2024 0 repositories listed
-
Self-Training Large Language Models for Improved Visual Program Synthesis With Visual Reinforcement6 Apr 2024 0 repositories listed
-
Improving Retrieval for RAG based Question Answering Models on Financial Documents23 Mar 2024 0 repositories listed
-
Improving Adversarial Transferability of Vision-Language Pre-training Models through Collaborative Multimodal Interaction16 Mar 2024 0 repositories listed
-
LuoJiaHOG: A Hierarchy Oriented Geo-aware Image Caption Dataset for Remote Sensing Image-Text Retrival16 Mar 2024 0 repositories listed
-
Refining Knowledge Transfer on Audio-Image Temporal Agreement for Audio-Text Cross Retrieval16 Mar 2024 0 repositories listed
-
Multiscale Matching Driven by Cross-Modal Similarity Consistency for Audio-Text Retrieval15 Mar 2024 0 repositories listed
-
CLIP the Bias: How Useful is Balancing Data in Multimodal Learning?7 Mar 2024 0 repositories listed
-
Multi-Task Contrastive Learning for 8192-Token Bilingual Text Embeddings26 Feb 2024 0 repositories listed
-
Unifying Latent and Lexicon Representations for Effective Video-Text Retrieval26 Feb 2024 0 repositories listed
-
MORE: Multi-mOdal REtrieval Augmented Generative Commonsense Reasoning21 Feb 2024 0 repositories listed
-
PIRB: A Comprehensive Benchmark of Polish Dense and Hybrid Text Retrieval Methods20 Feb 2024 0 repositories listed
-
Multimodal Learned Sparse Retrieval for Image Suggestion12 Feb 2024 0 repositories listed
-
Video Editing for Video Retrieval4 Feb 2024 0 repositories listed
-
4 Jan 2024 0 repositories listed
-
BEV-TSR: Text-Scene Retrieval in BEV Space for Autonomous Driving2 Jan 2024 0 repositories listed
-
Accept the Modality Gap: An Exploration in the Hyperbolic Space1 Jan 2024 0 repositories listed
-
Building Vision-Language Models on Solid Foundations with Masked Distillation1 Jan 2024 0 repositories listed
-
OTE: Exploring Accurate Scene Text Recognition Using One Token1 Jan 2024 0 repositories listed
-
Filter & Align: Leveraging Human Knowledge to Curate Image-Text Data11 Dec 2023 0 repositories listed
-
Leveraging Generative Language Models for Weakly Supervised Sentence Component Analysis in Video-Language Joint Learning10 Dec 2023 0 repositories listed
-
LightCLIP: Learning Multi-Level Interaction for Lightweight Vision-Language Models1 Dec 2023 0 repositories listed
-
IG Captioner: Information Gain Captioners are Strong Zero-shot Classifiers27 Nov 2023 0 repositories listed
-
Noisy Pair Corrector for Dense Retrieval7 Nov 2023 0 repositories listed
-
A New Fine-grained Alignment Method for Image-text Matching3 Nov 2023 0 repositories listed
-
FLAP: Fast Language-Audio Pre-training2 Nov 2023 0 repositories listed
-
MCAD: Multi-teacher Cross-modal Alignment Distillation for efficient image-text retrieval30 Oct 2023 0 repositories listed
-
End-to-End Autoregressive Retrieval via Bootstrapping for Smart Reply Systems29 Oct 2023 0 repositories listed
-
20 Oct 2023 0 repositories listed
-
12 Oct 2023 0 repositories listed
-
Ziya-Visual: Bilingual Large Vision-Language Model via Multi-Task Instruction Tuning12 Oct 2023 0 repositories listed
-
Policy-Gradient Training of Language Models for Ranking6 Oct 2023 0 repositories listed
-
Constructing Image-Text Pair Dataset from Books3 Oct 2023 0 repositories listed
-
Uncertainty-Aware Alignment Network for Cross-Domain Video-Text Retrieval21 Sep 2023 0 repositories listed
-
Uncertainty-Aware Alignment Network for Cross-Domain Video-Text Retrieval21 Sep 2023 0 repositories listed
-
Dynamic Visual Semantic Sub-Embeddings and Fast Re-Ranking15 Sep 2023 0 repositories listed
-
Dual Relation Alignment for Composed Image Retrieval5 Sep 2023 0 repositories listed
-
2 Sep 2023 0 repositories listed
-
Killing two birds with one stone: Can an audio captioning system also be used for audio-text retrieval?29 Aug 2023 0 repositories listed
-
DLIP: Distilling Language-Image Pre-training24 Aug 2023 0 repositories listed
-
EVE: Efficient Vision-Language Pre-training with Masked Prediction and Modality-Aware MoE23 Aug 2023 0 repositories listed
-
Hybrid Retrieval and Multi-stage Text Ranking Solution at TREC 2022 Deep Learning Track23 Aug 2023 0 repositories listed
-
Free-ATM: Exploring Unsupervised Learning on Diffusion-Generated Images with Free Attention Masks13 Aug 2023 0 repositories listed
-
Embedding-based Retrieval with LLM for Effective Agriculture Information Extracting from Unstructured Data6 Aug 2023 0 repositories listed
-
Defense of Adversarial Ranking Attack in Text Retrieval: Benchmark and Baseline via Detection31 Jul 2023 0 repositories listed
-
Towards a Visual-Language Foundation Model for Computational Pathology24 Jul 2023 0 repositories listed
-
Extracting Molecular Properties from Natural Language with Multimodal Contrastive Learning22 Jul 2023 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.