Browse State-of-the-Art › Text Matching
Text Matching
161 papers with code · 0 benchmarks · 7 datasets archive 2025-07-28
Matching a target text to a source text based on their meaning.
Description from the archive archive 2025-07-28.
Benchmarks archive 2025-07-28
No benchmark for this task in the archive.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
7 datasets whose archive record lists this task, ordered by the archive's paper count.
Subtasks archive 2025-07-28
No subtask under this task in the archive's task tree.
Most implemented papers archive 2025-07-28
30 shown of 161 papers with code (364 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
28 Nov 2017 20 repositories listed Syntology ran 3 of 3 samples · 0 unverifiedIn this paper, we propose an Attentional Generative Adversarial Network (AttnGAN) that allows attention-driven, multi-stage refinement for fine-grained text-to-image generation.
-
10 Nov 2019 7 repositories listed Syntology ran 4 of 6 samples · 2 unverified · 4 pointer-only (licence)We introduce an approach for open-domain question answering (QA) that retrieves and reads a passage graph, where vertices are passages of text and edges represent relationships that are derived from an external…
-
25 Sep 2019 7 repositories listed Syntology ran 3 of 3 samples · 0 unverified · 2 pointer-only (licence)Different from previous work that applies joint random masking to both modalities, we use conditional masking on pre-training tasks (i.
-
20 Feb 2016 7 repositories listedAn effective way is to extract meaningful matching patterns from words, phrases, and sentences to produce the matching score.
-
21 Mar 2018 6 repositories listed Syntology ran 7 of 16 samples · 9 unverified · 1 pointer-only (licence)Prior work either simply aggregates the similarity of all possible pairs of regions and words without attending differentially to more and less important words or regions, or uses a multi-step attentional process to…
-
27 Jun 2024 5 repositories listed Syntology ran 2 of 4 samples · 2 unverifiedDocuments are visually rich structures that convey information through text, but also figures, page layouts, tables, or even fonts.
-
1 Mar 2020 4 repositories listed Syntology ran 0 of 7 samples · 7 unverifiedTo improve fine-grained video-text retrieval, we propose a Hierarchical Graph Reasoning (HGR) model, which decomposes video-text matching into global-to-local levels.
-
6 May 2023 3 repositories listed Syntology ran 0 of 9 samples · 9 unverifiedIn this paper, we present an end-to-end framework Structure-CLIP, which integrates Scene Graph Knowledge (SGK) to enhance multi-modal structured representations.
-
13 Aug 2020 3 repositories listedTo these ends, we propose a simpler but more effective Deep Fusion Generative Adversarial Networks (DF-GAN).
-
1 Aug 2019 3 repositories listed Syntology ran 1 of 1 samples · 0 unverifiedIn this paper, we present a fast and strong neural approach for general purpose text matching applications.
-
24 Jul 2023 2 repositories listedThis paper aims to provide a comprehensive survey of cutting-edge research in prompt engineering on three types of vision-language models: multimodal-to-text generation models (e.
-
24 Nov 2022 2 repositories listed Syntology ran 6 of 10 samples · 4 unverifiedMedical image visual question answering (VQA) is a task to answer clinical questions, given a radiographic image, which is a challenging problem that requires a model to integrate both vision and language information.
-
21 Oct 2022 2 repositories listedIn the event that the gradients are not integrable to a valid loss function, we implement our proposed objectives such that they would directly operate in the gradient space instead of on the losses in the embedding…
-
7 Apr 2022 2 repositories listed Syntology ran 1 of 3 samples · 2 unverified · 3 pointer-only (licence)Image-Text matching (ITM) is a common task for evaluating the quality of Vision and Language (VL) models.
-
17 Sep 2021 2 repositories listed Syntology ran 5 of 9 samples · 4 unverifiedMoreover, to handle the deficiency of label texts and make use of tremendous web data, we propose a new paradigm based on this multimodal learning framework for action recognition, which we dub "pre-train, prompt and…
-
22 Mar 2021 2 repositories listedEmploying paraphrasing tools to conceal plagiarized text is a severe threat to academic integrity.
-
7 Oct 2020 2 repositories listed Syntology ran 3 of 11 samples · 8 unverifiedFurthermore, we introduce a new polynomial loss under the universal weighting framework, which defines a weight function for the positive and negative informative pairs respectively.
-
19 Apr 2020 2 repositories listedThis paper creates a paradigm shift with regard to the way we build neural extractive summarization systems.
-
6 Sep 2019 2 repositories listed Syntology ran 1 of 1 samples · 0 unverified · 1 pointer-only (licence)It outperforms the current best method by 6.
-
12 Aug 2019 2 repositories listedWe propose a novel framework that achieves remarkable matching performance with acceptable model complexity.
-
2 Nov 2016 2 repositories listedWe propose Dual Attention Networks (DANs) which jointly leverage visual and textual attention mechanisms to capture fine-grained interplay between vision and language.
-
10 Jun 2025 1 repository listedALTA achieves superior performance in vision-language matching tasks like retrieval and zero-shot classification by adapting the pretrained vision model from masked record modeling.
-
19 Mar 2025 1 repository listedEnabling Visual Semantic Models to effectively handle multi-view description matching has been a longstanding challenge.
-
5 Mar 2025 1 repository listed Syntology ran 2 of 4 samples · 2 unverified · 4 pointer-only (licence)Our paradigm is simple and training-free, providing the first method to defend CLIP from adversarial attacks at test time, which is orthogonal to existing methods aiming to boost zero-shot adversarial robustness of CLIP.
-
2 Mar 2025 1 repository listedTo address these challenges, we propose IteRPrimE (Iterative Grad-CAM Refinement and Primary word Emphasis), which leverages a saliency heatmap through Grad-CAM from a Vision-Language Pre-trained (VLP) model for…
-
27 Feb 2025 1 repository listed Syntology ran 5 of 5 samples · 0 unverified · 5 pointer-only (licence)Contrastive Language-Image Pre-training (CLIP) models excel in zero-shot classification, yet face challenges in complex multi-object scenarios.
-
17 Jan 2025 1 repository listedHowever, their handcrafted generic descriptions fail to capture the diverse range of anomalies that may emerge in different objects, and simple patch-level image-text matching often struggles to localize anomalous…
-
9 Dec 2024 1 repository listed Syntology ran 0 of 1 samples · 1 unverifiedIn the peer review process of top-tier machine learning (ML) and artificial intelligence (AI) conferences, reviewers are assigned to papers through automated methods.
-
16 Nov 2024 1 repository listedIn zero-shot skeleton-based action recognition, aligning skeleton features with the text features of action labels is essential for accurately predicting unseen actions.
-
10 Oct 2024 1 repository listed Syntology ran 5 of 7 samples · 2 unverified · 7 pointer-only (licence)To this end, in this paper, we propose DISCO, a hierarchical Disentanglement based Cognitive diagnosis framework, aimed at flexibly accommodating the underlying representation learning model for effective and…
Syntology lines on 17 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections