Browse State-of-the-Art › Retrieval
Retrieval
5,274 papers with code · 11 benchmarks · 38 datasets archive 2025-07-28
A methodology that involves selecting relevant data or examples from a large dataset to support tasks like prediction, learning, or inference. It enhances models by providing context or additional information, often used in systems like retrieval-augmented generation or in-context learning.
Description from the archive archive 2025-07-28.
Benchmarks archive 2025-07-28
11 leaderboard tables shown for this task, 11 with rows (a “benchmark” on this site is a table with at least one row, as on /sota), ordered by row count. “Best model” is the first row in the archive's own order at snapshot; nothing is re-ranked here and metric direction is not recorded in the archive. PwC's Trend sparklines are not in the archive, so that column is omitted. 10 shown of 11 until expanded.
Syntology column: samples harvested from the paper's repositories and executed on synthesized fixtures; “ran” is not a correctness claim and does not order the table. A dash means no Syntology record for that paper, not a recorded non-run. Read from the graph 2026-09-24.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
38 datasets whose archive record lists this task, ordered by the archive's paper count. 30 shown of 38 until expanded.
Subtasks archive 2025-07-28
3 subtasks in the archive's task tree.
Most implemented papers archive 2025-07-28
30 shown of 5,274 papers with code (14,297 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
Beyond Part Models: Person Retrieval with Refined Part Pooling (and a Strong Convolutional Baseline)26 Nov 2017 29 repositories listedRPP re-assigns these outliers to the parts they are closest to, resulting in refined parts with enhanced within-part consistency.
-
17 Mar 2017 27 repositories listed Syntology ran 10 of 32 samples · 22 unverified · 15 pointer-only (licence)We demonstrate the effectiveness of R-GCNs as a stand-alone model for entity classification.
-
23 Jan 2019 23 repositories listed Syntology ran 8 of 26 samples · 18 unverified · 7 pointer-only (licence)We introduce a new approach to generative data-driven dialogue systems (e.
-
10 Apr 2020 19 repositories listed Syntology ran 10 of 14 samples · 4 unverified · 9 pointer-only (licence)Open-domain question answering relies on efficient passage retrieval to select candidate contexts, where traditional sparse vector space models, such as TF-IDF or BM25, are the de facto method.
-
22 May 2020 18 repositories listed Syntology ran 4 of 6 samples · 2 unverifiedLarge pre-trained language models have been shown to store factual knowledge in their parameters, and achieve state-of-the-art results when fine-tuned on downstream NLP tasks.
-
25 Feb 2020 16 repositories listed Syntology ran 2 of 2 samples · 0 unverified · 1 pointer-only (licence)This paper provides a pair similarity optimization viewpoint on deep feature learning, aiming to maximize the within-class similarity sₚ and minimize the between-class similarity sₙ.
-
5 May 2018 15 repositories listed Syntology ran 19 of 23 samples · 4 unverified · 16 pointer-only (licence)Neural net classifiers trained on data with annotated class labels can also capture apparent visual similarity among categories without being directed to do so.
-
20 Sep 2019 14 repositories listed Syntology ran 8 of 17 samples · 9 unverifiedTo enable evaluation of progress on code search, we are releasing the CodeSearchNet Corpus and are presenting the CodeSearchNet Challenge, which consists of 99 natural language queries with about 4k expert relevance…
-
3 Nov 2017 14 repositories listed Syntology ran 3 of 24 samples · 21 unverified · 2 pointer-only (licence)We show that both hard-positive and hard-negative examples, selected by exploiting the geometry and the camera positions available from the 3D models, enhance the performance of particular-object retrieval.
-
23 Nov 2015 14 repositories listed Syntology ran 0 of 12 samples · 12 unverifiedWe tackle the problem of large scale visual place recognition, where the task is to quickly and accurately recognize the location of a given query photograph.
-
19 Dec 2016 13 repositories listed Syntology ran 3 of 12 samples · 9 unverifiedWe propose an attentive local feature descriptor suitable for large-scale image retrieval, referred to as DELF (DEep Local Feature).
-
7 Feb 2018 12 repositories listedThis work addresses the problem of billion-scale nearest neighbor search.
-
6 Aug 2019 11 repositories listed Syntology ran 10 of 34 samples · 24 unverified · 34 pointer-only (licence)We present ViLBERT (short for Vision-and-Language BERT), a model for learning task-agnostic joint representations of image content and natural language.
-
26 Nov 2016 11 repositories listed Syntology ran 2 of 2 samples · 0 unverifiedWe introduce the task of Visual Dialog, which requires an AI agent to hold a meaningful dialog with humans in natural, conversational language about visual content.
-
4 Jun 2013 11 repositories listed Syntology ran 23 of 30 samples · 7 unverified · 23 pointer-only (licence)Optimal transportation distances are a fundamental family of parameterized distances for histograms.
-
1 Jul 2018 10 repositories listedUser response prediction is a crucial component for personalized information retrieval and filtering scenarios, such as recommender system and web search.
-
28 Nov 2017 10 repositories listed Syntology ran 0 of 3 samples · 3 unverified · 3 pointer-only (licence)In this paper, we explicitly consider this challenge by introducing camera style (CamStyle) adaptation.
-
18 Jul 2017 10 repositories listed Syntology ran 2 of 5 samples · 3 unverified · 4 pointer-only (licence)We present a new technique for learning visual-semantic embeddings for cross-modal retrieval.
-
31 Mar 2017 10 repositories listed Syntology ran 1 of 1 samples · 0 unverified · 1 pointer-only (licence)This paper proposes to tackle open- domain question answering using Wikipedia as the unique knowledge source: the answer to any factoid question is a text span in a Wikipedia article.
-
28 Jan 2022 9 repositories listedFurthermore, performance improvement has been largely achieved by scaling up the dataset with noisy image-text pairs collected from the web, which is a suboptimal source of supervision.
-
22 Feb 2021 9 repositories listed Syntology ran 2 of 2 samples · 0 unverified · 1 pointer-only (licence)We show the formal equivalence of linearised self-attention mechanisms and fast weight controllers from the early '90s, where a ``slow" neural net learns by gradient descent to program the ``fast weights" of another net…
-
28 Jul 2020 9 repositories listedThe advent of deep machine learning platforms such as Tensorflow and Pytorch, developed in expressive high-level languages such as Python, have allowed more expressive representations of deep neural network…
-
27 Apr 2020 9 repositories listed Syntology ran 1 of 1 samples · 0 unverifiedColBERT introduces a late interaction architecture that independently encodes the query and the document using BERT and then employs a cheap yet powerful interaction step that models their fine-grained similarity.
-
3 Nov 2018 9 repositories listedMachine translation is highly sensitive to the size and quality of the training data, which has led to an increasing interest in collecting and filtering large parallel corpora.
-
27 May 2017 9 repositories listed Syntology ran 3 of 6 samples · 3 unverified · 3 pointer-only (licence)Despite their attractive properties and potential for opening up entirely new neural architectures, complex-valued deep neural networks have been marginalized due to the absence of the building blocks required to design…
-
4 Mar 2017 9 repositories listedThis paper extends fully-convolutional neural networks (FCN) for the clothing parsing problem.
-
17 May 2016 9 repositories listedState-of-the-art methods for zero-shot visual recognition formulate learning as a joint embedding problem of images and side information.
-
2 Jul 2020 8 repositories listedGenerative models for open domain question answering have proven to be competitive, without resorting to external knowledge.
-
14 Nov 2019 8 repositories listedThe works in the domain of visual semantic embeddings address this problem by first constructing a semantic embedding space based on some external knowledge and projecting image embeddings onto this fixed semantic…
-
24 Jul 2017 8 repositories listedThis article aims to provide a comprehensive review of recent research efforts on deep learning based recommender systems.
Syntology lines on 19 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections