Browse State-of-the-Art › Natural Language Queries
Natural Language Queries
135 papers with code · 1 benchmark · 3 datasets archive 2025-07-28
Benchmarks archive 2025-07-28
1 leaderboard table shown for this task, 1 with rows (a “benchmark” on this site is a table with at least one row, as on /sota), ordered by row count. “Best model” is the first row in the archive's own order at snapshot; nothing is re-ranked here and metric direction is not recorded in the archive. PwC's Trend sparklines are not in the archive, so that column is omitted.
| Dataset | Best model (first row in archive order) | Paper | Code | Syntology | Compare |
|---|---|---|---|---|---|
| Ego4D (10 rows) | EgoVideo | EgoVideo: Exploring Egocentric Foundation Model and Downstream Adaptation | code | Syntology ran 11 of 15 samples · 4 unverified | Compare |
Syntology column: samples harvested from the paper's repositories and executed on synthesized fixtures; “ran” is not a correctness claim and does not order the table. A dash means no Syntology record for that paper, not a recorded non-run. Read from the graph 2026-09-24.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
3 datasets whose archive record lists this task, ordered by the archive's paper count.
Subtasks archive 2025-07-28
1 subtask in the archive's task tree.
Most implemented papers archive 2025-07-28
30 shown of 135 papers with code (337 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
20 Sep 2019 14 repositories listed Syntology ran 8 of 17 samples · 9 unverifiedTo enable evaluation of progress on code search, we are releasing the CodeSearchNet Corpus and are presenting the CodeSearchNet Challenge, which consists of 99 natural language queries with about 4k expert relevance…
-
5 May 2017 12 repositories listedFor evaluation, we adopt TaCoS dataset, and build a new dataset for this task on top of Charades by adding sentence temporal annotations, called Charades-STA.
-
20 Jul 2021 4 repositories listed Syntology ran 2 of 2 samples · 0 unverified · 1 pointer-only (licence)Each video in the dataset is annotated with: (1) a human-written free-form NL query, (2) relevant moments in the video w.
-
26 Apr 2023 3 repositories listedV3CTRON is an open source vector database that allows users to upload text based documents & document collections, which are automatically embedded for super-accurate semantic search & retrieval using natural language…
-
23 Mar 2022 3 repositories listedFinding relevant moments and highlights in videos according to natural language queries is a natural and highly valuable common need in the current video content explosion era.
-
10 Dec 2021 3 repositories listedThe first computes a textual representation of a given question, the second combines it with the entity embeddings for entities involved in the question, and the third generates question-specific time embeddings.
-
10 Feb 2020 3 repositories listedIt has recently been observed that neural language models trained on unstructured text can implicitly store and retrieve knowledge using natural language queries.
-
31 Jul 2019 3 repositories listed Syntology ran 0 of 5 samples · 5 unverified · 1 pointer-only (licence)The rapid growth of video on the internet has made searching for video content using natural language queries a significant challenge.
-
11 Dec 2023 2 repositories listedAdapting existing short video (5-30 seconds) grounding methods to this problem yields poor performance.
-
6 Dec 2023 2 repositories listed Syntology ran 12 of 24 samples · 12 unverified · 1 pointer-only (licence)In this work, we go one step further: in addition to radiance field rendering, we enable 3D Gaussian splatting on arbitrary-dimension semantic features via 2D foundation model distillation.
-
1 Aug 2023 2 repositories listedExperimental results show that Late Fusion contrastive learning for Neural RIR outperforms all other contrastive IR configurations, Neural IR, and sparse retrieval baselines, thus demonstrating the power of exploiting…
-
10 Apr 2023 2 repositories listed Syntology ran 0 of 3 samples · 3 unverifiedThis capability is vital for Artificial Intelligence (AI) and should be embedded in comprehensive AI Agents, enabling them to harness expert models for complex task-solving towards Artificial General Intelligence (AGI).
-
17 Nov 2022 2 repositories listedIn this report, we present our champion solutions to five tracks at Ego4D challenge.
-
3 Jun 2022 2 repositories listed Syntology ran 3 of 3 samples · 0 unverified · 2 pointer-only (licence)Video-Language Pretraining (VLP), which aims to learn transferable representation to advance a wide range of video-text downstream tasks, has recently received increasing attention.
-
12 Sep 2021 2 repositories listed Syntology ran 1 of 1 samples · 0 unverifiedInternet search affects people's cognition of the world, so mitigating biases in search results and learning fair models is imperative for social good.
-
7 Oct 2020 2 repositories listed Syntology ran 0 of 7 samples · 7 unverifiedHowever, the extent to which these representations learned for link prediction generalize to other tasks is unclear.
-
1 Jul 2020 2 repositories listedIn a separate line of research, KG embedding methods have been proposed to reduce KG sparsity by performing missing link prediction.
-
9 May 2019 2 repositories listedOur evaluation shows that: 1.
-
7 Sep 2018 2 repositories listedIn this work, we introduce a general purpose transfer-learnable NLI with the goal of learning one model that can be used as NLI for any relational database.
-
6 Jul 2018 2 repositories listedWe address the problem of segmenting an object given a natural language expression that describes it.
-
4 Aug 2017 2 repositories listed Syntology ran 1 of 1 samples · 0 unverified · 1 pointer-only (licence)A key obstacle to training our MCN model is that current video datasets do not include pairs of localized video segments and referring expressions, or text descriptions which uniquely identify a corresponding moment.
-
1 Jun 2017 2 repositories listedFor a given text query and background API, the tool finds candidate functions by performing a translation from the text to known representations in the API using the semantic parsing approach of Richardson and Kuhn…
-
28 Nov 2016 2 repositories listedThe main experimental result in this paper is that a single Neural Programmer model achieves 34.
-
9 Aug 2016 2 repositories listedThis paper explores the task of translating natural language queries into regular expressions which embody their meaning.
-
11 Jun 2025 1 repository listedTo evaluate how well general knowledge is preserved in a finetuned representation, we introduce a metric that measures image retrieval accuracy based on captions generated by a vision language model (VLM).
-
9 Jun 2025 1 repository listedWe evaluated SEED on BIRD and Spider, demonstrating that it significantly improves SQL generation accuracy in the no-evidence scenario, and in some cases, even outperforms the setting where BIRD evidence is provided.
-
4 Jun 2025 1 repository listedIn this report, we present our champion solutions for the three egocentric video localization tracks of the Ego4D Episodic Memory Challenge at CVPR 2025.
-
2 Jun 2025 1 repository listedWe introduce DualMap, an online open-vocabulary mapping system that enables robots to understand and navigate dynamically changing environments through natural language queries.
-
1 Jun 2025 1 repository listedTo address this, we propose the concept of sensitivity awareness (SA), which enables LLMs to adhere to predefined access rights rules.
-
27 May 2025 1 repository listedAutonomous Vehicles (AVs) collect and pseudo-label terabytes of multi-modal data localized to HD maps during normal fleet testing.
Syntology lines on 9 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections