Browse State-of-the-Art › Zero-shot Generalization
Zero-shot Generalization
301 papers with code · 1 benchmark · 1 dataset archive 2025-07-28
Benchmarks archive 2025-07-28
1 leaderboard table shown for this task, 1 with rows (a “benchmark” on this site is a table with at least one row, as on /sota), ordered by row count. “Best model” is the first row in the archive's own order at snapshot; nothing is re-ranked here and metric direction is not recorded in the archive. PwC's Trend sparklines are not in the archive, so that column is omitted.
| Dataset | Best model (first row in archive order) | Paper | Code | Syntology | Compare |
|---|---|---|---|---|---|
| CALVIN (5 rows) | GR-MG | GR-MG: Leveraging Partially Annotated Data via Multi-Modal... | code | Syntology ran 8 of 10 samples · 2 unverified | Compare |
Syntology column: samples harvested from the paper's repositories and executed on synthesized fixtures; “ran” is not a correctness claim and does not order the table. A dash means no Syntology record for that paper, not a recorded non-run. Read from the graph 2026-09-24.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
1 dataset whose archive record lists this task, ordered by the archive's paper count.
Subtasks archive 2025-07-28
No subtask under this task in the archive's task tree.
Most implemented papers archive 2025-07-28
30 shown of 301 papers with code (572 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
15 Oct 2021 8 repositories listed Syntology ran 8 of 15 samples · 7 unverifiedLarge language models have recently been shown to attain reasonable zero-shot generalization on a diverse set of tasks (Brown et al., 2020).
-
31 Oct 2017 7 repositories listedHumans can understand and produce new utterances effortlessly, thanks to their compositional skills.
-
23 Feb 2023 6 repositories listed Syntology ran 10 of 16 samples · 6 unverified · 1 pointer-only (licence)Finally, ZoeD-M12-NK is the first model that can jointly train on multiple datasets (NYU Depth v2 and KITTI) without a significant drop in performance and achieve unprecedented zero-shot generalization performance to…
-
4 Dec 2023 4 repositories listed Syntology ran 14 of 26 samples · 12 unverifiedMonocular depth estimation is a fundamental computer vision task.
-
12 Jun 2020 4 repositories listed Syntology ran 3 of 7 samples · 4 unverifiedEnd-to-end training of neural network solvers for graph combinatorial optimization problems such as the Travelling Salesperson Problem (TSP) have seen a surge of interest recently, but remain intractable and inefficient…
-
12 Sep 2019 4 repositories listedIn this work, we introduce the the Schema-Guided Dialogue (SGD) dataset, containing over 16k multi-domain conversations spanning 16 domains.
-
4 Jun 2019 4 repositories listed Syntology ran 0 of 2 samples · 2 unverifiedWhile multi-agent interactions can be naturally modeled as a graph, the environment has traditionally been considered as a black box.
-
9 Mar 2025 3 repositories listedTraditional methods for reasoning segmentation rely on supervised fine-tuning with categorical labels and simple descriptions, limiting its out-of-domain generalization and lacking explicit reasoning processes.
-
3 Apr 2024 3 repositories listed Syntology ran 5 of 11 samples · 6 unverifiedWe present Visual AutoRegressive modeling (VAR), a new generation paradigm that redefines the autoregressive learning on images as coarse-to-fine "next-scale prediction" or "next-resolution prediction", diverging from…
-
8 Feb 2024 3 repositories listed Syntology ran 4 of 4 samples · 0 unverifiedUnlike past methods that learn to route among specialized models, PHATGOOSE explores the possibility that zero-shot generalization will be improved if different experts can be adaptively chosen for each token and at…
-
20 Dec 2023 3 repositories listed Syntology ran 0 of 1 samples · 1 unverifiedIn this paper, we extend the scope of this effectiveness by showing that visual robot manipulation can significantly benefit from large-scale video generative pre-training.
-
17 May 2023 3 repositories listed Syntology ran 3 of 4 samples · 1 unverifiedTested on 14 previously unseen datasets, the One-Prompt Model showcases superior zero-shot segmentation capabilities, outperforming a wide range of related methods.
-
21 Dec 2022 3 repositories listedTo address this issue, we propose \emph{Img2Prompt}, a plug-and-play module that provides the prompts that can bridge the aforementioned modality and task disconnections, so that LLMs can perform zero-shot VQA tasks…
-
22 Mar 2022 3 repositories listed Syntology ran 0 of 3 samples · 3 unverifiedMultiple domains like vision, natural language, and audio are witnessing tremendous progress by leveraging Transformers for large scale pre-training followed by task specific fine tuning.
-
8 Jul 2020 3 repositories listed Syntology ran 5 of 5 samples · 0 unverified · 2 pointer-only (licence)In this work, we focus on an analogical reasoning task that contains rich compositional structures, Raven's Progressive Matrices (RPM).
-
5 Nov 2019 3 repositories listed Syntology ran 0 of 6 samples · 6 unverifiedWe study compositional generalization, viz., the problem of zero-shot generalization to novel compositions of concepts in a domain.
-
29 Oct 2019 3 repositories listedWe introduce the Convolutional Conditional Neural Process (ConvCNP), a new member of the Neural Process family that models translation equivariance in the data.
-
3 Jul 2025 2 repositories listed Syntology ran 0 of 1 samples · 1 unverified · 1 pointer-only (licence)We introduce DeSTA2.
-
15 Apr 2025 2 repositories listed Syntology ran 16 of 31 samples · 15 unverified · 31 pointer-only (licence)Unsupervised reinforcement learning (RL) aims at pre-training agents that can solve a wide range of downstream tasks in complex environments.
-
Efficient Alignment of Unconditioned Action Prior for Language-conditioned Pick and Place in Clutter12 Mar 2025 2 repositories listedWe study the task of language-conditioned pick and place in clutter, where a robot should grasp a target object in open clutter and move it to a specified place.
-
17 Jan 2025 2 repositories listed Syntology ran 36 of 46 samples · 10 unverified · 46 pointer-only (licence)However, achieving strong zero-shot generalization - a hallmark of foundation models in other computer vision tasks - remains challenging for stereo matching.
-
9 Jan 2025 2 repositories listedDetecting object-level changes between two images across possibly different views is a core task in many applications that involve visual inspection or camera surveillance.
-
7 Jan 2025 2 repositories listedDespite the considerable performance improvements of face recognition algorithms in recent years, the same scientific advances responsible for this progress can also be used to create efficient ways to attack them,…
-
20 Dec 2024 2 repositories listed Syntology ran 2 of 13 samples · 11 unverifiedDiffusion Transformers (DiT) have become a leading architecture in image generation.
-
24 Sep 2024 2 repositories listed Syntology ran 16 of 26 samples · 10 unverified · 26 pointer-only (licence)Instruction tuning has emerged as an effective strategy for achieving zero-shot generalization by finetuning pretrained models on diverse multimodal tasks.
-
31 Jul 2024 2 repositories listedTo address this gap, this work conducts a systematic review on SAM for videos in the era of foundation models.
-
4 Jun 2024 2 repositories listed Syntology ran 5 of 7 samples · 2 unverifiedWe consider the task of active geo-localization (AGL) in which an agent uses a sequence of visual cues observed during aerial navigation to find a target specified through multiple possible modalities.
-
2 May 2024 2 repositories listed Syntology ran 4 of 5 samples · 1 unverifiedLearning to solve vehicle routing problems (VRPs) has garnered much attention.
-
14 Mar 2024 2 repositories listed Syntology ran 5 of 12 samples · 7 unverifiedSegment anything models (SAMs) are gaining attention for their zero-shot generalization capability in segmenting objects of unseen classes and in unseen domains when properly prompted.
-
12 Mar 2024 2 repositories listedRSBuilding is designed to enhance cross-scene generalization and task universality.
Syntology lines on 20 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections