Browse State-of-the-Art › Zero-shot Generalization › Papers, page 4
Zero-shot Generalization
Papers archive 2025-07-28
archive papers tagged: 572 · with a code link: 301 · where Syntology ran a sample: 128 (119 with a run with no instrument failure, 9 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (128 of 572 tagged: 119 with a run with no instrument failure, 9 where every run was a failure of Syntology's instrument)
Page 4 of 6: papers 301 to 400 of 572, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
23 Apr 2018 1 repository listed
-
SAMST: A Transformer framework based on SAM pseudo label filtering for remote sensing semi-supervised semantic segmentation16 Jul 2025 0 repositories listed
-
Towards Depth Foundation Model: Recent Trends in Vision-Based Depth Estimation15 Jul 2025 0 repositories listed
-
PoseLLM: Enhancing Language-Guided Human Pose Estimation with MLP Alignment12 Jul 2025 0 repositories listed
-
Go to Zero: Towards Zero-shot Motion Generation with Million-scale Data9 Jul 2025 0 repositories listed
-
Video Event Reasoning and Prediction by Fusing World Knowledge from LLMs with Vision Foundation Models8 Jul 2025 0 repositories listed
-
Helping CLIP See Both the Forest and the Trees: A Decomposition and Description Approach4 Jul 2025 0 repositories listed
-
RobuSTereo: Robust Zero-Shot Stereo Matching under Adverse Weather2 Jul 2025 0 repositories listed
-
VisLanding: Monocular 3D Perception for UAV Safe Landing via Depth-Normal Synergy17 Jun 2025 0 repositories listed
-
LeVERB: Humanoid Whole-Body Control with Latent Vision-Language Instruction16 Jun 2025 0 repositories listed
-
DEAL: Disentangling Transformer Head Activations for LLM Steering10 Jun 2025 0 repositories listed
-
CXR-LT 2024: A MICCAI challenge on long-tailed, multi-label, and zero-shot disease classification from chest X-ray9 Jun 2025 0 repositories listed
-
Deep Equivariant Multi-Agent Control Barrier Functions9 Jun 2025 0 repositories listed
-
ZeroVO: Visual Odometry with Minimal Assumptions9 Jun 2025 0 repositories listed
-
Latent Diffusion Model Based Denoising Receiver for 6G Semantic Communication: From Stochastic Differential Theory to Application6 Jun 2025 0 repositories listed
-
Generating Synthetic Stereo Datasets using 3D Gaussian Splatting and Expert Knowledge Transfer5 Jun 2025 0 repositories listed
-
Towards Vision-Language-Garment Models For Web Knowledge Garment Understanding and Generation5 Jun 2025 0 repositories listed
-
Language-Guided Multi-Agent Learning in Simulations: A Unified Framework and Evaluation1 Jun 2025 0 repositories listed
-
ViTaPEs: Visuotactile Position Encodings for Cross-Modal Alignment in Multimodal Transformers26 May 2025 0 repositories listed
-
WHISTRESS: Enriching Transcriptions with Sentence Stress Detection25 May 2025 0 repositories listed
-
Anchored Diffusion Language Model24 May 2025 0 repositories listed
-
G1: Teaching LLMs to Reason on Graphs with Reinforcement Learning24 May 2025 0 repositories listed
-
CoMo: Learning Continuous Latent Motion from Internet Videos for Scalable Robot Learning22 May 2025 0 repositories listed
-
EasyInsert: A Data-Efficient and Generalizable Insertion Policy22 May 2025 0 repositories listed
-
AnyBody: A Benchmark Suite for Cross-Embodiment Manipulation21 May 2025 0 repositories listed
-
EndoVLA: Dual-Phase Vision-Language-Action Model for Autonomous Tracking in Endoscopy21 May 2025 0 repositories listed
-
gen2seg: Generative Models Enable Generalizable Instance Segmentation21 May 2025 0 repositories listed
-
ORQA: A Benchmark and Foundation Model for Holistic Operating Room Modeling19 May 2025 0 repositories listed
-
AoP-SAM: Automation of Prompts for Efficient Segmentation17 May 2025 0 repositories listed
-
Depth Anything with Any Prior15 May 2025 0 repositories listed
-
NVSPolicy: Adaptive Novel-View Synthesis for Generalizable Language-Conditioned Policy Learning15 May 2025 0 repositories listed
-
Denoising and Alignment: Rethinking Domain Generalization for Multimodal Face Anti-Spoofing14 May 2025 0 repositories listed
-
Visual Image Reconstruction from Brain Activity via Latent Representation13 May 2025 0 repositories listed
-
Towards Artificial General or Personalized Intelligence? A Survey on Foundation Models for Personalized Federated Intelligence11 May 2025 0 repositories listed
-
Pro2SAM: Mask Prompt to SAM with Grid Points for Weakly Supervised Object Localization8 May 2025 0 repositories listed
-
A Review of 3D Object Detection with Vision-Language Models25 Apr 2025 0 repositories listed
-
Text-to-Decision Agent: Learning Generalist Policies from Natural Language Supervision21 Apr 2025 0 repositories listed
-
Evolutionary Prompt Optimization Discovers Emergent Multimodal Reasoning Strategies in Vision-Language Models30 Mar 2025 0 repositories listed
-
Zero-shot Domain Generalization of Foundational Models for 3D Medical Image Segmentation: An Experimental Study28 Mar 2025 0 repositories listed
-
Thinking agents for zero-shot generalization to qualitatively novel tasks25 Mar 2025 0 repositories listed
-
Unpaired Object-Level SAR-to-Optical Image Translation for Aircraft with Keypoints-Guided Diffusion Models25 Mar 2025 0 repositories listed
-
Aether: Geometric-Aware Unified World Modeling24 Mar 2025 0 repositories listed
-
Enhancing Zero-Shot Image Recognition in Vision-Language Models through Human-like Concept Guidance20 Mar 2025 0 repositories listed
-
20 Mar 2025 0 repositories listed
-
GenM³: Generative Pretrained Multi-path Motion Model for Text Conditional Human Motion Generation19 Mar 2025 0 repositories listed
-
Good Actions Succeed, Bad Actions Generalize: A Case Study on Why RL Generalizes Better19 Mar 2025 0 repositories listed
-
Foundation Feature-Driven Online End-Effector Pose Estimation: A Marker-Free and Learning-Free Approach18 Mar 2025 0 repositories listed
-
Compound Expression Recognition via Large Vision-Language Models14 Mar 2025 0 repositories listed
-
A Recipe for Improving Remote Sensing VLM Zero Shot Generalization10 Mar 2025 0 repositories listed
-
PoseLess: Depth-Free Vision-to-Joint Control via Direct Image Mapping with VLM10 Mar 2025 0 repositories listed
-
OTTER: A Vision-Language-Action Model with Text-Aware Visual Feature Extraction5 Mar 2025 0 repositories listed
-
RAILGUN: A Unified Convolutional Policy for Multi-Agent Path Finding Across Different Environments and Tasks4 Mar 2025 0 repositories listed
-
Contrastive Learning of English Language and Crystal Graphs for Multimodal Representation of Materials Knowledge23 Feb 2025 0 repositories listed
-
Learning from Reward-Free Offline Data: A Case for Planning with Latent Dynamics Models20 Feb 2025 0 repositories listed
-
WRT-SAM: Foundation Model-Driven Segmentation for Generalized Weld Radiographic Testing17 Feb 2025 0 repositories listed
-
Salience-Invariant Consistent Policy Learning for Generalization in Visual Reinforcement Learning12 Feb 2025 0 repositories listed
-
Mechanistic Understandings of Representation Vulnerabilities and Engineering Robust Vision Transformers7 Feb 2025 0 repositories listed
-
SimSort: A Data-Driven Framework for Spike Sorting by Large-Scale Electrophysiology Simulation5 Feb 2025 0 repositories listed
-
Toward Task Generalization via Memory Augmentation in Meta-Reinforcement Learning3 Feb 2025 0 repositories listed
-
A Zero-Shot Generalization Framework for LLM-Driven Cross-Domain Sequential Recommendation31 Jan 2025 0 repositories listed
-
FlexiCrackNet: A Flexible Pipeline for Enhanced Crack Segmentation with General Features Transfered from SAM31 Jan 2025 0 repositories listed
-
Test-time Loss Landscape Adaptation for Zero-Shot Generalization in Vision-Language Models31 Jan 2025 0 repositories listed
-
DynaPrompt: Dynamic Test-Time Prompt Tuning27 Jan 2025 0 repositories listed
-
Zero-Shot Trajectory Planning for Signal Temporal Logic Tasks23 Jan 2025 0 repositories listed
-
State Combinatorial Generalization In Decision Making With Conditional Diffusion Models22 Jan 2025 0 repositories listed
-
Survey on Monocular Metric Depth Estimation21 Jan 2025 0 repositories listed
-
MIFNet: Learning Modality-Invariant Features for Generalizable Multimodal Image Matching20 Jan 2025 0 repositories listed
-
Chain-of-Reasoning: Towards Unified Mathematical Reasoning in Large Language Models via a Multi-Paradigm Perspective19 Jan 2025 0 repositories listed
-
Zero-Shot Monocular Scene Flow Estimation in the Wild17 Jan 2025 0 repositories listed
-
StereoGen: High-quality Stereo Image Generation from a Single Image15 Jan 2025 0 repositories listed
-
Robotic Programmer: Video Instructed Policy Code Generation for Robotic Manipulation8 Jan 2025 0 repositories listed
-
Spot Risks Before Speaking! Unraveling Safety Attention Heads in Large Vision-Language Models3 Jan 2025 0 repositories listed
-
On the Out-Of-Distribution Generalization of Large Multimodal Models1 Jan 2025 0 repositories listed
-
On the Zero-shot Adversarial Robustness of Vision-Language Models: A Truly Zero-shot and Training-free Approach1 Jan 2025 0 repositories listed
-
From Pixels to Predicates: Learning Symbolic World Models via Pretrained Vision-Language Models31 Dec 2024 0 repositories listed
-
EC-Diffuser: Multi-Object Manipulation via Entity-Centric Behavior Generation25 Dec 2024 0 repositories listed
-
Multiple Consistency-guided Test-Time Adaptation for Contrastive Audio-Language Models with Unlabeled Audio23 Dec 2024 0 repositories listed
-
Efficient Fine-Tuning of Single-Cell Foundation Models Enables Zero-Shot Molecular Perturbation Prediction18 Dec 2024 0 repositories listed
-
Marigold-DC: Zero-Shot Monocular Depth Completion with Guided Diffusion18 Dec 2024 0 repositories listed
-
Zero-Shot Generalization for Blockage Localization in mmWave Communication18 Dec 2024 0 repositories listed
-
EasyRef: Omni-Generalized Group Image Reference for Diffusion Models via Multimodal LLM12 Dec 2024 0 repositories listed
-
WiFo: Wireless Foundation Model for Channel Prediction12 Dec 2024 0 repositories listed
-
Disentanglement and Compositionality of Letter Identity and Letter Position in Variational Auto-Encoder Vision Models11 Dec 2024 0 repositories listed
-
Lightweight Method for Interactive 3D Medical Image Segmentation with Multi-Round Result Fusion11 Dec 2024 0 repositories listed
-
S³: Synonymous Semantic Space for Improving Zero-Shot Generalization of Vision-Language Models6 Dec 2024 0 repositories listed
-
CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance5 Dec 2024 0 repositories listed
-
The Matrix: Infinite-Horizon World Generation with Real-Time Moving Control4 Dec 2024 0 repositories listed
-
UTSD: Unified Time Series Diffusion Model4 Dec 2024 0 repositories listed
-
Visatronic: A Multimodal Decoder-Only Model for Speech Synthesis26 Nov 2024 0 repositories listed
-
Generating Out-Of-Distribution Scenarios Using Language Models25 Nov 2024 0 repositories listed
-
Style-Pro: Style-Guided Prompt Learning for Generalizable Vision-Language Models25 Nov 2024 0 repositories listed
-
Context-Aware Multimodal Pretraining22 Nov 2024 0 repositories listed
-
HEIGHT: Heterogeneous Interaction Graph Transformer for Robot Navigation in Crowded and Constrained Environments19 Nov 2024 0 repositories listed
-
Scalable Autoregressive Monocular Depth Estimation18 Nov 2024 0 repositories listed
-
Mono2Stereo: Monocular Knowledge Transfer for Enhanced Stereo Matching14 Nov 2024 0 repositories listed
-
Self-Supervised Monocular 4D Scene Reconstruction for Egocentric Videos14 Nov 2024 0 repositories listed
-
In the Era of Prompt Learning with Vision-Language Models7 Nov 2024 0 repositories listed
-
Compositional Automata Embeddings for Goal-Conditioned Reinforcement Learning31 Oct 2024 0 repositories listed
-
JudgeRank: Leveraging Large Language Models for Reasoning-Intensive Reranking31 Oct 2024 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.