Methods › Computer Vision › Vision and Language Pre-Trained Models › CLIP › Papers where code ran, page 1
Contrastive Language-Image Pre-training
CLIP
Papers archive 2025-07-28
archive papers tagged: 3,094 · with a code link: 1,617 · where Syntology ran a sample: 649 (554 with a run with no instrument failure, 95 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (649 of 3,094 tagged: 554 with a run with no instrument failure, 95 where every run was a failure of Syntology's instrument)
Syntology We ran code from the paper's repository; we did not isolate this method inside it.
Page 1 of 7: papers 1 to 100 of the 649 tagged papers where Syntology ran at least one harvested sample (554 with a run with no instrument failure, 95 where every run was a failure of Syntology's instrument), newest first by the archive's date (ties by arXiv id). This is a filter on Syntology's record ordered by date only, not a ranking; a run is not a correctness claim. A paper missing from this list is not a recorded non-run: it may have no arXiv id, no harvested code, or only samples that have not run yet.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Test-Time Canonicalization by Foundation Models for Robust Perception 14 Jul 2025 · 1 repository · arXiv:2507.10375Syntology official (archive's flag): 5 ran · 5 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
CultureCLIP: Empowering CLIP with Cultural Awareness through Synthetic Images and Contextualized Captions 8 Jul 2025 · 1 repository · arXiv:2507.06210Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
pFedMMA: Personalized Federated Fine-Tuning with Multi-Modal Adapter for Vision-Language Models 7 Jul 2025 · 1 repository · arXiv:2507.05394Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 7 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
SharpZO: Hybrid Sharpness-Aware Vision Language Model Prompt Tuning via Forward-Only Passes 26 Jun 2025 · 1 repository · arXiv:2506.20990Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 4 where Syntology's instrument failed) · 4 unverified (of 10 harvested samples) · 4 pointer-only (licence)
-
Evolutionary Caching to Accelerate Your Off-the-Shelf Diffusion Model 18 Jun 2025 · 1 repository · arXiv:2506.15682Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval 10 Jun 2025 · 1 repository · arXiv:2506.08887Syntology official (archive's flag): 2 ran · 2 ran (of which 1 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Vision Transformers Don't Need Trained Registers 9 Jun 2025 · 1 repository · arXiv:2506.08010Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Robustness in Both Domains: CLIP Needs a Robust Text Encoder 3 Jun 2025 · 0 repositories · arXiv:2506.03355Syntology 9 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 8 where Syntology's instrument failed) · 3 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
Conformal Prediction for Zero-Shot Models 30 May 2025 · 1 repository · arXiv:2505.24693Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Correlating instruction-tuning (in multimodal models) with vision-language processing (in the brain) 26 May 2025 · 1 repository · arXiv:2505.20029Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
MoRE-Brain: Routed Mixture of Experts for Interpretable and Generalizable Cross-Subject fMRI Visual Decoding 21 May 2025 · 0 repositories · arXiv:2505.15946Syntology 2 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified; every one of the 2 samples that ran constructed an object rather than computing a result (of 4 harvested samples) · 4 pointer-only (licence)
-
GenZSL: Generative Zero-Shot Learning Via Inductive Variational Autoencoder 17 May 2025 · 1 repository · arXiv:2505.11882Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
MSCI: Addressing CLIP's Inherent Limitations for Compositional Zero-Shot Learning 15 May 2025 · 1 repository · arXiv:2505.10289Syntology official (archive's flag): 22 ran · 23 ran (of which 13 constructed an object rather than computing a result; 19 with no instrument failure: 0 honoured, 0 violated, 19 with no contract checked; 4 where Syntology's instrument failed) · 5 unverified (of 28 harvested samples) · 28 pointer-only (licence)
-
BLIP3-o: A Family of Fully Open Unified Multimodal Models-Architecture, Training and Dataset 14 May 2025 · 1 repository · arXiv:2505.09568Syntology official (archive's flag): 1 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Visually Guided Decoding: Gradient-Free Hard Prompt Inversion with Language Models 13 May 2025 · 1 repository · arXiv:2505.08622Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
DanceGRPO: Unleashing GRPO on Visual Generation 12 May 2025 · 1 repository · arXiv:2505.07818Syntology 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
FG-CLIP: Fine-Grained Visual and Textual Alignment 8 May 2025 · 1 repository · arXiv:2505.05071Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
X-Transfer Attacks: Towards Super Transferable Adversarial Attacks on CLIP 8 May 2025 · 1 repository · arXiv:2505.05528Syntology official (archive's flag): 5 ran · 5 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models 18 Apr 2025 · 2 repositories · arXiv:2504.14032Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 2 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 1 pointer-only (licence)
-
Logits DeConfusion with CLIP for Few-Shot Learning 16 Apr 2025 · 1 repository · arXiv:2504.12104Syntology official (archive's flag): 6 ran · 6 ran (of which 6 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; every one of the 6 samples that ran constructed an object rather than computing a result (of 7 harvested samples) · 7 pointer-only (licence)
-
Towards Efficient Partially Relevant Video Retrieval with Active Moment Discovering 15 Apr 2025 · 1 repository · arXiv:2504.10920Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
R-TPT: Improving Adversarial Robustness of Vision-Language Models through Test-Time Prompt Tuning 15 Apr 2025 · 1 repository · arXiv:2504.11195Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
HyperCore: The Core Framework for Building Hyperbolic Foundation Models with Comprehensive Modules 11 Apr 2025 · 1 repository · arXiv:2504.08912Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Sparse Autoencoders Learn Monosemantic Features in Vision-Language Models 3 Apr 2025 · 1 repository · arXiv:2504.02821Syntology official (archive's flag): 8 ran · 8 ran (of which 3 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 5 where Syntology's instrument failed) · 2 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
Rethinking Vision-Language Model in Face Forensics: Multi-Modal Interpretable Forged Face Detector 26 Mar 2025 · 1 repository · arXiv:2503.20188Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 1 violated, 7 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples)
-
Exploring Semantic Feature Discrimination for Perceptual Image Super-Resolution and Opinion-Unaware No-Reference Image Quality Assessment 25 Mar 2025 · 1 repository · arXiv:2503.19295Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
GOAL: Global-local Object Alignment Learning 22 Mar 2025 · 1 repository · arXiv:2503.17782Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Cross-Modal and Uncertainty-Aware Agglomeration for Open-Vocabulary 3D Scene Understanding 20 Mar 2025 · 1 repository · arXiv:2503.16707Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Recover and Match: Open-Vocabulary Multi-Label Recognition through Knowledge-Constrained Optimal Transport 19 Mar 2025 · 1 repository · arXiv:2503.15337Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
FP4DiT: Towards Effective Floating Point Quantization for Diffusion Transformers 19 Mar 2025 · 1 repository · arXiv:2503.15465Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
Hyperbolic Safety-Aware Vision-Language Models 15 Mar 2025 · 1 repository · arXiv:2503.12127Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
WISE: A World Knowledge-Informed Semantic Evaluation for Text-to-Image Generation 10 Mar 2025 · 2 repositories · arXiv:2503.07265Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
SEED: Towards More Accurate Semantic Evaluation for Visual Brain Decoding 9 Mar 2025 · 0 repositories · arXiv:2503.06437Syntology 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
DiffCLIP: Differential Attention Meets CLIP 9 Mar 2025 · 1 repository · arXiv:2503.06626Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
CLIP is Strong Enough to Fight Back: Test-time Counterattacks towards Zero-shot Adversarial Robustness of CLIP 5 Mar 2025 · 1 repository · arXiv:2503.03613Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Language-Assisted Feature Transformation for Anomaly Detection 3 Mar 2025 · 1 repository · arXiv:2503.01184Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Extrapolating and Decoupling Image-to-Video Generation Models: Motion Modeling is Easier Than You Think 2 Mar 2025 · 1 repository · arXiv:2503.00948Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 1 violated, 7 with no contract checked; 4 where Syntology's instrument failed) · 3 unverified (of 15 harvested samples) · 4 pointer-only (licence)
-
CLIP Under the Microscope: A Fine-Grained Analysis of Multi-Object Representation 27 Feb 2025 · 1 repository · arXiv:2502.19842Syntology official (archive's flag): 4 ran · 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 2 honoured, 2 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Learning to Generalize without Bias for Open-Vocabulary Action Recognition 27 Feb 2025 · 0 repositories · arXiv:2502.20158Syntology 4 ran (of which 4 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified; every one of the 4 samples that ran constructed an object rather than computing a result (of 7 harvested samples)
-
UniTok: A Unified Tokenizer for Visual Generation and Understanding 27 Feb 2025 · 1 repository · arXiv:2502.20321Syntology official (archive's flag): 11 ran · 12 ran (of which 9 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 3 where Syntology's instrument failed) · 3 unverified (of 15 harvested samples) · 1 pointer-only (licence)
-
CLIPure: Purification in Latent Space via CLIP for Adversarially Robust Zero-Shot Classification 25 Feb 2025 · 1 repository · arXiv:2502.18176Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
FeatSharp: Your Vision Model Features, Sharper 22 Feb 2025 · 1 repository · arXiv:2502.16025Syntology 11 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 2 where Syntology's instrument failed) · 9 unverified (of 20 harvested samples) · 20 pointer-only (licence)
-
Modality-Aware Neuron Pruning for Unlearning in Multimodal Large Language Models 21 Feb 2025 · 1 repository · arXiv:2502.15910Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 6 where Syntology's instrument failed) · 12 unverified (of 18 harvested samples) · 18 pointer-only (licence)
-
SongGen: A Single Stage Auto-regressive Transformer for Text-to-Song Generation 18 Feb 2025 · 1 repository · arXiv:2502.13128Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
When and How Does CLIP Enable Domain and Compositional Generalization? 13 Feb 2025 · 0 repositories · arXiv:2502.09507Syntology 24 ran (of which 0 constructed an object rather than computing a result; 16 with no instrument failure: 0 honoured, 0 violated, 16 with no contract checked; 8 where Syntology's instrument failed) · 8 unverified (of 32 harvested samples) · 5 pointer-only (licence)
-
ConceptAttention: Diffusion Transformers Learn Highly Interpretable Features 6 Feb 2025 · 1 repository · arXiv:2502.04320Syntology official (archive's flag): 4 ran · 4 ran (of which 2 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Kronecker Mask and Interpretive Prompts are Language-Action Video Learners 5 Feb 2025 · 1 repository · arXiv:2502.03549Syntology official (archive's flag): 9 ran · 9 ran (of which 9 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified; every one of the 9 samples that ran constructed an object rather than computing a result (of 11 harvested samples) · 11 pointer-only (licence)
-
CLIP Behaves like a Bag-of-Words Model Cross-modally but not Uni-modally 5 Feb 2025 · 1 repository · arXiv:2502.03566Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Detecting Backdoor Samples in Contrastive Language Image Pretraining 3 Feb 2025 · 1 repository · arXiv:2502.01385Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Mechanistic understanding and validation of large AI models with SemanticLens 9 Jan 2025 · 1 repository · arXiv:2501.05398Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 10 harvested samples)
-
MVREC: A General Few-shot Defect Classification Model Using Multi-View Region-Context 22 Dec 2024 · 1 repository · arXiv:2412.16897Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 12 harvested samples)
-
Does VLM Classification Benefit from LLM Description Semantics? 16 Dec 2024 · 1 repository · arXiv:2412.11917Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Enhance Vision-Language Alignment with Noise 14 Dec 2024 · 1 repository · arXiv:2412.10817Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 3 pointer-only (licence)
-
Can Graph Neural Networks Learn Language with Extremely Weak Text Supervision? 11 Dec 2024 · 1 repository · arXiv:2412.08174Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples)
-
Post-hoc Probabilistic Vision-Language Models 8 Dec 2024 · 1 repository · arXiv:2412.06014Syntology official (archive's flag): 4 ran · 7 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 5 where Syntology's instrument failed) · 6 unverified (of 13 harvested samples) · 5 pointer-only (licence)
-
Sparse autoencoders reveal selective remapping of visual concepts during adaptation 6 Dec 2024 · 1 repository · arXiv:2412.05276Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples)
-
Liquid: Language Models are Scalable Multi-modal Generators 5 Dec 2024 · 1 repository · arXiv:2412.04332Syntology official (archive's flag): 1 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 1 pointer-only (licence)
-
VisionZip: Longer is Better but Not Necessary in Vision Language Models 5 Dec 2024 · 1 repository · arXiv:2412.04467Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 11 harvested samples)
-
FLAIR: VLM with Fine-grained Language-informed Image Representations 4 Dec 2024 · 2 repositories · arXiv:2412.03561Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 1 violated, 10 with no contract checked; 2 where Syntology's instrument failed) · 7 unverified (of 20 harvested samples) · 20 pointer-only (licence)
-
NLPrompt: Noise-Label Prompt Learning for Vision-Language Models 2 Dec 2024 · 1 repository · arXiv:2412.01256Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Bootstraping Clustering of Gaussians for View-consistent 3D Scene Understanding 29 Nov 2024 · 1 repository · arXiv:2411.19551Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 2 honoured, 0 violated, 9 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
Dual Risk Minimization: Towards Next-Level Robustness in Fine-tuning Zero-Shot Models 29 Nov 2024 · 1 repository · arXiv:2411.19757Syntology official (archive's flag): 6 ran · 6 ran (of which 1 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 4 where Syntology's instrument failed) · 5 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
Talking to DINO: Bridging Self-Supervised Vision Backbones with Language for Open-Vocabulary Segmentation 28 Nov 2024 · 1 repository · arXiv:2411.19331Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 10 harvested samples) · 2 pointer-only (licence)
-
Self-Calibrated CLIP for Training-Free Open-Vocabulary Segmentation 24 Nov 2024 · 1 repository · arXiv:2411.15869Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples) · 2 pointer-only (licence)
-
MUNBa: Machine Unlearning via Nash Bargaining 23 Nov 2024 · 1 repository · arXiv:2411.15537Syntology official (archive's flag): 2 ran · 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 2 honoured, 2 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Multimodal Autoregressive Pre-training of Large Vision Encoders 21 Nov 2024 · 1 repository · arXiv:2411.14402Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
REDUCIO! Generating 1024×1024 Video within 16 Seconds using Extremely Compressed Motion Latents 20 Nov 2024 · 1 repository · arXiv:2411.13552Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
CorrCLIP: Reconstructing Correlations in CLIP with Off-the-Shelf Foundation Models for Open-Vocabulary Semantic Segmentation 15 Nov 2024 · 1 repository · arXiv:2411.10086Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
UMFC: Unsupervised Multi-Domain Feature Calibration for Vision-Language Models 11 Nov 2024 · 1 repository · arXiv:2411.06921Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Robust Fine-tuning of Zero-shot Models via Variance Reduction 11 Nov 2024 · 1 repository · arXiv:2411.06966Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 6 where Syntology's instrument failed) · 4 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
On Erroneous Agreements of CLIP Image Embeddings 7 Nov 2024 · 1 repository · arXiv:2411.05195Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 1 violated, 4 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
Customized Multiple Clustering via Multi-Modal Subspace Proxy Learning 6 Nov 2024 · 1 repository · arXiv:2411.03978Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Classification Done Right for Vision-Language Pre-Training 5 Nov 2024 · 1 repository · arXiv:2411.03313Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 5 where Syntology's instrument failed) · 8 unverified (of 18 harvested samples)
-
PPLLaVA: Varied Video Sequence Understanding With Prompt Guidance 4 Nov 2024 · 1 repository · arXiv:2411.02327Syntology official (archive's flag): 4 ran · 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 1 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (of 9 harvested samples) · 3 pointer-only (licence)
-
B-cosification: Transforming Deep Neural Networks to be Inherently Interpretable 1 Nov 2024 · 1 repository · arXiv:2411.00715Syntology official (archive's flag): 8 ran · 12 ran (of which 2 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 13 harvested samples)
-
Contrasting with Symile: Simple Model-Agnostic Representation Learning for Unlimited Modalities 1 Nov 2024 · 1 repository · arXiv:2411.01053Syntology official (archive's flag): 2 ran · 2 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Text-Guided Attention is All You Need for Zero-Shot Robustness in Vision-Language Models 29 Oct 2024 · 1 repository · arXiv:2410.21802Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
Enhancing Zero-Shot Vision Models by Label-Free Prompt Distribution Learning and Bias Correcting 25 Oct 2024 · 0 repositories · arXiv:2410.19294Syntology 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
NeuroClips: Towards High-fidelity and Smooth fMRI-to-Video Reconstruction 25 Oct 2024 · 1 repository · arXiv:2410.19452Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Scene Graph Generation with Role-Playing Large Language Models 20 Oct 2024 · 1 repository · arXiv:2410.15364Syntology official (archive's flag): 2 ran · 2 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; every one of the 2 samples that ran constructed an object rather than computing a result (of 3 harvested samples)
-
LoRA-IR: Taming Low-Rank Experts for Efficient All-in-One Image Restoration 20 Oct 2024 · 1 repository · arXiv:2410.15385Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 8 harvested samples) · 2 pointer-only (licence)
-
IPO: Interpretable Prompt Optimization for Vision-Language Models 20 Oct 2024 · 1 repository · arXiv:2410.15397Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 6 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
BoostAdapter: Improving Vision-Language Test-Time Adaptation via Regional Bootstrapping 20 Oct 2024 · 1 repository · arXiv:2410.15430Syntology official (archive's flag): 4 ran · 5 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 5 where Syntology's instrument failed) · 5 unverified (of 10 harvested samples) · 1 pointer-only (licence)
-
TransAgent: Transfer Vision-Language Foundation Models with Heterogeneous Agent Collaboration 16 Oct 2024 · 1 repository · arXiv:2410.12183Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 8 harvested samples) · 2 pointer-only (licence)
-
Interpreting and Analysing CLIP's Zero-Shot Image Classification via Mutual Knowledge 16 Oct 2024 · 1 repository · arXiv:2410.13016Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Improving Long-Text Alignment for Text-to-Image Diffusion Models 15 Oct 2024 · 1 repository · arXiv:2410.11817Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 6 unverified (of 10 harvested samples) · 1 pointer-only (licence)
-
Mixture of Experts Made Personalized: Federated Prompt Learning for Vision-Language Models 14 Oct 2024 · 1 repository · arXiv:2410.10114Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 2 pointer-only (licence)
-
HART: Efficient Visual Generation with Hybrid Autoregressive Transformer 14 Oct 2024 · 2 repositories · arXiv:2410.10812Syntology official (archive's flag): 23 ran · 23 ran (of which 0 constructed an object rather than computing a result; 15 with no instrument failure: 2 honoured, 0 violated, 13 with no contract checked; 8 where Syntology's instrument failed) · 4 unverified (of 27 harvested samples) · 4 pointer-only (licence)
-
Locality Alignment Improves Vision-Language Models 14 Oct 2024 · 1 repository · arXiv:2410.11087Syntology 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Generating Intermediate Representations for Compositional Text-To-Image Generation 13 Oct 2024 · 1 repository · arXiv:2410.09792Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
TULIP: Token-length Upgraded CLIP 13 Oct 2024 · 1 repository · arXiv:2410.10034Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 1 violated, 8 with no contract checked; 4 where Syntology's instrument failed) · 7 unverified (of 20 harvested samples) · 7 pointer-only (licence)
-
Progressive Autoregressive Video Diffusion Models 10 Oct 2024 · 1 repository · arXiv:2410.08151Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples) · 2 pointer-only (licence)
-
Compositional Entailment Learning for Hyperbolic Vision-Language Models 9 Oct 2024 · 1 repository · arXiv:2410.06912Syntology official (archive's flag): 7 ran · 7 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 5 where Syntology's instrument failed) · 6 unverified (of 13 harvested samples) · 13 pointer-only (licence)
-
SELECT: A Large-Scale Benchmark of Data Curation Strategies for Image Classification 7 Oct 2024 · 1 repository · arXiv:2410.05057Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 13 harvested samples)
-
Preserving Multi-Modal Capabilities of Pre-trained VLMs for Improving Vision-Linguistic Compositionality 7 Oct 2024 · 1 repository · arXiv:2410.05210Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 1 violated, 1 with no contract checked; 7 where Syntology's instrument failed) · 3 unverified (of 13 harvested samples) · 13 pointer-only (licence)
-
Redefining Temporal Modeling in Video Diffusion: The Vectorized Timestep Approach 4 Oct 2024 · 1 repository · arXiv:2410.03160Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 1 honoured, 0 violated, 6 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 10 harvested samples)
-
Investigating and Mitigating Object Hallucinations in Pretrained Vision-Language (CLIP) Models 4 Oct 2024 · 1 repository · arXiv:2410.03176Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
A Retention-Centric Framework for Continual Learning with Guaranteed Model Developmental Safety 4 Oct 2024 · 1 repository · arXiv:2410.03955Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 4 where Syntology's instrument failed) · 4 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Understanding and Mitigating Miscalibration in Prompt Tuning for Vision-Language Models 3 Oct 2024 · 1 repository · arXiv:2410.02681Syntology 4 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 3 where Syntology's instrument failed) · 5 unverified (of 9 harvested samples) · 6 pointer-only (licence)
-
Edge-preserving noise for diffusion models 2 Oct 2024 · 1 repository · arXiv:2410.01540Syntology 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)