Methods › Computer Vision › Vision and Language Pre-Trained Models › CLIP › Papers where code ran, page 2
Contrastive Language-Image Pre-training
CLIP
Papers archive 2025-07-28
archive papers tagged: 3,094 · with a code link: 1,617 · where Syntology ran a sample: 649 (554 with a run with no instrument failure, 95 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (649 of 3,094 tagged: 554 with a run with no instrument failure, 95 where every run was a failure of Syntology's instrument)
Syntology We ran code from the paper's repository; we did not isolate this method inside it.
Page 2 of 7: papers 101 to 200 of the 649 tagged papers where Syntology ran at least one harvested sample (554 with a run with no instrument failure, 95 where every run was a failure of Syntology's instrument), newest first by the archive's date (ties by arXiv id). This is a filter on Syntology's record ordered by date only, not a ranking; a run is not a correctness claim. A paper missing from this list is not a recorded non-run: it may have no arXiv id, no harvested code, or only samples that have not run yet.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
SegEarth-OV: Towards Training-Free Open-Vocabulary Segmentation for Remote Sensing Images 2 Oct 2024 · 2 repositories · arXiv:2410.01768Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 1 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
PointAD: Comprehending 3D Anomalies from Points and Pixels for Zero-shot 3D Anomaly Detection 1 Oct 2024 · 1 repository · arXiv:2410.00320Syntology official (archive's flag): 15 ran · 17 ran (of which 2 constructed an object rather than computing a result; 11 with no instrument failure: 1 honoured, 0 violated, 10 with no contract checked; 6 where Syntology's instrument failed) · 9 unverified (of 26 harvested samples) · 23 pointer-only (licence)
-
Magnet: We Never Know How Text-to-Image Diffusion Models Work, Until We Learn How Vision-Language Models Function 30 Sep 2024 · 1 repository · arXiv:2409.19967Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 1 honoured, 0 violated, 9 with no contract checked; 1 where Syntology's instrument failed) · 6 unverified (of 17 harvested samples) · 17 pointer-only (licence)
-
Federated Learning from Vision-Language Foundation Models: Theoretical Analysis and Method 29 Sep 2024 · 1 repository · arXiv:2409.19610Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 4 where Syntology's instrument failed) · 3 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
CLIP-MoE: Towards Building Mixture of Experts for CLIP with Diversified Multiplet Upcycling 28 Sep 2024 · 1 repository · arXiv:2409.19291Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Emu3: Next-Token Prediction is All You Need 27 Sep 2024 · 2 repositories · arXiv:2409.18869Syntology 6 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
The Hard Positive Truth about Vision-Language Compositionality 26 Sep 2024 · 1 repository · arXiv:2409.17958Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
MultiClimate: Multimodal Stance Detection on Climate Change Videos 26 Sep 2024 · 1 repository · arXiv:2409.18346Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Vision-Language Model Fine-Tuning via Simple Parameter-Efficient Modification 25 Sep 2024 · 1 repository · arXiv:2409.16718Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Attention Prompting on Image for Large Vision-Language Models 25 Sep 2024 · 1 repository · arXiv:2409.17143Syntology official (archive's flag): 6 ran · 8 ran (of which 3 constructed an object rather than computing a result; 5 with no instrument failure: 2 honoured, 0 violated, 3 with no contract checked; 3 where Syntology's instrument failed) · 4 unverified (of 12 harvested samples) · 2 pointer-only (licence)
-
Embedding Geometries of Contrastive Language-Image Pre-Training 19 Sep 2024 · 1 repository · arXiv:2409.13079Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 1 violated, 5 with no contract checked; 3 where Syntology's instrument failed) · 6 unverified (of 15 harvested samples) · 15 pointer-only (licence)
-
Less is More: A Simple yet Effective Token Reduction Method for Efficient Multi-modal LLMs 17 Sep 2024 · 1 repository · arXiv:2409.10994Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 1 violated, 2 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 8 harvested samples) · 1 pointer-only (licence)
-
Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models 16 Sep 2024 · 1 repository · arXiv:2409.10695Syntology 18 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 4 honoured, 1 violated, 4 with no contract checked; 9 where Syntology's instrument failed) · 3 unverified (of 21 harvested samples) · 21 pointer-only (licence)
-
Improving Virtual Try-On with Garment-focused Diffusion Models 12 Sep 2024 · 1 repository · arXiv:2409.08258Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 1 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
DetailCLIP: Detail-Oriented CLIP for Fine-Grained Tasks 10 Sep 2024 · 1 repository · arXiv:2409.06809Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 4 where Syntology's instrument failed) · 5 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
Optimizing CLIP Models for Image Retrieval with Maintained Joint-Embedding Alignment 3 Sep 2024 · 1 repository · arXiv:2409.01936Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 3 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 14 harvested samples) · 11 pointer-only (licence)
-
TempMe: Video Temporal Token Merging for Efficient Text-Video Retrieval 2 Sep 2024 · 1 repository · arXiv:2409.01156Syntology 7 ran (of which 3 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 0 violated, 3 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples)
-
Nemesis: Normalizing the Soft-prompt Vectors of Vision-Language Models 26 Aug 2024 · 1 repository · arXiv:2408.13979Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 3 unverified (of 6 harvested samples) · 3 pointer-only (licence)
-
SwiftBrush v2: Make Your One-step Diffusion Model Better Than Its Teacher 26 Aug 2024 · 1 repository · arXiv:2408.14176Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Social perception of faces in a vision-language model 26 Aug 2024 · 1 repository · arXiv:2408.14435Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples)
-
Generalizable Facial Expression Recognition 20 Aug 2024 · 1 repository · arXiv:2408.10614Syntology official (archive's flag): 5 ran · 5 ran (of which 2 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
C2P-CLIP: Injecting Category Common Prompt in CLIP to Enhance Generalization in Deepfake Detection 19 Aug 2024 · 1 repository · arXiv:2408.09647Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
ReCLIP++: Learn to Rectify the Bias of CLIP for Unsupervised Semantic Segmentation 13 Aug 2024 · 1 repository · arXiv:2408.06747Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
ProxyCLIP: Proxy Attention Improves CLIP for Open-Vocabulary Segmentation 9 Aug 2024 · 1 repository · arXiv:2408.04883Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 1 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Ensemble everything everywhere: Multi-scale aggregation for adversarial robustness 8 Aug 2024 · 2 repositories · arXiv:2408.05446Syntology official: no sample here; runs from other or unrecorded repositories · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
MoExtend: Tuning New Experts for Modality and Task Extension 7 Aug 2024 · 1 repository · arXiv:2408.03511Syntology official (archive's flag): 3 ran · 3 ran (of which 2 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
Exploring Conditional Multi-Modal Prompts for Zero-shot HOI Detection 5 Aug 2024 · 1 repository · arXiv:2408.02484Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 2 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
AdvQDet: Detecting Query-Based Adversarial Attacks with Adversarial Contrastive Prompt Tuning 4 Aug 2024 · 1 repository · arXiv:2408.01978Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 1 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
VAR-CLIP: Text-to-Image Generator with Visual Auto-Regressive Modeling 2 Aug 2024 · 1 repository · arXiv:2408.01181Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 5 where Syntology's instrument failed) · 4 unverified (of 13 harvested samples) · 13 pointer-only (licence)
-
Focus, Distinguish, and Prompt: Unleashing CLIP for Efficient and Flexible Scene Text Retrieval 1 Aug 2024 · 1 repository · arXiv:2408.00441Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Diffusion Feedback Helps CLIP See Better 29 Jul 2024 · 1 repository · arXiv:2407.20171Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples)
-
Contrasting Deepfakes Diffusion via Contrastive Learning and Global-Local Similarities 29 Jul 2024 · 1 repository · arXiv:2407.20337Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples)
-
Adversarial Robustification via Text-to-Image Diffusion Models 26 Jul 2024 · 1 repository · arXiv:2407.18658Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples) · 4 pointer-only (licence)
-
Multi-label Cluster Discrimination for Visual Representation Learning 24 Jul 2024 · 1 repository · arXiv:2407.17331Syntology official (archive's flag): 7 ran · 7 ran (of which 7 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified; every one of the 7 samples that ran constructed an object rather than computing a result (of 11 harvested samples)
-
SEDS: Semantically Enhanced Dual-Stream Encoder for Sign Language Retrieval 23 Jul 2024 · 1 repository · arXiv:2407.16394Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
Category-Extensible Out-of-Distribution Detection via Hierarchical Context Descriptions 23 Jul 2024 · 1 repository · arXiv:2407.16725Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 2 pointer-only (licence)
-
MLLM-CompBench: A Comparative Reasoning Benchmark for Multimodal LLMs 23 Jul 2024 · 1 repository · arXiv:2407.16837Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
CLIP with Generative Latent Replay: a Strong Baseline for Incremental Learning 22 Jul 2024 · 1 repository · arXiv:2407.15793Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 8 harvested samples) · 2 pointer-only (licence)
-
AdaCLIP: Adapting CLIP with Hybrid Learnable Prompts for Zero-Shot Anomaly Detection 22 Jul 2024 · 1 repository · arXiv:2407.15795Syntology official (archive's flag): 20 ran · 20 ran (of which 10 constructed an object rather than computing a result; 16 with no instrument failure: 0 honoured, 1 violated, 15 with no contract checked; 4 where Syntology's instrument failed) · 8 unverified (of 28 harvested samples) · 4 pointer-only (licence)
-
Prior Knowledge Integration via LLM Encoding and Pseudo Event Regulation for Video Moment Retrieval 21 Jul 2024 · 1 repository · arXiv:2407.15051Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 1 violated, 4 with no contract checked; 3 where Syntology's instrument failed) · 3 unverified (of 11 harvested samples) · 9 pointer-only (licence)
-
Class-Incremental Learning with CLIP: Adaptive Representation Adjustment and Parameter Fusion 19 Jul 2024 · 1 repository · arXiv:2407.14143Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Robust Calibration of Large Vision-Language Adapters 18 Jul 2024 · 1 repository · arXiv:2407.13588Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample)
-
IMAGDressing-v1: Customizable Virtual Dressing 17 Jul 2024 · 1 repository · arXiv:2407.12705Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 1 pointer-only (licence)
-
LaMI-DETR: Open-Vocabulary Detection with Language Model Instruction 16 Jul 2024 · 1 repository · arXiv:2407.11335Syntology official (archive's flag): 3 ran · 3 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
Unlearning Targeted Information via Single Layer Unlearning Gradient 16 Jul 2024 · 1 repository · arXiv:2407.11867Syntology official (archive's flag): 16 ran · 16 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 0 violated, 12 with no contract checked; 4 where Syntology's instrument failed) · 8 unverified (of 24 harvested samples) · 24 pointer-only (licence)
-
Accessing Vision Foundation Models at ImageNet-level Costs 15 Jul 2024 · 1 repository · arXiv:2407.10366Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 6 unverified (of 10 harvested samples)
-
DataDream: Few-shot Guided Dataset Generation 15 Jul 2024 · 2 repositories · arXiv:2407.10910Syntology official (archive's flag): 17 ran · 17 ran (of which 0 constructed an object rather than computing a result; 16 with no instrument failure: 1 honoured, 0 violated, 15 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (of 21 harvested samples) · 21 pointer-only (licence)
-
CLIP-Guided Generative Networks for Transferable Targeted Adversarial Attacks 14 Jul 2024 · 1 repository · arXiv:2407.10179Syntology official (archive's flag): 2 ran · 3 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
LAPT: Label-driven Automated Prompt Tuning for OOD Detection with Vision-Language Models 12 Jul 2024 · 2 repositories · arXiv:2407.08966Syntology community repositories only · 15 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 2 violated, 9 with no contract checked; 4 where Syntology's instrument failed) · 2 unverified (of 17 harvested samples) · 4 pointer-only (licence)
-
Explore the Potential of CLIP for Training-Free Open Vocabulary Semantic Segmentation 11 Jul 2024 · 1 repository · arXiv:2407.08268Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples)
-
Emergent Visual-Semantic Hierarchies in Image-Text Representations 11 Jul 2024 · 1 repository · arXiv:2407.08521Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
CoLA: Conditional Dropout and Language-driven Robust Dual-modal Salient Object Detection 9 Jul 2024 · 1 repository · arXiv:2407.06780Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 5 where Syntology's instrument failed) · 3 unverified (of 11 harvested samples) · 4 pointer-only (licence)
-
Fine-Tuning Attention Modules Only: Enhancing Weight Disentanglement in Task Arithmetic 9 Jul 2024 · 2 repositories · arXiv:2407.07089Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples)
-
AWT: Transferring Vision-Language Models via Augmentation, Weighting, and Transportation 5 Jul 2024 · 1 repository · arXiv:2407.04603Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Boosting Consistency in Story Visualization with Rich-Contextual Conditional Diffusion Models 2 Jul 2024 · 1 repository · arXiv:2407.02482Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
GalLoP: Learning Global and Local Prompts for Vision-Language Models 1 Jul 2024 · 1 repository · arXiv:2407.01400Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
EVF-SAM: Early Vision-Language Fusion for Text-Prompted Segment Anything Model 28 Jun 2024 · 1 repository · arXiv:2406.20076Syntology official (archive's flag): 1 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 2 pointer-only (licence)
-
A Sanity Check for AI-generated Image Detection 27 Jun 2024 · 2 repositories · arXiv:2406.19435Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 8 harvested samples) · 3 pointer-only (licence)
-
Mitigate the Gap: Investigating Approaches for Improving Cross-Modal Alignment in CLIP 25 Jun 2024 · 1 repository · arXiv:2406.17639Syntology official (archive's flag): 16 ran · 16 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 3 violated, 7 with no contract checked; 6 where Syntology's instrument failed) · 6 unverified (of 22 harvested samples) · 22 pointer-only (licence)
-
BioTrove: A Large Curated Image Dataset Enabling AI for Biodiversity 25 Jun 2024 · 2 repositories · arXiv:2406.17720Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
CLIP-Decoder : ZeroShot Multilabel Classification using Multimodal CLIP Aligned Representation 21 Jun 2024 · 1 repository · arXiv:2406.14830Syntology 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
African or European Swallow? Benchmarking Large Vision-Language Models for Fine-Grained Object Classification 20 Jun 2024 · 1 repository · arXiv:2406.14496Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples)
-
WATT: Weight Average Test-Time Adaptation of CLIP 19 Jun 2024 · 1 repository · arXiv:2406.13875Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 5 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples) · 3 pointer-only (licence)
-
Efficient and Long-Tailed Generalization for Pre-trained Vision-Language Model 18 Jun 2024 · 1 repository · arXiv:2406.12638Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Dissecting Adversarial Robustness of Multimodal LM Agents 18 Jun 2024 · 1 repository · arXiv:2406.12814Syntology official (archive's flag): 16 ran · 16 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 5 where Syntology's instrument failed) · 5 unverified (of 21 harvested samples) · 1 pointer-only (licence)
-
Frozen CLIP: A Strong Backbone for Weakly Supervised Semantic Segmentation 17 Jun 2024 · 1 repository · arXiv:2406.11189Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Visually Consistent Hierarchical Image Classification 17 Jun 2024 · 0 repositories · arXiv:2406.11608Syntology 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Not All Prompts Are Made Equal: Prompt-based Pruning of Text-to-Image Diffusion Models 17 Jun 2024 · 1 repository · arXiv:2406.12042Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples)
-
Open-Vocabulary Semantic Segmentation with Image Embedding Balancing 14 Jun 2024 · 1 repository · arXiv:2406.09829Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 12 harvested samples)
-
Understanding and Mitigating Compositional Issues in Text-to-Image Generative Models 12 Jun 2024 · 1 repository · arXiv:2406.07844Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
AsyncDiff: Parallelizing Diffusion Models by Asynchronous Denoising 11 Jun 2024 · 2 repositories · arXiv:2406.06911Syntology official (archive's flag): 3 ran · 5 ran (of which 1 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples)
-
RWKV-CLIP: A Robust Vision-Language Representation Learner 11 Jun 2024 · 2 repositories · arXiv:2406.06973Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 4 where Syntology's instrument failed) · 7 unverified (of 14 harvested samples) · 3 pointer-only (licence)
-
Let Go of Your Labels with Unsupervised Transfer 11 Jun 2024 · 1 repository · arXiv:2406.07236Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Vision Model Pre-training on Interleaved Image-Text Data via Latent Compression Learning 11 Jun 2024 · 1 repository · arXiv:2406.07543Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 1 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 6 unverified (of 12 harvested samples) · 8 pointer-only (licence)
-
Vript: A Video Is Worth Thousands of Words 10 Jun 2024 · 1 repository · arXiv:2406.06040Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
TLCM: Training-efficient Latent Consistency Model for Image Generation with 2-8 Steps 9 Jun 2024 · 1 repository · arXiv:2406.05768Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Visual-Text Cross Alignment: Refining the Similarity Score in Vision-Language Models 5 Jun 2024 · 1 repository · arXiv:2406.02915Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 8 where Syntology's instrument failed) · 5 unverified (of 13 harvested samples) · 3 pointer-only (licence)
-
CSS: Contrastive Semantic Similarity for Uncertainty Quantification of LLMs 5 Jun 2024 · 1 repository · arXiv:2406.03158Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
Open-YOLO 3D: Towards Fast and Accurate Open-Vocabulary 3D Instance Segmentation 4 Jun 2024 · 1 repository · arXiv:2406.02548Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 7 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
Long and Short Guidance in Score identity Distillation for One-Step Text-to-Image Generation 3 Jun 2024 · 2 repositories · arXiv:2406.01561Syntology official (archive's flag): 17 ran · 19 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 3 honoured, 0 violated, 10 with no contract checked; 6 where Syntology's instrument failed) · 5 unverified (of 24 harvested samples) · 7 pointer-only (licence)
-
Decomposing and Interpreting Image Representations via Text in ViTs Beyond CLIP 3 Jun 2024 · 1 repository · arXiv:2406.01583Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 1 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 13 harvested samples) · 13 pointer-only (licence)
-
Cascade-CLIP: Cascaded Vision-Language Embeddings Alignment for Zero-Shot Semantic Segmentation 2 Jun 2024 · 1 repository · arXiv:2406.00670Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 1 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples) · 1 pointer-only (licence)
-
Envisioning Outlier Exposure by Large Language Models for Out-of-Distribution Detection 2 Jun 2024 · 1 repository · arXiv:2406.00806Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
DeCoOp: Robust Prompt Tuning with Out-of-Distribution Detection 1 Jun 2024 · 1 repository · arXiv:2406.00345Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 3 where Syntology's instrument failed) · 3 unverified (of 9 harvested samples) · 4 pointer-only (licence)
-
MegActor: Harness the Power of Raw Video for Vivid Portrait Animation 31 May 2024 · 2 repositories · arXiv:2405.20851Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 17 harvested samples) · 2 pointer-only (licence)
-
What Makes CLIP More Robust to Long-Tailed Pre-Training Data? A Controlled Study for Transferable Insights 31 May 2024 · 1 repository · arXiv:2405.21070Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 7 unverified (of 12 harvested samples) · 1 pointer-only (licence)
-
Enhancing Zero-Shot Facial Expression Recognition by LLM Knowledge Transfer 29 May 2024 · 1 repository · arXiv:2405.19100Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 7 unverified (of 10 harvested samples) · 3 pointer-only (licence)
-
CLIPLoss and Norm-Based Data Selection Methods for Multimodal Contrastive Learning 29 May 2024 · 2 repositories · arXiv:2405.19547Syntology official (archive's flag): 14 ran · 14 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 0 honoured, 3 violated, 10 with no contract checked; 1 where Syntology's instrument failed) · 5 unverified (of 19 harvested samples) · 19 pointer-only (licence)
-
Why are Visually-Grounded Language Models Bad at Image Classification? 28 May 2024 · 1 repository · arXiv:2405.18415Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Benchmarking and Improving Bird's Eye View Perception Robustness in Autonomous Driving 27 May 2024 · 1 repository · arXiv:2405.17426Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Accelerating Transformers with Spectrum-Preserving Token Merging 25 May 2024 · 1 repository · arXiv:2405.16148Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
Learning Invariant Causal Mechanism from Vision-Language Models 24 May 2024 · 0 repositories · arXiv:2405.15289Syntology 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
CLIPScope: Enhancing Zero-Shot OOD Detection with Bayesian Scoring 23 May 2024 · 1 repository · arXiv:2405.14737Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
GMMFormer v2: An Uncertainty-aware Framework for Partially Relevant Video Retrieval 22 May 2024 · 1 repository · arXiv:2405.13824Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
TOPA: Extending Large Language Models for Video Understanding via Text-Only Pre-Alignment 22 May 2024 · 1 repository · arXiv:2405.13911Syntology official (archive's flag): 3 ran · 6 ran (of which 3 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 1 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 9 harvested samples) · 3 pointer-only (licence)
-
Harmonizing Generalization and Personalization in Federated Prompt Learning 16 May 2024 · 1 repository · arXiv:2405.09771Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 4 where Syntology's instrument failed) · 7 unverified (of 14 harvested samples) · 14 pointer-only (licence)
-
DiffAM: Diffusion-based Adversarial Makeup Transfer for Facial Privacy Protection 16 May 2024 · 2 repositories · arXiv:2405.09882Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 14 harvested samples) · 14 pointer-only (licence)
-
Pre-trained Text-to-Image Diffusion Models Are Versatile Representation Learners for Control 9 May 2024 · 1 repository · arXiv:2405.05852Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
OpenESS: Event-based Semantic Scene Understanding with Open Vocabularies 8 May 2024 · 1 repository · arXiv:2405.05259Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
iSEARLE: Improving Textual Inversion for Zero-Shot Composed Image Retrieval 5 May 2024 · 2 repositories · arXiv:2405.02951Syntology official: no sample here; runs from other or unrecorded repositories · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)