Methods › Computer Vision › Vision and Language Pre-Trained Models › CLIP › Papers, page 13
Contrastive Language-Image Pre-training
CLIP
Papers archive 2025-07-28
archive papers tagged: 3,094 · with a code link: 1,617 · where Syntology ran a sample: 649 (554 with a run with no instrument failure, 95 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (649 of 3,094 tagged: 554 with a run with no instrument failure, 95 where every run was a failure of Syntology's instrument)
Page 13 of 31: papers 1,201 to 1,300 of 3,094, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
WATT: Weight Average Test-Time Adaptation of CLIP 19 Jun 2024 · 1 repository · arXiv:2406.13875Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 5 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples) · 3 pointer-only (licence)
-
Dissecting Adversarial Robustness of Multimodal LM Agents 18 Jun 2024 · 1 repository · arXiv:2406.12814Syntology official (archive's flag): 16 ran · 16 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 5 where Syntology's instrument failed) · 5 unverified (of 21 harvested samples) · 1 pointer-only (licence)
-
Efficient and Long-Tailed Generalization for Pre-trained Vision-Language Model 18 Jun 2024 · 1 repository · arXiv:2406.12638Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Multilingual Synopses of Movie Narratives: A Dataset for Vision-Language Story Understanding 18 Jun 2024 · 1 repository · arXiv:2406.13092
-
Symmetric Multi-Similarity Loss for EPIC-KITCHENS-100 Multi-Instance Retrieval Challenge 2024 18 Jun 2024 · 1 repository · arXiv:2406.12256
-
BaFTA: Backprop-Free Test-Time Adaptation For Zero-Shot Vision-Language Models 17 Jun 2024 · 0 repositories · arXiv:2406.11309
-
DistillNeRF: Perceiving 3D Scenes from Single-Glance Images by Distilling Neural Fields and Foundation Model Features 17 Jun 2024 · 0 repositories · arXiv:2406.12095
-
Duoduo CLIP: Efficient 3D Understanding with Multi-View Images 17 Jun 2024 · 1 repository · arXiv:2406.11579Syntology official: harvested, nothing ran · 0 ran · 2 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Exploring the Role of Large Language Models in Prompt Encoding for Diffusion Models 17 Jun 2024 · 0 repositories · arXiv:2406.11831
-
Frozen CLIP: A Strong Backbone for Weakly Supervised Semantic Segmentation 17 Jun 2024 · 1 repository · arXiv:2406.11189Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Hallucination Mitigation Prompts Long-term Video Understanding 17 Jun 2024 · 0 repositories · arXiv:2406.11333
-
Visually Consistent Hierarchical Image Classification 17 Jun 2024 · 0 repositories · arXiv:2406.11608Syntology 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Mining Open Semantics from CLIP: A Relation Transition Perspective for Few-Shot Learning 17 Jun 2024 · 0 repositories · arXiv:2406.11252
-
Not All Prompts Are Made Equal: Prompt-based Pruning of Text-to-Image Diffusion Models 17 Jun 2024 · 1 repository · arXiv:2406.12042Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples)
-
They're All Doctors: Synthesizing Diverse Counterfactuals to Mitigate Associative Bias 17 Jun 2024 · 0 repositories · arXiv:2406.11331
-
VideoLLM-online: Online Video Large Language Model for Streaming Video 17 Jun 2024 · 0 repositories · arXiv:2406.11816
-
Exploiting Diffusion Prior for Out-of-Distribution Detection 16 Jun 2024 · 0 repositories · arXiv:2406.11105
-
Open-Vocabulary X-ray Prohibited Item Detection via Fine-tuning CLIP 16 Jun 2024 · 0 repositories · arXiv:2406.10961
-
Reconsidering Sentence-Level Sign Language Translation 16 Jun 2024 · 0 repositories · arXiv:2406.11049
-
Enhancing Anomaly Detection Generalization through Knowledge Exposure: The Dual Effects of Augmentation 15 Jun 2024 · 1 repository · arXiv:2406.10617
-
Enhancing Fake News Detection in Social Media via Label Propagation on Cross-modal Tweet Graph 14 Jun 2024 · 1 repository · arXiv:2406.09884
-
Open-Vocabulary Semantic Segmentation with Image Embedding Balancing 14 Jun 2024 · 1 repository · arXiv:2406.09829Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 12 harvested samples)
-
Adaptive Temporal Motion Guided Graph Convolution Network for Micro-expression Recognition 13 Jun 2024 · 0 repositories · arXiv:2406.08997
-
CLIPAway: Harmonizing Focused Embeddings for Removing Objects via Diffusion Models 13 Jun 2024 · 1 repository · arXiv:2406.09368Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Exploring the Spectrum of Visio-Linguistic Compositionality and Recognition 13 Jun 2024 · 1 repository · arXiv:2406.09388
-
Interpreting the structure of multi-object representations in vision encoders 13 Jun 2024 · 0 repositories · arXiv:2406.09067
-
An Efficient Post-hoc Framework for Reducing Task Discrepancy of Text Encoders for Composed Image Retrieval 13 Jun 2024 · 1 repository · arXiv:2406.09188
-
From a Social Cognitive Perspective: Context-aware Visual Social Relationship Recognition 12 Jun 2024 · 0 repositories · arXiv:2406.08358
-
Understanding and Mitigating Compositional Issues in Text-to-Image Generative Models 12 Jun 2024 · 1 repository · arXiv:2406.07844Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Updating CLIP to Prefer Descriptions Over Captions 12 Jun 2024 · 1 repository · arXiv:2406.09458
-
What If We Recaption Billions of Web Images with LLaMA-3? 12 Jun 2024 · 0 repositories · arXiv:2406.08478
-
Words Worth a Thousand Pictures: Measuring and Understanding Perceptual Variability in Text-to-Image Generation 12 Jun 2024 · 0 repositories · arXiv:2406.08482
-
AsyncDiff: Parallelizing Diffusion Models by Asynchronous Denoising 11 Jun 2024 · 2 repositories · arXiv:2406.06911Syntology official (archive's flag): 3 ran · 5 ran (of which 1 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples)
-
Bridging Language Gaps in Audio-Text Retrieval 11 Jun 2024 · 1 repository · arXiv:2406.07012
-
Let Go of Your Labels with Unsupervised Transfer 11 Jun 2024 · 1 repository · arXiv:2406.07236Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
RWKV-CLIP: A Robust Vision-Language Representation Learner 11 Jun 2024 · 2 repositories · arXiv:2406.06973Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 4 where Syntology's instrument failed) · 7 unverified (of 14 harvested samples) · 3 pointer-only (licence)
-
UVIS: Unsupervised Video Instance Segmentation 11 Jun 2024 · 0 repositories · arXiv:2406.06908
-
Vision Model Pre-training on Interleaved Image-Text Data via Latent Compression Learning 11 Jun 2024 · 1 repository · arXiv:2406.07543Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 1 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 6 unverified (of 12 harvested samples) · 8 pointer-only (licence)
-
Vript: A Video Is Worth Thousands of Words 10 Jun 2024 · 1 repository · arXiv:2406.06040Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
Set-CLIP: Exploring Aligned Semantic From Low-Alignment Multimodal Data Through A Distribution View 9 Jun 2024 · 0 repositories · arXiv:2406.05766
-
TLCM: Training-efficient Latent Consistency Model for Image Generation with 2-8 Steps 9 Jun 2024 · 1 repository · arXiv:2406.05768Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
VP-LLM: Text-Driven 3D Volume Completion with Large Language Models through Patchification 8 Jun 2024 · 0 repositories · arXiv:2406.05543
-
3rd Place Solution for MeViS Track in CVPR 2024 PVUW workshop: Motion Expression guided Video Segmentation 7 Jun 2024 · 0 repositories · arXiv:2406.04842
-
CoNo: Consistency Noise Injection for Tuning-free Long Video Diffusion 7 Jun 2024 · 0 repositories · arXiv:2406.05082
-
A Survey on 3D Human Avatar Modeling -- From Reconstruction to Generation 6 Jun 2024 · 0 repositories · arXiv:2406.04253
-
Attribute-Aware Implicit Modality Alignment for Text Attribute Person Search 6 Jun 2024 · 0 repositories · arXiv:2406.03721
-
GenAI Arena: An Open Evaluation Platform for Generative Models 6 Jun 2024 · 1 repository · arXiv:2406.04485
-
Interpreting the Second-Order Effects of Neurons in CLIP 6 Jun 2024 · 0 repositories · arXiv:2406.04341
-
VISTA: Visualized Text Embedding For Universal Multi-Modal Retrieval 6 Jun 2024 · 1 repository · arXiv:2406.04292Syntology official: harvested, nothing ran · 0 ran · 2 unverified (of 2 harvested samples)
-
Alignment Calibration: Machine Unlearning for Contrastive Learning under Auditing 5 Jun 2024 · 0 repositories · arXiv:2406.03603
-
CountCLIP -- [Re] Teaching CLIP to Count to Ten 5 Jun 2024 · 1 repository · arXiv:2406.03586
-
CSS: Contrastive Semantic Similarity for Uncertainty Quantification of LLMs 5 Jun 2024 · 1 repository · arXiv:2406.03158Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
Exploiting LMM-based knowledge for image classification tasks 5 Jun 2024 · 0 repositories · arXiv:2406.03071
-
Speech-based Clinical Depression Screening: An Empirical Study 5 Jun 2024 · 0 repositories · arXiv:2406.03510
-
Visual-Text Cross Alignment: Refining the Similarity Score in Vision-Language Models 5 Jun 2024 · 1 repository · arXiv:2406.02915Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 8 where Syntology's instrument failed) · 5 unverified (of 13 harvested samples) · 3 pointer-only (licence)
-
Analyzing the Feature Extractor Networks for Face Image Synthesis 4 Jun 2024 · 1 repository · arXiv:2406.02153
-
No Captions, No Problem: Captionless 3D-CLIP Alignment with Hard Negatives via CLIP Knowledge and LLMs 4 Jun 2024 · 0 repositories · arXiv:2406.02202
-
EUFCC-340K: A Faceted Hierarchical Dataset for Metadata Annotation in GLAM Collections 4 Jun 2024 · 1 repository · arXiv:2406.02380
-
FastLGS: Speeding up Language Embedded Gaussians with Feature Grid Mapping 4 Jun 2024 · 0 repositories · arXiv:2406.01916
-
M3DM-NR: RGB-3D Noisy-Resistant Industrial Anomaly Detection via Multimodal Denoising 4 Jun 2024 · 0 repositories · arXiv:2406.02263
-
Open-YOLO 3D: Towards Fast and Accurate Open-Vocabulary 3D Instance Segmentation 4 Jun 2024 · 1 repository · arXiv:2406.02548Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 7 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
OpenGaussian: Towards Point-Level 3D Gaussian-based Open Vocabulary Understanding 4 Jun 2024 · 0 repositories · arXiv:2406.02058
-
ProGEO: Generating Prompts through Image-Text Contrastive Learning for Visual Geo-localization 4 Jun 2024 · 1 repository · arXiv:2406.01906
-
Advancing Weakly-Supervised Audio-Visual Video Parsing via Segment-wise Pseudo Labeling 3 Jun 2024 · 0 repositories · arXiv:2406.00919
-
Decomposing and Interpreting Image Representations via Text in ViTs Beyond CLIP 3 Jun 2024 · 1 repository · arXiv:2406.01583Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 1 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 13 harvested samples) · 13 pointer-only (licence)
-
Learning Temporally Consistent Video Depth from Video Diffusion Priors 3 Jun 2024 · 0 repositories · arXiv:2406.01493
-
Long and Short Guidance in Score identity Distillation for One-Step Text-to-Image Generation 3 Jun 2024 · 2 repositories · arXiv:2406.01561Syntology official (archive's flag): 17 ran · 19 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 3 honoured, 0 violated, 10 with no contract checked; 6 where Syntology's instrument failed) · 5 unverified (of 24 harvested samples) · 7 pointer-only (licence)
-
MLIP: Efficient Multi-Perspective Language-Image Pretraining with Exhaustive Data Utilization 3 Jun 2024 · 0 repositories · arXiv:2406.01460
-
pOps: Photo-Inspired Diffusion Operators 3 Jun 2024 · 0 repositories · arXiv:2406.01300
-
Zero-Shot Out-of-Distribution Detection with Outlier Label Exposure 3 Jun 2024 · 1 repository · arXiv:2406.01170
-
Cascade-CLIP: Cascaded Vision-Language Embeddings Alignment for Zero-Shot Semantic Segmentation 2 Jun 2024 · 1 repository · arXiv:2406.00670Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 1 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples) · 1 pointer-only (licence)
-
Envisioning Outlier Exposure by Large Language Models for Out-of-Distribution Detection 2 Jun 2024 · 1 repository · arXiv:2406.00806Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
DeCoOp: Robust Prompt Tuning with Out-of-Distribution Detection 1 Jun 2024 · 1 repository · arXiv:2406.00345Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 3 where Syntology's instrument failed) · 3 unverified (of 9 harvested samples) · 4 pointer-only (licence)
-
Effectiveness of Vision Language Models for Open-world Single Image Test Time Adaptation 1 Jun 2024 · 0 repositories · arXiv:2406.00481
-
The Curious Case of End Token: A Zero-Shot Disentangled Image Editing using CLIP 1 Jun 2024 · 0 repositories · arXiv:2406.00457
-
Enhancing Vision Models for Text-Heavy Content Understanding and Interaction 31 May 2024 · 0 repositories · arXiv:2405.20906
-
What Makes CLIP More Robust to Long-Tailed Pre-Training Data? A Controlled Study for Transferable Insights 31 May 2024 · 1 repository · arXiv:2405.21070Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 7 unverified (of 12 harvested samples) · 1 pointer-only (licence)
-
MegActor: Harness the Power of Raw Video for Vivid Portrait Animation 31 May 2024 · 2 repositories · arXiv:2405.20851Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 17 harvested samples) · 2 pointer-only (licence)
-
Jina CLIP: Your CLIP Model Is Also Your Text Retriever 30 May 2024 · 0 repositories · arXiv:2405.20204
-
KerasCV and KerasNLP: Vision and Language Power-Ups 30 May 2024 · 0 repositories · arXiv:2405.20247
-
Learning Robust Correlation with Foundation Model for Weakly-Supervised Few-Shot Segmentation 30 May 2024 · 0 repositories · arXiv:2405.19638
-
RTGen: Generating Region-Text Pairs for Open-Vocabulary Object Detection 30 May 2024 · 1 repository · arXiv:2405.19854
-
CLIPLoss and Norm-Based Data Selection Methods for Multimodal Contrastive Learning 29 May 2024 · 2 repositories · arXiv:2405.19547Syntology official (archive's flag): 14 ran · 14 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 0 honoured, 3 violated, 10 with no contract checked; 1 where Syntology's instrument failed) · 5 unverified (of 19 harvested samples) · 19 pointer-only (licence)
-
Correctable Landmark Discovery via Large Models for Vision-Language Navigation 29 May 2024 · 1 repository · arXiv:2405.18721
-
Enhancing Vision-Language Model with Unmasked Token Alignment 29 May 2024 · 1 repository · arXiv:2405.19009
-
Enhancing Zero-Shot Facial Expression Recognition by LLM Knowledge Transfer 29 May 2024 · 1 repository · arXiv:2405.19100Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 7 unverified (of 10 harvested samples) · 3 pointer-only (licence)
-
I Bet You Did Not Mean That: Testing Semantic Importance via Betting 29 May 2024 · 1 repository · arXiv:2405.19146
-
LLM-based Hierarchical Concept Decomposition for Interpretable Fine-Grained Image Classification 29 May 2024 · 0 repositories · arXiv:2405.18672
-
Parameter-efficient Fine-tuning in Hyperspherical Space for Open-vocabulary Semantic Segmentation 29 May 2024 · 0 repositories · arXiv:2405.18840
-
RAP: Efficient Text-Video Retrieval with Sparse-and-Correlated Adapter 29 May 2024 · 0 repositories · arXiv:2405.19465
-
Topological Perspectives on Optimal Multimodal Embedding Spaces 29 May 2024 · 0 repositories · arXiv:2405.18867
-
3DitScene: Editing Any Scene via Language-guided Disentangled Gaussian Splatting 28 May 2024 · 0 repositories · arXiv:2405.18424
-
It's Not a Modality Gap: Characterizing and Addressing the Contrastive Gap 28 May 2024 · 0 repositories · arXiv:2405.18570
-
Multi-modal Generation via Cross-Modal In-Context Learning 28 May 2024 · 1 repository · arXiv:2405.18304
-
The Impacts of Data, Ordering, and Intrinsic Dimensionality on Recall in Hierarchical Navigable Small Worlds 28 May 2024 · 0 repositories · arXiv:2405.17813
-
Why are Visually-Grounded Language Models Bad at Image Classification? 28 May 2024 · 1 repository · arXiv:2405.18415Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
WIDIn: Wording Image for Domain-Invariant Representation in Single-Source Domain Generalization 28 May 2024 · 0 repositories · arXiv:2405.18405
-
Benchmarking and Improving Bird's Eye View Perception Robustness in Autonomous Driving 27 May 2024 · 1 repository · arXiv:2405.17426Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Synergy and Diversity in CLIP: Enhancing Performance Through Adaptive Backbone Ensembling 27 May 2024 · 0 repositories · arXiv:2405.17139
-
TIMA: Text-Image Mutual Awareness for Balancing Zero-Shot Adversarial Robustness and Generalization Ability 27 May 2024 · 0 repositories · arXiv:2405.17678