| Bridging Vision Foundation Model Priors with CLIP for Spatial-aware Few-shot Anomaly Detection in Medical Images added by Syntology |
2026-09 (from id) |
JuzhengMiao/Spatial-FAD/open_clip_SpatialFAD/coca_model.py fb651d0a97fd4d3f |
unverified |
MIT (permissive) |
| Powering Up Zeroth-Order Training via Subspace Gradient Orthogonalization added by Syntology |
2026-02 (from id) |
OPTML-Group/ZO-Muon/llm/prefix.py 84920a8095dca4fa |
unverified |
MIT (permissive) |
| Expert Knowledge-Guided Decision Calibration for Accurate Fine-Grained Tree Species Classification added by Syntology |
2026-01 (from id) |
WHU-USI3DV/TreeCLS/models/open_clip/coca_model.py fb651d0a97fd4d3f |
unverified |
no licence file found · pointer only |
| arXiv:2507.12998 |
2025-07 (from id) |
MediaBrain-SJTU/DISSect/src/open_clip/coca_model.py fb651d0a97fd4d3f |
unverified |
no licence file found · pointer only |
| MedITok: A Unified Tokenizer for Medical Image Synthesis and Interpretation |
25 May 2025 |
masaaki-75/meditok/local_openclip/coca_model.py fb651d0a97fd4d3f |
unverified |
MIT (permissive) |
| SimpleAR: Pushing the Frontier of Autoregressive Visual Generation through Pretraining, SFT, and RL |
15 Apr 2025 |
wdrink/simplear/hpsv2/src/open_clip/coca_model.py fb651d0a97fd4d3f |
unverified |
MIT (permissive) |
| GeoLangBind: Unifying Earth Observation with Agglomerative Vision-Language Foundation Models |
8 Mar 2025 |
xiong-zhitong/geolb-siglip/open_clip/src/open_clip/coca_model.py fb651d0a97fd4d3f |
unverified |
Apache-2.0 (permissive) |
| GEM: Empowering MLLM for Grounded ECG Understanding with Time Series and Images |
8 Mar 2025 |
lanxiang1017/gem/ecg_coca/open_clip/coca_model.py af7bd4078c114f3d |
ran · fixture could not drive it
|
Apache-2.0 (permissive) |
| Hidden in the Noise: Two-Stage Robust Watermarking for Images |
5 Dec 2024 |
Kasraarabi/Hidden-in-the-Noise/open_clip/coca_model.py fb651d0a97fd4d3f |
unverified |
no licence file found · pointer only |
| CorrCLIP: Reconstructing Correlations in CLIP with Off-the-Shelf Foundation Models for Open-Vocabulary Semantic Segmentation |
15 Nov 2024 |
zdk258/CorrCLIP/CorrCLIPv1/open_clip/coca_model.py fb651d0a97fd4d3f |
unverified |
no licence file found · pointer only |
| On Erroneous Agreements of CLIP Image Embeddings |
7 Nov 2024 |
lst627/CLIP-Embeds/open_clip/src/open_clip/coca_model.py fb651d0a97fd4d3f |
unverified |
no licence file found · pointer only |
| ROBIN: Robust and Invisible Watermarks for Diffusion Models with Adversarial Optimization |
6 Nov 2024 |
Hannah1102/ROBIN/open_clip/coca_model.py fb651d0a97fd4d3f |
unverified |
no licence file found · pointer only |
| What If the Input is Expanded in OOD Detection? |
24 Oct 2024 |
tmlr-group/CoVer/open_clip/coca_model.py fb651d0a97fd4d3f |
unverified |
no licence file found · pointer only |
| TULIP: Token-length Upgraded CLIP |
13 Oct 2024 |
ivonajdenkoska/tulip/open_clip/coca_model.py fb651d0a97fd4d3f |
unverified |
Apache-2.0 (permissive) |
| Zeroth-Order Fine-Tuning of LLMs in Random Subspaces |
11 Oct 2024 |
zimingyy/subzero/large_models/prefix_tuning.py 84920a8095dca4fa |
unverified |
GPL-3.0 (copyleft) · pointer only |
| SePPO: Semi-Policy Preference Optimization for Diffusion Alignment |
7 Oct 2024 |
dwanzhang-ai/seppo/utils/open_clip/coca_model.py fb651d0a97fd4d3f |
unverified |
no licence file found · pointer only |
| SegEarth-OV: Towards Training-Free Open-Vocabulary Segmentation for Remote Sensing Images |
2 Oct 2024 |
likyoo/SegEarth-OV/open_clip/coca_model.py fb651d0a97fd4d3f |
unverified |
no licence file found · pointer only |
| Embedding Geometries of Contrastive Language-Image Pre-Training |
19 Sep 2024 |
eify/open_clip/src/open_clip/coca_model.py fb651d0a97fd4d3f |
unverified |
no licence file found · pointer only |
| ECG-Chat: A Large ECG-Language Model for Cardiac Disease Diagnosis |
16 Aug 2024 |
YubaoZhao/ECG-Chat/open_clip/open_clip/coca_model.py 82c3cd0a861f967a |
ran
|
no licence file found · pointer only |
| ProxyCLIP: Proxy Attention Improves CLIP for Open-Vocabulary Segmentation |
9 Aug 2024 |
mc-lan/proxyclip/open_clip/coca_model.py fb651d0a97fd4d3f |
unverified |
no licence file found · pointer only |
| VSSD: Vision Mamba with Non-Causal State Space Duality |
26 Jul 2024 |
YuHengsss/Trident/open_clip/coca_model.py fb651d0a97fd4d3f |
unverified |
Apache-2.0 (permissive) |
| Audio Prompt Adapter: Unleashing Music Editing Abilities for Text-to-Music with Lightweight Finetuning |
23 Jul 2024 |
fundwotsai2001/ap-adapter/pipeline/pipeline_audioldm2.py d5a09007aef6b7ff |
ran
|
licence not identified · pointer only |
| Vision Model Pre-training on Interleaved Image-Text Data via Latent Compression Learning |
11 Jun 2024 |
opengvlab/lcl/src/open_clip/coca_model.py fb651d0a97fd4d3f |
unverified |
MIT (permissive) |
| VideoLLaMA 2: Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs |
11 Jun 2024 |
damo-nlp-sg/inf-clip/inf_clip/models/coca_arch.py fb651d0a97fd4d3f |
unverified |
Apache-2.0 (permissive) |
| Order-Independence Without Fine Tuning |
4 Jun 2024 |
reidmcy/set-based-prompting/set_based_prompting/input_processing.py 47fbcb27d06f5149 |
ran · fixture could not drive it
|
MIT (permissive) |
| Efficient Vision-Language Pre-training by Cluster Masking |
14 May 2024 |
zi-hao-wei/efficient-vision-language-pre-training-by-cluster-masking/open_clip/coca_model.py fb651d0a97fd4d3f |
unverified |
no licence file found · pointer only |
| RingID: Rethinking Tree-Ring Watermarking for Enhanced Multi-Key Identification |
22 Apr 2024 |
showlab/ringid/open_clip/coca_model.py fb651d0a97fd4d3f |
unverified |
no licence file found · pointer only |
| FiLo: Zero-Shot Anomaly Detection by Fine-Grained Description and High-Quality Localization |
21 Apr 2024 |
casia-iva-lab/filo/models/vv_open_clip/coca_model.py fb651d0a97fd4d3f |
unverified |
Apache-2.0 (permissive) |
| PromptAD: Learning Prompts with only Normal Samples for Few-Shot Anomaly Detection |
8 Apr 2024 |
funz-0/promptad/PromptAD/CLIPAD/coca_model.py fb651d0a97fd4d3f |
unverified |
Unlicense (permissive) |
| Gaussian Shading: Provable Performance-Lossless Image Watermarking for Diffusion Models |
7 Apr 2024 |
bsmhmmlf/Gaussian-Shading/open_clip/coca_model.py fb651d0a97fd4d3f |
unverified |
MIT (permissive) |
| ViTamin: Designing Scalable Vision Models in the Vision-Language Era |
2 Apr 2024 |
beckschen/vitamin/ViTamin/open_clip/coca_model.py fb651d0a97fd4d3f |
unverified |
Apache-2.0 (permissive) |
| DreamLIP: Language-Image Pre-training with Long Captions |
25 Mar 2024 |
zyf0619sjtu/DreamLIP/open_clip/coca_model.py fb651d0a97fd4d3f |
unverified |
CC-BY-4.0 · pointer only |
| UrbanVLP: Multi-Granularity Vision-Language Pretraining for Urban Socioeconomic Indicator Prediction |
25 Mar 2024 |
citymind-lab/urbanvlp/open_clip_mine/coca_model.py fb651d0a97fd4d3f |
unverified |
no licence file found · pointer only |
| Long-CLIP: Unlocking the Long-Text Capability of CLIP |
22 Mar 2024 |
beichenzbc/long-clip/open_clip_long/coca_model.py fb651d0a97fd4d3f |
unverified |
Apache-2.0 (permissive) |
| Toward Generalist Anomaly Detection via In-context Residual Learning with Few-shot Sample Prompts |
11 Mar 2024 |
mala-lab/WinCLIP/open_clip/coca_model.py fb651d0a97fd4d3f |
unverified |
GPL-3.0 (copyleft) · pointer only |
| Revisiting Zeroth-Order Optimization for Memory-Efficient LLM Fine-Tuning: A Benchmark |
18 Feb 2024 |
zo-bench/zo-llm/zo-bench/prefix_tuning.py 84920a8095dca4fa |
unverified |
GPL-3.0 (copyleft) · pointer only |
| LoRETTA: Low-Rank Economic Tensor-Train Adaptation for Ultra-Low-Parameter Fine-Tuning of Large Language Models |
18 Feb 2024 |
yifanycc/loretta/large_models/prefix.py 84920a8095dca4fa |
unverified |
GPL-3.0 (copyleft) · pointer only |
| Open-Vocabulary SAM: Segment and Recognize Twenty-thousand Classes Interactively |
5 Jan 2024 |
harboryuan/ovsam/ext/open_clip/coca_model.py fb651d0a97fd4d3f |
unverified |
no licence file found · pointer only |
| ZO-AdaMU Optimizer: Adapting Perturbation by the Momentum and Uncertainty in Zeroth-order Optimization |
23 Dec 2023 |
mathisall/zo-adamu/prefix.py 84920a8095dca4fa |
unverified |
no licence file found · pointer only |
| SkyScript: A Large and Semantically Diverse Vision-Language Dataset for Remote Sensing |
20 Dec 2023 |
wangzhecheng/skyscript/src/open_clip/coca_model.py fb651d0a97fd4d3f |
unverified |
MIT (permissive) |
| CLIM: Contrastive Language-Image Mosaic for Region Representation |
18 Dec 2023 |
wusize/clim/src/open_clip/coca_model.py fb651d0a97fd4d3f |
unverified |
no licence file found · pointer only |
| Diffusion Model Alignment Using Direct Preference Optimization |
21 Nov 2023 |
SalesforceAIResearch/DiffusionDPO/utils/open_clip/coca_model.py fb651d0a97fd4d3f |
unverified |
Apache-2.0 recorded; this copy not marked cleared · pointer only |
| AnomalyCLIP: Object-agnostic Prompt Learning for Zero-shot Anomaly Detection |
29 Oct 2023 |
zqhang/WinCLIP-pytorch/src/open_clip/coca_model.py fb651d0a97fd4d3f |
unverified |
MIT (permissive) |
| DPZero: Private Fine-Tuning of Language Models without Backpropagation |
14 Oct 2023 |
Liang137/DPZero/opt/src/prefix.py 84920a8095dca4fa |
unverified |
MIT (permissive) |
| CLIPSelf: Vision Transformer Distills Itself for Open-Vocabulary Dense Prediction |
2 Oct 2023 |
wusize/clipself/src/open_clip/coca_model.py fb651d0a97fd4d3f |
unverified |
no licence file found · pointer only |
| Understanding and Mitigating the Label Noise in Pre-training on Downstream Tasks |
29 Sep 2023 |
Hhhhhhao/Noisy-Model-Learning/open_clip/open_clip/coca_model.py fb651d0a97fd4d3f |
unverified |
no licence file found · pointer only |
| Investigating Personalization Methods in Text to Music Generation |
20 Sep 2023 |
zelaki/DreamSound/pipeline/pipeline_audioldm2.py d5a09007aef6b7ff |
ran
|
no licence file found · pointer only |
| Bootstrap Fine-Grained Vision-Language Alignment for Unified Zero-Shot Anomaly Localization |
30 Aug 2023 |
hq-deng/AnoVL/open_clip/coca_model.py fb651d0a97fd4d3f |
unverified |
MIT (permissive) |
| APRIL-GAN: A Zero-/Few-Shot Anomaly Classification and Segmentation Method for CVPR 2023 VAND Workshop Challenge Tracks 1&2: 1st Place on Zero-shot AD and 4th Place on Few-shot AD |
27 May 2023 |
bychelsea/vand-april-gan/open_clip/coca_model.py fb651d0a97fd4d3f |
unverified |
MIT (permissive) |
| Fine-Tuning Language Models with Just Forward Passes |
27 May 2023 |
princeton-nlp/mezo/large_models/prefix.py 84920a8095dca4fa |
unverified |
MIT (permissive) |
| An Inverse Scaling Law for CLIP Training |
11 May 2023 |
UCSC-VLAA/CLIPA/clipa_torch/open_clip/coca_model.py fb651d0a97fd4d3f |
unverified |
Apache-2.0 (permissive) |
| WinCLIP: Zero-/Few-Shot Anomaly Classification and Segmentation |
26 Mar 2023 |
caoyunkang/WinClip/WinCLIP/CLIPAD/coca_model.py fb651d0a97fd4d3f |
unverified |
MIT (permissive) |
| Learning Customized Visual Models with Retrieval-Augmented Knowledge |
17 Jan 2023 |
microsoft/react/react_customization/src/open_clip/coca_model.py fb651d0a97fd4d3f |
unverified |
MIT (permissive) |
| Retrieval-Free Knowledge-Grounded Dialogue Response Generation with Adapters |
13 May 2021 |
hltchkust/knowexpert/inference_task.py 67735cd057723d31 |
unverified |
MIT (permissive) |
| SummEval: Re-evaluating Summarization Evaluation |
24 Jul 2020 |
PrimerAI/blanc/blanc/shannon.py 56f81e0b7efc8ed8 |
unverified |
MIT (permissive) |
| arXiv:Ma_ReMP-AD_Retrieval-enhanced_Multi-modal_Prompt_Fusion_for_Few-Shot_Industrial_Visual_Anomaly_ICCV_2025_paper |
|
cshcma/ReMP-AD/open_clip/coca_model.py fb651d0a97fd4d3f |
unverified |
MIT (permissive) |
| arXiv:Hertz_Style_Aligned_Image_Generation_via_Shared_Attention_CVPR_2024_paper |
|
aim-uofa/StyleDrop-PyTorch/open_clip/coca_model.py fb651d0a97fd4d3f |
unverified |
MIT (permissive) |
| arXiv:Guo_SCAN_Bootstrapping_Contrastive_Pre-training_for_Data_Efficiency_ICCV_2025_paper |
|
guoyang9/SCAN/open_clip/src/open_clip/coca_model.py fb651d0a97fd4d3f |
unverified |
Apache-2.0 (permissive) |