| Cross-Modal Masked Compositional Concept Modeling for Enhancing Visio-Linguistic Compositionality added by Syntology |
2026-06 (from id) |
hiker-lw/MACCO/src/open_clip_code/MACCO_variant/MACCO_text_image.py b2bb1279ceae186b |
ran
|
MIT (permissive) |
| MAM-CLIP: Vision-Language Pretraining on Mammography Atlases for BI-RADS Classification added by Syntology |
2026-05 (from id) |
igulluk/MAM-CLIP/train/files/model.py ee9965b63eba0564 |
ran · metamorphic tier: invariant
fingerprinted |
licence not identified · pointer only |
| TIPS Over Tricks: Simple Prompts for Effective Zero-shot Anomaly Detection added by Syntology |
2026-02 (from id) |
AlirezaSalehy/Tipsomaly/model/tips/text_encoder.py bec621a25d6d8e86 |
ran · metamorphic tier: deterministic
|
MIT (permissive) |
| TOMCAT : Test-time Comprehensive Knowledge Accumulation for Compositional Zero-Shot Learning added by Syntology |
2025-10 (from id) |
xud-yan/TOMCAT/model/tomcat_bm.py d9787281b177005b |
ran · metamorphic tier: invariant
|
MIT (permissive) |
| CovMatch: Cross-Covariance Guided Multimodal Dataset Distillation with Trainable Text Encoder added by Syntology |
2025-10 (from id) |
Yongalls/CovMatch/src/model.py 6088793f94190d8d |
ran · metamorphic tier: deterministic
|
no licence file found · pointer only |
| InfMasking: Unleashing Synergistic Information by Contrastive Multimodal Interactions added by Syntology |
2025-09 (from id) |
brightest66/InfMasking/pl_modules/infmasking.py 00744d2555cd4f61 |
ran
|
no licence file found · pointer only |
| AliTok: Towards Sequence Modeling Alignment between Tokenizer and Autoregressive Model |
5 Jun 2025 |
ali-vilab/alitok/modeling/alitok.py 71a046a9abe78001 |
ran · metamorphic tier: invariant
fingerprinted |
no licence file found · pointer only |
| MSCI: Addressing CLIP's Inherent Limitations for Compositional Zero-Shot Learning |
15 May 2025 |
ltpwy/MSCI/MSCI/code/model/Mutifuse_new.py 1a767496bd62876c |
ran · metamorphic tier: deterministic
|
no licence file found · pointer only |
| Logits DeConfusion with CLIP for Few-Shot Learning |
16 Apr 2025 |
LiShuo1001/LDC/clip_ldc/model.py 7279d08798402859 |
ran
|
no licence file found · pointer only |
| Exploring CLIP's Dense Knowledge for Weakly Supervised Semantic Segmentation |
26 Mar 2025 |
zwyang6/ExCEL/model/model_excel.py 2c6a4eb599ddebf9 |
unverified |
no licence file found · pointer only |
| HiMTok: Learning Hierarchical Mask Tokens for Image Segmentation with Large Multimodal Model |
17 Mar 2025 |
yayafengzi/LMM-HiMTok/himt/himt.py 8d70fe4156ce8962 |
ran · metamorphic tier: deterministic
|
Apache-2.0 (permissive) |
| FlowTok: Flowing Seamlessly Across Text and Image Tokens |
13 Mar 2025 |
bytedance/1d-tokenizer/modeling/titok.py f9805e2cf90696a2 |
ran
|
Apache-2.0 (permissive) |
| Representing 3D Shapes With 64 Latent Vectors for 3D Diffusion Models |
11 Mar 2025 |
join16/COD-VAE/cod/models/vae/vae.py f4f6c8934a1b368d |
unverified |
no licence file found · pointer only |
| LongProLIP: A Probabilistic Vision-Language Model with Long Context Text |
11 Mar 2025 |
naver-ai/prolip/src/prolip/model.py 022a5391f0e240ca |
ran · metamorphic tier: invariant
fingerprinted |
licence not identified · pointer only |
| Learning to Generalize without Bias for Open-Vocabulary Action Recognition |
27 Feb 2025 |
Mia-YatingYu/Open-MeDe/slowfast/models/customize_visiontransformer.py b84374673409547e |
ran · metamorphic tier: deterministic
|
Apache-2.0 (permissive) |
| Kronecker Mask and Interpretive Prompts are Language-Action Video Learners |
5 Feb 2025 |
yjyddq/CLAVER/models/claver.py 23a40303b56bf120 |
ran
|
no licence file found · pointer only |
| Classification Done Right for Vision-Language Pre-Training |
5 Nov 2024 |
x-cls/superclass/opencls/open_clip/cls_model.py 7000ad070b2b4c5c |
ran · metamorphic tier: invariant
fingerprinted |
Apache-2.0 (permissive) |
| PointAD: Comprehending 3D Anomalies from Points and Pixels for Zero-shot 3D Anomaly Detection |
1 Oct 2024 |
zqhang/pointad/AnomalyCLIP_lib/AnomalyCLIP.py 6cbaa01bb223f295 |
unverified |
no licence file found · pointer only |
| TempMe: Video Temporal Token Merging for Efficient Text-Video Retrieval |
2 Sep 2024 |
LunarShen/TempMe/tvr/models/module_tome_patch.py 07221ceabbfc85ec |
ran · metamorphic tier: deterministic
|
MIT (permissive) |
| AdaCLIP: Adapting CLIP with Hybrid Learnable Prompts for Zero-Shot Anomaly Detection |
22 Jul 2024 |
caoyunkang/AdaCLIP/method/adaclip.py 7affcf897896a315 |
ran
|
MIT (permissive) |
| Weighted Point Cloud Embedding for Multimodal Contrastive Learning Toward Optimal Similarity Metric |
30 Apr 2024 |
sony/wpse/models.py 7f9563e94f649f1c |
ran
|
MIT (permissive) |
| Faceptor: A Generalist Model for Face Perception |
14 Mar 2024 |
lxq1000/Faceptor/faceptor_project/core/model/backbone/farl_vit.py 9ed400c6ff095fd9 |
ran · metamorphic tier: deterministic
|
MIT (permissive) |
| Accurate and Fast Compressed Video Captioning |
22 Sep 2023 |
acherstyx/CoCap/cocap/modules/compressed_video/compressed_video_transformer.py 3e212daf932326fa |
ran · metamorphic tier: deterministic
|
MIT (permissive) |
| Chinese Text Recognition with A Pre-Trained CLIP-Like Model Through Image-IDS Aligning |
3 Sep 2023 |
FudanVI/FudanOCR/image-ids-CTR/CCR-CLIP/model.py 177010bd20c75960 |
ran · metamorphic tier: deterministic
|
no licence file found · pointer only |
| ALIP: Adaptive Language-Image Pre-training with Synthetic Caption |
16 Aug 2023 |
deepglint/alip/src/open_alip/model.py 84ad28ccef1c7735 |
ran · metamorphic tier: invariant
fingerprinted |
no licence file found · pointer only |
| ChessGPT: Bridging Policy Learning and Language Modeling |
15 Jun 2023 |
waterhorse1/chessgpt/chessclip/src/open_clip/model.py ace306ca4e4fe383 |
ran
|
Apache-2.0 (permissive) |
| CrossGET: Cross-Guided Ensemble of Tokens for Accelerating Vision-Language Transformers |
27 May 2023 |
sdc17/crossget/CLIP/clip/model.py 5564a2cac36e0f33 |
unverified |
BSD-3-Clause (permissive) |
| Learning Emotion Representations from Verbal and Nonverbal Communication |
22 May 2023 |
Xeaver/EmotionCLIP/src/models/base.py 72ae1815d1a81517 |
ran
|
MIT (permissive) |
| InfoMetIC: An Informative Metric for Reference-free Image Caption Evaluation |
10 May 2023 |
hawlyq/infometic/infometic/model/ClipSeq.py 6ee9db68c1c1e0ab |
ran
|
no licence file found · pointer only |
| Learning Procedure-aware Video Representation from Instructional Videos and Their Narrations |
31 Mar 2023 |
facebookresearch/ProcedureVRL/lib/models/tfm_model.py 1385cee74dfe54ab |
ran · metamorphic tier: invariant
fingerprinted |
licence not identified · pointer only |
| Video-Text as Game Players: Hierarchical Banzhaf Interaction for Cross-Modal Representation Learning |
25 Mar 2023 |
jpthu17/dicosa/tvr/models/modeling.py 99083d92c6568c29 |
ran · metamorphic tier: deterministic
|
Apache-2.0 (permissive) |
| Unbiased Multiple Instance Learning for Weakly Supervised Video Anomaly Detection |
22 Mar 2023 |
ktr-hubrt/UMIL/models/mit.py 99145f709504d6ad |
ran · metamorphic tier: deterministic
|
no licence file found · pointer only |
| Lana: A Language-Capable Navigator for Instruction Following and Generation |
15 Mar 2023 |
wxh1996/lana-vln/finetune_src/models/vilmodel_cmt_lana.py e60a700608513530 |
ran
|
MIT (permissive) |
| Fine-Grained Object Classification via Self-Supervised Pose Alignment |
30 Mar 2022 |
salwaalkhatib/p2p-net/model.py 456fa1217672f45e |
ran · metamorphic tier: deterministic
|
no licence file found · pointer only |
| GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion Models |
20 Dec 2021 |
openai/glide-text2im/glide_text2im/text2im_model.py 240a712d20a01f00 |
ran · metamorphic tier: deterministic
|
MIT (permissive) |
| Scaling Up Visual and Vision-Language Representation Learning With Noisy Text Supervision |
11 Feb 2021 |
facebookresearch/metaclip/src/mini_clip/model.py 19aedb18f8c2e8ad |
ran · metamorphic tier: deterministic
|
licence not identified · pointer only |
| Exploring Simple Siamese Representation Learning |
20 Nov 2020 |
facebookresearch/clip-rocket/models.py 009322dbbfd98763 |
ran · metamorphic tier: deterministic
|
licence not identified · pointer only |
| arXiv:2025.findings-emnlp.28 |
|
BUAAPY/ProPy/modules/clip_propy.py cd9f8414477e8811 |
ran · metamorphic tier: deterministic
|
no licence file found · pointer only |