| Driving on Memory added by Syntology |
2026-08 (from id) |
boschresearch/MemoryDrivoR/bench2drive/navsim/agents/drivoR/timm_layers.py 87577b3ff9d32712 |
ran · fixture could not drive it
|
AGPL-3.0 (copyleft) · pointer only |
| GET: Generative Embedding Translation for Medical Image Segmentation added by Syntology |
2026-08 (from id) |
maklachur/GET/networks/embedding_translation.py 68b7ffb36f0515ff |
ran · fixture could not drive it
|
no licence file found · pointer only |
| Acoustic Prompting via Stage-wise Modulation for Few-Shot Learning in Audio Language Models added by Syntology |
2026-06 (from id) |
hyebin-c/aspl/pengi/models/htsat.py 3ac6b7d76e8e3584 |
ran · fixture could not drive it
fingerprinted |
MIT (permissive) |
| FlexiBrain: Resolution-Agnostic Voxel-Level Encoding for Native fMRI added by Syntology |
2026-06 (from id) |
OneMore1/FlexiBrain/flexibrain/models/transformer_block.py 4bfe3da8a35995ea |
ran
|
licence not identified · pointer only |
| Dial HEALTHDIAL for Advice: A Multilingual and Multi-Parallel Spoken Dialogue Dataset for Knowledge-Grounded Information Seeking added by Syntology |
2026-05 (from id) |
LAION-AI/CLAP/src/laion_clap/clap_module/htsat.py 3ac6b7d76e8e3584 |
ran · fixture could not drive it
fingerprinted |
CC0-1.0 (permissive) |
| Design Your Ad: Personalized Advertising Image and Text Generation with Unified Autoregressive Models added by Syntology |
2026-05 (from id) |
JD-GenX/Uni-AdGen/PBS_metrics/model/vision_transformer.py 55120f2026b56aa2 |
ran · fixture could not drive it
fingerprinted |
no licence file found · pointer only |
| VARestorer: One-Step VAR Distillation for Real-World Image Super-Resolution added by Syntology |
2026-04 (from id) |
EternalEvan/VARestorer/infinity/models/helpers.py 971ae8d2d8313e30 |
ran · fixture could not drive it
|
MIT (permissive) |
| On Neural Scaling Laws for Weather Emulation through Continual Training added by Syntology |
2026-03 (from id) |
ShashankSubramanian/neural-scaling-weather/models/timm_helpers.py 1c864612eb3b612b |
unverified |
licence not identified · pointer only |
| FILT3R: Latent State Adaptive Kalman Filter for Streaming 3D Reconstruction added by Syntology |
2026-03 (from id) |
jinotter3/FILT3R/src/dust3r/model.py aea279a2d78d1352 |
ran · fixture could not drive it
|
licence not identified · pointer only |
| FedBCGD: Communication-Efficient Accelerated Block Coordinate Gradient Descent for Federated Learning added by Syntology |
5 Mar 2026 |
junkangLiu0/FedBCGD/vit_model.py fe7d4321dbeef661 |
ran · fixture could not drive it
|
no licence file found · pointer only |
| Learning from Complexity: Exploring Dynamic Sample Pruning of Spatio-Temporal Training added by Syntology |
2026-02 (from id) |
identical code first harvested elsewhere 9bdf2492c6e0a8de |
ran · fixture could not drive it
|
licence of this copy not recorded |
| Enabling Progressive Whole-slide Image Analysis with Multi-scale Pyramidal Network added by Syntology |
2026-02 (from id) |
mahmoodlab/HIPT/HIPT_4K/vision_transformer.py 55120f2026b56aa2 |
ran · fixture could not drive it
fingerprinted |
no licence file found · pointer only |
| Spatial-Regularization-Aware Dual-Branch Collaborative Inference for Training-Free OVSS in Remote Sensing Imagery added by Syntology |
2026-01 (from id) |
yu-ni1989/SDCI/modified_clip/vision_transformer.py 55120f2026b56aa2 |
ran · fixture could not drive it
fingerprinted |
MIT (permissive) |
| EmoLat: Text-driven Image Sentiment Transfer via Emotion Latent Space added by Syntology |
2026-01 (from id) |
JingVIPLab/EmoLat/model/ViT_helper.py 3ac6b7d76e8e3584 |
ran · fixture could not drive it
fingerprinted |
no licence file found · pointer only |
| Inference-Time Scaling for Visual AutoRegressive modeling by Searching Representative Samples added by Syntology |
2026-01 (from id) |
WD7ang/VAR-Scaling/VAR-main/models/helpers.py 971ae8d2d8313e30 |
ran · fixture could not drive it
|
MIT (permissive) |
| Pixel-Perfect Visual Geometry Estimation added by Syntology |
2026-01 (from id) |
gangweix/pixel-perfect-depth/ppd/models/depth_anything_v2/dinov2_layers/drop_path.py c157f5b112b3a392 |
ran · fixture could not drive it
|
Apache-2.0 (permissive) |
| Open Ad-hoc Categorization with Contextualized Feature Learning added by Syntology |
2025-12 (from id) |
Wayne2Wang/OAK/src/models/dino_vision_transformer.py 55120f2026b56aa2 |
ran · fixture could not drive it
fingerprinted |
MIT (permissive) |
| BioBench: A Blueprint to Move Beyond ImageNet for Scientific ML Benchmarks added by Syntology |
2025-11 (from id) |
samuelstevens/biobench/src/biobench/webssl.py c157f5b112b3a392 |
ran · fixture could not drive it
|
MIT (permissive) |
| Brain Harmony: A Multimodal Foundation Model Unifying Morphology and Function into 1D Tokens added by Syntology |
2025-09 (from id) |
identical code first harvested elsewhere 55120f2026b56aa2 |
ran · fixture could not drive it
fingerprinted |
licence of this copy not recorded |
| Streaming 4D Visual Geometry Transformer |
15 Jul 2025 |
wzzheng/streamvggt/src/streamvggt/layers/drop_path.py c157f5b112b3a392 |
ran · fixture could not drive it
|
no licence file found · pointer only |
| U-RWKV: Lightweight medical image segmentation with direction-adaptive RWKV |
15 Jul 2025 |
hbyecoding/u-rwkv/models/cmunext/cmunext_rwkv_test_bk.py 87577b3ff9d32712 |
ran · fixture could not drive it
|
MIT (permissive) |
| arXiv:2507.03779 |
2025-07 (from id) |
KevinZ0217/fast_dinov2/dinov2/layers/drop_path.py c157f5b112b3a392 |
ran · fixture could not drive it
|
no licence file found · pointer only |
| Seg-R1: Segmentation Can Be Surprisingly Simple with Reinforcement Learning |
27 Jun 2025 |
geshang777/FOCUS/focus/modeling/edge_enhancer/edge_enhancer.py 4134f673786a2abe |
unverified |
Apache-2.0 (permissive) |
| Test3R: Learning to Reconstruct 3D at Test Time |
16 Jun 2025 |
nopqaq/test3r/croco/models/blocks.py f9a1900525331e0f |
ran · fixture could not drive it
|
no licence file found · pointer only |
| MedITok: A Unified Tokenizer for Medical Image Synthesis and Interpretation |
25 May 2025 |
masaaki-75/meditok/layers/drop_path.py c157f5b112b3a392 |
ran · fixture could not drive it
|
MIT (permissive) |
| MVAR: Visual Autoregressive Modeling with Scale and Spatial Markovian Conditioning |
19 May 2025 |
labshuhanggu/mvar/models/helpers.py 971ae8d2d8313e30 |
ran · fixture could not drive it
|
MIT (permissive) |
| Understanding the Capabilities of Molecular Graph Neural Networks in Materials Science Through Multimodal Learning and Physical Context Encoding |
17 May 2025 |
kurbanintelligencelab/understandingmultimodalgnns/models/equiformer/drop.py 3ac6b7d76e8e3584 |
ran · fixture could not drive it
fingerprinted |
MIT (permissive) |
| Object-Shot Enhanced Grounding Network for Egocentric Video |
7 May 2025 |
Yisen-Feng/OSGNet/libs/modeling/blocks.py a34c005ba2203f35 |
ran · fixture could not drive it
|
MIT (permissive) |
| Perception Encoder: The best visual embeddings are not at the output of the network |
17 Apr 2025 |
facebookresearch/perception_models/apps/detection/DETA_pe/models/swin.py 87577b3ff9d32712 |
ran · fixture could not drive it
|
Apache-2.0 (permissive) |
| F$^3$Set: Towards Analyzing Fast, Frequent, and Fine-grained Events from Videos |
11 Apr 2025 |
f3set/f3set/model/impl/actionformer.py a34c005ba2203f35 |
ran · fixture could not drive it
|
no licence file found · pointer only |
| ProtoGCD: Unified and Unbiased Prototype Learning for Generalized Category Discovery |
2 Apr 2025 |
mashijie1028/protogcd/models/vision_transformer.py 55120f2026b56aa2 |
ran · fixture could not drive it
fingerprinted |
MIT (permissive) |
| Modeling Multiple Normal Action Representations for Error Detection in Procedural Tasks |
28 Mar 2025 |
iSEE-Laboratory/AMNAR/libs/modeling/blocks.py a34c005ba2203f35 |
ran · fixture could not drive it
|
MIT (permissive) |
| TreeMeshGPT: Artistic Mesh Generation with Autoregressive Tree Sequencing |
14 Mar 2025 |
sail-sg/treemeshgpt/model/pc_encoder.py 87577b3ff9d32712 |
ran · fixture could not drive it
|
MIT (permissive) |
| SVIP: Semantically Contextualized Visual Patches for Zero-Shot Learning |
13 Mar 2025 |
identical code first harvested elsewhere fe7d4321dbeef661 |
ran · fixture could not drive it
|
licence of this copy not recorded |
| VLog: Video-Language Models by Generative Retrieval of Narration Vocabulary |
12 Mar 2025 |
showlab/VLog/VLog/model/models.py 10e67144852e5fce |
ran · fixture could not drive it
fingerprinted |
no licence file found · pointer only |
| "Principal Components" Enable A New Language of Images |
11 Mar 2025 |
visual-gen/semanticist/semanticist/stage1/vision_transformer.py 4134f673786a2abe |
unverified |
MIT (permissive) |
| Mellow: a small audio language model for reasoning |
11 Mar 2025 |
soham97/mellow/mellow/model/htsat.py 3ac6b7d76e8e3584 |
ran · fixture could not drive it
fingerprinted |
MIT (permissive) |
| FlexVAR: Flexible Visual Autoregressive Modeling without Residual Prediction |
27 Feb 2025 |
jiaosiyu1999/FlexVAR/models/helpers.py 971ae8d2d8313e30 |
ran · fixture could not drive it
|
MIT (permissive) |
| Fractal Generative Models |
24 Feb 2025 |
LTH14/fractalgen/models/ar.py 87577b3ff9d32712 |
ran · fixture could not drive it
|
MIT (permissive) |
| ADIFF: Explaining audio difference using natural language |
6 Feb 2025 |
soham97/adiff/model/htsat.py 3ac6b7d76e8e3584 |
ran · fixture could not drive it
fingerprinted |
MIT (permissive) |
| Visual Autoregressive Modeling for Image Super-Resolution |
31 Jan 2025 |
qyp2000/varsr/models/helpers.py 971ae8d2d8313e30 |
ran · fixture could not drive it
|
MIT (permissive) |
| LLMDet: Learning Strong Open-Vocabulary Object Detectors under the Supervision of Large Language Models |
31 Jan 2025 |
iSEE-Laboratory/LLMDet/hf_model/modeling_grounding_dino.py b6822336792e87db |
unverified |
Apache-2.0 recorded; this copy not marked cleared · pointer only |
| Video Depth Anything: Consistent Depth Estimation for Super-Long Videos |
21 Jan 2025 |
DepthAnything/Video-Depth-Anything/video_depth_anything/dinov2_layers/drop_path.py c157f5b112b3a392 |
ran · fixture could not drive it
|
Apache-2.0 (permissive) |
| Personalized Representation from Personalized Generation |
20 Dec 2024 |
ssundaram21/personalized-rep/models/vision_transformer.py 55120f2026b56aa2 |
ran · fixture could not drive it
fingerprinted |
MIT (permissive) |
| Dense Audio-Visual Event Localization under Cross-Modal Consistency and Multi-Temporal Granularity Collaboration |
17 Dec 2024 |
zzhhfut/ccnet-aaai2025/libs/modeling/blocks.py 019e6b38831f1474 |
unverified |
no licence file found · pointer only |
| SLAM3R: Real-Time Dense Scene Reconstruction from Monocular RGB Videos |
12 Dec 2024 |
pku-vcl-3dv/slam3r/slam3r/blocks/basic_blocks.py f9a1900525331e0f |
ran · fixture could not drive it
|
no licence file found · pointer only |
| Speech-Forensics: Towards Comprehensive Synthetic Speech Dataset Establishment and Analysis |
12 Dec 2024 |
ring-zl/Speech-Forensics/libs/modeling/blocks.py a34c005ba2203f35 |
ran · fixture could not drive it
|
no licence file found · pointer only |
| Infinity: Scaling Bitwise AutoRegressive Modeling for High-Resolution Image Synthesis |
5 Dec 2024 |
FoundationVision/VAR/models/helpers.py 971ae8d2d8313e30 |
ran · fixture could not drive it
|
MIT (permissive) |
| Utilizing Uncertainty in 2D Pose Detectors for Probabilistic 3D Human Mesh Recovery |
25 Nov 2024 |
twehrbein/humr/easy_vitpose/vit.py 87577b3ff9d32712 |
ran · fixture could not drive it
|
MIT (permissive) |
| PromptHSI: Universal Hyperspectral Image Restoration with Vision-Language Modulated Frequency Adaptation |
24 Nov 2024 |
chingheng0808/PromptHSI/utils/DRCT.py 52d96aa31ed74a56 |
ran · fixture could not drive it
|
MIT (permissive) |
| PanoLlama: Generating Endless and Coherent Panoramas with Next-Token-Prediction LLMs |
24 Nov 2024 |
0606zt/panollama/utils/drop_path.py e3aa4e8e74369506 |
ran · fixture could not drive it
|
no licence file found · pointer only |
| Sample- and Parameter-Efficient Auto-Regressive Image Models |
23 Nov 2024 |
elad-amrani/xtra/src/modules/vit.py 55120f2026b56aa2 |
ran · fixture could not drive it
fingerprinted |
no licence file found · pointer only |
| MambaIRv2: Attentive State Space Restoration |
22 Nov 2024 |
csguoh/mambair/analysis/model_zoo/hat.py 52d96aa31ed74a56 |
ran · fixture could not drive it
|
Apache-2.0 (permissive) |
| Generalizable Person Re-identification via Balancing Alignment and Uniformity |
18 Nov 2024 |
yoonkicho/BAU/bau/models/vit.py 39eace7e2822504f |
ran · fixture could not drive it
|
MIT (permissive) |
| IKEA Manuals at Work: 4D Grounding of Assembly Instructions on Internet Videos |
18 Nov 2024 |
yunongLiu1/IKEA-Manuals-at-Work/src/IKEAVideo/featurizers/DINO.py 55120f2026b56aa2 |
ran · fixture could not drive it
fingerprinted |
no licence file found · pointer only |
| IKEA Manuals at Work: 4D Grounding of Assembly Instructions on Internet Videos |
18 Nov 2024 |
yunongLiu1/IKEA-Manuals-at-Work/src/IKEAVideo/featurizers/DINOv2.py 8a105009df710b93 |
unverified |
no licence file found · pointer only |
| M-VAR: Decoupled Scale-wise Autoregressive Modeling for High-Quality Image Generation |
15 Nov 2024 |
oliverrensu/mvar/models/helpers.py 971ae8d2d8313e30 |
ran · fixture could not drive it
|
no licence file found · pointer only |
| Test-Time Dynamic Image Fusion |
5 Nov 2024 |
Yinan-Xia/TTD/net.py fe7d4321dbeef661 |
ran · fixture could not drive it
|
Apache-2.0 (permissive) |
| PPLLaVA: Varied Video Sequence Understanding With Prompt Guidance |
4 Nov 2024 |
farewellthree/ppllava/ppllava/models/clip_btadapter.py 9a83c0663b6afcd4 |
unverified |
Apache-2.0 (permissive) |
| Expanding Sparse Tuning for Low Memory Usage |
4 Nov 2024 |
ssfgunner/SNELL/model/utils.py 39eace7e2822504f |
ran · fixture could not drive it
|
MIT (permissive) |
| LARP: Tokenizing Videos with a Learned Autoregressive Generative Prior |
28 Oct 2024 |
hywang66/LARP/models/larp_ar.py 87577b3ff9d32712 |
ran · fixture could not drive it
|
MIT (permissive) |
| On Occlusions in Video Action Detection: Benchmark Datasets And Training Recipes |
25 Oct 2024 |
rajatmodi62/OccludedActionBenchmark/codebase/codebase_islands/slowfast/models/common.py bcc1cdae3bb3212c |
unverified |
no licence file found · pointer only |
| Prototypical Hash Encoding for On-the-Fly Fine-Grained Category Discovery |
24 Oct 2024 |
HaiyangZheng/PHE/vision_transformer.py 55120f2026b56aa2 |
ran · fixture could not drive it
fingerprinted |
no licence file found · pointer only |
| Open Materials 2024 (OMat24) Inorganic Materials Dataset and Models |
16 Oct 2024 |
atomicarchitects/dens/model/equiformer_v2/drop.py 3ac6b7d76e8e3584 |
ran · fixture could not drive it
fingerprinted |
MIT (permissive) |
| HART: Efficient Visual Generation with Hybrid Autoregressive Transformer |
14 Oct 2024 |
mit-han-lab/hart/hart/modules/networks/utils.py 971ae8d2d8313e30 |
ran · fixture could not drive it
|
MIT (permissive) |
| MoTE: Reconciling Generalization with Specialization for Visual-Language to Video Knowledge Transfer |
14 Oct 2024 |
ZMHH-H/MoTE/clip/model.py 3ac6b7d76e8e3584 |
ran · fixture could not drive it
fingerprinted |
Apache-2.0 (permissive) |
| Happy: A Debiased Learning Framework for Continual Generalized Category Discovery |
9 Oct 2024 |
mashijie1028/Happy-CGCD/models/lora_vision_transformer.py 55120f2026b56aa2 |
ran · fixture could not drive it
fingerprinted |
MIT (permissive) |
| Temporally Aligned Audio for Video with Autoregression |
20 Sep 2024 |
ilpoviertola/V-AURA/utils/drop_path.py 87577b3ff9d32712 |
ran · fixture could not drive it
|
MIT (permissive) |
| Prithvi WxC: Foundation Model for Weather and Climate |
20 Sep 2024 |
nasa-impact/prithvi-wxc/PrithviWxC/model.py 9fb52d0e32d427ed |
ran
|
MIT (permissive) |
| Mamba-ST: State Space Model for Efficient Style Transfer |
16 Sep 2024 |
filippobotti/mambast/models/ViT_helper.py 3ac6b7d76e8e3584 |
ran · fixture could not drive it
fingerprinted |
no licence file found · pointer only |
| SelEx: Self-Expertise in Fine-Grained Generalized Category Discovery |
26 Aug 2024 |
sarahrastegar/selex/models/vision_transformer.py 55120f2026b56aa2 |
ran · fixture could not drive it
fingerprinted |
MIT (permissive) |
| Online Continuous Generalized Category Discovery |
24 Aug 2024 |
khu-agi/ocgcd/net/vision_transformer.py 55120f2026b56aa2 |
ran · fixture could not drive it
fingerprinted |
no licence file found · pointer only |
| OpenCity: Open Spatio-Temporal Foundation Models for Traffic Prediction |
16 Aug 2024 |
hkuds/opencity/model/OpenCity/OpenCity.py 9bdf2492c6e0a8de |
ran · fixture could not drive it
|
MIT (permissive) |
| SpectralEarth: Training Hyperspectral Foundation Models at Scale |
15 Aug 2024 |
panopticon-fm/panopticon/dinov2/layers/drop_path.py c157f5b112b3a392 |
ran · fixture could not drive it
|
Apache-2.0 (permissive) |
| Advancing Multi-grained Alignment for Contrastive Language-Audio Pre-training |
15 Aug 2024 |
ming-er/mga-clap/models/htsat.py 3ac6b7d76e8e3584 |
ran · fixture could not drive it
fingerprinted |
no licence file found · pointer only |
| VAR-CLIP: Text-to-Image Generator with Visual Auto-Regressive Modeling |
2 Aug 2024 |
daixiangzi/var-clip/models/helpers.py 971ae8d2d8313e30 |
ran · fixture could not drive it
|
no licence file found · pointer only |
| Mixture of Nested Experts: Adaptive Processing of Visual Tokens |
29 Jul 2024 |
usryokousha/mone-pytorch/mone_pytorch/layers/block.py c157f5b112b3a392 |
ran · fixture could not drive it
|
MIT (permissive) |
| Learning from Memory: Non-Parametric Memory Augmented Self-Supervised Learning of Visual Features |
3 Jul 2024 |
sthalles/MaSSL/models/vision_transformer.py 55120f2026b56aa2 |
ran · fixture could not drive it
fingerprinted |
no licence file found · pointer only |
| Early Preparation Pays Off: New Classifier Pre-tuning for Class Incremental Semantic Segmentation |
19 Jul 2024 |
zhengyuan-xie/ECCV24_NeST/modules/backbone.py f5ab76486eeab7b9 |
ran · fixture could not drive it
fingerprinted |
no licence file found · pointer only |
| SCAPE: A Simple and Strong Category-Agnostic Pose Estimator |
18 Jul 2024 |
tiny-smart/SCAPE/scape/models/layers/drop_path.py c157f5b112b3a392 |
ran · fixture could not drive it
|
no licence file found · pointer only |
| Rethinking the Architecture Design for Efficient Generic Event Boundary Detection |
17 Jul 2024 |
ziwei-zheng/efficientgebd/EffSoccerNet/diff_former.py 971ae8d2d8313e30 |
ran · fixture could not drive it
|
no licence file found · pointer only |
| Accessing Vision Foundation Models at ImageNet-level Costs |
15 Jul 2024 |
bespontaneous/proteus-pytorch/pretrain/models_dinov2.py c157f5b112b3a392 |
ran · fixture could not drive it
|
MIT (permissive) |
| Retrospective for the Dynamic Sensorium Competition for predicting large-scale mouse primary visual cortex activity from videos |
12 Jul 2024 |
lRomul/sensorium/src/models/dwiseneuro.py 971ae8d2d8313e30 |
ran · fixture could not drive it
|
MIT (permissive) |
| WildGaussians: 3D Gaussian Splatting in the Wild |
11 Jul 2024 |
jkulhanek/wild-gaussians/wildgaussians/dinov2.py c157f5b112b3a392 |
ran · fixture could not drive it
|
no licence file found · pointer only |
| A Clinical Benchmark of Public Self-Supervised Pathology Foundation Models |
9 Jul 2024 |
fuchs-lab-public/opal/SSL_benchmarks/code/feature_extraction/vision_transformer.py 55120f2026b56aa2 |
ran · fixture could not drive it
fingerprinted |
MIT (permissive) |
| CLAMP-ViT: Contrastive Data-Free Learning for Adaptive Post-Training Quantization of ViTs |
7 Jul 2024 |
georgia-tech-synergy-lab/clamp-vit/models/layers_quant.py 3ac6b7d76e8e3584 |
ran · fixture could not drive it
fingerprinted |
MIT (permissive) |
| DyFADet: Dynamic Feature Aggregation for Temporal Action Detection |
3 Jul 2024 |
yangle15/DyFADet-pytorch/libs/modeling/blocks.py a34c005ba2203f35 |
ran · fixture could not drive it
|
no licence file found · pointer only |
| Explicitly Guided Information Interaction Network for Cross-modal Point Cloud Completion |
3 Jul 2024 |
WHU-USI3DV/EGIInet/models/layers/drop.py 87577b3ff9d32712 |
ran · fixture could not drive it
|
MIT (permissive) |
| Semantically Guided Representation Learning For Action Anticipation |
2 Jul 2024 |
ADiko1997/S-GEAR/models/base_model_ts.py 314973a77b15948e |
ran · fixture could not drive it
fingerprinted |
Apache-2.0 (permissive) |
| ObjectNLQ @ Ego4D Episodic Memory Challenge 2024 |
22 Jun 2024 |
yisen-feng/objectnlq/libs/modeling/blocks.py a34c005ba2203f35 |
ran · fixture could not drive it
|
MIT (permissive) |
| Stylebreeder: Exploring and Democratizing Artistic Styles through Text-to-Image Models |
20 Jun 2024 |
stylebreeder/stylebreeder-code/models/dino_vits.py 55120f2026b56aa2 |
ran · fixture could not drive it
fingerprinted |
MIT (permissive) |
| Strengthening Layer Interaction via Dynamic Layer Attention |
19 Jun 2024 |
tunantu/dynamic-layer-attention/image_classification/cifar/models/cifar/dla_b.py 39eace7e2822504f |
ran · fixture could not drive it
|
no licence file found · pointer only |
| Enhancing Automated Audio Captioning via Large Language Models with Optimized Audio Encoding |
19 Jun 2024 |
frankenliu/LOAE/models/ced/layers.py 87577b3ff9d32712 |
ran · fixture could not drive it
|
Apache-2.0 (permissive) |
| Depth Anything V2 |
13 Jun 2024 |
DepthAnything/Depth-Anything-V2/depth_anything_v2/dinov2_layers/drop_path.py c157f5b112b3a392 |
ran · fixture could not drive it
|
Apache-2.0 (permissive) |
| Depth Anything V2 |
13 Jun 2024 |
fabio-sim/Depth-Anything-ONNX/depth_anything_v2/dinov2_layers/drop_path.py 7fd11117910b59f9 |
unverified |
Apache-2.0 (permissive) |
| GLAD: Towards Better Reconstruction with Global and Local Adaptive Diffusion Models for Unsupervised Anomaly Detection |
11 Jun 2024 |
hyao1/glad/dino/vision_transformer.py 55120f2026b56aa2 |
ran · fixture could not drive it
fingerprinted |
MIT (permissive) |
| RWKV-CLIP: A Robust Vision-Language Representation Learner |
11 Jun 2024 |
deepglint/RWKV-CLIP/model/utils_vision_rwkv/drop.py 87577b3ff9d32712 |
ran · fixture could not drive it
|
MIT (permissive) |
| Weakly Supervised Set-Consistency Learning Improves Morphological Profiling of Single-Cell Images |
8 Jun 2024 |
Genentech/set-dino/vision_transformer.py 55120f2026b56aa2 |
ran · fixture could not drive it
fingerprinted |
Apache-2.0 (permissive) |
| Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis |
31 May 2024 |
PhysGame/PhysGame/physvlm/models/clip_btadapter.py 9a83c0663b6afcd4 |
unverified |
Apache-2.0 (permissive) |
| Automatic Jailbreaking of the Text-to-Image Generative AI Systems |
26 May 2024 |
Kim-Minseon/APGP/vision_transformer.py 55120f2026b56aa2 |
ran · fixture could not drive it
fingerprinted |
no licence file found · pointer only |
| Sparse-Tuning: Adapting Vision Transformers with Efficient Fine-tuning and Inference |
23 May 2024 |
liuting20/sparse-tuning/models/vit_image.py 39eace7e2822504f |
ran · fixture could not drive it
|
no licence file found · pointer only |
| Dinomaly: The Less Is More Philosophy in Multi-Class Unsupervised Anomaly Detection |
23 May 2024 |
guojiajeremy/dinomaly/models/vision_transformer.py 55120f2026b56aa2 |
ran · fixture could not drive it
fingerprinted |
Apache-2.0 (permissive) |