| GET: Generative Embedding Translation for Medical Image Segmentation added by Syntology |
2026-08 (from id) |
maklachur/GET/networks/embedding_translation.py 6970a40ad31697f9 |
ran · metamorphic tier: deterministic
fingerprinted |
no licence file found · pointer only |
| Learning the Context of Errors: Black-Box Online Adaptation of Time Series Foundation Models added by Syntology |
2026-06 (from id) |
Fifthky/ORCA/core/refiner_attn.py ea28ec451a5d07bc |
ran · metamorphic tier: deterministic
fingerprinted |
no licence file found · pointer only |
| Graph Set Transformer added by Syntology |
2026-06 (from id) |
daenuprobst/gst-conference/src/graph_set_transformer/models/gst.py 8d650e598e1402c8 |
ran · metamorphic tier: deterministic
fingerprinted |
no licence file found · pointer only |
| PQDT: Pseudo-Query Dual Transformer for Robust Point Cloud Restoration added by Syntology |
2026-05 (from id) |
ins-uni-bonn/PQDT/pqdt/models/pq_transformer.py ae51a68c68f169d8 |
ran · metamorphic tier: deterministic
fingerprinted |
no licence file found · pointer only |
| FILT3R: Latent State Adaptive Kalman Filter for Streaming 3D Reconstruction added by Syntology |
2026-03 (from id) |
jinotter3/FILT3R/src/dust3r/model.py 52ab34ca8fd2e134 |
ran
fingerprinted |
licence not identified · pointer only |
| SpecMoE: Spectral Mixture-of-Experts Foundation Model for Cross-Species EEG Decoding added by Syntology |
2026-03 (from id) |
935963004/LaBraM/modeling_pretrain.py ea7880551588ee9f |
ran · metamorphic tier: deterministic
fingerprinted |
MIT (permissive) |
| Autoregressive Visual Decoding from EEG Signals added by Syntology |
2026-02 (from id) |
ddicee/avde/models/labram.py 84590d675e94d536 |
ran
fingerprinted |
no licence file found · pointer only |
| ECHOSAT: Estimating Canopy Height Over Space and Time added by Syntology |
2026-02 (from id) |
ai4forest/echosat/fine-tuning/models/swin_video_unet.py bdfb44252c4a4200 |
ran · metamorphic tier: deterministic
fingerprinted |
Apache-2.0 (permissive) |
| T1: One-to-One Channel-Head Binding for Multivariate Time-Series Imputation added by Syntology |
2026-02 (from id) |
Oppenheimerdinger/T1/models/T1.py 96592475595fbed1 |
ran · metamorphic tier: invariant
fingerprinted |
MIT (permissive) |
| Learning from Complexity: Exploring Dynamic Sample Pruning of Spatio-Temporal Training added by Syntology |
2026-02 (from id) |
HKUDS/OpenCity/model/OpenCity/OpenCity.py dbb44455f4f7cf38 |
ran · metamorphic tier: deterministic
fingerprinted |
MIT (permissive) |
| BrainRVQ: A High-Fidelity EEG Foundation Model via Dual-Domain Residual Quantization and Hierarchical Autoregression added by Syntology |
2026-02 (from id) |
keqicmz/BrainRVQ/DDRVQ/modeling_ddrvq.py 8c1361edd265ca8a |
ran
fingerprinted |
no licence file found · pointer only |
| EchoJEPA: A Latent Predictive Foundation Model for Echocardiography added by Syntology |
2026-02 (from id) |
bowang-lab/EchoJEPA/src/models/predictor.py 367c364be6605f4c |
ran
fingerprinted |
Apache-2.0 (permissive) |
| EEG-FM-Compass: Progress, Benchmarking, and Future Directions for EEG Foundation Models added by Syntology |
2026-01 (from id) |
Dingkun0817/EEG-FM-Benchmark/models/FM/EEGPT/Model_EEGPT.py 8e464eea7a15e985 |
ran
fingerprinted |
MIT (permissive) |
| CliffordNet: All You Need is Geometric Algebra added by Syntology |
2026-01 (from id) |
ParaMind2025/CAN/model.py 2a0ce6ede68d0fc3 |
ran
fingerprinted |
no licence file found · pointer only |
| Omni-View: Unlocking How Generation Facilitates Understanding in Unified 3D Model based on Multiview images added by Syntology |
2025-11 (from id) |
AIDC-AI/Omni-View/modeling/bagel/bagel.py 9e7503f9df29310e |
ran
fingerprinted |
Apache-2.0 (permissive) |
| Brain Harmony: A Multimodal Foundation Model Unifying Morphology and Function into 1D Tokens added by Syntology |
2025-09 (from id) |
hzlab/Brain-Harmony/modules/harmonizer/stage1_pretrain/models.py 862cf06383f8bde7 |
ran · metamorphic tier: deterministic
|
no licence file found · pointer only |
| Weakly-Supervised Affordance Grounding Guided by Part-Level Semantic Priors |
30 May 2025 |
woyut/wsag-plsp/codes/models/decoder_affordance.py 7bf42bf0e6458cb1 |
ran · metamorphic tier: invariant
fingerprinted |
no licence file found · pointer only |
| AutoSSVH: Exploring Automated Frame Sampling for Efficient Self-Supervised Video Hashing |
4 Apr 2025 |
EliSpectre/CVPR25-AutoSSVH/model/AutoSSVH.py b7e965b5c4fbf144 |
ran · metamorphic tier: deterministic
fingerprinted |
no licence file found · pointer only |
| ChA-MAEViT: Unifying Channel-Aware Masked Autoencoders and Multi-Channel Vision Transformers for Improved Cross-Channel Learning |
25 Mar 2025 |
chaudatascience/cha_mae_vit/models/cha_mae_vit.py 4d93211cfa9bdffc |
ran · metamorphic tier: deterministic
|
MIT (permissive) |
| arXiv:2503.15141 |
2025-03 (from id) |
djukicn/ocebo/models/ocebo.py 15090130e3a1e904 |
ran · metamorphic tier: invariant
fingerprinted |
Apache-2.0 (permissive) |
| SVIP: Semantically Contextualized Visual Patches for Zero-Shot Learning |
13 Mar 2025 |
uqzhichen/SVIP/models/vit_model.py e8fe59329c1a7b8a |
ran
fingerprinted |
no licence file found · pointer only |
| VLog: Video-Language Models by Generative Retrieval of Narration Vocabulary |
12 Mar 2025 |
showlab/VLog/VLog/model/models.py eb3438f3ff649ddf |
ran · metamorphic tier: invariant
fingerprinted |
no licence file found · pointer only |
| Kronecker Mask and Interpretive Prompts are Language-Action Video Learners |
5 Feb 2025 |
yjyddq/CLAVER/models/claver.py 02c3d66d7386b4ce |
ran
fingerprinted |
no licence file found · pointer only |
| ControlAR: Controllable Image Generation with Autoregressive Models |
3 Oct 2024 |
hustvl/controlar/autoregressive/models/gpt.py 76f9e42efc0d4711 |
ran
fingerprinted |
Apache-2.0 (permissive) |
| CountGD: Multi-Modal Open-World Counting |
5 Jul 2024 |
niki-amini-naieni/countx/models_counting_network.py 59c07cf4d49e5e3c |
ran · metamorphic tier: invariant
fingerprinted |
MIT (permissive) |
| Semantically Guided Representation Learning For Action Anticipation |
2 Jul 2024 |
ADiko1997/S-GEAR/models/base_model_ts.py f4388cc6db5bce14 |
ran
fingerprinted |
Apache-2.0 (permissive) |
| Vision-LSTM: xLSTM as Generic Vision Backbone |
6 Jun 2024 |
NX-AI/vision-lstm/vision_lstm/vision_lstm.py 5fa540c6a1346171 |
ran · metamorphic tier: invariant
|
Apache-2.0 (permissive) |
| Enhancing Feature Diversity Boosts Channel-Adaptive Vision Transformers |
26 May 2024 |
chaudatascience/diverse_channel_vit/models/dichavit.py 07bed269ee55adea |
ran
fingerprinted |
MIT (permissive) |
| TSLANet: Rethinking Transformers for Time Series Representation Learning |
12 Apr 2024 |
WenjieDu/PyPOTS/pypots/nn/modules/tslanet/backbone.py 43c44182d498a36d |
ran · metamorphic tier: deterministic
fingerprinted |
BSD-3-Clause (permissive) |
| Bridging the Gap Between End-to-End and Two-Step Text Spotting |
6 Apr 2024 |
mxin262/bridging-text-spotting/adet/modeling/bridge.py 7d9a8a01e3d8b2af |
ran · metamorphic tier: deterministic
fingerprinted |
licence not identified · pointer only |
| On Train-Test Class Overlap and Detection for Image Retrieval |
1 Apr 2024 |
MCC-WH/Token/networks/RetrievalNet.py 938eafe395be5d19 |
ran · metamorphic tier: invariant
fingerprinted |
MIT (permissive) |
| Deep Generative Model based Rate-Distortion for Image Downscaling Assessment |
22 Mar 2024 |
IIGROUP/MANIQA/models/maniqa.py 93f0fde3bc773068 |
ran · metamorphic tier: invariant
fingerprinted |
Apache-2.0 (permissive) |
| Perceiving Longer Sequences With Bi-Directional Cross-Attention Transformers |
19 Feb 2024 |
mrkshllr/bixt/timm/models/bixt.py 59f48e4236839efb |
ran · metamorphic tier: deterministic
fingerprinted |
licence not identified · pointer only |
| Guiding Masked Representation Learning to Capture Spatio-Temporal Relationship of Electrocardiogram |
2 Feb 2024 |
bakqui/st-mem/models/st_mem.py 2b3233292ba7ed4f |
ran
fingerprinted |
licence not identified · pointer only |
| 4M: Massively Multimodal Masked Modeling |
11 Dec 2023 |
apple/ml-4m/fourm/models/fm.py 0e191e0cd5b78fae |
ran · metamorphic tier: deterministic
fingerprinted |
Apache-2.0 (permissive) |
| Eventful Transformers: Leveraging Temporal Redundancy in Vision Transformers |
25 Aug 2023 |
WISION-Lab/eventful-transformer/eventful_transformer/blocks.py f70a7b357e415e5e |
unverified |
MIT (permissive) |
| Adaptive Frequency Filters As Efficient Global Token Mixers |
26 Jul 2023 |
NWPU-Li/AFFNet/aff_block_LL.py 99f368d7aea9ef52 |
ran · metamorphic tier: invariant
fingerprinted |
no licence file found · pointer only |
| Efficient Video Action Detection with Token Dropout and Context Refinement |
17 Apr 2023 |
MCG-NJU/EVAD/projects/evad/models/vit_model.py f1ba368c5af15b64 |
ran · metamorphic tier: invariant
fingerprinted |
licence not identified · pointer only |
| DINOv2: Learning Robust Visual Features without Supervision |
14 Apr 2023 |
ByungKwanLee/Causal-Unsupervised-Segmentation/models/dinov2vit.py 22e81d05c408426c |
ran · metamorphic tier: deterministic
fingerprinted |
no licence file found · pointer only |
| Towards Unified Scene Text Spotting based on Sequence Generation |
7 Apr 2023 |
clovaai/units/units/models/model.py 2dfdf4c6074881e3 |
ran · metamorphic tier: deterministic
fingerprinted |
Apache-2.0 (permissive) |
| Efficient Computation Sharing for Multi-Task Visual Scene Understanding |
16 Mar 2023 |
sarashoouri/EfficientMTL/Codes/multimae/multimae.py e6e9f6bf7f63ce5e |
ran · metamorphic tier: invariant
fingerprinted |
no licence file found · pointer only |
| PiMAE: Point Cloud and Image Interactive Masked Autoencoders for 3D Object Detection |
14 Mar 2023 |
BLVLab/PiMAE/Pretrain/models/pimae.py 81da177eaf52ee01 |
ran
fingerprinted |
no licence file found · pointer only |
| Contrastive Model Adaptation for Cross-Condition Robustness in Semantic Segmentation |
9 Mar 2023 |
brdav/cma/models/model.py a6b6a8722e93e121 |
ran · metamorphic tier: deterministic
fingerprinted |
MIT (permissive) |
| Where We Are and What We're Looking At: Query Based Worldwide Image Geo-localization Using Hierarchies and Scenes |
7 Mar 2023 |
AHKerrigan/GeoGuessNet/networks.py 43b361f444ae05e1 |
ran · metamorphic tier: deterministic
fingerprinted |
no licence file found · pointer only |
| Using a Waffle Iron for Automotive Point Cloud Semantic Segmentation |
24 Jan 2023 |
valeoai/WaffleIron/waffleiron/backbone.py e058da1955e6f991 |
ran
fingerprinted |
licence not identified · pointer only |
| Scalable Adaptive Computation for Iterative Generation |
22 Dec 2022 |
google-research/pix2seq/architectures/transformers.py 80161ca59f374d2a |
ran
|
Apache-2.0 (permissive) |
| CroCo: Self-Supervised Pre-training for 3D Vision Tasks by Cross-View Completion |
19 Oct 2022 |
naver/croco/models/croco.py 99b1932affbf5fdf |
ran · metamorphic tier: deterministic
fingerprinted |
licence not identified · pointer only |
| Panoramic Vision Transformer for Saliency Detection in 360° Videos |
19 Sep 2022 |
hs-yn/paver/code/model/decoder.py 511f72cb4c20504a |
ran · metamorphic tier: invariant
fingerprinted |
MIT (permissive) |
| TinyViT: Fast Pretraining Distillation for Small Vision Transformers |
21 Jul 2022 |
microsoft/cream/TinyViT/models/tiny_vit.py a8aec28fef3b8a76 |
ran
fingerprinted |
MIT (permissive) |
| Global Context Vision Transformers |
20 Jun 2022 |
awsaf49/gcvit-tf/gcvit/models/gcvit.py cb662ef431014eb8 |
ran
|
MIT (permissive) |
| StarGraph: Knowledge Representation Learning based on Incomplete Two-hop Subgraph |
27 May 2022 |
hzli-ucas/stargraph/model.py f3aa96e814dfeea9 |
ran
fingerprinted |
no licence file found · pointer only |
| ViTPose: Simple Vision Transformer Baselines for Human Pose Estimation |
26 Apr 2022 |
gpastal24/ViTPose-Pytorch/src/vitpose_infer/builder/backbones/vit.py cd0e7be073ad785f |
ran · metamorphic tier: deterministic
fingerprinted |
GPL-3.0 (copyleft) · pointer only |
| ViTPose: Simple Vision Transformer Baselines for Human Pose Estimation |
26 Apr 2022 |
JunkyByte/easy_ViTPose/easy_ViTPose/vit_models/model.py bcec6a4434dcfdb8 |
ran
fingerprinted |
Apache-2.0 (permissive) |
| CRAFT: Cross-Attentional Flow Transformer for Robust Optical Flow |
31 Mar 2022 |
askerlee/craft/core/setrans.py 691dde70dfd8362b |
ran · metamorphic tier: deterministic
fingerprinted |
WTFPL · pointer only |
| End-to-End Transformer Based Model for Image Captioning |
29 Mar 2022 |
jchenghu/expansionnet_v2/models/End_ExpansionNet_v2.py d14d5773d2c8e320 |
ran
fingerprinted |
MIT (permissive) |
| Adaptive Patch Exiting for Scalable Single Image Super-Resolution |
22 Mar 2022 |
littlepure2333/APE/model/swinir_ape.py 2c28ad5498af1d70 |
ran · metamorphic tier: deterministic
fingerprinted |
no licence file found · pointer only |
| Global Matching with Overlapping Attention for Optical Flow Estimation |
21 Mar 2022 |
xiaofeng94/gmflownet/core/gmflownet_model.py b20207e9c582c326 |
ran · metamorphic tier: invariant
fingerprinted |
MIT (permissive) |
| UniFormer: Unified Transformer for Efficient Spatiotemporal Representation Learning |
12 Jan 2022 |
towhee-io/towhee/towhee/models/uniformer/uniformer.py 8f059bf3db11537b |
ran · metamorphic tier: deterministic
fingerprinted |
Apache-2.0 (permissive) |
| A ConvNet for the 2020s |
10 Jan 2022 |
murufeng/awesome_lightweight_networks/light_cnns/Transformer/ConvNeXt.py 6255e860ca403ee6 |
ran
fingerprinted |
MIT (permissive) |
| A ConvNet for the 2020s |
10 Jan 2022 |
DarshanDeshpande/jax-models/jax_models/models/convnext.py 983cc65c35c3d898 |
unverified |
Apache-2.0 (permissive) |
| A ConvNet for the 2020s |
10 Jan 2022 |
SarthakYadav/audax/audax/models/convnext.py 0c184b9cfad87e20 |
unverified |
BSD-2-Clause (permissive) |
| Masked Autoencoders Are Scalable Vision Learners |
11 Nov 2021 |
DarshanDeshpande/jax-models/jax_models/models/masked_autoencoder.py a84a5b2f271cf15c |
ran · metamorphic tier: deterministic
|
Apache-2.0 (permissive) |
| Masked Autoencoders Are Scalable Vision Learners |
11 Nov 2021 |
BUPT-PRIV/MAE-priv/mae/modeling_pretrain.py 762514af1d1eeb94 |
ran · metamorphic tier: invariant
fingerprinted |
licence not identified · pointer only |
| Contextual Similarity Aggregation with Self-attention for Visual Re-ranking |
26 Oct 2021 |
mcc-wh/csa/network/RerankTransformer.py 8b3435b073f323ec |
ran
fingerprinted |
MIT (permissive) |
| BEiT: BERT Pre-Training of Image Transformers |
15 Jun 2021 |
facebookresearch/vissl/vissl/models/trunks/beit_transformer.py f4d723c4060a9c88 |
ran · metamorphic tier: invariant
fingerprinted |
MIT (permissive) |
| RegionViT: Regional-to-Local Attention for Vision Transformers |
4 Jun 2021 |
dc3ea9f/RegionViT/models/region_vit.py 67186f084bc62090 |
ran
fingerprinted |
MIT (permissive) |
| SegFormer: Simple and Efficient Design for Semantic Segmentation with Transformers |
31 May 2021 |
IMvision12/SegFormer-tf/models/segformer.py 4b04a139b83d832e |
ran
fingerprinted |
no licence file found · pointer only |
| SegFormer: Simple and Efficient Design for Semantic Segmentation with Transformers |
31 May 2021 |
DarshanDeshpande/jax-models/jax_models/models/segformer.py 8085cbc641b30993 |
ran
|
Apache-2.0 (permissive) |
| TransMatcher: Deep Image Matching Through Transformers for Generalizable Person Re-identification |
30 May 2021 |
JDAI-CV/fast-reid/fastreid/modeling/backbones/vision_transformer.py 75f5c156267fccbd |
ran · metamorphic tier: deterministic
fingerprinted |
Apache-2.0 (permissive) |
| MLP-Mixer: An all-MLP Architecture for Vision |
4 May 2021 |
martinsbruveris/tensorflow-image-models/tfimm/architectures/mlp_mixer.py d011b7b14c9b2500 |
ran · metamorphic tier: invariant
fingerprinted |
Apache-2.0 (permissive) |
| Multiscale Vision Transformers |
22 Apr 2021 |
towhee-io/towhee/towhee/models/multiscale_vision_transformers/mvit.py 8eadaa524165fa27 |
ran
fingerprinted |
Apache-2.0 (permissive) |
| Swin Transformer: Hierarchical Vision Transformer using Shifted Windows |
25 Mar 2021 |
shkarupa-alex/tfswin/tfswin/model.py c0a6db98075e5a32 |
ran
|
MIT (permissive) |
| Swin Transformer: Hierarchical Vision Transformer using Shifted Windows |
25 Mar 2021 |
innat/VideoSwin/videoswin/blocks/swin_transformer.py 889677dfb48a1e7a |
ran
|
Apache-2.0 (permissive) |
| Swin Transformer: Hierarchical Vision Transformer using Shifted Windows |
25 Mar 2021 |
Burf/tfdetection/tfdet/model/backbone/swin_transformer.py 1f71103e68295350 |
ran
|
Apache-2.0 (permissive) |
| Swin Transformer: Hierarchical Vision Transformer using Shifted Windows |
25 Mar 2021 |
innat/HybridModel-GradCAM/layers/swin_blocks.py 5bcf14becf4f18ba |
ran
|
Apache-2.0 (permissive) |
| Swin Transformer: Hierarchical Vision Transformer using Shifted Windows |
25 Mar 2021 |
yangyangxu0/demt/src/model/backbones/swin.py 89aae91283eb968d |
ran
fingerprinted |
no licence file found · pointer only |
| Swin Transformer: Hierarchical Vision Transformer using Shifted Windows |
25 Mar 2021 |
rishigami/Swin-Transformer-TF/swintransformer/model.py 0fda0888f3c18df0 |
ran · metamorphic tier: invariant
fingerprinted |
Apache-2.0 (permissive) |
| Swin Transformer: Hierarchical Vision Transformer using Shifted Windows |
25 Mar 2021 |
DarshanDeshpande/jax-models/jax_models/models/swin_transformer.py 0267e8763dbcd8ad |
unverified |
Apache-2.0 (permissive) |
| Training data-efficient image transformers & distillation through attention |
23 Dec 2020 |
alibaba/EasyCV/easycv/models/backbones/vision_transformer.py c713760d26bf0a45 |
ran · metamorphic tier: invariant
fingerprinted |
Apache-2.0 (permissive) |
| Exploring Simple Siamese Representation Learning |
20 Nov 2020 |
facebookresearch/clip-rocket/models.py 6bd663b5be82dccd |
ran · metamorphic tier: deterministic
fingerprinted |
licence not identified · pointer only |
| An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale |
22 Oct 2020 |
mahmoodlab/hipt/HIPT_4K/vision_transformer.py ad5f80b54573c0e0 |
ran · metamorphic tier: deterministic
fingerprinted |
licence not identified · pointer only |
| An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale |
22 Oct 2020 |
OML-Team/open-metric-learning/oml/models/vit_dino/external_v2/vision_transformer.py 1d3ee770ca7b30c9 |
ran · metamorphic tier: deterministic
fingerprinted |
Apache-2.0 (permissive) |
| An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale |
22 Oct 2020 |
SHI-Labs/Compact-Transformers/src/vit.py e96ac00ca379ee0e |
ran
fingerprinted |
Apache-2.0 (permissive) |
| Attention Is All You Need |
12 Jun 2017 |
HzcIrving/DeepLearning_PlayGround/Swin-Transformer/Model.py fd903c745e98dffd |
ran
fingerprinted |
MIT (permissive) |
| Deep Networks with Stochastic Depth |
30 Mar 2016 |
nachiket273/pytorch_resnet_rs/model/base.py bd6c794dd7be5c10 |
ran
fingerprinted |
MIT (permissive) |
| Deep Networks with Stochastic Depth |
30 Mar 2016 |
DarshanDeshpande/jax-models/jax_models/layers/drop.py 7e0f6db857f04fb0 |
ran
|
Apache-2.0 (permissive) |
| arXiv:aaai_28529 |
|
AlienZhang1996/S2WAT/model/s2wat.py 7d33ccdf4bdeadbf |
ran · metamorphic tier: invariant
fingerprinted |
no licence file found · pointer only |
| arXiv:aaai_28388 |
|
924973292/TOP-ReID/modeling/fusion_part/CRM.py cb0614ea84e3c9e8 |
ran · metamorphic tier: invariant
fingerprinted |
MIT (permissive) |
| arXiv:Zhou_PanoLlama_Generating_Endless_and_Coherent_Panoramas_with_Next-Token-Prediction_LLMs_ICCV_2025_paper |
|
0606zt/PanoLlama/token_generator/gpt.py 1e7c67ed5e41cb0d |
unverified |
no licence file found · pointer only |