| RecPFN: Prior-Fitted Networks for In-Context-Based Recommendations added by Syntology |
2026-08 (from id) |
SAP-samples/tabular-ai-recpfn/src/architecture/recpfn.py 8153b54d2abfc622 |
unverified |
Apache-2.0 (permissive) |
| Hybrid-Adaptive Thread Tuning to Mitigate Simulation Execution Bottlenecks in High-Performance Reinforcement Learning Inference added by Syntology |
2026-08 (from id) |
suchenjm/AutoThread/python_algorithm/PINN.py 83d537ec97c49bd0 |
ran · metamorphic tier: invariant
|
MIT (permissive) |
| Text2Sign: A Single-GPU Diffusion Baseline for Text-to-Sign Language Video Generation added by Syntology |
2026-07 (from id) |
xiaruize0911/text2sign/pipeline.py c3d09298e5fbcef2 |
ran
fingerprinted |
no licence file found · pointer only |
| PACT: Self-Evolving Physical Safety Alignment for Diffusion Policies in Embodied Manipulation added by Syntology |
2026-06 (from id) |
thu-ml/RDT2/models/rdt/model.py 071967d102ce1b68 |
unverified |
Apache-2.0 (permissive) |
| TS-ICL: A Flexible Time-Indexed Foundation Model for Time Series via In-Context Learning added by Syntology |
2026-06 (from id) |
EDF-Lab/ts-icl/src/tsicl/model/network/ts_icl.py dbad070fabb35fcb |
ran · metamorphic tier: deterministic
|
licence not identified · pointer only |
| Envisioning Beyond the Few: Disentangled Semantics and Primitives for Few-Shot Atypical Layout-to-Image Generation added by Syntology |
2026-05 (from id) |
iCVTEAM/DSP/models/dsp/layers.py a820b721e43b1ab9 |
ran · metamorphic tier: invariant
|
MIT (permissive) |
| PQDT: Pseudo-Query Dual Transformer for Robust Point Cloud Restoration added by Syntology |
2026-05 (from id) |
ins-uni-bonn/PQDT/pqdt/models/pq_transformer.py 607eff327c3cb5c5 |
ran · metamorphic tier: deterministic
|
no licence file found · pointer only |
| λSPLIT: SELF-SUPERVISED CONTENT-AWARE SPECTRAL UNMIXING FOR FLUORESCENCE MICROSCOPY added by Syntology |
2026-03 (from id) |
preetam22n/DeepTrans-HSU/Trans_mod.py 38e30b6b7893e8d0 |
ran · metamorphic tier: invariant
|
no licence file found · pointer only |
| FILT3R: Latent State Adaptive Kalman Filter for Streaming 3D Reconstruction added by Syntology |
2026-03 (from id) |
jinotter3/FILT3R/src/dust3r/model.py 775287cd7a9a7949 |
ran
fingerprinted |
licence not identified · pointer only |
| Learning Transferable Sensor Models via Language-Informed Pretraining added by Syntology |
2026-03 (from id) |
yuc0805/SLIP/modeling_slip.py d936fd21ebf8cc5b |
ran
|
no licence file found · pointer only |
| Escaping the BLEU Trap: A Signal-Grounded Framework with Decoupled Semantic Guidance for EEG-to-Text Decoding added by Syntology |
2026-03 (from id) |
xmed-lab/SemKey/model/semkey_parallel.py 8ad54fe671ae56e3 |
ran
|
MIT (permissive) |
| Cortex-Grounded Diffusion Models for Brain Image Generation added by Syntology |
2026-01 (from id) |
ai-med/Cor2Vox/model/BrownianBridge/BrownianBridgeModel_c2v.py d4fae2e24ad5c07e |
ran
|
GPL-3.0 (copyleft) · pointer only |
| Cross360: 360°Monocular Depth Estimation via Cross Projections Across Scales added by Syntology |
2026-01 (from id) |
huangkun101230/Cross360/Networks/network.py 544581ee39313b81 |
ran
|
no licence file found · pointer only |
| TIMEPERCEIVER: An Encoder-Decoder Framework for Generalized Time-Series Forecasting added by Syntology |
2025-12 (from id) |
efficient-learning-lab/TimePerceiver/models/TimePerceiver.py ba95d81f47fd8823 |
ran · metamorphic tier: invariant
|
MIT (permissive) |
| AnyUp: Universal Feature Upsampling added by Syntology |
2025-10 (from id) |
wimmerth/anyup/anyup/model.py 05254353888b6fff |
ran
|
CC-BY-4.0 · pointer only |
| Latent Representation Learning in Heavy-Ion Collisions with MaskPoint Transformer added by Syntology |
2025-10 (from id) |
Giovanni-Sforza/MaskPoint-AMPT/models/MaskPoint.py eecd5279cc0854b7 |
ran · metamorphic tier: invariant
|
BSD-3-Clause (permissive) |
| Language Model Based Text-to-Audio Generation: Anti-Causally Aligned Collaborative Residual Transformers added by Syntology |
2025-10 (from id) |
wjc2830/Siren/src/siren/model.py cda35aded4b4d4c8 |
unverified |
no licence file found · pointer only |
| ReferSplat: Referring Segmentation in 3D Gaussian Splatting added by Syntology |
2025-08 (from id) |
heshuting555/ReferSplat/scene/cross_attention.py a116fcfa48d9680f |
unverified |
no licence file found · pointer only |
| Beyond Isolated Words: Diffusion Brush for Handwritten Text-Line Generation added by Syntology |
2025-08 (from id) |
dailenson/DiffBrush/models/unet.py aaf73d2a3d84cc5b |
ran
|
MIT (permissive) |
| Modulated Diffusion: Accelerating Generative Modeling with Modulated Quantization |
18 Jun 2025 |
WeizhiGao/MoDiff/qdiff/quant_model.py edbc6b070aba0b51 |
ran · metamorphic tier: deterministic
|
no licence file found · pointer only |
| Weakly-Supervised Affordance Grounding Guided by Part-Level Semantic Priors |
30 May 2025 |
woyut/wsag-plsp/codes/models/decoder_affordance.py 02f1b10113bdeb6d |
ran · metamorphic tier: invariant
|
no licence file found · pointer only |
| ZPressor: Bottleneck-Aware Compression for Scalable Feed-Forward 3DGS |
29 May 2025 |
ziplab/ZPressor/zpressor/zpressor.py 154432c6c478da51 |
ran
|
MIT (permissive) |
| LLaMA-Omni2: LLM-based Real-time Spoken Chatbot with Autoregressive Streaming Speech Synthesis |
5 May 2025 |
rainbowluocs/openomni/openomni/model/speech_generator_ar/speech_generator.py d5a46d215b2b4700 |
ran · metamorphic tier: deterministic
|
no licence file found · pointer only |
| arXiv:2504.20860 |
2025-04 (from id) |
mainaksingha01/FedMVP/model/FedMVP.py 172f846498770187 |
ran · metamorphic tier: invariant
|
MIT (permissive) |
| SeCap: Self-Calibrating and Adaptive Prompts for Cross-view Person Re-Identification in Aerial-Ground Networks |
10 Mar 2025 |
wangshining681/SeCap-AGPReID/fastreid/modeling/backbones/vision_transformer_multiview_SeCap.py 4ea2e6c979e1783b |
ran · metamorphic tier: deterministic
|
GPL-3.0 (copyleft) · pointer only |
| Generative Modeling and Data Augmentation for Power System Production Simulation |
10 Dec 2024 |
Becklishious/NeurIPS2024/TS-Diffusion/Models/interpretable_diffusion/gaussian_diffusion.py 70413232706855f0 |
ran · metamorphic tier: deterministic
|
no licence file found · pointer only |
| FlexDiT: Dynamic Token Density Control for Diffusion Transformer |
8 Dec 2024 |
changsn/FlexDiT/models.py ba088b40a6c22823 |
ran
|
no licence file found · pointer only |
| Scene Graph Generation with Role-Playing Large Language Models |
20 Oct 2024 |
guikunchen/sdsgg/maskrcnn_benchmark/modeling/roi_heads/relation_head/roi_relation_predictors.py c3b3f92371d4c055 |
ran
|
MIT (permissive) |
| MonoWAD: Weather-Adaptive Diffusion Model for Robust Monocular 3D Object Detection |
23 Jul 2024 |
VisualAIKHU/MonoWAD/visualDet3D/networks/detectors/MonoWAD.py 056357c296407ea2 |
unverified |
MIT (permissive) |
| Towards Latent Masked Image Modeling for Self-Supervised Visual Representation Learning |
22 Jul 2024 |
yibingwei-1/LatentMIM/models_lmim.py 0aa2506ff4669ad9 |
ran
|
no licence file found · pointer only |
| CountGD: Multi-Modal Open-World Counting |
5 Jul 2024 |
niki-amini-naieni/countx/models_counting_network.py 4cb253256cb32632 |
ran · metamorphic tier: invariant
|
MIT (permissive) |
| Cambrian-1: A Fully Open, Vision-Centric Exploration of Multimodal LLMs |
24 Jun 2024 |
cambrian-mllm/cambrian/cambrian/model/vision_sampler.py 7f1838302ff8b2bc |
ran
|
Apache-2.0 (permissive) |
| LDMol: Text-to-Molecule Diffusion Model with Structurally Informative Latent Space |
28 May 2024 |
jinhojsk515/ldmol/models.py 97fc89954c17d8f4 |
ran · metamorphic tier: deterministic
|
Apache-2.0 (permissive) |
| Diff-BGM: A Diffusion Model for Video Background Music Generation |
20 May 2024 |
sizhelee/Diff-BGM/diffbgm/stable_diffusion/model/unet.py ecb181d8f481c786 |
ran
|
no licence file found · pointer only |
| Structured Click Control in Transformer-based Interactive Segmentation |
7 May 2024 |
hahamyt/scc/isegm/model/modeling/gatlayers.py df5cfe9ca081ba57 |
ran · metamorphic tier: deterministic
|
GPL-3.0 (copyleft) · pointer only |
| Towards Highly Realistic Artistic Style Transfer via Stable Diffusion with Step-aware and Layer-aware Prompt |
17 Apr 2024 |
Jamie-Cheung/LSAST/cldm/cldm.py edd0dbdde64951f7 |
ran · metamorphic tier: deterministic
|
Apache-2.0 (permissive) |
| More Room for Language: Investigating the Effect of Retrieval on Language Models |
16 Apr 2024 |
ltgoslo/more-room-for-language/pretraining/model.py bb2cd17b7a096164 |
ran
|
GPL-3.0 (copyleft) · pointer only |
| Masked Autoencoders for Microscopy are Scalable Learners of Cellular Biology |
16 Apr 2024 |
recursionpharma/maes_microscopy/mae_modules.py 80cfc8ee2b854473 |
unverified |
licence not identified · pointer only |
| Weakly-supervised Audio Separation via Bi-modal Semantic Similarity |
2 Apr 2024 |
microsoft/bimodalaudioseparation/models/cond_unet_attn.py 837dd2dcaeaaded4 |
ran
|
MIT (permissive) |
| From Similarity to Superiority: Channel Clustering for Time Series Forecasting |
31 Mar 2024 |
graph-and-geometric-learning/timeseriesccm/models/layers.py cfa6cff859120675 |
unverified |
no licence file found · pointer only |
| LAKE-RED: Camouflaged Images Generation by Latent Background Knowledge Retrieval-Augmented Diffusion |
30 Mar 2024 |
PanchengZhao/LAKE-RED/ldm/ldm/models/diffusion/ddpm.py ff2c454a70dac1f9 |
ran · metamorphic tier: deterministic
|
Apache-2.0 (permissive) |
| EquiAV: Leveraging Equivariance for Audio-Visual Contrastive Learning |
14 Mar 2024 |
jongsuk1/equiav/models/pt_EquiAV.py d4d9aba462915ffd |
ran · metamorphic tier: deterministic
|
MIT (permissive) |
| Beyond Text: Frozen Large Language Models in Visual Signal Comprehension |
12 Mar 2024 |
zh460045050/V2L-Tokenizer/models/models_v2l.py 43b458c02b16d6b9 |
ran · metamorphic tier: invariant
|
no licence file found · pointer only |
| UPS: Efficiently Building Foundation Models for PDE Solving via Cross-Modal Adaptation |
11 Mar 2024 |
sjunhongshen/unifiedpdesolvers/embedder.py ad5b233df0fa6774 |
ran
|
MIT (permissive) |
| DEADiff: An Efficient Stylization Diffusion Model with Disentangled Representations |
11 Mar 2024 |
Tianhao-Qi/DEADiff_code/ldm/modules/new_attention.py e602decd5226f079 |
unverified |
Apache-2.0 (permissive) |
| MoST: Motion Style Transformer between Diverse Action Contents |
10 Mar 2024 |
Boeun-Kim/MoST/model/networks.py 1a6a8c5cc76b3eb9 |
ran
|
MIT (permissive) |
| UniTS: A Unified Multi-Task Time Series Model |
29 Feb 2024 |
mims-harvard/UniTS/models/UniTS.py ef2961b050617953 |
ran · metamorphic tier: invariant
|
MIT (permissive) |
| SeD: Semantic-Aware Discriminator for Image Super-Resolution |
29 Feb 2024 |
lbc12345/sed/models/sed.py b46f5f90925902c0 |
ran
|
no licence file found · pointer only |
| Perceiving Longer Sequences With Bi-Directional Cross-Attention Transformers |
19 Feb 2024 |
mrkshllr/bixt/timm/models/bixt.py 406496ff28d67b77 |
ran · metamorphic tier: deterministic
|
licence not identified · pointer only |
| NfgTransformer: Equivariant Representation Learning for Normal-form Games |
2024-02 (from id) |
google-deepmind/nfg_transformer/nfg_transformer/network.py 32711727edad3841 |
unverified |
Apache-2.0 (permissive) |
| MIGC: Multi-Instance Generation Controller for Text-to-Image Synthesis |
8 Feb 2024 |
limuloo/migc/migc/migc_arch.py c14fdb4777f76063 |
ran · metamorphic tier: invariant
fingerprinted |
licence not identified · pointer only |
| 4M: Massively Multimodal Masked Modeling |
11 Dec 2023 |
apple/ml-4m/fourm/models/fm.py 92c05b09aa8b432d |
ran · metamorphic tier: deterministic
|
Apache-2.0 (permissive) |
| Object-Centric Learning with Slot Mixture Module |
8 Nov 2023 |
airi-institute/smm/modules/smm.py ca05fae56d0a5950 |
ran · metamorphic tier: invariant
|
no licence file found · pointer only |
| CROMA: Remote Sensing Representations with Contrastive Radar-Optical Masked Autoencoders |
1 Nov 2023 |
antofuller/croma/pretrain_croma.py ba24c9bdfa918f05 |
ran · metamorphic tier: invariant
fingerprinted |
MIT (permissive) |
| Enhancing High-Resolution 3D Generation through Pixel-wise Gradient Clipping |
19 Oct 2023 |
fudan-zvg/pgc-3d/ldm/modules/diffusionmodules/openaimodel.py 93dcd06df2fc1896 |
ran · metamorphic tier: invariant
|
Apache-2.0 (permissive) |
| Global-correlated 3D-decoupling Transformer for Clothed Avatar Reconstruction |
24 Sep 2023 |
river-zhang/gta/lib/net/Transformer.py c770dc5d61e03875 |
ran · metamorphic tier: deterministic
|
no licence file found · pointer only |
| LayoutLLM-T2I: Eliciting Layout Guidance from LLM for Text-to-Image Generation |
9 Aug 2023 |
layoutllm-t2i/layoutllm-t2i/GLIGEN/ldm/modules/diffusionmodules/gligen_combine_layout.py 62dab2e5d80c65f2 |
ran
|
no licence file found · pointer only |
| SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis |
4 Jul 2023 |
stability-ai/generative-models/sgm/modules/diffusionmodules/openaimodel.py 75e42984360b91c5 |
unverified |
MIT (permissive) |
| MVDiffusion: Enabling Holistic Multi-view Image Generation with Correspondence-Aware Diffusion |
3 Jul 2023 |
Tangshitao/MVDiffusion/src/models/pano/MVGenModel.py 105b547188750534 |
ran
|
no licence file found · pointer only |
| GlyphControl: Glyph Conditional Control for Visual Text Generation |
29 May 2023 |
aigtext/glyphcontrol-release/cldm/cldm.py 8f64e619c23619c5 |
ran
|
MIT (permissive) |
| CommonScenes: Generating Commonsense 3D Indoor Scenes with Scene Graph Diffusion |
25 May 2023 |
ymxlzgy/commonscenes/model/networks/diffusion_networks/sg_diff.py feb6629ec7db268f |
ran · metamorphic tier: deterministic
|
no licence file found · pointer only |
| Harnessing the Spatial-Temporal Attention of Diffusion Models for High-Fidelity Text-to-Image Synthesis |
7 Apr 2023 |
ucsb-nlp-chang/diffusion-spacetime-attn/attention_optimization/stable-diffusion/ldm/modules/attention.py dbb6a642cadd75ab |
ran · metamorphic tier: deterministic
|
MIT (permissive) |
| Diffusion Models as Masked Autoencoders |
6 Apr 2023 |
kimdanni/DiffMAE/models_cross.py 563ed599da513e37 |
ran
|
licence not identified · pointer only |
| EGC: Image Generation and Classification via a Diffusion Energy-Based Model |
4 Apr 2023 |
GuoQiushan/EGC/guided_diffusion/unet.py 351f5f1606153b33 |
ran
|
MIT (permissive) |
| Freestyle Layout-to-Image Synthesis |
25 Mar 2023 |
essunny310/FreestyleNet/ldm/modules/attention_FLIS.py 3e38161420ae652c |
ran · metamorphic tier: deterministic
|
MIT (permissive) |
| Correlational Image Modeling for Self-Supervised Visual Pre-Training |
22 Mar 2023 |
weivision/correlational-image-modeling/models/cim.py 2e7d0ca574d0d946 |
ran
|
licence not identified · pointer only |
| DDFM: Denoising Diffusion Model for Multi-Modality Image Fusion |
13 Mar 2023 |
zhaozixiang1228/if-film/net/Film.py 553a596819206f47 |
ran · metamorphic tier: deterministic
|
no licence file found · pointer only |
| Domain Expansion of Image Generators |
12 Jan 2023 |
lllyasviel/controlnet/cldm/cldm.py 87dcf207c5c9187c |
ran · metamorphic tier: deterministic
|
Apache-2.0 (permissive) |
| TinyMIM: An Empirical Study of Distilling MIM Pre-trained Models |
3 Jan 2023 |
oliverrensu/d-igpt/DiGPT_torch/models_digpt.py ce2a200723aec3f8 |
ran · metamorphic tier: deterministic
|
no licence file found · pointer only |
| Training-Free Structured Diffusion Guidance for Compositional Text-to-Image Synthesis |
9 Dec 2022 |
weixi-feng/structured-diffusion-guidance/structured_stable_diffusion/modules/attention.py 3efa08f77f40bfd5 |
ran
|
licence not identified · pointer only |
| Generating astronomical spectra from photometry with conditional diffusion models |
2022-11 (from id) |
larsdoorenbos/generate-spectra/generative/unet.py 7eb48a00e07ad73f |
ran
|
MIT (permissive) |
| CroCo: Self-Supervised Pre-training for 3D Vision Tasks by Cross-View Completion |
19 Oct 2022 |
naver/croco/models/croco.py 863ed6dfef2b047a |
ran · metamorphic tier: deterministic
|
licence not identified · pointer only |
| Domain Randomization-Enhanced Depth Simulation and Restoration for Perceiving and Grasping Specular and Transparent Objects |
7 Aug 2022 |
PKU-EPIC/DREDS/CatePoseEstimation/networks/SwinDRNet.py 10375f1d1a2e05b2 |
unverified |
no licence file found · pointer only |
| Prompt-to-Prompt Image Editing with Cross Attention Control |
2 Aug 2022 |
chenwu98/cycle-diffusion/model/lib/stable_diffusion/ldm/modules/attention.py eb721f1e5050ad01 |
ran · metamorphic tier: deterministic
|
licence not identified · pointer only |
| PaLM: Scaling Language Modeling with Pathways |
5 Apr 2022 |
lucidrains/CoCa-pytorch/coca_pytorch/coca_pytorch.py 3e79daee4d06c337 |
ran · metamorphic tier: deterministic
|
MIT (permissive) |
| Bridging Video-text Retrieval with Multiple Choice Questions |
13 Jan 2022 |
towhee-io/towhee/towhee/models/bridgeformer/bridge_former_training_block.py 697a1de2168a5a59 |
ran · metamorphic tier: deterministic
|
Apache-2.0 (permissive) |
| Bridging Video-text Retrieval with Multiple Choice Questions |
13 Jan 2022 |
tencentarc/mcq/model/video_bridge_former.py fe726d5b3aa825e3 |
unverified |
no licence file found · pointer only |
| Improving language models by retrieving from trillions of tokens |
8 Dec 2021 |
labmlai/annotated_deep_learning_paper_implementations/labml_nn/transformers/retro/model.py c3e56d6f2af858c1 |
ran · metamorphic tier: invariant
|
MIT (permissive) |
| Learning to Prompt for Vision-Language Models |
2 Sep 2021 |
hhenryd/tap/trainers/TAP.py aa2102c5842a8045 |
ran · metamorphic tier: invariant
|
MIT (permissive) |
| Perceiver IO: A General Architecture for Structured Inputs & Outputs |
30 Jul 2021 |
esceptico/perceiver-io/src/perceiver_io/perceiver.py cf3adb72e7c5e13d |
ran
|
MIT (permissive) |
| CrossViT: Cross-Attention Multi-Scale Vision Transformer for Image Classification |
27 Mar 2021 |
rishikksh20/CrossViT-pytorch/crossvit.py 8abca80f7f8a5dda |
ran
|
MIT (permissive) |
| CrossViT: Cross-Attention Multi-Scale Vision Transformer for Image Classification |
27 Mar 2021 |
IBM/CrossViT/models/crossvit.py 9811d06fb8557fa5 |
ran · metamorphic tier: deterministic
|
Apache-2.0 (permissive) |
| Perceiver: General Perception with Iterative Attention |
4 Mar 2021 |
kietngt00/hmdb51-recognition/model/attention.py 122a2bb1f3342e5c |
ran
|
no licence file found · pointer only |
| Perceiver: General Perception with Iterative Attention |
4 Mar 2021 |
towhee-io/towhee/towhee/models/perceiver/cross_attention.py ce5db0a13e19e003 |
ran · metamorphic tier: invariant
fingerprinted |
Apache-2.0 (permissive) |
| Pre-training via Paraphrasing |
26 Jun 2020 |
lucidrains/marge-pytorch/marge_pytorch/marge_pytorch.py fb7817bba2fe0a0b |
ran · metamorphic tier: deterministic
|
MIT (permissive) |
| arXiv:Lu_FedHCA2_Towards_Hetero-Client_Federated_Multi-Task_Learning_CVPR_2024_paper |
|
innovator-zero/FedHCA2/models/hyperweight.py 0bbea3081c4b97ef |
unverified |
MIT (permissive) |
| arXiv:Koo_PartGlot_Learning_Shape_Part_Segmentation_From_Language_Reference_Games_CVPR_2022_paper |
|
63days/PartGlot/partglot/models/pn_agnostic.py 69f36807fe0a43bd |
unverified |
no licence file found · pointer only |
| arXiv:2024.acl-long.332 |
|
BUAAw-ML/SKP/lavis/models/adapt_module/amortized_encdec.py aa1cf692d15daa47 |
ran · metamorphic tier: deterministic
|
BSD-3-Clause (permissive) |