| Learning task-specific subspaces via interventional post-training of speech foundation models added by Syntology |
2026-06 (from id) |
s3prl/s3prl/s3prl/upstream/distiller/model.py 6f76d015b2817409 |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| CausalMoE: A Billion-Scale Multimodal Foundation Model for Granger Causal Discovery with Pattern-Routed Heterogeneous Experts added by Syntology |
2026-06 (from id) |
liubolab/CausalMoE/PatchTST_layers.py 72b9272542f279c6 |
ran
|
no licence file found · pointer only |
| Quantifying the Pre-training Dividend: Generative versus Latent Self-Supervised Learning for Time Series Foundation Models added by Syntology |
2026-05 (from id) |
noammajor/Models/NTP/models/layers/basics.py 6db91a5d47822c1a |
ran
|
no licence file found · pointer only |
| A Geometric Analysis of Sign-Magnitude Asymmetry in a ReLU + RMSNorm Block under Ternary Quantization added by Syntology |
2026-05 (from id) |
donglei628/sign-magnitude-pre-norm/experiments/exp5_multilayer_toy.py 5fec7274c7f22bcc |
ran
|
MIT (permissive) |
| Temporal Patch Shuffle (TPS): Leveraging Patch-Level Shuffling to Boost Generalization and Robustness in Time Series Forecasting added by Syntology |
2026-04 (from id) |
jafarbakhshaliyev/TPS/time_series_forecasting/layers/PatchTST_layers.py 72b9272542f279c6 |
ran
|
no licence file found · pointer only |
| Scaling Open Discrete Audio Foundation Models with Interleaved Semantic, Acoustic, and Text Tokens added by Syntology |
2026-02 (from id) |
BytedanceSpeech/seed-tts-eval/thirdparty/UniSpeech/WavLM/modules.py 01233f4856456c5f |
unverified |
no licence file found · pointer only |
| ZeroSyl: Simple Zero-Resource Syllable Tokenization for Spoken Language Modeling added by Syntology |
2026-02 (from id) |
nicolvisser/ZeroSyl/zerosyl/zerosyl.py 591bd7145f129c23 |
ran · our draft was wrong
|
MIT (permissive) |
| Video-based Music Generation added by Syntology |
2026-02 (from id) |
serkansulun/video-emotion/src/beats/modules.py 01233f4856456c5f |
unverified |
no licence file found · pointer only |
| Vanish into Thin Air: Cross-prompt Universal Adversarial Attacks for SAM2 added by Syntology |
2025-10 (from id) |
CGCL-codes/UAP-SAM2/sam2/modeling/sam2_utils.py f2e72d81f39d65cd |
unverified |
no licence file found · pointer only |
| Abstain Mask Retain Core: Time Series Prediction by Adaptive Masking Loss with Representation Consistency added by Syntology |
2025-10 (from id) |
MazelTovy/AMRC/PatchTST/layers/PatchTST_layers.py 72b9272542f279c6 |
ran
|
no licence file found · pointer only |
| SEMPO: Lightweight Foundation Models for Time Series Forecasting added by Syntology |
2025-10 (from id) |
mala-lab/SEMPO/models/SEMPO.py cf01255b103815b4 |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| Situat3DChange: Situated 3D Change Understanding Dataset for Multimodal Large Language Model added by Syntology |
2025-10 (from id) |
RuipingL/Situat3DChange/SCReasoner/model/utils.py 165417eb35c55e89 |
unverified |
CC-BY-4.0 · pointer only |
| Robust Ego-Exo Correspondence with Long-Term Memory added by Syntology |
2025-10 (from id) |
juneyeeHu/LM-EEC/sam2/modeling/sam2_utils.py f2e72d81f39d65cd |
unverified |
no licence file found · pointer only |
| Synthetic Series-Symbol Data Generation for Time Series Foundation Models added by Syntology |
2025-10 (from id) |
wwhenxuan/SymTime/models/pretrain_model.py bb11ef32f48c0c09 |
ran · our draft was wrong
|
MIT (permissive) |
| ReNF: Rethinking the Design of Neural Long-Term Time Series Forecasters added by Syntology |
2025-09 (from id) |
Luoauoa/ReNF/layers/PatchTST_layers.py 72b9272542f279c6 |
ran
|
MIT (permissive) |
| arXiv:2507.22699 |
2025-07 (from id) |
dvttran/nsft/nsft/models/base_model.py e4b131bffbe89293 |
unverified |
MIT (permissive) |
| arXiv:2506.23596 |
2025-06 (from id) |
KU-VGI/AP/layers/PatchTST_layers.py 72b9272542f279c6 |
ran
|
Apache-2.0 (permissive) |
| MiniMax-M1: Scaling Test-Time Compute Efficiently with Lightning Attention |
16 Jun 2025 |
minimax-ai/minimax-m1/modeling_minimax_m1.py 5e3dd6e650adb72c |
unverified |
Apache-2.0 (permissive) |
| Prisma: An Open Source Toolkit for Mechanistic Interpretability in Vision and Video |
28 Apr 2025 |
prisma-multimodal/vit-prisma/src/vit_prisma/sae/sae.py bd6e1088e609f3eb |
ran · our draft was wrong
|
MIT (permissive) |
| Caption Anything in Video: Fine-grained Object-centric Captioning via Spatiotemporal Multimodal Prompting |
7 Apr 2025 |
yunlong10/CAT-V/sam2/modeling/sam2_utils.py f2e72d81f39d65cd |
unverified |
BSD-3-Clause (permissive) |
| Multi-Flow: Multi-View-Enriched Normalizing Flows for Industrial Anomaly Detection |
4 Apr 2025 |
m-kruse98/multi-flow/models/MVANet.py f2e72d81f39d65cd |
unverified |
MIT (permissive) |
| MolSpectra: Pre-training 3D Molecular Representation with Multi-modal Energy Spectra |
22 Feb 2025 |
azureleon1/molspectra/torchmdnet/models/SpecFormer_layers.py 2d15324019001de4 |
ran · our draft was wrong
|
no licence file found · pointer only |
| MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation |
23 Jan 2025 |
rongfu-dsb/MPG-SAM2/models/sam2/modeling/sam2_utils.py f2e72d81f39d65cd |
unverified |
Apache-2.0 (permissive) |
| MetaLA: Unified Optimal Linear Approximation to Softmax Attention Map |
16 Nov 2024 |
BICLab/MetaLA/metala/utils.py cc7630717878e956 |
unverified |
no licence file found · pointer only |
| SAM2Long: Enhancing SAM 2 for Long Video Segmentation with a Training-Free Memory Tree |
21 Oct 2024 |
mark12ding/sam2long/sam2/modeling/sam2_utils.py f2e72d81f39d65cd |
unverified |
no licence file found · pointer only |
| Anonymising Elderly and Pathological Speech: Voice Conversion Using DDSP and Query-by-Example |
20 Oct 2024 |
suhitaghosh10/ddsp-qbe/wavlm/modules.py 01233f4856456c5f |
unverified |
no licence file found · pointer only |
| Simultaneous Weight and Architecture Optimization for Neural Networks |
10 Oct 2024 |
zitonghuangcynthia/Simultaneous-Weight-and-Architecture-Optimization/training_autoencoder/train_autoencoder.py 47f13a28826d6cf6 |
ran
|
no licence file found · pointer only |
| MixLinear: Extreme Low Resource Multivariate Time Series Forecasting with 0.1K Parameters |
2 Oct 2024 |
lss-1138/SparseTSF/layers/PatchTST_layers.py 72b9272542f279c6 |
ran
|
Apache-2.0 (permissive) |
| SAM2-UNet: Segment Anything 2 Makes Strong Encoder for Natural and Medical Image Segmentation |
16 Aug 2024 |
wzh0120/sam2-unet/sam2/modeling/sam2_utils.py f2e72d81f39d65cd |
unverified |
Apache-2.0 (permissive) |
| Surgical SAM 2: Real-time Segment Anything in Surgical Video by Efficient Frame Pruning |
15 Aug 2024 |
jinlab-imvr/surgical-sam-2/sam2/modeling/sam2_utils.py f2e72d81f39d65cd |
unverified |
Apache-2.0 (permissive) |
| Medical SAM 2: Segment medical images as video via Segment Anything Model 2 |
1 Aug 2024 |
medicinetoken/medical-sam2/sam2_train/modeling/sam2_utils.py f2e72d81f39d65cd |
unverified |
Apache-2.0 (permissive) |
| Score matching for bridges without learning time-reversals |
22 Jul 2024 |
libbylbaker/forward_bridge/src/models/neuralop.py eb153707eee11c99 |
ran
|
MIT (permissive) |
| SEED: A Simple and Effective 3D DETR in Point Clouds |
15 Jul 2024 |
happinesslz/SEED/pcdet/models/model_utils/seed_transformer.py f2e72d81f39d65cd |
unverified |
no licence file found · pointer only |
| TimeCMA: Towards LLM-Empowered Multivariate Time Series Forecasting via Cross-Modality Alignment |
3 Jun 2024 |
chenxiliu-hnu/timecma/layers/TS_Pos_Enc.py 72b9272542f279c6 |
ran
|
no licence file found · pointer only |
| SST: Multi-Scale Hybrid Mamba-Transformer Experts for Long-Short Range Time Series Forecasting |
23 Apr 2024 |
xiongxiaoxu/sst/layers/LWT_layers.py 72b9272542f279c6 |
ran
|
no licence file found · pointer only |
| Multi-view Aggregation Network for Dichotomous Image Segmentation |
11 Apr 2024 |
qianyu-dlut/mvanet/model/MVANet.py f2e72d81f39d65cd |
unverified |
MIT (permissive) |
| SERVAL: Synergy Learning between Vertical Models and LLMs towards Oracle-Level Zero-shot Medical Prediction |
3 Mar 2024 |
jyansir/sersal/models/t2g.py 7627d2c02619ab87 |
ran · our draft was wrong
|
MIT (permissive) |
| ConvTimeNet: A Deep Hierarchical Fully Convolutional Model for Multivariate Time Series Analysis |
3 Mar 2024 |
mingyue-cheng/convtimenet/TSForecasting/layers/ConvTimeNet_backbone.py 72b9272542f279c6 |
ran
|
no licence file found · pointer only |
| ConvTimeNet: A Deep Hierarchical Fully Convolutional Model for Multivariate Time Series Analysis |
3 Mar 2024 |
mingyue-cheng/convtimenet/TSClassification/models/ConvTimeNet_backbone.py 5c170505ad8d13d2 |
ran
|
no licence file found · pointer only |
| Rethinking Channel Dependence for Multivariate Time Series Forecasting: Learning from Leading Indicators |
31 Jan 2024 |
SJTU-DMTai/LIFT/layers/PatchTST_layers.py 72b9272542f279c6 |
ran
|
no licence file found · pointer only |
| An Embodied Generalist Agent in 3D World |
18 Nov 2023 |
embodied-generalist/embodied-generalist/model/utils.py 165417eb35c55e89 |
unverified |
MIT (permissive) |
| Multi-resolution Time-Series Transformer for Long-term Forecasting |
7 Nov 2023 |
Yitiann/MTST/layers/PatchTST_layers.py 72b9272542f279c6 |
ran
|
no licence file found · pointer only |
| LocoMuJoCo: A Comprehensive Imitation Learning Benchmark for Locomotion |
4 Nov 2023 |
robfiras/loco-mujoco/loco_mujoco/algorithms/common/networks.py 7efd09ca70f8e666 |
ran
|
MIT (permissive) |
| 19 Parameters Is All You Need: Tiny Neural Networks for Particle Physics |
24 Oct 2023 |
abogatskiy/PELICAN-nano/src/layers/generic_layers.py c3b4e5ac5f9963cc |
ran
|
MIT (permissive) |
| Calibration of Time-Series Forecasting: Detecting and Adapting Context-Driven Distribution Shift |
23 Oct 2023 |
half111/calibration_cds/layers/PatchTST_layers.py 72b9272542f279c6 |
ran
|
MIT (permissive) |
| iTransformer: Inverted Transformers Are Effective for Time Series Forecasting |
10 Oct 2023 |
lss-1138/SegRNN/layers/PatchTST_layers.py 72b9272542f279c6 |
ran
|
Apache-2.0 (permissive) |
| Cross-modal Cognitive Consensus guided Audio-Visual Segmentation |
10 Oct 2023 |
zhaofengshi/avs-c3n/C3N_beats/avs_scripts/avs_ms3/beats/modules.py 01233f4856456c5f |
unverified |
no licence file found · pointer only |
| PatchMixer: A Patch-Mixing Architecture for Long-Term Time Series Forecasting |
1 Oct 2023 |
Zeying-Gong/PatchMixer/layers/PatchTST_layers.py 72b9272542f279c6 |
ran
|
MIT (permissive) |
| Improving Audio Captioning Models with Fine-grained Audio Features, Text Embedding Supervision, and LLM Mix-up Augmentation |
2023-09 (from id) |
slseanwu/beats-conformer-bart-audio-captioner/model/modules.py 01233f4856456c5f |
unverified |
Apache-2.0 (permissive) |
| Diverse and Aligned Audio-to-Video Generation via Text-to-Video Model Adaptation |
28 Sep 2023 |
guyyariv/TempoTokens/modules/beats/modules.py 01233f4856456c5f |
unverified |
MIT (permissive) |
| Unsupervised Open-Vocabulary Object Localization in Videos |
18 Sep 2023 |
amazon-science/object-centric-vol/models/model_grouping.py 94db904018002114 |
ran
|
Apache-2.0 (permissive) |
| PETformer: Long-term Time Series Forecasting via Placeholder-enhanced Transformer |
9 Aug 2023 |
ACAT-SCUT/PETformer/layers/PatchTST_layers.py 72b9272542f279c6 |
ran
|
no licence file found · pointer only |
| AudioToken: Adaptation of Text-Conditioned Diffusion Models for Audio-to-Image Generation |
22 May 2023 |
guyyariv/AudioToken/modules/BEATs/modules.py 01233f4856456c5f |
unverified |
MIT (permissive) |
| Toeplitz Neural Network for Sequence Modeling |
8 May 2023 |
Doraemonzzz/tnn-pytorch/tnn_pytorch/tno.py 9a369f7c4e13695c |
ran · our draft was wrong
|
MIT (permissive) |
| Heterogeneous Multi-Robot Reinforcement Learning |
17 Jan 2023 |
proroklab/hetgppo/models/gppo.py 4a6e23b2b8a26945 |
ran · our draft was wrong
|
no licence file found · pointer only |
| BEATs: Audio Pre-Training with Acoustic Tokenizers |
18 Dec 2022 |
phuriches/genrepasd/beats/modules.py 01233f4856456c5f |
unverified |
MIT (permissive) |
| PELICAN: Permutation Equivariant and Lorentz Invariant or Covariant Aggregator Network for Particle Physics |
1 Nov 2022 |
abogatskiy/pelican/src/layers/generic_layers.py c3b4e5ac5f9963cc |
ran
|
MIT (permissive) |
| Wespeaker: A Research and Production oriented Speaker Embedding Learning Toolkit |
2022-10 (from id) |
BUTSpeechFIT/wespeaker_ssl_public/wespeaker/models/ssl/modules.py 01233f4856456c5f |
unverified |
Apache-2.0 (permissive) |
| Multi-Objective GFlowNets |
23 Oct 2022 |
samuelstanton/lambo/lambo/models/shared_elements.py 79857710f232bbc6 |
unverified |
Apache-2.0 (permissive) |
| Bridging the Gap to Real-World Object-Centric Learning |
29 Sep 2022 |
amazon-science/object-centric-learning-framework/ocl/neural_networks/convenience.py 1d2c0e37157e1caf |
unverified |
Apache-2.0 (permissive) |
| OSSGAN: Open-Set Semi-Supervised Image Generation |
29 Apr 2022 |
raven38/ossgan/src/models/big_resnet.py 8945689abdf62ec1 |
ran · our draft was wrong
|
licence not identified · pointer only |
| Boosting Self-Supervised Embeddings for Speech Enhancement |
7 Apr 2022 |
khhungg/BSSE-SE/models/modules.py 01233f4856456c5f |
unverified |
MIT (permissive) |
| UTMOS: UTokyo-SaruLab System for VoiceMOS Challenge 2022 |
2022-04 (from id) |
sarulab-speech/utmos22/strong/modules.py 01233f4856456c5f |
unverified |
MIT (permissive) |
| GradMax: Growing Neural Networks using Gradient Information |
13 Jan 2022 |
google-research/growneuron/growneuron/layers.py f5060e86ebb597e4 |
unverified |
Apache-2.0 (permissive) |
| Multimeasurement Generative Models |
18 Dec 2021 |
nnaisense/mems/src/mems/archs/resnet.py 8ba4f0d9f8238d90 |
unverified |
BSD-3-Clause (permissive) |
| Revisiting Deep Learning Models for Tabular Data |
22 Jun 2021 |
jrzaurin/pytorch-widedeep/pytorch_widedeep/models/_get_activation_fn.py 755bddefe05fb03e |
unverified |
Apache-2.0 (permissive) |
| SimCSE: Simple Contrastive Learning of Sentence Embeddings |
18 Apr 2021 |
jeongukjae/KR-BERT-SimCSE/model.py f16e63fb25fd4d54 |
ran · our draft was wrong
|
MIT (permissive) |
| arXiv:aaai_29295 |
|
Jack24658735/FedLGT/models/utils.py efb77cb95297b034 |
unverified |
Apache-2.0 (permissive) |