| Diagnosing Temporal Misalignment in Multichannel Time-Series Classification via Minimum Description Length added by Syntology |
2026-09 (from id) |
sbuschjaeger/mdl-temporal-misalignment/src/data.py ce8615a3bc5c4230 |
unverified |
no licence file found · pointer only |
| Do General NLP Embeddings Capture Ontological Reasoning? added by Syntology |
2026-09 (from id) |
sciknoworg/AVA/src/utils.py a7ce6b13fccd9850 |
unverified |
MIT (permissive) |
| QASM-Eval: A Dataset to Train and Evaluate LLMs on OpenQASM-3 Beyond Quantum Circuits added by Syntology |
2026-05 (from id) |
fuzhenxiao/QASM-Eval/dataset_factory/generate_dataset.py b35350dd508a219b |
unverified |
no licence file found · pointer only |
| Muon in Vision Transformers: Optimizer-Recipe Interactions and Gradient Spectra added by Syntology |
2026-05 (from id) |
facebookresearch/mae/util/datasets.py 27f8c4c706fd43ec |
ran
|
licence not identified · pointer only |
| Balancing Knowledge Distillation for Imbalance Learning with Bilevel Optimization added by Syntology |
2026-05 (from id) |
phan-tho/Bilevel-balancing-kd/cifar.py 64027512b5c071fd |
unverified |
no licence file found · pointer only |
| ESGLens: An LLM-Based RAG Framework for Interactive ESG Report Analysis and Score Prediction added by Syntology |
2026-04 (from id) |
tyy4ng/ESGLens/esg_2_dataprocessing.py 05addca96cbd33e4 |
unverified |
no licence file found · pointer only |
| Efficient Bilevel Optimization with KFAC-Based Hypergradients added by Syntology |
2026-03 (from id) |
xjtushujun/Meta-weight-net_class-imbalance/data_utils.py 8f873dc2f47b8f1c |
unverified |
MIT (permissive) |
| Mixture-of-Top-k Attention: Efficient Attention via Scalable Fast Weights added by Syntology |
2026-02 (from id) |
QishuaiWen/MiTA/MiTA-DeiT/datasets.py d58967bba23da4e6 |
unverified |
MIT (permissive) |
| Understanding the Transfer Limits of Vision Foundation Models added by Syntology |
2026-01 (from id) |
pimed/ProViCNet/ProViCNet/ModelArchitectures/LeViTUnet/datasets.py 1a770e29157e6b41 |
unverified |
MIT (permissive) |
| Generalizing Abstention for Noise-Robust Learning in Medical Image Segmentation added by Syntology |
2026-01 (from id) |
wemous/abstention-for-segmentation/datasets/cadis.py 07c94b759b069d9f |
unverified |
MIT (permissive) |
| Generalizing Abstention for Noise-Robust Learning in Medical Image Segmentation added by Syntology |
2026-01 (from id) |
wemous/abstention-for-segmentation/datasets/dsad.py f64d395578fc446c |
unverified |
MIT (permissive) |
| Inference-Time Scaling for Visual AutoRegressive modeling by Searching Representative Samples added by Syntology |
2026-01 (from id) |
WD7ang/VAR-Scaling/VAR-main/utils/data.py 8ff7703f731423de |
unverified |
MIT (permissive) |
| Crowded Video Individual Counting Informed by Social Grouping and Spatial-Temporal Displacement Priors added by Syntology |
2026-01 (from id) |
tiny-smart/OMAN/datasets/Sense_dataset.py 0628e16af7cd6425 |
unverified |
MIT (permissive) |
| Eliminating Inductive Bias in Reward Models with Information-Theoretic Guidance added by Syntology |
2025-12 (from id) |
Qwen-Applications/DIR/reward_models/load_datasets.py 47ee1e29f8388bb2 |
unverified |
Apache-2.0 (permissive) |
| APLOT: Robust Reward Modeling via Adaptive Preference Learning with Optimal Transport added by Syntology |
2025-10 (from id) |
BIRlz/APLOT/reward_models/load_datasets.py 47ee1e29f8388bb2 |
unverified |
no licence file found · pointer only |
| arXiv:2507.22264 |
2025-07 (from id) |
LAION-AI/CLIP_benchmark/clip_benchmark/datasets/builder.py 33133229b84c5784 |
unverified |
MIT (permissive) |
| AgentStealth: Reinforcing Large Language Model for Anonymizing User-generated Text |
26 Jun 2025 |
tsinghua-fib-lab/AgentStealth/rl/train_single.py 17b0f353027f5a92 |
unverified |
MIT (permissive) |
| SpikeVideoFormer: An Efficient Spike-Driven Video Transformer with Hamming Attention and $\mathcal{O}(T)$ Complexity |
15 May 2025 |
jimmyzou/spikevideoformer/classification/dataloader/datasets.py 9bcbccb90a4be753 |
unverified |
Apache-2.0 (permissive) |
| Efficient Token Compression for Vision Transformer with Spatial Information Preserved |
30 Mar 2025 |
nust-machine-intelligence-laboratory/prune_and_merge/pm-vit/datasets.py 43f5d3f52fed74d4 |
unverified |
Apache-2.0 recorded; this copy not marked cleared · pointer only |
| SeisMoLLM: Advancing Seismic Monitoring via Cross-modal Transfer with Pre-trained Large Language Model |
27 Feb 2025 |
StarMoonWang/SeisMoLLM/datasets/_factory.py 06f95911e0f86891 |
unverified |
MIT (permissive) |
| Battling the Non-stationarity in Time Series Forecasting via Test-time Adaptation |
9 Jan 2025 |
kimanki/tafas/tta/tafas.py bdf4a9030fd490ad |
unverified |
licence not identified · pointer only |
| Infinity: Scaling Bitwise AutoRegressive Modeling for High-Resolution Image Synthesis |
5 Dec 2024 |
FoundationVision/VAR/utils/data.py 8ff7703f731423de |
unverified |
MIT (permissive) |
| Revisiting the Integration of Convolution and Attention for Vision Backbone |
21 Nov 2024 |
rayleizhu/GLMix/datasets.py 9153241c4331eb1a |
unverified |
MIT (permissive) |
| Distill the Best, Ignore the Rest: Improving Dataset Distillation with Loss-Value-Based Pruning |
18 Nov 2024 |
Brian-Moser/prune_and_distill/src/glad_utils.py 8ac263b783faa918 |
unverified |
no licence file found · pointer only |
| M-VAR: Decoupled Scale-wise Autoregressive Modeling for High-Quality Image Generation |
15 Nov 2024 |
oliverrensu/mvar/utils/data.py 8ff7703f731423de |
unverified |
no licence file found · pointer only |
| Coevolving with the Other You: Fine-Tuning LLM with Sequential Cooperative Multi-Agent Reinforcement Learning |
8 Oct 2024 |
Harry67Hu/CORY/gsm8k_utils/gsm8k_eval.py 0efa9af73116efc0 |
unverified |
MIT (permissive) |
| Famba-V: Fast Vision Mamba with Cross-Layer Token Fusion |
15 Sep 2024 |
aiot-mlsys-lab/famba-v/fambav/datasets.py 80ba95ca0da5ce5f |
unverified |
no licence file found · pointer only |
| Domain-Adaptive 2D Human Pose Estimation via Dual Teachers in Extremely Low-Light Conditions |
22 Jul 2024 |
ayh015-dev/da-llpose/lib/dataset/build.py 983e45bff55979f5 |
unverified |
no licence file found · pointer only |
| On the Role of Discrete Tokenization in Visual Representation Learning |
12 Jul 2024 |
PKU-ML/ClusterMIM/util/datasets.py 27f8c4c706fd43ec |
ran
|
no licence file found · pointer only |
| LLMBox: A Comprehensive Library for Large Language Models |
8 Jul 2024 |
RUCAIBox/LLMBox/training/ppo.py c53aa72d3d68339b |
unverified |
MIT (permissive) |
| Elliptical Attention |
19 Jun 2024 |
stefvk/Elliptical-Attention/ImageNet/datasets.py 9153241c4331eb1a |
unverified |
no licence file found · pointer only |
| Unveiling the Hidden Structure of Self-Attention via Kernel Principal Component Analysis |
19 Jun 2024 |
rachtsy/KPCA_code/Attack/datasets.py 9153241c4331eb1a |
unverified |
no licence file found · pointer only |
| Scaling Efficient Masked Image Modeling on Large Remote Sensing Dataset |
17 Jun 2024 |
Fengxiang23/SelectiveMAE/SelectiveMAE/util/datasets.py 519bbbff76c0ed7c |
ran
|
MIT (permissive) |
| Regularizing Hidden States Enables Learning Generalizable Reward Model for LLMs |
14 Jun 2024 |
yangrui2015/generalizable-reward-model/reward_models/load_datasets.py 6a00527237189ca7 |
unverified |
MIT (permissive) |
| Autoregressive Pretraining with Mamba in Vision |
11 Jun 2024 |
oliverrensu/arm/Finetuning/util/datasets.py affa54d3baff9b37 |
ran
|
no licence file found · pointer only |
| OphNet: A Large-Scale Video Benchmark for Ophthalmic Surgical Workflow Understanding |
11 Jun 2024 |
minghu0830/ophnet-benchmark/baselines/task2/backbone/videomaev2/dataset/build.py 283de3b7dcb770a3 |
unverified |
MIT (permissive) |
| Towards noise contrastive estimation with soft targets for conditional models |
22 Apr 2024 |
uhlmanngroup/soft-target-infonce/util/data.py 1f2ddd312fd8be3d |
unverified |
no licence file found · pointer only |
| Adapting LLaMA Decoder to Vision Transformer |
10 Apr 2024 |
techmonsterwang/illama/datasets.py 9f7ea37592e1e7c7 |
unverified |
no licence file found · pointer only |
| Improving Visual Recognition with Hyperbolical Visual Hierarchy Mapping |
1 Apr 2024 |
kwonjunn01/Hi-Mapper/datasets.py d58967bba23da4e6 |
unverified |
Apache-2.0 recorded; this copy not marked cleared · pointer only |
| A General and Efficient Training for Transformer via Token Expansion |
31 Mar 2024 |
osilly/tokenexpansion/ToE/EfficientTrain/datasets.py cd3815bc733c3af8 |
ran
|
MIT (permissive) |
| Benchmarking the Robustness of Temporal Action Detection Models Against Temporal Corruptions |
29 Mar 2024 |
Alvin-Zeng/temporal-robustness-benchmark/extract_corrupted_feature_code/videomae_v2/dataset/build.py 283de3b7dcb770a3 |
unverified |
no licence file found · pointer only |
| Residual-based Language Models are Free Boosters for Biomedical Imaging |
26 Mar 2024 |
zhixinlai/llmboostmedical/2D_classification/datasets.py 8b43480472330eaf |
unverified |
MIT (permissive) |
| QKFormer: Hierarchical Spiking Transformer using Q-K Attention |
25 Mar 2024 |
Fancyssc/Spiking-Transformers/datasets.py 99c6a11c5039db76 |
unverified |
MIT (permissive) |
| Confidence Self-Calibration for Multi-Label Class-Incremental Learning |
19 Mar 2024 |
Kaile-Du/CSC/CSC/src/helper_functions/IncrementalDataset.py 0db6441e7882770b |
unverified |
no licence file found · pointer only |
| PYRA: Parallel Yielding Re-Activation for Training-Inference Efficient Task Adaptation |
14 Mar 2024 |
thu-mig/pyra/lib/datasets.py 4d972cb75e0ce17d |
ran
|
MIT (permissive) |
| An Empirical Study of LLM-as-a-Judge for LLM Evaluation: Fine-tuned Judge Model is not a General Substitute for GPT-4 |
5 Mar 2024 |
huihuichyan/unlimitedjudge/src/build_dataset.py 08ef47c526ea281d |
ran
|
no licence file found · pointer only |
| The Hidden Attention of Mamba Models |
3 Mar 2024 |
ameenali/hiddenmambaattn/vim/datasets.py d58967bba23da4e6 |
unverified |
no licence file found · pointer only |
| Spike-driven Transformer V2: Meta Spiking Neural Network Architecture Inspiring the Design of Next-generation Neuromorphic Chips |
15 Feb 2024 |
biclab/spike-driven-transformer-v2/classification/util/datasets.py 27f8c4c706fd43ec |
ran
|
no licence file found · pointer only |
| Key Patch Proposer: Key Patches Contain Rich Information |
18 Feb 2024 |
CA-TT-AC/key-patch-proposer/util/datasets.py 27f8c4c706fd43ec |
ran
|
no licence file found · pointer only |
| FViT: A Focal Vision Transformer with Gabor Filter |
17 Feb 2024 |
nkusyl/fvit/datasets.py ed28cb59239cd994 |
unverified |
MIT (permissive) |
| LF-ViT: Reducing Spatial Redundancy in Vision Transformer for Efficient Image Recognition |
8 Jan 2024 |
edgeai1/lf-vit/deit/datasets.py 72c99ead58ffc65e |
unverified |
no licence file found · pointer only |
| Context-Guided Spatio-Temporal Video Grounding |
3 Jan 2024 |
henglan/cgstvg/datasets/build.py ea91f100bf0515a9 |
unverified |
no licence file found · pointer only |
| From Static to Dynamic: Adapting Landmark-Aware Image Models for Facial Expression Recognition in Videos |
9 Dec 2023 |
FER-LMC/S2D/datasets/datasets.py f00a047305ea9335 |
unverified |
Apache-2.0 (permissive) |
| Mamba: Linear-Time Sequence Modeling with Selective State Spaces |
1 Dec 2023 |
thearkaprava/ms-temba/vim/datasets.py d58967bba23da4e6 |
unverified |
Apache-2.0 (permissive) |
| CAST: Cross-Attention in Space and Time for Video Action Recognition |
30 Nov 2023 |
KHU-VLL/CAST/dataset/datasets.py bc71b534fe126be4 |
unverified |
licence not identified · pointer only |
| A Coefficient Makes SVRG Effective |
9 Nov 2023 |
davidyyd/alpha-SVRG/datasets.py 911ca22b0cc8ef70 |
unverified |
no licence file found · pointer only |
| Learning with Unmasked Tokens Drives Stronger Vision Learners |
20 Oct 2023 |
naver-ai/lut/util/datasets.py 27f8c4c706fd43ec |
ran
|
no licence file found · pointer only |
| Frozen Transformers in Language Models Are Effective Visual Encoder Layers |
19 Oct 2023 |
ziqipang/lm4visualencoding/image_classification/datasets.py 8b43480472330eaf |
unverified |
MIT (permissive) |
| Personalized Soups: Personalized Large Language Model Alignment via Post-hoc Parameter Merging |
17 Oct 2023 |
joeljang/rlphf/training/pmorl.py fc2e047ee0f8e902 |
unverified |
no licence file found · pointer only |
| Towards Robust Multi-Modal Reasoning via Model Selection |
12 Oct 2023 |
LINs-lab/M3/MS-GQA/code/run_metagl.py d8d9625bb4bcbddf |
unverified |
Apache-2.0 (permissive) |
| Conformal Prediction for Deep Classifier via Label Ranking |
10 Oct 2023 |
ml-stat-Sustech/conformal_prediction_via_label_ranking/datasets/utils.py ffb029c3ba50dd29 |
ran
|
no licence file found · pointer only |
| PPT: Token Pruning and Pooling for Efficient Vision Transformers |
3 Oct 2023 |
xjwu1024/PPT/datasets.py d58967bba23da4e6 |
unverified |
Apache-2.0 (permissive) |
| Streaming Motion Forecasting for Autonomous Driving |
2 Oct 2023 |
ziqipang/streamingforecasting/streaming_forecasting/forecaster/builder.py ff108e9c703a507f |
unverified |
MIT (permissive) |
| Convolutional Networks with Oriented 1D Kernels |
27 Sep 2023 |
princeton-vl/oriented1d/datasets.py fdbba1a8d4185180 |
unverified |
MIT (permissive) |
| DropPos: Pre-Training Vision Transformers by Reconstructing Dropped Positions |
7 Sep 2023 |
Haochen-Wang409/DropPos/util/datasets.py 01d32b851b217f4d |
ran
|
Apache-2.0 (permissive) |
| Prototype-based Dataset Comparison |
5 Sep 2023 |
nanne/protosim/protosim_utils.py 56795f695e6eb702 |
ran
|
no licence file found · pointer only |
| NLLB-CLIP -- train performant multilingual image retrieval model on a budget |
4 Sep 2023 |
eify/clip_benchmark/clip_benchmark/datasets/builder.py 63b80a251f3277e4 |
unverified |
MIT (permissive) |
| Exploring Format Consistency for Instruction Tuning |
28 Jul 2023 |
thunlp/unifiedinstructiontuning/model_center/dataset/distributed_dataset.py 2610bd6913783d7c |
ran
|
no licence file found · pointer only |
| MOCA: Self-supervised Representation Learning by Predicting Masked Online Codebook Assignments |
18 Jul 2023 |
valeoai/moca/eval_finetune.py 9d24fe78509853ec |
ran
|
licence not identified · pointer only |
| MobileViG: Graph-Based Sparse Attention for Mobile Vision Applications |
1 Jul 2023 |
sldgroup/mobilevig/util/datasets.py e07e90c2695fc1f1 |
unverified |
Apache-2.0 (permissive) |
| Document Understanding Dataset and Evaluation (DUDE) |
15 May 2023 |
rubenpt91/MP-DocVQA-Framework/build_utils.py 88a281f34cff4e96 |
unverified |
MIT (permissive) |
| MultiModal-GPT: A Vision and Language Model for Dialogue with Humans |
8 May 2023 |
open-mmlab/multimodal-gpt/mmgpt/datasets/builder.py 298782a0d6335bbc |
unverified |
Apache-2.0 (permissive) |
| LMPT: Prompt Tuning with Class-Specific Embedding Loss for Long-tailed Multi-Label Visual Recognition |
8 May 2023 |
richard-peng-xia/LMPT/lmpt/datasets.py af60b4a83d7a2c24 |
unverified |
Apache-2.0 (permissive) |
| PreNAS: Preferred One-Shot Learning Towards Efficient Neural Architecture Search |
2023-04 (from id) |
tinyvision/PreNAS/lib/datasets.py afa7efdfbea334df |
unverified |
Apache-2.0 recorded; this copy not marked cleared · pointer only |
| Joint Token Pruning and Squeezing Towards More Aggressive Compression of Vision Transformers |
21 Apr 2023 |
megvii-research/TPS-CVPR2023/torch_codebase/datasets.py 0d5eb6177974e3b2 |
unverified |
Apache-2.0 (permissive) |
| SegGPT: Segmenting Everything In Context |
6 Apr 2023 |
baaivision/painter/Painter/util/datasets.py 27f8c4c706fd43ec |
ran
|
MIT (permissive) |
| SMPConv: Self-moving Point Representations for Continuous Convolution |
5 Apr 2023 |
sangnekim/SMPConv/smp_imagenet/datasets.py fdbba1a8d4185180 |
unverified |
MIT (permissive) |
| AdPE: Adversarial Positional Embeddings for Pretraining Vision Transformers via MAE+ |
14 Mar 2023 |
maple-research-lab/AdPE/data_processing/build_dataset.py 27f8c4c706fd43ec |
ran
|
MIT (permissive) |
| Open-Vocabulary Affordance Detection in 3D Point Clouds |
4 Mar 2023 |
Fsoft-AIC/Open-Vocabulary-Affordance-Detection-in-3D-Point-Clouds/utils/builder.py f4261a73cded9e8b |
unverified |
MIT (permissive) |
| Chemically Transferable Generative Backmapping of Coarse-Grained Proteins |
2 Mar 2023 |
learningmatter-mit/genzprot/scripts/inference.py 0408760c5846262f |
unverified |
no licence file found · pointer only |
| MedViT: A Robust Vision Transformer for Generalized Medical Image Classification |
19 Feb 2023 |
Omid-Nejati/MedViT/CustomDataset/datasets.py 22ed072362b1f173 |
unverified |
MIT (permissive) |
| Stitchable Neural Networks |
13 Feb 2023 |
ziplab/SN-Net/stitching_deit/datasets.py 9153241c4331eb1a |
unverified |
Apache-2.0 (permissive) |
| CEDNet: A Cascade Encoder-Decoder Network for Dense Prediction |
13 Feb 2023 |
zhanggang001/cfnet/datasets.py fdbba1a8d4185180 |
unverified |
MIT recorded; this copy not marked cleared · pointer only |
| Dynamic Grained Encoder for Vision Transformers |
10 Jan 2023 |
stevengrove/vtpack/vtpack/engine/datasets.py 9153241c4331eb1a |
unverified |
Apache-2.0 recorded; this copy not marked cleared · pointer only |
| Masked autoencoders are effective solution to transformer data-hungry |
12 Dec 2022 |
talented-q/sdmae/util/datasets.py 4eb159b13199f984 |
unverified |
MIT (permissive) |
| Deep Incubation: Training Large Models by Divide-and-Conquering |
8 Dec 2022 |
leaplabthu/model-assembling/datasets.py 9153241c4331eb1a |
unverified |
MIT (permissive) |
| Towards Good Practices for Missing Modality Robust Action Recognition |
25 Nov 2022 |
sangminwoo/actionmae/lib/dataset/dataset_builder.py 5d298d5f727c6422 |
unverified |
MIT (permissive) |
| ViTALiTy: Unifying Low-rank and Sparse Approximation for Vision Transformer Acceleration with a Linear Taylor Attention |
9 Nov 2022 |
GATECH-EIC/ViTaLiTy/src/datasets.py 9153241c4331eb1a |
unverified |
Apache-2.0 recorded; this copy not marked cleared · pointer only |
| Towards a Unified View on Visual Parameter-Efficient Transfer Learning |
3 Oct 2022 |
bruceyo/V-PETL/util/datasets.py 27f8c4c706fd43ec |
ran
|
MIT (permissive) |
| Exploring Target Representations for Masked Autoencoders |
8 Sep 2022 |
liuxingbin/dbot/evaluation/eval_cls.py ac882d65ee949545 |
unverified |
Apache-2.0 (permissive) |
| Deep Coarse-grained Potentials via Relative Entropy Minimization |
2022-08 (from id) |
tummfm/relative-entropy/chemtrain/force_matching.py b8c0cc579811e6f6 |
unverified |
Apache-2.0 (permissive) |
| ScaleNet: Searching for the Model to Scale |
15 Jul 2022 |
luminolx/ScaleNet/core/dataset/build_dataloader.py 9e325fc12ea59580 |
unverified |
Apache-2.0 (permissive) |
| Next-ViT: Next Generation Vision Transformer for Efficient Deployment in Realistic Industrial Scenarios |
12 Jul 2022 |
bytedance/next-vit/classification/datasets.py 2317eda6ad7d0b3b |
unverified |
Apache-2.0 (permissive) |
| SimA: Simple Softmax-free Attention for Vision Transformers |
17 Jun 2022 |
ucdvision/sima/datasets.py 18119fb5f5299851 |
unverified |
MIT (permissive) |
| Neural Prompt Search |
9 Jun 2022 |
ZhangYuanhan-AI/NOAH/lib/datasets.py 3d5d962b181c73fb |
unverified |
MIT (permissive) |
| Masked Unsupervised Self-training for Label-free Image Classification |
7 Jun 2022 |
salesforce/must/build_dataset.py 5fc280732189fd24 |
unverified |
BSD-3-Clause (permissive) |
| ConvMAE: Masked Convolution Meets Masked Autoencoders |
8 May 2022 |
mx-mark/dmjd/util/datasets.py 27f8c4c706fd43ec |
ran
|
MIT (permissive) |
| Mugs: A Multi-Granular Self-Supervised Learning Framework |
27 Mar 2022 |
sail-sg/mugs/eval/eval_finetuning/eval_finetuning.py 961c92c8136b0f58 |
unverified |
Apache-2.0 (permissive) |
| Smoothing Matters: Momentum Transformer for Domain Adaptive Semantic Segmentation |
15 Mar 2022 |
alpc91/transda/core/datasets/build.py 2f921ad8e71df973 |
unverified |
MIT (permissive) |
| The Principle of Diversity: Training Stronger Vision Transformers Calls for Reducing All Levels of Redundancy |
12 Mar 2022 |
VITA-Group/Diverse-ViT/datasets.py 49c67c60672b93ae |
unverified |
MIT (permissive) |
| Anti-Oversmoothing in Deep Vision Transformers via the Fourier Domain Analysis: From Theory to Practice |
9 Mar 2022 |
VITA-Group/ViT-Anti-Oversmoothing/datasets.py ccfff191e1bbe653 |
unverified |
MIT (permissive) |