Methods › Computer Vision › Image Models › DeiT
Data-efficient Image Transformer
DeiT
archive 2025-07-28 Description, source and code snippet are the archive's method entry.
A Data-Efficient Image Transformer is a type of Vision Transformer for image classification tasks. The model is trained using a teacher-student strategy specific to transformers. It relies on a distillation token ensuring that the student learns from the teacher through attention.
Papers archive 2025-07-28
30 shown of 93, newest first. Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code; it is per sample and not a correctness claim.
-
DART: Differentiable Dynamic Adaptive Region Tokenizer for Vision Transformer and Mamba 12 Jun 2025 · 1 repository · arXiv:2506.10390
-
Is Attention Required for Transformer Inference? Explore Function-preserving Attention Replacement 24 May 2025 · 0 repositories · arXiv:2505.21535
-
Deep learning-based identification of precipitation clouds from all-sky camera data for observatory safety 24 Mar 2025 · 0 repositories · arXiv:2503.18670
-
CAE-Net: Generalized Deepfake Image Detection using Convolution and Attention Mechanisms with Spatial and Frequency Domain Features 15 Feb 2025 · 0 repositories · arXiv:2502.10682
-
Mix-QViT: Mixed-Precision Vision Transformer Quantization Driven by Layer Importance and Quantization Sensitivity 10 Jan 2025 · 0 repositories · arXiv:2501.06357
-
Adventurer: Optimizing Vision Mamba Architecture Designs for Efficiency 1 Jan 2025 · 0 repositories
-
Training Noise Token Pruning 27 Nov 2024 · 1 repository · arXiv:2411.18092
-
Scalable iterative pruning of large language and vision models using block coordinate descent 26 Nov 2024 · 0 repositories · arXiv:2411.17796
-
PQV-Mobile: A Combined Pruning and Quantization Toolkit to Optimize Vision Transformers for Mobile Applications 15 Aug 2024 · 1 repository · arXiv:2408.08437
-
Mixed Non-linear Quantization for Vision Transformers 26 Jul 2024 · 1 repository · arXiv:2407.18437
-
Knowledge distillation to effectively attain both region-of-interest and global semantics from an image where multiple objects appear 11 Jul 2024 · 1 repository · arXiv:2407.08257
-
Leveraging Transformers for Weakly Supervised Object Localization in Unconstrained Videos 8 Jul 2024 · 1 repository · arXiv:2407.06018
-
Predicting Depression and Anxiety Risk in Dutch Neighborhoods from Street-View Images 27 Jun 2024 · 0 repositories · arXiv:2407.09547
-
LieRE: Generalizing Rotary Position Encodings 14 Jun 2024 · 1 repository · arXiv:2406.10322Syntology ran 1 of 3 samples · 2 unverified
-
Decomposing and Interpreting Image Representations via Text in ViTs Beyond CLIP 3 Jun 2024 · 1 repository · arXiv:2406.01583Syntology ran 10 of 13 samples · 3 unverified · 13 pointer-only (licence)
-
A General and Efficient Training for Transformer via Token Expansion 31 Mar 2024 · 1 repository · arXiv:2404.00672Syntology ran 7 of 8 samples · 1 unverified
-
Multi-criteria Token Fusion with One-step-ahead Attention for Efficient Vision Transformers 15 Mar 2024 · 1 repository · arXiv:2403.10030Syntology ran 3 of 5 samples · 2 unverified
-
MPTQ-ViT: Mixed-Precision Post-Training Quantization for Vision Transformer 26 Jan 2024 · 0 repositories · arXiv:2401.14895
-
Detecting and recognizing characters in Greek papyri with YOLOv8, DeiT and SimCLR 23 Jan 2024 · 0 repositories · arXiv:2401.12513
-
LRP-QViT: Mixed-Precision Vision Transformer Quantization via Layer-wise Relevance Propagation 20 Jan 2024 · 0 repositories · arXiv:2401.11243
-
Vision Mamba: Efficient Visual Representation Learning with Bidirectional State Space Model 17 Jan 2024 · 15 repositories · arXiv:2401.09417Syntology ran 2 of 13 samples · 11 unverified · 1 pointer-only (licence)
-
TPC-ViT: Token Propagation Controller for Efficient Vision Transformer 3 Jan 2024 · 0 repositories · arXiv:2401.01470
-
ROI-Aware Multiscale Cross-Attention Vision Transformer for Pest Image Identification 28 Dec 2023 · 0 repositories · arXiv:2312.16914
-
Bridging The Gaps Between Token Pruning and Full Pre-training via Masked Fine-tuning 26 Oct 2023 · 0 repositories · arXiv:2310.17177
-
USDC: Unified Static and Dynamic Compression for Visual Transformer 17 Oct 2023 · 0 repositories · arXiv:2310.11117
-
Accelerating Vision Transformers Based on Heterogeneous Attention Patterns 11 Oct 2023 · 0 repositories · arXiv:2310.07664
-
EWasteNet: A Two-Stream Data Efficient Image Transformer Approach for E-Waste Classification 28 Sep 2023 · 0 repositories · arXiv:2311.12823
-
Multi-Dimensional Hyena for Spatial Inductive Bias 24 Sep 2023 · 0 repositories · arXiv:2309.13600
-
Weight Averaging Improves Knowledge Distillation under Domain Shift 20 Sep 2023 · 1 repository · arXiv:2309.11446
-
Tuning Pre-trained Model via Moment Probing 21 Jul 2023 · 1 repository · arXiv:2307.11342Syntology ran 8 of 13 samples · 5 unverified · 13 pointer-only (licence)
Tasks archive 2025-07-28
20 shown of 78 tasks the archive attaches to papers tagged with this method, by distinct papers. A task without a page in the catalog is plain text.
Usage over time archive 2025-07-28
Components: the archive holds no method-to-method composition, so PwC's Components table cannot be rebuilt; the Papers list carries no Results column for the same reason (the archive does not join its leaderboard rows to method tags).
Categories archive 2025-07-28
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections