Methods › Natural Language Processing › Autoregressive Transformers › Linformer
Linformer
Introduced by Sinong Wang et al. in Linformer: Self-Attention with Linear Complexity
archive 2025-07-28 Description, source and code snippet are the archive's method entry.
Linformer is a linear Transformer that utilises a linear self-attention mechanism to tackle the self-attention bottleneck with Transformer models. The original scaled dot-product attention is decomposed into multiple smaller attentions through linear projections, such that the combination of these operations forms a low-rank factorization of the original attention.
Papers archive 2025-07-28
17 shown of 17, newest first. Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code; it is per sample and not a correctness claim.
-
CacheFormer: High Attention-Based Segment Caching 18 Apr 2025 · 0 repositories · arXiv:2504.13981
-
LinFormer: A Linear-based Lightweight Transformer Architecture For Time-Aware MIMO Channel Prediction 28 Oct 2024 · 0 repositories · arXiv:2410.21351
-
Joint Fine-tuning and Conversion of Pretrained Speech and Language Models towards Linear Complexity 9 Oct 2024 · 1 repository · arXiv:2410.06846Syntology ran 5 of 5 samples · 0 unverified
-
GLMHA A Guided Low-rank Multi-Head Self-Attention for Efficient Image Restoration and Spectral Reconstruction 1 Oct 2024 · 0 repositories · arXiv:2410.00380
-
Sumformer: Universal Approximation for Efficient Transformers 5 Jul 2023 · 0 repositories · arXiv:2307.02301
-
MPCViT: Searching for Accurate and Efficient MPC-Friendly Vision Transformer with Heterogeneous Attention 25 Nov 2022 · 1 repository · arXiv:2211.13955
-
Treeformer: Dense Gradient Trees for Efficient Attention Computation 18 Aug 2022 · 0 repositories · arXiv:2208.09015
-
Linearizing Transformer with Key-Value Memory 23 Mar 2022 · 0 repositories · arXiv:2203.12644
-
Sketching as a Tool for Understanding and Accelerating Self-attention for Long Sequences 10 Dec 2021 · 1 repository · arXiv:2112.05359
-
Greenformers: Improving Computation and Memory Efficiency in Transformer Models via Low-Rank Approximation 24 Aug 2021 · 0 repositories · arXiv:2108.10808
-
Vision Xformers: Efficient Attention for Image Classification 5 Jul 2021 · 2 repositories · arXiv:2107.02239
-
Styleformer: Transformer based Generative Adversarial Networks with Style Vector 13 Jun 2021 · 3 repositories · arXiv:2106.07023Syntology ran 4 of 4 samples · 0 unverified
-
Self-supervised Depth Estimation Leveraging Global Perception and Geometric Smoothness Using On-board Videos 7 Jun 2021 · 0 repositories · arXiv:2106.03505
-
A Practical Survey on Faster and Lighter Transformers 26 Mar 2021 · 0 repositories · arXiv:2103.14636
-
Revisiting Linformer with a modified self-attention with linear complexity 16 Dec 2020 · 0 repositories · arXiv:2101.10277
-
Efficient Transformers: A Survey 14 Sep 2020 · 0 repositories · arXiv:2009.06732
-
Linformer: Self-Attention with Linear Complexity 8 Jun 2020 · 3 repositories · arXiv:2006.04768Syntology ran 3 of 4 samples · 1 unverified · 2 pointer-only (licence)
Tasks archive 2025-07-28
20 shown of 30 tasks the archive attaches to papers tagged with this method, by distinct papers. A task without a page in the catalog is plain text.
Usage over time archive 2025-07-28
Components: the archive holds no method-to-method composition, so PwC's Components table cannot be rebuilt; the Papers list carries no Results column for the same reason (the archive does not join its leaderboard rows to method tags).
Categories archive 2025-07-28
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections