Methods › General › Feedforward Networks › Position-Wise Feed-Forward Layer
Position-Wise Feed-Forward Layer
archive 2025-07-28 Description, source and code snippet are the archive's method entry.
Position-Wise Feed-Forward Layer is a type of feedforward layer consisting of two dense layers that applies to the last dimension, which means the same dense layers are used for each position item in the sequence, so called position-wise.
Papers archive 2025-07-28
30 shown of 13,895, newest first. Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code; it is per sample and not a correctness claim.
-
Towards Robust Multimodal Emotion Recognition under Missing Modalities and Distribution Shifts 12 Jun 2025 · 1 repository · arXiv:2506.10452
-
SparseSSM: Efficient Selective Structured State Space Models Can Be Pruned in One-Shot 11 Jun 2025 · 0 repositories · arXiv:2506.09613
-
Hierarchical Neural Collapse Detection Transformer for Class Incremental Object Detection 10 Jun 2025 · 0 repositories · arXiv:2506.08562
-
Hyperspectral Image Classification via Transformer-based Spectral-Spatial Attention Decoupling and Adaptive Gating 10 Jun 2025 · 1 repository · arXiv:2506.08324
-
MedMoE: Modality-Specialized Mixture of Experts for Medical Vision-Language Understanding 10 Jun 2025 · 0 repositories · arXiv:2506.08356
-
PatchGuard: Adversarially Robust Anomaly Detection and Localization through Vision Transformers and Pseudo Anomalies 10 Jun 2025 · 2 repositories · arXiv:2506.09237
-
Robust Visual Localization via Semantic-Guided Multi-Scale Transformer 10 Jun 2025 · 0 repositories · arXiv:2506.08526
-
TACTIC: Translation Agents with Cognitive-Theoretic Interactive Collaboration 10 Jun 2025 · 1 repository · arXiv:2506.08403
-
4DGT: Learning a 4D Gaussian Transformer Using Real-World Monocular Videos 9 Jun 2025 · 0 repositories · arXiv:2506.08015
-
Lightweight Sequential Transformers for Blood Glucose Level Prediction in Type-1 Diabetes 9 Jun 2025 · 0 repositories · arXiv:2506.07864
-
MADFormer: Mixed Autoregressive and Diffusion Transformers for Continuous Image Generation 9 Jun 2025 · 0 repositories · arXiv:2506.07999
-
Quantum Graph Transformer for NLP Sentiment Classification 9 Jun 2025 · 0 repositories · arXiv:2506.07937
-
Breaking Data Silos: Towards Open and Scalable Mobility Foundation Models via Generative Continual Learning 7 Jun 2025 · 0 repositories · arXiv:2506.06694
-
Can In-Context Reinforcement Learning Recover From Reward Poisoning Attacks? 7 Jun 2025 · 0 repositories · arXiv:2506.06891
-
BEAST: Efficient Tokenization of B-Splines Encoded Action Sequences for Imitation Learning 6 Jun 2025 · 0 repositories · arXiv:2506.06072
-
Intentionally Unintentional: GenAI Exceptionalism and the First Amendment 5 Jun 2025 · 0 repositories · arXiv:2506.05211
-
Enhancing Automatic PT Tagging for MEDLINE Citations Using Transformer-Based Models 3 Jun 2025 · 0 repositories · arXiv:2506.03321
-
Rhythm Controllable and Efficient Zero-Shot Voice Conversion via Shortcut Flow Matching 1 Jun 2025 · 0 repositories · arXiv:2506.01014
-
Channel-Imposed Fusion: A Simple yet Effective Method for Medical Time Series Classification 31 May 2025 · 0 repositories · arXiv:2506.00337
-
Evaluating Robot Policies in a World Model 31 May 2025 · 0 repositories · arXiv:2506.00613
-
Machine vs Machine: Using AI to Tackle Generative AI Threats in Assessment 31 May 2025 · 0 repositories · arXiv:2506.02046
-
Translate With Care: Addressing Gender Bias, Neutrality, and Reasoning in Large Language Model Translations 31 May 2025 · 1 repository · arXiv:2506.00748
-
Cross-Attention Speculative Decoding 30 May 2025 · 0 repositories · arXiv:2505.24544
-
D2AF: A Dual-Driven Annotation and Filtering Framework for Visual Grounding 30 May 2025 · 0 repositories · arXiv:2505.24372
-
Leveraging Intermediate Features of Vision Transformer for Face Anti-Spoofing 30 May 2025 · 0 repositories · arXiv:2505.24402
-
Mamba Knockout for Unraveling Factual Information Flow 30 May 2025 · 1 repository · arXiv:2505.24244Syntology ran 2 of 6 samples · 4 unverified
-
Mastering Massive Multi-Task Reinforcement Learning via Mixture-of-Expert Decision Transformer 30 May 2025 · 1 repository · arXiv:2505.24378
-
PCIE_Pose Solution for EgoExo4D Pose and Proficiency Estimation Challenge 30 May 2025 · 0 repositories · arXiv:2505.24411
-
PersianMedQA: Language-Centric Evaluation of LLMs in the Persian Medical Domain 30 May 2025 · 0 repositories · arXiv:2506.00250
-
SPPSFormer: High-quality Superpoint-based Transformer for Roof Plane Instance Segmentation from Point Clouds 30 May 2025 · 0 repositories · arXiv:2505.24475
Tasks archive 2025-07-28
20 shown of 2,147 tasks the archive attaches to papers tagged with this method, by distinct papers. A task without a page in the catalog is plain text.
| Task | Papers |
|---|---|
| Language Modelling | 1,235 |
| Decoder | 1,067 |
| Language Modeling | 946 |
| Translation | 757 |
| Machine Translation | 717 |
| Semantic Segmentation | 685 |
| Object Detection | 548 |
| Image Classification | 516 |
| Question Answering | 510 |
| object-detection | 494 |
| Segmentation | 464 |
| Retrieval | 453 |
| Sentence | 452 |
| Representation Learning | 426 |
| image-classification | 410 |
| Large Language Model | 406 |
| Time Series | 368 |
| Object | 348 |
| Text Generation | 315 |
| Transfer Learning | 297 |
Usage over time archive 2025-07-28
Components: the archive holds no method-to-method composition, so PwC's Components table cannot be rebuilt; the Papers list carries no Results column for the same reason (the archive does not join its leaderboard rows to method tags).
Categories archive 2025-07-28
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections