Methods › General › Feedforward Networks › Position-Wise Feed-Forward Layer › Papers, page 107
Position-Wise Feed-Forward Layer
Papers archive 2025-07-28
archive papers tagged: 13,895 · with a code link: 6,514 · where Syntology ran a sample: 2,229 (1,902 with a run with no instrument failure, 327 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,229 of 13,895 tagged: 1,902 with a run with no instrument failure, 327 where every run was a failure of Syntology's instrument)
Page 107 of 139: papers 10,601 to 10,700 of 13,895, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Voxel Set Transformer: A Set-to-Set Approach to 3D Object Detection from Point Clouds 19 Mar 2022 · 1 repository · arXiv:2203.10314Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples)
-
AlignTransformer: Hierarchical Alignment of Visual Regions and Disease Tags for Medical Report Generation 18 Mar 2022 · 0 repositories · arXiv:2203.10095
-
Local-Global Context Aware Transformer for Language-Guided Video Segmentation 18 Mar 2022 · 1 repository · arXiv:2203.09773Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
M2TS: Multi-Scale Multi-Modal Approach Based on Transformer for Source Code Summarization 18 Mar 2022 · 1 repository · arXiv:2203.09707
-
Cascade Transformers for End-to-End Person Search 17 Mar 2022 · 1 repository · arXiv:2203.09642
-
Fine- and Coarse-Granularity Hybrid Self-Attention for Efficient BERT 17 Mar 2022 · 1 repository · arXiv:2203.09055
-
HiStruct+: Improving Extractive Text Summarization with Hierarchical Structure Information 17 Mar 2022 · 0 repositories · arXiv:2203.09629
-
Look Outside the Room: Synthesizing A Consistent Long-Term 3D Scene Video from A Single Image 17 Mar 2022 · 1 repository · arXiv:2203.09457Syntology 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
ODE Transformer: An Ordinary Differential Equation-Inspired Model for Sequence Generation 17 Mar 2022 · 1 repository · arXiv:2203.09176Syntology official (archive's flag): 3 ran · 3 ran (of which 3 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified; every one of the 3 samples that ran constructed an object rather than computing a result (of 3 harvested samples) · 3 pointer-only (licence)
-
On Vision Features in Multimodal Machine Translation 17 Mar 2022 · 2 repositories · arXiv:2203.09173Syntology official (archive's flag): 2 ran · 2 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
PLANET: Dynamic Content Planning in Autoregressive Transformers for Long-form Text Generation 17 Mar 2022 · 0 repositories · arXiv:2203.09100
-
PreTR: Spatio-Temporal Non-Autoregressive Trajectory Prediction Transformer 17 Mar 2022 · 0 repositories · arXiv:2203.09293
-
Semantic-aligned Fusion Transformer for One-shot Object Detection 17 Mar 2022 · 0 repositories · arXiv:2203.09093
-
SepTr: Separable Transformer for Audio Spectrogram Processing 17 Mar 2022 · 1 repository · arXiv:2203.09581
-
Transframer: Arbitrary Frame Prediction with Generative Models 17 Mar 2022 · 0 repositories · arXiv:2203.09494
-
UNIMO-2: End-to-End Unified Vision-Language Grounded Learning 17 Mar 2022 · 1 repository · arXiv:2203.09067
-
A Squeeze-and-Excitation and Transformer based Cross-task System for Environmental Sound Recognition 16 Mar 2022 · 0 repositories · arXiv:2203.08350
-
DeciWatch: A Simple Baseline for 10x Efficient 2D and 3D Pose Estimation 16 Mar 2022 · 1 repository · arXiv:2203.08713Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 8 harvested samples) · 1 pointer-only (licence)
-
The Devil Is in the Details: Window-based Attention for Image Compression 16 Mar 2022 · 2 repositories · arXiv:2203.08450Syntology official (archive's flag): 15 ran · 15 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 3 honoured, 0 violated, 8 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 15 harvested samples) · 11 pointer-only (licence)
-
Towards Practical Certifiable Patch Defense with Vision Transformer 16 Mar 2022 · 0 repositories · arXiv:2203.08519
-
WegFormer: Transformers for Weakly Supervised Semantic Segmentation 16 Mar 2022 · 0 repositories · arXiv:2203.08421
-
ActFormer: A GAN-based Transformer towards General Action-Conditioned 3D Human Motion Generation 15 Mar 2022 · 0 repositories · arXiv:2203.07706
-
Compressing Sentence Representation for Semantic Retrieval via Homomorphic Projective Distillation 15 Mar 2022 · 1 repository · arXiv:2203.07687
-
Efficient Long Sequence Encoding via Synchronization 15 Mar 2022 · 0 repositories · arXiv:2203.07644
-
Fast Autofocusing using Tiny Transformer Networks for Digital Holographic Microscopy 15 Mar 2022 · 1 repository · arXiv:2203.07772
-
Generating Privacy-Preserving Process Data with Deep Generative Models 15 Mar 2022 · 0 repositories · arXiv:2203.07949
-
HUMUS-Net: Hybrid unrolled multi-scale network architecture for accelerated MRI reconstruction 15 Mar 2022 · 2 repositories · arXiv:2203.08213Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 3 pointer-only (licence)
-
InvPT: Inverted Pyramid Multi-task Transformer for Dense Scene Understanding 15 Mar 2022 · 1 repository · arXiv:2203.07997Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples) · 1 pointer-only (licence)
-
Smoothing Matters: Momentum Transformer for Domain Adaptive Semantic Segmentation 15 Mar 2022 · 1 repository · arXiv:2203.07988Syntology official (archive's flag): 1 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 5 unverified (of 8 harvested samples) · 2 pointer-only (licence)
-
Unified Visual Transformer Compression 15 Mar 2022 · 1 repository · arXiv:2203.08243Syntology official (archive's flag): 7 ran · 9 ran (of which 3 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 1 violated, 5 with no contract checked; 3 where Syntology's instrument failed) · 7 unverified (of 16 harvested samples) · 3 pointer-only (licence)
-
Accelerating DETR Convergence via Semantic-Aligned Matching 14 Mar 2022 · 1 repository · arXiv:2203.06883Syntology official (archive's flag): 3 ran · 3 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples)
-
All in One: Exploring Unified Video-Language Pre-training 14 Mar 2022 · 1 repository · arXiv:2203.07303
-
Efficient Language Modeling with Sparse all-MLP 14 Mar 2022 · 0 repositories · arXiv:2203.06850
-
Deep Transformers Thirst for Comprehensive-Frequency Data 14 Mar 2022 · 1 repository · arXiv:2203.07116
-
Switch Trajectory Transformer with Distributional Value Approximation for Multi-Task Reinforcement Learning 14 Mar 2022 · 0 repositories · arXiv:2203.07413
-
CMKD: CNN/Transformer-Based Cross-Model Knowledge Distillation for Audio Classification 13 Mar 2022 · 2 repositories · arXiv:2203.06760
-
Masked Autoencoders for Point Cloud Self-supervised Learning 13 Mar 2022 · 4 repositories · arXiv:2203.06604Syntology official (archive's flag): 4 ran · 9 ran (of which 1 constructed an object rather than computing a result; 6 with no instrument failure: 2 honoured, 0 violated, 4 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples) · 7 pointer-only (licence)
-
SATr: Slice Attention with Transformer for Universal Lesion Detection 13 Mar 2022 · 0 repositories · arXiv:2203.07373
-
Scaling Up Your Kernels to 31x31: Revisiting Large Kernel Design in CNNs 13 Mar 2022 · 8 repositories · arXiv:2203.06717Syntology official (archive's flag): 4 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples) · 1 pointer-only (licence)
-
Sparse Local Patch Transformer for Robust Face Alignment and Landmarks Inherent Relation Learning 13 Mar 2022 · 1 repository · arXiv:2203.06541
-
DFTR: Depth-supervised Fusion Transformer for Salient Object Detection 12 Mar 2022 · 0 repositories · arXiv:2203.06429
-
Joint CNN and Transformer Network via weakly supervised Learning for efficient crowd counting 12 Mar 2022 · 0 repositories · arXiv:2203.06388
-
One-stage Video Instance Segmentation: From Frame-in Frame-out to Clip-in Clip-out 12 Mar 2022 · 0 repositories · arXiv:2203.06421
-
Wasserstein Adversarial Transformer for Cloud Workload Prediction 12 Mar 2022 · 1 repository · arXiv:2203.06501
-
Block-Recurrent Transformers 11 Mar 2022 · 3 repositories · arXiv:2203.07852Syntology official: no sample here; runs from other or unrecorded repositories · 18 ran (of which 2 constructed an object rather than computing a result; 9 with no instrument failure: 2 honoured, 3 violated, 4 with no contract checked; 9 where Syntology's instrument failed) · 5 unverified (of 23 harvested samples) · 4 pointer-only (licence)
-
Font Shape-to-Impression Translation 11 Mar 2022 · 0 repositories · arXiv:2203.05808
-
PathSAGE: Spatial Graph Attention Neural Networks With Random Path Sampling 11 Mar 2022 · 0 repositories · arXiv:2203.05793
-
Transformer-based Streaming ASR with Cumulative Attention 11 Mar 2022 · 0 repositories · arXiv:2203.05736
-
Visualizing and Understanding Patch Interactions in Vision Transformer 11 Mar 2022 · 0 repositories · arXiv:2203.05922
-
Look Backward and Forward: Self-Knowledge Distillation with Bidirectional Decoder for Neural Machine Translation 10 Mar 2022 · 0 repositories · arXiv:2203.05248
-
Parameter-Free Attentive Scoring for Speaker Verification 10 Mar 2022 · 1 repository · arXiv:2203.05642
-
StyleBabel: Artistic Style Tagging and Captioning 10 Mar 2022 · 0 repositories · arXiv:2203.05321
-
TrueType Transformer: Character and Font Style Recognition in Outline Format 10 Mar 2022 · 1 repository · arXiv:2203.05338
-
Anti-Oversmoothing in Deep Vision Transformers via the Fourier Domain Analysis: From Theory to Practice 9 Mar 2022 · 1 repository · arXiv:2203.05962Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 2 pointer-only (licence)
-
Coarse-to-Fine Sparse Transformer for Hyperspectral Image Reconstruction 9 Mar 2022 · 1 repository · arXiv:2203.04845Syntology official (archive's flag): 13 ran · 13 ran (of which 8 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 4 where Syntology's instrument failed) · 3 unverified (of 16 harvested samples)
-
Multiscale Convolutional Transformer with Center Mask Pretraining for Hyperspectral Image Classification 9 Mar 2022 · 0 repositories · arXiv:2203.04771
-
PHTrans: Parallelly Aggregating Global and Local Representations for Medical Image Segmentation 9 Mar 2022 · 2 repositories · arXiv:2203.04568
-
Region-Aware Face Swapping 9 Mar 2022 · 0 repositories · arXiv:2203.04564
-
The evolution, evolvability and engineering of gene regulatory DNA 9 Mar 2022 · 1 repository
-
Uni4Eye: Unified 2D and 3D Self-supervised Pre-training via Masked Image Modeling Transformer for Ophthalmic Image Classification 9 Mar 2022 · 0 repositories · arXiv:2203.04614
-
CaSS: A Channel-aware Self-supervised Representation Learning Framework for Multivariate Time Series Classification 8 Mar 2022 · 0 repositories · arXiv:2203.04298
-
DuMLP-Pin: A Dual-MLP-dot-product Permutation-invariant Network for Set Feature Extraction 8 Mar 2022 · 1 repository · arXiv:2203.04007
-
Dynamic Group Transformer: A General Vision Transformer Backbone with Dynamic Group Attention 8 Mar 2022 · 0 repositories · arXiv:2203.03937
-
Graph Attention Transformer Network for Multi-Label Image Classification 8 Mar 2022 · 1 repository · arXiv:2203.04049
-
Joint rotational invariance and adversarial training of a dual-stream Transformer yields state of the art Brain-Score for Area V4 8 Mar 2022 · 1 repository · arXiv:2203.06649
-
Lane Detection with Versatile AtrousFormer and Local Semantic Guidance 8 Mar 2022 · 0 repositories · arXiv:2203.04067
-
Measuring the Mixing of Contextual Information in the Transformer 8 Mar 2022 · 2 repositories · arXiv:2203.04212Syntology community repositories only · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Neural Face Identification in a 2D Wireframe Projection of a Manifold Object 8 Mar 2022 · 1 repository · arXiv:2203.04229Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 1 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 9 harvested samples)
-
Monocular Robot Navigation with Self-Supervised Pretrained Vision Transformers 7 Mar 2022 · 0 repositories · arXiv:2203.03682
-
SkillNet-NLU: A Sparsely Activated Model for General-Purpose Natural Language Understanding 7 Mar 2022 · 0 repositories · arXiv:2203.03312
-
Stepwise Feature Fusion: Local Guides Global 7 Mar 2022 · 1 repository · arXiv:2203.03635
-
Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer 7 Mar 2022 · 7 repositories · arXiv:2203.03466Syntology official (archive's flag): 1 ran · 3 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples)
-
Conditional Bilingual Mutual Information Based Adaptive Training for Neural Machine Translation 6 Mar 2022 · 1 repository · arXiv:2203.02951
-
Exploring Dual-task Correlation for Pose Guided Person Image Generation 6 Mar 2022 · 1 repository · arXiv:2203.02910Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; the one sample that ran constructed an object rather than computing a result (of 2 harvested samples) · 2 pointer-only (licence)
-
Learnable Irrelevant Modality Dropout for Multimodal Action Recognition on Modality-Specific Annotated Videos 6 Mar 2022 · 0 repositories · arXiv:2203.03014
-
Multi-class Token Transformer for Weakly Supervised Semantic Segmentation 6 Mar 2022 · 1 repository · arXiv:2203.02891Syntology official (archive's flag): 4 ran · 4 ran (of which 4 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; every one of the 4 samples that ran constructed an object rather than computing a result (of 5 harvested samples) · 5 pointer-only (licence)
-
PanFormer: a Transformer Based Model for Pan-sharpening 6 Mar 2022 · 1 repository · arXiv:2203.02916
-
ClarET: Pre-training a Correlation-Aware Context-To-Event Transformer for Event-Centric Generation and Classification 4 Mar 2022 · 1 repository · arXiv:2203.02225
-
DiT: Self-supervised Pre-training for Document Image Transformer 4 Mar 2022 · 4 repositories · arXiv:2203.02378Syntology community repositories only · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples)
-
LiteTransformerSearch: Training-free Neural Architecture Search for Efficient Language Models 4 Mar 2022 · 1 repository · arXiv:2203.02094Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 1 pointer-only (licence)
-
UVCGAN: UNet Vision Transformer cycle-consistent GAN for unpaired image-to-image translation 4 Mar 2022 · 2 repositories · arXiv:2203.02557
-
ETCetera: beyond Event-Triggered Control 3 Mar 2022 · 0 repositories · arXiv:2203.01623
-
LGT-Net: Indoor Panoramic Room Layout Estimation with Geometry-Aware Transformer Network 3 Mar 2022 · 1 repository · arXiv:2203.01824Syntology official (archive's flag): 17 ran · 29 ran (of which 5 constructed an object rather than computing a result; 20 with no instrument failure: 2 honoured, 1 violated, 17 with no contract checked; 9 where Syntology's instrument failed) · 12 unverified (of 41 harvested samples)
-
Multi-Tailed Vision Transformer for Efficient Inference 3 Mar 2022 · 0 repositories · arXiv:2203.01587
-
ViTransPAD: Video Transformer using convolution and self-attention for Face Presentation Attack Detection 3 Mar 2022 · 0 repositories · arXiv:2203.01562
-
3DCTN: 3D Convolution-Transformer Network for Point Cloud Classification 2 Mar 2022 · 1 repository · arXiv:2203.00828
-
Aggregated Pyramid Vision Transformer: Split-transform-merge Strategy for Image Recognition without Convolutions 2 Mar 2022 · 0 repositories · arXiv:2203.00960
-
Bending Reality: Distortion-aware Transformers for Adapting to Panoramic Semantic Segmentation 2 Mar 2022 · 1 repository · arXiv:2203.01452
-
Contextual Attention Network: Transformer Meets U-Net 2 Mar 2022 · 3 repositories · arXiv:2203.01932
-
D^2ETR: Decoder-Only DETR with Computationally Efficient Cross-Scale Attention 2 Mar 2022 · 0 repositories · arXiv:2203.00860
-
DN-DETR: Accelerate DETR Training by Introducing Query DeNoising 2 Mar 2022 · 17 repositories · arXiv:2203.01305Syntology official (archive's flag): 10 ran · 17 ran (of which 3 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 1 violated, 5 with no contract checked; 11 where Syntology's instrument failed) · 5 unverified (of 22 harvested samples) · 9 pointer-only (licence)
-
FastFold: Reducing AlphaFold Training Time from 11 Days to 67 Hours 2 Mar 2022 · 1 repository · arXiv:2203.00854
-
MTet: Multi-domain Translation for English-Vietnamese 2 Mar 2022 · 1 repository
-
Protecting Celebrities from DeepFake with Identity Consistency Transformer 2 Mar 2022 · 1 repository · arXiv:2203.01318Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Temporal Context Matters: Enhancing Single Image Prediction with Disease Progression Representations 2 Mar 2022 · 0 repositories · arXiv:2203.01933
-
X-Trans2Cap: Cross-Modal Knowledge Transfer using Transformer for 3D Dense Captioning 2 Mar 2022 · 1 repository · arXiv:2203.00843
-
DeepNet: Scaling Transformers to 1,000 Layers 1 Mar 2022 · 6 repositories · arXiv:2203.00555Syntology official (archive's flag): 1 ran · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 3 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 3 pointer-only (licence)
-
Fast-R2D2: A Pretrained Recursive Neural Network based on Pruned CKY for Grammar Induction and Text Representation 1 Mar 2022 · 2 repositories · arXiv:2203.00281Syntology official: harvested for another paper · 0 ran · 1 unverified (of 1 harvested sample)
-
Temporal Perceiver: A General Architecture for Arbitrary Boundary Detection 1 Mar 2022 · 0 repositories · arXiv:2203.00307
-
Transformer Grammars: Augmenting Transformer Language Models with Syntactic Inductive Biases at Scale 1 Mar 2022 · 0 repositories · arXiv:2203.00633