Methods › General › Feedforward Networks › Position-Wise Feed-Forward Layer › Papers, page 111
Position-Wise Feed-Forward Layer
Papers archive 2025-07-28
archive papers tagged: 13,895 · with a code link: 6,514 · where Syntology ran a sample: 2,229 (1,902 with a run with no instrument failure, 327 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,229 of 13,895 tagged: 1,902 with a run with no instrument failure, 327 where every run was a failure of Syntology's instrument)
Page 111 of 139: papers 11,001 to 11,100 of 13,895, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
5th Place Solution for VSPW 2021 Challenge 13 Dec 2021 · 0 repositories · arXiv:2112.06379
-
Dependency Learning for Legal Judgment Prediction with a Unified Text-to-Text Transformer 13 Dec 2021 · 1 repository · arXiv:2112.06370
-
Embracing Single Stride 3D Object Detector with Sparse Transformer 13 Dec 2021 · 2 repositories · arXiv:2112.06375
-
Hformer: Hybrid CNN-Transformer for Fringe Order Prediction in Phase Unwrapping of Fringe Projection 13 Dec 2021 · 0 repositories · arXiv:2112.06759
-
Pedestrian Trajectory Prediction via Spatial Interaction Transformer Network 13 Dec 2021 · 0 repositories · arXiv:2112.06624
-
Improving Sequential Recommendations via Bidirectional Temporal Data Augmentation with Pre-training 13 Dec 2021 · 1 repository · arXiv:2112.06460
-
Implicit Transformer Network for Screen Content Image Continuous Super-Resolution 12 Dec 2021 · 1 repository · arXiv:2112.06174Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Improving Vision Transformers for Incremental Learning 12 Dec 2021 · 0 repositories · arXiv:2112.06103
-
Towards More Efficient Insertion Transformer with Fractional Positional Encoding 12 Dec 2021 · 1 repository · arXiv:2112.06295
-
COMPOSER: Compositional Reasoning of Group Activity in Videos with Keypoint-Only Modality 11 Dec 2021 · 1 repository · arXiv:2112.05892
-
Building a great multi-lingual teacher with sparsely-gated mixture of experts for speech recognition 10 Dec 2021 · 0 repositories · arXiv:2112.05820
-
Couplformer:Rethinking Vision Transformer with Coupling Attention Map 10 Dec 2021 · 1 repository · arXiv:2112.05425Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 6 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
Deep ViT Features as Dense Visual Descriptors 10 Dec 2021 · 1 repository · arXiv:2112.05814Syntology 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 1 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples)
-
Self-Supervised Transformers for fMRI representation 10 Dec 2021 · 2 repositories · arXiv:2112.05761
-
VUT: Versatile UI Transformer for Multi-Modal Multi-Task User Interface Modeling 10 Dec 2021 · 0 repositories · arXiv:2112.05692
-
3D Medical Point Transformer: Introducing Convolution to Attention Networks for Medical Point Cloud Analysis 9 Dec 2021 · 1 repository · arXiv:2112.04863
-
A Bilingual, OpenWorld Video Text Dataset and End-to-end Video Text Spotter with Transformer 9 Dec 2021 · 3 repositories · arXiv:2112.04888Syntology official (archive's flag): 16 ran · 16 ran (of which 0 constructed an object rather than computing a result; 14 with no instrument failure: 9 honoured, 1 violated, 4 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 19 harvested samples) · 19 pointer-only (licence)
-
Extending AdamW by Leveraging Its Second Moment and Magnitude 9 Dec 2021 · 0 repositories · arXiv:2112.06125
-
Fast Point Transformer 9 Dec 2021 · 1 repository · arXiv:2112.04702Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 1 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
PE-former: Pose Estimation Transformer 9 Dec 2021 · 1 repository · arXiv:2112.04981
-
Recurrent Glimpse-based Decoder for Detection with Transformer 9 Dec 2021 · 1 repository · arXiv:2112.04632
-
Semi-Supervised Medical Image Segmentation via Cross Teaching between CNN and Transformer 9 Dec 2021 · 1 repository · arXiv:2112.04894Syntology official (archive's flag): 1 ran · 5 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Garment4D: Garment Reconstruction from Point Cloud Sequences 8 Dec 2021 · 1 repository · arXiv:2112.04159Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Improving language models by retrieving from trillions of tokens 8 Dec 2021 · 2 repositories · arXiv:2112.04426Syntology 16 ran (of which 5 constructed an object rather than computing a result; 14 with no instrument failure: 0 honoured, 3 violated, 11 with no contract checked; 2 where Syntology's instrument failed) · 7 unverified (of 23 harvested samples) · 3 pointer-only (licence)
-
Joint Global and Local Hierarchical Priors for Learned Image Compression 8 Dec 2021 · 1 repository · arXiv:2112.04487
-
MASTAF: A Model-Agnostic Spatio-Temporal Attention Fusion Network for Few-shot Video Classification 8 Dec 2021 · 1 repository · arXiv:2112.04585
-
Transformaly -- Two (Feature Spaces) Are Better Than One 8 Dec 2021 · 1 repository · arXiv:2112.04185
-
A deep language model to predict metabolic network equilibria 7 Dec 2021 · 0 repositories · arXiv:2112.03588
-
Attention-Based Model and Deep Reinforcement Learning for Distribution of Event Processing Tasks 7 Dec 2021 · 1 repository · arXiv:2112.03835
-
Bootstrapping ViTs: Towards Liberating Vision Transformers from Pre-training 7 Dec 2021 · 1 repository · arXiv:2112.03552
-
Emulating Spatio-Temporal Realizations of Three-Dimensional Isotropic Turbulence via Deep Sequence Learning Models 7 Dec 2021 · 1 repository · arXiv:2112.03469
-
Regularity Learning via Explicit Distribution Modeling for Skeletal Video Anomaly Detection 7 Dec 2021 · 1 repository · arXiv:2112.03649
-
Relating transformers to models and neural representations of the hippocampal formation 7 Dec 2021 · 0 repositories · arXiv:2112.04035
-
SSAT: A Symmetric Semantic-Aware Transformer Network for Makeup Transfer and Removal 7 Dec 2021 · 2 repositories · arXiv:2112.03631
-
GETAM: Gradient-weighted Element-wise Transformer Attention Map for Weakly-supervised Semantic segmentation 6 Dec 2021 · 1 repository · arXiv:2112.02841
-
Offline Pre-trained Multi-Agent Decision Transformer: One Big Sequence Model Tackles All SMAC Tasks 6 Dec 2021 · 1 repository · arXiv:2112.02845
-
One-shot Talking Face Generation from Single-speaker Audio-Visual Correlation Learning 6 Dec 2021 · 0 repositories · arXiv:2112.02749
-
PTTR: Relational 3D Point Cloud Object Tracking with Transformer 6 Dec 2021 · 1 repository · arXiv:2112.02857
-
Scaling Up Influence Functions 6 Dec 2021 · 2 repositories · arXiv:2112.03052
-
Spatio-Temporal meets Wavelet: Disentangled Traffic Flow Forecasting via Efficient Spectral Graph Attention Network 6 Dec 2021 · 0 repositories · arXiv:2112.02740
-
Dynamic Token Normalization Improves Vision Transformers 5 Dec 2021 · 1 repository · arXiv:2112.02624Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample) · 1 pointer-only (licence)
-
Learning Tracking Representations via Dual-Branch Fully Transformer Networks 5 Dec 2021 · 1 repository · arXiv:2112.02571
-
Pose-guided Feature Disentangling for Occluded Person Re-identification Based on Transformer 5 Dec 2021 · 1 repository · arXiv:2112.02466
-
3rd Place: A Global and Local Dual Retrieval Solution to Facebook AI Image Similarity Challenge 4 Dec 2021 · 1 repository · arXiv:2112.02373
-
A Multi-Strategy based Pre-Training Method for Cold-Start Recommendation 4 Dec 2021 · 0 repositories · arXiv:2112.02275
-
LAVT: Language-Aware Vision Transformer for Referring Image Segmentation 4 Dec 2021 · 1 repository · arXiv:2112.02244
-
U2-Former: A Nested U-shaped Transformer for Image Restoration 4 Dec 2021 · 0 repositories · arXiv:2112.02279
-
CTIN: Robust Contextual Transformer Network for Inertial Navigation 3 Dec 2021 · 1 repository · arXiv:2112.02143
-
Efficient Two-Stage Detection of Human-Object Interactions with a Novel Unary-Pairwise Transformer 3 Dec 2021 · 1 repository · arXiv:2112.01838Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 2 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 1 pointer-only (licence)
-
NN-LUT: Neural Approximation of Non-Linear Operations for Efficient Transformer Inference 3 Dec 2021 · 0 repositories · arXiv:2112.02191
-
Single-Shot Black-Box Adversarial Attacks Against Malware Detectors: A Causal Language Model Approach 3 Dec 2021 · 0 repositories · arXiv:2112.01724
-
TransZero: Attribute-guided Transformer for Zero-Shot Learning 3 Dec 2021 · 1 repository · arXiv:2112.01683
-
Masked-attention Mask Transformer for Universal Image Segmentation 2 Dec 2021 · 7 repositories · arXiv:2112.01527Syntology community repositories only · 7 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples)
-
MutualFormer: Multi-Modality Representation Learning via Cross-Diffusion Attention 2 Dec 2021 · 1 repository · arXiv:2112.01177
-
PLSUM: Generating PT-BR Wikipedia by Summarizing Multiple Websites 2 Dec 2021 · 1 repository · arXiv:2112.01591
-
ScaleVLAD: Improving Multimodal Sentiment Analysis via Multi-Scale Fusion of Locally Descriptors 2 Dec 2021 · 0 repositories · arXiv:2112.01368
-
Self-supervised Video Transformer 2 Dec 2021 · 1 repository · arXiv:2112.01514Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
SwinTrack: A Simple and Strong Baseline for Transformer Tracking 2 Dec 2021 · 1 repository · arXiv:2112.00995Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
TBN-ViT: Temporal Bilateral Network with Vision Transformer for Video Scene Parsing 2 Dec 2021 · 0 repositories · arXiv:2112.01033
-
PTCT: Patches with 3D-Temporal Convolutional Transformer Network for Precipitation Nowcasting 2 Dec 2021 · 1 repository · arXiv:2112.01085
-
Uni-Perceiver: Pre-training Unified Architecture for Generic Perception for Zero-shot and Few-shot Tasks 2 Dec 2021 · 1 repository · arXiv:2112.01522
-
Visual-Semantic Transformer for Scene Text Recognition 2 Dec 2021 · 0 repositories · arXiv:2112.00948
-
Co-evolution Transformer for Protein Contact Prediction 1 Dec 2021 · 1 repository
-
Container: Context Aggregation Networks 1 Dec 2021 · 2 repositories
-
Cross-view Geo-localization with Layer-to-Layer Transformer 1 Dec 2021 · 0 repositories
-
Detecting Extratropical Cyclones of the Northern Hemisphere with Single Shot Detector 1 Dec 2021 · 0 repositories · arXiv:2112.01283
-
Do Transformers Really Perform Badly for Graph Representation? 1 Dec 2021 · 0 repositories
-
Federated Split Task-Agnostic Vision Transformer for COVID-19 CXR Diagnosis 1 Dec 2021 · 0 repositories
-
Focal Attention for Long-Range Interactions in Vision Transformers 1 Dec 2021 · 1 repository
-
Gauge Equivariant Transformer 1 Dec 2021 · 0 repositories
-
HRFormer: High-Resolution Vision Transformer for Dense Predict 1 Dec 2021 · 2 repositories
-
Integrating Tree Path in Transformer for Code Representation 1 Dec 2021 · 1 repository
-
Multi-View Stereo with Transformer 1 Dec 2021 · 0 repositories · arXiv:2112.00336
-
Raw Nav-merge Seismic Data to Subsurface Properties with MLP based Multi-Modal Information Unscrambler 1 Dec 2021 · 0 repositories
-
Score Transformer: Generating Musical Score from Note-level Representation 1 Dec 2021 · 1 repository · arXiv:2112.00355
-
Searching for Efficient Transformers for Language Modeling 1 Dec 2021 · 0 repositories
-
Shapeshifter: a Parameter-efficient Transformer using Factorized Reshaped Matrices 1 Dec 2021 · 1 repository
-
Speech-T: Transducer for Text to Speech and Beyond 1 Dec 2021 · 0 repositories
-
Systematic Generalization with Edge Transformers 1 Dec 2021 · 1 repository · arXiv:2112.00578
-
TEDGE-Caching: Transformer-based Edge Caching Towards 6G Networks 1 Dec 2021 · 0 repositories · arXiv:2112.00633
-
Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer 1 Dec 2021 · 1 repository
-
UniDoc: Unified Pretraining Framework for Document Understanding 1 Dec 2021 · 0 repositories
-
A Comparative Study of Transformers on Word Sense Disambiguation 30 Nov 2021 · 0 repositories · arXiv:2111.15417
-
HEAT: Holistic Edge Attention Transformer for Structured Reconstruction 30 Nov 2021 · 1 repository · arXiv:2111.15143
-
KARL-Trans-NER: Knowledge Aware Representation Learning for Named Entity Recognition using Transformers 30 Nov 2021 · 0 repositories · arXiv:2111.15436
-
Pixelated Butterfly: Simple and Efficient Sparse training for Neural Network Models 30 Nov 2021 · 1 repository · arXiv:2112.00029Syntology official: harvested, nothing ran · 4 ran (of which 4 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified; every one of the 4 samples that ran constructed an object rather than computing a result (of 6 harvested samples) · 5 pointer-only (licence)
-
Pyramid Adversarial Training Improves ViT Performance 30 Nov 2021 · 1 repository · arXiv:2111.15121
-
Robust Partial-to-Partial Point Cloud Registration in a Full Range 30 Nov 2021 · 1 repository · arXiv:2111.15606
-
Shunted Self-Attention via Multi-Scale Token Aggregation 30 Nov 2021 · 1 repository · arXiv:2111.15193
-
Building extraction with vision transformer 29 Nov 2021 · 0 repositories · arXiv:2111.15637
-
DAFormer: Improving Network Architectures and Training Strategies for Domain-Adaptive Semantic Segmentation 29 Nov 2021 · 3 repositories · arXiv:2111.14887Syntology official: no sample here; runs from other or unrecorded repositories · 6 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
End-to-End Referring Video Object Segmentation with Multimodal Transformers 29 Nov 2021 · 2 repositories · arXiv:2111.14821Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 5 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples) · 6 pointer-only (licence)
-
Mixed Precision Low-bit Quantization of Neural Network Language Models for Speech Recognition 29 Nov 2021 · 0 repositories · arXiv:2112.11438
-
Mixed Precision of Quantization of Transformer Language Models for Speech Recognition 29 Nov 2021 · 0 repositories · arXiv:2112.11540
-
On the rate of convergence of a classifier based on a Transformer encoder 29 Nov 2021 · 0 repositories · arXiv:2111.14574
-
Point-BERT: Pre-training 3D Point Cloud Transformers with Masked Point Modeling 29 Nov 2021 · 3 repositories · arXiv:2111.14819Syntology official (archive's flag): 4 ran · 5 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 0 violated, 1 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 2 pointer-only (licence)
-
Recurrent Vision Transformer for Solving Visual Reasoning Problems 29 Nov 2021 · 0 repositories · arXiv:2111.14576
-
Searching the Search Space of Vision Transformer 29 Nov 2021 · 2 repositories · arXiv:2111.14725
-
Sparse DETR: Efficient End-to-End Object Detection with Learnable Sparsity 29 Nov 2021 · 1 repository · arXiv:2111.14330Syntology official (archive's flag): 3 ran · 3 ran (of which 3 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified; every one of the 3 samples that ran constructed an object rather than computing a result (of 7 harvested samples)
-
TransMVSNet: Global Context-aware Multi-view Stereo Network with Transformers 29 Nov 2021 · 1 repository · arXiv:2111.14600