Methods › General › Feedforward Networks › Position-Wise Feed-Forward Layer › Papers, page 116
Position-Wise Feed-Forward Layer
Papers archive 2025-07-28
archive papers tagged: 13,895 · with a code link: 6,514 · where Syntology ran a sample: 2,229 (1,902 with a run with no instrument failure, 327 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,229 of 13,895 tagged: 1,902 with a run with no instrument failure, 327 where every run was a failure of Syntology's instrument)
Page 116 of 139: papers 11,501 to 11,600 of 13,895, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Primer: Searching for Efficient Transformers for Language Modeling 17 Sep 2021 · 4 repositories · arXiv:2109.08668Syntology official: no sample here; runs from other or unrecorded repositories · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 2 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Scaling Laws vs Model Architectures: How does Inductive Bias Influence Scaling? An Extensive Empirical Study on Language Tasks 17 Sep 2021 · 0 repositories
-
The JHU-Microsoft Submission for WMT21 Quality Estimation Shared Task 17 Sep 2021 · 0 repositories · arXiv:2109.08724
-
An End-to-End Transformer Model for 3D Object Detection 16 Sep 2021 · 1 repository · arXiv:2109.08141Syntology 6 ran (of which 3 constructed an object rather than computing a result; 5 with no instrument failure: 1 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 1 pointer-only (licence)
-
Fast-Slow Transformer for Visually Grounding Speech 16 Sep 2021 · 1 repository · arXiv:2109.08186
-
Label-Attention Transformer with Geometrically Coherent Objects for Image Captioning 16 Sep 2021 · 1 repository · arXiv:2109.07799
-
MeLT: Message-Level Transformer with Masked Document Representations as Pre-Training for Stance Detection 16 Sep 2021 · 1 repository · arXiv:2109.08113
-
Scaling Laws for Neural Machine Translation 16 Sep 2021 · 0 repositories · arXiv:2109.07740
-
Sparse Factorization of Large Square Matrices 16 Sep 2021 · 1 repository · arXiv:2109.08184
-
TANet: A new Paradigm for Global Face Super-resolution via Transformer-CNN Aggregation Network 16 Sep 2021 · 0 repositories · arXiv:2109.08174
-
The NiuTrans System for the WMT21 Efficiency Task 16 Sep 2021 · 1 repository · arXiv:2109.08003
-
The NiuTrans System for WNGT 2020 Efficiency Task 16 Sep 2021 · 2 repositories · arXiv:2109.08008Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Utterance-level neural confidence measure for end-to-end children speech recognition 16 Sep 2021 · 0 repositories · arXiv:2109.07750
-
Anchor DETR: Query Design for Transformer-Based Object Detection 15 Sep 2021 · 2 repositories · arXiv:2109.07107
-
Complementary Feature Enhanced Network with Vision Transformer for Image Dehazing 15 Sep 2021 · 1 repository · arXiv:2109.07100
-
Incorporating Residual and Normalization Layers into Analysis of Masked Language Models 15 Sep 2021 · 2 repositories · arXiv:2109.07152Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
MISSFormer: An Effective Medical Image Segmentation Transformer 15 Sep 2021 · 1 repository · arXiv:2109.07162
-
PnP-DETR: Towards Efficient Visual Analysis with Transformers 15 Sep 2021 · 1 repository · arXiv:2109.07036Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample)
-
Pose Transformers (POTR): Human Motion Prediction with Non-Autoregressive Transformers 15 Sep 2021 · 1 repository · arXiv:2109.07531
-
RetroPrime: A Diverse, plausible and Transformer-based method for Single-Step retrosynthesis predictions 15 Sep 2021 · 1 repository
-
Sequence Length is a Domain: Length-based Overfitting in Transformer Models 15 Sep 2021 · 1 repository · arXiv:2109.07276
-
SupCL-Seq: Supervised Contrastive Learning for Downstream Optimized Sequence Representations 15 Sep 2021 · 1 repository · arXiv:2109.07424
-
Towards Incremental Transformers: An Empirical Analysis of Transformer Models for Incremental NLU 15 Sep 2021 · 1 repository · arXiv:2109.07364
-
Transformer-based Lexically Constrained Headline Generation 15 Sep 2021 · 1 repository · arXiv:2109.07080
-
A Three Step Training Approach with Data Augmentation for Morphological Inflection 14 Sep 2021 · 0 repositories · arXiv:2109.07006
-
Evaluating Biomedical BERT Models for Vocabulary Alignment at Scale in the UMLS Metathesaurus 14 Sep 2021 · 0 repositories · arXiv:2109.13348
-
Structure-Enhanced Pop Music Generation via Harmony-Aware Learning 14 Sep 2021 · 1 repository · arXiv:2109.06441
-
Vision Transformer for Learning Driving Policies in Complex Multi-Agent Environments 14 Sep 2021 · 0 repositories · arXiv:2109.06514
-
Attention Weights in Transformer NMT Fail Aligning Words Between Sequences but Largely Explain Model Predictions 13 Sep 2021 · 0 repositories · arXiv:2109.05853
-
CDTrans: Cross-domain Transformer for Unsupervised Domain Adaptation 13 Sep 2021 · 2 repositories · arXiv:2109.06165Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 8 harvested samples) · 2 pointer-only (licence)
-
CPT: A Pre-Trained Unbalanced Transformer for Both Chinese Language Understanding and Generation 13 Sep 2021 · 1 repository · arXiv:2109.05729
-
KroneckerBERT: Learning Kronecker Decomposition for Pre-trained Language Models via Knowledge Distillation 13 Sep 2021 · 0 repositories · arXiv:2109.06243
-
On Pursuit of Designing Multi-modal Transformer for Video Grounding 13 Sep 2021 · 0 repositories · arXiv:2109.06085
-
Packed Levitated Marker for Entity and Relation Extraction 13 Sep 2021 · 2 repositories · arXiv:2109.06067Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 9 harvested samples) · 3 pointer-only (licence)
-
ArtiBoost: Boosting Articulated 3D Hand-Object Pose Estimation via Online Exploration and Synthesis 12 Sep 2021 · 2 repositories · arXiv:2109.05488
-
Constructing Phrase-level Semantic Labels to Form Multi-Grained Supervision for Image-Text Retrieval 12 Sep 2021 · 0 repositories · arXiv:2109.05523
-
Levenshtein Training for Word-level Quality Estimation 12 Sep 2021 · 1 repository · arXiv:2109.05611
-
Single-Read Reconstruction for DNA Data Storage Using Transformers 12 Sep 2021 · 0 repositories · arXiv:2109.05478
-
Sparse MLP for Image Recognition: Is Self-Attention Really Necessary? 12 Sep 2021 · 2 repositories · arXiv:2109.05422
-
TEASEL: A Transformer-Based Speech-Prefixed Language Model 12 Sep 2021 · 1 repository · arXiv:2109.05522Syntology 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples)
-
Bornon: Bengali Image Captioning with Transformer-based Deep learning approach 11 Sep 2021 · 0 repositories · arXiv:2109.05218
-
Empirical Analysis of Training Strategies of Transformer-based Japanese Chit-chat Systems 11 Sep 2021 · 1 repository · arXiv:2109.05217
-
Multilingual Translation via Grafting Pre-trained Language Models 11 Sep 2021 · 1 repository · arXiv:2109.05256
-
Real-time multimodal image registration with partial intraoperative point-set data 10 Sep 2021 · 0 repositories · arXiv:2109.05023
-
Temporal Pyramid Transformer with Multimodal Interaction for Video Question Answering 10 Sep 2021 · 1 repository · arXiv:2109.04735
-
A Three-Stage Learning Framework for Low-Resource Knowledge-Grounded Dialogue Generation 9 Sep 2021 · 1 repository · arXiv:2109.04096Syntology official (archive's flag): 18 ran · 18 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 1 violated, 11 with no contract checked; 6 where Syntology's instrument failed) · 9 unverified (of 27 harvested samples) · 5 pointer-only (licence)
-
Bag of Tricks for Optimizing Transformer Efficiency 9 Sep 2021 · 1 repository · arXiv:2109.04030
-
DAN: Decentralized Attention-based Neural Network for the MinMax Multiple Traveling Salesman Problem 9 Sep 2021 · 0 repositories · arXiv:2109.04205
-
ESimCSE: Enhanced Sample Building Method for Contrastive Learning of Unsupervised Sentence Embedding 9 Sep 2021 · 2 repositories · arXiv:2109.04380
-
Graph-Based Decoding for Task Oriented Semantic Parsing 9 Sep 2021 · 0 repositories · arXiv:2109.04587
-
MATE: Multi-view Attention for Table Transformer Efficiency 9 Sep 2021 · 1 repository · arXiv:2109.04312
-
Thinking Clearly, Talking Fast: Concept-Guided Non-Autoregressive Generation for Open-Domain Dialogue Systems 9 Sep 2021 · 1 repository · arXiv:2109.04084Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 1 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 10 harvested samples) · 2 pointer-only (licence)
-
UCTransNet: Rethinking the Skip Connections in U-Net from a Channel-wise Perspective with Transformer 9 Sep 2021 · 3 repositories · arXiv:2109.04335Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Variational Latent-State GPT for Semi-Supervised Task-Oriented Dialog Systems 9 Sep 2021 · 2 repositories · arXiv:2109.04314
-
Label Verbalization and Entailment for Effective Zero- and Few-Shot Relation Extraction 8 Sep 2021 · 1 repository · arXiv:2109.03659
-
Panoptic SegFormer: Delving Deeper into Panoptic Segmentation with Transformers 8 Sep 2021 · 3 repositories · arXiv:2109.03814
-
Retrieve, Caption, Generate: Visual Grounding for Enhancing Commonsense in Text Generation Models 8 Sep 2021 · 0 repositories · arXiv:2109.03892
-
What's Hidden in a One-layer Randomly Weighted Transformer? 8 Sep 2021 · 1 repository · arXiv:2109.03939
-
FuseFormer: Fusing Fine-Grained Information in Transformers for Video Inpainting 7 Sep 2021 · 1 repository · arXiv:2109.02974Syntology official (archive's flag): 10 ran · 10 ran (of which 8 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
Hierarchical Graph Convolutional Skeleton Transformer for Action Recognition 7 Sep 2021 · 0 repositories · arXiv:2109.02860
-
Mixed Attention Transformer for Leveraging Word-Level Knowledge to Neural Cross-Lingual Information Retrieval 7 Sep 2021 · 0 repositories · arXiv:2109.02789
-
nnFormer: Interleaved Transformer for Volumetric Segmentation 7 Sep 2021 · 2 repositories · arXiv:2109.03201Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Puzzle Solving without Search or Human Knowledge: An Unnatural Language Approach 7 Sep 2021 · 0 repositories · arXiv:2109.02797
-
Rendezvous: Attention Mechanisms for the Recognition of Surgical Action Triplets in Endoscopic Videos 7 Sep 2021 · 8 repositories · arXiv:2109.03223
-
3D Human Texture Estimation from a Single Image with Transformers 6 Sep 2021 · 1 repository · arXiv:2109.02563Syntology official (archive's flag): 14 ran · 14 ran (of which 7 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 7 where Syntology's instrument failed) · 2 unverified (of 16 harvested samples) · 16 pointer-only (licence)
-
Eliminating Sentiment Bias for Aspect-Level Sentiment Classification with Unsupervised Opinion Extraction 6 Sep 2021 · 1 repository · arXiv:2109.02403
-
Enhancing Natural Language Representation with Large-Scale Out-of-Domain Commonsense 6 Sep 2021 · 1 repository · arXiv:2109.02572
-
PermuteFormer: Efficient Relative Position Encoding for Long Sequences 6 Sep 2021 · 1 repository · arXiv:2109.02377Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample) · 1 pointer-only (licence)
-
The Animation Transformer: Visual Correspondence via Segment Matching 6 Sep 2021 · 0 repositories · arXiv:2109.02614
-
Vision Transformers For Weeds and Crops Classification Of High Resolution UAV Images 6 Sep 2021 · 0 repositories · arXiv:2109.02716
-
Voxel Transformer for 3D Object Detection 6 Sep 2021 · 1 repository · arXiv:2109.02497Syntology 0 ran · 7 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
Transformer Models for Text Coherence Assessment 5 Sep 2021 · 2 repositories · arXiv:2109.02176
-
Error Detection in Large-Scale Natural Language Understanding Systems Using Transformer Models 4 Sep 2021 · 0 repositories · arXiv:2109.01754
-
Contextualized Embeddings based Convolutional Neural Networks for Duplicate Question Identification 3 Sep 2021 · 0 repositories · arXiv:2109.01560
-
IMG2SMI: Translating Molecular Structure Images to Simplified Molecular-input Line-entry System 3 Sep 2021 · 0 repositories · arXiv:2109.04202
-
Semantic Segmentation on VSPW Dataset through Aggregation of Transformer Models 3 Sep 2021 · 0 repositories · arXiv:2109.01316
-
CodeT5: Identifier-aware Unified Pre-trained Encoder-Decoder Models for Code Understanding and Generation 2 Sep 2021 · 5 repositories · arXiv:2109.00859Syntology official: harvested, nothing ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 5 unverified (of 11 harvested samples)
-
Common Objects in 3D: Large-Scale Learning and Evaluation of Real-life 3D Category Reconstruction 1 Sep 2021 · 1 repository · arXiv:2109.00512
-
CTAL: Pre-training Cross-modal Transformer for Audio-and-Language Representations 1 Sep 2021 · 1 repository · arXiv:2109.00181Syntology official (archive's flag): 13 ran · 15 ran (of which 2 constructed an object rather than computing a result; 15 with no instrument failure: 1 honoured, 0 violated, 14 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 19 harvested samples) · 6 pointer-only (licence)
-
∞-former: Infinite Memory Transformer 1 Sep 2021 · 1 repository · arXiv:2109.00301
-
Searching for Efficient Multi-Stage Vision Transformers 1 Sep 2021 · 1 repository · arXiv:2109.00642
-
Stochastic Transformer Networks with Linear Competing Units: Application to end-to-end SL Translation 1 Sep 2021 · 1 repository · arXiv:2109.13318Syntology official (archive's flag): 5 ran · 5 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
AraT5: Text-to-Text Transformers for Arabic Language Generation 31 Aug 2021 · 1 repository · arXiv:2109.12068
-
SANSformers: Self-Supervised Forecasting in Electronic Health Records with Attention-Free Models 31 Aug 2021 · 0 repositories · arXiv:2108.13672
-
A Battle of Network Structures: An Empirical Study of CNN, Transformer, and MLP 30 Aug 2021 · 1 repository · arXiv:2108.13002
-
From General to Specific: Informative Scene Graph Generation via Balance Adjustment 30 Aug 2021 · 1 repository · arXiv:2108.13129Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Multi-Channel Transformer Transducer for Speech Recognition 30 Aug 2021 · 0 repositories · arXiv:2108.12953
-
Scheduled Sampling Based on Decoding Steps for Neural Machine Translation 30 Aug 2021 · 1 repository · arXiv:2108.12963
-
Shatter: An Efficient Transformer Encoder with Single-Headed Self-Attention and Relative Sequence Partitioning 30 Aug 2021 · 0 repositories · arXiv:2108.13032
-
TCCT: Tightly-Coupled Convolutional Transformer on Time Series Forecasting 29 Aug 2021 · 2 repositories · arXiv:2108.12784
-
AMMASurv: Asymmetrical Multi-Modal Attention for Accurate Survival Analysis with Whole Slide Images and Gene Expression Data 28 Aug 2021 · 0 repositories · arXiv:2108.12565
-
GroupFormer: Group Activity Recognition with Clustered Spatial-Temporal Transformer 28 Aug 2021 · 1 repository · arXiv:2108.12630Syntology official (archive's flag): 20 ran · 22 ran (of which 2 constructed an object rather than computing a result; 8 with no instrument failure: 3 honoured, 0 violated, 5 with no contract checked; 14 where Syntology's instrument failed) · 5 unverified (of 27 harvested samples) · 2 pointer-only (licence)
-
Towards Fine-grained Image Classification with Generative Adversarial Networks and Facial Landmark Detection 28 Aug 2021 · 1 repository · arXiv:2109.00891
-
Lyra: A Benchmark for Turducken-Style Code Generation 27 Aug 2021 · 1 repository · arXiv:2108.12144
-
Can the Transformer Be Used as a Drop-in Replacement for RNNs in Text-Generating GANs? 26 Aug 2021 · 0 repositories · arXiv:2108.12275
-
Evaluating Transformer-based Semantic Segmentation Networks for Pathological Image Segmentation 26 Aug 2021 · 0 repositories · arXiv:2108.11993
-
DeepGene Transformer: Transformer for the gene expression-based classification of cancer subtypes 26 Aug 2021 · 0 repositories · arXiv:2108.11833
-
LayoutReader: Pre-training of Text and Layout for Reading Order Detection 26 Aug 2021 · 1 repository · arXiv:2108.11591
-
Reiterative Domain Aware Multi-Target Adaptation 26 Aug 2021 · 0 repositories · arXiv:2109.00919
-
Shifted Chunk Transformer for Spatio-Temporal Representational Learning 26 Aug 2021 · 0 repositories · arXiv:2108.11575