Methods › General › Feedforward Networks › Position-Wise Feed-Forward Layer › Papers, page 96
Position-Wise Feed-Forward Layer
Papers archive 2025-07-28
archive papers tagged: 13,895 · with a code link: 6,514 · where Syntology ran a sample: 2,229 (1,902 with a run with no instrument failure, 327 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,229 of 13,895 tagged: 1,902 with a run with no instrument failure, 327 where every run was a failure of Syntology's instrument)
Page 96 of 139: papers 9,501 to 9,600 of 13,895, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Long-Form Video-Language Pre-Training with Multimodal Temporal Contrastive Learning 12 Oct 2022 · 1 repository · arXiv:2210.06031Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
MotionBERT: A Unified Perspective on Learning Human Motion Representations 12 Oct 2022 · 1 repository · arXiv:2210.06551Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 11 harvested samples) · 4 pointer-only (licence)
-
Integrating Translation Memories into Non-Autoregressive Machine Translation 12 Oct 2022 · 1 repository · arXiv:2210.06020
-
Prepended Domain Transformer: Heterogeneous Face Recognition without Bells and Whistles 12 Oct 2022 · 2 repositories · arXiv:2210.06529
-
S4ND: Modeling Images and Videos as Multidimensional Signals Using State Spaces 12 Oct 2022 · 1 repository · arXiv:2210.06583Syntology community repositories only · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Towards Theoretically Inspired Neural Initialization Optimization 12 Oct 2022 · 1 repository · arXiv:2210.05956
-
Uplift and Upsample: Efficient 3D Human Pose Estimation with Uplifting Transformers 12 Oct 2022 · 2 repositories · arXiv:2210.06110Syntology official (archive's flag): 7 ran · 15 ran (of which 0 constructed an object rather than computing a result; 15 with no instrument failure: 0 honoured, 0 violated, 15 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 17 harvested samples) · 2 pointer-only (licence)
-
ZITS++: Image Inpainting by Improving the Incremental Transformer on Structural Priors 12 Oct 2022 · 2 repositories · arXiv:2210.05950Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples) · 2 pointer-only (licence)
-
An Exploration of Hierarchical Attention Transformers for Efficient Long Document Classification 11 Oct 2022 · 0 repositories · arXiv:2210.05529
-
Reliable Conditioning of Behavioral Cloning for Offline Reinforcement Learning 11 Oct 2022 · 1 repository · arXiv:2210.05158
-
Enriching Biomedical Knowledge for Low-resource Language Through Large-Scale Translation 11 Oct 2022 · 1 repository · arXiv:2210.05598
-
Mixture of Attention Heads: Selecting Attention Heads Per Token 11 Oct 2022 · 2 repositories · arXiv:2210.05144Syntology official (archive's flag): 5 ran · 13 ran (of which 2 constructed an object rather than computing a result; 10 with no instrument failure: 1 honoured, 0 violated, 9 with no contract checked; 3 where Syntology's instrument failed) · 4 unverified (of 17 harvested samples) · 4 pointer-only (licence)
-
Point Transformer V2: Grouped Vector Attention and Partition-based Pooling 11 Oct 2022 · 2 repositories · arXiv:2210.05666
-
SaiT: Sparse Vision Transformers through Adaptive Token Pruning 11 Oct 2022 · 1 repository · arXiv:2210.05832
-
Streaming Punctuation for Long-form Dictation with Transformers 11 Oct 2022 · 0 repositories · arXiv:2210.05756
-
Understanding the Failure of Batch Normalization for Transformers in NLP 11 Oct 2022 · 1 repository · arXiv:2210.05153Syntology official (archive's flag): 9 ran · 9 ran (of which 8 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
Viterbi Decoding of Directed Acyclic Transformer for Non-Autoregressive Machine Translation 11 Oct 2022 · 1 repository · arXiv:2210.05193Syntology official (archive's flag): 2 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
A Memory Transformer Network for Incremental Learning 10 Oct 2022 · 0 repositories · arXiv:2210.04485
-
Characterization of anomalous diffusion through convolutional transformers 10 Oct 2022 · 0 repositories · arXiv:2210.04959
-
DCVQE: A Hierarchical Transformer for Video Quality Assessment 10 Oct 2022 · 0 repositories · arXiv:2210.04377
-
Ensemble Learning using Transformers and Convolutional Networks for Masked Face Recognition 10 Oct 2022 · 1 repository · arXiv:2210.04816
-
FS-DETR: Few-Shot DEtection TRansformer with prompting and without re-training 10 Oct 2022 · 0 repositories · arXiv:2210.04845
-
LAPFormer: A Light and Accurate Polyp Segmentation Transformer 10 Oct 2022 · 0 repositories · arXiv:2210.04393
-
LMQFormer: A Laplace-Prior-Guided Mask Query Transformer for Lightweight Snow Removal 10 Oct 2022 · 1 repository · arXiv:2210.04787
-
MMT: Image-guided Story Ending Generation with Multimodal Memory Transformer 10 Oct 2022 · 1 repository
-
Revisiting adapters with adversarial training 10 Oct 2022 · 0 repositories · arXiv:2210.04886
-
SCAM! Transferring humans between images with Semantic Cross Attention Modulation 10 Oct 2022 · 1 repository · arXiv:2210.04883
-
Visual Prompt Tuning for Test-time Domain Adaptation 10 Oct 2022 · 0 repositories · arXiv:2210.04831
-
A Transformer-based deep neural network model for SSVEP classification 9 Oct 2022 · 2 repositories · arXiv:2210.04172
-
AMPose: Alternately Mixed Global-Local Attention Model for 3D Human Pose Estimation 9 Oct 2022 · 0 repositories · arXiv:2210.04216
-
ConTra: (Con)text (Tra)nsformer for Cross-Modal Video Retrieval 9 Oct 2022 · 1 repository · arXiv:2210.04341
-
Deep Span Representations for Named Entity Recognition 9 Oct 2022 · 1 repository · arXiv:2210.04182
-
KSAT: Knowledge-infused Self Attention Transformer -- Integrating Multiple Domain-Specific Contexts 9 Oct 2022 · 0 repositories · arXiv:2210.04307
-
Learning Texture Transformer Network for Light Field Super-Resolution 9 Oct 2022 · 0 repositories · arXiv:2210.09293
-
Strong Gravitational Lensing Parameter Estimation with Vision Transformer 9 Oct 2022 · 1 repository · arXiv:2210.04143Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples)
-
Transformer-based Flood Scene Segmentation for Developing Countries 9 Oct 2022 · 0 repositories · arXiv:2210.04218
-
VoLTA: Vision-Language Transformer with Weakly-Supervised Local-Feature Alignment 9 Oct 2022 · 1 repository · arXiv:2210.04135Syntology official (archive's flag): 10 ran · 12 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 1 honoured, 0 violated, 8 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 14 harvested samples) · 4 pointer-only (licence)
-
FBNet: Feedback Network for Point Cloud Completion 8 Oct 2022 · 1 repository · arXiv:2210.03974
-
Hierarchical Graph Transformer with Adaptive Node Sampling 8 Oct 2022 · 1 repository · arXiv:2210.03930Syntology official (archive's flag): 6 ran · 6 ran (of which 2 constructed an object rather than computing a result; 5 with no instrument failure: 2 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Short Text Pre-training with Extended Token Classification for E-commerce Query Understanding 8 Oct 2022 · 0 repositories · arXiv:2210.03915
-
Towards Light Weight Object Detection System 8 Oct 2022 · 0 repositories · arXiv:2210.03861
-
Time-Space Transformers for Video Panoptic Segmentation 7 Oct 2022 · 0 repositories · arXiv:2210.03546
-
ByteTransformer: A High-Performance Transformer Boosted for Variable-Length Inputs 6 Oct 2022 · 1 repository · arXiv:2210.03052
-
Focal and Global Spatial-Temporal Transformer for Skeleton-based Action Recognition 6 Oct 2022 · 0 repositories · arXiv:2210.02693
-
Forecasting Bitcoin volatility spikes from whale transactions and CryptoQuant data using Synthesizer Transformer models 6 Oct 2022 · 1 repository · arXiv:2211.08281
-
Interpreting County Level COVID-19 Infection and Feature Sensitivity using Deep Learning Time Series Models 6 Oct 2022 · 1 repository · arXiv:2210.03258
-
Mask3D: Mask Transformer for 3D Semantic Instance Segmentation 6 Oct 2022 · 1 repository · arXiv:2210.03105Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 10 harvested samples)
-
MechRetro is a chemical-mechanism-driven graph learning framework for interpretable retrosynthesis prediction and pathway planning 6 Oct 2022 · 1 repository · arXiv:2210.02630
-
Melody Infilling with User-Provided Structural Context 6 Oct 2022 · 1 repository · arXiv:2210.02829
-
MuRAG: Multimodal Retrieval-Augmented Generator for Open Question Answering over Images and Text 6 Oct 2022 · 0 repositories · arXiv:2210.02928
-
Structure Representation Network and Uncertainty Feedback Learning for Dense Non-Uniform Fog Removal 6 Oct 2022 · 1 repository · arXiv:2210.03061Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Vision Transformer Based Model for Describing a Set of Images as a Story 6 Oct 2022 · 0 repositories · arXiv:2210.02762
-
VLSNR:Vision-Linguistics Coordination Time Sequence-aware News Recommendation 6 Oct 2022 · 2 repositories · arXiv:2210.02946
-
XDoc: Unified Pre-training for Cross-Format Document Understanding 6 Oct 2022 · 1 repository · arXiv:2210.02849
-
Exploring The Role of Mean Teachers in Self-supervised Masked Auto-Encoders 5 Oct 2022 · 1 repository · arXiv:2210.02077
-
FQDet: Fast-converging Query-based Detector 5 Oct 2022 · 2 repositories · arXiv:2210.02318Syntology official (archive's flag): 2 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 6 harvested samples) · 3 pointer-only (licence)
-
Medical Image Retrieval via Nearest Neighbor Search on Pre-trained Image Features 5 Oct 2022 · 1 repository · arXiv:2210.02401
-
Temporally Consistent Transformers for Video Generation 5 Oct 2022 · 2 repositories · arXiv:2210.02396
-
TgDLF2.0: Theory-guided deep-learning for electrical load forecasting via Transformer and transfer learning 5 Oct 2022 · 0 repositories · arXiv:2210.02448
-
WavSpA: Wavelet Space Attention for Boosting Transformers' Long Sequence Learning Ability 5 Oct 2022 · 1 repository · arXiv:2210.01989
-
A Perceptual Quality Metric for Video Frame Interpolation 4 Oct 2022 · 1 repository · arXiv:2210.01879Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Accurate Image Restoration with Attention Retractable Transformer 4 Oct 2022 · 1 repository · arXiv:2210.01427Syntology official (archive's flag): 4 ran · 4 ran (of which 4 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified; every one of the 4 samples that ran constructed an object rather than computing a result (of 8 harvested samples)
-
Bridged Transformer for Vision and Point Cloud 3D Object Detection 4 Oct 2022 · 2 repositories · arXiv:2210.01391
-
ImmFusion: Robust mmWave-RGB Fusion for 3D Human Body Reconstruction in All Weather Conditions 4 Oct 2022 · 0 repositories · arXiv:2210.01346
-
Improving Label-Deficient Keyword Spotting Through Self-Supervised Pretraining 4 Oct 2022 · 2 repositories · arXiv:2210.01703
-
Memory in humans and deep language models: Linking hypotheses for model augmentation 4 Oct 2022 · 0 repositories · arXiv:2210.01869
-
MOAT: Alternating Mobile Convolution and Attention Brings Strong Vision Models 4 Oct 2022 · 2 repositories · arXiv:2210.01820
-
MTSMAE: Masked Autoencoders for Multivariate Time-Series Forecasting 4 Oct 2022 · 0 repositories · arXiv:2210.02199
-
One Transformer Can Understand Both 2D & 3D Molecular Data 4 Oct 2022 · 1 repository · arXiv:2210.01765Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Early or Late Fusion Matters: Efficient RGB-D Fusion in Vision Transformers for 3D Object Recognition 3 Oct 2022 · 0 repositories · arXiv:2210.00843
-
Dual-former: Hybrid Self-attention Transformer for Efficient Image Restoration 3 Oct 2022 · 0 repositories · arXiv:2210.01069
-
Masked Spiking Transformer 3 Oct 2022 · 1 repository · arXiv:2210.01208
-
Enhancing Fine-Grained 3D Object Recognition using Hybrid Multi-Modal Vision Transformer-CNN Models 3 Oct 2022 · 1 repository · arXiv:2210.04613
-
Fully Transformer Network for Change Detection of Remote Sensing Images 3 Oct 2022 · 1 repository · arXiv:2210.00757
-
Introducing Vision Transformer for Alzheimer's Disease classification task with 3D input 3 Oct 2022 · 0 repositories · arXiv:2210.01177
-
Smooth image-to-image translations with latent space interpolations 3 Oct 2022 · 1 repository · arXiv:2210.00841
-
Contrastive Audio-Visual Masked Autoencoder 2 Oct 2022 · 1 repository · arXiv:2210.07839Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
DARTFormer: Finding The Best Type Of Attention 2 Oct 2022 · 0 repositories · arXiv:2210.00641
-
Deep-OCTA: Ensemble Deep Learning Approaches for Diabetic Retinopathy Analysis on OCTA Images 2 Oct 2022 · 1 repository · arXiv:2210.00515
-
Seeing Through the Noisy Dark: Towards Real-world Low-Light Image Enhancement and Denoising 2 Oct 2022 · 0 repositories · arXiv:2210.00545
-
Wide Attention Is The Way Forward For Transformers? 2 Oct 2022 · 0 repositories · arXiv:2210.00640
-
A Comparison of Transformer, Convolutional, and Recurrent Neural Networks on Phoneme Recognition 1 Oct 2022 · 0 repositories · arXiv:2210.00367
-
Cascaded Multi-Modal Mixing Transformers for Alzheimer's Disease Classification with Incomplete Data 1 Oct 2022 · 0 repositories · arXiv:2210.00255
-
Multimodal Analogical Reasoning over Knowledge Graphs 1 Oct 2022 · 2 repositories · arXiv:2210.00312Syntology official (archive's flag): 5 ran · 17 ran (of which 2 constructed an object rather than computing a result; 14 with no instrument failure: 1 honoured, 1 violated, 12 with no contract checked; 3 where Syntology's instrument failed) · 10 unverified (of 27 harvested samples) · 2 pointer-only (licence)
-
Adaptive Sparse and Monotonic Attention for Transformer-based Automatic Speech Recognition 30 Sep 2022 · 0 repositories · arXiv:2209.15176
-
Diffusion-based Image Translation using Disentangled Style and Content Representation 30 Sep 2022 · 1 repository · arXiv:2209.15264Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Dual Progressive Transformations for Weakly Supervised Semantic Segmentation 30 Sep 2022 · 1 repository · arXiv:2209.15211
-
State-specific protein-ligand complex structure prediction with a multi-scale deep generative model 30 Sep 2022 · 2 repositories · arXiv:2209.15171
-
Impact of Face Image Quality Estimation on Presentation Attack Detection 30 Sep 2022 · 0 repositories · arXiv:2209.15489
-
Improving Local Features with Relevant Spatial Information by Vision Transformer for Crowd Counting 30 Sep 2022 · 1 repository
-
FusionRetro: Molecule Representation Fusion via In-Context Learning for Retrosynthetic Planning 30 Sep 2022 · 1 repository · arXiv:2209.15315Syntology official (archive's flag): 5 ran · 5 ran (of which 3 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 11 unverified (of 16 harvested samples) · 16 pointer-only (licence)
-
PART: Pre-trained Authorship Representation Transformer 30 Sep 2022 · 1 repository · arXiv:2209.15373
-
Self-Distillation for Further Pre-training of Transformers 30 Sep 2022 · 0 repositories · arXiv:2210.02871
-
SpeechLM: Enhanced Speech Pre-Training with Unpaired Textual Data 30 Sep 2022 · 1 repository · arXiv:2209.15329
-
Visuo-Tactile Transformers for Manipulation 30 Sep 2022 · 1 repository · arXiv:2210.00121
-
3D UX-Net: A Large Kernel Volumetric ConvNet Modernizing Hierarchical Transformer for Medical Image Segmentation 29 Sep 2022 · 2 repositories · arXiv:2209.15076Syntology official (archive's flag): 2 ran · 5 ran (of which 3 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples) · 4 pointer-only (licence)
-
ConvRNN-T: Convolutional Augmented Recurrent Neural Network Transducers for Streaming Speech Recognition 29 Sep 2022 · 0 repositories · arXiv:2209.14868
-
DiGress: Discrete Denoising diffusion for graph generation 29 Sep 2022 · 3 repositories · arXiv:2209.14734Syntology official: harvested, nothing ran · 15 ran (of which 10 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 5 where Syntology's instrument failed) · 10 unverified (of 25 harvested samples)
-
Spikformer: When Spiking Neural Network Meets Transformer 29 Sep 2022 · 2 repositories · arXiv:2209.15425
-
360FusionNeRF: Panoramic Neural Radiance Fields with Joint Guidance 28 Sep 2022 · 1 repository · arXiv:2209.14265