Methods › General › Feedforward Networks › Position-Wise Feed-Forward Layer › Papers, page 92
Position-Wise Feed-Forward Layer
Papers archive 2025-07-28
archive papers tagged: 13,895 · with a code link: 6,514 · where Syntology ran a sample: 2,229 (1,902 with a run with no instrument failure, 327 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,229 of 13,895 tagged: 1,902 with a run with no instrument failure, 327 where every run was a failure of Syntology's instrument)
Page 92 of 139: papers 9,101 to 9,200 of 13,895, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Automated ICD Coding using Extreme Multi-label Long Text Transformer-based Models 12 Dec 2022 · 1 repository · arXiv:2212.05857
-
BeautyREC: Robust, Efficient, and Content-preserving Makeup Transfer 12 Dec 2022 · 0 repositories · arXiv:2212.05855
-
CTT-Net: A Multi-view Cross-token Transformer for Cataract Postoperative Visual Acuity Prediction 12 Dec 2022 · 1 repository · arXiv:2212.05794
-
Deep learning approaches to building rooftop thermal bridge detection from aerial images 12 Dec 2022 · 1 repository
-
NMS Strikes Back 12 Dec 2022 · 1 repository · arXiv:2212.06137
-
P-Transformer: Towards Better Document-to-Document Neural Machine Translation 12 Dec 2022 · 1 repository · arXiv:2212.05830
-
ROIFormer: Semantic-Aware Region of Interest Transformer for Efficient Self-Supervised Monocular Depth Estimation 12 Dec 2022 · 0 repositories · arXiv:2212.05729
-
Video Prediction by Efficient Transformers 12 Dec 2022 · 1 repository · arXiv:2212.06026Syntology official: harvested, nothing ran · 0 ran · 7 unverified (of 7 harvested samples)
-
Extending TrOCR for Text Localization-Free OCR of Full-Page Scanned Receipt Images 11 Dec 2022 · 0 repositories · arXiv:2212.05525
-
Joint Spatio-Temporal Modeling for the Semantic Change Detection in Remote Sensing Images 10 Dec 2022 · 3 repositories · arXiv:2212.05245Syntology official: no sample here; runs from other or unrecorded repositories · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Thinking Fast and Slow in Large Language Models 10 Dec 2022 · 0 repositories · arXiv:2212.05206
-
MAGVIT: Masked Generative Video Transformer 10 Dec 2022 · 1 repository · arXiv:2212.05199Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 10 harvested samples) · 1 pointer-only (licence)
-
Position Embedding Needs an Independent Layer Normalization 10 Dec 2022 · 1 repository · arXiv:2212.05262
-
SMILE: Scaling Mixture-of-Experts with Efficient Bi-level Routing 10 Dec 2022 · 0 repositories · arXiv:2212.05191
-
Dynamic Test-Time Augmentation via Differentiable Functions 9 Dec 2022 · 1 repository · arXiv:2212.04681
-
Masked Lip-Sync Prediction by Audio-Visual Contextual Exploitation in Transformers 9 Dec 2022 · 0 repositories · arXiv:2212.04970
-
MIMO Is All You Need : A Strong Multi-In-Multi-Out Baseline for Video Prediction 9 Dec 2022 · 1 repository · arXiv:2212.04655
-
Cross-Domain Synthetic-to-Real In-the-Wild Depth and Normal Estimation for 3D Scene Understanding 9 Dec 2022 · 0 repositories · arXiv:2212.05040
-
RCDT: Relational Remote Sensing Change Detection with Transformer 9 Dec 2022 · 1 repository · arXiv:2212.04869
-
Sparse Upcycling: Training Mixture-of-Experts from Dense Checkpoints 9 Dec 2022 · 1 repository · arXiv:2212.05055Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
TRBLLmaker -- Transformer Reads Between Lyrics Lines maker 9 Dec 2022 · 0 repositories · arXiv:2212.04917
-
Federated Learning for Inference at Anytime and Anywhere 8 Dec 2022 · 0 repositories · arXiv:2212.04084
-
Group Generalized Mean Pooling for Vision Transformer 8 Dec 2022 · 0 repositories · arXiv:2212.04114
-
Harnessing the Power of Multi-Task Pretraining for Ground-Truth Level Natural Language Explanations 8 Dec 2022 · 1 repository · arXiv:2212.04231Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 1 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples) · 6 pointer-only (licence)
-
NRTR: Neuron Reconstruction with Transformer from 3D Optical Microscopy Images 8 Dec 2022 · 0 repositories · arXiv:2212.04163
-
Gaussian Radar Transformer for Semantic Segmentation in Noisy Radar Data 7 Dec 2022 · 0 repositories · arXiv:2212.03690
-
Learning-To-Embed: Adopting Transformer based models for E-commerce Products Representation Learning 7 Dec 2022 · 0 repositories · arXiv:2212.03725
-
Multimodal Vision Transformers with Forced Attention for Behavior Analysis 7 Dec 2022 · 0 repositories · arXiv:2212.03968
-
A K-variate Time Series Is Worth K Words: Evolution of the Vanilla Transformer Architecture for Long-term Multivariate Time Series Forecasting 6 Dec 2022 · 0 repositories · arXiv:2212.02789
-
AbHE: All Attention-based Homography Estimation 6 Dec 2022 · 0 repositories · arXiv:2212.03029
-
Document-Level Abstractive Summarization 6 Dec 2022 · 1 repository · arXiv:2212.03013
-
IncepFormer: Efficient Inception Transformer with Pyramid Pooling for Semantic Segmentation 6 Dec 2022 · 1 repository · arXiv:2212.03035
-
Open World DETR: Transformer based Open World Object Detection 6 Dec 2022 · 0 repositories · arXiv:2212.02969
-
Pretrained Diffusion Models for Unified Human Motion Synthesis 6 Dec 2022 · 0 repositories · arXiv:2212.02837
-
Semantic-Conditional Diffusion Networks for Image Captioning 6 Dec 2022 · 2 repositories · arXiv:2212.03099
-
Simple Baseline for Weather Forecasting Using Spatiotemporal Context Aggregation Network 6 Dec 2022 · 1 repository · arXiv:2212.02952Syntology official (archive's flag): 5 ran · 5 ran (of which 4 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
UniGeo: Unifying Geometry Logical Reasoning via Reformulating Mathematical Expression 6 Dec 2022 · 2 repositories · arXiv:2212.02746Syntology official (archive's flag): 2 ran · 2 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified; every one of the 2 samples that ran constructed an object rather than computing a result (of 4 harvested samples) · 4 pointer-only (licence)
-
Video Object of Interest Segmentation 6 Dec 2022 · 0 repositories · arXiv:2212.02871
-
3D-LatentMapper: View Agnostic Single-View Reconstruction of 3D Shapes 5 Dec 2022 · 0 repositories · arXiv:2212.02184
-
FBLNet: FeedBack Loop Network for Driver Attention Prediction 5 Dec 2022 · 0 repositories · arXiv:2212.02096
-
MAP-Music2Vec: A Simple and Effective Baseline for Self-Supervised Music Audio Representation Learning 5 Dec 2022 · 0 repositories · arXiv:2212.02508
-
Mask Matching Transformer for Few-Shot Segmentation 5 Dec 2022 · 1 repository · arXiv:2301.01208
-
Retrieval as Attention: End-to-end Learning of Retrieval and Reading within a Single Transformer 5 Dec 2022 · 1 repository · arXiv:2212.02027
-
Unifying Vision, Text, and Layout for Universal Document Processing 5 Dec 2022 · 5 repositories · arXiv:2212.02623Syntology official: no sample here; runs from other or unrecorded repositories · 15 ran (of which 0 constructed an object rather than computing a result; 14 with no instrument failure: 0 honoured, 2 violated, 12 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 17 harvested samples) · 4 pointer-only (licence)
-
Joint Self-Supervised Image-Volume Representation Learning with Intra-Inter Contrastive Clustering 4 Dec 2022 · 0 repositories · arXiv:2212.01893
-
Melody transcription via generative pre-training 4 Dec 2022 · 1 repository · arXiv:2212.01884
-
Recognition and Prediction of Surgical Gestures and Trajectories Using Transformer Models in Robot-Assisted Surgery 3 Dec 2022 · 0 repositories · arXiv:2212.01683
-
FECAM: Frequency Enhanced Channel Attention Mechanism for Time Series Forecasting 2 Dec 2022 · 1 repository · arXiv:2212.01209
-
Relation-Aware Language-Graph Transformer for Question Answering 2 Dec 2022 · 1 repository · arXiv:2212.00975
-
Multi-scale Transformer Network with Edge-aware Pre-training for Cross-Modality MR Image Synthesis 2 Dec 2022 · 2 repositories · arXiv:2212.01108
-
Tackling Low-Resourced Sign Language Translation: UPC at WMT-SLT 22 2 Dec 2022 · 1 repository · arXiv:2212.01140
-
Towards Diverse, Relevant and Coherent Open-Domain Dialogue Generation via Hybrid Latent Variables 2 Dec 2022 · 0 repositories · arXiv:2212.01145
-
CHAPTER: Exploiting Convolutional Neural Network Adapters for Self-supervised Speech Models 1 Dec 2022 · 0 repositories · arXiv:2212.01282
-
Concealed Object Detection for Passive Millimeter-Wave Security Imaging Based on Task-Aligned Detection Transformer 1 Dec 2022 · 0 repositories · arXiv:2212.00313
-
CUNI Non-Autoregressive System for the WMT 22 Efficient Translation Shared Task 1 Dec 2022 · 0 repositories · arXiv:2212.00477
-
Explainable Artificial Intelligence for Improved Modeling of Processes 1 Dec 2022 · 1 repository · arXiv:2212.00695
-
Ghost-free High Dynamic Range Imaging via Hybrid CNN-Transformer and Structure Tensor 1 Dec 2022 · 1 repository · arXiv:2212.00595
-
Learning Progressive Modality-shared Transformers for Effective Visible-Infrared Person Re-identification 1 Dec 2022 · 1 repository · arXiv:2212.00226
-
DSNet: a simple yet efficient network with dual-stream attention for lesion segmentation 30 Nov 2022 · 0 repositories · arXiv:2211.16950
-
Part-based Face Recognition with Vision Transformers 30 Nov 2022 · 1 repository · arXiv:2212.00057Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Pattern Attention Transformer with Doughnut Kernel 30 Nov 2022 · 0 repositories · arXiv:2211.16961
-
Rephrasing the Reference for Non-Autoregressive Machine Translation 30 Nov 2022 · 0 repositories · arXiv:2211.16863
-
T2G-Former: Organizing Tabular Features into Relation Graphs Promotes Heterogeneous Feature Interaction 30 Nov 2022 · 1 repository · arXiv:2211.16887Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 1 violated, 9 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 13 harvested samples) · 1 pointer-only (licence)
-
Task-Specific Embeddings for Ante-Hoc Explainable Text Classification 30 Nov 2022 · 0 repositories · arXiv:2212.00086
-
Topological Data Analysis for Speech Processing 30 Nov 2022 · 0 repositories · arXiv:2211.17223
-
AirFormer: Predicting Nationwide Air Quality in China with Transformers 29 Nov 2022 · 1 repository · arXiv:2211.15979
-
Attribute De-biased Vision Transformer (AD-ViT) for Long-Term Person Re-identification 29 Nov 2022 · 1 repository
-
Hierarchical Transformer for Survival Prediction Using Multimodality Whole Slide Images and Genomics 29 Nov 2022 · 1 repository · arXiv:2211.16632
-
NoisyQuant: Noisy Bias-Enhanced Post-Training Activation Quantization for Vision Transformers 29 Nov 2022 · 1 repository · arXiv:2211.16056Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
PhraseTransformer: An Incorporation of Local Context Information into Sequence-to-sequence Semantic Parsing 29 Nov 2022 · 1 repository
-
SPARTAN: Sparse Hierarchical Memory for Parameter-Efficient Transformers 29 Nov 2022 · 1 repository · arXiv:2211.16634
-
A Light Touch Approach to Teaching Transformers Multi-view Geometry 28 Nov 2022 · 0 repositories · arXiv:2211.15107
-
BJTU-WeChat's Systems for the WMT22 Chat Translation Task 28 Nov 2022 · 0 repositories · arXiv:2211.15009
-
Connecting the Dots: Floorplan Reconstruction Using Two-Level Queries 28 Nov 2022 · 1 repository · arXiv:2211.15658
-
DQ-DETR: Dual Query Detection Transformer for Phrase Extraction and Grounding 28 Nov 2022 · 1 repository · arXiv:2211.15516
-
Summer: WeChat Neural Machine Translation Systems for the WMT22 Biomedical Translation Task 28 Nov 2022 · 0 repositories · arXiv:2211.15022
-
Superpoint Transformer for 3D Scene Instance Segmentation 28 Nov 2022 · 1 repository · arXiv:2211.15766
-
3DPPE: 3D Point Positional Encoding for Multi-Camera 3D Object Detection Transformers 27 Nov 2022 · 1 repository · arXiv:2211.14710
-
A Time Series is Worth 64 Words: Long-term Forecasting with Transformers 27 Nov 2022 · 8 repositories · arXiv:2211.14730Syntology official (archive's flag): 2 ran · 17 ran (of which 2 constructed an object rather than computing a result; 16 with no instrument failure: 0 honoured, 0 violated, 16 with no contract checked; 1 where Syntology's instrument failed) · 13 unverified (of 30 harvested samples)
-
Prototype as Query for Few Shot Semantic Segmentation 27 Nov 2022 · 1 repository · arXiv:2211.14764Syntology official: no sample here; runs from other or unrecorded repositories · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Semantic-Aware Local-Global Vision Transformer 27 Nov 2022 · 0 repositories · arXiv:2211.14705
-
CDDFuse: Correlation-Driven Dual-Branch Feature Decomposition for Multi-Modality Image Fusion 26 Nov 2022 · 3 repositories · arXiv:2211.14461Syntology official (archive's flag): 1 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Cross-Field Transformer for Diabetic Retinopathy Grading on Two-field Fundus Images 26 Nov 2022 · 1 repository · arXiv:2211.14552
-
How Crucial is Transformer in Decision Transformer? 26 Nov 2022 · 1 repository · arXiv:2211.14655Syntology official (archive's flag): 1 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 5 harvested samples) · 2 pointer-only (licence)
-
PatchGT: Transformer over Non-trainable Clusters for Learning Graph Representations 26 Nov 2022 · 1 repository · arXiv:2211.14425
-
Transformer-based Model for Word Level Language Identification in Code-mixed Kannada-English Texts 26 Nov 2022 · 0 repositories · arXiv:2211.14459
-
A System for Morphology-Task Generalization via Unified Representation and Behavior Distillation 25 Nov 2022 · 1 repository · arXiv:2211.14296
-
Aggregated Text Transformer for Scene Text Detection 25 Nov 2022 · 0 repositories · arXiv:2211.13984
-
Asynchronous Event-Triggered Control for Non-Linear Systems 25 Nov 2022 · 0 repositories · arXiv:2211.13846
-
BatmanNet: Bi-branch Masked Graph Transformer Autoencoder for Molecular Representation 25 Nov 2022 · 1 repository · arXiv:2211.13979
-
Degenerate Swin to Win: Plain Window-based Transformer without Sophisticated Operations 25 Nov 2022 · 0 repositories · arXiv:2211.14255
-
Galvatron: Efficient Transformer Training over Multiple GPUs Using Automatic Parallelism 25 Nov 2022 · 3 repositories · arXiv:2211.13878Syntology official (archive's flag): 2 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Interaction Region Visual Transformer for Egocentric Action Anticipation 25 Nov 2022 · 1 repository · arXiv:2211.14154
-
Molecular Joint Representation Learning via Multi-modal Information 25 Nov 2022 · 0 repositories · arXiv:2211.14042
-
RUST: Latent Neural Scene Representations from Unposed Imagery 25 Nov 2022 · 0 repositories · arXiv:2211.14306
-
Spatial-Spectral Transformer for Hyperspectral Image Denoising 25 Nov 2022 · 3 repositories · arXiv:2211.14090Syntology official: not harvested · 0 ran · 2 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
TAOTF: A Two-stage Approximately Orthogonal Training Framework in Deep Neural Networks 25 Nov 2022 · 0 repositories · arXiv:2211.13902
-
The Naughtyformer: A Transformer Understands Offensive Humor 25 Nov 2022 · 0 repositories · arXiv:2211.14369
-
MUSTER: A Multi-scale Transformer-based Decoder for Semantic Segmentation 25 Nov 2022 · 2 repositories · arXiv:2211.13928
-
A Self-Attention Ansatz for Ab-initio Quantum Chemistry 24 Nov 2022 · 3 repositories · arXiv:2211.13672Syntology community repositories only · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples)