Methods › General › Normalization › Layer Normalization › Papers, page 187
Layer Normalization
Papers archive 2025-07-28
archive papers tagged: 24,980 · with a code link: 11,273 · where Syntology ran a sample: 3,471 (2,923 with a run with no instrument failure, 548 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,471 of 24,980 tagged: 2,923 with a run with no instrument failure, 548 where every run was a failure of Syntology's instrument)
Page 187 of 250: papers 18,601 to 18,700 of 24,980, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Pedestrian Trajectory Prediction via Spatial Interaction Transformer Network 13 Dec 2021 · 0 repositories · arXiv:2112.06624
-
Roof-Transformer: Divided and Joined Understanding with Knowledge Enhancement 13 Dec 2021 · 0 repositories · arXiv:2112.06736
-
Improving Sequential Recommendations via Bidirectional Temporal Data Augmentation with Pre-training 13 Dec 2021 · 1 repository · arXiv:2112.06460
-
WECHSEL: Effective initialization of subword embeddings for cross-lingual transfer of monolingual language models 13 Dec 2021 · 1 repository · arXiv:2112.06598Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Implicit Transformer Network for Screen Content Image Continuous Super-Resolution 12 Dec 2021 · 1 repository · arXiv:2112.06174Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Improving Logical-Level Natural Language Generation with Topic-Conditioned Data Augmentation and Logical Form Generation 12 Dec 2021 · 0 repositories · arXiv:2112.06240
-
Improving Vision Transformers for Incremental Learning 12 Dec 2021 · 0 repositories · arXiv:2112.06103
-
Towards More Efficient Insertion Transformer with Fractional Positional Encoding 12 Dec 2021 · 1 repository · arXiv:2112.06295
-
COMPOSER: Compositional Reasoning of Group Activity in Videos with Keypoint-Only Modality 11 Dec 2021 · 1 repository · arXiv:2112.05892
-
Building a great multi-lingual teacher with sparsely-gated mixture of experts for speech recognition 10 Dec 2021 · 0 repositories · arXiv:2112.05820
-
Couplformer:Rethinking Vision Transformer with Coupling Attention Map 10 Dec 2021 · 1 repository · arXiv:2112.05425Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 6 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
Deep ViT Features as Dense Visual Descriptors 10 Dec 2021 · 1 repository · arXiv:2112.05814Syntology 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 1 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples)
-
Findings on Conversation Disentanglement 10 Dec 2021 · 0 repositories · arXiv:2112.05346
-
Multimodal Interactions Using Pretrained Unimodal Models for SIMMC 2.0 10 Dec 2021 · 1 repository · arXiv:2112.05328
-
Self-Supervised Transformers for fMRI representation 10 Dec 2021 · 2 repositories · arXiv:2112.05761
-
Sketching as a Tool for Understanding and Accelerating Self-attention for Long Sequences 10 Dec 2021 · 1 repository · arXiv:2112.05359
-
VUT: Versatile UI Transformer for Multi-Modal Multi-Task User Interface Modeling 10 Dec 2021 · 0 repositories · arXiv:2112.05692
-
3D Medical Point Transformer: Introducing Convolution to Attention Networks for Medical Point Cloud Analysis 9 Dec 2021 · 1 repository · arXiv:2112.04863
-
A Bilingual, OpenWorld Video Text Dataset and End-to-end Video Text Spotter with Transformer 9 Dec 2021 · 3 repositories · arXiv:2112.04888Syntology official (archive's flag): 16 ran · 16 ran (of which 0 constructed an object rather than computing a result; 14 with no instrument failure: 9 honoured, 1 violated, 4 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 19 harvested samples) · 19 pointer-only (licence)
-
Detecting potentially harmful and protective suicide-related content on twitter: A machine learning approach 9 Dec 2021 · 2 repositories · arXiv:2112.04796
-
Extending AdamW by Leveraging Its Second Moment and Magnitude 9 Dec 2021 · 0 repositories · arXiv:2112.06125
-
Fast Point Transformer 9 Dec 2021 · 1 repository · arXiv:2112.04702Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 1 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
From Scattered Sources to Comprehensive Technology Landscape: A Recommendation-based Retrieval Approach 9 Dec 2021 · 0 repositories · arXiv:2112.04810
-
Injecting Semantic Concepts into End-to-End Image Captioning 9 Dec 2021 · 1 repository · arXiv:2112.05230Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
PE-former: Pose Estimation Transformer 9 Dec 2021 · 1 repository · arXiv:2112.04981
-
Recurrent Glimpse-based Decoder for Detection with Transformer 9 Dec 2021 · 1 repository · arXiv:2112.04632
-
Semantic Search as Extractive Paraphrase Span Detection 9 Dec 2021 · 1 repository · arXiv:2112.04886
-
Semi-Supervised Medical Image Segmentation via Cross Teaching between CNN and Transformer 9 Dec 2021 · 1 repository · arXiv:2112.04894Syntology official (archive's flag): 1 ran · 5 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Towards Neural Functional Program Evaluation 9 Dec 2021 · 0 repositories · arXiv:2112.04630
-
Garment4D: Garment Reconstruction from Point Cloud Sequences 8 Dec 2021 · 1 repository · arXiv:2112.04159Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Improving language models by retrieving from trillions of tokens 8 Dec 2021 · 2 repositories · arXiv:2112.04426Syntology 16 ran (of which 5 constructed an object rather than computing a result; 14 with no instrument failure: 0 honoured, 3 violated, 11 with no contract checked; 2 where Syntology's instrument failed) · 7 unverified (of 23 harvested samples) · 3 pointer-only (licence)
-
JABER and SABER: Junior and Senior Arabic BERt 8 Dec 2021 · 1 repository · arXiv:2112.04329
-
Joint Global and Local Hierarchical Priors for Learned Image Compression 8 Dec 2021 · 1 repository · arXiv:2112.04487
-
MASTAF: A Model-Agnostic Spatio-Temporal Attention Fusion Network for Few-shot Video Classification 8 Dec 2021 · 1 repository · arXiv:2112.04585
-
Transformaly -- Two (Feature Spaces) Are Better Than One 8 Dec 2021 · 1 repository · arXiv:2112.04185
-
A deep language model to predict metabolic network equilibria 7 Dec 2021 · 0 repositories · arXiv:2112.03588
-
A Transferable Approach for Partitioning Machine Learning Models on Multi-Chip-Modules 7 Dec 2021 · 0 repositories · arXiv:2112.04041
-
Attention-Based Model and Deep Reinforcement Learning for Distribution of Event Processing Tasks 7 Dec 2021 · 1 repository · arXiv:2112.03835
-
Bootstrapping ViTs: Towards Liberating Vision Transformers from Pre-training 7 Dec 2021 · 1 repository · arXiv:2112.03552
-
Emulating Spatio-Temporal Realizations of Three-Dimensional Isotropic Turbulence via Deep Sequence Learning Models 7 Dec 2021 · 1 repository · arXiv:2112.03469
-
raceBERT -- A Transformer-based Model for Predicting Race and Ethnicity from Names 7 Dec 2021 · 1 repository · arXiv:2112.03807
-
Regularity Learning via Explicit Distribution Modeling for Skeletal Video Anomaly Detection 7 Dec 2021 · 1 repository · arXiv:2112.03649
-
Relating transformers to models and neural representations of the hippocampal formation 7 Dec 2021 · 0 repositories · arXiv:2112.04035
-
SSAT: A Symmetric Semantic-Aware Transformer Network for Makeup Transfer and Removal 7 Dec 2021 · 2 repositories · arXiv:2112.03631
-
GETAM: Gradient-weighted Element-wise Transformer Attention Map for Weakly-supervised Semantic segmentation 6 Dec 2021 · 1 repository · arXiv:2112.02841
-
Offline Pre-trained Multi-Agent Decision Transformer: One Big Sequence Model Tackles All SMAC Tasks 6 Dec 2021 · 1 repository · arXiv:2112.02845
-
One-shot Talking Face Generation from Single-speaker Audio-Visual Correlation Learning 6 Dec 2021 · 0 repositories · arXiv:2112.02749
-
PTTR: Relational 3D Point Cloud Object Tracking with Transformer 6 Dec 2021 · 1 repository · arXiv:2112.02857
-
Scaling Up Influence Functions 6 Dec 2021 · 2 repositories · arXiv:2112.03052
-
Spatio-Temporal meets Wavelet: Disentangled Traffic Flow Forecasting via Efficient Spectral Graph Attention Network 6 Dec 2021 · 0 repositories · arXiv:2112.02740
-
Team Hitachi @ AutoMin 2021: Reference-free Automatic Minuting Pipeline with Argument Structure Construction over Topic-based Summarization 6 Dec 2021 · 0 repositories · arXiv:2112.02741
-
BERTMap: A BERT-based Ontology Alignment System 5 Dec 2021 · 1 repository · arXiv:2112.02682
-
Causal Distillation for Language Models 5 Dec 2021 · 1 repository · arXiv:2112.02505Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
DIBERT: Dependency Injected Bidirectional Encoder Representations from Transformers 5 Dec 2021 · 1 repository
-
Dynamic Token Normalization Improves Vision Transformers 5 Dec 2021 · 1 repository · arXiv:2112.02624Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample) · 1 pointer-only (licence)
-
Gaudí: Conversational Interactions with Deep Representations to Generate Image Collections 5 Dec 2021 · 0 repositories · arXiv:2112.04404
-
Learning Tracking Representations via Dual-Branch Fully Transformer Networks 5 Dec 2021 · 1 repository · arXiv:2112.02571
-
PolyphonicFormer: Unified Query Learning for Depth-aware Video Panoptic Segmentation 5 Dec 2021 · 1 repository · arXiv:2112.02582
-
Pose-guided Feature Disentangling for Occluded Person Re-identification Based on Transformer 5 Dec 2021 · 1 repository · arXiv:2112.02466
-
VarCLR: Variable Semantic Representation Pre-training via Contrastive Learning 5 Dec 2021 · 1 repository · arXiv:2112.02650Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
3rd Place: A Global and Local Dual Retrieval Solution to Facebook AI Image Similarity Challenge 4 Dec 2021 · 1 repository · arXiv:2112.02373
-
A Multi-Strategy based Pre-Training Method for Cold-Start Recommendation 4 Dec 2021 · 0 repositories · arXiv:2112.02275
-
Bridging Pre-trained Models and Downstream Tasks for Source Code Understanding 4 Dec 2021 · 1 repository · arXiv:2112.02268Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 0 violated, 12 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 13 harvested samples)
-
LAVT: Language-Aware Vision Transformer for Referring Image Segmentation 4 Dec 2021 · 1 repository · arXiv:2112.02244
-
Representation Learning for Conversational Data using Discourse Mutual Information Maximization 4 Dec 2021 · 0 repositories · arXiv:2112.05787
-
U2-Former: A Nested U-shaped Transformer for Image Restoration 4 Dec 2021 · 0 repositories · arXiv:2112.02279
-
Unraveling Social Perceptions & Behaviors towards Migrants on Twitter 4 Dec 2021 · 0 repositories · arXiv:2112.06642
-
A Novel Deep Parallel Time-series Relation Network for Fault Diagnosis 3 Dec 2021 · 0 repositories · arXiv:2112.03405
-
Augmenting Customer Support with an NLP-based Receptionist 3 Dec 2021 · 0 repositories · arXiv:2112.01959
-
CTIN: Robust Contextual Transformer Network for Inertial Navigation 3 Dec 2021 · 1 repository · arXiv:2112.02143
-
Efficient Two-Stage Detection of Human-Object Interactions with a Novel Unary-Pairwise Transformer 3 Dec 2021 · 1 repository · arXiv:2112.01838Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 2 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 1 pointer-only (licence)
-
Given Users Recommendations Based on Reviews on Yelp 3 Dec 2021 · 1 repository · arXiv:2112.01762
-
Make A Long Image Short: Adaptive Token Length for Vision Transformers 3 Dec 2021 · 0 repositories · arXiv:2112.01686
-
NN-LUT: Neural Approximation of Non-Linear Operations for Efficient Transformer Inference 3 Dec 2021 · 0 repositories · arXiv:2112.02191
-
Siamese BERT-based Model for Web Search Relevance Ranking Evaluated on a New Czech Dataset 3 Dec 2021 · 1 repository · arXiv:2112.01810
-
Single-Shot Black-Box Adversarial Attacks Against Malware Detectors: A Causal Language Model Approach 3 Dec 2021 · 0 repositories · arXiv:2112.01724
-
TransZero: Attribute-guided Transformer for Zero-Shot Learning 3 Dec 2021 · 1 repository · arXiv:2112.01683
-
BEVT: BERT Pretraining of Video Transformers 2 Dec 2021 · 1 repository · arXiv:2112.01529
-
Masked-attention Mask Transformer for Universal Image Segmentation 2 Dec 2021 · 7 repositories · arXiv:2112.01527Syntology community repositories only · 7 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples)
-
MutualFormer: Multi-Modality Representation Learning via Cross-Diffusion Attention 2 Dec 2021 · 1 repository · arXiv:2112.01177
-
PLSUM: Generating PT-BR Wikipedia by Summarizing Multiple Websites 2 Dec 2021 · 1 repository · arXiv:2112.01591
-
ScaleVLAD: Improving Multimodal Sentiment Analysis via Multi-Scale Fusion of Locally Descriptors 2 Dec 2021 · 0 repositories · arXiv:2112.01368
-
Self-supervised Video Transformer 2 Dec 2021 · 1 repository · arXiv:2112.01514Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
SwinTrack: A Simple and Strong Baseline for Transformer Tracking 2 Dec 2021 · 1 repository · arXiv:2112.00995Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
TBN-ViT: Temporal Bilateral Network with Vision Transformer for Video Scene Parsing 2 Dec 2021 · 0 repositories · arXiv:2112.01033
-
PTCT: Patches with 3D-Temporal Convolutional Transformer Network for Precipitation Nowcasting 2 Dec 2021 · 1 repository · arXiv:2112.01085
-
Uni-Perceiver: Pre-training Unified Architecture for Generic Perception for Zero-shot and Few-shot Tasks 2 Dec 2021 · 1 repository · arXiv:2112.01522
-
Unsupervised Law Article Mining based on Deep Pre-Trained Language Representation Models with Application to the Italian Civil Code 2 Dec 2021 · 0 repositories · arXiv:2112.03033
-
Visual-Semantic Transformer for Scene Text Recognition 2 Dec 2021 · 0 repositories · arXiv:2112.00948
-
Co-evolution Transformer for Protein Contact Prediction 1 Dec 2021 · 1 repository
-
Combining Global and Local Attention with Positional Encoding for Video Summarization 1 Dec 2021 · 1 repository
-
Container: Context Aggregation Networks 1 Dec 2021 · 2 repositories
-
Controlling Conditional Language Models without Catastrophic Forgetting 1 Dec 2021 · 2 repositories · arXiv:2112.00791Syntology official (archive's flag): 3 ran · 7 ran (of which 1 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples) · 4 pointer-only (licence)
-
Cross-view Geo-localization with Layer-to-Layer Transformer 1 Dec 2021 · 0 repositories
-
Detecting Extratropical Cyclones of the Northern Hemisphere with Single Shot Detector 1 Dec 2021 · 0 repositories · arXiv:2112.01283
-
Do Transformers Really Perform Badly for Graph Representation? 1 Dec 2021 · 0 repositories
-
Domain-oriented Language Pre-training with Adaptive Hybrid Masking and Optimal Transport Alignment 1 Dec 2021 · 0 repositories · arXiv:2112.03024
-
DRONE: Data-aware Low-rank Compression for Large NLP Models 1 Dec 2021 · 0 repositories
-
Federated Split Task-Agnostic Vision Transformer for COVID-19 CXR Diagnosis 1 Dec 2021 · 0 repositories
-
Focal Attention for Long-Range Interactions in Vision Transformers 1 Dec 2021 · 1 repository