Methods › General › Normalization › Layer Normalization › Papers, page 199
Layer Normalization
Papers archive 2025-07-28
archive papers tagged: 24,980 · with a code link: 11,273 · where Syntology ran a sample: 3,471 (2,923 with a run with no instrument failure, 548 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,471 of 24,980 tagged: 2,923 with a run with no instrument failure, 548 where every run was a failure of Syntology's instrument)
Page 199 of 250: papers 19,801 to 19,900 of 24,980, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Rethinking Why Intermediate-Task Fine-Tuning Works 26 Aug 2021 · 1 repository · arXiv:2108.11696
-
Shifted Chunk Transformer for Spatio-Temporal Representational Learning 26 Aug 2021 · 0 repositories · arXiv:2108.11575
-
SLIM: Explicit Slot-Intent Mapping with BERT for Joint Multi-Intent Detection and Slot Filling 26 Aug 2021 · 1 repository · arXiv:2108.11711
-
The Devil is in the Detail: Simple Tricks Improve Systematic Generalization of Transformers 26 Aug 2021 · 2 repositories · arXiv:2108.12284Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
TPH-YOLOv5: Improved YOLOv5 Based on Transformer Prediction Head for Object Detection on Drone-captured Scenarios 26 Aug 2021 · 3 repositories · arXiv:2108.11539
-
Multilingual Multi-Aspect Explainability Analyses on Machine Reading Comprehension Models 26 Aug 2021 · 1 repository · arXiv:2108.11574
-
CancerBERT: a BERT model for Extracting Breast Cancer Phenotypes from Electronic Health Records 25 Aug 2021 · 0 repositories · arXiv:2108.11303
-
Transformer for Single Image Super-Resolution 25 Aug 2021 · 1 repository · arXiv:2108.11084Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 10 harvested samples) · 4 pointer-only (licence)
-
Models In a Spelling Bee: Language Models Implicitly Learn the Character Composition of Tokens 25 Aug 2021 · 1 repository · arXiv:2108.11193Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
On Approximate Nearest Neighbour Selection for Multi-Stage Dense Retrieval 25 Aug 2021 · 1 repository · arXiv:2108.11480
-
Ontology-Enhanced Slot Filling 25 Aug 2021 · 0 repositories · arXiv:2108.11275
-
What do pre-trained code models know about code? 25 Aug 2021 · 1 repository · arXiv:2108.11308
-
Auto-Parsing Network for Image Captioning and Visual Question Answering 24 Aug 2021 · 0 repositories · arXiv:2108.10568
-
Greenformers: Improving Computation and Memory Efficiency in Transformer Models via Low-Rank Approximation 24 Aug 2021 · 0 repositories · arXiv:2108.10808
-
sigmoidF1: A Smooth F1 Score Surrogate Loss for Multilabel Classification 24 Aug 2021 · 1 repository · arXiv:2108.10566
-
Towards Offensive Language Identification for Tamil Code-Mixed YouTube Comments and Posts 24 Aug 2021 · 1 repository · arXiv:2108.10939
-
Using BERT Encoding and Sentence-Level Language Model for Sentence Ordering 24 Aug 2021 · 0 repositories · arXiv:2108.10986
-
Weakly Supervised Cross-platform Teenager Detection with Adversarial BERT 24 Aug 2021 · 0 repositories · arXiv:2108.10619
-
CGEMs: A Metric Model for Automatic Code Generation using GPT-3 23 Aug 2021 · 0 repositories · arXiv:2108.10168
-
Deploying a BERT-based Query-Title Relevance Classifier in a Production System: a View from the Trenches 23 Aug 2021 · 0 repositories · arXiv:2108.10197
-
Improving 3D Object Detection with Channel-wise Transformer 23 Aug 2021 · 1 repository · arXiv:2108.10723Syntology official (archive's flag): 6 ran · 6 ran (of which 4 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 6 unverified (of 12 harvested samples)
-
One TTS Alignment To Rule Them All 23 Aug 2021 · 3 repositories · arXiv:2108.10447
-
Query Embedding Pruning for Dense Retrieval 23 Aug 2021 · 1 repository · arXiv:2108.10341
-
Recurrent multiple shared layers in Depth for Neural Machine Translation 23 Aug 2021 · 0 repositories · arXiv:2108.10417
-
Regularizing Transformers With Deep Probabilistic Layers 23 Aug 2021 · 0 repositories · arXiv:2108.10764
-
Sarcasm Detection in Twitter -- Performance Impact while using Data Augmentation: Word Embeddings 23 Aug 2021 · 1 repository · arXiv:2108.09924
-
SwinIR: Image Restoration Using Swin Transformer 23 Aug 2021 · 9 repositories · arXiv:2108.10257Syntology official (archive's flag): 3 ran · 30 ran (of which 14 constructed an object rather than computing a result; 16 with no instrument failure: 0 honoured, 0 violated, 16 with no contract checked; 14 where Syntology's instrument failed) · 15 unverified (of 45 harvested samples) · 5 pointer-only (licence)
-
ZS-SLR: Zero-Shot Sign Language Recognition from RGB-D Videos 23 Aug 2021 · 0 repositories · arXiv:2108.10059
-
Guiding Query Position and Performing Similar Attention for Transformer-Based Detection Heads 22 Aug 2021 · 0 repositories · arXiv:2108.09691
-
Spatial Transformer Networks for Curriculum Learning 22 Aug 2021 · 0 repositories · arXiv:2108.09696
-
StarVQA: Space-Time Attention for Video Quality Assessment 22 Aug 2021 · 0 repositories · arXiv:2108.09635
-
Using Large Pre-Trained Models with Cross-Modal Attention for Multi-Modal Emotion Recognition 22 Aug 2021 · 0 repositories · arXiv:2108.09669
-
UzBERT: pretraining a BERT model for Uzbek 22 Aug 2021 · 0 repositories · arXiv:2108.09814
-
Construction material classification on imbalanced datasets using Vision Transformer (ViT) architecture 21 Aug 2021 · 0 repositories · arXiv:2108.09527
-
Approximate Bayesian Neural Doppler Imaging 20 Aug 2021 · 1 repository · arXiv:2108.09266
-
Convolutional Neural Network (CNN) vs Vision Transformer (ViT) for Digital Holography 20 Aug 2021 · 0 repositories · arXiv:2108.09147
-
Extracting Radiological Findings With Normalized Anatomical Information Using a Span-Based BERT Relation Extraction Model 20 Aug 2021 · 0 repositories · arXiv:2108.09211
-
Fastformer: Additive Attention Can Be All You Need 20 Aug 2021 · 13 repositories · arXiv:2108.09084Syntology official: no sample here; runs from other or unrecorded repositories · 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 2 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 3 pointer-only (licence)
-
Frozen Pretrained Transformers for Neural Sign Language Translation 20 Aug 2021 · 1 repository
-
MM-ViT: Multi-Modal Video Transformer for Compressed Video Action Recognition 20 Aug 2021 · 0 repositories · arXiv:2108.09322
-
One Chatbot Per Person: Creating Personalized Chatbots based on Implicit User Profiles 20 Aug 2021 · 1 repository · arXiv:2108.09355
-
Pre-training for Ad-hoc Retrieval: Hyperlink is Also You Need 20 Aug 2021 · 1 repository · arXiv:2108.09346
-
Semantic Communication with Adaptive Universal Transformer 20 Aug 2021 · 0 repositories · arXiv:2108.09119
-
Smart Bird: Learnable Sparse Attention for Efficient and Effective Transformer 20 Aug 2021 · 0 repositories · arXiv:2108.09193
-
Trans4Trans: Efficient Transformer for Transparent Object and Semantic Scene Segmentation in Real-World Navigation Assistance 20 Aug 2021 · 1 repository · arXiv:2108.09174
-
Uncertainties and output feedback in rollout event-triggered control 20 Aug 2021 · 0 repositories · arXiv:2108.09125
-
A Framework for Neural Topic Modeling of Text Corpora 19 Aug 2021 · 1 repository · arXiv:2108.08946
-
Causal Attention for Unbiased Visual Recognition 19 Aug 2021 · 1 repository · arXiv:2108.08782Syntology official (archive's flag): 2 ran · 2 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; every one of the 2 samples that ran constructed an object rather than computing a result (of 2 harvested samples) · 2 pointer-only (licence)
-
Contrastive Language-Image Pre-training for the Italian Language 19 Aug 2021 · 1 repository · arXiv:2108.08688
-
Detection of Illicit Drug Trafficking Events on Instagram: A Deep Multimodal Multilabel Learning Approach 19 Aug 2021 · 0 repositories · arXiv:2108.08920
-
Do Vision Transformers See Like Convolutional Neural Networks? 19 Aug 2021 · 4 repositories · arXiv:2108.08810Syntology 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Fast Passage Re-ranking with Contextualized Exact Term Matching and Efficient Passage Expansion 19 Aug 2021 · 1 repository · arXiv:2108.08513
-
Fine-Grained Element Identification in Complaint Text of Internet Fraud 19 Aug 2021 · 0 repositories · arXiv:2108.08676
-
How Hateful are Movies? A Study and Prediction on Movie Subtitles 19 Aug 2021 · 1 repository · arXiv:2108.10724
-
MvSR-NAT: Multi-view Subset Regularization for Non-Autoregressive Machine Translation 19 Aug 2021 · 0 repositories · arXiv:2108.08447
-
Sentence-T5: Scalable Sentence Encoders from Pre-trained Text-to-Text Models 19 Aug 2021 · 2 repositories · arXiv:2108.08877
-
UNIQORN: Unified Question Answering over RDF Knowledge Graphs and Natural Language Text 19 Aug 2021 · 1 repository · arXiv:2108.08614
-
Video Relation Detection via Tracklet based Visual Transformer 19 Aug 2021 · 1 repository · arXiv:2108.08669
-
Contributions of Transformer Attention Heads in Multi- and Cross-lingual Tasks 18 Aug 2021 · 0 repositories · arXiv:2108.08375
-
Integrating Dialog History into End-to-End Spoken Language Understanding Systems 18 Aug 2021 · 0 repositories · arXiv:2108.08405
-
SHAQ: Single Headed Attention with Quasi-Recurrence 18 Aug 2021 · 0 repositories · arXiv:2108.08207
-
SIFN: A Sentiment-aware Interactive Fusion Network for Review-based Item Recommendation 18 Aug 2021 · 0 repositories · arXiv:2108.08022
-
Table Caption Generation in Scholarly Documents Leveraging Pre-trained Language Models 18 Aug 2021 · 1 repository · arXiv:2108.08111
-
Transformers predicting the future. Applying attention in next-frame and time series forecasting 18 Aug 2021 · 1 repository · arXiv:2108.08224
-
TSI: an Ad Text Strength Indicator using Text-to-CTR and Semantic-Ad-Similarity 18 Aug 2021 · 0 repositories · arXiv:2108.08226
-
A Multi-level Acoustic Feature Extraction Framework for Transformer Based End-to-End Speech Recognition 18 Aug 2021 · 0 repositories · arXiv:2108.07980
-
Boosting Salient Object Detection with Transformer-based Asymmetric Bilateral U-Net 17 Aug 2021 · 1 repository · arXiv:2108.07851
-
EncT5: Fine-tuning T5 Encoder for Discriminative Tasks 17 Aug 2021 · 0 repositories
-
Learning C to x86 Translation: An Experiment in Neural Compilation 17 Aug 2021 · 1 repository · arXiv:2108.07639Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
Light Field Image Super-Resolution with Transformers 17 Aug 2021 · 1 repository · arXiv:2108.07597
-
Modulating Language Models with Emotions 17 Aug 2021 · 0 repositories · arXiv:2108.07886
-
MOI-Mixer: Improving MLP-Mixer with Multi Order Interactions in Sequential Recommendation 17 Aug 2021 · 0 repositories · arXiv:2108.07505
-
Response Ranking with Multi-types of Deep Interactive Representations in Retrieval-based Dialogues 17 Aug 2021 · 1 repository
-
Semantics-aware Attention Improves Neural Machine Translation 17 Aug 2021 · 0 repositories
-
Accounting for shared covariates in semi-parametric Bayesian additive regression trees 17 Aug 2021 · 2 repositories · arXiv:2108.07636
-
An Effective Non-Autoregressive Model for Spoken Language Understanding 16 Aug 2021 · 0 repositories · arXiv:2108.07005
-
Deep Natural Language Processing for LinkedIn Search 16 Aug 2021 · 0 repositories · arXiv:2108.13300
-
Misleading the Covid-19 vaccination discourse on Twitter: An exploratory study of infodemic around the pandemic 16 Aug 2021 · 1 repository · arXiv:2108.10735
-
No-Reference Image Quality Assessment via Transformers, Relative Ranking, and Self-Consistency 16 Aug 2021 · 1 repository · arXiv:2108.06858
-
On the Opportunities and Risks of Foundation Models 16 Aug 2021 · 2 repositories · arXiv:2108.07258
-
Scene Designer: a Unified Model for Scene Search and Synthesis from Sketch 16 Aug 2021 · 1 repository · arXiv:2108.07353
-
Exploring Generalization Ability of Pretrained Language Models on Arithmetic and Logical Reasoning 15 Aug 2021 · 0 repositories · arXiv:2108.06743
-
Exploring Temporal Coherence for More General Video Face Forgery Detection 15 Aug 2021 · 1 repository · arXiv:2108.06693Syntology 6 ran (of which 4 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Maps Search Misspelling Detection Leveraging Domain-Augmented Contextual Representations 15 Aug 2021 · 0 repositories · arXiv:2108.06842
-
SAPPHIRE: Approaches for Enhanced Concept-to-Text Generation 15 Aug 2021 · 0 repositories · arXiv:2108.06643
-
SOTR: Segmenting Objects with Transformers 15 Aug 2021 · 1 repository · arXiv:2108.06747Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 2 violated, 4 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples) · 2 pointer-only (licence)
-
heterogeneous temporal graph transformer: an intelligent system for evolving android malware detection 14 Aug 2021 · 1 repository
-
Conditional DETR for Fast Training Convergence 13 Aug 2021 · 4 repositories · arXiv:2108.06152Syntology official (archive's flag): 5 ran · 6 ran (of which 1 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 1 violated, 2 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 9 harvested samples) · 4 pointer-only (licence)
-
The Stability-Efficiency Dilemma: Investigating Sequence Length Warmup for Training GPT Models 13 Aug 2021 · 1 repository · arXiv:2108.06084
-
PVT: Point-Voxel Transformer for Point Cloud Learning 13 Aug 2021 · 2 repositories · arXiv:2108.06076
-
Towards Structured Dynamic Sparse Pre-Training of BERT 13 Aug 2021 · 0 repositories · arXiv:2108.06277
-
AMMUS : A Survey of Transformer-based Pretrained Models in Natural Language Processing 12 Aug 2021 · 1 repository · arXiv:2108.05542
-
How Optimal is Greedy Decoding for Extractive Question Answering? 12 Aug 2021 · 1 repository · arXiv:2108.05857Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples)
-
Mobile-Former: Bridging MobileNet and Transformer 12 Aug 2021 · 4 repositories · arXiv:2108.05895
-
Modeling Relevance Ranking under the Pre-training and Fine-tuning Paradigm 12 Aug 2021 · 0 repositories · arXiv:2108.05652
-
Multimodal analysis of the predictability of hand-gesture properties 12 Aug 2021 · 0 repositories · arXiv:2108.05762
-
MUSIQ: Multi-scale Image Quality Transformer 12 Aug 2021 · 2 repositories · arXiv:2108.05997Syntology community repositories only · 5 ran (of which 1 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Overview of the HASOC track at FIRE 2020: Hate Speech and Offensive Content Identification in Indo-European Languages 12 Aug 2021 · 0 repositories · arXiv:2108.05927
-
PatrickStar: Parallel Training of Pre-trained Models via Chunk-based Memory Management 12 Aug 2021 · 1 repository · arXiv:2108.05818
-
TVT: Transferable Vision Transformer for Unsupervised Domain Adaptation 12 Aug 2021 · 1 repository · arXiv:2108.05988Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)