Methods › General › Normalization › Layer Normalization › Papers, page 194
Layer Normalization
Papers archive 2025-07-28
archive papers tagged: 24,980 · with a code link: 11,273 · where Syntology ran a sample: 3,471 (2,923 with a run with no instrument failure, 548 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,471 of 24,980 tagged: 2,923 with a run with no instrument failure, 548 where every run was a failure of Syntology's instrument)
Page 194 of 250: papers 19,301 to 19,400 of 24,980, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
The Layout Generation Algorithm of Graphic Design Based on Transformer-CVAE 8 Oct 2021 · 0 repositories · arXiv:2110.06794
-
Adversarial Token Attacks on Vision Transformers 8 Oct 2021 · 0 repositories · arXiv:2110.04337
-
ALL-IN-ONE: Multi-Task Learning BERT models for Evaluating Peer Assessments 8 Oct 2021 · 0 repositories · arXiv:2110.03895
-
Context-LGM: Leveraging Object-Context Relation for Context-Aware Object Recognition 8 Oct 2021 · 0 repositories · arXiv:2110.04042
-
Towards Learning (Dis)-Similarity of Source Code from Program Contrasts 8 Oct 2021 · 0 repositories · arXiv:2110.03868
-
Cross-speaker Emotion Transfer Based on Speaker Condition Layer Normalization and Semi-Supervised Training in Text-To-Speech 8 Oct 2021 · 1 repository · arXiv:2110.04153
-
Development of an Extractive Title Generation System Using Titles of Papers of Top Conferences for Intermediate English Students 8 Oct 2021 · 0 repositories · arXiv:2110.04204
-
HydraSum: Disentangling Stylistic Features in Text Summarization using Multi-Decoder Models 8 Oct 2021 · 1 repository · arXiv:2110.04400Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified (of 9 harvested samples)
-
KG-FiD: Infusing Knowledge Graph in Fusion-in-Decoder for Open-Domain Question Answering 8 Oct 2021 · 0 repositories · arXiv:2110.04330
-
Knowledge-Enhanced Hierarchical Graph Transformer Network for Multi-Behavior Recommendation 8 Oct 2021 · 1 repository · arXiv:2110.04000
-
Learning Adaptive Control Flow in Transformers for Improved Systematic Generalization 8 Oct 2021 · 0 repositories
-
Local and Global Context-Based Pairwise Models for Sentence Ordering 8 Oct 2021 · 1 repository · arXiv:2110.04291
-
M6-10T: A Sharing-Delinking Paradigm for Efficient Multi-Trillion Parameter Pretraining 8 Oct 2021 · 0 repositories · arXiv:2110.03888
-
Multiplex Behavioral Relation Learning for Recommendation via Memory Augmented Transformer Network 8 Oct 2021 · 1 repository · arXiv:2110.04002
-
RPT: Toward Transferable Model on Heterogeneous Researcher Data via Pre-Training 8 Oct 2021 · 1 repository · arXiv:2110.07336
-
Speeding up Deep Model Training by Sharing Weights and Then Unsharing 8 Oct 2021 · 0 repositories · arXiv:2110.03848
-
Taming Sparsely Activated Transformer with Stochastic Experts 8 Oct 2021 · 1 repository · arXiv:2110.04260Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Text analysis and deep learning: A network approach 8 Oct 2021 · 0 repositories · arXiv:2110.04151
-
ViDT: An Efficient and Effective Fully Transformer-based Object Detector 8 Oct 2021 · 1 repository · arXiv:2110.03921
-
A Comparative Study of Transformer-Based Language Models on Extractive Question Answering 7 Oct 2021 · 0 repositories · arXiv:2110.03142
-
Attention is All You Need? Good Embeddings with Statistics are enough:Large Scale Audio Understanding without Transformers/ Convolutions/ BERTs/ Mixers/ Attention/ RNNs or .... 7 Oct 2021 · 0 repositories · arXiv:2110.03183
-
Cross-Language Learning for Entity Matching 7 Oct 2021 · 1 repository · arXiv:2110.03338
-
End-to-End Supermask Pruning: Learning to Prune Image Captioning Models 7 Oct 2021 · 1 repository · arXiv:2110.03298Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 1 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples) · 1 pointer-only (licence)
-
Generative Pre-Trained Transformer for Cardiac Abnormality Detection 7 Oct 2021 · 0 repositories · arXiv:2110.04071
-
Investigating Growth at Risk Using a Multi-country Non-parametric Quantile Factor Model 7 Oct 2021 · 0 repositories · arXiv:2110.03411
-
Layer-wise Pruning of Transformer Attention Heads for Efficient Language Modeling 7 Oct 2021 · 1 repository · arXiv:2110.03252
-
Minimum word error training for non-autoregressive Transformer-based code-switching ASR 7 Oct 2021 · 0 repositories · arXiv:2110.03573
-
Mixer-TTS: non-autoregressive, fast and compact text-to-speech model conditioned on language model embeddings 7 Oct 2021 · 1 repository · arXiv:2110.03584
-
Universality of Winning Tickets: A Renormalization Group Perspective 7 Oct 2021 · 0 repositories · arXiv:2110.03210
-
8-bit Optimizers via Block-wise Quantization 6 Oct 2021 · 3 repositories · arXiv:2110.02861
-
Adversarial Robustness Comparison of Vision Transformer and MLP-Mixer to CNNs 6 Oct 2021 · 1 repository · arXiv:2110.02797
-
Anomaly Transformer: Time Series Anomaly Detection with Association Discrepancy 6 Oct 2021 · 3 repositories · arXiv:2110.02642Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
Dynamically Decoding Source Domain Knowledge for Domain Generalization 6 Oct 2021 · 0 repositories · arXiv:2110.03027
-
Geometric Transformers for Protein Interface Contact Prediction 6 Oct 2021 · 2 repositories · arXiv:2110.02423
-
How BPE Affects Memorization in Transformers 6 Oct 2021 · 0 repositories · arXiv:2110.02782
-
Learning to Iteratively Solve Routing Problems with Dual-Aspect Collaborative Transformer 6 Oct 2021 · 2 repositories · arXiv:2110.02544Syntology official (archive's flag): 7 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 7 unverified (of 16 harvested samples)
-
LIDSNet: A Lightweight on-device Intent Detection model using Deep Siamese Network 6 Oct 2021 · 0 repositories · arXiv:2110.15717
-
NUS-IDS at FinCausal 2021: Dependency Tree in Graph Neural Network for Better Cause-Effect Span Detection 6 Oct 2021 · 1 repository · arXiv:2110.02991
-
On Neurons Invariant to Sentence Structural Changes in Neural Machine Translation 6 Oct 2021 · 1 repository · arXiv:2110.03067
-
PoNet: Pooling Network for Efficient Token Mixing in Long Sequences 6 Oct 2021 · 1 repository · arXiv:2110.02442Syntology official: no sample here; runs from other or unrecorded repositories · 11 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
Pretrained Transformers for Offensive Language Identification in Tanglish 6 Oct 2021 · 1 repository · arXiv:2110.02852
-
Semantic Prediction: Which One Should Come First, Recognition or Prediction? 6 Oct 2021 · 1 repository · arXiv:2110.02829
-
Analyzing the Impact of COVID-19 on Economy from the Perspective of Users Reviews 5 Oct 2021 · 0 repositories · arXiv:2110.02198
-
ASR Rescoring and Confidence Estimation with ELECTRA 5 Oct 2021 · 0 repositories · arXiv:2110.01857
-
BERT Attends the Conversation: Improving Low-Resource Conversational ASR 5 Oct 2021 · 1 repository · arXiv:2110.02267
-
DistilHuBERT: Speech Representation Learning by Layer-wise Distillation of Hidden-unit BERT 5 Oct 2021 · 1 repository · arXiv:2110.01900
-
Exploiting Twitter as Source of Large Corpora of Weakly Similar Pairs for Semantic Sentence Embeddings 5 Oct 2021 · 1 repository · arXiv:2110.02030
-
FoodChem: A food-chemical relation extraction model 5 Oct 2021 · 1 repository · arXiv:2110.02019
-
Learning Sense-Specific Static Embeddings using Contextualised Word Embeddings as a Proxy 5 Oct 2021 · 0 repositories · arXiv:2110.02204
-
Leveraging the Inductive Bias of Large Language Models for Abstract Textual Reasoning 5 Oct 2021 · 0 repositories · arXiv:2110.02370
-
MobileViT: Light-weight, General-purpose, and Mobile-friendly Vision Transformer 5 Oct 2021 · 31 repositories · arXiv:2110.02178Syntology official: no sample here; runs from other or unrecorded repositories · 53 ran (of which 31 constructed an object rather than computing a result; 44 with no instrument failure: 1 honoured, 0 violated, 43 with no contract checked; 9 where Syntology's instrument failed) · 15 unverified (of 68 harvested samples) · 18 pointer-only (licence)
-
Sicilian Translator: A Recipe for Low-Resource NMT 5 Oct 2021 · 1 repository · arXiv:2110.01938
-
Sound Event Detection Transformer: An Event-based End-to-End Model for Sound Event Detection 5 Oct 2021 · 1 repository · arXiv:2110.02011
-
Top-N: Equivariant set and graph generation without exchangeability 5 Oct 2021 · 1 repository · arXiv:2110.02096Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
ur-iw-hnt at GermEval 2021: An Ensembling Strategy with Multiple BERT Models 5 Oct 2021 · 0 repositories · arXiv:2110.02042
-
Word Acquisition in Neural Language Models 5 Oct 2021 · 1 repository · arXiv:2110.02406
-
Molformer: Motif-based Transformer on 3D Heterogeneous Molecular Graphs 4 Oct 2021 · 2 repositories · arXiv:2110.01191Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
A free lunch from ViT:Adaptive Attention Multi-scale Fusion Transformer for Fine-grained Visual Recognition 4 Oct 2021 · 0 repositories · arXiv:2110.01240
-
DeepA2: A Modular Framework for Deep Argument Analysis with Pretrained Neural Text2Text Language Models 4 Oct 2021 · 1 repository · arXiv:2110.01509
-
Exploiting Pre-Trained ASR Models for Alzheimer's Disease Recognition Through Spontaneous Speech 4 Oct 2021 · 0 repositories · arXiv:2110.01493
-
JuriBERT: A Masked-Language Model Adaptation for French Legal Text 4 Oct 2021 · 1 repository · arXiv:2110.01485
-
Perhaps PTLMs Should Go to School -- A Task to Assess Open Book and Closed Book QA 4 Oct 2021 · 0 repositories · arXiv:2110.01552
-
VTAMIQ: Transformers for Attention Modulated Image Quality Assessment 4 Oct 2021 · 1 repository · arXiv:2110.01655
-
Adversarial Examples Generation for Reducing Implicit Gender Bias in Pre-trained Models 3 Oct 2021 · 0 repositories · arXiv:2110.01094
-
Music Playlist Title Generation: A Machine-Translation Approach 3 Oct 2021 · 0 repositories · arXiv:2110.07354
-
Unsupervised paradigm for information extraction from transcripts using BERT 3 Oct 2021 · 0 repositories · arXiv:2110.00949
-
Artificial intelligence for Sustainable Energy: A Contextual Topic Modeling and Content Analysis 2 Oct 2021 · 0 repositories · arXiv:2110.00828
-
Implicit and Explicit Attention for Zero-Shot Learning 2 Oct 2021 · 1 repository · arXiv:2110.00860
-
ProTo: Program-Guided Transformer for Program-Guided Tasks 2 Oct 2021 · 1 repository · arXiv:2110.00804Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Swiss-Judgment-Prediction: A Multilingual Legal Judgment Prediction Benchmark 2 Oct 2021 · 1 repository · arXiv:2110.00806
-
BERT4GCN: Using BERT Intermediate Layers to Augment GCN for Aspect-based Sentiment Classification 1 Oct 2021 · 0 repositories · arXiv:2110.00171
-
3rd Place Scheme on Instance Segmentation Track of ICCV 2021 VIPriors Challenges 1 Oct 2021 · 0 repositories · arXiv:2110.00242
-
Geometry Attention Transformer with Position-aware LSTMs for Image Captioning 1 Oct 2021 · 1 repository · arXiv:2110.00335
-
Improving Punctuation Restoration for Speech Transcripts via External Data 1 Oct 2021 · 0 repositories · arXiv:2110.00560
-
Low Frequency Names Exhibit Bias and Overfitting in Contextualizing Language Models 1 Oct 2021 · 0 repositories · arXiv:2110.00672
-
Span Labeling Approach for Vietnamese and Chinese Word Segmentation 1 Oct 2021 · 0 repositories · arXiv:2110.00156
-
Unpacking the Interdependent Systems of Discrimination: Ableist Bias in NLP Systems through an Intersectional Lens 1 Oct 2021 · 0 repositories · arXiv:2110.00521
-
BERT got a Date: Introducing Transformers to Temporal Tagging 30 Sep 2021 · 1 repository · arXiv:2109.14927
-
Bitcoin Transaction Strategy Construction Based on Deep Reinforcement Learning 30 Sep 2021 · 0 repositories · arXiv:2109.14789
-
COVID-19 Fake News Detection Using Bidirectional Encoder Representations from Transformers Based Models 30 Sep 2021 · 1 repository · arXiv:2109.14816
-
GT U-Net: A U-Net Like Group Transformer Network for Tooth Root Segmentation 30 Sep 2021 · 1 repository · arXiv:2109.14813
-
Inducing Transformer's Compositional Generalization Ability via Auxiliary Sequence Prediction Tasks 30 Sep 2021 · 1 repository · arXiv:2109.15256
-
Learning to Predict Trustworthiness with Steep Slope Loss 30 Sep 2021 · 1 repository · arXiv:2110.00054Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample)
-
MobTCast: Leveraging Auxiliary Trajectory Forecasting for Human Mobility Prediction 30 Sep 2021 · 0 repositories · arXiv:2110.01401
-
PortaSpeech: Portable and High-Quality Generative Text-to-Speech 30 Sep 2021 · 4 repositories · arXiv:2109.15166Syntology official (archive's flag): 4 ran · 12 ran (of which 1 constructed an object rather than computing a result; 9 with no instrument failure: 3 honoured, 0 violated, 6 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 12 harvested samples) · 6 pointer-only (licence)
-
Prose2Poem: The Blessing of Transformers in Translating Prose to Persian Poetry 30 Sep 2021 · 1 repository · arXiv:2109.14934
-
Redesigning the Transformer Architecture with Insights from Multi-particle Dynamical Systems 30 Sep 2021 · 1 repository · arXiv:2109.15142
-
PubTables-1M: Towards comprehensive table extraction from unstructured documents 30 Sep 2021 · 2 repositories · arXiv:2110.00061Syntology official (archive's flag): 3 ran · 11 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 3 honoured, 0 violated, 6 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 13 harvested samples) · 10 pointer-only (licence)
-
Structural Persistence in Language Models: Priming as a Window into Abstract Language Representations 30 Sep 2021 · 1 repository · arXiv:2109.14989
-
A Collaborative Attention Adaptive Network for Financial Market Forecasting 29 Sep 2021 · 0 repositories
-
A Dot Product Attention Free Transformer 29 Sep 2021 · 0 repositories
-
Adaptive Control Flow in Transformers Improves Systematic Generalization 29 Sep 2021 · 0 repositories
-
Adaptive Wavelet Transformer Network for 3D Shape Representation Learning 29 Sep 2021 · 0 repositories
-
An Investigation on Hardware-Aware Vision Transformer Scaling 29 Sep 2021 · 0 repositories
-
An object-centric sensitivity analysis of deep learning based instance segmentation 29 Sep 2021 · 0 repositories
-
Analyzing the Implicit Position Encoding Ability of Transformer Decoder 29 Sep 2021 · 0 repositories
-
Are BERT Families Zero-Shot Learners? A Study on Their Potential and Limitations 29 Sep 2021 · 0 repositories
-
Are Vision Transformers Robust to Patch-wise Perturbations? 29 Sep 2021 · 0 repositories
-
Associated Learning: an Alternative to End-to-End Backpropagation that Works on CNN, RNN, and Transformer 29 Sep 2021 · 0 repositories
-
Attention-based Interpretability with Concept Transformers 29 Sep 2021 · 0 repositories