Methods › General › Normalization › Layer Normalization › Papers, page 233
Layer Normalization
Papers archive 2025-07-28
archive papers tagged: 24,980 · with a code link: 11,273 · where Syntology ran a sample: 3,471 (2,923 with a run with no instrument failure, 548 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,471 of 24,980 tagged: 2,923 with a run with no instrument failure, 548 where every run was a failure of Syntology's instrument)
Page 233 of 250: papers 23,201 to 23,300 of 24,980, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
BERT-kNN: Adding a kNN Search Component to Pretrained Language Models for Better QA 2 May 2020 · 1 repository · arXiv:2005.00766
-
Birds have four legs?! NumerSense: Probing Numerical Commonsense Knowledge of Pre-trained Language Models 2 May 2020 · 0 repositories · arXiv:2005.00683
-
Contrastive Self-Supervised Learning for Commonsense Reasoning 2 May 2020 · 3 repositories · arXiv:2005.00669Syntology official (archive's flag): 3 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples) · 1 pointer-only (licence)
-
DeFormer: Decomposing Pre-trained Transformers for Faster Question Answering 2 May 2020 · 1 repository · arXiv:2005.00697Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 6 harvested samples)
-
DagoBERT: Generating Derivational Morphology with a Pretrained Language Model 2 May 2020 · 1 repository · arXiv:2005.00672
-
Hard-Coded Gaussian Attention for Neural Machine Translation 2 May 2020 · 1 repository · arXiv:2005.00742Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
IsoBN: Fine-Tuning BERT with Isotropic Batch Normalization 2 May 2020 · 1 repository · arXiv:2005.02178
-
Is Multihop QA in DiRe Condition? Measuring and Reducing Disconnected Reasoning 2 May 2020 · 1 repository · arXiv:2005.00789Syntology official: harvested, nothing ran · 0 ran · 3 unverified (of 3 harvested samples)
-
Quantifying Attention Flow in Transformers 2 May 2020 · 7 repositories · arXiv:2005.00928Syntology official: harvested, nothing ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 6 unverified (of 8 harvested samples) · 6 pointer-only (licence)
-
Synthesizer: Rethinking Self-Attention in Transformer Models 2 May 2020 · 1 repository · arXiv:2005.00743Syntology 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample)
-
A Controllable Model of Grounded Response Generation 1 May 2020 · 1 repository · arXiv:2005.00613
-
``A Passage to India'': Pre-trained Word Embeddings for Indian Languages 1 May 2020 · 0 repositories
-
A Summarization Dataset of Slovak News Articles 1 May 2020 · 0 repositories
-
A Transformer-based Approach for Source Code Summarization 1 May 2020 · 9 repositories · arXiv:2005.00653
-
Abusive language in Spanish children and young teenager's conversations: data preparation and short text classification with contextual word embeddings 1 May 2020 · 0 repositories
-
Adaptation of Deep Bidirectional Transformers for Afrikaans Language 1 May 2020 · 0 repositories
-
Adapting BERT to Implicit Discourse Relation Classification with a Focus on Discourse Connectives 1 May 2020 · 0 repositories
-
Aggression and Misogyny Detection using BERT: A Multi-Task Approach 1 May 2020 · 1 repository
-
Aggression Identification in English, Hindi and Bangla Text using BERT, RoBERTa and SVM 1 May 2020 · 1 repository
-
Aggression Identification in Social Media: a Transfer Learning Based Approach 1 May 2020 · 0 repositories
-
AIA-BDE: A Corpus of FAQs in Portuguese and their Variations 1 May 2020 · 0 repositories
-
An Evaluation Dataset for Identifying Communicative Functions of Sentences in English Scholarly Papers 1 May 2020 · 0 repositories
-
Analyzing ELMo and DistilBERT on Socio-political News Classification 1 May 2020 · 0 repositories
-
ASU_OPTO at OSACT4 - Offensive Language Detection for Arabic text 1 May 2020 · 0 repositories
-
Automated Essay Scoring System for Nonnative Japanese Learners 1 May 2020 · 0 repositories
-
Bagging BERT Models for Robust Aggression Identification 1 May 2020 · 1 repository
-
Building a Task-oriented Dialog System for Languages with no Training Data: the Case for Basque 1 May 2020 · 0 repositories
-
Chinese Discourse Parsing: Model and Evaluation 1 May 2020 · 0 repositories
-
Clinical Reading Comprehension: A Thorough Analysis of the emrQA Dataset 1 May 2020 · 1 repository · arXiv:2005.00574Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
Contextualized Embeddings based Transformer Encoder for Sentence Similarity Modeling in Answer Selection Task 1 May 2020 · 1 repository
-
Corpora for Document-Level Neural Machine Translation 1 May 2020 · 0 repositories
-
Cross-lingual and Cross-domain Evaluation of Machine Reading Comprehension with Squad and CALOR-Quest Corpora 1 May 2020 · 0 repositories
-
Cross-lingual Zero Pronoun Resolution 1 May 2020 · 0 repositories
-
Cross-Linguistic Syntactic Evaluation of Word Prediction Models 1 May 2020 · 2 repositories · arXiv:2005.00187
-
DaNE: A Named Entity Resource for Danish 1 May 2020 · 0 repositories
-
DecOp: A Multilingual and Multi-domain Corpus For Detecting Deception In Typed Text 1 May 2020 · 0 repositories
-
Detecting Direct Speech in Multilingual Collection of 19th-century Novels 1 May 2020 · 0 repositories
-
Development and Validation of a Corpus for Machine Humor Comprehension 1 May 2020 · 0 repositories
-
Evaluation Metrics for Headline Generation Using Deep Pre-Trained Embeddings 1 May 2020 · 0 repositories
-
Evaluation of Dataset Selection for Pre-Training and Fine-Tuning Transformer Language Models for Clinical Question Answering 1 May 2020 · 0 repositories
-
Event Clustering within News Articles 1 May 2020 · 0 repositories
-
Exploring Transformer Text Generation for Medical Dataset Augmentation 1 May 2020 · 2 repositories
-
FrameNet Annotations Alignment using Attention-based Machine Translation 1 May 2020 · 0 repositories
-
From Web Crawl to Clean Register-Annotated Corpora 1 May 2020 · 0 repositories
-
Global Relational Models of Source Code 1 May 2020 · 1 repository
-
HERO: Hierarchical Encoder for Video+Language Omni-representation Pre-training 1 May 2020 · 3 repositories · arXiv:2005.00200Syntology official (archive's flag): 8 ran · 10 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 2 honoured, 0 violated, 7 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 13 harvested samples) · 8 pointer-only (licence)
-
Discourse-Aware Unsupervised Summarization of Long Scientific Documents 1 May 2020 · 1 repository · arXiv:2005.00513
-
Hitachi at SemEval-2020 Task 12: Offensive Language Identification with Noisy Labels using Statistical Sampling and Post-Processing 1 May 2020 · 0 repositories · arXiv:2005.00295
-
Identifying Necessary Elements for BERT's Multilinguality 1 May 2020 · 1 repository · arXiv:2005.00396
-
Implementation of Supervised Training Approaches for Monolingual Word Sense Alignment: ACDH-CH System Description for the MWSA Shared Task at GlobaLex 2020 1 May 2020 · 0 repositories
-
Improving Neural Language Generation with Spectrum Control 1 May 2020 · 0 repositories
-
Information Seeking in the Spirit of Learning: a Dataset for Conversational Curiosity 1 May 2020 · 1 repository · arXiv:2005.00172
-
Intermediate-Task Transfer Learning with Pretrained Models for Natural Language Understanding: When and Why Does It Work? 1 May 2020 · 0 repositories · arXiv:2005.00628
-
Introducing a Large-Scale Dataset for Vietnamese POS Tagging on Conversational Texts 1 May 2020 · 0 repositories
-
IRIT at TRAC 2020 1 May 2020 · 0 repositories
-
Is Language Modeling Enough? Evaluating Effective Embedding Combinations 1 May 2020 · 0 repositories
-
Joint Learning of Syntactic Features Helps Discourse Segmentation 1 May 2020 · 1 repository
-
KLEJ: Comprehensive Benchmark for Polish Language Understanding 1 May 2020 · 1 repository · arXiv:2005.00630Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Linguistically Informed Hindi-English Neural Machine Translation 1 May 2020 · 0 repositories
-
Logic and the 2-Simplicial Transformer 1 May 2020 · 1 repository
-
Massive vs. Curated Embeddings for Low-Resourced Languages: the Case of Yorùbá and Twi 1 May 2020 · 0 repositories
-
Minority Positive Sampling for Switching Points - an Anecdote for the Code-Mixing Language Modeling 1 May 2020 · 0 repositories
-
Much Ado About Nothing -- Identification of Zero Copulas in Hungarian Using an NMT Model 1 May 2020 · 0 repositories
-
Multi-scale Transformer Language Models 1 May 2020 · 0 repositories · arXiv:2005.00581
-
Multilingual Corpus Creation for Multilingual Semantic Similarity Task 1 May 2020 · 0 repositories
-
Multilingual Joint Fine-tuning of Transformer models for identifying Trolling, Aggression and Cyberbullying at TRAC 2020 1 May 2020 · 1 repository
-
Neural Symbolic Reader: Scalable Integration of Distributed and Symbolic Representations for Reading Comprehension 1 May 2020 · 0 repositories
-
On the Influence of Coreference Resolution on Word Embeddings in Lexical-semantic Evaluation Tasks 1 May 2020 · 0 repositories
-
One Classifier for All Ambiguous Words: Overcoming Data Sparsity by Utilizing Sense Correlations Across Words 1 May 2020 · 0 repositories
-
Paraphrase Generation and Evaluation on Colloquial-Style Sentences 1 May 2020 · 0 repositories
-
ParlVote: A Corpus for Sentiment Analysis of Political Debates 1 May 2020 · 0 repositories
-
Parsing as Tagging 1 May 2020 · 0 repositories
-
POINTER: Constrained Progressive Text Generation via Insertion-based Generative Pre-training 1 May 2020 · 1 repository · arXiv:2005.00558Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 1 honoured, 1 violated, 4 with no contract checked; 5 where Syntology's instrument failed) · 7 unverified (of 18 harvested samples) · 1 pointer-only (licence)
-
Probing Contextual Language Models for Common Ground with Visual Representations 1 May 2020 · 0 repositories · arXiv:2005.00619
-
Region-Based Self-Triggered Control for Perturbed and Uncertain Nonlinear Systems 1 May 2020 · 0 repositories · arXiv:2005.00473
-
Scaling Language Data Import/Export with a Data Transformer Interface 1 May 2020 · 0 repositories
-
SciREX: A Challenge Dataset for Document-Level Information Extraction 1 May 2020 · 1 repository · arXiv:2005.00512Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Scmhl5 at TRAC-2 Shared Task on Aggression Identification: Bert Based Ensemble Learning Approach 1 May 2020 · 0 repositories
-
Seq2SeqPy: A Lightweight and Customizable Toolkit for Neural Sequence-to-Sequence Modeling 1 May 2020 · 0 repositories
-
SiBert: Enhanced Chinese Pre-trained Language Model with Sentence Insertion 1 May 2020 · 1 repository
-
TermEval 2020: TALN-LS2N System for Automatic Term Extraction 1 May 2020 · 0 repositories
-
Text Categorization for Conflict Event Annotation 1 May 2020 · 0 repositories
-
TF-IDF Character N-grams versus Word Embedding-based Models for Fine-grained Event Classification: A Preliminary Study 1 May 2020 · 0 repositories
-
The AVA-Kinetics Localized Human Actions Video Dataset 1 May 2020 · 0 repositories · arXiv:2005.00214
-
Transfer learning applied to text classification in Spanish radiological reports 1 May 2020 · 0 repositories
-
Understanding User Utterances in a Dialog System for Caregiving 1 May 2020 · 0 repositories
-
When BERT Plays the Lottery, All Tickets Are Winning 1 May 2020 · 0 repositories · arXiv:2005.00561
-
An Empirical Study of Pre-trained Transformers for Arabic Information Extraction 30 Apr 2020 · 1 repository · arXiv:2004.14519
-
A Matter of Framing: The Impact of Linguistic Formalism on Probing Results 30 Apr 2020 · 0 repositories · arXiv:2004.14999
-
Accurate Word Alignment Induction from Neural Machine Translation 30 Apr 2020 · 1 repository · arXiv:2004.14837
-
Addressing Zero-Resource Domains Using Document-Level Context in Neural Machine Translation 30 Apr 2020 · 0 repositories · arXiv:2004.14927
-
Breaking (Global) Barriers in Parallel Stochastic Optimization with Wait-Avoiding Group Averaging 30 Apr 2020 · 0 repositories · arXiv:2005.00124
-
Capsule-Transformer for Neural Machine Translation 30 Apr 2020 · 0 repositories · arXiv:2004.14649
-
Character-Level Translation with Self-attention 30 Apr 2020 · 0 repositories · arXiv:2004.14788
-
End-to-End Neural Word Alignment Outperforms GIZA++ 30 Apr 2020 · 0 repositories · arXiv:2004.14675
-
Enriched Pre-trained Transformers for Joint Slot Filling and Intent Detection 30 Apr 2020 · 0 repositories · arXiv:2004.14848
-
Exploring Contextualized Neural Language Models for Temporal Dependency Parsing 30 Apr 2020 · 1 repository · arXiv:2004.14577
-
Few-Shot Learning for Opinion Summarization 30 Apr 2020 · 1 repository · arXiv:2004.14884
-
How do Decisions Emerge across Layers in Neural Models? Interpretation with Differentiable Masking 30 Apr 2020 · 2 repositories · arXiv:2004.14992Syntology official (archive's flag): 4 ran · 6 ran (of which 4 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 10 harvested samples)
-
Interpretable Entity Representations through Large-Scale Typing 30 Apr 2020 · 0 repositories · arXiv:2005.00147