Methods › General › Normalization › Layer Normalization › Papers, page 226
Layer Normalization
Papers archive 2025-07-28
archive papers tagged: 24,980 · with a code link: 11,273 · where Syntology ran a sample: 3,471 (2,923 with a run with no instrument failure, 548 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,471 of 24,980 tagged: 2,923 with a run with no instrument failure, 548 where every run was a failure of Syntology's instrument)
Page 226 of 250: papers 22,501 to 22,600 of 24,980, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Will it Unblend? 18 Sep 2020 · 1 repository · arXiv:2009.09123
-
A Multimodal Memes Classification: A Survey and Open Research Issues 17 Sep 2020 · 0 repositories · arXiv:2009.08395
-
Compositional and Lexical Semantics in RoBERTa, BERT and DistilBERT: A Case Study on CoQA 17 Sep 2020 · 0 repositories · arXiv:2009.08257
-
Cross-Modal Alignment with Mixture Experts Neural Network for Intral-City Retail Recommendation 17 Sep 2020 · 0 repositories · arXiv:2009.09926
-
Distilled One-Shot Federated Learning 17 Sep 2020 · 1 repository · arXiv:2009.07999
-
DSC IIT-ISM at SemEval-2020 Task 6: Boosting BERT with Dependencies for Definition Extraction 17 Sep 2020 · 1 repository · arXiv:2009.08180
-
Efficient Transformer-based Large Scale Language Representations using Hardware-friendly Block Structured Pruning 17 Sep 2020 · 0 repositories · arXiv:2009.08065
-
GraphCodeBERT: Pre-training Code Representations with Data Flow 17 Sep 2020 · 1 repository · arXiv:2009.08366
-
Multi²OIE: Multilingual Open Information Extraction Based on Multi-Head Attention with BERT 17 Sep 2020 · 1 repository · arXiv:2009.08128Syntology official (archive's flag): 6 ran · 6 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 4 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples)
-
Towards Fully 8-bit Integer Inference for the Transformer Model 17 Sep 2020 · 0 repositories · arXiv:2009.08034
-
Automated Source Code Generation and Auto-completion Using Deep Learning: Comparing and Discussing Current Language-Model-Related Approaches 16 Sep 2020 · 1 repository · arXiv:2009.07740
-
CogTree: Cognition Tree Loss for Unbiased Scene Graph Generation 16 Sep 2020 · 1 repository · arXiv:2009.07526
-
Deep Learning Approaches for Extracting Adverse Events and Indications of Dietary Supplements from Clinical Text 16 Sep 2020 · 0 repositories · arXiv:2009.07780
-
Document-level Neural Machine Translation with Document Embeddings 16 Sep 2020 · 0 repositories · arXiv:2009.08775
-
Extremely Low Bit Transformer Quantization for On-Device Neural Machine Translation 16 Sep 2020 · 0 repositories · arXiv:2009.07453
-
Graph-to-Sequence Neural Machine Translation 16 Sep 2020 · 0 repositories · arXiv:2009.07489
-
NABU - Multilingual Graph-based Neural RDF Verbalizer 16 Sep 2020 · 0 repositories · arXiv:2009.07728
-
Retrofitting Structure-aware Transformer Language Model for End Tasks 16 Sep 2020 · 0 repositories · arXiv:2009.07408
-
Simplified TinyBERT: Knowledge Distillation for Document Retrieval 16 Sep 2020 · 4 repositories · arXiv:2009.07531
-
Solomon at SemEval-2020 Task 11: Ensemble Architecture for Fine-Tuned Propaganda Detection in News Articles 16 Sep 2020 · 0 repositories · arXiv:2009.07473
-
UNION: An Unreferenced Metric for Evaluating Open-ended Story Generation 16 Sep 2020 · 1 repository · arXiv:2009.07602Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Real-Time Execution of Large-scale Language Models on Mobile 15 Sep 2020 · 0 repositories · arXiv:2009.06823
-
Global-aware Beam Search for Neural Abstractive Summarization 15 Sep 2020 · 2 repositories · arXiv:2009.06891
-
Augmented Natural Language for Generative Sequence Labeling 15 Sep 2020 · 0 repositories · arXiv:2009.13272
-
BERT-QE: Contextualized Query Expansion for Document Re-ranking 15 Sep 2020 · 1 repository · arXiv:2009.07258Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 9 unverified (of 11 harvested samples)
-
Critical Thinking for Language Models 15 Sep 2020 · 1 repository · arXiv:2009.07185Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
DeNERT-KG: Named Entity and Relation Extraction Model Using DQN, Knowledge Graph, and BERT 15 Sep 2020 · 0 repositories
-
Dialogue Response Ranking Training with Large-Scale Human Feedback Data 15 Sep 2020 · 2 repositories · arXiv:2009.06978Syntology 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples)
-
Event Presence Prediction Helps Trigger Detection Across Languages 15 Sep 2020 · 0 repositories · arXiv:2009.07188
-
It's Not Just Size That Matters: Small Language Models Are Also Few-Shot Learners 15 Sep 2020 · 5 repositories · arXiv:2009.07118
-
Lessons Learned from Applying off-the-shelf BERT: There is no Silver Bullet 15 Sep 2020 · 0 repositories · arXiv:2009.07238
-
MLMLM: Link Prediction with Mean Likelihood Masked Language Model 15 Sep 2020 · 0 repositories · arXiv:2009.07058
-
The Radicalization Risks of GPT-3 and Advanced Neural Language Models 15 Sep 2020 · 0 repositories · arXiv:2009.06807
-
Beyond Accuracy: ROI-driven Data Analytics of Empirical Data 14 Sep 2020 · 0 repositories · arXiv:2009.06492
-
On Robustness and Bias Analysis of BERT-based Relation Extraction 14 Sep 2020 · 1 repository · arXiv:2009.06206
-
Efficient Transformers: A Survey 14 Sep 2020 · 0 repositories · arXiv:2009.06732
-
Filling the Gap of Utterance-aware and Speaker-aware Representation for Multi-turn Dialogue 14 Sep 2020 · 1 repository · arXiv:2009.06504Syntology 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
GeDi: Generative Discriminator Guided Sequence Generation 14 Sep 2020 · 3 repositories · arXiv:2009.06367Syntology official (archive's flag): 4 ran · 6 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 0 violated, 2 with no contract checked; 3 where Syntology's instrument failed) · 5 unverified (of 11 harvested samples)
-
BoostingBERT:Integrating Multi-Class Boosting into BERT for NLP Tasks 13 Sep 2020 · 0 repositories · arXiv:2009.05959
-
Cluster-Former: Clustering-based Sparse Transformer for Long-Range Dependency Encoding 13 Sep 2020 · 0 repositories · arXiv:2009.06097
-
CIA_NITT at WNUT-2020 Task 2: Classification of COVID-19 Tweets Using Pre-trained Language Models 12 Sep 2020 · 0 repositories · arXiv:2009.05782
-
Country Image in COVID-19 Pandemic: A Case Study of China 12 Sep 2020 · 1 repository · arXiv:2009.05817
-
Fine-tuning Pre-trained Contextual Embeddings for Citation Content Analysis in Scholarly Publication 12 Sep 2020 · 0 repositories · arXiv:2009.05836
-
A Comparison of LSTM and BERT for Small Corpus 11 Sep 2020 · 0 repositories · arXiv:2009.05451
-
Compressed Deep Networks: Goodbye SVD, Hello Robust Low-Rank Approximation 11 Sep 2020 · 1 repository · arXiv:2009.05647
-
GTEA: Inductive Representation Learning on Temporal Interaction Graphs via Temporal Edge Aggregation 11 Sep 2020 · 2 repositories · arXiv:2009.05266
-
Unit Test Case Generation with Transformers and Focal Context 11 Sep 2020 · 1 repository · arXiv:2009.05617
-
UPB at SemEval-2020 Task 11: Propaganda Detection with Domain-Specific Trained BERT 11 Sep 2020 · 0 repositories · arXiv:2009.05289
-
UPB at SemEval-2020 Task 6: Pretrained Language Models for Definition Extraction 11 Sep 2020 · 3 repositories · arXiv:2009.05603
-
Brain2Word: Decoding Brain Activity for Language Generation 10 Sep 2020 · 1 repository · arXiv:2009.04765
-
Do Response Selection Models Really Know What's Next? Utterance Manipulation Strategies for Multi-turn Response Selection 10 Sep 2020 · 1 repository · arXiv:2009.04703
-
FILTER: An Enhanced Fusion Method for Cross-lingual Language Understanding 10 Sep 2020 · 1 repository · arXiv:2009.05166Syntology 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 1 pointer-only (licence)
-
Investigating Gender Bias in BERT 10 Sep 2020 · 0 repositories · arXiv:2009.05021
-
Learning Universal Representations from Word to Sentence 10 Sep 2020 · 0 repositories · arXiv:2009.04656
-
Modern Methods for Text Generation 10 Sep 2020 · 2 repositories · arXiv:2009.04968
-
Rank over Class: The Untapped Potential of Ranking in Natural Language Processing 10 Sep 2020 · 1 repository · arXiv:2009.05160Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
Sparsifying Transformer Models with Trainable Representation Pooling 10 Sep 2020 · 1 repository · arXiv:2009.05169
-
Comparative Study of Language Models on Cross-Domain Data with Model Agnostic Explainability 9 Sep 2020 · 0 repositories · arXiv:2009.04095
-
Pay Attention when Required 9 Sep 2020 · 2 repositories · arXiv:2009.04534
-
ERNIE at SemEval-2020 Task 10: Learning Word Emphasis Selection by Pre-trained Language Model 8 Sep 2020 · 0 repositories · arXiv:2009.03706
-
Masked Label Prediction: Unified Message Passing Model for Semi-Supervised Classification 8 Sep 2020 · 3 repositories · arXiv:2009.03509Syntology official (archive's flag): 3 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 2 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples) · 2 pointer-only (licence)
-
Adversarial Watermarking Transformer: Towards Tracing Text Provenance with Data Hiding 7 Sep 2020 · 1 repository · arXiv:2009.03015Syntology 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Black Box to White Box: Discover Model Characteristics Based on Strategic Probing 7 Sep 2020 · 0 repositories · arXiv:2009.03136
-
E-BERT: A Phrase and Product Knowledge Enhanced Language Model for E-commerce 7 Sep 2020 · 0 repositories · arXiv:2009.02835
-
Improving Language Generation with Sentence Coherence Objective 7 Sep 2020 · 1 repository · arXiv:2009.06358
-
Measuring Massive Multitask Language Understanding 7 Sep 2020 · 18 repositories · arXiv:2009.03300Syntology community repositories only · 19 ran (of which 0 constructed an object rather than computing a result; 15 with no instrument failure: 0 honoured, 0 violated, 15 with no contract checked; 4 where Syntology's instrument failed) · 7 unverified (of 26 harvested samples) · 1 pointer-only (licence)
-
Robust Conversational AI with Grounded Text Generation 7 Sep 2020 · 0 repositories · arXiv:2009.03457
-
TransModality: An End2End Fusion Method with Transformer for Multimodal Sentiment Analysis 7 Sep 2020 · 0 repositories · arXiv:2009.02902
-
EdinburghNLP at WNUT-2020 Task 2: Leveraging Transformers with Generalized Augmentation for Identifying Informativeness in COVID-19 Tweets 6 Sep 2020 · 0 repositories · arXiv:2009.06375
-
QiaoNing at SemEval-2020 Task 4: Commonsense Validation and Explanation system based on ensemble of language model 6 Sep 2020 · 0 repositories · arXiv:2009.02645
-
UPB at SemEval-2020 Task 8: Joint Textual and Visual Modeling in a Multi-Task Learning Architecture for Memotion Analysis 6 Sep 2020 · 0 repositories · arXiv:2009.02779
-
Accenture at CheckThat! 2020: If you say so: Post-hoc fact-checking of claims using transformer-based models 5 Sep 2020 · 0 repositories · arXiv:2009.02431
-
Voice Conversion by Cascading Automatic Speech Recognition and Text-to-Speech Synthesis with Prosody Transfer 3 Sep 2020 · 0 repositories · arXiv:2009.01475
-
Comparative Evaluation of Pretrained Transfer Learning Models on Automatic Short Answer Grading 2 Sep 2020 · 1 repository · arXiv:2009.01303
-
Automatic Assignment of Radiology Examination Protocols Using Pre-trained Language Models with Knowledge Distillation 1 Sep 2020 · 1 repository · arXiv:2009.00694
-
LiftFormer: 3D Human Pose Estimation using attention models 1 Sep 2020 · 0 repositories · arXiv:2009.00348
-
Sentimental LIAR: Extended Corpus and Deep Learning Models for Fake Claim Classification 1 Sep 2020 · 2 repositories · arXiv:2009.01047
-
A Bidirectional Tree Tagging Scheme for Joint Medical Relation Extraction 31 Aug 2020 · 0 repositories · arXiv:2008.13339
-
Parallel Rescoring with Transformer for Streaming On-Device Speech Recognition 30 Aug 2020 · 0 repositories · arXiv:2008.13093
-
SocCogCom at SemEval-2020 Task 11: Characterizing and Detecting Propaganda using Sentence-Level Emotional Salience Features 29 Aug 2020 · 1 repository · arXiv:2008.13012
-
HittER: Hierarchical Transformers for Knowledge Graph Embeddings 28 Aug 2020 · 3 repositories · arXiv:2008.12813Syntology 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples) · 1 pointer-only (licence)
-
Knowledge Efficient Deep Learning for Natural Language Processing 28 Aug 2020 · 0 repositories · arXiv:2008.12878
-
Rethinking the Objectives of Extractive Question Answering 28 Aug 2020 · 1 repository · arXiv:2008.12804Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
TATL at W-NUT 2020 Task 2: A Transformer-based Baseline System for Identification of Informative COVID-19 English Tweets 28 Aug 2020 · 0 repositories · arXiv:2008.12854
-
Text-Conditioned Transformer for Automatic Pronunciation Error Detection 28 Aug 2020 · 0 repositories · arXiv:2008.12424
-
A Fast and Robust BERT-based Dialogue State Tracker for Schema-Guided Dialogue Dataset 27 Aug 2020 · 1 repository · arXiv:2008.12335
-
AMBERT: A Pre-trained Language Model with Multi-Grained Tokenization 27 Aug 2020 · 0 repositories · arXiv:2008.11869
-
DAVE: Deriving Automatically Verilog from English 27 Aug 2020 · 0 repositories · arXiv:2009.01026
-
Entity and Evidence Guided Relation Extraction for DocRED 27 Aug 2020 · 0 repositories · arXiv:2008.12283
-
GREEK-BERT: The Greeks visiting Sesame Street 27 Aug 2020 · 1 repository · arXiv:2008.12014Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
Improvement of a dedicated model for open domain persona-aware dialogue generation 27 Aug 2020 · 1 repository · arXiv:2008.11970
-
MultiGBS: A multi-layer graph approach to biomedical summarization 27 Aug 2020 · 0 repositories · arXiv:2008.11908
-
Query Focused Multi-document Summarisation of Biomedical Texts 27 Aug 2020 · 1 repository · arXiv:2008.11986
-
Query Focused Multi-document Summarisation of Biomedical Texts: Macquarie Universiy and the Australian National University at BioASQ8b 27 Aug 2020 · 1 repository
-
A Multitask Deep Learning Approach for User Depression Detection on Sina Weibo 26 Aug 2020 · 0 repositories · arXiv:2008.11708
-
APMSqueeze: A Communication Efficient Adam-Preconditioned Momentum SGD Algorithm 26 Aug 2020 · 0 repositories · arXiv:2008.11343
-
Discrete Word Embedding for Logical Natural Language Understanding 26 Aug 2020 · 0 repositories · arXiv:2008.11649
-
Analysis and Evaluation of Language Models for Word Sense Disambiguation 26 Aug 2020 · 1 repository · arXiv:2008.11608
-
Conceptualized Representation Learning for Chinese Biomedical Text Mining 25 Aug 2020 · 0 repositories · arXiv:2008.10813
-
ETC-NLG: End-to-end Topic-Conditioned Natural Language Generation 25 Aug 2020 · 1 repository · arXiv:2008.10875