Methods › General › Normalization › Layer Normalization › Papers, page 218
Layer Normalization
Papers archive 2025-07-28
archive papers tagged: 24,980 · with a code link: 11,273 · where Syntology ran a sample: 3,471 (2,923 with a run with no instrument failure, 548 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,471 of 24,980 tagged: 2,923 with a run with no instrument failure, 548 where every run was a failure of Syntology's instrument)
Page 218 of 250: papers 21,701 to 21,800 of 24,980, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
DialogXL: All-in-One XLNet for Multi-Party Conversation Emotion Recognition 16 Dec 2020 · 4 repositories · arXiv:2012.08695
-
Learning from Mistakes: Using Mis-predictions as Harm Alerts in Language Pre-Training 16 Dec 2020 · 0 repositories · arXiv:2012.08789
-
Point Transformer 16 Dec 2020 · 24 repositories · arXiv:2012.09164
-
Query expansion with artificially generated texts 16 Dec 2020 · 0 repositories · arXiv:2012.08787
-
R²-Net: Relation of Relation Learning Network for Sentence Semantic Matching 16 Dec 2020 · 0 repositories · arXiv:2012.08920
-
Revisiting Linformer with a modified self-attention with linear complexity 16 Dec 2020 · 0 repositories · arXiv:2101.10277
-
High throughput screening with machine learning 15 Dec 2020 · 0 repositories · arXiv:2012.08275
-
Pre-Training Transformers as Energy-Based Cloze Models 15 Dec 2020 · 1 repository · arXiv:2012.08561
-
RecipeNLG: A Cooking Recipes Dataset for Semi-Structured Text Generation 15 Dec 2020 · 1 repository
-
Traditional IR rivals neural models on the MS MARCO Document Ranking Leaderboard 15 Dec 2020 · 2 repositories · arXiv:2012.08020
-
Contrastive Learning with Adversarial Perturbations for Conditional Text Generation 14 Dec 2020 · 1 repository · arXiv:2012.07280Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Extracting Training Data from Large Language Models 14 Dec 2020 · 3 repositories · arXiv:2012.07805Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
Informer: Beyond Efficient Transformer for Long Sequence Time-Series Forecasting 14 Dec 2020 · 14 repositories · arXiv:2012.07436Syntology official: harvested, nothing ran · 62 ran (of which 53 constructed an object rather than computing a result; 62 with no instrument failure: 0 honoured, 1 violated, 61 with no contract checked; 0 where Syntology's instrument failed) · 14 unverified (of 76 harvested samples) · 13 pointer-only (licence)
-
LRC-BERT: Latent-representation Contrastive Knowledge Distillation for Natural Language Understanding 14 Dec 2020 · 0 repositories · arXiv:2012.07335
-
Reasoning in Dialog: Improving Response Generation by Context Reading Comprehension 14 Dec 2020 · 1 repository · arXiv:2012.07410
-
Vartani Spellcheck -- Automatic Context-Sensitive Spelling Correction of OCR-generated Hindi Text Using BERT and Levenshtein Distance 14 Dec 2020 · 0 repositories · arXiv:2012.07652
-
Discriminative Pre-training for Low Resource Title Compression in Conversational Grocery 13 Dec 2020 · 0 repositories · arXiv:2012.06943
-
Improving Image Captioning by Leveraging Intra- and Inter-layer Global Representation in Transformer Network 13 Dec 2020 · 1 repository · arXiv:2012.07061
-
KVL-BERT: Knowledge Enhanced Visual-and-Linguistic BERT for Visual Commonsense Reasoning 13 Dec 2020 · 0 repositories · arXiv:2012.07000
-
MiniVLM: A Smaller and Faster Vision-Language Model 13 Dec 2020 · 0 repositories · arXiv:2012.06946
-
CogALex-VI Shared Task: Transrelation - A Robust Multilingual Language Model for Multilingual Relation Identification 12 Dec 2020 · 1 repository
-
DETR for Crowd Pedestrian Detection 12 Dec 2020 · 1 repository · arXiv:2012.06785
-
Yelp Review Rating Prediction: Machine Learning and Deep Learning Models 12 Dec 2020 · 1 repository · arXiv:2012.06690
-
Hardware Beyond Backpropagation: a Photonic Co-Processor for Direct Feedback Alignment 11 Dec 2020 · 0 repositories · arXiv:2012.06373
-
Improving Task-Agnostic BERT Distillation with Layer Mapping Search 11 Dec 2020 · 0 repositories · arXiv:2012.06153
-
Spatial Temporal Transformer Network for Skeleton-based Action Recognition 11 Dec 2020 · 1 repository · arXiv:2012.06399
-
TabTransformer: Tabular Data Modeling Using Contextual Embeddings 11 Dec 2020 · 12 repositories · arXiv:2012.06678Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
A Practical Approach towards Causality Mining in Clinical Text using Active Transfer Learning 10 Dec 2020 · 0 repositories · arXiv:2012.07563
-
As Good as New. How to Successfully Recycle English GPT-2 to Make Models for Other Languages 10 Dec 2020 · 1 repository · arXiv:2012.05628
-
GDA-HIN: A Generalized Domain Adaptive Model across Heterogeneous Information Networks 10 Dec 2020 · 0 repositories · arXiv:2012.05688
-
Towards Neural Programming Interfaces 10 Dec 2020 · 1 repository · arXiv:2012.05983
-
Cross-lingual Word Sense Disambiguation using mBERT Embeddings with Syntactic Dependencies 9 Dec 2020 · 0 repositories · arXiv:2012.05300
-
Mapping the Space of Chemical Reactions Using Attention-Based Neural Networks 9 Dec 2020 · 1 repository · arXiv:2012.06051
-
Simple is not Easy: A Simple Strong Baseline for TextVQA and TextCaps 9 Dec 2020 · 1 repository · arXiv:2012.05153
-
Discourse Parsing of Contentious, Non-Convergent Online Discussions 8 Dec 2020 · 0 repositories · arXiv:2012.04585
-
Large-scale Quantitative Evidence of Media Impact on Public Opinion toward China 8 Dec 2020 · 0 repositories · arXiv:2012.07575
-
Efficient Estimation of Influence of a Training Instance 8 Dec 2020 · 0 repositories · arXiv:2012.04207
-
Extractive Opinion Summarization in Quantized Transformer Spaces 8 Dec 2020 · 2 repositories · arXiv:2012.04443
-
From Bag of Sentences to Document: Distantly Supervised Relation Extraction via Machine Reading Comprehension 8 Dec 2020 · 1 repository · arXiv:2012.04334
-
Parameter Efficient Multimodal Transformers for Video Representation Learning 8 Dec 2020 · 0 repositories · arXiv:2012.04124
-
TADO: Time-varying Attention with Dual-Optimizer Model 8 Dec 2020 · 1 repository · arXiv:2012.04558
-
An Empirical Survey of Unsupervised Text Representation Methods on Twitter Data 7 Dec 2020 · 0 repositories · arXiv:2012.03468
-
CX DB8: A queryable extractive summarizer and semantic search engine 7 Dec 2020 · 2 repositories · arXiv:2012.03942
-
Dartmouth CS at WNUT-2020 Task 2: Informative COVID-19 Tweet Classification Using BERT 7 Dec 2020 · 0 repositories · arXiv:2012.04539
-
Deep Policy Networks for NPC Behaviors that Adapt to Changing Design Parameters in Roguelike Games 7 Dec 2020 · 0 repositories · arXiv:2012.03532
-
Detecting Insincere Questions from Text: A Transfer Learning Approach 7 Dec 2020 · 1 repository · arXiv:2012.07587
-
Document Graph for Neural Machine Translation 7 Dec 2020 · 0 repositories · arXiv:2012.03477
-
KgPLM: Knowledge-guided Language Model Pre-training via Generative and Discriminative Learning 7 Dec 2020 · 0 repositories · arXiv:2012.03551
-
UBAR: Towards Fully End-to-End Task-Oriented Dialog Systems with GPT-2 7 Dec 2020 · 1 repository · arXiv:2012.03539
-
[Re] Satellite Image Time Series Classification with Pixel-Set Encoders and Temporal Self-Attention 6 Dec 2020 · 1 repository
-
Data-Efficient Methods for Dialogue Systems 5 Dec 2020 · 0 repositories · arXiv:2012.02929
-
Enhanced Offensive Language Detection Through Data Augmentation 5 Dec 2020 · 0 repositories · arXiv:2012.02954
-
Pre-training Protein Language Models with Label-Agnostic Binding Pairs Enhances Performance in Downstream Tasks 5 Dec 2020 · 1 repository · arXiv:2012.03084
-
Automated Detection of Cyberbullying Against Women and Immigrants and Cross-domain Adaptability 4 Dec 2020 · 0 repositories · arXiv:2012.02565
-
Batch Group Normalization 4 Dec 2020 · 0 repositories · arXiv:2012.02782
-
CUED_speech at TREC 2020 Podcast Summarisation Track 4 Dec 2020 · 0 repositories · arXiv:2012.02535
-
EchoBERT: A Transformer-Based Approach for Behavior Detection in Echograms 4 Dec 2020 · 1 repository
-
Fine-tuning BERT for Low-Resource Natural Language Understanding via Active Learning 4 Dec 2020 · 0 repositories · arXiv:2012.02462
-
Modelling General Properties of Nouns by Selectively Averaging Contextualised Embeddings 4 Dec 2020 · 0 repositories · arXiv:2012.07580
-
Playing Text-Based Games with Common Sense 4 Dec 2020 · 0 repositories · arXiv:2012.02757
-
Pre-trained language models as knowledge bases for Automotive Complaint Analysis 4 Dec 2020 · 0 repositories · arXiv:2012.02558
-
RPT: Relational Pre-trained Transformer Is Almost All You Need towards Democratizing Data Preparation 4 Dec 2020 · 0 repositories · arXiv:2012.02469
-
Spread Mechanism and Influence Measurement of Online Rumors in China During the COVID-19 Pandemic 4 Dec 2020 · 0 repositories · arXiv:2012.02446
-
BERT-hLSTMs: BERT and Hierarchical LSTMs for Visual Storytelling 3 Dec 2020 · 0 repositories · arXiv:2012.02128
-
Circles are like Ellipses, or Ellipses are like Circles? Measuring the Degree of Asymmetry of Static and Contextual Embeddings and the Implications to Representation Learning 3 Dec 2020 · 0 repositories · arXiv:2012.01631
-
DialogBERT: Discourse-Aware Response Generation via Learning to Recover and Rank Utterances 3 Dec 2020 · 1 repository · arXiv:2012.01775Syntology 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Federated Learning for Personalized Humor Recognition 3 Dec 2020 · 0 repositories · arXiv:2012.01675
-
GottBERT: a pure German Language Model 3 Dec 2020 · 0 repositories · arXiv:2012.02110
-
Sentiment analysis in Bengali via transfer learning using multi-lingual BERT 3 Dec 2020 · 1 repository · arXiv:2012.07538
-
TRACE: Early Detection of Chronic Kidney Disease Onset with Transformer-Enhanced Feature Embedding 3 Dec 2020 · 0 repositories · arXiv:2012.03729
-
What Makes a Star Teacher? A Hierarchical BERT Model for Evaluating Teacher's Performance in Online Education 3 Dec 2020 · 0 repositories · arXiv:2012.01633
-
A Framework and Dataset for Abstract Art Generation via CalligraphyGAN 2 Dec 2020 · 0 repositories · arXiv:2012.00744
-
Contour Transformer Network for One-shot Segmentation of Anatomical Structures 2 Dec 2020 · 1 repository · arXiv:2012.01480
-
Exploiting BERT to improve aspect-based sentiment analysis performance on Persian language 2 Dec 2020 · 1 repository · arXiv:2012.07510
-
How Can We Know When Language Models Know? On the Calibration of Language Models for Question Answering 2 Dec 2020 · 1 repository · arXiv:2012.00955
-
Two-Stage Single Image Reflection Removal with Reflection-Aware Guidance 2 Dec 2020 · 1 repository · arXiv:2012.00945
-
A Deep Generative Approach to Native Language Identification 1 Dec 2020 · 0 repositories
-
A Large-Scale Corpus of E-mail Conversations with Standard and Two-Level Dialogue Act Annotations 1 Dec 2020 · 0 repositories
-
A Neural Local Coherence Analysis Model for Clarity Text Scoring 1 Dec 2020 · 0 repositories
-
Adversarial Sparse Transformer for Time Series Forecasting 1 Dec 2020 · 1 repository
-
Affective and Contextual Embedding for Sarcasm Detection 1 Dec 2020 · 1 repository
-
AlexU-AUX-BERT at SemEval-2020 Task 3: Improving BERT Contextual Similarity Using Multiple Auxiliary Contexts 1 Dec 2020 · 0 repositories
-
AlexU-BackTranslation-TL at SemEval-2020 Task 12: Improving Offensive Language Detection Using Data Augmentation and Transfer Learning 1 Dec 2020 · 0 repositories
-
ALT at SemEval-2020 Task 12: Arabic and English Offensive Language Identification in Social Media 1 Dec 2020 · 0 repositories
-
Analogy Models for Neural Word Inflection 1 Dec 2020 · 1 repository
-
Arabizi Language Models for Sentiment Analysis 1 Dec 2020 · 0 repositories
-
Assessing Polyseme Sense Similarity through Co-predication Acceptability and Contextualised Embedding Distance 1 Dec 2020 · 0 repositories
-
Attentively Embracing Noise for Robust Latent Representation in BERT 1 Dec 2020 · 1 repository
-
BERT at SemEval-2020 Task 8: Using BERT to Analyse Meme Emotions 1 Dec 2020 · 0 repositories
-
BERT-based Cohesion Analysis of Japanese Texts 1 Dec 2020 · 1 repository
-
BERT-Based Neural Collaborative Filtering and Fixed-Length Contiguous Tokens Explanation 1 Dec 2020 · 0 repositories
-
BERTatDE at SemEval-2020 Task 6: Extracting Term-definition Pairs in Free Text Using Pre-trained Model 1 Dec 2020 · 0 repositories
-
Bilingual Subword Segmentation for Neural Machine Translation 1 Dec 2020 · 0 repositories
-
BLCU-NLP at SemEval-2020 Task 5: Data Augmentation for Efficient Counterfactual Detecting 1 Dec 2020 · 0 repositories
-
BYteam at SemEval-2020 Task 5: Detecting Counterfactual Statements with BERT and Ensembles 1 Dec 2020 · 0 repositories
-
Cardiff University at SemEval-2020 Task 6: Fine-tuning BERT for Domain-Specific Definition Classification 1 Dec 2020 · 0 repositories
-
CitiusNLP at SemEval-2020 Task 3: Comparing Two Approaches for Word Vector Contextualization 1 Dec 2020 · 0 repositories
-
Classifier Probes May Just Learn from Linear Context Features 1 Dec 2020 · 1 repository
-
ClimaText: A Dataset for Climate Change Topic Detection 1 Dec 2020 · 0 repositories · arXiv:2012.00483
-
CN-HIT-MI.T at SemEval-2020 Task 8: Memotion Analysis Based on BERT 1 Dec 2020 · 0 repositories