Methods › General › Normalization › Layer Normalization › Papers, page 176
Layer Normalization
Papers archive 2025-07-28
archive papers tagged: 24,980 · with a code link: 11,273 · where Syntology ran a sample: 3,471 (2,923 with a run with no instrument failure, 548 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,471 of 24,980 tagged: 2,923 with a run with no instrument failure, 548 where every run was a failure of Syntology's instrument)
Page 176 of 250: papers 17,501 to 17,600 of 24,980, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
ResT V2: Simpler, Faster and Stronger 15 Apr 2022 · 2 repositories · arXiv:2204.07366
-
Text Revision by On-the-Fly Representation Optimization 15 Apr 2022 · 1 repository · arXiv:2204.07359
-
Unconditional Image-Text Pair Generation with Multimodal Cross Quantizer 15 Apr 2022 · 1 repository · arXiv:2204.07537
-
3D Shuffle-Mixer: An Efficient Context-Aware Vision Learner of Transformer-MLP Paradigm for Dense Prediction in Medical Volume 14 Apr 2022 · 0 repositories · arXiv:2204.06779
-
Activation Regression for Continuous Domain Generalization with Applications to Crop Classification 14 Apr 2022 · 1 repository · arXiv:2204.07030
-
Analysing similarities between legal court documents using natural language processing approaches based on Transformers 14 Apr 2022 · 0 repositories · arXiv:2204.07182
-
CalBERT - Code-mixed Adaptive Language representations using BERT 14 Apr 2022 · 1 repository
-
Causal Transformer for Estimating Counterfactual Outcomes 14 Apr 2022 · 1 repository · arXiv:2204.07258
-
Challenges for Open-domain Targeted Sentiment Analysis 14 Apr 2022 · 0 repositories · arXiv:2204.06893
-
DeiT III: Revenge of the ViT 14 Apr 2022 · 12 repositories · arXiv:2204.07118
-
Does BERT really agree ? Fine-grained Analysis of Lexical Dependence on a Syntactic Task 14 Apr 2022 · 0 repositories · arXiv:2204.06889
-
Generative power of a protein language model trained on multiple sequence alignments 14 Apr 2022 · 1 repository · arXiv:2204.07110
-
GPT-NeoX-20B: An Open-Source Autoregressive Language Model 14 Apr 2022 · 11 repositories · arXiv:2204.06745Syntology community repositories only · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Hierarchical Embedded Bayesian Additive Regression Trees 14 Apr 2022 · 0 repositories · arXiv:2204.07207
-
Latent Aspect Detection from Online Unsolicited Customer Reviews 14 Apr 2022 · 1 repository · arXiv:2204.06964
-
MiniViT: Compressing Vision Transformers with Weight Multiplexing 14 Apr 2022 · 2 repositories · arXiv:2204.07154Syntology official: harvested for another paper · 0 ran · 4 unverified (of 4 harvested samples)
-
Multi-label topic classification for COVID-19 literature with Bioformer 14 Apr 2022 · 0 repositories · arXiv:2204.06758
-
Residual Swin Transformer Channel Attention Network for Image Demosaicing 14 Apr 2022 · 0 repositories · arXiv:2204.07098
-
Rows from Many Sources: Enriching row completions from Wikidata with a pre-trained Language Model 14 Apr 2022 · 0 repositories · arXiv:2204.07014
-
Deep Relation Learning for Regression and Its Application to Brain Age Estimation 13 Apr 2022 · 0 repositories · arXiv:2204.06598
-
Fix Bugs with Transformer through a Neural-Symbolic Edit Grammar 13 Apr 2022 · 0 repositories · arXiv:2204.06643
-
Formal Language Recognition by Hard Attention Transformers: Perspectives from Circuit Complexity 13 Apr 2022 · 0 repositories · arXiv:2204.06618
-
HuBERT-EE: Early Exiting HuBERT for Efficient Speech Recognition 13 Apr 2022 · 1 repository · arXiv:2204.06328
-
IIITDWD-ShankarB@ Dravidian-CodeMixi-HASOC2021: mBERT based model for identification of offensive content in south Indian languages 13 Apr 2022 · 0 repositories · arXiv:2204.10195
-
METRO: Efficient Denoising Pretraining of Large Scale Autoencoding Language Models with Model Generated Signals 13 Apr 2022 · 0 repositories · arXiv:2204.06644
-
Probing for Constituency Structure in Neural Language Models 13 Apr 2022 · 1 repository · arXiv:2204.06201Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Recognition of Freely Selected Keypoints on Human Limbs 13 Apr 2022 · 0 repositories · arXiv:2204.06326
-
Building Markovian Generative Architectures over Pretrained LM Backbones for Efficient Task-Oriented Dialog Systems 13 Apr 2022 · 2 repositories · arXiv:2204.06452
-
TangoBERT: Reducing Inference Cost by using Cascaded Architecture 13 Apr 2022 · 0 repositories · arXiv:2204.06271
-
Are Multimodal Transformers Robust to Missing Modality? 12 Apr 2022 · 0 repositories · arXiv:2204.05454
-
Enhancement of Pitch Controllability using Timbre-Preserving Pitch Augmentation in FastPitch 12 Apr 2022 · 0 repositories · arXiv:2204.05753
-
Explore More Guidance: A Task-aware Instruction Network for Sign Language Translation Enhanced with Data Augmentation 12 Apr 2022 · 1 repository · arXiv:2204.05953
-
Few-shot Learning with Noisy Labels 12 Apr 2022 · 1 repository · arXiv:2204.05494
-
HiTPR: Hierarchical Transformer for Place Recognition in Point Cloud 12 Apr 2022 · 0 repositories · arXiv:2204.05481
-
L3Cube-MahaNER: A Marathi Named Entity Recognition Dataset and BERT models 12 Apr 2022 · 1 repository · arXiv:2204.06029
-
SwinNet: Swin Transformer drives edge-aware RGB-D and RGB-T salient object detection 12 Apr 2022 · 1 repository · arXiv:2204.05585
-
What Language Model Architecture and Pretraining Objective Work Best for Zero-Shot Generalization? 12 Apr 2022 · 1 repository · arXiv:2204.05832Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 1 honoured, 0 violated, 11 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 16 harvested samples)
-
Category-Aware Transformer Network for Better Human-Object Interaction Detection 11 Apr 2022 · 0 repositories · arXiv:2204.04911
-
A Token-level Contrastive Framework for Sign Language Translation 11 Apr 2022 · 1 repository · arXiv:2204.04916Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
Event Transformer 11 Apr 2022 · 1 repository · arXiv:2204.05172
-
HiMODE: A Hybrid Monocular Omnidirectional Depth Estimation Model 11 Apr 2022 · 0 repositories · arXiv:2204.05007
-
Large-Scale Streaming End-to-End Speech Translation with Neural Transducers 11 Apr 2022 · 1 repository · arXiv:2204.05352Syntology 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Permutation-Invariant Relational Network for Multi-person 3D Pose Estimation 11 Apr 2022 · 0 repositories · arXiv:2204.04913
-
Self-supervised Vision Transformers for Joint SAR-optical Representation Learning 11 Apr 2022 · 2 repositories · arXiv:2204.05381
-
SUMD: Super U-shaped Matrix Decomposition Convolutional neural network for Image denoising 11 Apr 2022 · 0 repositories · arXiv:2204.04861
-
Team ÚFAL at CMCL 2022 Shared Task: Figuring out the correct recipe for predicting Eye-Tracking features using Pretrained Language Models 11 Apr 2022 · 0 repositories · arXiv:2204.04998
-
Tokenwise Contrastive Pretraining for Finer Speech-to-BERT Alignment in End-to-End Speech-to-Intent Systems 11 Apr 2022 · 0 repositories · arXiv:2204.05188
-
Towards Generalizable Semantic Product Search by Text Similarity Pre-training on Search Click Logs 11 Apr 2022 · 0 repositories · arXiv:2204.05231
-
Uniform Complexity for Text Generation 11 Apr 2022 · 1 repository · arXiv:2204.05185
-
Confidence Estimation Transformer for Long-term Renewable Energy Forecasting in Reinforcement Learning-based Power Grid Dispatching 10 Apr 2022 · 1 repository · arXiv:2204.04612
-
Representation Learning by Detecting Incorrect Location Embeddings 10 Apr 2022 · 1 repository · arXiv:2204.04788
-
Fake news detection using parallel BERT deep neural networks 10 Apr 2022 · 0 repositories · arXiv:2204.04793
-
Fashionformer: A simple, Effective and Unified Baseline for Human Fashion Segmentation and Recognition 10 Apr 2022 · 1 repository · arXiv:2204.04654
-
Few-Shot Cross-lingual Transfer for Coarse-grained De-identification of Code-Mixed Clinical Texts 10 Apr 2022 · 1 repository · arXiv:2204.04775
-
Panoptic-PartFormer: Learning a Unified Model for Panoptic Part Segmentation 10 Apr 2022 · 1 repository · arXiv:2204.04655
-
Pushing on Personality Detection from Verbal Behavior: A Transformer Meets Text Contours of Psycholinguistic Features 10 Apr 2022 · 0 repositories · arXiv:2204.04629
-
Self-Supervised Audio-and-Text Pre-training with Extremely Low-Resource Parallel Data 10 Apr 2022 · 1 repository · arXiv:2204.04645
-
Efficient Extraction of Pathologies from C-Spine Radiology Reports using Multi-Task Learning 9 Apr 2022 · 0 repositories · arXiv:2204.04544
-
FoundationLayerNorm: Scaling BERT and GPT to 1,000 Layers 9 Apr 2022 · 0 repositories · arXiv:2204.04477
-
Modeling Multi-Granularity Hierarchical Features for Relation Extraction 9 Apr 2022 · 1 repository · arXiv:2204.04437Syntology official: no sample here; runs from other or unrecorded repositories · 2 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Are We Really Making Much Progress in Text Classification? A Comparative Review 8 Apr 2022 · 1 repository · arXiv:2204.03954
-
BioBART: Pretraining and Evaluation of A Biomedical Generative Language Model 8 Apr 2022 · 1 repository · arXiv:2204.03905
-
Contextual Representation Learning beyond Masked Language Modeling 8 Apr 2022 · 1 repository · arXiv:2204.04163
-
Does Robustness on ImageNet Transfer to Downstream Tasks? 8 Apr 2022 · 0 repositories · arXiv:2204.03934
-
Efficient tracking of team sport players with few game-specific annotations 8 Apr 2022 · 0 repositories · arXiv:2204.04049
-
Enhance Incomplete Utterance Restoration by Joint Learning Token Extraction and Text Generation 8 Apr 2022 · 1 repository · arXiv:2204.03958
-
Exploring Transformer's potential on automatic piano transcription 8 Apr 2022 · 0 repositories · arXiv:2204.03898
-
Infusing Knowledge from Wikipedia to Enhance Stance Detection 8 Apr 2022 · 2 repositories · arXiv:2204.03839
-
Learning Trajectory-Aware Transformer for Video Super-Resolution 8 Apr 2022 · 1 repository · arXiv:2204.04216Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 1 pointer-only (licence)
-
MMTAfrica: Multilingual Machine Translation for African Languages 8 Apr 2022 · 1 repository · arXiv:2204.04306
-
Points to Patches: Enabling the Use of Self-Attention for 3D Shape Recognition 8 Apr 2022 · 1 repository · arXiv:2204.03957
-
A Survey of Supernet Optimization and its Applications: Spatial and Temporal Optimization for Neural Architecture Search 8 Apr 2022 · 0 repositories · arXiv:2204.03916
-
Towards Understanding Large-Scale Discourse Structures in Pre-Trained and Fine-Tuned Language Models 8 Apr 2022 · 0 repositories · arXiv:2204.04289
-
Transformer-Based Self-Supervised Learning for Emotion Recognition 8 Apr 2022 · 0 repositories · arXiv:2204.05103
-
Vision Transformers for Single Image Dehazing 8 Apr 2022 · 1 repository · arXiv:2204.03883Syntology 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 0 violated, 2 with no contract checked; 2 where Syntology's instrument failed) · 5 unverified (of 10 harvested samples) · 3 pointer-only (licence)
-
Accelerating Attention through Gradient-Based Learned Runtime Pruning 7 Apr 2022 · 0 repositories · arXiv:2204.03227
-
Autoencoding Language Model Based Ensemble Learning for Commonsense Validation and Explanation 7 Apr 2022 · 0 repositories · arXiv:2204.03324
-
BERTuit: Understanding Spanish language in Twitter through a native transformer 7 Apr 2022 · 0 repositories · arXiv:2204.03465
-
Compositional Generalization and Decomposition in Neural Program Synthesis 7 Apr 2022 · 0 repositories · arXiv:2204.03758
-
DaViT: Dual Attention Vision Transformers 7 Apr 2022 · 4 repositories · arXiv:2204.03645Syntology official (archive's flag): 8 ran · 8 ran (of which 5 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 2 where Syntology's instrument failed) · 7 unverified (of 15 harvested samples)
-
CoCoSoDa: Effective Contrastive Learning for Code Search 7 Apr 2022 · 0 repositories · arXiv:2204.03293
-
Event Transformer. A sparse-aware solution for efficient event data processing 7 Apr 2022 · 1 repository · arXiv:2204.03355
-
Low-Dose CT Denoising via Sinogram Inner-Structure Transformer 7 Apr 2022 · 0 repositories · arXiv:2204.03163
-
Multi-Task Distributed Learning using Vision Transformer with Random Patch Permutation 7 Apr 2022 · 0 repositories · arXiv:2204.03500
-
PALBERT: Teaching ALBERT to Ponder 7 Apr 2022 · 1 repository · arXiv:2204.03276Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; the one sample that ran constructed an object rather than computing a result (of 2 harvested samples)
-
Pretraining Text Encoders with Adversarial Mixture of Training Signal Generators 7 Apr 2022 · 1 repository · arXiv:2204.03243Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 8 harvested samples) · 2 pointer-only (licence)
-
T4PdM: a Deep Neural Network based on the Transformer Architecture for Fault Diagnosis of Rotating Machinery 7 Apr 2022 · 0 repositories · arXiv:2204.03725
-
Testing the limits of natural language models for predicting human language judgments 7 Apr 2022 · 1 repository · arXiv:2204.03592
-
Unified Contrastive Learning in Image-Text-Label Space 7 Apr 2022 · 1 repository · arXiv:2204.03610Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Zero-Shot Category-Level Object Pose Estimation 7 Apr 2022 · 1 repository · arXiv:2204.03635
-
An Empirical Study of Remote Sensing Pretraining 6 Apr 2022 · 2 repositories · arXiv:2204.02825
-
ByT5 model for massively multilingual grapheme-to-phoneme conversion 6 Apr 2022 · 1 repository · arXiv:2204.03067
-
CCAT-NET: A Novel Transformer Based Semi-supervised Framework for Covid-19 Lung Lesion Segmentation 6 Apr 2022 · 0 repositories · arXiv:2204.02839
-
DAGAM: Data Augmentation with Generation And Modification 6 Apr 2022 · 1 repository · arXiv:2204.02633
-
Domain Specific Fine-tuning of Denoising Sequence-to-Sequence Models for Natural Language Summarization 6 Apr 2022 · 0 repositories · arXiv:2204.09716
-
drsphelps at SemEval-2022 Task 2: Learning idiom representations using BERTRAM 6 Apr 2022 · 0 repositories · arXiv:2204.02821
-
Knowledge Infused Decoding 6 Apr 2022 · 1 repository · arXiv:2204.03084Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
MixFormer: Mixing Features across Windows and Dimensions 6 Apr 2022 · 3 repositories · arXiv:2204.02557Syntology community repositories only · 10 ran (of which 3 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 14 harvested samples)
-
Paying More Attention to Self-attention: Improving Pre-trained Language Models via Attention Guiding 6 Apr 2022 · 0 repositories · arXiv:2204.02922
-
SMDT: Cross-View Geo-Localization with Image Alignment and Transformer 6 Apr 2022 · 4 repositories