Methods › General › Normalization › Layer Normalization › Papers, page 160
Layer Normalization
Papers archive 2025-07-28
archive papers tagged: 24,980 · with a code link: 11,273 · where Syntology ran a sample: 3,471 (2,923 with a run with no instrument failure, 548 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,471 of 24,980 tagged: 2,923 with a run with no instrument failure, 548 where every run was a failure of Syntology's instrument)
Page 160 of 250: papers 15,901 to 16,000 of 24,980, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Self-Repetition in Abstractive Neural Summarizers 14 Oct 2022 · 0 repositories · arXiv:2210.08145
-
TestAug: A Framework for Augmenting Capability-based NLP Tests 14 Oct 2022 · 1 repository · arXiv:2210.08097
-
Improving Transfer Learning with a Dual Image and Video Transformer for Multi-label Movie Trailer Genre Classification 14 Oct 2022 · 1 repository · arXiv:2210.07983
-
Automotive Multilingual Fault Diagnosis 13 Oct 2022 · 0 repositories · arXiv:2210.06918
-
Brain Network Transformer 13 Oct 2022 · 2 repositories · arXiv:2210.06681Syntology official (archive's flag): 3 ran · 3 ran (of which 1 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 7 harvested samples)
-
Categorizing Semantic Representations for Neural Machine Translation 13 Oct 2022 · 0 repositories · arXiv:2210.06709
-
Saliency Map Verbalization: Comparing Feature Importance Representations from Model-free and Instruction-based Methods 13 Oct 2022 · 1 repository · arXiv:2210.07222
-
ConvTransSeg: A Multi-resolution Convolution-Transformer Network for Medical Image Segmentation 13 Oct 2022 · 1 repository · arXiv:2210.07072
-
An Embarrassingly Simple Backdoor Attack on Self-supervised Learning 13 Oct 2022 · 4 repositories · arXiv:2210.07346Syntology community repositories only · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Explanations from Large Language Models Make Small Reasoners Better 13 Oct 2022 · 0 repositories · arXiv:2210.06726
-
Feature-Proxy Transformer for Few-Shot Segmentation 13 Oct 2022 · 2 repositories · arXiv:2210.06908
-
How to Train Vision Transformer on Small-scale Datasets? 13 Oct 2022 · 2 repositories · arXiv:2210.07240
-
Intermediate Prototype Mining Transformer for Few-Shot Semantic Segmentation 13 Oct 2022 · 1 repository · arXiv:2210.06780Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 3 where Syntology's instrument failed) · 4 unverified (of 13 harvested samples) · 10 pointer-only (licence)
-
Joint Reasoning on Hybrid-knowledge sources for Task-Oriented Dialog 13 Oct 2022 · 1 repository · arXiv:2210.07295
-
Jointly Reinforced User Simulator and Task-oriented Dialog System with Simplified Generative Architecture 13 Oct 2022 · 0 repositories · arXiv:2210.06706
-
Language Models of Code are Few-Shot Commonsense Learners 13 Oct 2022 · 2 repositories · arXiv:2210.07128Syntology community repositories only · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Large Language Models are few(1)-shot Table Reasoners 13 Oct 2022 · 1 repository · arXiv:2210.06710
-
Overlooked Video Classification in Weakly Supervised Video Anomaly Detection 13 Oct 2022 · 1 repository · arXiv:2210.06688
-
Scene Text Image Super-Resolution via Content Perceptual Loss and Criss-Cross Transformer Blocks 13 Oct 2022 · 0 repositories · arXiv:2210.06924
-
SQuAT: Sharpness- and Quantization-Aware Training for BERT 13 Oct 2022 · 0 repositories · arXiv:2210.07171
-
SWFormer: Sparse Window Transformer for 3D Object Detection in Point Clouds 13 Oct 2022 · 0 repositories · arXiv:2210.07372
-
Tone prediction and orthographic conversion for Basaa 13 Oct 2022 · 0 repositories · arXiv:2210.06986
-
Are Sample-Efficient NLP Models More Robust? 12 Oct 2022 · 0 repositories · arXiv:2210.06456
-
Bridging the Gap Between Vision Transformers and Convolutional Neural Networks on Small Datasets 12 Oct 2022 · 1 repository · arXiv:2210.05958Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 10 harvested samples) · 5 pointer-only (licence)
-
CTL++: Evaluating Generalization on Never-Seen Compositional Patterns of Known Functions, and Compatibility of Neural Representations 12 Oct 2022 · 1 repository · arXiv:2210.06350Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
DATScore: Evaluating Translation with Data Augmented Translations 12 Oct 2022 · 0 repositories · arXiv:2210.06576
-
ACSeg: Adaptive Conceptualization for Unsupervised Semantic Segmentation 12 Oct 2022 · 0 repositories · arXiv:2210.05944
-
Flare7K: A Phenomenological Nighttime Flare Removal Dataset 12 Oct 2022 · 1 repository · arXiv:2210.06570
-
FontTransformer: Few-shot High-resolution Chinese Glyph Image Synthesis via Stacked Transformers 12 Oct 2022 · 0 repositories · arXiv:2210.06301
-
Foundation Transformers 12 Oct 2022 · 4 repositories · arXiv:2210.06423Syntology community repositories only · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
Entity Tracking via Effective Use of Multi-Task Learning Model and Mention-guided Decoding 12 Oct 2022 · 2 repositories · arXiv:2210.06444Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
GMP*: Well-Tuned Gradual Magnitude Pruning Can Outperform Most BERT-Pruning Methods 12 Oct 2022 · 0 repositories · arXiv:2210.06384
-
Hate-CLIPper: Multimodal Hateful Meme Classification based on Cross-modal Interaction of CLIP Features 12 Oct 2022 · 1 repository · arXiv:2210.05916Syntology official: harvested, nothing ran · 0 ran · 5 unverified (of 5 harvested samples)
-
Improved Data Augmentation for Translation Suggestion 12 Oct 2022 · 0 repositories · arXiv:2210.06138
-
Instruction Tuning for Few-Shot Aspect-Based Sentiment Analysis 12 Oct 2022 · 1 repository · arXiv:2210.06629
-
JukeDrummer: Conditional Beat-aware Audio-domain Drum Accompaniment Generation via Transformer VQ-VAE 12 Oct 2022 · 1 repository · arXiv:2210.06007
-
The Lazy Neuron Phenomenon: On Emergence of Activation Sparsity in Transformers 12 Oct 2022 · 0 repositories · arXiv:2210.06313
-
Long-Form Video-Language Pre-Training with Multimodal Temporal Contrastive Learning 12 Oct 2022 · 1 repository · arXiv:2210.06031Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
MotionBERT: A Unified Perspective on Learning Human Motion Representations 12 Oct 2022 · 1 repository · arXiv:2210.06551Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 11 harvested samples) · 4 pointer-only (licence)
-
Integrating Translation Memories into Non-Autoregressive Machine Translation 12 Oct 2022 · 1 repository · arXiv:2210.06020
-
On Text Style Transfer via Style Masked Language Models 12 Oct 2022 · 0 repositories · arXiv:2210.06394
-
Predictive Querying for Autoregressive Neural Sequence Models 12 Oct 2022 · 1 repository · arXiv:2210.06464Syntology official (archive's flag): 7 ran · 7 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 5 where Syntology's instrument failed) · 6 unverified (of 13 harvested samples)
-
Prepended Domain Transformer: Heterogeneous Face Recognition without Bells and Whistles 12 Oct 2022 · 2 repositories · arXiv:2210.06529
-
Probing Commonsense Knowledge in Pre-trained Language Models with Sense-level Precision and Expanded Vocabulary 12 Oct 2022 · 1 repository · arXiv:2210.06376
-
RankT5: Fine-Tuning T5 for Text Ranking with Ranking Losses 12 Oct 2022 · 0 repositories · arXiv:2210.10634
-
S4ND: Modeling Images and Videos as Multidimensional Signals Using State Spaces 12 Oct 2022 · 1 repository · arXiv:2210.06583Syntology community repositories only · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
SUMBot: Summarizing Context in Open-Domain Dialogue Systems 12 Oct 2022 · 0 repositories · arXiv:2210.06496
-
Towards Theoretically Inspired Neural Initialization Optimization 12 Oct 2022 · 1 repository · arXiv:2210.05956
-
Uplift and Upsample: Efficient 3D Human Pose Estimation with Uplifting Transformers 12 Oct 2022 · 2 repositories · arXiv:2210.06110Syntology official (archive's flag): 7 ran · 15 ran (of which 0 constructed an object rather than computing a result; 15 with no instrument failure: 0 honoured, 0 violated, 15 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 17 harvested samples) · 2 pointer-only (licence)
-
ZITS++: Image Inpainting by Improving the Incremental Transformer on Structural Priors 12 Oct 2022 · 2 repositories · arXiv:2210.05950Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples) · 2 pointer-only (licence)
-
A Win-win Deal: Towards Sparse and Robust Pre-trained Language Models 11 Oct 2022 · 1 repository · arXiv:2210.05211Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
An Exploration of Hierarchical Attention Transformers for Efficient Long Document Classification 11 Oct 2022 · 0 repositories · arXiv:2210.05529
-
Are Pretrained Multilingual Models Equally Fair Across Languages? 11 Oct 2022 · 1 repository · arXiv:2210.05457
-
CLIP also Understands Text: Prompting CLIP for Phrase Understanding 11 Oct 2022 · 0 repositories · arXiv:2210.05836
-
Reliable Conditioning of Behavioral Cloning for Offline Reinforcement Learning 11 Oct 2022 · 1 repository · arXiv:2210.05158
-
Enriching Biomedical Knowledge for Low-resource Language Through Large-Scale Translation 11 Oct 2022 · 1 repository · arXiv:2210.05598
-
Memory transformers for full context and high-resolution 3D Medical Segmentation 11 Oct 2022 · 0 repositories · arXiv:2210.05313
-
Mixture of Attention Heads: Selecting Attention Heads Per Token 11 Oct 2022 · 2 repositories · arXiv:2210.05144Syntology official (archive's flag): 5 ran · 13 ran (of which 2 constructed an object rather than computing a result; 10 with no instrument failure: 1 honoured, 0 violated, 9 with no contract checked; 3 where Syntology's instrument failed) · 4 unverified (of 17 harvested samples) · 4 pointer-only (licence)
-
Multilingual BERT has an accent: Evaluating English influences on fluency in multilingual models 11 Oct 2022 · 0 repositories · arXiv:2210.05619
-
On the Interpolation of Contextualized Term-based Ranking with BM25 for Query-by-Example Retrieval 11 Oct 2022 · 1 repository · arXiv:2210.05512
-
On the Use of Semantically-Aligned Speech Representations for Spoken Language Understanding 11 Oct 2022 · 0 repositories · arXiv:2210.05291
-
Point Transformer V2: Grouped Vector Attention and Partition-based Pooling 11 Oct 2022 · 2 repositories · arXiv:2210.05666
-
Reflection of Thought: Inversely Eliciting Numerical Reasoning in Language Models via Solving Linear Systems 11 Oct 2022 · 0 repositories · arXiv:2210.05075
-
SaiT: Sparse Vision Transformers through Adaptive Token Pruning 11 Oct 2022 · 1 repository · arXiv:2210.05832
-
Streaming Punctuation for Long-form Dictation with Transformers 11 Oct 2022 · 0 repositories · arXiv:2210.05756
-
T5 for Hate Speech, Augmented Data and Ensemble 11 Oct 2022 · 1 repository · arXiv:2210.05480
-
Understanding the Failure of Batch Normalization for Transformers in NLP 11 Oct 2022 · 1 repository · arXiv:2210.05153Syntology official (archive's flag): 9 ran · 9 ran (of which 8 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
Viterbi Decoding of Directed Acyclic Transformer for Non-Autoregressive Machine Translation 11 Oct 2022 · 1 repository · arXiv:2210.05193Syntology official (archive's flag): 2 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Vote'n'Rank: Revision of Benchmarking with Social Choice Theory 11 Oct 2022 · 1 repository · arXiv:2210.05769Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 11 unverified (of 17 harvested samples)
-
A Memory Transformer Network for Incremental Learning 10 Oct 2022 · 0 repositories · arXiv:2210.04485
-
Characterization of anomalous diffusion through convolutional transformers 10 Oct 2022 · 0 repositories · arXiv:2210.04959
-
DCVQE: A Hierarchical Transformer for Video Quality Assessment 10 Oct 2022 · 0 repositories · arXiv:2210.04377
-
DEPTWEET: A Typology for Social Media Texts to Detect Depression Severities 10 Oct 2022 · 1 repository · arXiv:2210.05372
-
Empowering the Fact-checkers! Automatic Identification of Claim Spans on Twitter 10 Oct 2022 · 1 repository · arXiv:2210.04710
-
Ensemble Learning using Transformers and Convolutional Networks for Masked Face Recognition 10 Oct 2022 · 1 repository · arXiv:2210.04816
-
FS-DETR: Few-Shot DEtection TRansformer with prompting and without re-training 10 Oct 2022 · 0 repositories · arXiv:2210.04845
-
LAPFormer: A Light and Accurate Polyp Segmentation Transformer 10 Oct 2022 · 0 repositories · arXiv:2210.04393
-
MMT: Image-guided Story Ending Generation with Multimodal Memory Transformer 10 Oct 2022 · 1 repository
-
Multi-CLS BERT: An Efficient Alternative to Traditional Ensembling 10 Oct 2022 · 1 repository · arXiv:2210.05043
-
REV: Information-Theoretic Evaluation of Free-Text Rationales 10 Oct 2022 · 1 repository · arXiv:2210.04982Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Revisiting adapters with adversarial training 10 Oct 2022 · 0 repositories · arXiv:2210.04886
-
SCAM! Transferring humans between images with Semantic Cross Attention Modulation 10 Oct 2022 · 1 repository · arXiv:2210.04883
-
The Minimum Wage as an Anchor: Effects on Determinations of Fairness by Humans and AI 10 Oct 2022 · 0 repositories · arXiv:2210.10585
-
Uncertainty Quantification with Pre-trained Language Models: A Large-Scale Empirical Analysis 10 Oct 2022 · 1 repository · arXiv:2210.04714Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Visual Prompt Tuning for Test-time Domain Adaptation 10 Oct 2022 · 0 repositories · arXiv:2210.04831
-
A Transformer-based deep neural network model for SSVEP classification 9 Oct 2022 · 2 repositories · arXiv:2210.04172
-
AMPose: Alternately Mixed Global-Local Attention Model for 3D Human Pose Estimation 9 Oct 2022 · 0 repositories · arXiv:2210.04216
-
ASDOT: Any-Shot Data-to-Text Generation with Pretrained Language Models 9 Oct 2022 · 1 repository · arXiv:2210.04325
-
CHARD: Clinical Health-Aware Reasoning Across Dimensions for Text Generation Models 9 Oct 2022 · 1 repository · arXiv:2210.04191
-
ConTra: (Con)text (Tra)nsformer for Cross-Modal Video Retrieval 9 Oct 2022 · 1 repository · arXiv:2210.04341
-
Controllable Dialogue Simulation with In-Context Learning 9 Oct 2022 · 1 repository · arXiv:2210.04185
-
Deep Span Representations for Named Entity Recognition 9 Oct 2022 · 1 repository · arXiv:2210.04182
-
Fine-Grained Detection of Solidarity for Women and Migrants in 155 Years of German Parliamentary Debates 9 Oct 2022 · 2 repositories · arXiv:2210.04359
-
Fine-Tuning Pre-trained Transformers into Decaying Fast Weights 9 Oct 2022 · 1 repository · arXiv:2210.04243Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Better Pre-Training by Reducing Representation Confusion 9 Oct 2022 · 0 repositories · arXiv:2210.04246
-
KSAT: Knowledge-infused Self Attention Transformer -- Integrating Multiple Domain-Specific Contexts 9 Oct 2022 · 0 repositories · arXiv:2210.04307
-
Learning Texture Transformer Network for Light Field Super-Resolution 9 Oct 2022 · 0 repositories · arXiv:2210.09293
-
Spread Love Not Hate: Undermining the Importance of Hateful Pre-training for Hate Speech Detection 9 Oct 2022 · 1 repository · arXiv:2210.04267
-
Strong Gravitational Lensing Parameter Estimation with Vision Transformer 9 Oct 2022 · 1 repository · arXiv:2210.04143Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples)
-
Transformer-based Flood Scene Segmentation for Developing Countries 9 Oct 2022 · 0 repositories · arXiv:2210.04218