Methods › General › Normalization › Layer Normalization › Papers, page 181
Layer Normalization
Papers archive 2025-07-28
archive papers tagged: 24,980 · with a code link: 11,273 · where Syntology ran a sample: 3,471 (2,923 with a run with no instrument failure, 548 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,471 of 24,980 tagged: 2,923 with a run with no instrument failure, 548 where every run was a failure of Syntology's instrument)
Page 181 of 250: papers 18,001 to 18,100 of 24,980, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
From Natural Language to Simulations: Applying GPT-3 Codex to Automate Simulation Modeling of Logistics Systems 24 Feb 2022 · 1 repository · arXiv:2202.12107
-
Instantaneous Physiological Estimation using Video Transformers 24 Feb 2022 · 1 repository · arXiv:2202.12368
-
openFEAT: Improving Speaker Identification by Open-set Few-shot Embedding Adaptation with Transformer 24 Feb 2022 · 0 repositories · arXiv:2202.12349
-
Overcoming a Theoretical Limitation of Self-Attention 24 Feb 2022 · 1 repository · arXiv:2202.12172
-
Pretraining without Wordpieces: Learning Over a Vocabulary of Millions of Words 24 Feb 2022 · 0 repositories · arXiv:2202.12142
-
Probing BERT's priors with serial reproduction chains 24 Feb 2022 · 1 repository · arXiv:2202.12226Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample) · 1 pointer-only (licence)
-
Sky Computing: Accelerating Geo-distributed Computing in Federated Learning 24 Feb 2022 · 1 repository · arXiv:2202.11836
-
Transformers in Medical Image Analysis: A Review 24 Feb 2022 · 0 repositories · arXiv:2202.12165
-
TrimBERT: Tailoring BERT for Trade-offs 24 Feb 2022 · 0 repositories · arXiv:2202.12411
-
Using calibrator to improve robustness in Machine Reading Comprehension 24 Feb 2022 · 0 repositories · arXiv:2202.11865
-
A Differential Attention Fusion Model Based on Transformer for Time Series Forecasting 23 Feb 2022 · 0 repositories · arXiv:2202.11402
-
Consistent Dropout for Policy Gradient Reinforcement Learning 23 Feb 2022 · 0 repositories · arXiv:2202.11818
-
FastRPB: a Scalable Relative Positional Encoding for Long Sequence Tasks 23 Feb 2022 · 1 repository · arXiv:2202.11364
-
Integration of neural network and fuzzy logic decision making compared with bilayered neural network in the simulation of daily dew point temperature 23 Feb 2022 · 0 repositories · arXiv:2202.12256
-
Paying U-Attention to Textures: Multi-Stage Hourglass Vision Transformer for Universal Texture Synthesis 23 Feb 2022 · 0 repositories · arXiv:2202.11703
-
Refining the state-of-the-art in Machine Translation, optimizing NMT for the JA <-> EN language pair by leveraging personal domain expertise 23 Feb 2022 · 0 repositories · arXiv:2202.11669
-
Think Global, Act Local: Dual-scale Graph Transformer for Vision-and-Language Navigation 23 Feb 2022 · 1 repository · arXiv:2202.11742Syntology 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 1 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
A New Generation of Perspective API: Efficient Multilingual Character-level Transformers 22 Feb 2022 · 0 repositories · arXiv:2202.11176
-
GroupViT: Semantic Segmentation Emerges from Text Supervision 22 Feb 2022 · 6 repositories · arXiv:2202.11094
-
Improving CTC-based speech recognition via knowledge transferring from pre-trained language models 22 Feb 2022 · 1 repository · arXiv:2203.03582
-
JAMES: Normalizing Job Titles with Multi-Aspect Graph Embeddings and Reasoning 22 Feb 2022 · 0 repositories · arXiv:2202.10739
-
Learning Cluster Patterns for Abstractive Summarization 22 Feb 2022 · 0 repositories · arXiv:2202.10967
-
One-shot Scene Graph Generation 22 Feb 2022 · 1 repository · arXiv:2202.10824
-
Social Computational Design Method for Generating Product Shapes with GAN and Transformer Models 22 Feb 2022 · 0 repositories · arXiv:2202.10774
-
Socialformer: Social Network Inspired Long Document Modeling for Document Ranking 22 Feb 2022 · 1 repository · arXiv:2202.10870
-
Embarrassingly Simple Performance Prediction for Abductive Natural Language Inference 21 Feb 2022 · 1 repository · arXiv:2202.10408
-
Formal Analysis of the Sampling Behaviour of Stochastic Event-Triggered Control 21 Feb 2022 · 0 repositories · arXiv:2202.10178
-
Items from Psychometric Tests as Training Data for Personality Profiling Models of Twitter Users 21 Feb 2022 · 0 repositories · arXiv:2202.10415
-
Rethinking the Zigzag Flattening for Image Reading 21 Feb 2022 · 0 repositories · arXiv:2202.10240
-
S3T: Self-Supervised Pre-training with Swin Transformer for Music Classification 21 Feb 2022 · 1 repository · arXiv:2202.10139
-
ViTAEv2: Vision Transformer Advanced by Exploring Inductive Bias for Image Recognition and Beyond 21 Feb 2022 · 8 repositories · arXiv:2202.10108Syntology community repositories only · 22 ran (of which 8 constructed an object rather than computing a result; 16 with no instrument failure: 0 honoured, 0 violated, 16 with no contract checked; 6 where Syntology's instrument failed) · 3 unverified (of 25 harvested samples) · 4 pointer-only (licence)
-
ARM3D: Attention-based relation module for indoor 3D object detection 20 Feb 2022 · 1 repository · arXiv:2202.09715
-
Contextual Semantic Embeddings for Ontology Subsumption Prediction 20 Feb 2022 · 2 repositories · arXiv:2202.09791
-
Generalized Bayesian Additive Regression Trees Models: Beyond Conditional Conjugacy 20 Feb 2022 · 0 repositories · arXiv:2202.09924
-
Do Transformers know symbolic rules, and would we know if they did? 19 Feb 2022 · 0 repositories · arXiv:2203.00162
-
GCNET: graph-based prediction of stock price movement using graph convolutional network 19 Feb 2022 · 1 repository · arXiv:2203.11091
-
TransDreamer: Reinforcement Learning with Transformer World Models 19 Feb 2022 · 0 repositories · arXiv:2202.09481Syntology 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
A Survey of Vision-Language Pre-Trained Models 18 Feb 2022 · 0 repositories · arXiv:2202.10936
-
Evaluating the Construct Validity of Text Embeddings with Application to Survey Questions 18 Feb 2022 · 1 repository · arXiv:2202.09166
-
Mixture-of-Experts with Expert Choice Routing 18 Feb 2022 · 0 repositories · arXiv:2202.09368
-
Task Specific Attention is one more thing you need for object detection 18 Feb 2022 · 1 repository · arXiv:2202.09048
-
Unleashing the Power of Transformer for Graphs 18 Feb 2022 · 0 repositories · arXiv:2202.10581
-
Multi-Scale Hybrid Vision Transformer for Learning Gastric Histology: AI-Based Decision Support System for Gastric Cancer Treatment 17 Feb 2022 · 0 repositories · arXiv:2202.08510
-
ST-MoE: Designing Stable and Transferable Sparse Expert Models 17 Feb 2022 · 3 repositories · arXiv:2202.08906Syntology community repositories only · 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 3 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Graph Masked Autoencoders with Transformers 17 Feb 2022 · 1 repository · arXiv:2202.08391
-
Improving English to Sinhala Neural Machine Translation using Part-of-Speech Tag 17 Feb 2022 · 0 repositories · arXiv:2202.08882
-
Revisiting Over-smoothing in BERT from the Perspective of Graph 17 Feb 2022 · 0 repositories · arXiv:2202.08625
-
SGPT: GPT Sentence Embeddings for Semantic Search 17 Feb 2022 · 1 repository · arXiv:2202.08904Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Transformer for Graphs: An Overview from Architecture Perspective 17 Feb 2022 · 1 repository · arXiv:2202.08455Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
TraSeTR: Track-to-Segment Transformer with Contrastive Query for Instance-level Instrument Segmentation in Robotic Surgery 17 Feb 2022 · 0 repositories · arXiv:2202.08453
-
When BERT Meets Quantum Temporal Convolution Learning for Text Classification in Heterogeneous Computing 17 Feb 2022 · 0 repositories · arXiv:2203.03550
-
A Survey of Pretraining on Graphs: Taxonomy, Methods, and Applications 16 Feb 2022 · 3 repositories · arXiv:2202.07893
-
ActionFormer: Localizing Moments of Actions with Transformers 16 Feb 2022 · 1 repository · arXiv:2202.07925
-
EdgeFormer: A Parameter-Efficient Transformer for On-Device Seq2seq Generation 16 Feb 2022 · 1 repository · arXiv:2202.07959Syntology official: no sample here; runs from other or unrecorded repositories · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
No One Left Behind: Inclusive Federated Learning over Heterogeneous Devices 16 Feb 2022 · 0 repositories · arXiv:2202.08036
-
Probing Pretrained Models of Source Code 16 Feb 2022 · 1 repository · arXiv:2202.08975
-
The NLP Task Effectiveness of Long-Range Transformers 16 Feb 2022 · 0 repositories · arXiv:2202.07856
-
A Survey on Dynamic Neural Networks for Natural Language Processing 15 Feb 2022 · 0 repositories · arXiv:2202.07101
-
A Survey on Model Compression and Acceleration for Pretrained Language Models 15 Feb 2022 · 0 repositories · arXiv:2202.07105
-
BLUE at Memotion 2.0 2022: You have my Image, my Text and my Transformer 15 Feb 2022 · 0 repositories · arXiv:2202.07543
-
Defending against Reconstruction Attacks with Rényi Differential Privacy 15 Feb 2022 · 0 repositories · arXiv:2202.07623
-
One Configuration to Rule Them All? Towards Hyperparameter Transfer in Topic Models using Multi-Objective Bayesian Optimization 15 Feb 2022 · 1 repository · arXiv:2202.07631
-
Personalized Prompt Learning for Explainable Recommendation 15 Feb 2022 · 1 repository · arXiv:2202.07371
-
Predicting on the Edge: Identifying Where a Larger Model Does Better 15 Feb 2022 · 0 repositories · arXiv:2202.07652
-
Tomayto, Tomahto. Beyond Token-level Answer Equivalence for Question Answering Evaluation 15 Feb 2022 · 1 repository · arXiv:2202.07654
-
Toxic Comments Hunter : Score Severity of Toxic Comments 15 Feb 2022 · 0 repositories · arXiv:2203.03548
-
Transformers in Time Series: A Survey 15 Feb 2022 · 11 repositories · arXiv:2202.07125
-
ViNTER: Image Narrative Generation with Emotion-Arc-Aware Transformer 15 Feb 2022 · 0 repositories · arXiv:2202.07305
-
XAI for Transformers: Better Explanations through Conservative Propagation 15 Feb 2022 · 1 repository · arXiv:2202.07304Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 1 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 12 harvested samples)
-
CodeFill: Multi-token Code Completion by Jointly Learning from Structure and Naming Sequences 14 Feb 2022 · 1 repository · arXiv:2202.06689
-
Geometric Transformer for Fast and Robust Point Cloud Registration 14 Feb 2022 · 2 repositories · arXiv:2202.06688Syntology community repositories only · 4 ran (of which 4 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified; every one of the 4 samples that ran constructed an object rather than computing a result (of 9 harvested samples) · 9 pointer-only (licence)
-
Handcrafted Histological Transformer (H2T): Unsupervised Representation of Whole Slide Images 14 Feb 2022 · 1 repository · arXiv:2202.07001Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
Mixing and Shifting: Exploiting Global and Local Dependencies in Vision MLPs 14 Feb 2022 · 2 repositories · arXiv:2202.06510Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 1 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 10 harvested samples) · 6 pointer-only (licence)
-
Punctuation restoration in Swedish through fine-tuned KB-BERT 14 Feb 2022 · 0 repositories · arXiv:2202.06769
-
QA4QG: Using Question Answering to Constrain Multi-Hop Question Generation 14 Feb 2022 · 1 repository · arXiv:2202.06538
-
Sequence-to-Sequence Resources for Catalan 14 Feb 2022 · 1 repository · arXiv:2202.06871
-
Source Code Summarization with Structural Relative Position Guided Transformer 14 Feb 2022 · 1 repository · arXiv:2202.06521
-
Transformer Memory as a Differentiable Search Index 14 Feb 2022 · 1 repository · arXiv:2202.06991
-
UserBERT: Modeling Long- and Short-Term User Preferences via Self-Supervision 14 Feb 2022 · 0 repositories · arXiv:2202.07605
-
What Do They Capture? -- A Structural Analysis of Pre-Trained Language Models for Source Code 14 Feb 2022 · 1 repository · arXiv:2202.06840
-
Assessment of contextualised representations in detecting outcome phrases in clinical trials 13 Feb 2022 · 0 repositories · arXiv:2203.03547
-
BViT: Broad Attention based Vision Transformer 13 Feb 2022 · 1 repository · arXiv:2202.06268
-
ET-BERT: A Contextualized Datagram Representation with Pre-training Transformers for Encrypted Traffic Classification 13 Feb 2022 · 1 repository · arXiv:2202.06335
-
LighTN: Light-weight Transformer Network for Performance-overhead Tradeoff in Point Cloud Downsampling 13 Feb 2022 · 0 repositories · arXiv:2202.06263
-
LMN at SemEval-2022 Task 11: A Transformer-based System for English Named Entity Recognition 13 Feb 2022 · 0 repositories · arXiv:2203.03546
-
A multi-task semi-supervised framework for Text2Graph & Graph2Text 12 Feb 2022 · 1 repository · arXiv:2202.06041
-
Automatic Issue Classifier: A Transfer Learning Framework for Classifying Issue Reports 12 Feb 2022 · 1 repository · arXiv:2202.06149
-
Benchmark Assessment for DeepSpeed Optimization Library 12 Feb 2022 · 0 repositories · arXiv:2202.12831
-
Maximizing Communication Efficiency for Large-scale Training via 0/1 Adam 12 Feb 2022 · 1 repository · arXiv:2202.06009
-
Multi-direction and Multi-scale Pyramid in Transformer for Video-based Pedestrian Retrieval 12 Feb 2022 · 1 repository · arXiv:2202.06014
-
Multi-direction and Multi-scale Pyramid in Transformer for Video-based Pedestrian Retrieval 12 Feb 2022 · 1 repository
-
HaT5: Hate Language Identification using Text-to-Text Transfer Transformer 11 Feb 2022 · 0 repositories · arXiv:2202.05690
-
Including Facial Expressions in Contextual Embeddings for Sign Language Generation 11 Feb 2022 · 0 repositories · arXiv:2202.05383
-
White-Box Attacks on Hate-speech BERT Classifiers in German with Explicit and Implicit Character Level Defense 11 Feb 2022 · 1 repository · arXiv:2202.05778
-
A Multi-task Learning Framework for Product Ranking with BERT 10 Feb 2022 · 0 repositories · arXiv:2202.05317
-
AA-TransUNet: Attention Augmented TransUNet For Nowcasting Tasks 10 Feb 2022 · 1 repository · arXiv:2202.04996
-
Slovene SuperGLUE Benchmark: Translation and Evaluation 10 Feb 2022 · 0 repositories · arXiv:2202.04994
-
Can Open Domain Question Answering Systems Answer Visual Knowledge Questions? 9 Feb 2022 · 0 repositories · arXiv:2202.04306
-
Social Media as an Instant Source of Feedback on Water Quality 9 Feb 2022 · 0 repositories · arXiv:2202.04462
-
pNLP-Mixer: an Efficient all-MLP Architecture for Language 9 Feb 2022 · 1 repository · arXiv:2202.04350