Methods › General › Normalization › Layer Normalization › Papers, page 157
Layer Normalization
Papers archive 2025-07-28
archive papers tagged: 24,980 · with a code link: 11,273 · where Syntology ran a sample: 3,471 (2,923 with a run with no instrument failure, 548 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,471 of 24,980 tagged: 2,923 with a run with no instrument failure, 548 where every run was a failure of Syntology's instrument)
Page 157 of 250: papers 15,601 to 15,700 of 24,980, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Pure Transformer with Integrated Experts for Scene Text Recognition 9 Nov 2022 · 0 repositories · arXiv:2211.04963
-
Sentiment Analysis of Persian Language: Review of Algorithms, Approaches and Datasets 9 Nov 2022 · 0 repositories · arXiv:2212.06041
-
SG-Shuffle: Multi-aspect Shuffle Transformer for Scene Graph Generation 9 Nov 2022 · 0 repositories · arXiv:2211.04773
-
Towards Reasoning-Aware Explainable VQA 9 Nov 2022 · 0 repositories · arXiv:2211.05190
-
Training a Vision Transformer from scratch in less than 24 hours with 1 GPU 9 Nov 2022 · 1 repository · arXiv:2211.05187Syntology 4 ran (of which 4 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified; every one of the 4 samples that ran constructed an object rather than computing a result (of 6 harvested samples) · 6 pointer-only (licence)
-
Transformers Meet Small Datasets 9 Nov 2022 · 0 repositories
-
ViTALiTy: Unifying Low-rank and Sparse Approximation for Vision Transformer Acceleration with a Linear Taylor Attention 9 Nov 2022 · 1 repository · arXiv:2211.05109Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 1 honoured, 0 violated, 10 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 14 harvested samples) · 14 pointer-only (licence)
-
A Multimodal Approach for Dementia Detection from Spontaneous Speech with Tensor Fusion Layer 8 Nov 2022 · 0 repositories · arXiv:2211.04368
-
flexBART: Flexible Bayesian regression trees with categorical predictors 8 Nov 2022 · 1 repository · arXiv:2211.04459
-
Active Example Selection for In-Context Learning 8 Nov 2022 · 1 repository · arXiv:2211.04486Syntology official (archive's flag): 10 ran · 10 ran (of which 3 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 0 violated, 3 with no contract checked; 6 where Syntology's instrument failed) · 6 unverified (of 16 harvested samples)
-
Conciseness: An Overlooked Language Task 8 Nov 2022 · 0 repositories · arXiv:2211.04126
-
Discover, Explanation, Improvement: An Automatic Slice Detection Framework for Natural Language Processing 8 Nov 2022 · 0 repositories · arXiv:2211.04476
-
Linear Self-Attention Approximation via Trainable Feedforward Kernel 8 Nov 2022 · 0 repositories · arXiv:2211.04076
-
Pushing the limits of self-supervised speaker verification using regularized distillation framework 8 Nov 2022 · 1 repository · arXiv:2211.04168
-
SimOn: A Simple Framework for Online Temporal Action Localization 8 Nov 2022 · 1 repository · arXiv:2211.04905
-
Splitting expands the application range of Vision Transformer -- variable Vision Transformer (vViT) 8 Nov 2022 · 0 repositories · arXiv:2211.03992
-
AD-BERT: Using Pre-trained contextualized embeddings to Predict the Progression from Mild Cognitive Impairment to Alzheimer's Disease 7 Nov 2022 · 0 repositories · arXiv:2212.06042
-
Retrieval augmentation of large language models for lay language generation 7 Nov 2022 · 1 repository · arXiv:2211.03818Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 7 harvested samples)
-
CoNMix for Source-free Single and Multi-target Domain Adaptation 7 Nov 2022 · 1 repository · arXiv:2211.03876Syntology official: harvested, nothing ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Group DETR v2: Strong Object Detector with Encoder-Decoder Pretraining 7 Nov 2022 · 0 repositories · arXiv:2211.03594
-
How Much Does Attention Actually Attend? Questioning the Importance of Attention in Pretrained Transformers 7 Nov 2022 · 1 repository · arXiv:2211.03495
-
Sequential Transformer for End-to-End Person Search 6 Nov 2022 · 0 repositories · arXiv:2211.04323
-
Suffix Retrieval-Augmented Language Modeling 6 Nov 2022 · 1 repository · arXiv:2211.03053
-
Wall Street Tree Search: Risk-Aware Planning for Offline Reinforcement Learning 6 Nov 2022 · 0 repositories · arXiv:2211.04583
-
Inductive Graph Transformer for Delivery Time Estimation 5 Nov 2022 · 1 repository · arXiv:2211.02863
-
Learning to Infer from Unlabeled Data: A Semi-supervised Learning Approach for Robust Natural Language Inference 5 Nov 2022 · 1 repository · arXiv:2211.02971
-
A Transformer Architecture for Online Gesture Recognition of Mathematical Expressions 4 Nov 2022 · 0 repositories · arXiv:2211.02643
-
A Weakly-Supervised Streaming Multilingual Speech Model with Truly Zero-Shot Capability 4 Nov 2022 · 0 repositories · arXiv:2211.02499
-
BERT-Deep CNN: State-of-the-Art for Sentiment Analysis of COVID-19 Tweets 4 Nov 2022 · 0 repositories · arXiv:2211.09733
-
BERT for Long Documents: A Case Study of Automated ICD Coding 4 Nov 2022 · 0 repositories · arXiv:2211.02519
-
CCATMos: Convolutional Context-aware Transformer Network for Non-intrusive Speech Quality Assessment 4 Nov 2022 · 0 repositories · arXiv:2211.02577
-
Continuous Prompt Tuning Based Textual Entailment Model for E-commerce Entity Typing 4 Nov 2022 · 1 repository · arXiv:2211.02483
-
Deep learning for structural health monitoring: An application to heritage structures 4 Nov 2022 · 0 repositories · arXiv:2211.10351
-
Fraudulent User Detection Via Behavior Information Aggregation Network (BIAN) On Large-Scale Financial Social Network 4 Nov 2022 · 0 repositories · arXiv:2211.06315
-
Generation of Chinese classical poetry based on pre-trained model 4 Nov 2022 · 0 repositories · arXiv:2211.02541
-
OSIC: A New One-Stage Image Captioner Coined 4 Nov 2022 · 0 repositories · arXiv:2211.02321
-
Patch DCT vs LeNet 4 Nov 2022 · 0 repositories · arXiv:2211.02392
-
RCDPT: Radar-Camera fusion Dense Prediction Transformer 4 Nov 2022 · 1 repository · arXiv:2211.02432
-
Real-Time Target Sound Extraction 4 Nov 2022 · 1 repository · arXiv:2211.02250Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples)
-
SPEAKER VGG CCT: Cross-corpus Speech Emotion Recognition with Speaker Embedding and Vision Transformers 4 Nov 2022 · 1 repository · arXiv:2211.02366
-
Alternative formulations for gilthead seabream diets: towards a more sustainable production 3 Nov 2022 · 0 repositories · arXiv:2211.02430
-
Channel-Aware Pretraining of Joint Encoder-Decoder Self-Supervised Model for Telephonic-Speech ASR 3 Nov 2022 · 0 repositories · arXiv:2211.01669
-
Crosslingual Generalization through Multitask Finetuning 3 Nov 2022 · 1 repository · arXiv:2211.01786Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
Evaluating a Synthetic Image Dataset Generated with Stable Diffusion 3 Nov 2022 · 0 repositories · arXiv:2211.01777
-
Transformers on Multilingual Clause-Level Morphology 3 Nov 2022 · 1 repository · arXiv:2211.01736
-
FedTP: Federated Learning by Transformer Personalization 3 Nov 2022 · 1 repository · arXiv:2211.01572Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 1 honoured, 2 violated, 2 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 10 harvested samples) · 1 pointer-only (licence)
-
Fine-Tuning Language Models via Epistemic Neural Networks 3 Nov 2022 · 1 repository · arXiv:2211.01568Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Pangu-Weather: A 3D High-Resolution Model for Fast and Accurate Global Weather Forecast 3 Nov 2022 · 6 repositories · arXiv:2211.02556
-
PolyBuilding: Polygon Transformer for End-to-End Building Extraction 3 Nov 2022 · 0 repositories · arXiv:2211.01589
-
Rethinking Hierarchies in Pre-trained Plain Vision Transformer 3 Nov 2022 · 0 repositories · arXiv:2211.01785
-
SAP-DETR: Bridging the Gap Between Salient Points and Queries-Based Transformer Detector for Fast Model Convergency 3 Nov 2022 · 1 repository · arXiv:2211.02006
-
Scaling Multimodal Pre-Training via Cross-Modality Gradient Harmonization 3 Nov 2022 · 0 repositories · arXiv:2211.02077
-
Using Large Pre-Trained Language Model to Assist FDA in Premarket Medical Device 3 Nov 2022 · 0 repositories · arXiv:2212.01217
-
Attention-based Neural Cellular Automata 2 Nov 2022 · 0 repositories · arXiv:2211.01233
-
BECTRA: Transducer-based End-to-End ASR with BERT-Enhanced Encoder 2 Nov 2022 · 0 repositories · arXiv:2211.00792
-
eDiff-I: Text-to-Image Diffusion Models with an Ensemble of Expert Denoisers 2 Nov 2022 · 2 repositories · arXiv:2211.01324Syntology 11 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 1 honoured, 0 violated, 9 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 13 harvested samples)
-
MAST: Multiscale Audio Spectrogram Transformers 2 Nov 2022 · 1 repository · arXiv:2211.01515
-
MPCFormer: fast, performant and private Transformer inference with MPC 2 Nov 2022 · 1 repository · arXiv:2211.01452
-
Multi-level Distillation of Semantic Knowledge for Pre-training Multilingual Language Model 2 Nov 2022 · 0 repositories · arXiv:2211.01200
-
Pop2Piano : Pop Audio-based Piano Cover Generation 2 Nov 2022 · 4 repositories · arXiv:2211.00895Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Processing Long Legal Documents with Pre-trained Transformers: Modding LegalBERT and Longformer 2 Nov 2022 · 0 repositories · arXiv:2211.00974
-
RegCLR: A Self-Supervised Framework for Tabular Representation Learning in the Wild 2 Nov 2022 · 0 repositories · arXiv:2211.01165
-
Transformer-based encoder-encoder architecture for Spoken Term Detection 2 Nov 2022 · 0 repositories · arXiv:2211.01089
-
WITT: A Wireless Image Transmission Transformer for Semantic Communications 2 Nov 2022 · 2 repositories · arXiv:2211.00937
-
ClassActionPrediction: A Challenging Benchmark for Legal Judgment Prediction of Class Action Cases in the US 1 Nov 2022 · 1 repository · arXiv:2211.00582
-
FRSUM: Towards Faithful Abstractive Summarization via Enhancing Factual Robustness 1 Nov 2022 · 0 repositories · arXiv:2211.00294
-
Interpretability in the Wild: a Circuit for Indirect Object Identification in GPT-2 small 1 Nov 2022 · 7 repositories · arXiv:2211.00593Syntology community repositories only · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 13 harvested samples)
-
Investigating Content-Aware Neural Text-To-Speech MOS Prediction Using Prosodic and Linguistic Features 1 Nov 2022 · 0 repositories · arXiv:2211.00342
-
An Empirical Study on Data Leakage and Generalizability of Link Prediction Models for Issues and Commits 1 Nov 2022 · 0 repositories · arXiv:2211.00381
-
Two-stage LLM Fine-tuning with Less Specialization and More Generalization 1 Nov 2022 · 0 repositories · arXiv:2211.00635
-
Reduce, Reuse, Recycle: Improving Training Efficiency with Distillation 1 Nov 2022 · 0 repositories · arXiv:2211.00683
-
T5lephone: Bridging Speech and Text Self-supervised Models for Spoken Language Understanding via Phoneme level T5 1 Nov 2022 · 1 repository · arXiv:2211.00586
-
Text-Only Training for Image Captioning using Noise-Injected CLIP 1 Nov 2022 · 4 repositories · arXiv:2211.00575Syntology official (archive's flag): 2 ran · 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 1 pointer-only (licence)
-
VID-Trans-ReID: Enhanced Video Transformers for Person Re-identification 1 Nov 2022 · 1 repository
-
ViT-DeiT: An Ensemble Model for Breast Cancer Histopathological Images Classification 1 Nov 2022 · 0 repositories · arXiv:2211.00749
-
AdaMix: Mixture-of-Adaptations for Parameter-efficient Model Tuning 31 Oct 2022 · 1 repository · arXiv:2210.17451Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Controllable Factuality in Document-Grounded Dialog Systems Using a Noisy Channel Model 31 Oct 2022 · 1 repository · arXiv:2210.17418
-
Efficient Document Retrieval by End-to-End Refining and Quantizing BERT Embedding with Contrastive Product Quantization 31 Oct 2022 · 1 repository · arXiv:2210.17170
-
GPTQ: Accurate Post-Training Quantization for Generative Pre-trained Transformers 31 Oct 2022 · 17 repositories · arXiv:2210.17323Syntology official (archive's flag): 1 ran · 5 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 3 where Syntology's instrument failed) · 10 unverified (of 15 harvested samples) · 1 pointer-only (licence)
-
Joint Audio/Text Training for Transformer Rescorer of Streaming Speech Recognition 31 Oct 2022 · 0 repositories · arXiv:2211.00174
-
Leveraging Pre-trained Models for Failure Analysis Triplets Generation 31 Oct 2022 · 0 repositories · arXiv:2210.17497
-
Multi-Camera Calibration Free BEV Representation for 3D Object Detection 31 Oct 2022 · 0 repositories · arXiv:2210.17252
-
Probabilistic Decomposition Transformer for Time Series Forecasting 31 Oct 2022 · 1 repository · arXiv:2210.17393
-
QNet: A Quantum-native Sequence Encoder Architecture 31 Oct 2022 · 0 repositories · arXiv:2210.17262
-
QuaLA-MiniLM: a Quantized Length Adaptive MiniLM 31 Oct 2022 · 2 repositories · arXiv:2210.17114
-
SDCL: Self-Distillation Contrastive Learning for Chinese Spell Checking 31 Oct 2022 · 0 repositories · arXiv:2210.17168
-
Spatial-Temporal Synchronous Graph Transformer network (STSGT) for COVID-19 forecasting 31 Oct 2022 · 1 repository · arXiv:2211.00082
-
SSD-LM: Semi-autoregressive Simplex-based Diffusion Language Model for Text Generation and Modular Control 31 Oct 2022 · 2 repositories · arXiv:2210.17432
-
Structured State Space Decoder for Speech Recognition and Synthesis 31 Oct 2022 · 0 repositories · arXiv:2210.17098
-
Towards Zero-Shot and Few-Shot Table Question Answering using GPT-3 31 Oct 2022 · 0 repositories · arXiv:2210.17284
-
ViT-LSLA: Vision Transformer with Light Self-Limited-Attention 31 Oct 2022 · 0 repositories · arXiv:2210.17115
-
An Efficient Memory-Augmented Transformer for Knowledge-Intensive NLP Tasks 30 Oct 2022 · 1 repository · arXiv:2210.16773
-
Attention Swin U-Net: Cross-Contextual Attention Mechanism for Skin Lesion Segmentation 30 Oct 2022 · 1 repository · arXiv:2210.16898
-
Exemplar Guided Deep Neural Network for Spatial Transcriptomics Analysis of Gene Expression Prediction 30 Oct 2022 · 1 repository · arXiv:2210.16721
-
Foreign Object Debris Detection for Airport Pavement Images based on Self-supervised Localization and Vision Transformer 30 Oct 2022 · 1 repository · arXiv:2210.16901
-
Learning to Decompose: Hypothetical Question Decomposition Based on Comparable Texts 30 Oct 2022 · 0 repositories · arXiv:2210.16865
-
Multi-view Multi-label Anomaly Network Traffic Classification based on MLP-Mixer Neural Network 30 Oct 2022 · 0 repositories · arXiv:2210.16719
-
Parameter-Efficient Tuning Makes a Good Classification Head 30 Oct 2022 · 1 repository · arXiv:2210.16771
-
QuEst: Graph Transformer for Quantum Circuit Reliability Estimation 30 Oct 2022 · 1 repository · arXiv:2210.16724
-
Time-rEversed diffusioN tEnsor Transformer: A new TENET of Few-Shot Object Detection 30 Oct 2022 · 0 repositories · arXiv:2210.16897