Methods › General › Normalization › Layer Normalization › Papers, page 131
Layer Normalization
Papers archive 2025-07-28
archive papers tagged: 24,980 · with a code link: 11,273 · where Syntology ran a sample: 3,471 (2,923 with a run with no instrument failure, 548 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,471 of 24,980 tagged: 2,923 with a run with no instrument failure, 548 where every run was a failure of Syntology's instrument)
Page 131 of 250: papers 13,001 to 13,100 of 24,980, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
InvPT++: Inverted Pyramid Multi-Task Transformer for Visual Scene Understanding 8 Jun 2023 · 1 repository · arXiv:2306.04842
-
Large-scale Dataset Pruning with Dynamic Uncertainty 8 Jun 2023 · 2 repositories · arXiv:2306.05175Syntology official (archive's flag): 10 ran · 11 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 1 honoured, 0 violated, 7 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 11 harvested samples) · 7 pointer-only (licence)
-
Leveraging Language Identification to Enhance Code-Mixed Text Classification 8 Jun 2023 · 0 repositories · arXiv:2306.04964
-
M3Exam: A Multilingual, Multimodal, Multilevel Benchmark for Examining Large Language Models 8 Jun 2023 · 1 repository · arXiv:2306.05179
-
Mapping the Challenges of HCI: An Application and Evaluation of ChatGPT and GPT-4 for Mining Insights at Scale 8 Jun 2023 · 0 repositories · arXiv:2306.05036
-
A Task-driven Network for Mesh Classification and Semantic Part Segmentation 8 Jun 2023 · 0 repositories · arXiv:2306.05246
-
Mixture-of-Domain-Adapters: Decoupling and Injecting Domain Knowledge to Pre-trained Language Models Memories 8 Jun 2023 · 1 repository · arXiv:2306.05406Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample)
-
Mixture-of-Supernets: Improving Weight-Sharing Supernet Training with Architecture-Routed Mixture-of-Experts 8 Jun 2023 · 1 repository · arXiv:2306.04845
-
Multi-level Multiple Instance Learning with Transformer for Whole Slide Image Classification 8 Jun 2023 · 1 repository · arXiv:2306.05029
-
Multi-Scale And Token Mergence: Make Your ViT More Efficient 8 Jun 2023 · 0 repositories · arXiv:2306.04897
-
NOWJ at COLIEE 2023 -- Multi-Task and Ensemble Approaches in Legal Information Processing 8 Jun 2023 · 0 repositories · arXiv:2306.04903
-
PandaLM: An Automatic Evaluation Benchmark for LLM Instruction Tuning Optimization 8 Jun 2023 · 2 repositories · arXiv:2306.05087Syntology official (archive's flag): 2 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Point-LGMask: Local and Global Contexts Embedding for Point Cloud Pre-training with Multi-Ratio Masking 8 Jun 2023 · 1 repository
-
Prefer to Classify: Improving Text Classifiers via Auxiliary Preference Learning 8 Jun 2023 · 1 repository · arXiv:2306.04925Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample) · 1 pointer-only (licence)
-
Rate Forecaster based Energy Aware Band Assignment in Multiband Networks 8 Jun 2023 · 0 repositories · arXiv:2306.05369
-
The ADAIO System at the BEA-2023 Shared Task on Generating AI Teacher Responses in Educational Dialogues 8 Jun 2023 · 0 repositories · arXiv:2306.05360
-
ToolAlpaca: Generalized Tool Learning for Language Models with 3000 Simulated Cases 8 Jun 2023 · 3 repositories · arXiv:2306.05301Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 8 harvested samples)
-
TRIGS: Trojan Identification from Gradient-based Signatures 8 Jun 2023 · 1 repository · arXiv:2306.04877
-
Object Detection with Transformers: A Review 7 Jun 2023 · 2 repositories · arXiv:2306.04670
-
On the Detectability of ChatGPT Content: Benchmarking, Methodology, and Evaluation through the Lens of Academic Writing 7 Jun 2023 · 2 repositories · arXiv:2306.05524
-
Examining Bias in Opinion Summarisation Through the Perspective of Opinion Diversity 7 Jun 2023 · 0 repositories · arXiv:2306.04424
-
Good Data, Large Data, or No Data? Comparing Three Approaches in Developing Research Aspect Classifiers for Biomedical Papers 7 Jun 2023 · 1 repository · arXiv:2306.04820
-
GPT Self-Supervision for a Better Data Annotator 7 Jun 2023 · 0 repositories · arXiv:2306.04349
-
How Far Can Camels Go? Exploring the State of Instruction Tuning on Open Resources 7 Jun 2023 · 4 repositories · arXiv:2306.04751Syntology official (archive's flag): 5 ran · 6 ran (of which 1 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 0 violated, 1 with no contract checked; 4 where Syntology's instrument failed) · 6 unverified (of 12 harvested samples) · 2 pointer-only (licence)
-
INSTRUCTEVAL: Towards Holistic Evaluation of Instruction-Tuned Large Language Models 7 Jun 2023 · 2 repositories · arXiv:2306.04757Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Normalization Layers Are All That Sharpness-Aware Minimization Needs 7 Jun 2023 · 1 repository · arXiv:2306.04226
-
Personality testing of Large Language Models: Limited temporal stability, but highlighted prosociality 7 Jun 2023 · 0 repositories · arXiv:2306.04308
-
ScienceBenchmark: A Complex Real-World Benchmark for Evaluating Natural Language to SQL Systems 7 Jun 2023 · 0 repositories · arXiv:2306.04743
-
Self-supervised Audio Teacher-Student Transformer for Both Clip-level and Frame-level Tasks 7 Jun 2023 · 2 repositories · arXiv:2306.04186
-
TEC-Net: Vision Transformer Embrace Convolutional Neural Networks for Medical Image Segmentation 7 Jun 2023 · 1 repository · arXiv:2306.04086
-
The Two Word Test: A Semantic Benchmark for Large Language Models 7 Jun 2023 · 1 repository · arXiv:2306.04610
-
An Empirical Analysis of Parameter-Efficient Methods for Debiasing Pre-Trained Language Models 6 Jun 2023 · 1 repository · arXiv:2306.04067
-
Certified Deductive Reasoning with Language Models 6 Jun 2023 · 1 repository · arXiv:2306.04031Syntology 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
GCD-DDPM: A Generative Change Detection Model Based on Difference-Feature Guided DDPM 6 Jun 2023 · 1 repository · arXiv:2306.03424
-
CiT-Net: Convolutional Neural Networks Hand in Hand with Vision Transformers for Medical Image Segmentation 6 Jun 2023 · 1 repository · arXiv:2306.03373Syntology official (archive's flag): 15 ran · 15 ran (of which 15 constructed an object rather than computing a result; 15 with no instrument failure: 0 honoured, 0 violated, 15 with no contract checked; 0 where Syntology's instrument failed) · 8 unverified; every one of the 15 samples that ran constructed an object rather than computing a result (of 23 harvested samples) · 23 pointer-only (licence)
-
LESS: Label-efficient Multi-scale Learning for Cytological Whole Slide Image Screening 6 Jun 2023 · 0 repositories · arXiv:2306.03407
-
DenseDINO: Boosting Dense Self-Supervised Learning with Token-Based Point-Level Consistency 6 Jun 2023 · 0 repositories · arXiv:2306.04654
-
Detecting Human Rights Violations on Social Media during Russia-Ukraine War 6 Jun 2023 · 0 repositories · arXiv:2306.05370
-
Industrial Anomaly Detection and Localization Using Weakly-Supervised Residual Transformers 6 Jun 2023 · 0 repositories · arXiv:2306.03492
-
Emergent Correspondence from Image Diffusion 6 Jun 2023 · 2 repositories · arXiv:2306.03881Syntology 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples) · 2 pointer-only (licence)
-
Human-imperceptible, Machine-recognizable Images 6 Jun 2023 · 1 repository · arXiv:2306.03679
-
Iterative Translation Refinement with Large Language Models 6 Jun 2023 · 0 repositories · arXiv:2306.03856
-
Language acquisition: do children and language models follow similar learning stages? 6 Jun 2023 · 0 repositories · arXiv:2306.03586
-
LEACE: Perfect linear concept erasure in closed form 6 Jun 2023 · 2 repositories · arXiv:2306.03819Syntology official (archive's flag): 3 ran · 12 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 1 honoured, 0 violated, 9 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 14 harvested samples)
-
Long term 5G network traffic forecasting via modeling non-stationarity with deep learning 6 Jun 2023 · 1 repository
-
On the Difference of BERT-style and CLIP-style Text Encoders 6 Jun 2023 · 1 repository · arXiv:2306.03678
-
PGformer: Proxy-Bridged Game Transformer for Multi-Person Highly Interactive Extreme Motion Prediction 6 Jun 2023 · 0 repositories · arXiv:2306.03374
-
SGAT4PASS: Spherical Geometry-Aware Transformer for PAnoramic Semantic Segmentation 6 Jun 2023 · 1 repository · arXiv:2306.03403Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 1 honoured, 1 violated, 3 with no contract checked; 4 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
TextFormer: A Query-based End-to-End Text Spotter with Mixed Supervision 6 Jun 2023 · 0 repositories · arXiv:2306.03377
-
Triggering Multi-Hop Reasoning for Question Answering in Language Models using Soft Prompts and Random Walks 6 Jun 2023 · 0 repositories · arXiv:2306.04009
-
PEARL: Zero-shot Cross-task Preference Alignment and Robust Reward Learning for Robotic Manipulation 6 Jun 2023 · 0 repositories · arXiv:2306.03615
-
A Vessel-Segmentation-Based CycleGAN for Unpaired Multi-modal Retinal Image Synthesis 5 Jun 2023 · 0 repositories · arXiv:2306.02901
-
Analyzing Syntactic Generalization Capacity of Pre-trained Language Models on Japanese Honorific Conversion 5 Jun 2023 · 0 repositories · arXiv:2306.03055
-
Benchmarking Large Language Models on CMExam -- A Comprehensive Chinese Medical Exam Dataset 5 Jun 2023 · 1 repository · arXiv:2306.03030Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
ChatGPT as a mapping assistant: A novel method to enrich maps with generative AI and content derived from street-level photographs 5 Jun 2023 · 0 repositories · arXiv:2306.03204
-
COMET: Learning Cardinality Constrained Mixture of Experts with Trees and Local Search 5 Jun 2023 · 2 repositories · arXiv:2306.02824
-
Cross-Drone Transformer Network for Robust Single Object Tracking 5 Jun 2023 · 1 repository
-
CTRL: Connect Collaborative and Language Model for CTR Prediction 5 Jun 2023 · 0 repositories · arXiv:2306.02841
-
Efficient GPT Model Pre-training using Tensor Train Matrix Representation 5 Jun 2023 · 0 repositories · arXiv:2306.02697
-
On "Scientific Debt" in NLP: A Case for More Rigour in Language Model Pre-Training Research 5 Jun 2023 · 0 repositories · arXiv:2306.02870
-
Orca: Progressive Learning from Complex Explanation Traces of GPT-4 5 Jun 2023 · 4 repositories · arXiv:2306.02707
-
Permutation Decision Trees 5 Jun 2023 · 0 repositories · arXiv:2306.02617
-
SelfEvolve: A Code Evolution Framework via Large Language Models 5 Jun 2023 · 0 repositories · arXiv:2306.02907
-
Equity-Transformer: Solving NP-hard Min-Max Routing Problems as Sequential Generation with Equity Context 5 Jun 2023 · 2 repositories · arXiv:2306.02689Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Skill over Scale: The Case for Medium, Domain-Specific Models for SE 5 Jun 2023 · 0 repositories · arXiv:2306.03268
-
Transformer-Based UNet with Multi-Headed Cross-Attention Skip Connections to Eliminate Artifacts in Scanned Documents 5 Jun 2023 · 0 repositories · arXiv:2306.02815
-
Using Sequences of Life-events to Predict Human Lives 5 Jun 2023 · 2 repositories · arXiv:2306.03009
-
A Mathematical Abstraction for Balancing the Trade-off Between Creativity and Reality in Large Language Models 4 Jun 2023 · 0 repositories · arXiv:2306.02295
-
Auto-GPT for Online Decision Making: Benchmarks and Additional Opinions 4 Jun 2023 · 1 repository · arXiv:2306.02224Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 5 unverified (of 11 harvested samples) · 4 pointer-only (licence)
-
Detector Guidance for Multi-Object Text-to-Image Generation 4 Jun 2023 · 1 repository · arXiv:2306.02236
-
Graph Transformer for Recommendation 4 Jun 2023 · 1 repository · arXiv:2306.02330Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Modular Transformers: Compressing Transformers into Modularized Layers for Flexible Efficient Inference 4 Jun 2023 · 0 repositories · arXiv:2306.02379
-
SpellMapper: A non-autoregressive neural spellchecker for ASR customization with candidate retrieval based on n-gram mappings 4 Jun 2023 · 1 repository · arXiv:2306.02317
-
Financial sentiment analysis using FinBERT with application in predicting stock movement 3 Jun 2023 · 0 repositories · arXiv:2306.02136
-
GAT-GAN : A Graph-Attention-based Time-Series Generative Adversarial Network 3 Jun 2023 · 0 repositories · arXiv:2306.01999
-
Lightweight Structure-aware Transformer Network for VHR Remote Sensing Image Change Detection 3 Jun 2023 · 0 repositories · arXiv:2306.01988
-
Memorization Capacity of Multi-Head Attention in Transformers 3 Jun 2023 · 1 repository · arXiv:2306.02010
-
MultiLegalPile: A 689GB Multilingual Legal Corpus 3 Jun 2023 · 0 repositories · arXiv:2306.02069
-
Question-Context Alignment and Answer-Context Dependencies for Effective Answer Sentence Selection 3 Jun 2023 · 0 repositories · arXiv:2306.02196
-
Towards Coding Social Science Datasets with Language Models 3 Jun 2023 · 0 repositories · arXiv:2306.02177
-
TransRUPNet for Improved Polyp Segmentation 3 Jun 2023 · 1 repository · arXiv:2306.02176
-
Unsupervised Low Light Image Enhancement Using SNR-Aware Swin Transformer 3 Jun 2023 · 0 repositories · arXiv:2306.02082
-
5IDER: Unified Query Rewriting for Steering, Intent Carryover, Disfluencies, Entity Carryover and Repair 2 Jun 2023 · 0 repositories · arXiv:2306.01855
-
A Novel Vision Transformer with Residual in Self-attention for Biomedical Image Classification 2 Jun 2023 · 0 repositories · arXiv:2306.01594
-
MathChat: Converse to Tackle Challenging Math Problems with LLM Agents 2 Jun 2023 · 2 repositories · arXiv:2306.01337
-
Binary and Ternary Natural Language Generation 2 Jun 2023 · 1 repository · arXiv:2306.01841Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Can Contextual Biasing Remain Effective with Whisper and GPT-2? 2 Jun 2023 · 1 repository · arXiv:2306.01942
-
Can LLMs like GPT-4 outperform traditional AI tools in dementia diagnosis? Maybe, but not today 2 Jun 2023 · 0 repositories · arXiv:2306.01499
-
Concurrent Classifier Error Detection (CCED) in Large Scale Machine Learning Systems 2 Jun 2023 · 0 repositories · arXiv:2306.01820
-
Establishment of NLP-Based Greenwashing Pattern Detection Service 2 Jun 2023 · 0 repositories
-
Evaluating Language Models for Mathematics through Interactions 2 Jun 2023 · 1 repository · arXiv:2306.01694Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Explainability of Speech Recognition Transformers via Gradient-based Attention Visualization 2 Jun 2023 · 1 repository
-
Context selectivity with dynamic availability enables lifelong continual learning 2 Jun 2023 · 1 repository · arXiv:2306.01690
-
Generalist Equivariant Transformer Towards 3D Molecular Interaction Learning 2 Jun 2023 · 1 repository · arXiv:2306.01474Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
Improved Training for End-to-End Streaming Automatic Speech Recognition Model with Punctuation 2 Jun 2023 · 0 repositories · arXiv:2306.01296
-
Understanding MLP-Mixer as a Wide and Sparse MLP 2 Jun 2023 · 0 repositories · arXiv:2306.01470
-
nnMobileNet: Rethinking CNN for Retinopathy Research 2 Jun 2023 · 2 repositories · arXiv:2306.01289
-
RITA: Group Attention is All You Need for Timeseries Analytics 2 Jun 2023 · 0 repositories · arXiv:2306.01926
-
Towards Robust FastSpeech 2 by Modelling Residual Multimodality 2 Jun 2023 · 1 repository · arXiv:2306.01442
-
Transformer-based Annotation Bias-aware Medical Image Segmentation 2 Jun 2023 · 1 repository · arXiv:2306.01340