Methods › General › Normalization › Layer Normalization › Papers, page 46
Layer Normalization
Papers archive 2025-07-28
archive papers tagged: 24,980 · with a code link: 11,273 · where Syntology ran a sample: 3,471 (2,923 with a run with no instrument failure, 548 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,471 of 24,980 tagged: 2,923 with a run with no instrument failure, 548 where every run was a failure of Syntology's instrument)
Page 46 of 250: papers 4,501 to 4,600 of 24,980, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Can LLMs Reliably Simulate Human Learner Actions? A Simulation Authoring Framework for Open-Ended Learning Environments 3 Oct 2024 · 1 repository · arXiv:2410.02110
-
CAX: Cellular Automata Accelerated in JAX 3 Oct 2024 · 1 repository · arXiv:2410.02651
-
Coal Mining Question Answering with LLMs 3 Oct 2024 · 0 repositories · arXiv:2410.02959
-
CodeJudge: Evaluating Code Generation with Large Language Models 3 Oct 2024 · 1 repository · arXiv:2410.02184Syntology official (archive's flag): 13 ran · 15 ran (of which 0 constructed an object rather than computing a result; 14 with no instrument failure: 2 honoured, 0 violated, 12 with no contract checked; 1 where Syntology's instrument failed) · 8 unverified (of 23 harvested samples) · 2 pointer-only (licence)
-
Adversarial Decoding: Generating Readable Documents for Adversarial Objectives 3 Oct 2024 · 1 repository · arXiv:2410.02163
-
Defining Knowledge: Bridging Epistemology and Large Language Models 3 Oct 2024 · 0 repositories · arXiv:2410.02499
-
Domain-Specific Retrieval-Augmented Generation Using Vector Stores, Knowledge Graphs, and Tensor Factorization 3 Oct 2024 · 0 repositories · arXiv:2410.02721
-
Efficient Semantic Segmentation via Lightweight Multiple-Information Interaction Network 3 Oct 2024 · 0 repositories · arXiv:2410.02224
-
FAN: Fourier Analysis Networks 3 Oct 2024 · 2 repositories · arXiv:2410.02675Syntology official (archive's flag): 3 ran · 6 ran (of which 3 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 3 pointer-only (licence)
-
From Pixels to Tokens: Byte-Pair Encoding on Quantized Visual Modalities 3 Oct 2024 · 0 repositories · arXiv:2410.02155
-
Grounding Large Language Models In Embodied Environment With Imperfect World Models 3 Oct 2024 · 0 repositories · arXiv:2410.02742
-
HELMET: How to Evaluate Long-Context Language Models Effectively and Thoroughly 3 Oct 2024 · 1 repository · arXiv:2410.02694Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 10 harvested samples) · 2 pointer-only (licence)
-
HiFiSeg: High-Frequency Information Enhanced Polyp Segmentation with Global-Local Vision Transformer 3 Oct 2024 · 0 repositories · arXiv:2410.02528
-
How Much Can RAG Help the Reasoning of LLM? 3 Oct 2024 · 0 repositories · arXiv:2410.02338
-
IndicSentEval: How Effectively do Multilingual Transformer Models encode Linguistic Properties for Indic Languages? 3 Oct 2024 · 0 repositories · arXiv:2410.02611
-
Intrinsic Evaluation of RAG Systems for Deep-Logic Questions 3 Oct 2024 · 0 repositories · arXiv:2410.02932
-
IoT-LLM: Enhancing Real-World IoT Task Reasoning with Large Language Models 3 Oct 2024 · 0 repositories · arXiv:2410.02429
-
L-CiteEval: Do Long-Context Models Truly Leverage Context for Responding? 3 Oct 2024 · 2 repositories · arXiv:2410.02115
-
LLaVA-Critic: Learning to Evaluate Multimodal Models 3 Oct 2024 · 0 repositories · arXiv:2410.02712
-
Morphological evaluation of subwords vocabulary used by BETO language model 3 Oct 2024 · 0 repositories · arXiv:2410.02283
-
Plots Unlock Time-Series Understanding in Multimodal Models 3 Oct 2024 · 0 repositories · arXiv:2410.02637
-
Reward-RAG: Enhancing RAG with Reward Driven Supervision 3 Oct 2024 · 0 repositories · arXiv:2410.03780
-
SC-CDM: Enhancing Quality of Image Semantic Communication with a Compact Diffusion Model 3 Oct 2024 · 0 repositories · arXiv:2410.02121
-
GPT-4o as the Gold Standard: A Scalable and General Purpose Approach to Filter Language Model Pretraining Data 3 Oct 2024 · 0 repositories · arXiv:2410.02755
-
Theoretical Insights into Fine-Tuning Attention Mechanism: Generalization and Optimization 3 Oct 2024 · 2 repositories · arXiv:2410.02247Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 5 harvested samples) · 1 pointer-only (licence)
-
Towards Understanding the Universality of Transformers for Next-Token Prediction 3 Oct 2024 · 0 repositories · arXiv:2410.03011
-
Training Language Models on Synthetic Edit Sequences Improves Code Synthesis 3 Oct 2024 · 1 repository · arXiv:2410.02749Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Training Nonlinear Transformers for Chain-of-Thought Inference: A Theoretical Generalization Analysis 3 Oct 2024 · 0 repositories · arXiv:2410.02167
-
TrajGPT: Irregular Time-Series Representation Learning for Health Trajectory Analysis 3 Oct 2024 · 0 repositories · arXiv:2410.02133
-
UncertaintyRAG: Span-Level Uncertainty Enhanced Long-Context Modeling for Retrieval-Augmented Generation 3 Oct 2024 · 0 repositories · arXiv:2410.02719
-
Visual Editing with LLM-based Tool Chaining: An Efficient Distillation Approach for Real-Time Applications 3 Oct 2024 · 1 repository · arXiv:2410.02952
-
A Spark of Vision-Language Intelligence: 2-Dimensional Autoregressive Transformer for Efficient Finegrained Image Generation 2 Oct 2024 · 1 repository · arXiv:2410.01912
-
A versatile machine learning workflow for high-throughput analysis of supported metal catalyst particles 2 Oct 2024 · 1 repository · arXiv:2410.01213
-
AHP-Powered LLM Reasoning for Multi-Criteria Evaluation of Open-Ended Responses 2 Oct 2024 · 0 repositories · arXiv:2410.01246
-
Attention layers provably solve single-location regression 2 Oct 2024 · 1 repository · arXiv:2410.01537Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Automated Red Teaming with GOAT: the Generative Offensive Agent Tester 2 Oct 2024 · 0 repositories · arXiv:2410.01606
-
Automatic deductive coding in discourse analysis: an application of large language models in learning analytics 2 Oct 2024 · 1 repository · arXiv:2410.01240
-
BordIRlines: A Dataset for Evaluating Cross-lingual Retrieval-Augmented Generation 2 Oct 2024 · 1 repository · arXiv:2410.01171Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 0 violated, 12 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 16 harvested samples)
-
DeepProtein: Deep Learning Library and Benchmark for Protein Sequence Learning 2 Oct 2024 · 1 repository · arXiv:2410.02023
-
Depth Pro: Sharp Monocular Metric Depth in Less Than a Second 2 Oct 2024 · 1 repository · arXiv:2410.02073Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Emotion-Aware Embedding Fusion in LLMs (Flan-T5, LLAMA 2, DeepSeek-R1, and ChatGPT 4) for Intelligent Response Generation 2 Oct 2024 · 0 repositories · arXiv:2410.01306
-
Enhancing LLM Fine-tuning for Text-to-SQLs by SQL Quality Measurement 2 Oct 2024 · 0 repositories · arXiv:2410.01869
-
Enhancing Retrieval in QA Systems with Derived Feature Association 2 Oct 2024 · 1 repository · arXiv:2410.03754
-
ENTP: Encoder-only Next Token Prediction 2 Oct 2024 · 0 repositories · arXiv:2410.01600
-
ET-Plan-Bench: Embodied Task-level Planning Benchmark Towards Spatial-Temporal Cognition with Foundation Models 2 Oct 2024 · 0 repositories · arXiv:2410.14682
-
Financial Sentiment Analysis on News and Reports Using Large Language Models and FinBERT 2 Oct 2024 · 0 repositories · arXiv:2410.01987
-
FlashMask: Efficient and Rich Mask Extension of FlashAttention 2 Oct 2024 · 1 repository · arXiv:2410.01359Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Getting Free Bits Back from Rotational Symmetries in LLMs 2 Oct 2024 · 0 repositories · arXiv:2410.01309
-
Imaging foundation model for universal enhancement of non-ideal measurement CT 2 Oct 2024 · 1 repository · arXiv:2410.01591
-
MARPLE: A Benchmark for Long-Horizon Inference 2 Oct 2024 · 1 repository · arXiv:2410.01926Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
Mind Scramble: Unveiling Large Language Model Psychology Via Typoglycemia 2 Oct 2024 · 1 repository · arXiv:2410.01677
-
On The Adaptation of Unlimiformer for Decoder-Only Transformers 2 Oct 2024 · 0 repositories · arXiv:2410.01637
-
Open-RAG: Enhanced Retrieval-Augmented Reasoning with Open-Source Large Language Models 2 Oct 2024 · 1 repository · arXiv:2410.01782
-
Quantifying Generalization Complexity for Large Language Models 2 Oct 2024 · 1 repository · arXiv:2410.01769Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
RADAR: Robust Two-stage Modality-incomplete Industrial Anomaly Detection 2 Oct 2024 · 0 repositories · arXiv:2410.01737
-
RS-FME-SwinT: A Novel Feature Map Enhancement Framework Integrating Customized SwinT with Residual and Spatial CNN for Monkeypox Diagnosis 2 Oct 2024 · 0 repositories · arXiv:2410.01216
-
Saliency-Guided DETR for Moment Retrieval and Highlight Detection 2 Oct 2024 · 1 repository · arXiv:2410.01615
-
Seeing Eye to AI: Human Alignment via Gaze-Based Response Rewards for Large Language Models 2 Oct 2024 · 1 repository · arXiv:2410.01532Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
UlcerGPT: A Multimodal Approach Leveraging Large Language and Vision Models for Diabetic Foot Ulcer Image Transcription 2 Oct 2024 · 0 repositories · arXiv:2410.01989
-
Advanced Arabic Alphabet Sign Language Recognition Using Transfer Learning and Transformer Models 1 Oct 2024 · 0 repositories · arXiv:2410.00681
-
Unleashing the Unseen: Harnessing Benign Datasets for Jailbreaking Large Language Models 1 Oct 2024 · 1 repository · arXiv:2410.00451
-
AlignSum: Data Pyramid Hierarchical Fine-tuning for Aligning with Human Summarization Preference 1 Oct 2024 · 1 repository · arXiv:2410.00409Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Creative and Context-Aware Translation of East Asian Idioms with GPT-4 1 Oct 2024 · 1 repository · arXiv:2410.00988
-
Decoding Hate: Exploring Language Models' Reactions to Hate Speech 1 Oct 2024 · 0 repositories · arXiv:2410.00775
-
Deep Multimodal Fusion for Semantic Segmentation of Remote Sensing Earth Observation Data 1 Oct 2024 · 0 repositories · arXiv:2410.00469
-
Domain Aware Multi-Task Pretraining of 3D Swin Transformer for T1-weighted Brain MRI 1 Oct 2024 · 1 repository · arXiv:2410.00410
-
End-to-End Speech Recognition with Pre-trained Masked Language Model 1 Oct 2024 · 1 repository · arXiv:2410.00528
-
Exploring the Learning Capabilities of Language Models using LEVERWORLDS 1 Oct 2024 · 0 repositories · arXiv:2410.00519
-
Pediatric Wrist Fracture Detection Using Feature Context Excitation Modules in X-ray Images 1 Oct 2024 · 1 repository · arXiv:2410.01031
-
GLMHA A Guided Low-rank Multi-Head Self-Attention for Efficient Image Restoration and Spectral Reconstruction 1 Oct 2024 · 0 repositories · arXiv:2410.00380
-
Insight: A Multi-Modal Diagnostic Pipeline using LLMs for Ocular Surface Disease Diagnosis 1 Oct 2024 · 0 repositories · arXiv:2410.00292
-
Language Enhanced Model for Eye (LEME): An Open-Source Ophthalmology-Specific Large Language Model 1 Oct 2024 · 0 repositories · arXiv:2410.03740
-
MAP: Unleashing Hybrid Mamba-Transformer Vision Backbone's Potential with Masked Autoregressive Pretraining 1 Oct 2024 · 0 repositories · arXiv:2410.00871
-
Multi-Scale Temporal Transformer For Speech Emotion Recognition 1 Oct 2024 · 0 repositories · arXiv:2410.00390
-
nGPT: Normalized Transformer with Representation Learning on the Hypersphere 1 Oct 2024 · 0 repositories · arXiv:2410.01131
-
Optimizing and Evaluating Enterprise Retrieval-Augmented Generation (RAG): A Content Design Perspective 1 Oct 2024 · 1 repository · arXiv:2410.12812
-
Quantifying reliance on external information over parametric knowledge during Retrieval Augmented Generation (RAG) using mechanistic analysis 1 Oct 2024 · 0 repositories · arXiv:2410.00857
-
RATIONALYST: Pre-training Process-Supervision for Improving Reasoning 1 Oct 2024 · 1 repository · arXiv:2410.01044
-
Robust Traffic Forecasting against Spatial Shift over Years 1 Oct 2024 · 1 repository · arXiv:2410.00373Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 1 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 13 harvested samples)
-
Sparse Attention Decomposition Applied to Circuit Tracing 1 Oct 2024 · 1 repository · arXiv:2410.00340
-
STGformer: Efficient Spatiotemporal Graph Transformer for Traffic Forecasting 1 Oct 2024 · 1 repository · arXiv:2410.00385
-
TFCT-I2P: Three stream fusion network with color aware transformer for image-to-point cloud registration 1 Oct 2024 · 1 repository · arXiv:2410.00360
-
TransResNet: Integrating the Strengths of ViTs and CNNs for High Resolution Medical Image Segmentation via Feature Grafting 1 Oct 2024 · 1 repository · arXiv:2410.00986
-
A Looming Replication Crisis in Evaluating Behavior in Language Models? Evidence and Solutions 30 Sep 2024 · 0 repositories · arXiv:2409.20303
-
A Methodology for Explainable Large Language Models with Integrated Gradients and Linguistic Analysis in Text Classification 30 Sep 2024 · 0 repositories · arXiv:2410.00250
-
ACE: All-round Creator and Editor Following Instructions via Diffusion Transformer 30 Sep 2024 · 0 repositories · arXiv:2410.00086
-
Adapting LLMs for the Medical Domain in Portuguese: A Study on Fine-Tuning and Model Evaluation 30 Sep 2024 · 0 repositories · arXiv:2410.00163
-
ASQuery: A Query-based Model for Action Segmentation 30 Sep 2024 · 1 repository
-
BSharedRAG: Backbone Shared Retrieval-Augmented Generation for the E-commerce Domain 30 Sep 2024 · 0 repositories · arXiv:2409.20075
-
CBAM-SwinT-BL: Small Rail Surface Defect Detection Method Based on Swin Transformer with Block Level CBAM Enhancement 30 Sep 2024 · 0 repositories · arXiv:2409.20113
-
CliMB: An AI-enabled Partner for Clinical Predictive Modeling 30 Sep 2024 · 1 repository · arXiv:2410.03736Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 10 harvested samples)
-
Depression detection in social media posts using transformer-based models and auxiliary features 30 Sep 2024 · 0 repositories · arXiv:2409.20048
-
Evaluating the fairness of task-adaptive pretraining on unlabeled test data before few-shot text classification 30 Sep 2024 · 1 repository · arXiv:2410.00179
-
GTransPDM: A Graph-embedded Transformer with Positional Decoupling for Pedestrian Crossing Intention Prediction 30 Sep 2024 · 0 repositories · arXiv:2409.20223
-
Ingest-And-Ground: Dispelling Hallucinations from Continually-Pretrained LLMs with RAG 30 Sep 2024 · 0 repositories · arXiv:2410.02825
-
MaskMamba: A Hybrid Mamba-Transformer Model for Masked Image Generation 30 Sep 2024 · 0 repositories · arXiv:2409.19937
-
Modelando procesos cognitivos de la lectura natural con GPT-2 30 Sep 2024 · 0 repositories · arXiv:2409.20174
-
Exploring Social Media Image Categorization Using Large Models with Different Adaptation Methods: A Case Study on Cultural Nature's Contributions to People 30 Sep 2024 · 0 repositories · arXiv:2410.00275
-
On The Planning Abilities of OpenAI's o1 Models: Feasibility, Optimality, and Generalizability 30 Sep 2024 · 2 repositories · arXiv:2409.19924
-
QAEncoder: Towards Aligned Representation Learning in Question Answering System 30 Sep 2024 · 1 repository · arXiv:2409.20434