Methods › General › Normalization › Layer Normalization › Papers, page 31
Layer Normalization
Papers archive 2025-07-28
archive papers tagged: 24,980 · with a code link: 11,273 · where Syntology ran a sample: 3,471 (2,923 with a run with no instrument failure, 548 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,471 of 24,980 tagged: 2,923 with a run with no instrument failure, 548 where every run was a failure of Syntology's instrument)
Page 31 of 250: papers 3,001 to 3,100 of 24,980, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Enhancing Masked Time-Series Modeling via Dropping Patches 19 Dec 2024 · 1 repository · arXiv:2412.15315
-
Graph-Convolutional Networks: Named Entity Recognition and Large Language Model Embedding in Document Clustering 19 Dec 2024 · 0 repositories · arXiv:2412.14867
-
How good is GPT at writing political speeches for the White House? 19 Dec 2024 · 0 repositories · arXiv:2412.14617
-
Jet: A Modern Transformer-Based Normalizing Flow 19 Dec 2024 · 0 repositories · arXiv:2412.15129
-
Joint Models for Handling Non-Ignorable Missing Data using Bayesian Additive Regression Trees: Application to Leaf Photosynthetic Traits Data 19 Dec 2024 · 0 repositories · arXiv:2412.14946
-
Knowledge Injection via Prompt Distillation 19 Dec 2024 · 0 repositories · arXiv:2412.14964
-
LLMs as mediators: Can they diagnose conflicts accurately? 19 Dec 2024 · 0 repositories · arXiv:2412.14675
-
Mention Attention for Pronoun Translation 19 Dec 2024 · 0 repositories · arXiv:2412.14829
-
MIETT: Multi-Instance Encrypted Traffic Transformer for Encrypted Traffic Classification 19 Dec 2024 · 1 repository · arXiv:2412.15306
-
MSA-GCN: Exploiting Multi-Scale Temporal Dynamics With Adaptive Graph Convolution for Skeleton-Based Action Recognition 19 Dec 2024 · 0 repositories
-
PA-RAG: RAG Alignment via Multi-Perspective Preference Optimization 19 Dec 2024 · 1 repository · arXiv:2412.14510
-
Qua²SeDiMo: Quantifiable Quantization Sensitivity of Diffusion Models 19 Dec 2024 · 0 repositories · arXiv:2412.14628
-
Query pipeline optimization for cancer patient question answering systems 19 Dec 2024 · 0 repositories · arXiv:2412.14751
-
Relational Programming with Foundation Models 19 Dec 2024 · 0 repositories · arXiv:2412.14515
-
ResoFilter: Fine-grained Synthetic Data Filtering for Large Language Models through Data-Parameter Resonance Analysis 19 Dec 2024 · 1 repository · arXiv:2412.14809
-
Review-Then-Refine: A Dynamic Framework for Multi-Hop Question Answering with Temporal Adaptability 19 Dec 2024 · 0 repositories · arXiv:2412.15101
-
SKETCH: Structured Knowledge Enhanced Text Comprehension for Holistic Retrieval 19 Dec 2024 · 0 repositories · arXiv:2412.15443
-
Systematic Evaluation of Long-Context LLMs on Financial Concepts 19 Dec 2024 · 0 repositories · arXiv:2412.15386
-
Till the Layers Collapse: Compressing a Deep Neural Network through the Lenses of Batch Normalization Layers 19 Dec 2024 · 1 repository · arXiv:2412.15077
-
Tokenphormer: Structure-aware Multi-token Graph Transformer for Node Classification 19 Dec 2024 · 1 repository · arXiv:2412.15302
-
TOMG-Bench: Evaluating LLMs on Text-based Open Molecule Generation 19 Dec 2024 · 1 repository · arXiv:2412.14642
-
VISA: Retrieval Augmented Generation with Visual Source Attribution 19 Dec 2024 · 0 repositories · arXiv:2412.14457
-
Autonomous Microscopy Experiments through Large Language Model Agents 18 Dec 2024 · 1 repository · arXiv:2501.10385
-
Combining Aggregated Attention and Transformer Architecture for Accurate and Efficient Performance of Spiking Neural Networks 18 Dec 2024 · 0 repositories · arXiv:2412.13553
-
Distilled Pooling Transformer Encoder for Efficient Realistic Image Dehazing 18 Dec 2024 · 1 repository · arXiv:2412.14220
-
Enhancing Rhetorical Figure Annotation: An Ontology-Based Web Application with RAG Integration 18 Dec 2024 · 1 repository · arXiv:2412.13799
-
EvoWiki: Evaluating LLMs on Evolving Knowledge 18 Dec 2024 · 0 repositories · arXiv:2412.13582
-
Exploring Transformer-Augmented LSTM for Temporal and Spatial Feature Learning in Trajectory Prediction 18 Dec 2024 · 0 repositories · arXiv:2412.13419
-
Fake News Detection: Comparative Evaluation of BERT-like Models and Large Language Models with Generative AI-Annotated Data 18 Dec 2024 · 1 repository · arXiv:2412.14276
-
FarExStance: Explainable Stance Detection for Farsi 18 Dec 2024 · 2 repositories · arXiv:2412.14008
-
Federated Learning and RAG Integration: A Scalable Approach for Medical Large Language Models 18 Dec 2024 · 0 repositories · arXiv:2412.13720
-
GNN-Transformer Cooperative Architecture for Trustworthy Graph Contrastive Learning 18 Dec 2024 · 1 repository · arXiv:2412.16218
-
JoVALE: Detecting Human Actions in Video Using Audiovisual and Language Contexts 18 Dec 2024 · 1 repository · arXiv:2412.13708
-
MCMat: Multiview-Consistent and Physically Accurate PBR Material Generation 18 Dec 2024 · 0 repositories · arXiv:2412.14148
-
Memorizing SAM: 3D Medical Segment Anything Model with Memorizing Transformer 18 Dec 2024 · 1 repository · arXiv:2412.13908
-
Mix-LN: Unleashing the Power of Deeper Layers by Combining Pre-LN and Post-LN 18 Dec 2024 · 1 repository · arXiv:2412.13795
-
MMHMR: Generative Masked Modeling for Hand Mesh Recovery 18 Dec 2024 · 0 repositories · arXiv:2412.13393
-
Modality-Independent Graph Neural Networks with Global Transformers for Multimodal Recommendation 18 Dec 2024 · 1 repository · arXiv:2412.13994
-
Model Decides How to Tokenize: Adaptive DNA Sequence Tokenization with MxDNA 18 Dec 2024 · 1 repository · arXiv:2412.13716Syntology official (archive's flag): 2 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 2 pointer-only (licence)
-
Policy Decorator: Model-Agnostic Online Refinement for Large Policy Model 18 Dec 2024 · 0 repositories · arXiv:2412.13630
-
PsyDT: Using LLMs to Construct the Digital Twin of Psychological Counselor with Personalized Counseling Style for Psychological Counseling 18 Dec 2024 · 1 repository · arXiv:2412.13660
-
RAG-RewardBench: Benchmarking Reward Models in Retrieval Augmented Generation for Preference Alignment 18 Dec 2024 · 1 repository · arXiv:2412.13746
-
Reinforcement Learning from Automatic Feedback for High-Quality Unit Test Generation 18 Dec 2024 · 0 repositories · arXiv:2412.14308
-
Self-attentive Transformer for Fast and Accurate Postprocessing of Temperature and Wind Speed Forecasts 18 Dec 2024 · 1 repository · arXiv:2412.13957
-
Smarter, Better, Faster, Longer: A Modern Bidirectional Encoder for Fast, Memory Efficient, and Long Context Finetuning and Inference 18 Dec 2024 · 2 repositories · arXiv:2412.13663Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
A MapReduce Approach to Effectively Utilize Long Context Information in Retrieval Augmented Language Models 17 Dec 2024 · 0 repositories · arXiv:2412.15271
-
Adaptations of AI models for querying the LandMatrix database in natural language 17 Dec 2024 · 1 repository · arXiv:2412.12961
-
C-FedRAG: A Confidential Federated Retrieval-Augmented Generation System 17 Dec 2024 · 0 repositories · arXiv:2412.13163
-
Chinese SafetyQA: A Safety Short-form Factuality Benchmark for Large Language Models 17 Dec 2024 · 0 repositories · arXiv:2412.15265
-
CovNet: Covariance Information-Assisted CSI Feedback for FDD Massive MIMO Systems 17 Dec 2024 · 0 repositories · arXiv:2412.12875
-
Detecting Document-level Paraphrased Machine Generated Content: Mimicking Human Writing Style and Involving Discourse Features 17 Dec 2024 · 0 repositories · arXiv:2412.12679
-
Efficient Diffusion Transformer Policies with Mixture of Expert Denoisers for Multitask Learning 17 Dec 2024 · 1 repository · arXiv:2412.12953Syntology 10 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 11 harvested samples) · 3 pointer-only (licence)
-
Enhanced Momentum with Momentum Transformers 17 Dec 2024 · 0 repositories · arXiv:2412.12516
-
EXIT: Context-Aware Extractive Compression for Enhancing Retrieval-Augmented Generation 17 Dec 2024 · 1 repository · arXiv:2412.12559
-
Falcon: Faster and Parallel Inference of Large Language Models through Enhanced Semi-Autoregressive Drafting and Custom-Designed Decoding Tree 17 Dec 2024 · 0 repositories · arXiv:2412.12639
-
GaussTR: Foundation Model-Aligned Gaussian Transformer for Self-Supervised 3D Spatial Understanding 17 Dec 2024 · 1 repository · arXiv:2412.13193Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Harnessing Event Sensory Data for Error Pattern Prediction in Vehicles: A Language Model Approach 17 Dec 2024 · 1 repository · arXiv:2412.13041
-
JudgeBlender: Ensembling Judgments for Automatic Relevance Assessment 17 Dec 2024 · 1 repository · arXiv:2412.13268
-
LLM-based Discriminative Reasoning for Knowledge Graph Question Answering 17 Dec 2024 · 0 repositories · arXiv:2412.12643
-
LLMCL-GEC: Advancing Grammatical Error Correction with LLM-Driven Curriculum Learning 17 Dec 2024 · 0 repositories · arXiv:2412.12541
-
LLMs are Also Effective Embedding Models: An In-depth Overview 17 Dec 2024 · 0 repositories · arXiv:2412.12591
-
OmniEval: An Omnidirectional and Automatic RAG Evaluation Benchmark in Financial Domain 17 Dec 2024 · 1 repository · arXiv:2412.13018
-
PERC: Plan-As-Query Example Retrieval for Underrepresented Code Generation 17 Dec 2024 · 0 repositories · arXiv:2412.12447
-
PT: A Plain Transformer is Good Hospital Readmission Predictor 17 Dec 2024 · 0 repositories · arXiv:2412.12909
-
RAG-Star: Enhancing Deliberative Reasoning with Retrieval Augmented Verification and Refinement 17 Dec 2024 · 0 repositories · arXiv:2412.12881
-
RCTrans: Radar-Camera Transformer via Radar Densifier and Sequential Decoder for 3D Object Detection 17 Dec 2024 · 1 repository · arXiv:2412.12799
-
RemoteRAG: A Privacy-Preserving LLM Cloud RAG Service 17 Dec 2024 · 0 repositories · arXiv:2412.12775
-
SimGRAG: Leveraging Similar Subgraphs for Knowledge Graphs Driven Retrieval-Augmented Generation 17 Dec 2024 · 1 repository · arXiv:2412.15272
-
TimeCHEAT: A Channel Harmony Strategy for Irregularly Sampled Multivariate Time Series Analysis 17 Dec 2024 · 1 repository · arXiv:2412.12886Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples)
-
What External Knowledge is Preferred by LLMs? Characterizing and Exploring Chain of Evidence in Imperfect Context 17 Dec 2024 · 0 repositories · arXiv:2412.12632
-
A Benchmark and Robustness Study of In-Context-Learning with Large Language Models in Music Entity Detection 16 Dec 2024 · 1 repository · arXiv:2412.11851
-
A LoRA is Worth a Thousand Pictures 16 Dec 2024 · 0 repositories · arXiv:2412.12048
-
BioBridge: Unified Bio-Embedding with Bridging Modality in Code-Switched EMR 16 Dec 2024 · 1 repository · arXiv:2412.11671
-
Can Language Models Rival Mathematics Students? Evaluating Mathematical Reasoning through Textual Manipulation and Human Experiments 16 Dec 2024 · 0 repositories · arXiv:2412.11908
-
Causal Diffusion Transformers for Generative Modeling 16 Dec 2024 · 1 repository · arXiv:2412.12095Syntology official (archive's flag): 7 ran · 8 ran (of which 3 constructed an object rather than computing a result; 6 with no instrument failure: 1 honoured, 0 violated, 5 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
EDformer: Embedded Decomposition Transformer for Interpretable Multivariate Time Series Predictions 16 Dec 2024 · 1 repository · arXiv:2412.12227Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
GeoX: Geometric Problem Solving Through Unified Formalized Vision-Language Pre-training 16 Dec 2024 · 2 repositories · arXiv:2412.11863Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
Glimpse: Enabling White-Box Methods to Use Proprietary Models for Zero-Shot LLM-Generated Text Detection 16 Dec 2024 · 1 repository · arXiv:2412.11506Syntology official: not harvested · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Graph-Guided Textual Explanation Generation Framework 16 Dec 2024 · 0 repositories · arXiv:2412.12318
-
HResFormer: Hybrid Residual Transformer for Volumetric Medical Image Segmentation 16 Dec 2024 · 0 repositories · arXiv:2412.11458
-
Investigating Mixture of Experts in Dense Retrieval 16 Dec 2024 · 0 repositories · arXiv:2412.11864
-
Look Ahead Text Understanding and LLM Stitching 16 Dec 2024 · 1 repository · arXiv:2412.17836
-
Magnetic Field Data Calibration with Transformer Model Using Physical Constraints: A Scalable Method for Satellite Missions, Illustrated by Tianwen-1 16 Dec 2024 · 0 repositories · arXiv:2501.00020
-
No More Adam: Learning Rate Scaling at Initialization is All You Need 16 Dec 2024 · 1 repository · arXiv:2412.11768
-
OpenReviewer: A Specialized Large Language Model for Generating Critical Scientific Paper Reviews 16 Dec 2024 · 0 repositories · arXiv:2412.11948
-
Optimized Quran Passage Retrieval Using an Expanded QA Dataset and Fine-Tuned Language Models 16 Dec 2024 · 0 repositories · arXiv:2412.11431
-
Priority-Aware Model-Distributed Inference at Edge Networks 16 Dec 2024 · 0 repositories · arXiv:2412.12371
-
RAG Playground: A Framework for Systematic Evaluation of Retrieval Strategies and Prompt Engineering in RAG Systems 16 Dec 2024 · 1 repository · arXiv:2412.12322
-
Second Language (Arabic) Acquisition of LLMs via Progressive Vocabulary Expansion 16 Dec 2024 · 0 repositories · arXiv:2412.12310
-
The Impact of AI Assistance on Radiology Reporting: A Pilot Study Using Simulated AI Draft Reports 16 Dec 2024 · 0 repositories · arXiv:2412.12042
-
The Open Source Advantage in Large Language Models (LLMs) 16 Dec 2024 · 0 repositories · arXiv:2412.12004
-
Unanswerability Evaluation for Retrieval Augmented Generation 16 Dec 2024 · 0 repositories · arXiv:2412.12300
-
UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer 16 Dec 2024 · 0 repositories · arXiv:2412.11836
-
A Contextualized BERT model for Knowledge Graph Completion 15 Dec 2024 · 0 repositories · arXiv:2412.11016
-
MoRe: Class Patch Attention Needs Regularization for Weakly Supervised Semantic Segmentation 15 Dec 2024 · 1 repository · arXiv:2412.11076
-
One-Shot Multilingual Font Generation Via ViT 15 Dec 2024 · 0 repositories · arXiv:2412.11342
-
RoLargeSum: A Large Dialect-Aware Romanian News Dataset for Summary, Headline, and Keyword Generation 15 Dec 2024 · 1 repository · arXiv:2412.11317
-
Smaller Language Models Are Better Instruction Evolvers 15 Dec 2024 · 1 repository · arXiv:2412.11231
-
Towards Context-aware Convolutional Network for Image Restoration 15 Dec 2024 · 0 repositories · arXiv:2412.11008
-
Transformer-Based Bearing Fault Detection using Temporal Decomposition Attention Mechanism 15 Dec 2024 · 0 repositories · arXiv:2412.11245