Methods › General › Normalization › Layer Normalization › Papers, page 27
Layer Normalization
Papers archive 2025-07-28
archive papers tagged: 24,980 · with a code link: 11,273 · where Syntology ran a sample: 3,471 (2,923 with a run with no instrument failure, 548 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,471 of 24,980 tagged: 2,923 with a run with no instrument failure, 548 where every run was a failure of Syntology's instrument)
Page 27 of 250: papers 2,601 to 2,700 of 24,980, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Investigating Energy Efficiency and Performance Trade-offs in LLM Inference Across Tasks and DVFS Settings 14 Jan 2025 · 0 repositories · arXiv:2501.08219
-
Large Language Models For Text Classification: Case Study And Comprehensive Review 14 Jan 2025 · 0 repositories · arXiv:2501.08457
-
Optimizing Language Models for Grammatical Acceptability: A Comparative Study of Fine-Tuning Techniques 14 Jan 2025 · 0 repositories · arXiv:2501.07853
-
PokerBench: Training Large Language Models to become Professional Poker Players 14 Jan 2025 · 1 repository · arXiv:2501.08328
-
PSReg: Prior-guided Sparse Mixture of Experts for Point Cloud Registration 14 Jan 2025 · 0 repositories · arXiv:2501.07762
-
READ: Reinforcement-based Adversarial Learning for Text Classification with Limited Labeled Data 14 Jan 2025 · 0 repositories · arXiv:2501.08035
-
ReARTeR: Retrieval-Augmented Reasoning with Trustworthy Process Rewarding 14 Jan 2025 · 1 repository · arXiv:2501.07861
-
Towards Lightweight Time Series Forecasting: a Patch-wise Transformer with Weak Data Enriching 14 Jan 2025 · 0 repositories · arXiv:2501.10448
-
Transforming Indoor Localization: Advanced Transformer Architecture for NLOS Dominated Wireless Environments with Distributed Sensors 14 Jan 2025 · 0 repositories · arXiv:2501.07774
-
UFGraphFR: An attempt at a federated recommendation system based on user text characteristics 14 Jan 2025 · 1 repository · arXiv:2501.08044
-
Comparative analysis of optical character recognition methods for Sámi texts from the National Library of Norway 13 Jan 2025 · 2 repositories · arXiv:2501.07300
-
D3MES: Diffusion Transformer with multihead equivariant self-attention for 3D molecule generation 13 Jan 2025 · 1 repository · arXiv:2501.07077
-
EdgeTAM: On-Device Track Anything Model 13 Jan 2025 · 1 repository · arXiv:2501.07256Syntology official: no sample here; runs from other or unrecorded repositories · 7 ran (of which 2 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 1 violated, 2 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples)
-
Enhancing Retrieval-Augmented Generation: A Study of Best Practices 13 Jan 2025 · 1 repository · arXiv:2501.07391Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Enhancing Talent Employment Insights Through Feature Extraction with LLM Finetuning 13 Jan 2025 · 0 repositories · arXiv:2501.07663
-
Estimating Musical Surprisal in Audio 13 Jan 2025 · 1 repository · arXiv:2501.07474
-
FinerWeb-10BT: Refining Web Data with LLM-Based Line-Level Filtering 13 Jan 2025 · 1 repository · arXiv:2501.07314
-
Future-Conditioned Recommendations with Multi-Objective Controllable Decision Transformer 13 Jan 2025 · 0 repositories · arXiv:2501.07212
-
GPT as a Monte Carlo Language Tree: A Probabilistic Perspective 13 Jan 2025 · 0 repositories · arXiv:2501.07641
-
How GPT learns layer by layer 13 Jan 2025 · 1 repository · arXiv:2501.07108
-
MathReader : Text-to-Speech for Mathematical Documents 13 Jan 2025 · 1 repository · arXiv:2501.07088
-
Parallel Key-Value Cache Fusion for Position Invariant RAG 13 Jan 2025 · 0 repositories · arXiv:2501.07523
-
PRKAN: Parameter-Reduced Kolmogorov-Arnold Networks 13 Jan 2025 · 1 repository · arXiv:2501.07032
-
Scaling Up ESM2 Architectures for Long Protein Sequences Analysis: Long and Quantized Approaches 13 Jan 2025 · 0 repositories · arXiv:2501.07747
-
SST-EM: Advanced Metrics for Evaluating Semantic, Spatial and Temporal Aspects in Video Editing 13 Jan 2025 · 1 repository · arXiv:2501.07554
-
WebWalker: Benchmarking LLMs in Web Traversal 13 Jan 2025 · 2 repositories · arXiv:2501.07572
-
Better Prompt Compression Without Multi-Layer Perceptrons 12 Jan 2025 · 0 repositories · arXiv:2501.06730
-
DRDT3: Diffusion-Refined Decision Test-Time Training Model 12 Jan 2025 · 0 repositories · arXiv:2501.06718
-
Eliza: A Web3 friendly AI Agent Operating System 12 Jan 2025 · 2 repositories · arXiv:2501.06781
-
Generative Artificial Intelligence-Supported Pentesting: A Comparison between Claude Opus, GPT-4, and Copilot 12 Jan 2025 · 0 repositories · arXiv:2501.06963
-
MFConvTr: Multi-Frequency Convolutional Transformer for Fetal Arrhythmia Detection in Non-Invasive fECG 12 Jan 2025 · 0 repositories · arXiv:2501.16649
-
MiniRAG: Towards Extremely Simple Retrieval-Augmented Generation 12 Jan 2025 · 1 repository · arXiv:2501.06713Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Transforming Vision Transformer: Towards Efficient Multi-Task Asynchronous Learning 12 Jan 2025 · 1 repository · arXiv:2501.06884Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
ZNO-Eval: Benchmarking reasoning capabilities of large language models in Ukrainian 12 Jan 2025 · 1 repository · arXiv:2501.06715
-
A Comparative Performance Analysis of Classification and Segmentation Models on Bangladeshi Pothole Dataset 11 Jan 2025 · 0 repositories · arXiv:2501.06602
-
Assessing instructor-AI cooperation for grading essay-type questions in an introductory sociology course 11 Jan 2025 · 1 repository · arXiv:2501.06461
-
CeViT: Copula-Enhanced Vision Transformer in multi-task learning and bi-group image covariates with an application to myopia screening 11 Jan 2025 · 1 repository · arXiv:2501.06540
-
First Token Probability Guided RAG for Telecom Question Answering 11 Jan 2025 · 0 repositories · arXiv:2501.06468
-
Flash Window Attention: speedup the attention computation for Swin Transformer 11 Jan 2025 · 2 repositories · arXiv:2501.06480
-
FocusDD: Real-World Scene Infusion for Robust Dataset Distillation 11 Jan 2025 · 0 repositories · arXiv:2501.06405
-
Ladder-residual: parallelism-aware architecture for accelerating large model inference with communication overlapping 11 Jan 2025 · 1 repository · arXiv:2501.06589Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 0 violated, 3 with no contract checked; 5 where Syntology's instrument failed) · 1 unverified (of 10 harvested samples) · 1 pointer-only (licence)
-
Tensor Product Attention Is All You Need 11 Jan 2025 · 1 repository · arXiv:2501.06425Syntology official (archive's flag): 5 ran · 8 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 1 violated, 2 with no contract checked; 5 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples) · 2 pointer-only (licence)
-
Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization 11 Jan 2025 · 0 repositories · arXiv:2501.06663
-
A Holistically Point-guided Text Framework for Weakly-Supervised Camouflaged Object Detection 10 Jan 2025 · 0 repositories · arXiv:2501.06038
-
An Attention-Guided Deep Learning Approach for Classifying 39 Skin Lesion Types 10 Jan 2025 · 1 repository · arXiv:2501.05991
-
Analyzing Spatio-Temporal Dynamics of Dissolved Oxygen for the River Thames using Superstatistical Methods and Machine Learning 10 Jan 2025 · 0 repositories · arXiv:2501.07599
-
Binary Event-Driven Spiking Transformer 10 Jan 2025 · 0 repositories · arXiv:2501.05904
-
Bridging Dialects: Translating Standard Bangla to Regional Variants Using Neural Models 10 Jan 2025 · 0 repositories · arXiv:2501.05749
-
CognoSpeak: an automatic, remote assessment of early cognitive decline in real-world conversational speech 10 Jan 2025 · 0 repositories · arXiv:2501.05755
-
Iconicity in Large Language Models 10 Jan 2025 · 0 repositories · arXiv:2501.05643
-
Merging Feed-Forward Sublayers for Compressed Transformers 10 Jan 2025 · 1 repository · arXiv:2501.06126
-
Mix-QViT: Mixed-Precision Vision Transformer Quantization Driven by Layer Importance and Quantization Sensitivity 10 Jan 2025 · 0 repositories · arXiv:2501.06357
-
Model Inversion in Split Learning for Personalized LLMs: New Insights from Information Bottleneck Theory 10 Jan 2025 · 0 repositories · arXiv:2501.05965
-
MSCViT: A Small-size ViT architecture with Multi-Scale Self-Attention Mechanism for Tiny Datasets 10 Jan 2025 · 0 repositories · arXiv:2501.06040
-
Multi-subject Open-set Personalization in Video Generation 10 Jan 2025 · 0 repositories · arXiv:2501.06187
-
Aligning Brain Activity with Advanced Transformer Models: Exploring the Role of Punctuation in Semantic Processing 10 Jan 2025 · 1 repository · arXiv:2501.06278
-
Swin-X2S: Reconstructing 3D Shape from 2D Biplanar X-ray with Swin Transformers 10 Jan 2025 · 0 repositories · arXiv:2501.05961
-
TTS-Transducer: End-to-End Speech Synthesis with Neural Transducer 10 Jan 2025 · 0 repositories · arXiv:2501.06320
-
VideoRAG: Retrieval-Augmented Generation over Video Corpus 10 Jan 2025 · 1 repository · arXiv:2501.05874Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Weakly Supervised Segmentation of Hyper-Reflective Foci with Compact Convolutional Transformers and SAM2 10 Jan 2025 · 0 repositories · arXiv:2501.05933
-
A General Retrieval-Augmented Generation Framework for Multimodal Case-Based Reasoning Applications 9 Jan 2025 · 0 repositories · arXiv:2501.05030
-
Analyzing Memorization in Large Language Models through the Lens of Model Attribution 9 Jan 2025 · 1 repository · arXiv:2501.05078
-
Biomedical Relation Extraction via Adaptive Document-Relation Cross-Mapping and Concept Unique Identifier 9 Jan 2025 · 0 repositories · arXiv:2501.05155
-
DisSim-FinBERT: Text Simplification for Core Message Extraction in Complex Financial Texts 9 Jan 2025 · 0 repositories · arXiv:2501.04959
-
Enhancing Plagiarism Detection in Marathi with a Weighted Ensemble of TF-IDF and BERT Embeddings for Low-Resource Language Processing 9 Jan 2025 · 1 repository · arXiv:2501.05260
-
Large language models streamline automated systematic review: A preliminary study 9 Jan 2025 · 0 repositories · arXiv:2502.15702
-
LLMQuoter: Enhancing RAG Capabilities Through Efficient Quote Extraction From Large Contexts 9 Jan 2025 · 1 repository · arXiv:2501.05554
-
LongViTU: Instruction Tuning for Long-Form Video Understanding 9 Jan 2025 · 0 repositories · arXiv:2501.05037
-
OpenAI ChatGPT interprets Radiological Images: GPT-4 as a Medical Doctor for a Fast Check-Up 9 Jan 2025 · 0 repositories · arXiv:2501.06269
-
Optimizing Multitask Industrial Processes with Predictive Action Guidance 9 Jan 2025 · 0 repositories · arXiv:2501.05108
-
RAG-WM: An Efficient Black-Box Watermarking Approach for Retrieval-Augmented Generation of Large Language Models 9 Jan 2025 · 0 repositories · arXiv:2501.05249
-
SpecTf: Transformers Enable Data-Driven Imaging Spectroscopy Cloud Detection 9 Jan 2025 · 1 repository · arXiv:2501.04916
-
The dynamics of meaning through time: Assessment of Large Language Models 9 Jan 2025 · 0 repositories · arXiv:2501.05552
-
The more polypersonal the better -- a short look on space geometry of fine-tuned layers 9 Jan 2025 · 0 repositories · arXiv:2501.05503
-
UAV-VLA: Vision-Language-Action System for Large Scale Aerial Mission Generation 9 Jan 2025 · 1 repository · arXiv:2501.05014
-
Advancing Retrieval-Augmented Generation for Persian: Development of Language Models, Comprehensive Benchmarks, and Best Practices for Optimization 8 Jan 2025 · 0 repositories · arXiv:2501.04858
-
Circuit Complexity Bounds for Visual Autoregressive Model 8 Jan 2025 · 0 repositories · arXiv:2501.04299
-
Integrating LLMs with ITS: Recent Advances, Potentials, Challenges, and Future Directions 8 Jan 2025 · 0 repositories · arXiv:2501.04437
-
Knowledge Retrieval Based on Generative AI 8 Jan 2025 · 0 repositories · arXiv:2501.04635
-
MB-TaylorFormer V2: Improved Multi-branch Linear Transformer Expanded by Taylor Formula for Image Restoration 8 Jan 2025 · 2 repositories · arXiv:2501.04486
-
Multi-task retriever fine-tuning for domain-specific and efficient RAG 8 Jan 2025 · 0 repositories · arXiv:2501.04652
-
Quantum-inspired Embeddings Projection and Similarity Metrics for Representation Learning 8 Jan 2025 · 1 repository · arXiv:2501.04591
-
Re-ranking the Context for Multimodal Retrieval Augmented Generation 8 Jan 2025 · 0 repositories · arXiv:2501.04695
-
Scaling Large Language Model Training on Frontier with Low-Bandwidth Partitioning 8 Jan 2025 · 0 repositories · arXiv:2501.04266
-
AuxDepthNet: Real-Time Monocular 3D Object Detection with Depth-Sensitive Features 7 Jan 2025 · 0 repositories · arXiv:2501.03700
-
CFFormer: Cross CNN-Transformer Channel Attention and Spatial Feature Fusion for Improved Segmentation of Low Quality Medical Images 7 Jan 2025 · 0 repositories · arXiv:2501.03629
-
Efficient and Accurate Tuberculosis Diagnosis: Attention Residual U-Net and Vision Transformer Based Detection Framework 7 Jan 2025 · 0 repositories · arXiv:2501.03538
-
Entropy-Guided Attention for Private LLMs 7 Jan 2025 · 1 repository · arXiv:2501.03489
-
Finding A Voice: Evaluating African American Dialect Generation for Chatbot Technology 7 Jan 2025 · 1 repository · arXiv:2501.03441
-
How to Select Pre-Trained Code Models for Reuse? A Learning Perspective 7 Jan 2025 · 1 repository · arXiv:2501.03783
-
HP-BERT: A framework for longitudinal study of Hinduphobia on social media via LLMs 7 Jan 2025 · 1 repository · arXiv:2501.05482
-
IntegrityAI at GenAI Detection Task 2: Detecting Machine-Generated Academic Essays in English and Arabic Using ELECTRA and Stylometry 7 Jan 2025 · 0 repositories · arXiv:2501.05476
-
Language and Planning in Robotic Navigation: A Multilingual Evaluation of State-of-the-Art Models 7 Jan 2025 · 0 repositories · arXiv:2501.05478
-
LM-Net: A Light-weight and Multi-scale Network for Medical Image Segmentation 7 Jan 2025 · 1 repository · arXiv:2501.03838
-
MTRAG: A Multi-Turn Conversational Benchmark for Evaluating Retrieval-Augmented Generation Systems 7 Jan 2025 · 1 repository · arXiv:2501.03468
-
Practical Design and Benchmarking of Generative AI Applications for Surgical Billing and Coding 7 Jan 2025 · 0 repositories · arXiv:2501.05479
-
RAG-Check: Evaluating Multimodal Retrieval Augmented Generation Performance 7 Jan 2025 · 0 repositories · arXiv:2501.03995
-
Reading with Intent -- Neutralizing Intent 7 Jan 2025 · 0 repositories · arXiv:2501.03475
-
SNR-EQ-JSCC: Joint Source-Channel Coding with SNR-Based Embedding and Query 7 Jan 2025 · 0 repositories · arXiv:2501.04732
-
Text to Band Gap: Pre-trained Language Models as Encoders for Semiconductor Band Gap Prediction 7 Jan 2025 · 1 repository · arXiv:2501.03456