Methods › General › Normalization › Layer Normalization › Papers, page 63
Layer Normalization
Papers archive 2025-07-28
archive papers tagged: 24,980 · with a code link: 11,273 · where Syntology ran a sample: 3,471 (2,923 with a run with no instrument failure, 548 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,471 of 24,980 tagged: 2,923 with a run with no instrument failure, 548 where every run was a failure of Syntology's instrument)
Page 63 of 250: papers 6,201 to 6,300 of 24,980, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
BioMNER: A Dataset for Biomedical Method Entity Recognition 28 Jun 2024 · 0 repositories · arXiv:2406.20038
-
Can GPT-4 Help Detect Quit Vaping Intentions? An Exploration of Automatic Data Annotation Approach 28 Jun 2024 · 0 repositories · arXiv:2407.00167
-
Covert Malicious Finetuning: Challenges in Safeguarding LLM Adaptation 28 Jun 2024 · 0 repositories · arXiv:2406.20053
-
FRED: Flexible REduction-Distribution Interconnect and Communication Implementation for Wafer-Scale Distributed Training of DNN Models 28 Jun 2024 · 0 repositories · arXiv:2406.19580
-
Generative Iris Prior Embedded Transformer for Iris Restoration 28 Jun 2024 · 1 repository · arXiv:2407.00261
-
InfiniGen: Efficient Generative Inference of Large Language Models with Dynamic KV Cache Management 28 Jun 2024 · 1 repository · arXiv:2406.19707Syntology 14 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 4 where Syntology's instrument failed) · 1 unverified (of 15 harvested samples) · 1 pointer-only (licence)
-
Machine Learning Predictors for Min-Entropy Estimation 28 Jun 2024 · 1 repository · arXiv:2406.19983
-
Multimodal Prototyping for cancer survival prediction 28 Jun 2024 · 1 repository · arXiv:2407.00224Syntology official (archive's flag): 7 ran · 7 ran (of which 3 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 3 where Syntology's instrument failed) · 5 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
PoliFormer: Scaling On-Policy RL with Transformers Results in Masterful Navigators 28 Jun 2024 · 0 repositories · arXiv:2406.20083
-
ScaleBiO: Scalable Bilevel Optimization for LLM Data Reweighting 28 Jun 2024 · 0 repositories · arXiv:2406.19976
-
ShortcutsBench: A Large-Scale Real-world Benchmark for API-based Agents 28 Jun 2024 · 1 repository · arXiv:2407.00132Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 1 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples)
-
The Computational Curse of Big Data for Bayesian Additive Regression Trees: A Hitting Time Analysis 28 Jun 2024 · 1 repository · arXiv:2406.19958
-
Uncertainty Quantification in Large Language Models Through Convex Hull Analysis 28 Jun 2024 · 0 repositories · arXiv:2406.19712
-
AutoPureData: Automated Filtering of Undesirable Web Data to Update LLM Knowledge 27 Jun 2024 · 1 repository · arXiv:2406.19271
-
AutoRAG-HP: Automatic Online Hyper-Parameter Tuning for Retrieval-Augmented Generation 27 Jun 2024 · 0 repositories · arXiv:2406.19251
-
Can Large Language Models Generate High-quality Patent Claims? 27 Jun 2024 · 1 repository · arXiv:2406.19465
-
Diminishing Stereotype Bias in Image Generation Model using Reinforcemenlent Learning Feedback 27 Jun 2024 · 0 repositories · arXiv:2407.09551
-
Enhancing Video-Language Representations with Structural Spatio-Temporal Alignment 27 Jun 2024 · 0 repositories · arXiv:2406.19255
-
Fibottention: Inceptive Visual Representation Learning with Diverse Attention Across Heads 27 Jun 2024 · 1 repository · arXiv:2406.19391
-
Fine-tuned network relies on generic representation to solve unseen cognitive task 27 Jun 2024 · 0 repositories · arXiv:2406.18926
-
From Artificial Needles to Real Haystacks: Improving Retrieval Capabilities in LLMs by Finetuning on Synthetic Data 27 Jun 2024 · 1 repository · arXiv:2406.19292
-
Granite-Function Calling Model: Introducing Function Calling Abilities via Multi-task Learning of Granular Tasks 27 Jun 2024 · 0 repositories · arXiv:2407.00121
-
Historia Magistra Vitae: Dynamic Topic Modeling of Roman Literature using Neural Embeddings 27 Jun 2024 · 0 repositories · arXiv:2406.18907
-
Human-Aware Vision-and-Language Navigation: Bridging Simulation to Reality with Dynamic Human Interactions 27 Jun 2024 · 1 repository · arXiv:2406.19236Syntology official (archive's flag): 7 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
IndoToxic2024: A Demographically-Enriched Dataset of Hate Speech and Toxicity Types for Indonesian Language 27 Jun 2024 · 0 repositories · arXiv:2406.19349
-
Leveraging Contrastive Learning for Enhanced Node Representations in Tokenized Graph Transformers 27 Jun 2024 · 0 repositories · arXiv:2406.19258
-
NTFormer: A Composite Node Tokenized Graph Transformer for Node Classification 27 Jun 2024 · 0 repositories · arXiv:2406.19249
-
RAVEN: Multitask Retrieval Augmented Vision-Language Learning 27 Jun 2024 · 0 repositories · arXiv:2406.19150
-
Retain, Blend, and Exchange: A Quality-aware Spatial-Stereo Fusion Approach for Event Stream Recognition 27 Jun 2024 · 1 repository · arXiv:2406.18845
-
SeaKR: Self-aware Knowledge Retrieval for Adaptive Retrieval Augmented Generation 27 Jun 2024 · 1 repository · arXiv:2406.19215Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
Generating Is Believing: Membership Inference Attacks against Retrieval-Augmented Generation 27 Jun 2024 · 0 repositories · arXiv:2406.19234
-
Segment Anything Model for automated image data annotation: empirical studies using text prompts from Grounding DINO 27 Jun 2024 · 0 repositories · arXiv:2406.19057
-
Sonnet or Not, Bot? Poetry Evaluation for Large Models and Datasets 27 Jun 2024 · 1 repository · arXiv:2406.18906Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 2 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Structural Attention: Rethinking Transformer for Unpaired Medical Image Synthesis 27 Jun 2024 · 1 repository · arXiv:2406.18967
-
The Model Arena for Cross-lingual Sentiment Analysis: A Comparative Study in the Era of Large Language Models 27 Jun 2024 · 0 repositories · arXiv:2406.19358
-
UniGen: A Unified Framework for Textual Dataset Generation Using Large Language Models 27 Jun 2024 · 1 repository · arXiv:2406.18966Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 13 harvested samples) · 13 pointer-only (licence)
-
YZS-model: A Predictive Model for Organic Drug Solubility Based on Graph Convolutional Networks and Transformer-Attention 27 Jun 2024 · 1 repository · arXiv:2406.19136
-
3D-MVP: 3D Multiview Pretraining for Robotic Manipulation 26 Jun 2024 · 0 repositories · arXiv:2406.18158
-
A Stem-Agnostic Single-Decoder System for Music Source Separation Beyond Four Stems 26 Jun 2024 · 1 repository · arXiv:2406.18747Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
Adversarial Search Engine Optimization for Large Language Models 26 Jun 2024 · 0 repositories · arXiv:2406.18382
-
APIGen: Automated Pipeline for Generating Verifiable and Diverse Function-Calling Datasets 26 Jun 2024 · 0 repositories · arXiv:2406.18518
-
BADGE: BADminton report Generation and Evaluation with LLM 26 Jun 2024 · 1 repository · arXiv:2406.18116
-
Evaluating Quality of Answers for Retrieval-Augmented Generation: A Strong LLM Is All You Need 26 Jun 2024 · 0 repositories · arXiv:2406.18064
-
FactFinders at CheckThat! 2024: Refining Check-worthy Statement Detection with LLMs through Data Pruning 26 Jun 2024 · 1 repository · arXiv:2406.18297
-
Geometric Features Enhanced Human-Object Interaction Detection 26 Jun 2024 · 1 repository · arXiv:2406.18691
-
"Glue pizza and eat rocks" -- Exploiting Vulnerabilities in Retrieval-Augmented Generative Models 26 Jun 2024 · 0 repositories · arXiv:2406.19417
-
Human-Free Automated Prompting for Vision-Language Anomaly Detection: Prompt Optimization with Meta-guiding Prompt Scheme 26 Jun 2024 · 0 repositories · arXiv:2406.18197
-
Improving Entity Recognition Using Ensembles of Deep Learning and Fine-tuned Large Language Models: A Case Study on Adverse Event Extraction from Multiple Sources 26 Jun 2024 · 0 repositories · arXiv:2406.18049
-
Jailbreaking LLMs with Arabic Transliteration and Arabizi 26 Jun 2024 · 1 repository · arXiv:2406.18725Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Knowledge graph enhanced retrieval-augmented generation for failure mode and effects analysis 26 Jun 2024 · 1 repository · arXiv:2406.18114
-
MathOdyssey: Benchmarking Mathematical Problem-Solving Skills in Large Language Models Using Odyssey Math Data 26 Jun 2024 · 3 repositories · arXiv:2406.18321Syntology 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
MFDNet: Multi-Frequency Deflare Network for Efficient Nighttime Flare Removal 26 Jun 2024 · 1 repository · arXiv:2406.18079Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
Multi-step Inference over Unstructured Data 26 Jun 2024 · 0 repositories · arXiv:2406.17987
-
Octo-planner: On-device Language Model for Planner-Action Agents 26 Jun 2024 · 0 repositories · arXiv:2406.18082
-
PianoBART: Symbolic Piano Music Generation and Understanding with Large-Scale Pre-Training 26 Jun 2024 · 1 repository · arXiv:2407.03361
-
Poisoned LangChain: Jailbreak LLMs by LangChain 26 Jun 2024 · 0 repositories · arXiv:2406.18122
-
Re-Ranking Step by Step: Investigating Pre-Filtering for Re-Ranking with Large Language Models 26 Jun 2024 · 0 repositories · arXiv:2406.18740
-
ResumeAtlas: Revisiting Resume Classification with Large-Scale Datasets and Large Language Models 26 Jun 2024 · 1 repository · arXiv:2406.18125
-
Step-DPO: Step-wise Preference Optimization for Long-chain Reasoning of LLMs 26 Jun 2024 · 1 repository · arXiv:2406.18629Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
Themis: A Reference-free NLG Evaluation Language Model with Flexibility and Interpretability 26 Jun 2024 · 1 repository · arXiv:2406.18365Syntology official (archive's flag): 5 ran · 5 ran (of which 3 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
Understand What LLM Needs: Dual Preference Alignment for Retrieval-Augmented Generation 26 Jun 2024 · 1 repository · arXiv:2406.18676Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Unveiling and Controlling Anomalous Attention Distribution in Transformers 26 Jun 2024 · 0 repositories · arXiv:2407.01601
-
WildGuard: Open One-Stop Moderation Tools for Safety Risks, Jailbreaks, and Refusals of LLMs 26 Jun 2024 · 4 repositories · arXiv:2406.18495Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified; the one sample that ran constructed an object rather than computing a result (of 3 harvested samples) · 3 pointer-only (licence)
-
Zero-shot prompt-based classification: topic labeling in times of foundation models in German Tweets 26 Jun 2024 · 0 repositories · arXiv:2406.18239
-
Accelerating Clinical Evidence Synthesis with Large Language Models 25 Jun 2024 · 0 repositories · arXiv:2406.17755
-
ARES: Alternating Reinforcement Learning and Supervised Fine-Tuning for Enhanced Multi-Modal Chain-of-Thought Reasoning Through Diverse AI Feedback 25 Jun 2024 · 1 repository · arXiv:2407.00087
-
Autonomous Prompt Engineering in Large Language Models 25 Jun 2024 · 0 repositories · arXiv:2407.11000
-
SetBERT: Enhancing Retrieval Performance for Boolean Logic and Set Operation Queries 25 Jun 2024 · 0 repositories · arXiv:2406.17282
-
Brain Tumor Classification using Vision Transformer with Selective Cross-Attention Mechanism and Feature Calibration 25 Jun 2024 · 0 repositories · arXiv:2406.17670
-
Cross-Modal Spherical Aggregation for Weakly Supervised Remote Sensing Shadow Removal 25 Jun 2024 · 1 repository · arXiv:2406.17469
-
CTBench: A Comprehensive Benchmark for Evaluating Language Model Capabilities in Clinical Trial Design 25 Jun 2024 · 1 repository · arXiv:2406.17888
-
Dark Transformer: A Video Transformer for Action Recognition in the Dark 25 Jun 2024 · 0 repositories · arXiv:2407.12805
-
Director3D: Real-world Camera Trajectory and 3D Scene Generation from Text 25 Jun 2024 · 1 repository · arXiv:2406.17601Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 3 honoured, 0 violated, 4 with no contract checked; 5 where Syntology's instrument failed) · 3 unverified (of 15 harvested samples) · 15 pointer-only (licence)
-
Discrete Diffusion Language Model for Long Text Summarization 25 Jun 2024 · 0 repositories · arXiv:2407.10998
-
ET tu, CLIP? Addressing Common Object Errors for Unseen Environments 25 Jun 2024 · 0 repositories · arXiv:2406.17876
-
Interpreting Attention Layer Outputs with Sparse Autoencoders 25 Jun 2024 · 1 repository · arXiv:2406.17759Syntology official: harvested, nothing ran · 0 ran · 2 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Knowledge Distillation in Automated Annotation: Supervised Text Classification with LLM-Generated Training Labels 25 Jun 2024 · 0 repositories · arXiv:2406.17633
-
LongIns: A Challenging Long-context Instruction-based Exam for LLMs 25 Jun 2024 · 0 repositories · arXiv:2406.17588
-
LumberChunker: Long-Form Narrative Document Segmentation 25 Jun 2024 · 1 repository · arXiv:2406.17526Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Point Tree Transformer for Point Cloud Registration 25 Jun 2024 · 0 repositories · arXiv:2406.17530
-
Q-DiT: Accurate Post-Training Quantization for Diffusion Transformers 25 Jun 2024 · 1 repository · arXiv:2406.17343Syntology official (archive's flag): 16 ran · 17 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 4 honoured, 0 violated, 7 with no contract checked; 6 where Syntology's instrument failed) · 2 unverified (of 19 harvested samples) · 19 pointer-only (licence)
-
RAGBench: Explainable Benchmark for Retrieval-Augmented Generation Systems 25 Jun 2024 · 0 repositories · arXiv:2407.11005
-
Revitalizing Convolutional Network for Image Restoration 25 Jun 2024 · 1 repository
-
Semi-supervised classification of dental conditions in panoramic radiographs using large language model and instance segmentation: A real-world dataset evaluation 25 Jun 2024 · 0 repositories · arXiv:2406.17915
-
Sound Tagging in Infant-centric Home Soundscapes 25 Jun 2024 · 0 repositories · arXiv:2406.17190
-
Structured Unrestricted-Rank Matrices for Parameter Efficient Fine-tuning 25 Jun 2024 · 1 repository · arXiv:2406.17740Syntology official (archive's flag): 2 ran · 2 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
SUM: Saliency Unification through Mamba for Visual Attention Modeling 25 Jun 2024 · 1 repository · arXiv:2406.17815Syntology official (archive's flag): 14 ran · 14 ran (of which 0 constructed an object rather than computing a result; 14 with no instrument failure: 0 honoured, 0 violated, 14 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 14 harvested samples)
-
Task-Agnostic Federated Learning 25 Jun 2024 · 0 repositories · arXiv:2406.17235
-
Temporal-Channel Modeling in Multi-head Self-Attention for Synthetic Speech Detection 25 Jun 2024 · 1 repository · arXiv:2406.17376
-
This Paper Had the Smartest Reviewers -- Flattery Detection Utilising an Audio-Textual Transformer-Based Approach 25 Jun 2024 · 1 repository · arXiv:2406.17667
-
Towards Optimal Trade-offs in Knowledge Distillation for CNNs and Vision Transformers at the Edge 25 Jun 2024 · 0 repositories · arXiv:2407.12808
-
Univariate Skeleton Prediction in Multivariate Systems Using Transformers 25 Jun 2024 · 1 repository · arXiv:2406.17834
-
Unlocking Continual Learning Abilities in Language Models 25 Jun 2024 · 1 repository · arXiv:2406.17245Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples)
-
Understanding Language Model Circuits through Knowledge Editing 25 Jun 2024 · 0 repositories · arXiv:2406.17241
-
Accelerating Phase Field Simulations Through a Hybrid Adaptive Fourier Neural Operator with U-Net Backbone 24 Jun 2024 · 0 repositories · arXiv:2406.17119
-
Anomaly Detection of Tabular Data Using LLMs 24 Jun 2024 · 0 repositories · arXiv:2406.16308
-
Attention Instruction: Amplifying Attention in the Middle via Prompting 24 Jun 2024 · 1 repository · arXiv:2406.17095
-
Building on Efficient Foundations: Effectively Training LLMs with Structured Feedforward Layers 24 Jun 2024 · 1 repository · arXiv:2406.16450
-
Classification of Geological Borehole Descriptions Using a Domain Adapted Large Language Model 24 Jun 2024 · 0 repositories · arXiv:2407.10991
-
Confidence Regulation Neurons in Language Models 24 Jun 2024 · 1 repository · arXiv:2406.16254Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 5 pointer-only (licence)