Methods › General › Normalization › Layer Normalization › Papers, page 58
Layer Normalization
Papers archive 2025-07-28
archive papers tagged: 24,980 · with a code link: 11,273 · where Syntology ran a sample: 3,471 (2,923 with a run with no instrument failure, 548 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,471 of 24,980 tagged: 2,923 with a run with no instrument failure, 548 where every run was a failure of Syntology's instrument)
Page 58 of 250: papers 5,701 to 5,800 of 24,980, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
I Could've Asked That: Reformulating Unanswerable Questions 24 Jul 2024 · 1 repository · arXiv:2407.17469
-
Improving ICD coding using Chapter based Named Entities and Attentional Models 24 Jul 2024 · 0 repositories · arXiv:2407.17230
-
LoFormer: Local Frequency Transformer for Image Deblurring 24 Jul 2024 · 2 repositories · arXiv:2407.16993
-
MuST: Multi-Scale Transformers for Surgical Phase Recognition 24 Jul 2024 · 1 repository · arXiv:2407.17361
-
Reporting and Analysing the Environmental Impact of Language Models on the Example of Commonsense Question Answering with External Knowledge 24 Jul 2024 · 0 repositories · arXiv:2408.01453
-
Testing Large Language Models on Driving Theory Knowledge and Skills for Connected Autonomous Vehicles 24 Jul 2024 · 0 repositories · arXiv:2407.17211
-
Trans2Unet: Neural fusion for Nuclei Semantic Segmentation 24 Jul 2024 · 0 repositories · arXiv:2407.17181
-
Analyzing Polysemy Evolution Using Semantic Cells 23 Jul 2024 · 0 repositories · arXiv:2407.16110
-
Artificial Intelligence in Extracting Diagnostic Data from Dental Records 23 Jul 2024 · 0 repositories · arXiv:2407.21050
-
Channel-Partitioned Windowed Attention And Frequency Learning for Single Image Super-Resolution 23 Jul 2024 · 0 repositories · arXiv:2407.16232
-
Data Mixture Inference: What do BPE Tokenizers Reveal about their Training Data? 23 Jul 2024 · 1 repository · arXiv:2407.16607Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Diffusion Transformer Captures Spatial-Temporal Dependencies: A Theory for Gaussian Process Data 23 Jul 2024 · 0 repositories · arXiv:2407.16134
-
Do LLMs Know When to NOT Answer? Investigating Abstention Abilities of Large Language Models 23 Jul 2024 · 0 repositories · arXiv:2407.16221
-
Enhancing LLM's Cognition via Structurization 23 Jul 2024 · 1 repository · arXiv:2407.16434Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples)
-
Exploring The Neural Burden In Pruned Models: An Insight Inspired By Neuroscience 23 Jul 2024 · 0 repositories · arXiv:2407.16716
-
HSVLT: Hierarchical Scale-Aware Vision-Language Transformer for Multi-Label Image Classification 23 Jul 2024 · 0 repositories · arXiv:2407.16244
-
HyTAS: A Hyperspectral Image Transformer Architecture Search Benchmark and Analysis 23 Jul 2024 · 1 repository · arXiv:2407.16269
-
LawLuo: A Multi-Agent Collaborative Framework for Multi-Round Chinese Legal Consultation 23 Jul 2024 · 0 repositories · arXiv:2407.16252
-
Lawma: The Power of Specialization for Legal Tasks 23 Jul 2024 · 0 repositories · arXiv:2407.16615Syntology 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
OriGen:Enhancing RTL Code Generation with Code-to-Code Augmentation and Self-Reflection 23 Jul 2024 · 1 repository · arXiv:2407.16237Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Patched RTC: evaluating LLMs for diverse software development tasks 23 Jul 2024 · 1 repository · arXiv:2407.16557
-
RedAgent: Red Teaming Large Language Models with Context-aware Autonomous Language Agent 23 Jul 2024 · 0 repositories · arXiv:2407.16667
-
Retrieval Augmented Generation or Long-Context LLMs? A Comprehensive Study and Hybrid Approach 23 Jul 2024 · 0 repositories · arXiv:2407.16833
-
Robust Privacy Amidst Innovation with Large Language Models Through a Critical Assessment of the Risks 23 Jul 2024 · 1 repository · arXiv:2407.16166
-
S-E Pipeline: A Vision Transformer (ViT) based Resilient Classification Pipeline for Medical Imaging Against Adversarial Attacks 23 Jul 2024 · 0 repositories · arXiv:2407.17587
-
SINDER: Repairing the Singular Defects of DINOv2 23 Jul 2024 · 1 repository · arXiv:2407.16826Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample)
-
Synthesizer Sound Matching Using Audio Spectrogram Transformers 23 Jul 2024 · 0 repositories · arXiv:2407.16643
-
TookaBERT: A Step Forward for Persian NLU 23 Jul 2024 · 0 repositories · arXiv:2407.16382
-
An Empirical Comparison of Video Frame Sampling Methods for Multi-Modal RAG Retrieval 22 Jul 2024 · 0 repositories · arXiv:2408.03340
-
Can GPT-4 learn to analyse moves in research article abstracts? 22 Jul 2024 · 0 repositories · arXiv:2407.15612
-
Customized Retrieval Augmented Generation and Benchmarking for EDA Tool Documentation QA 22 Jul 2024 · 0 repositories · arXiv:2407.15353
-
Dissecting Multiplication in Transformers: Insights into LLMs 22 Jul 2024 · 1 repository · arXiv:2407.15360
-
Efficient Multi-disparity Transformer for Light Field Image Super-resolution 22 Jul 2024 · 0 repositories · arXiv:2407.15329
-
Estimating Probability Densities with Transformer and Denoising Diffusion 22 Jul 2024 · 1 repository · arXiv:2407.15703Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 1 pointer-only (licence)
-
Impacts of Anthropomorphizing Large Language Models in Learning Environments 22 Jul 2024 · 0 repositories · arXiv:2408.03945
-
Imposter.AI: Adversarial Attacks with Hidden Intentions towards Aligned Large Language Models 22 Jul 2024 · 0 repositories · arXiv:2407.15399
-
Inverted Activations: Reducing Memory Footprint in Neural Network Training 22 Jul 2024 · 1 repository · arXiv:2407.15545
-
KWT-Tiny: RISC-V Accelerated, Embedded Keyword Spotting Transformer 22 Jul 2024 · 0 repositories · arXiv:2407.16026
-
Learning to Manipulate Anywhere: A Visual Generalizable Framework For Reinforcement Learning 22 Jul 2024 · 0 repositories · arXiv:2407.15815
-
LLMmap: Fingerprinting For Large Language Models 22 Jul 2024 · 1 repository · arXiv:2407.15847Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 7 unverified (of 16 harvested samples)
-
Local All-Pair Correspondence for Point Tracking 22 Jul 2024 · 2 repositories · arXiv:2407.15420Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 12 harvested samples) · 2 pointer-only (licence)
-
Mini-Sequence Transformer: Optimizing Intermediate Memory for Long Sequences Training 22 Jul 2024 · 1 repository · arXiv:2407.15892
-
MMInstruct: A High-Quality Multi-Modal Instruction Tuning Dataset with Extensive Diversity 22 Jul 2024 · 1 repository · arXiv:2407.15838Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
MoRSE: Bridging the Gap in Cybersecurity Expertise with Retrieval Augmented Generation 22 Jul 2024 · 0 repositories · arXiv:2407.15748
-
NV-Retriever: Improving text embedding models with effective hard-negative mining 22 Jul 2024 · 0 repositories · arXiv:2407.15831
-
Predicting the Best of N Visual Trackers 22 Jul 2024 · 1 repository · arXiv:2407.15707
-
Promises and Pitfalls of Generative Masked Language Modeling: Theoretical Framework and Practical Guidelines 22 Jul 2024 · 1 repository · arXiv:2407.21046
-
RadioRAG: Factual large language models for enhanced diagnostics in radiology using online retrieval augmented generation 22 Jul 2024 · 1 repository · arXiv:2407.15621
-
Stretching Each Dollar: Diffusion Training from Scratch on a Micro-Budget 22 Jul 2024 · 1 repository · arXiv:2407.15811Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples)
-
Unlocking the Potential: Benchmarking Large Language Models in Water Engineering and Research 22 Jul 2024 · 0 repositories · arXiv:2407.21045
-
ZZU-NLP at SIGHAN-2024 dimABSA Task: Aspect-Based Sentiment Analysis with Coarse-to-Fine In-context Learning 22 Jul 2024 · 0 repositories · arXiv:2407.15341
-
A multi-level multi-label text classification dataset of 19th century Ottoman and Russian literary and critical texts 21 Jul 2024 · 0 repositories · arXiv:2407.15136
-
Arondight: Red Teaming Large Vision Language Models with Auto-generated Multi-modal Jailbreak Prompts 21 Jul 2024 · 0 repositories · arXiv:2407.15050
-
Decoding Multilingual Moral Preferences: Unveiling LLM's Biases Through the Moral Machine Experiment 21 Jul 2024 · 1 repository · arXiv:2407.15184
-
Efficient Visual Transformer by Learnable Token Merging 21 Jul 2024 · 1 repository · arXiv:2407.15219
-
Point Transformer V3 Extreme: 1st Place Solution for 2024 Waymo Open Dataset Challenge in Semantic Segmentation 21 Jul 2024 · 0 repositories · arXiv:2407.15282
-
Prior Knowledge Integration via LLM Encoding and Pseudo Event Regulation for Video Moment Retrieval 21 Jul 2024 · 1 repository · arXiv:2407.15051Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 1 violated, 4 with no contract checked; 3 where Syntology's instrument failed) · 3 unverified (of 11 harvested samples) · 9 pointer-only (licence)
-
Toward Adaptive Reasoning in Large Language Models with Thought Rollback 21 Jul 2024 · 1 repository
-
Golden-Retriever: High-Fidelity Agentic Retrieval Augmented Generation for Industrial Knowledge Base 20 Jul 2024 · 0 repositories · arXiv:2408.00798
-
Automatic Generation of Fashion Images using Prompting in Generative Machine Learning Models 20 Jul 2024 · 1 repository · arXiv:2407.14944
-
Differential Privacy of Cross-Attention with Provable Guarantee 20 Jul 2024 · 0 repositories · arXiv:2407.14717
-
FairViT: Fair Vision Transformer via Adaptive Masking 20 Jul 2024 · 1 repository · arXiv:2407.14799
-
GaitMA: Pose-guided Multi-modal Feature Fusion for Gait Recognition 20 Jul 2024 · 0 repositories · arXiv:2407.14812
-
Improving Context-Aware Preference Modeling for Language Models 20 Jul 2024 · 0 repositories · arXiv:2407.14916
-
RGB2Point: 3D Point Cloud Generation from Single RGB Images 20 Jul 2024 · 0 repositories · arXiv:2407.14979
-
Step-by-Step Reasoning to Solve Grid Puzzles: Where do LLMs Falter? 20 Jul 2024 · 1 repository · arXiv:2407.14790
-
Technical report: Improving the properties of molecules generated by LIMO 20 Jul 2024 · 0 repositories · arXiv:2407.14968
-
TraveLLM: Could you plan my new public transit route in face of a network disruption? 20 Jul 2024 · 0 repositories · arXiv:2407.14926
-
Adversarial Databases Improve Success in Retrieval-based Large Language Models 19 Jul 2024 · 0 repositories · arXiv:2407.14609
-
ChatQA 2: Bridging the Gap to Proprietary LLMs in Long Context and RAG Capabilities 19 Jul 2024 · 0 repositories · arXiv:2407.14482
-
Comparing and Contrasting Deep Learning Weather Prediction Backbones on Navier-Stokes and Atmospheric Dynamics 19 Jul 2024 · 1 repository · arXiv:2407.14129Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
Unipa-GPT: Large Language Models for university-oriented QA in Italian 19 Jul 2024 · 1 repository · arXiv:2407.14246
-
Domain-Specific Pretraining of Language Models: A Comparative Study in the Medical Field 19 Jul 2024 · 0 repositories · arXiv:2407.14076
-
Double-Shot 3D Shape Measurement with a Dual-Branch Network for Structured Light Projection Profilometry 19 Jul 2024 · 0 repositories · arXiv:2407.14198
-
Generative Language Model for Catalyst Discovery 19 Jul 2024 · 0 repositories · arXiv:2407.14040
-
HeCiX: Integrating Knowledge Graphs and Large Language Models for Biomedical Research 19 Jul 2024 · 0 repositories · arXiv:2407.14030
-
Impact of Model Size on Fine-tuned LLM Performance in Data-to-Text Generation: A State-of-the-Art Investigation 19 Jul 2024 · 0 repositories · arXiv:2407.14088
-
Improving Representation of High-frequency Components for Medical Visual Foundation Models 19 Jul 2024 · 1 repository · arXiv:2407.14651Syntology official (archive's flag): 15 ran · 15 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 0 honoured, 0 violated, 13 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 17 harvested samples) · 1 pointer-only (licence)
-
Improving Retrieval in Sponsored Search by Leveraging Query Context Signals 19 Jul 2024 · 0 repositories · arXiv:2407.14346
-
LAPIS: Language Model-Augmented Police Investigation System 19 Jul 2024 · 0 repositories · arXiv:2407.20248
-
LLMs left, right, and center: Assessing GPT's capabilities to label political bias from web domains 19 Jul 2024 · 0 repositories · arXiv:2407.14344
-
Longhorn: State Space Models are Amortized Online Learners 19 Jul 2024 · 1 repository · arXiv:2407.14207Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
LORTSAR: Low-Rank Transformer for Skeleton-based Action Recognition 19 Jul 2024 · 0 repositories · arXiv:2407.14655
-
Modality-Order Matters! A Novel Hierarchical Feature Fusion Method for CoSAm: A Code-Switched Autism Corpus 19 Jul 2024 · 0 repositories · arXiv:2407.14328
-
MSCT: Addressing Time-Varying Confounding with Marginal Structural Causal Transformer for Counterfactual Post-Crash Traffic Prediction 19 Jul 2024 · 0 repositories · arXiv:2407.14065
-
On the use of Probabilistic Forecasting for Network Analysis in Open RAN 19 Jul 2024 · 0 repositories · arXiv:2407.14375
-
PolyFormer: Scalable Node-wise Filters via Polynomial Graph Transformer 19 Jul 2024 · 1 repository · arXiv:2407.14459Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 1 honoured, 0 violated, 7 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 14 harvested samples) · 14 pointer-only (licence)
-
SQLfuse: Enhancing Text-to-SQL Performance through Comprehensive LLM Synergy 19 Jul 2024 · 0 repositories · arXiv:2407.14568
-
TorchGT: A Holistic System for Large-scale Graph Transformer Training 19 Jul 2024 · 0 repositories · arXiv:2407.14106
-
A light-weight and efficient punctuation and word casing prediction model for on-device streaming ASR 18 Jul 2024 · 0 repositories · arXiv:2407.13142
-
Scalable Exploration via Ensemble++ 18 Jul 2024 · 2 repositories · arXiv:2407.13195Syntology official (archive's flag): 5 ran · 5 ran (of which 1 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
Black-Box Opinion Manipulation Attacks to Retrieval-Augmented Generation of Large Language Models 18 Jul 2024 · 0 repositories · arXiv:2407.13757
-
Can Open-Source LLMs Compete with Commercial Models? Exploring the Few-Shot Performance of Current GPT Models in Biomedical Tasks 18 Jul 2024 · 1 repository · arXiv:2407.13511
-
Continual Distillation Learning: Knowledge Distillation in Prompt-based Continual Learning 18 Jul 2024 · 0 repositories · arXiv:2407.13911
-
Autonomous self-evolving research on biomedical data: the DREAM paradigm 18 Jul 2024 · 0 repositories · arXiv:2407.13637
-
Evaluating Large Language Models for Anxiety and Depression Classification using Counseling and Psychotherapy Transcripts 18 Jul 2024 · 1 repository · arXiv:2407.13228
-
GPSFormer: A Global Perception and Local Structure Fitting-based Transformer for Point Cloud Understanding 18 Jul 2024 · 1 repository · arXiv:2407.13519
-
HHGT: Hierarchical Heterogeneous Graph Transformer for Heterogeneous Graph Representation Learning 18 Jul 2024 · 0 repositories · arXiv:2407.13158
-
How Reliable are LLMs as Knowledge Bases? Re-thinking Facutality and Consistency 18 Jul 2024 · 0 repositories · arXiv:2407.13578
-
Learning-From-Mistakes Prompting for Indigenous Language Translation 18 Jul 2024 · 0 repositories · arXiv:2407.13343