Methods › General › Attention Mechanisms › Attention › Papers, page 9
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 9 of 316: papers 801 to 900 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
HDLxGraph: Bridging Large Language Models and HDL Repositories via HDL Graph Databases 21 May 2025 · 1 repository · arXiv:2505.15701
-
Higher-order Structure Boosts Link Prediction on Temporal Graphs 21 May 2025 · 0 repositories · arXiv:2505.15746
-
Hunyuan-TurboS: Advancing Large Language Models through Mamba-Transformer Synergy and Adaptive Chain-of-Thought 21 May 2025 · 0 repositories · arXiv:2505.15431
-
InfoDeepSeek: Benchmarking Agentic Information Seeking for Retrieval-Augmented Generation 21 May 2025 · 0 repositories · arXiv:2505.15872
-
Internal and External Impacts of Natural Language Processing Papers 21 May 2025 · 0 repositories · arXiv:2505.16061
-
Interspatial Attention for Efficient 4D Human Video Generation 21 May 2025 · 0 repositories · arXiv:2505.15800
-
Leveraging Foundation Models for Multimodal Graph-Based Action Recognition 21 May 2025 · 0 repositories · arXiv:2505.15192
-
Leveraging Large Language Models for Command Injection Vulnerability Analysis in Python: An Empirical Study on Popular Open-Source Projects 21 May 2025 · 0 repositories · arXiv:2505.15088
-
Leveraging the Powerful Attention of a Pre-trained Diffusion Model for Exemplar-based Image Colorization 21 May 2025 · 1 repository · arXiv:2505.15812
-
LFTF: Locating First and Then Fine-Tuning for Mitigating Gender Bias in Large Language Models 21 May 2025 · 0 repositories · arXiv:2505.15475
-
LogiCase: Effective Test Case Generation from Logical Description in Competitive Programming 21 May 2025 · 0 repositories · arXiv:2505.15039Syntology 11 ran (of which 3 constructed an object rather than computing a result; 5 with no instrument failure: 2 honoured, 0 violated, 3 with no contract checked; 6 where Syntology's instrument failed) · 8 unverified (of 19 harvested samples) · 19 pointer-only (licence)
-
Lossless Token Merging Even Without Fine-Tuning in Vision Transformers 21 May 2025 · 0 repositories · arXiv:2505.15160
-
MaxPoolBERT: Enhancing BERT Classification via Layer- and Token-Wise Aggregation 21 May 2025 · 0 repositories · arXiv:2505.15696
-
Mechanistic Insights into Grokking from the Embedding Layer 21 May 2025 · 0 repositories · arXiv:2505.15624
-
Mitigating Spurious Correlations with Causal Logit Perturbation 21 May 2025 · 0 repositories · arXiv:2505.15246
-
MonoSplat: Generalizable 3D Gaussian Splatting from Monocular Depth Foundation Models 21 May 2025 · 1 repository · arXiv:2505.15185Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 6 harvested samples)
-
Moonbeam: A MIDI Foundation Model Using Both Absolute and Relative Music Attributes 21 May 2025 · 1 repository · arXiv:2505.15559Syntology 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 1 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (of 11 harvested samples) · 2 pointer-only (licence)
-
NL-Debugging: Exploiting Natural Language as an Intermediate Representation for Code Debugging 21 May 2025 · 0 repositories · arXiv:2505.15356
-
Oversmoothing, "Oversquashing", Heterophily, Long-Range, and more: Demystifying Common Beliefs in Graph Machine Learning 21 May 2025 · 0 repositories · arXiv:2505.15547
-
Ranking Free RAG: Replacing Re-ranking with Selection in RAG for Sensitive Domains 21 May 2025 · 0 repositories · arXiv:2505.16014
-
RePPL: Recalibrating Perplexity by Uncertainty in Semantic Propagation and Language Generation for Explainable QA Hallucination Detection 21 May 2025 · 0 repositories · arXiv:2505.15386
-
Reranking with Compressed Document Representation 21 May 2025 · 0 repositories · arXiv:2505.15394
-
RLBenchNet: The Right Network for the Right Reinforcement Learning Task 21 May 2025 · 1 repository · arXiv:2505.15040
-
Robo-DM: Data Management For Large Robot Datasets 21 May 2025 · 0 repositories · arXiv:2505.15558
-
RoT: Enhancing Table Reasoning with Iterative Row-Wise Traversals 21 May 2025 · 0 repositories · arXiv:2505.15110
-
SAMA-UNet: Enhancing Medical Image Segmentation with Self-Adaptive Mamba-Like Attention and Causal-Resonance Learning 21 May 2025 · 1 repository · arXiv:2505.15234
-
Scaling Diffusion Transformers Efficiently via μP 21 May 2025 · 1 repository · arXiv:2505.15270
-
Seeing the Trees for the Forest: Rethinking Weakly-Supervised Medical Visual Grounding 21 May 2025 · 0 repositories · arXiv:2505.15123
-
Set-LLM: A Permutation-Invariant LLM 21 May 2025 · 0 repositories · arXiv:2505.15433
-
Short-Range Dependency Effects on Transformer Instability and a Decomposed Attention Solution 21 May 2025 · 0 repositories · arXiv:2505.15548
-
Silent Leaks: Implicit Knowledge Extraction Attack on RAG Systems through Benign Queries 21 May 2025 · 1 repository · arXiv:2505.15420
-
Single LLM, Multiple Roles: A Unified Retrieval-Augmented Generation Framework Using Role-Specific Token Optimization 21 May 2025 · 0 repositories · arXiv:2505.15444
-
Small Language Models in the Real World: Insights from Industrial Text Classification 21 May 2025 · 0 repositories · arXiv:2505.16078
-
Sonnet: Spectral Operator Neural Network for Multivariable Time Series Forecasting 21 May 2025 · 1 repository · arXiv:2505.15312
-
SUS backprop: linear backpropagation algorithm for long inputs in transformers 21 May 2025 · 0 repositories · arXiv:2505.15080
-
The Atlas of In-Context Learning: How Attention Heads Shape In-Context Retrieval Augmentation 21 May 2025 · 1 repository · arXiv:2505.15807Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Time Tracker: Mixture-of-Experts-Enhanced Foundation Time Series Forecasting Model with Decoupled Training Pipelines 21 May 2025 · 0 repositories · arXiv:2505.15151
-
UAV-Flow Colosseo: A Real-World Benchmark for Flying-on-a-Word UAV Imitation Learning 21 May 2025 · 0 repositories · arXiv:2505.15725
-
Unified Cross-Modal Attention-Mixer Based Structural-Functional Connectomics Fusion for Neuropsychiatric Disorder Diagnosis 21 May 2025 · 0 repositories · arXiv:2505.15139
-
When Can Large Reasoning Models Save Thinking? Mechanistic Analysis of Behavioral Divergence in Reasoning 21 May 2025 · 0 repositories · arXiv:2505.15276
-
A Direct Comparison of Simultaneously Recorded Scalp, Around-Ear, and In-Ear EEG for Neural Selective Auditory Attention Decoding to Speech 20 May 2025 · 0 repositories · arXiv:2505.14478
-
A Logic of General Attention Using Edge-Conditioned Event Models (Extended Version) 20 May 2025 · 0 repositories · arXiv:2505.14539
-
AAPO: Enhance the Reasoning Capabilities of LLMs with Advantage Momentum 20 May 2025 · 0 repositories · arXiv:2505.14264
-
AI-empowered Channel Estimation for Block-based Active IRS-enhanced Hybrid-field IoT Network 20 May 2025 · 0 repositories · arXiv:2505.14098
-
Aligning Attention Distribution to Information Flow for Hallucination Mitigation in Large Vision-Language Models 20 May 2025 · 0 repositories · arXiv:2505.14257
-
Articulatory Feature Prediction from Surface EMG during Speech Production 20 May 2025 · 1 repository · arXiv:2505.13814
-
Automated Quality Evaluation of Cervical Cytopathology Whole Slide Images Based on Content Analysis 20 May 2025 · 0 repositories · arXiv:2505.13875
-
Automatic Dataset Generation for Knowledge Intensive Question Answering Tasks 20 May 2025 · 0 repositories · arXiv:2505.14212
-
Beyond Text: Unveiling Privacy Vulnerabilities in Multi-modal Retrieval-Augmented Generation 20 May 2025 · 0 repositories · arXiv:2505.13957
-
Breaking Bad Tokens: Detoxification of LLMs Using Sparse Autoencoders 20 May 2025 · 0 repositories · arXiv:2505.14536Syntology 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
CAD-Coder: An Open-Source Vision-Language Model for Computer-Aided Design Code Generation 20 May 2025 · 1 repository · arXiv:2505.14646Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Choosing a Model, Shaping a Future: Comparing LLM Perspectives on Sustainability and its Relationship with AI 20 May 2025 · 0 repositories · arXiv:2505.14435
-
Cost-Augmented Monte Carlo Tree Search for LLM-Assisted Planning 20 May 2025 · 0 repositories · arXiv:2505.14656
-
CSAGC-IDS: A Dual-Module Deep Learning Network Intrusion Detection Model for Complex and Imbalanced Data 20 May 2025 · 0 repositories · arXiv:2505.14027
-
Disentangled Multi-span Evolutionary Network against Temporal Knowledge Graph Reasoning 20 May 2025 · 0 repositories · arXiv:2505.14020
-
Divide by Question, Conquer by Agent: SPLIT-RAG with Question-Driven Graph Partitioning 20 May 2025 · 0 repositories · arXiv:2505.13994
-
Do Language Models Use Their Depth Efficiently? 20 May 2025 · 1 repository · arXiv:2505.13898Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples)
-
DSMentor: Enhancing Data Science Agents with Curriculum Learning and Online Knowledge Accumulation 20 May 2025 · 0 repositories · arXiv:2505.14163
-
EEG-to-Text Translation: A Model for Deciphering Human Brain Activity 20 May 2025 · 1 repository · arXiv:2505.13936
-
EfficientLLM: Efficiency in Large Language Models 20 May 2025 · 0 repositories · arXiv:2505.13840
-
Embedded Mean Field Reinforcement Learning for Perimeter-defense Game 20 May 2025 · 0 repositories · arXiv:2505.14209
-
Energy-Efficient Deep Reinforcement Learning with Spiking Transformers 20 May 2025 · 0 repositories · arXiv:2505.14533
-
Enhancing Abstractive Summarization of Scientific Papers Using Structure Information 20 May 2025 · 1 repository · arXiv:2505.14179
-
EVA: Red-Teaming GUI Agents via Evolving Indirect Prompt Injection 20 May 2025 · 0 repositories · arXiv:2505.14289
-
Every Pixel Tells a Story: End-to-End Urdu Newspaper OCR 20 May 2025 · 0 repositories · arXiv:2505.13943
-
Exploring Image Quality Assessment from a New Perspective: Pupil Size 20 May 2025 · 0 repositories · arXiv:2505.13841
-
Exploring Jailbreak Attacks on LLMs through Intent Concealment and Diversion 20 May 2025 · 0 repositories · arXiv:2505.14316
-
FLASH-D: FlashAttention with Hidden Softmax Division 20 May 2025 · 0 repositories · arXiv:2505.14201
-
FlashKAT: Understanding and Addressing Performance Bottlenecks in the Kolmogorov-Arnold Transformer 20 May 2025 · 1 repository · arXiv:2505.13813
-
FlowQ: Energy-Guided Flow Policies for Offline Reinforcement Learning 20 May 2025 · 0 repositories · arXiv:2505.14139
-
Generative AI at the Crossroads: Light Bulb, Dynamo, or Microscope? 20 May 2025 · 0 repositories · arXiv:2505.14588
-
Grouping First, Attending Smartly: Training-Free Acceleration for Diffusion Transformers 20 May 2025 · 1 repository · arXiv:2505.14687
-
HausaNLP: Current Status, Challenges and Future Directions for Hausa Natural Language Processing 20 May 2025 · 0 repositories · arXiv:2505.14311
-
Informatics for Food Processing 20 May 2025 · 1 repository · arXiv:2505.17087
-
JOLT-SQL: Joint Loss Tuning of Text-to-SQL with Confusion-aware Noisy Schema Sampling 20 May 2025 · 1 repository · arXiv:2505.14305Syntology official: no sample here; runs from other or unrecorded repositories · 16 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 3 violated, 9 with no contract checked; 4 where Syntology's instrument failed) · 2 unverified (of 18 harvested samples) · 1 pointer-only (licence)
-
Large Language Model-Driven Distributed Integrated Multimodal Sensing and Semantic Communications 20 May 2025 · 0 repositories · arXiv:2505.18194
-
Latent Flow Transformer 20 May 2025 · 1 repository · arXiv:2505.14513
-
LEANCODE: Understanding Models Better for Code Simplification of Pre-trained Large Language Models 20 May 2025 · 0 repositories · arXiv:2505.14759
-
Learning Spatio-Temporal Dynamics for Trajectory Recovery via Time-Aware Transformer 20 May 2025 · 2 repositories · arXiv:2505.13857
-
LOD1 3D City Model from LiDAR: The Impact of Segmentation Accuracy on Quality of Urban 3D Modeling and Morphology Extraction 20 May 2025 · 1 repository · arXiv:2505.14747
-
Low-Cost FlashAttention with Fused Exponential and Multiplication Hardware Operators 20 May 2025 · 0 repositories · arXiv:2505.14314
-
Mechanistic Fine-tuning for In-context Learning 20 May 2025 · 0 repositories · arXiv:2505.14233
-
MGStream: Motion-aware 3D Gaussian for Streamable Dynamic Scene Reconstruction 20 May 2025 · 1 repository · arXiv:2505.13839
-
ModRWKV: Transformer Multimodality in Linear Time 20 May 2025 · 1 repository · arXiv:2505.14505
-
MSDformer: Multi-scale Discrete Transformer For Time Series Generation 20 May 2025 · 0 repositories · arXiv:2505.14202
-
Multi-Channel Swin Transformer Framework for Bearing Remaining Useful Life Prediction 20 May 2025 · 0 repositories · arXiv:2505.14897
-
Multimodal RAG-driven Anomaly Detection and Classification in Laser Powder Bed Fusion using Large Language Models 20 May 2025 · 0 repositories · arXiv:2505.13828
-
OmniStyle: Filtering High Quality Style Transfer Data at Scale 20 May 2025 · 1 repository · arXiv:2505.14028
-
Out-of-Distribution Generalization of In-Context Learning: A Low-Dimensional Subspace Perspective 20 May 2025 · 0 repositories · arXiv:2505.14808
-
Pierce the Mists, Greet the Sky: Decipher Knowledge Overshadowing via Knowledge Circuit Analysis 20 May 2025 · 0 repositories · arXiv:2505.14406Syntology 7 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 5 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples)
-
Plane Geometry Problem Solving with Multi-modal Reasoning: A Survey 20 May 2025 · 0 repositories · arXiv:2505.14340
-
Polar Sparsity: High Throughput Batched LLM Inferencing with Scalable Contextual Sparsity 20 May 2025 · 1 repository · arXiv:2505.14884Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Predicting Neo-Adjuvant Chemotherapy Response in Triple-Negative Breast Cancer Using Pre-Treatment Histopathologic Images 20 May 2025 · 0 repositories · arXiv:2505.14730
-
Probing BERT for German Compound Semantics 20 May 2025 · 0 repositories · arXiv:2505.14130
-
Process vs. Outcome Reward: Which is Better for Agentic RAG Reinforcement Learning 20 May 2025 · 1 repository · arXiv:2505.14069Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
ReactDiff: Latent Diffusion for Facial Reaction Generation 20 May 2025 · 1 repository · arXiv:2505.14151
-
s3: You Don't Need That Much Data to Train a Search Agent via RL 20 May 2025 · 1 repository · arXiv:2505.14146Syntology official (archive's flag): 7 ran · 7 ran (of which 3 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 4 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples)
-
SAE-FiRE: Enhancing Earnings Surprise Predictions Through Sparse Autoencoder Feature Selection 20 May 2025 · 0 repositories · arXiv:2505.14420
-
SAFEPATH: Preventing Harmful Reasoning in Chain-of-Thought via Early Alignment 20 May 2025 · 0 repositories · arXiv:2505.14667
-
Sample and Computationally Efficient Continuous-Time Reinforcement Learning with General Function Approximation 20 May 2025 · 1 repository · arXiv:2505.14821