Methods › General › Attention Mechanisms › Attention › Papers, page 13
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 13 of 316: papers 1,201 to 1,300 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Knowledge-Informed Deep Learning for Irrigation Type Mapping from Remote Sensing 13 May 2025 · 0 repositories · arXiv:2505.08302
-
LaDi-WM: A Latent Diffusion-based World Model for Predictive Manipulation 13 May 2025 · 0 repositories · arXiv:2505.11528
-
Lie Group Symmetry Discovery and Enforcement Using Vector Fields 13 May 2025 · 0 repositories · arXiv:2505.08219
-
Lost in Transmission: When and Why LLMs Fail to Reason Globally 13 May 2025 · 0 repositories · arXiv:2505.08140Syntology 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
Memorization-Compression Cycles Improve Generalization 13 May 2025 · 0 repositories · arXiv:2505.08727
-
Multimodal Fusion of Glucose Monitoring and Food Imagery for Caloric Content Prediction 13 May 2025 · 0 repositories · arXiv:2505.09018
-
OpenThinkIMG: Learning to Think with Images via Visual Tool Reinforcement Learning 13 May 2025 · 1 repository · arXiv:2505.08617
-
Optimizing Retrieval-Augmented Generation: Analysis of Hyperparameter Impact on Performance and Efficiency 13 May 2025 · 0 repositories · arXiv:2505.08445
-
Probability Consistency in Large Language Models: Theoretical Foundations Meet Empirical Discrepancies 13 May 2025 · 1 repository · arXiv:2505.08739
-
SAR-GTR: Attributed Scattering Information Guided SAR Graph Transformer Recognition Algorithm 13 May 2025 · 0 repositories · arXiv:2505.08547
-
Scaling Context, Not Parameters: Training a Compact 7B Language Model for Efficient Long-Context Processing 13 May 2025 · 0 repositories · arXiv:2505.08651
-
Securing RAG: A Risk Assessment and Mitigation Framework 13 May 2025 · 0 repositories · arXiv:2505.08728
-
Skeleton-Guided Diffusion Model for Accurate Foot X-ray Synthesis in Hallux Valgus Diagnosis 13 May 2025 · 1 repository · arXiv:2505.08247
-
Small but Significant: On the Promise of Small Language Models for Accessible AIED 13 May 2025 · 0 repositories · arXiv:2505.08588
-
SPAT: Sensitivity-based Multihead-attention Pruning on Time Series Forecasting Models 13 May 2025 · 0 repositories · arXiv:2505.08768
-
Structural-Temporal Coupling Anomaly Detection with Dynamic Graph Transformer 13 May 2025 · 1 repository · arXiv:2505.08330
-
The Truth Becomes Clearer Through Debate! Multi-Agent Systems with Large Language Models Unmask Fake News 13 May 2025 · 0 repositories · arXiv:2505.08532
-
Thermal Detection of People with Mobility Restrictions for Barrier Reduction at Traffic Lights Controlled Intersections 13 May 2025 · 1 repository · arXiv:2505.08568
-
TiMo: Spatiotemporal Foundation Model for Satellite Image Time Series 13 May 2025 · 1 repository · arXiv:2505.08723
-
Ultrasound Report Generation with Multimodal Large Language Models for Standardized Texts 13 May 2025 · 0 repositories · arXiv:2505.08838
-
WaveGuard: Robust Deepfake Detection and Source Tracing via Dual-Tree Complex Wavelet and Graph Neural Networks 13 May 2025 · 1 repository · arXiv:2505.08614
-
WixQA: A Multi-Dataset Benchmark for Enterprise Retrieval-Augmented Generation 13 May 2025 · 0 repositories · arXiv:2505.08643
-
A Comparative Study of Transformer-Based Models for Multi-Horizon Blood Glucose Prediction 12 May 2025 · 1 repository · arXiv:2505.08821
-
A Generative Re-ranking Model for List-level Multi-objective Optimization at Taobao 12 May 2025 · 0 repositories · arXiv:2505.07197
-
A Multi-Dimensional Constraint Framework for Evaluating and Improving Instruction Following in Large Language Models 12 May 2025 · 1 repository · arXiv:2505.07591
-
AIS Data-Driven Maritime Monitoring Based on Transformer: A Comprehensive Review 12 May 2025 · 1 repository · arXiv:2505.07374
-
An Extra RMSNorm is All You Need for Fine Tuning to 1.58 Bits 12 May 2025 · 0 repositories · arXiv:2505.08823
-
Anatomical Attention Alignment representation for Radiology Report Generation 12 May 2025 · 1 repository · arXiv:2505.07689
-
AttentionInfluence: Adopting Attention Head Influence for Weak-to-Strong Pretraining Data Selection 12 May 2025 · 0 repositories · arXiv:2505.07293
-
Automated Visual Attention Detection using Mobile Eye Tracking in Behavioral Classroom Studies 12 May 2025 · 0 repositories · arXiv:2505.07552
-
Benchmarking Retrieval-Augmented Generation for Chemistry 12 May 2025 · 0 repositories · arXiv:2505.07671
-
Breast Cancer Classification in Deep Ultraviolet Fluorescence Images Using a Patch-Level Vision Transformer Framework 12 May 2025 · 0 repositories · arXiv:2505.07654
-
Comparative sentiment analysis of public perception: Monkeypox vs. COVID-19 behavioral insights 12 May 2025 · 0 repositories · arXiv:2505.07430
-
DynamicRAG: Leveraging Outputs of Large Language Model as Feedback for Dynamic Reranking in Retrieval-Augmented Generation 12 May 2025 · 1 repository · arXiv:2505.07233
-
Efficient and Reproducible Biomedical Question Answering using Retrieval Augmented Generation 12 May 2025 · 1 repository · arXiv:2505.07917
-
Examining the Role of LLM-Driven Interactions on Attention and Cognitive Engagement in Virtual Classrooms 12 May 2025 · 0 repositories · arXiv:2505.07377
-
Fused3S: Fast Sparse Attention on Tensor Cores 12 May 2025 · 1 repository · arXiv:2505.08098
-
Generative Pre-trained Autoregressive Diffusion Transformer 12 May 2025 · 0 repositories · arXiv:2505.07344
-
GIFStream: 4D Gaussian-based Immersive Video with Feature Stream 12 May 2025 · 0 repositories · arXiv:2505.07539
-
GRADA: Graph-based Reranker against Adversarial Documents Attack 12 May 2025 · 1 repository · arXiv:2505.07546
-
HALO: Half Life-Based Outdated Fact Filtering in Temporal Knowledge Graphs 12 May 2025 · 1 repository · arXiv:2505.07509
-
HAMLET: Healthcare-focused Adaptive Multilingual Learning Embedding-based Topic Modeling 12 May 2025 · 0 repositories · arXiv:2505.07157
-
Hierarchical Sparse Attention Framework for Computationally Efficient Classification of Biological Cells 12 May 2025 · 0 repositories · arXiv:2505.07661
-
Hybrid Spiking Vision Transformer for Object Detection with Event Cameras 12 May 2025 · 0 repositories · arXiv:2505.07715
-
KAQG: A Knowledge-Graph-Enhanced RAG for Difficulty-Controlled Question Generation 12 May 2025 · 0 repositories · arXiv:2505.07618
-
Lagrange Oscillatory Neural Networks for Constraint Satisfaction and Optimization 12 May 2025 · 1 repository · arXiv:2505.07179
-
LAMM-ViT: AI Face Detection via Layer-Aware Modulation of Region-Guided Attention 12 May 2025 · 0 repositories · arXiv:2505.07734
-
MAIS: Memory-Attention for Interactive Segmentation 12 May 2025 · 0 repositories · arXiv:2505.07511
-
Multi-Plane Vision Transformer for Hemorrhage Classification Using Axial and Sagittal MRI Data 12 May 2025 · 0 repositories · arXiv:2505.07349
-
Multimodal Assessment of Classroom Discourse Quality: A Text-Centered Attention-Based Multi-Task Learning Approach 12 May 2025 · 0 repositories · arXiv:2505.07902
-
No Query, No Access 12 May 2025 · 0 repositories · arXiv:2505.07258
-
Pre-training vs. Fine-tuning: A Reproducibility Study on Dense Retrieval Knowledge Acquisition 12 May 2025 · 1 repository · arXiv:2505.07166
-
RDD: Robust Feature Detector and Descriptor using Deformable Transformer 12 May 2025 · 0 repositories · arXiv:2505.08013
-
ReCDAP: Relation-Based Conditional Diffusion with Attention Pooling for Few-Shot Knowledge Graph Completion 12 May 2025 · 1 repository · arXiv:2505.07171
-
Representation Learning with Mutual Influence of Modalities for Node Classification in Multi-Modal Heterogeneous Networks 12 May 2025 · 1 repository · arXiv:2505.07895
-
SEReDeEP: Hallucination Detection in Retrieval-Augmented Models via Semantic Entropy and Context-Parameter Fusion 12 May 2025 · 0 repositories · arXiv:2505.07528
-
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models 12 May 2025 · 0 repositories · arXiv:2505.07652
-
Sleep Position Classification using Transfer Learning for Bed-based Pressure Sensors 12 May 2025 · 0 repositories · arXiv:2505.08111
-
Statistical CSI-Based Distributed Precoding Design for OFDM-Cooperative Multi-Satellite Systems 12 May 2025 · 0 repositories · arXiv:2505.08038
-
Task-Adaptive Semantic Communications with Controllable Diffusion-based Data Regeneration 12 May 2025 · 0 repositories · arXiv:2505.07980
-
The Geography of Transportation Cybersecurity: Visitor Flows, Industry Clusters, and Spatial Dynamics 12 May 2025 · 0 repositories · arXiv:2505.08822
-
The Influence of the Memory Capacity of Neural DDEs on the Universal Approximation Property 12 May 2025 · 0 repositories · arXiv:2505.07244
-
Topology-Guided Knowledge Distillation for Efficient Point Cloud Processing 12 May 2025 · 1 repository · arXiv:2505.08101
-
Towards Requirements Engineering for RAG Systems 12 May 2025 · 0 repositories · arXiv:2505.07553
-
UMoE: Unifying Attention and FFN with Shared Experts 12 May 2025 · 0 repositories · arXiv:2505.07260
-
Wasserstein Distributionally Robust Nonparametric Regression 12 May 2025 · 0 repositories · arXiv:2505.07967
-
Why Uncertainty Estimation Methods Fall Short in RAG: An Axiomatic Analysis 12 May 2025 · 0 repositories · arXiv:2505.07459
-
2025 TGRS A Self-Supervised Method for Seismic Random Noise Attenuation under Non-Pixelwise Independent Assumption 11 May 2025 · 1 repository
-
A Self-Supervised Method for Attenuating Seismic Random and Tracewise Coherent Noise under the Non-Pixelwise Independence Assumption 11 May 2025 · 1 repository
-
A Self-Supervised Method for Attenuating Seismic Random and Tracewise Coherent Noise under the Non-Pixelwise Independence Assumption 11 May 2025 · 1 repository
-
A systematic review of challenges and proposed solutions in modeling multimodal data 11 May 2025 · 0 repositories · arXiv:2505.06945
-
BridgeIV: Bridging Customized Image and Video Generation through Test-Time Autoregressive Identity Propagation 11 May 2025 · 0 repositories · arXiv:2505.06985
-
Efficient and Robust Multidimensional Attention in Remote Physiological Sensing through Target Signal Constrained Factorization 11 May 2025 · 0 repositories · arXiv:2505.07013
-
Evaluating Reasoning LLMs for Suicide Screening with the Columbia-Suicide Severity Rating Scale 11 May 2025 · 1 repository · arXiv:2505.13480
-
Explainable Artificial Intelligence Techniques for Software Development Lifecycle: A Phase-specific Survey 11 May 2025 · 0 repositories · arXiv:2505.07058
-
IM-BERT: Enhancing Robustness of BERT through the Implicit Euler Method 11 May 2025 · 0 repositories · arXiv:2505.06889
-
Image Classification Using a Diffusion Model as a Pre-Training Model 11 May 2025 · 0 repositories · arXiv:2505.06890
-
Matrix Is All You Need 11 May 2025 · 0 repositories · arXiv:2506.01966
-
NeuRN: Neuro-inspired Domain Generalization for Image Classification 11 May 2025 · 0 repositories · arXiv:2505.06881
-
RefPentester: A Knowledge-Informed Self-Reflective Penetration Testing Framework Based on Large Language Models 11 May 2025 · 0 repositories · arXiv:2505.07089
-
Technical Report for ICRA 2025 GOOSE 2D Semantic Segmentation Challenge: Leveraging Color Shift Correction, RoPE-Swin Backbone, and Quantile-based Label Denoising Strategy for Robust Outdoor Scene Understanding 11 May 2025 · 0 repositories · arXiv:2505.06991
-
The Distracting Effect: Understanding Irrelevant Passages in RAG 11 May 2025 · 0 repositories · arXiv:2505.06914
-
Transformer-Based Dual-Optical Attention Fusion Crowd Head Point Counting and Localization Network 11 May 2025 · 1 repository · arXiv:2505.06937
-
Attention Is Not All You Need: The Importance of Feedforward Networks in Transformer Models 10 May 2025 · 0 repositories · arXiv:2505.06633
-
Attention Mechanisms in Dynamical Systems: A Case Study with Predator-Prey Models 10 May 2025 · 0 repositories · arXiv:2505.06503
-
Boosting Neural Language Inference via Cascaded Interactive Reasoning 10 May 2025 · 0 repositories · arXiv:2505.06607
-
E2E-FANet: A Highly Generalizable Framework for Waves prediction Behind Floating Breakwaters via Exogenous-to-Endogenous Variable Attention 10 May 2025 · 0 repositories · arXiv:2505.06690
-
Gated Attention for Large Language Models: Non-linearity, Sparsity, and Attention-Sink-Free 10 May 2025 · 1 repository · arXiv:2505.06708Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
MacRAG: Compress, Slice, and Scale-up for Multi-Scale Adaptive Context RAG 10 May 2025 · 1 repository · arXiv:2505.06569
-
MultiTaskVIF: Segmentation-oriented visible and infrared image fusion via multi-task learning 10 May 2025 · 0 repositories · arXiv:2505.06665
-
OMGM: Orchestrate Multiple Granularities and Modalities for Efficient Multimodal Retrieval 10 May 2025 · 0 repositories · arXiv:2505.07879Syntology 4 ran (of which 4 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; every one of the 4 samples that ran constructed an object rather than computing a result (of 5 harvested samples) · 5 pointer-only (licence)
-
OptiGait-LGBM: An Efficient Approach of Gait-based Person Re-identification in Non-Overlapping Regions 10 May 2025 · 0 repositories · arXiv:2505.08801
-
Probing In-Context Learning: Impact of Task Complexity and Model Architecture on Generalization and Efficiency 10 May 2025 · 1 repository · arXiv:2505.06475
-
ProFashion: Prototype-guided Fashion Video Generation with Multiple Reference Images 10 May 2025 · 0 repositories · arXiv:2505.06537
-
QoS-Efficient Serving of Multiple Mixture-of-Expert LLMs Using Partial Runtime Reconfiguration 10 May 2025 · 0 repositories · arXiv:2505.06481
-
REFINE-AF: A Task-Agnostic Framework to Align Language Models via Self-Generated Instructions using Reinforcement Learning from Automated Feedback 10 May 2025 · 0 repositories · arXiv:2505.06548
-
RuleGenie: SIEM Detection Rule Set Optimization 10 May 2025 · 0 repositories · arXiv:2505.06701
-
TACFN: Transformer-based Adaptive Cross-modal Fusion Network for Multimodal Emotion Recognition 10 May 2025 · 1 repository · arXiv:2505.06536
-
The Sound of Populism: Distinct Linguistic Features Across Populist Variants 10 May 2025 · 0 repositories · arXiv:2505.07874
-
Underwater object detection in sonar imagery with detection transformer and Zero-shot neural architecture search 10 May 2025 · 0 repositories · arXiv:2505.06694