Methods › General › Attention Mechanisms › Attention › Papers, page 37
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 37 of 316: papers 3,601 to 3,700 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Using Machine Learning for move sequence visualization and generation in climbing 1 Mar 2025 · 0 repositories · arXiv:2503.00458
-
FANformer: Improving Large Language Models Through Effective Periodicity Modeling 28 Feb 2025 · 1 repository · arXiv:2502.21309Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
SafeAuto: Knowledge-Enhanced Safe Autonomous Driving with Multimodal Foundation Models 28 Feb 2025 · 1 repository · arXiv:2503.00211
-
A Compact Model for Large-Scale Time Series Forecasting 28 Feb 2025 · 0 repositories · arXiv:2502.20634
-
A novel Fourier Adjacency Transformer for advanced EEG emotion recognition 28 Feb 2025 · 1 repository · arXiv:2503.13465Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 1 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 14 harvested samples)
-
AMPLE: Event-Driven Accelerator for Mixed-Precision Inference of Graph Neural Networks 28 Feb 2025 · 0 repositories · arXiv:2502.21196
-
Beyond Words: A Latent Memory Approach to Internal Reasoning in LLMs 28 Feb 2025 · 0 repositories · arXiv:2502.21030
-
BST: Badminton Stroke-type Transformer for Skeleton-based Action Recognition in Racket Sports 28 Feb 2025 · 1 repository · arXiv:2502.21085
-
CODI: Compressing Chain-of-Thought into Continuous Space via Self-Distillation 28 Feb 2025 · 1 repository · arXiv:2502.21074Syntology 0 ran · 2 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Continual Learning-Aided Super-Resolution Scheme for Channel Reconstruction and Generalization in OFDM Systems 28 Feb 2025 · 0 repositories · arXiv:2503.01897
-
DiffBrush:Just Painting the Art by Your Hands 28 Feb 2025 · 0 repositories · arXiv:2502.20904
-
Dynamic Markov Blanket Detection for Macroscopic Physics Discovery 28 Feb 2025 · 1 repository · arXiv:2502.21217
-
EDENet: Echo Direction Encoding Network for Place Recognition Based on Ground Penetrating Radar 28 Feb 2025 · 1 repository · arXiv:2502.20643
-
Efficient Jailbreaking of Large Models by Freeze Training: Lower Layers Exhibit Greater Sensitivity to Harmful Content 28 Feb 2025 · 0 repositories · arXiv:2502.20952
-
Efficient Transformer-based Decoder for Varshamov-Tenengolts Codes 28 Feb 2025 · 0 repositories · arXiv:2502.21060
-
Ext2Gen: Alignment through Unified Extraction and Generation for Robust Retrieval-Augmented Generation 28 Feb 2025 · 0 repositories · arXiv:2503.04789
-
FASTer: Focal Token Acquiring-and-Scaling Transformer for Long-term 3D Object Detection 28 Feb 2025 · 1 repository · arXiv:2503.01899
-
FlexPrefill: A Context-Aware Sparse Attention Mechanism for Efficient Long-Sequence Inference 28 Feb 2025 · 1 repository · arXiv:2502.20766Syntology official (archive's flag): 9 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 1 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 11 harvested samples)
-
How Metacognitive Architectures Remember Their Own Thoughts: A Systematic Review 28 Feb 2025 · 0 repositories · arXiv:2503.13467
-
Indoor Localization for Autonomous Robot Navigation 28 Feb 2025 · 0 repositories · arXiv:2502.20731
-
Integrating convolutional layers and biformer network with forward-forward and backpropagation training 28 Feb 2025 · 1 repository
-
JiTTER: Jigsaw Temporal Transformer for Event Reconstruction for Self-Supervised Sound Event Detection 28 Feb 2025 · 1 repository · arXiv:2502.20857
-
LexRAG: Benchmarking Retrieval-Augmented Generation in Multi-Turn Legal Consultation Conversation 28 Feb 2025 · 1 repository · arXiv:2502.20640
-
Movable Antenna Aided Multiuser Communications: Antenna Position Optimization Based on Statistical Channel Information 28 Feb 2025 · 0 repositories · arXiv:2502.20856
-
Multimodal Learning for Just-In-Time Software Defect Prediction in Autonomous Driving Systems 28 Feb 2025 · 1 repository · arXiv:2502.20806
-
"No negatives needed": weakly-supervised regression for interpretable tumor detection in whole-slide histopathology images 28 Feb 2025 · 1 repository · arXiv:2502.21109
-
NutriGen: Personalized Meal Plan Generator Leveraging Large Language Models to Enhance Dietary and Nutritional Adherence 28 Feb 2025 · 1 repository · arXiv:2502.20601
-
Retrieval Augmented Generation for Topic Modeling in Organizational Research: An Introduction with Empirical Demonstration 28 Feb 2025 · 0 repositories · arXiv:2502.20963
-
Retrieval Backward Attention without Additional Training: Enhance Embeddings of Large Language Models via Repetition 28 Feb 2025 · 1 repository · arXiv:2502.20726
-
RuCCoD: Towards Automated ICD Coding in Russian 28 Feb 2025 · 1 repository · arXiv:2502.21263
-
Solar Multimodal Transformer: Intraday Solar Irradiance Predictor using Public Cameras and Time Series 28 Feb 2025 · 0 repositories · arXiv:2503.00250
-
SPD: Sync-Point Drop for Efficient Tensor Parallelism of Large Language Models 28 Feb 2025 · 0 repositories · arXiv:2502.20727
-
Spiking Transformer:Introducing Accurate Addition-Only Spiking Self-Attention for Transformer 28 Feb 2025 · 0 repositories · arXiv:2503.00226
-
SuperRAG: Beyond RAG with Layout-Aware Graph Modeling 28 Feb 2025 · 0 repositories · arXiv:2503.04790
-
TeleRAG: Efficient Retrieval-Augmented Generation Inference with Lookahead Retrieval 28 Feb 2025 · 0 repositories · arXiv:2502.20969
-
The Luce Model, Regularity, and Choice Overload 28 Feb 2025 · 0 repositories · arXiv:2502.21063
-
The RAG Paradox: A Black-Box Attack Exploiting Unintentional Vulnerabilities in Retrieval-Augmented Generation Systems 28 Feb 2025 · 0 repositories · arXiv:2502.20995
-
Transformers with Joint Tokens and Local-Global Attention for Efficient Human Pose Estimation 28 Feb 2025 · 0 repositories · arXiv:2503.00232
-
A Novel P-bit-based Probabilistic Computing Approach for Solving the 3-D Protein Folding Problem 27 Feb 2025 · 0 repositories · arXiv:2502.20050
-
Advanced Deep Learning Techniques for Analyzing Earnings Call Transcripts: Methodologies and Applications 27 Feb 2025 · 0 repositories · arXiv:2503.01886
-
An exploration of features to improve the generalisability of fake news detection models 27 Feb 2025 · 0 repositories · arXiv:2502.20299
-
An Integrated Deep Learning Framework Leveraging NASNet and Vision Transformer with MixProcessing for Accurate and Precise Diagnosis of Lung Diseases 27 Feb 2025 · 0 repositories · arXiv:2502.20570
-
Attention Distillation: A Unified Approach to Visual Characteristics Transfer 27 Feb 2025 · 1 repository · arXiv:2502.20235Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Beyond the Tip of Efficiency: Uncovering the Submerged Threats of Jailbreak Attacks in Small Language Models 27 Feb 2025 · 0 repositories · arXiv:2502.19883
-
Bridging Legal Knowledge and AI: Retrieval-Augmented Generation with Vector Stores, Knowledge Graphs, and Hierarchical Non-negative Matrix Factorization 27 Feb 2025 · 1 repository · arXiv:2502.20364
-
CirT: Global Subseasonal-to-Seasonal Forecasting with Geometry-inspired Transformer 27 Feb 2025 · 1 repository · arXiv:2502.19750Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
CNsum:Automatic Summarization for Chinese News Text 27 Feb 2025 · 0 repositories · arXiv:2502.19723
-
Conformal Tail Risk Control for Large Language Model Alignment 27 Feb 2025 · 0 repositories · arXiv:2502.20285
-
DIPSER: A Dataset for In-Person Student1 Engagement Recognition in the Wild 27 Feb 2025 · 0 repositories · arXiv:2502.20209
-
Do computer vision foundation models learn the low-level characteristics of the human visual system? 27 Feb 2025 · 0 repositories · arXiv:2502.20256
-
Educator Attention: How computational tools can systematically identify the distribution of a key resource for students 27 Feb 2025 · 0 repositories · arXiv:2502.20135
-
Efficient and Universal Neural-Network Decoder for Stabilizer-Based Quantum Error Correction 27 Feb 2025 · 0 repositories · arXiv:2502.19971
-
Enhancing 3D Gaze Estimation in the Wild using Weak Supervision with Gaze Following Labels 27 Feb 2025 · 0 repositories · arXiv:2502.20249
-
FedMentalCare: Towards Privacy-Preserving Fine-Tuned LLMs to Analyze Mental Health Status Using Federated Learning Framework 27 Feb 2025 · 0 repositories · arXiv:2503.05786
-
Implicit Search via Discrete Diffusion: A Study on Chess 27 Feb 2025 · 1 repository · arXiv:2502.19805Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples) · 1 pointer-only (licence)
-
Improving Adversarial Transferability in MLLMs via Dynamic Vision-Language Alignment Attack 27 Feb 2025 · 0 repositories · arXiv:2502.19672
-
Investigating and Enhancing Vision-Audio Capability in Omnimodal Large Language Models 27 Feb 2025 · 0 repositories · arXiv:2503.00059
-
Investigating Neurons and Heads in Transformer-based LLMs for Typographical Errors 27 Feb 2025 · 0 repositories · arXiv:2502.19669
-
Language-Informed Hyperspectral Image Synthesis for Imbalanced-Small Sample Classification via Semi-Supervised Conditional Diffusion Model 27 Feb 2025 · 0 repositories · arXiv:2502.19700
-
Lightweight Contrastive Distilled Hashing for Online Cross-modal Retrieval 27 Feb 2025 · 0 repositories · arXiv:2502.19751
-
LLM-driven Effective Knowledge Tracing by Integrating Dual-channel Difficulty 27 Feb 2025 · 0 repositories · arXiv:2502.19915
-
Long-Context Inference with Retrieval-Augmented Speculative Decoding 27 Feb 2025 · 1 repository · arXiv:2502.20330
-
Low-rank tensor completion via a novel minimax p-th order concave penalty function 27 Feb 2025 · 0 repositories · arXiv:2502.19979
-
M^3Builder: A Multi-Agent System for Automated Machine Learning in Medical Imaging 27 Feb 2025 · 0 repositories · arXiv:2502.20301
-
Med-RLVR: Emerging Medical Reasoning from a 3B base model via reinforcement Learning 27 Feb 2025 · 0 repositories · arXiv:2502.19655
-
Minds on the Move: Decoding Trajectory Prediction in Autonomous Driving with Cognitive Insights 27 Feb 2025 · 0 repositories · arXiv:2502.20084
-
MITracker: Multi-View Integration for Visual Object Tracking 27 Feb 2025 · 0 repositories · arXiv:2502.20111
-
Mixmamba-fewshot: mamba and attention mixer-based method with few-shot learning for bearing fault diagnosis 27 Feb 2025 · 1 repository
-
Multimodal Representation Alignment for Image Generation: Text-Image Interleaved Control Is Easier Than You Think 27 Feb 2025 · 1 repository · arXiv:2502.20172
-
On the Importance of Reward Design in Reinforcement Learning-based Dynamic Algorithm Configuration: A Case Study on OneMax with (1+(λ,λ))-GA 27 Feb 2025 · 1 repository · arXiv:2502.20265
-
OverLoCK: An Overview-first-Look-Closely-next ConvNet with Context-Mixing Dynamic Kernels 27 Feb 2025 · 1 repository · arXiv:2502.20087Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
PrimeK-Net: Multi-scale Spectral Learning via Group Prime-Kernel Convolutional Neural Networks for Single Channel Speech Enhancement 27 Feb 2025 · 1 repository · arXiv:2502.19906
-
Promote, Suppress, Iterate: How Language Models Answer One-to-Many Factual Queries 27 Feb 2025 · 1 repository · arXiv:2502.20475
-
QORT-Former: Query-optimized Real-time Transformer for Understanding Two Hands Manipulating Objects 27 Feb 2025 · 0 repositories · arXiv:2502.19769
-
Regional climate projections using a deep-learning-based model-ranking and downscaling framework: Application to European climate zones 27 Feb 2025 · 0 repositories · arXiv:2502.20132
-
Revisit the Stability of Vanilla Federated Learning Under Diverse Conditions 27 Feb 2025 · 0 repositories · arXiv:2502.19849
-
RURANET++: An Unsupervised Learning Method for Diabetic Macular Edema Based on SCSE Attention Mechanisms and Dynamic Multi-Projection Head Clustering 27 Feb 2025 · 0 repositories · arXiv:2502.20224
-
SAP-DIFF: Semantic Adversarial Patch Generation for Black-Box Face Recognition Models via Diffusion Models 27 Feb 2025 · 0 repositories · arXiv:2502.19710
-
Scalable Graph Attention-based Instance Selection via Mini-Batch Sampling and Hierarchical Hashing 27 Feb 2025 · 0 repositories · arXiv:2502.20293
-
SecureGaze: Defending Gaze Estimation Against Backdoor Attacks 27 Feb 2025 · 1 repository · arXiv:2502.20306
-
SeisMoLLM: Advancing Seismic Monitoring via Cross-modal Transfer with Pre-trained Large Language Model 27 Feb 2025 · 1 repository · arXiv:2502.19960Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 6 harvested samples)
-
Semiparametric Triple Difference Estimators 27 Feb 2025 · 0 repositories · arXiv:2502.19788
-
Space Rotation with Basis Transformation for Training-free Test-Time Adaptation 27 Feb 2025 · 0 repositories · arXiv:2502.19946
-
Thinking Slow, Fast: Scaling Inference Compute with Distilled Reasoners 27 Feb 2025 · 0 repositories · arXiv:2502.20339
-
Transient Stability Analysis and Fault Clearing Angle Estimation of VSG Based on Domain of Attraction Estimated by Trajectory Reversing Method 27 Feb 2025 · 0 repositories · arXiv:2502.19728
-
UIFace: Unleashing Inherent Model Capabilities to Enhance Intra-Class Diversity in Synthetic Face Recognition 27 Feb 2025 · 1 repository · arXiv:2502.19803
-
Unlocking Multi-Modal Potentials for Dynamic Text-Attributed Graph Representation 27 Feb 2025 · 0 repositories · arXiv:2502.19651
-
WalnutData: A UAV Remote Sensing Dataset of Green Walnuts and Model Evaluation 27 Feb 2025 · 1 repository · arXiv:2502.20092
-
3D Nephrographic Image Synthesis in CT Urography with the Diffusion Model and Swin Transformer 26 Feb 2025 · 0 repositories · arXiv:2502.19623
-
A Sliding Layer Merging Method for Efficient Depth-Wise Pruning in LLMs 26 Feb 2025 · 1 repository · arXiv:2502.19159
-
A Survey on Foundation-Model-Based Industrial Defect Detection 26 Feb 2025 · 0 repositories · arXiv:2502.19106
-
AKDT: Adaptive Kernel Dilation Transformer for Effective Image Denoising 26 Feb 2025 · 1 repository
-
Attention-Guided Integration of CLIP and SAM for Precise Object Masking in Robotic Manipulation 26 Feb 2025 · 0 repositories · arXiv:2502.18842
-
Brain-inspired analogical mixture prototypes for few-shot class-incremental learning 26 Feb 2025 · 0 repositories · arXiv:2502.18923
-
Clip-TTS: Contrastive Text-content and Mel-spectrogram, A High-Quality Text-to-Speech Method based on Contextual Semantic Understanding 26 Feb 2025 · 0 repositories · arXiv:2502.18889
-
Cognitive networks highlight differences and similarities in the STEM mindsets of human and LLM-simulated trainees, experts and academics 26 Feb 2025 · 0 repositories · arXiv:2502.19529
-
CS-Dialogue: A 104-Hour Dataset of Spontaneous Mandarin-English Code-Switching Dialogues for Speech Recognition 26 Feb 2025 · 0 repositories · arXiv:2502.18913
-
Deep-Bench: Deep Learning Benchmark Dataset for Code Generation 26 Feb 2025 · 0 repositories · arXiv:2502.18726
-
DualSpec: Text-to-spatial-audio Generation via Dual-Spectrogram Guided Diffusion Model 26 Feb 2025 · 0 repositories · arXiv:2502.18952
-
Efficient Federated Search for Retrieval-Augmented Generation 26 Feb 2025 · 0 repositories · arXiv:2502.19280