Methods › General › Attention Mechanisms › Attention › Papers, page 2
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 2 of 316: papers 101 to 200 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
LaDCast: A Latent Diffusion Model for Medium-Range Ensemble Weather Forecasting 10 Jun 2025 · 1 repository · arXiv:2506.09193
-
Learnable Spatial-Temporal Positional Encoding for Link Prediction 10 Jun 2025 · 1 repository · arXiv:2506.08309Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
LiftVSR: Lifting Image Diffusion to Video Super-Resolution via Hybrid Temporal Modeling with Only 4×RTX 4090s 10 Jun 2025 · 0 repositories · arXiv:2506.08529
-
MAC: An Efficient Gradient Preconditioning using Mean Activation Approximated Curvature 10 Jun 2025 · 1 repository · arXiv:2506.08464
-
MedMoE: Modality-Specialized Mixture of Experts for Medical Vision-Language Understanding 10 Jun 2025 · 0 repositories · arXiv:2506.08356
-
Mitigating Posterior Salience Attenuation in Long-Context LLMs with Positional Contrastive Decoding 10 Jun 2025 · 0 repositories · arXiv:2506.08371
-
MLVTG: Mamba-Based Feature Alignment and LLM-Driven Purification for Multi-Modal Video Temporal Grounding 10 Jun 2025 · 0 repositories · arXiv:2506.08512
-
NAM: A Normalization Attention Model for Personalized Product Search In Fliggy 10 Jun 2025 · 0 repositories · arXiv:2506.08382
-
Olica: Efficient Structured Pruning of Large Language Models without Retraining 10 Jun 2025 · 1 repository · arXiv:2506.08436Syntology official (archive's flag): 4 ran · 4 ran (of which 2 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 6 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
PatchGuard: Adversarially Robust Anomaly Detection and Localization through Vision Transformers and Pseudo Anomalies 10 Jun 2025 · 2 repositories · arXiv:2506.09237
-
PlantBert: An Open Source Language Model for Plant Science 10 Jun 2025 · 0 repositories · arXiv:2506.08897
-
Plug-and-Play Linear Attention for Pre-trained Image and Video Restoration Models 10 Jun 2025 · 1 repository · arXiv:2506.08520
-
Robust Visual Localization via Semantic-Guided Multi-Scale Transformer 10 Jun 2025 · 0 repositories · arXiv:2506.08526
-
ScalableHD: Scalable and High-Throughput Hyperdimensional Computing Inference on Multi-Core CPUs 10 Jun 2025 · 0 repositories · arXiv:2506.09282
-
SDMPrune: Self-Distillation MLP Pruning for Efficient Large Language Models 10 Jun 2025 · 1 repository · arXiv:2506.11120
-
SeerAttention-R: Sparse Attention Adaptation for Long Reasoning 10 Jun 2025 · 1 repository · arXiv:2506.08889
-
Self-Anchored Attention Model for Sample-Efficient Classification of Prosocial Text Chat 10 Jun 2025 · 0 repositories · arXiv:2506.09259
-
SEMA: a Scalable and Efficient Mamba like Attention via Token Localization and Averaging 10 Jun 2025 · 0 repositories · arXiv:2506.08297Syntology 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 1 honoured, 2 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 13 harvested samples) · 13 pointer-only (licence)
-
SurfR: Surface Reconstruction with Multi-scale Attention 10 Jun 2025 · 0 repositories · arXiv:2506.08635
-
TACTIC: Translation Agents with Cognitive-Theoretic Interactive Collaboration 10 Jun 2025 · 1 repository · arXiv:2506.08403
-
The Predictive Brain: Neural Correlates of Word Expectancy Align with Large Language Model Prediction Probabilities 10 Jun 2025 · 0 repositories · arXiv:2506.08511
-
Time Series Representations for Classification Lie Hidden in Pretrained Vision Transformers 10 Jun 2025 · 0 repositories · arXiv:2506.08641
-
Transformers Meet Hyperspectral Imaging: A Comprehensive Study of Models, Challenges and Open Problems 10 Jun 2025 · 0 repositories · arXiv:2506.08596
-
4DGT: Learning a 4D Gaussian Transformer Using Real-World Monocular Videos 9 Jun 2025 · 0 repositories · arXiv:2506.08015
-
A Comprehensive Study of Decoder-Only LLMs for Text-to-Image Generation 9 Jun 2025 · 0 repositories · arXiv:2506.08210
-
A Unified Anti-Jamming Design in Complex Environments Based on Cross-Modal Fusion and Intelligent Decision-Making 9 Jun 2025 · 0 repositories · arXiv:2506.07532
-
Adapter Naturally Serves as Decoupler for Cross-Domain Few-Shot Semantic Segmentation 9 Jun 2025 · 0 repositories · arXiv:2506.07376
-
Adaptive Blind Super-Resolution Network for Spatial-Specific and Spatial-Agnostic Degradations 9 Jun 2025 · 0 repositories · arXiv:2506.07705
-
Bingo: Boosting Efficient Reasoning of LLMs via Dynamic and Significance-based Reinforcement Learning 9 Jun 2025 · 0 repositories · arXiv:2506.08125
-
Can Hessian-Based Insights Support Fault Diagnosis in Attention-based Models? 9 Jun 2025 · 0 repositories · arXiv:2506.07871
-
CrosswalkNet: An Optimized Deep Learning Framework for Pedestrian Crosswalk Detection in Aerial Images with High-Performance Computing 9 Jun 2025 · 0 repositories · arXiv:2506.07885
-
CyberV: Cybernetics for Test-time Scaling in Video Understanding 9 Jun 2025 · 1 repository · arXiv:2506.07971
-
DLNet: Direction-Aware Feature Integration for Robust Lane Detection in Complex Environments 9 Jun 2025 · 1 repository
-
Evidential Spectrum-Aware Contrastive Learning for OOD Detection in Dynamic Graphs 9 Jun 2025 · 1 repository · arXiv:2506.07417
-
Generative Voice Bursts during Phone Call 9 Jun 2025 · 0 repositories · arXiv:2506.07526
-
Hierarchical Lexical Graph for Enhanced Multi-Hop Retrieval 9 Jun 2025 · 1 repository · arXiv:2506.08074
-
Lightweight Sequential Transformers for Blood Glucose Level Prediction in Type-1 Diabetes 9 Jun 2025 · 0 repositories · arXiv:2506.07864
-
LlamaRec-LKG-RAG: A Single-Pass, Learnable Knowledge Graph-RAG Framework for LLM-Based Ranking 9 Jun 2025 · 1 repository · arXiv:2506.07449
-
LLM-driven Indoor Scene Layout Generation via Scaled Human-aligned Data Synthesis and Multi-Stage Preference Optimization 9 Jun 2025 · 0 repositories · arXiv:2506.07570
-
LUCIFER: Language Understanding and Context-Infused Framework for Exploration and Behavior Refinement 9 Jun 2025 · 0 repositories · arXiv:2506.07915
-
M2Restore: Mixture-of-Experts-based Mamba-CNN Fusion Framework for All-in-One Image Restoration 9 Jun 2025 · 0 repositories · arXiv:2506.07814
-
MADFormer: Mixed Autoregressive and Diffusion Transformers for Continuous Image Generation 9 Jun 2025 · 0 repositories · arXiv:2506.07999
-
MiniCPM4: Ultra-Efficient LLMs on End Devices 9 Jun 2025 · 1 repository · arXiv:2506.07900Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Mondrian: Transformer Operators via Domain Decomposition 9 Jun 2025 · 0 repositories · arXiv:2506.08226
-
Multilingual Hate Speech Detection in Social Media Using Translation-Based Approaches with Large Language Models 9 Jun 2025 · 0 repositories · arXiv:2506.08147
-
Multiple Object Stitching for Unsupervised Representation Learning 9 Jun 2025 · 1 repository · arXiv:2506.07364
-
Nearness of Neighbors Attention for Regression in Supervised Finetuning 9 Jun 2025 · 1 repository · arXiv:2506.08139
-
OneIG-Bench: Omni-dimensional Nuanced Evaluation for Image Generation 9 Jun 2025 · 1 repository · arXiv:2506.07977
-
Quantum Graph Transformer for NLP Sentiment Classification 9 Jun 2025 · 0 repositories · arXiv:2506.07937
-
Rethinking Cross-Modal Interaction in Multimodal Diffusion Transformers 9 Jun 2025 · 1 repository · arXiv:2506.07986
-
SceneRAG: Scene-level Retrieval-Augmented Generation for Video Understanding 9 Jun 2025 · 0 repositories · arXiv:2506.07600
-
SoK: Data Reconstruction Attacks Against Machine Learning Models: Definition, Metrics, and Benchmark 9 Jun 2025 · 0 repositories · arXiv:2506.07888
-
ST-GraphNet: A Spatio-Temporal Graph Neural Network for Understanding and Predicting Automated Vehicle Crash Severity 9 Jun 2025 · 0 repositories · arXiv:2506.08051
-
STAMImputer: Spatio-Temporal Attention MoE for Traffic Data Imputation 9 Jun 2025 · 1 repository · arXiv:2506.08054
-
Vision Transformers Don't Need Trained Registers 9 Jun 2025 · 1 repository · arXiv:2506.08010Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
When Style Breaks Safety: Defending Language Models Against Superficial Style Alignment 9 Jun 2025 · 1 repository · arXiv:2506.07452Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
A dependently-typed calculus of event telicity and culminativity 8 Jun 2025 · 0 repositories · arXiv:2506.06968
-
Accelerating 3D Gaussian Splatting with Neural Sorting and Axis-Oriented Rasterization 8 Jun 2025 · 0 repositories · arXiv:2506.07069
-
Backdoor Attack on Vision Language Models with Stealthy Semantic Manipulation 8 Jun 2025 · 0 repositories · arXiv:2506.07214
-
Circuit-Based Modeling Approach for Channel Estimation in RIS-Assisted Communications 8 Jun 2025 · 0 repositories · arXiv:2506.07124
-
Joint Channel and Symbol Estimation for Communication Systems with Movable Antennas 8 Jun 2025 · 0 repositories · arXiv:2506.07183
-
MAGNet: A Multi-Scale Attention-Guided Graph Fusion Network for DRC Violation Detection 8 Jun 2025 · 0 repositories · arXiv:2506.07126
-
Quality-Diversity Red-Teaming: Automated Generation of High-Quality and Diverse Attackers for Large Language Models 8 Jun 2025 · 0 repositories · arXiv:2506.07121
-
RBA-FE: A Robust Brain-Inspired Audio Feature Extractor for Depression Diagnosis 8 Jun 2025 · 0 repositories · arXiv:2506.07118
-
Reward Model Interpretability via Optimal and Pessimal Tokens 8 Jun 2025 · 0 repositories · arXiv:2506.07326
-
SiliCoN: Simultaneous Nuclei Segmentation and Color Normalization of Histological Images 8 Jun 2025 · 0 repositories · arXiv:2506.07028
-
Breaking Data Silos: Towards Open and Scalable Mobility Foundation Models via Generative Continual Learning 7 Jun 2025 · 0 repositories · arXiv:2506.06694
-
Can In-Context Reinforcement Learning Recover From Reward Poisoning Attacks? 7 Jun 2025 · 0 repositories · arXiv:2506.06891
-
Deep Inertial Pose: A deep learning approach for human pose estimation 7 Jun 2025 · 0 repositories · arXiv:2506.06850
-
Exploring Length Generalization For Transformer-based Speech Enhancement 7 Jun 2025 · 0 repositories · arXiv:2506.06697
-
Graph Neural Networks in Modern AI-aided Drug Discovery 7 Jun 2025 · 0 repositories · arXiv:2506.06915
-
RoboCerebra: A Large-scale Benchmark for Long-horizon Robotic Manipulation Evaluation 7 Jun 2025 · 0 repositories · arXiv:2506.06677
-
Training-Free Identity Preservation in Stylized Image Generation Using Diffusion Models 7 Jun 2025 · 0 repositories · arXiv:2506.06802
-
BEAST: Efficient Tokenization of B-Splines Encoded Action Sequences for Imitation Learning 6 Jun 2025 · 0 repositories · arXiv:2506.06072
-
Direct Behavior Optimization: Unlocking the Potential of Lightweight LLMs 6 Jun 2025 · 0 repositories · arXiv:2506.06401
-
Domain Adaptation in Agricultural Image Analysis: A Comprehensive Review from Shallow Models to Deep Learning 6 Jun 2025 · 0 repositories · arXiv:2506.05972
-
FPDANet: A Multi-Section Classification Model for Intelligent Screening of Fetal Ultrasound 6 Jun 2025 · 0 repositories · arXiv:2506.06054
-
Joint-GCG: Unified Gradient-Based Poisoning Attacks on Retrieval-Augmented Generation Systems 6 Jun 2025 · 1 repository · arXiv:2506.06151
-
On-board Mission Replanning for Adaptive Cooperative Multi-Robot Systems 6 Jun 2025 · 0 repositories · arXiv:2506.06094
-
RecGPT: A Foundation Model for Sequential Recommendation 6 Jun 2025 · 1 repository · arXiv:2506.06270Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
TADA: Training-free Attribution and Out-of-Domain Detection of Audio Deepfakes 6 Jun 2025 · 1 repository · arXiv:2506.05802
-
Textile Analysis for Recycling Automation using Transfer Learning and Zero-Shot Foundation Models 6 Jun 2025 · 0 repositories · arXiv:2506.06569
-
The Lock-in Hypothesis: Stagnation by Algorithm 6 Jun 2025 · 0 repositories · arXiv:2506.06166
-
The Optimization Paradox in Clinical AI Multi-Agent Systems 6 Jun 2025 · 1 repository · arXiv:2506.06574
-
When Better Features Mean Greater Risks: The Performance-Privacy Trade-Off in Contrastive Learning 6 Jun 2025 · 0 repositories · arXiv:2506.05743
-
When to use Graphs in RAG: A Comprehensive Analysis for Graph Retrieval-Augmented Generation 6 Jun 2025 · 1 repository · arXiv:2506.05690Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
DACN: Dual-Attention Convolutional Network for Hyperspectral Image Super-Resolution 5 Jun 2025 · 1 repository · arXiv:2506.05041
-
Nonlinear Causal Discovery for Grouped Data 5 Jun 2025 · 0 repositories · arXiv:2506.05120
-
On the Convergence of Gradient Descent on Learning Transformers with Residual Connections 5 Jun 2025 · 0 repositories · arXiv:2506.05249
-
A Multi-Dataset Evaluation of Models for Automated Vulnerability Repair 5 Jun 2025 · 0 repositories · arXiv:2506.04987
-
A Neural Network Model of Spatial and Feature-Based Attention 5 Jun 2025 · 0 repositories · arXiv:2506.05487
-
AI-powered Contextual 3D Environment Generation: A Systematic Review 5 Jun 2025 · 0 repositories · arXiv:2506.05449
-
Associative Memory and Generative Diffusion in the Zero-noise Limit 5 Jun 2025 · 0 repositories · arXiv:2506.05178
-
Benchmarking Large Language Models on Homework Assessment in Circuit Analysis 5 Jun 2025 · 0 repositories · arXiv:2506.06390
-
Debatable Intelligence: Benchmarking LLM Judges via Debate Speech Evaluation 5 Jun 2025 · 1 repository · arXiv:2506.05062Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Demonstrations of Integrity Attacks in Multi-Agent Systems 5 Jun 2025 · 0 repositories · arXiv:2506.04572
-
Design of intelligent proofreading system for English translation based on CNN and BERT 5 Jun 2025 · 0 repositories · arXiv:2506.04811
-
Dissecting Bias in LLMs: A Mechanistic Interpretability Perspective 5 Jun 2025 · 0 repositories · arXiv:2506.05166
-
Distributionally Robust Auction Design with Deferred Inspection 5 Jun 2025 · 0 repositories · arXiv:2506.04767
-
Dynamic Context Tuning for Retrieval-Augmented Generation: Enhancing Multi-Turn Planning and Tool Adaptation 5 Jun 2025 · 0 repositories · arXiv:2506.11092