Methods › General › Attention Mechanisms › Attention › Papers, page 23
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 23 of 316: papers 2,201 to 2,300 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Multihead self-attention in cortico-thalamic circuits 8 Apr 2025 · 0 repositories · arXiv:2504.06354
-
NNN: Next-Generation Neural Networks for Marketing Measurement 8 Apr 2025 · 0 repositories · arXiv:2504.06212
-
OmniSVG: A Unified Scalable Vector Graphics Generation Model 8 Apr 2025 · 0 repositories · arXiv:2504.06263
-
Panoptic: True Joint mmWave Communication and Sensing with Compressive Sidelobe Forming 8 Apr 2025 · 0 repositories · arXiv:2504.06400
-
Parasite: A Steganography-based Backdoor Attack Framework for Diffusion Models 8 Apr 2025 · 0 repositories · arXiv:2504.05815
-
PathGPT: Leveraging Large Language Models for Personalized Route Generation 8 Apr 2025 · 0 repositories · arXiv:2504.05846
-
Physics-informed KAN PointNet: Deep learning for simultaneous solutions to inverse problems in incompressible flow on numerous irregular geometries 8 Apr 2025 · 1 repository · arXiv:2504.06327
-
PRIMEDrive-CoT: A Precognitive Chain-of-Thought Framework for Uncertainty-Aware Object Interaction in Driving Scene Scenario 8 Apr 2025 · 0 repositories · arXiv:2504.05908
-
Rethinking the Nested U-Net Approach: Enhancing Biomarker Segmentation with Attention Mechanisms and Multiscale Feature Fusion 8 Apr 2025 · 1 repository · arXiv:2504.06158
-
Retrieval Augmented Generation with Collaborative Filtering for Personalized Text Generation 8 Apr 2025 · 1 repository · arXiv:2504.05731
-
Separator Injection Attack: Uncovering Dialogue Biases in Large Language Models Caused by Role Separators 8 Apr 2025 · 0 repositories · arXiv:2504.05689
-
ShadowCoT: Cognitive Hijacking for Stealthy Reasoning Backdoors in LLMs 8 Apr 2025 · 0 repositories · arXiv:2504.05605
-
Storybooth: Training-free Multi-Subject Consistency for Improved Visual Storytelling 8 Apr 2025 · 0 repositories · arXiv:2504.05800
-
The Hall of AI Fears and Hopes: Comparing the Views of AI Influencers and those of Members of the U.S. Public Through an Interactive Platform 8 Apr 2025 · 0 repositories · arXiv:2504.06016
-
A moving target in AI-assisted decision-making: Dataset shift, model updating, and the problem of update opacity 7 Apr 2025 · 0 repositories · arXiv:2504.05210
-
AI for Climate Finance: Agentic Retrieval and Multi-Step Reasoning for Early Warning System Investments 7 Apr 2025 · 0 repositories · arXiv:2504.05104
-
AsyReC: A Multimodal Graph-based Framework for Spatio-Temporal Asymmetric Dyadic Relationship Classification 7 Apr 2025 · 1 repository · arXiv:2504.05030
-
Attention-Augmented Inverse Reinforcement Learning with Graph Convolutions for Multi-Agent Task Allocation 7 Apr 2025 · 0 repositories · arXiv:2504.05045
-
Attention-Based Multiscale Temporal Fusion Network for Uncertain-Mode Fault Diagnosis in Multimode Processes 7 Apr 2025 · 1 repository · arXiv:2504.05172
-
Bidirectional Hierarchical Protein Multi-Modal Representation Learning 7 Apr 2025 · 0 repositories · arXiv:2504.04770
-
Boundary representation learning via Transformer 7 Apr 2025 · 0 repositories · arXiv:2504.07134
-
CCSK:Cognitive Convection of Self-Knowledge Based Retrieval Augmentation for Large Language Models 7 Apr 2025 · 0 repositories · arXiv:2504.10498
-
Collab-RAG: Boosting Retrieval-Augmented Generation for Complex Question Answering via White-Box and Black-Box LLM Collaboration 7 Apr 2025 · 1 repository · arXiv:2504.04915Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Content-Aware Transformer for All-in-one Image Restoration 7 Apr 2025 · 1 repository · arXiv:2504.04869
-
DFormerv2: Geometry Self-Attention for RGBD Semantic Segmentation 7 Apr 2025 · 1 repository · arXiv:2504.04701Syntology official: harvested, nothing ran · 0 ran · 3 unverified (of 3 harvested samples)
-
Dynamic Vision Mamba 7 Apr 2025 · 1 repository · arXiv:2504.04787
-
FORCE: Feature-Oriented Representation with Clustering and Explanation 7 Apr 2025 · 0 repositories · arXiv:2504.05530
-
GAMDTP: Dynamic Trajectory Prediction with Graph Attention Mamba Network 7 Apr 2025 · 0 repositories · arXiv:2504.04862
-
GraphPINE: Graph Importance Propagation for Interpretable Drug Response Prediction 7 Apr 2025 · 0 repositories · arXiv:2504.05454
-
InstructionBench: An Instructional Video Understanding Benchmark 7 Apr 2025 · 0 repositories · arXiv:2504.05040
-
Interval-Valued Time Series Classification Using D_K-Distance 7 Apr 2025 · 0 repositories · arXiv:2504.04667
-
LagKV: Lag-Relative Information of the KV Cache Tells Which Tokens Are Important 7 Apr 2025 · 1 repository · arXiv:2504.04704
-
LDGNet: A Lightweight Difference Guiding Network for Remote Sensing Change Detection 7 Apr 2025 · 0 repositories · arXiv:2504.05062
-
Learning Affine Correspondences by Integrating Geometric Constraints 7 Apr 2025 · 1 repository · arXiv:2504.04834Syntology official: no sample here; runs from other or unrecorded repositories · 8 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 2 honoured, 1 violated, 2 with no contract checked; 3 where Syntology's instrument failed) · 5 unverified (of 13 harvested samples) · 13 pointer-only (licence)
-
Leveraging LLMs for Utility-Focused Annotation: Reducing Manual Effort for Retrieval and RAG 7 Apr 2025 · 0 repositories · arXiv:2504.05220
-
Leveraging State Space Models in Long Range Genomics 7 Apr 2025 · 0 repositories · arXiv:2504.06304
-
Lumina-OmniLV: A Unified Multimodal Framework for General Low-Level Vision 7 Apr 2025 · 0 repositories · arXiv:2504.04903
-
MSA-UNet3+: Multi-Scale Attention UNet3+ with New Supervised Prototypical Contrastive Loss for Coronary DSA Image Segmentation 7 Apr 2025 · 0 repositories · arXiv:2504.05184
-
OmniEcon Nexus: Global Microeconomic Simulation Engine 7 Apr 2025 · 1 repository
-
One-Minute Video Generation with Test-Time Training 7 Apr 2025 · 0 repositories · arXiv:2504.05298
-
One Quantizer is Enough: Toward a Lightweight Audio Codec 7 Apr 2025 · 1 repository · arXiv:2504.04949
-
Provable Failure of Language Models in Learning Majority Boolean Logic via Gradient Descent 7 Apr 2025 · 0 repositories · arXiv:2504.04702
-
REWIND: Real-Time Egocentric Whole-Body Motion Diffusion with Exemplar-Based Identity Conditioning 7 Apr 2025 · 0 repositories · arXiv:2504.04956
-
SAD-Net: a full spectral self-attention detail enhancement network for single image dehazing 7 Apr 2025 · 1 repository
-
SEAL: Steerable Reasoning Calibration of Large Language Models for Free 7 Apr 2025 · 1 repository · arXiv:2504.07986
-
TC-MGC: Text-Conditioned Multi-Grained Contrastive Learning for Text-Video Retrieval 7 Apr 2025 · 1 repository · arXiv:2504.04707Syntology official: harvested, nothing ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 1 pointer-only (licence)
-
Topic mining based on fine-tuning Sentence-BERT and LDA 7 Apr 2025 · 0 repositories · arXiv:2504.07984
-
Unleashing the Power of LLMs in Dense Retrieval with Query Likelihood Modeling 7 Apr 2025 · 0 repositories · arXiv:2504.05216
-
Weak-for-Strong: Training Weak Meta-Agent to Harness Strong Executors 7 Apr 2025 · 1 repository · arXiv:2504.04785
-
Multimodal Cinematic Video Synthesis Using Text-to-Image and Audio Generation Models 6 Apr 2025 · 0 repositories · arXiv:2506.10005
-
Attention-Driven LPLC2 Neural Ensemble Model for Multi-Target Looming Detection and Localization 6 Apr 2025 · 0 repositories · arXiv:2504.04477
-
Capturing AI's Attention: Physics of Repetition, Hallucination, Bias and Beyond 6 Apr 2025 · 0 repositories · arXiv:2504.04600
-
CO-Bench: Benchmarking Language Model Agents in Algorithm Search for Combinatorial Optimization 6 Apr 2025 · 1 repository · arXiv:2504.04310Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
DanceMosaic: High-Fidelity Dance Generation with Multimodal Editability 6 Apr 2025 · 0 repositories · arXiv:2504.04634
-
Driving-RAG: Driving Scenarios Embedding, Search, and RAG Applications 6 Apr 2025 · 0 repositories · arXiv:2504.04419
-
DynClean: Training Dynamics-based Label Cleaning for Distantly-Supervised Named Entity Recognition 6 Apr 2025 · 0 repositories · arXiv:2504.04616
-
Exploring Generative AI Techniques in Government: A Case Study 6 Apr 2025 · 0 repositories · arXiv:2504.10497
-
Gating is Weighting: Understanding Gated Linear Attention through In-context Learning 6 Apr 2025 · 0 repositories · arXiv:2504.04308
-
Hierarchical Planning for Complex Tasks with Knowledge Graph-RAG and Symbolic Verification 6 Apr 2025 · 0 repositories · arXiv:2504.04578
-
Modeling the Dynamics of Attentional Gamma Oscillations During the Encoding Process of Noise-Mixed Speech Signals 6 Apr 2025 · 0 repositories · arXiv:2504.04329
-
Saliency-driven Dynamic Token Pruning for Large Language Models 6 Apr 2025 · 0 repositories · arXiv:2504.04514
-
SolRPDS: A Dataset for Analyzing Rug Pulls in Solana Decentralized Finance 6 Apr 2025 · 1 repository · arXiv:2504.07132
-
Beyond the Hype: Embeddings vs. Prompting for Multiclass Classification Tasks 5 Apr 2025 · 0 repositories · arXiv:2504.04277
-
CATS: Mitigating Correlation Shift for Multivariate Time Series Classification 5 Apr 2025 · 0 repositories · arXiv:2504.04283
-
Computational Efficient Informative Nonignorable Matrix Completion: A Row- and Column-Wise Matrix U-Statistic Pseudo-Likelihood Approach 5 Apr 2025 · 0 repositories · arXiv:2504.04016
-
Deep-Learning-Directed Preventive Dynamic Security Control via Coordinated Demand Response 5 Apr 2025 · 0 repositories · arXiv:2504.04059
-
EMF: Event Meta Formers for Event-based Real-time Traffic Object Detection 5 Apr 2025 · 0 repositories · arXiv:2504.04124
-
Impact of Price Inflation on Algorithmic Collusion Through Reinforcement Learning Agents 5 Apr 2025 · 0 repositories · arXiv:2504.05335
-
Performance Analysis of Deep Learning Models for Femur Segmentation in MRI Scan 5 Apr 2025 · 0 repositories · arXiv:2504.04066
-
Psychological Health Knowledge-Enhanced LLM-based Social Network Crisis Intervention Text Transfer Recognition Method 5 Apr 2025 · 0 repositories · arXiv:2504.07983
-
QE-RAG: A Robust Retrieval-Augmented Generation Benchmark for Query Entry Errors 5 Apr 2025 · 0 repositories · arXiv:2504.04062
-
Quantum Adaptive Self-Attention for Quantum Transformer Models 5 Apr 2025 · 0 repositories · arXiv:2504.05336
-
Resilience of Vision Transformers for Domain Generalisation in the Presence of Out-of-Distribution Noisy Images 5 Apr 2025 · 0 repositories · arXiv:2504.04225
-
Sigma: A dataset for text-to-code semantic parsing with statistical analysis 5 Apr 2025 · 1 repository · arXiv:2504.04301
-
TARAC: Mitigating Hallucination in LVLMs via Temporal Attention Real-time Accumulative Connection 5 Apr 2025 · 0 repositories · arXiv:2504.04099
-
Towards Principled Learning for Re-ranking in Recommender Systems 5 Apr 2025 · 1 repository · arXiv:2504.04188
-
Transformer representation learning is necessary for dynamic multi-modal physiological data on small-cohort patients 5 Apr 2025 · 0 repositories · arXiv:2504.04120
-
Video4DGen: Enhancing Video and 4D Generation through Mutual Optimization 5 Apr 2025 · 1 repository · arXiv:2504.04153
-
Adaptive Classification of Interval-Valued Time Series 4 Apr 2025 · 0 repositories · arXiv:2504.03318
-
AdaViT: Adaptive Vision Transformer for Flexible Pretrain and Finetune with Variable 3D Medical Image Modalities 4 Apr 2025 · 0 repositories · arXiv:2504.03589
-
Beyond Progress Measures: Theoretical Insights into the Mechanism of Grokking 4 Apr 2025 · 1 repository · arXiv:2504.03162
-
Block Toeplitz Sparse Precision Matrix Estimation for Large-Scale Interval-Valued Time Series Forecasting 4 Apr 2025 · 0 repositories · arXiv:2504.03322
-
Detecting underdetermination in parameterized quantum circuits 4 Apr 2025 · 0 repositories · arXiv:2504.03315
-
Do LLM Evaluators Prefer Themselves for a Reason? 4 Apr 2025 · 1 repository · arXiv:2504.03846
-
DP-LET: An Efficient Spatio-Temporal Network Traffic Prediction Framework 4 Apr 2025 · 0 repositories · arXiv:2504.03792
-
Dynamic Importance in Diffusion U-Net for Enhanced Image Synthesis 4 Apr 2025 · 1 repository · arXiv:2504.03471
-
Efficient Dynamic Clustering-Based Document Compression for Retrieval-Augmented-Generation 4 Apr 2025 · 1 repository · arXiv:2504.03165
-
Electromyography-Based Gesture Recognition: Hierarchical Feature Extraction for Enhanced Spatial-Temporal Dynamics 4 Apr 2025 · 0 repositories · arXiv:2504.03221
-
FADConv: A Frequency-Aware Dynamic Convolution for Farmland Non-agriculturalization Identification and Segmentation 4 Apr 2025 · 0 repositories · arXiv:2504.03510
-
FaR: Enhancing Multi-Concept Text-to-Image Diffusion via Concept Fusion and Localized Refinement 4 Apr 2025 · 0 repositories · arXiv:2504.03292
-
Generating ensembles of spatially-coherent in-situ forecasts using flow matching 4 Apr 2025 · 0 repositories · arXiv:2504.03463
-
Generative AI Enhanced Financial Risk Management Information Retrieval 4 Apr 2025 · 1 repository · arXiv:2504.06293
-
HeterMoE: Efficient Training of Mixture-of-Experts Models on Heterogeneous GPUs 4 Apr 2025 · 0 repositories · arXiv:2504.03871
-
HumanDreamer-X: Photorealistic Single-image Human Avatars Reconstruction via Gaussian Restoration 4 Apr 2025 · 0 repositories · arXiv:2504.03536
-
Inherent and emergent liability issues in LLM-based agentic systems: a principal-agent perspective 4 Apr 2025 · 0 repositories · arXiv:2504.03255
-
JanusDDG: A Thermodynamics-Compliant Model for Sequence-Based Protein Stability via Two-Fronts Multi-Head Attention 4 Apr 2025 · 1 repository · arXiv:2504.03278
-
Joint Retrieval of Cloud properties using Attention-based Deep Learning Models 4 Apr 2025 · 0 repositories · arXiv:2504.03133
-
Learning Natural Language Constraints for Safe Reinforcement Learning of Language Agents 4 Apr 2025 · 0 repositories · arXiv:2504.03185
-
Mamba as a Bridge: Where Vision Foundation Models Meet Vision Language Models for Domain-Generalized Semantic Segmentation 4 Apr 2025 · 1 repository · arXiv:2504.03193
-
Meta-DAN: towards an efficient prediction strategy for page-level handwritten text recognition 4 Apr 2025 · 1 repository · arXiv:2504.03349