Methods › General › Attention Mechanisms › Attention › Papers, page 36
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 36 of 316: papers 3,501 to 3,600 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Wyckoff Transformer: Generation of Symmetric Crystals 4 Mar 2025 · 1 repository · arXiv:2503.02407
-
Zero-Shot Multi-Label Classification of Bangla Documents: Large Decoders Vs. Classic Encoders 4 Mar 2025 · 0 repositories · arXiv:2503.02993
-
From Claims to Evidence: A Unified Framework and Critical Analysis of CNN vs. Transformer vs. Mamba in Medical Image Segmentation 3 Mar 2025 · 1 repository · arXiv:2503.01306
-
Neural ODE Transformers: Analyzing Internal Dynamics and Adaptive Fine-tuning 3 Mar 2025 · 0 repositories · arXiv:2503.01329Syntology 0 ran · 1 unverified (of 1 harvested sample)
-
Enhancing Social Media Rumor Detection: A Semantic and Graph Neural Network Approach for the 2024 Global Election 3 Mar 2025 · 0 repositories · arXiv:2503.01394
-
AC-Lite : A Lightweight Image Captioning Model for Low-Resource Assamese Language 3 Mar 2025 · 0 repositories · arXiv:2503.01453
-
SrSv: Integrating Sequential Rollouts with Sequential Value Estimation for Multi-agent Reinforcement Learning 3 Mar 2025 · 0 repositories · arXiv:2503.01458
-
Liger: Linearizing Large Language Models to Gated Recurrent Structures 3 Mar 2025 · 1 repository · arXiv:2503.01496Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
An Efficient Approach to Detecting Lung Nodules Using Swin Transformer 3 Mar 2025 · 0 repositories · arXiv:2503.01592
-
Machine Learners Should Acknowledge the Legal Implications of Large Language Models as Personal Data 3 Mar 2025 · 0 repositories · arXiv:2503.01630
-
SAGE: A Framework of Precise Retrieval for RAG 3 Mar 2025 · 0 repositories · arXiv:2503.01713
-
LLMInit: A Free Lunch from Large Language Models for Selective Initialization of Recommendation 3 Mar 2025 · 0 repositories · arXiv:2503.01814
-
A Generalized Theory of Mixup for Structure-Preserving Synthetic Data 3 Mar 2025 · 1 repository · arXiv:2503.02645
-
A Hybrid CNN-Transformer Model for Heart Disease Prediction Using Life History Data 3 Mar 2025 · 0 repositories · arXiv:2503.02124
-
ACCORD: Alleviating Concept Coupling through Dependence Regularization for Text-to-Image Diffusion Personalization 3 Mar 2025 · 0 repositories · arXiv:2503.01122
-
Architectural and Inferential Inductive Biases For Exchangeable Sequence Modeling 3 Mar 2025 · 1 repository · arXiv:2503.01215
-
AskToAct: Enhancing LLMs Tool Use via Self-Correcting Clarification 3 Mar 2025 · 0 repositories · arXiv:2503.01940
-
Attention Condensation via Sparsity Induced Regularized Training 3 Mar 2025 · 0 repositories · arXiv:2503.01564
-
Boolean-aware Attention for Dense Retrieval 3 Mar 2025 · 0 repositories · arXiv:2503.01753
-
Cancer Type, Stage and Prognosis Assessment from Pathology Reports using LLMs 3 Mar 2025 · 1 repository · arXiv:2503.01194
-
Dementia Insights: A Context-Based MultiModal Approach 3 Mar 2025 · 0 repositories · arXiv:2503.01226
-
Efficient or Powerful? Trade-offs Between Machine Learning and Deep Learning for Mental Illness Detection on Social Media 3 Mar 2025 · 0 repositories · arXiv:2503.01082
-
Every SAM Drop Counts: Embracing Semantic Priors for Multi-Modality Image Fusion and Beyond 3 Mar 2025 · 0 repositories · arXiv:2503.01210
-
Fault Localization and State Estimation of Power Grid under Parallel Cyber-Physical Attacks 3 Mar 2025 · 0 repositories · arXiv:2503.05797
-
Forgetting Transformer: Softmax Attention with a Forget Gate 3 Mar 2025 · 1 repository · arXiv:2503.02130Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 13 harvested samples)
-
GRAIN: Exact Graph Reconstruction from Gradients 3 Mar 2025 · 1 repository · arXiv:2503.01838Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples)
-
HanDrawer: Leveraging Spatial Information to Render Realistic Hands Using a Conditional Diffusion Model in Single Stage 3 Mar 2025 · 0 repositories · arXiv:2503.02127
-
HeterRec: Heterogeneous Information Transformer for Scalable Sequential Recommendation 3 Mar 2025 · 0 repositories · arXiv:2503.01469
-
HoH: A Dynamic Benchmark for Evaluating the Impact of Outdated Information on Retrieval-Augmented Generation 3 Mar 2025 · 0 repositories · arXiv:2503.04800Syntology 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
HOP: Heterogeneous Topology-based Multimodal Entanglement for Co-Speech Gesture Generation 3 Mar 2025 · 0 repositories · arXiv:2503.01175
-
How simple can you go? An off-the-shelf transformer approach to molecular dynamics 3 Mar 2025 · 1 repository · arXiv:2503.01431
-
Interactive Gadolinium-Free MRI Synthesis: A Transformer with Localization Prompt Learning 3 Mar 2025 · 1 repository · arXiv:2503.01265
-
Label Ranker: Self-Aware Preference for Classification Label Position in Visual Masked Self-Supervised Pre-Trained Model 3 Mar 2025 · 1 repository
-
Linear Representations of Political Perspective Emerge in Large Language Models 3 Mar 2025 · 1 repository · arXiv:2503.02080
-
MAPS: Motivation-Aware Personalized Search via LLM-Driven Consultation Alignment 3 Mar 2025 · 1 repository · arXiv:2503.01711Syntology official (archive's flag): 5 ran · 5 ran (of which 4 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
MeshPad: Interactive Sketch-Conditioned Artist-Designed Mesh Generation and Editing 3 Mar 2025 · 0 repositories · arXiv:2503.01425
-
MI-DETR: An Object Detection Model with Multi-time Inquiries Mechanism 3 Mar 2025 · 1 repository · arXiv:2503.01463
-
MRI super-resolution reconstruction using efficient diffusion probabilistic model with residual shifting 3 Mar 2025 · 1 repository · arXiv:2503.01576
-
Object-Aware Video Matting with Cross-Frame Guidance 3 Mar 2025 · 0 repositories · arXiv:2503.01262
-
Primer C-VAE: An interpretable deep learning primer design method to detect emerging virus variants 3 Mar 2025 · 0 repositories · arXiv:2503.01459
-
Primus: Enforcing Attention Usage for 3D Medical Image Segmentation 3 Mar 2025 · 0 repositories · arXiv:2503.01835
-
Provable Benefits of Task-Specific Prompts for In-context Learning 3 Mar 2025 · 1 repository · arXiv:2503.02102
-
Retrieval-Augmented Perception: High-Resolution Image Perception Meets Visual RAG 3 Mar 2025 · 1 repository · arXiv:2503.01222
-
Rotary Outliers and Rotary Offset Features in Large Language Models 3 Mar 2025 · 0 repositories · arXiv:2503.01832
-
RSQ: Learning from Important Tokens Leads to Better Quantized LLMs 3 Mar 2025 · 1 repository · arXiv:2503.01820
-
SePer: Measure Retrieval Utility Through The Lens Of Semantic Perplexity Reduction 3 Mar 2025 · 1 repository · arXiv:2503.01478Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples)
-
SRAG: Structured Retrieval-Augmented Generation for Multi-Entity Question Answering over Wikipedia Graph 3 Mar 2025 · 0 repositories · arXiv:2503.01346
-
Streaming Piano Transcription Based on Consistent Onset and Offset Decoding with Sustain Pedal Detection 3 Mar 2025 · 0 repositories · arXiv:2503.01362
-
SVDC: Consistent Direct Time-of-Flight Video Depth Completion with Frequency Selective Fusion 3 Mar 2025 · 1 repository · arXiv:2503.01257Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 7 unverified (of 14 harvested samples) · 5 pointer-only (licence)
-
Syntactic Learnability of Echo State Neural Language Models at Scale 3 Mar 2025 · 0 repositories · arXiv:2503.01724
-
ToLo: A Two-Stage, Training-Free Layout-To-Image Generation Framework For High-Overlap Layouts 3 Mar 2025 · 1 repository · arXiv:2503.01667
-
Transferring between sparse and dense matching via probabilistic reweighting 3 Mar 2025 · 0 repositories · arXiv:2503.01472
-
Unify and Anchor: A Context-Aware Transformer for Cross-Domain Time Series Forecasting 3 Mar 2025 · 0 repositories · arXiv:2503.01157
-
Using (Not so) Large Language Models for Generating Simulation Models in a Formal DSL -- A Study on Reaction Networks 3 Mar 2025 · 0 repositories · arXiv:2503.01675
-
ViKANformer: Embedding Kolmogorov Arnold Networks in Vision Transformers for Pattern-Based Learning 3 Mar 2025 · 0 repositories · arXiv:2503.01124
-
WeightedKV: Attention Scores Weighted Key-Value Cache Merging for Large Language Models 3 Mar 2025 · 0 repositories · arXiv:2503.01330
-
Why Is Spatial Reasoning Hard for VLMs? An Attention Mechanism Perspective on Focus Areas 3 Mar 2025 · 1 repository · arXiv:2503.01773Syntology official (archive's flag): 6 ran · 7 ran (of which 4 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
SemViQA: A Semantic Question Answering System for Vietnamese Information Fact-Checking 2 Mar 2025 · 1 repository · arXiv:2503.00955
-
MedUnifier: Unifying Vision-and-Language Pre-training on Medical Data with Vision Generation Task using Discrete Visual Representations 2 Mar 2025 · 0 repositories · arXiv:2503.01019
-
CAGN-GAT Fusion: A Hybrid Contrastive Attentive Graph Neural Network for Network Intrusion Detection 2 Mar 2025 · 1 repository · arXiv:2503.00961
-
Confidence Based Asynchronous Integrated Communication and Localization Networks Using Pulsed UWB Signals 2 Mar 2025 · 0 repositories · arXiv:2503.00922
-
Enhancing Text Editing for Grammatical Error Correction: Arabic as a Case Study 2 Mar 2025 · 0 repositories · arXiv:2503.00985
-
ER-RAG: Enhance RAG with ER-Based Unified Modeling of Heterogeneous Data Sources 2 Mar 2025 · 0 repositories · arXiv:2504.06271
-
FACROC: a fairness measure for FAir Clustering through ROC curves 2 Mar 2025 · 1 repository · arXiv:2503.00854
-
Forecasting realized volatility in the stock market: a path-dependent perspective 2 Mar 2025 · 0 repositories · arXiv:2503.00851
-
Hierarchical graph sampling based minibatch learning with chain preservation and variance reduction 2 Mar 2025 · 1 repository · arXiv:2503.00860
-
Spiking World Model with Multi-Compartment Neurons for Model-based Reinforcement Learning 2 Mar 2025 · 1 repository · arXiv:2503.00713
-
LightEndoStereo: A Real-time Lightweight Stereo Matching Method for Endoscopy Images 2 Mar 2025 · 1 repository · arXiv:2503.00731
-
Optimizing Multi-Hop Document Retrieval Through Intermediate Representations 2 Mar 2025 · 0 repositories · arXiv:2503.04796
-
Patch-wise Structural Loss for Time Series Forecasting 2 Mar 2025 · 1 repository · arXiv:2503.00877
-
Revisiting CAD Model Generation by Learning Raster Sketch 2 Mar 2025 · 0 repositories · arXiv:2503.00928
-
Training-Free Dataset Pruning for Instance Segmentation 2 Mar 2025 · 1 repository · arXiv:2503.00828Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 4 where Syntology's instrument failed) · 3 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
Transformer Based Self-Context Aware Prediction for Few-Shot Anomaly Detection in Videos 2 Mar 2025 · 0 repositories · arXiv:2503.00670
-
Transformer Meets Twicing: Harnessing Unattended Residual Information 2 Mar 2025 · 1 repository · arXiv:2503.00687Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample) · 1 pointer-only (licence)
-
Unmasking Digital Falsehoods: A Comparative Analysis of LLM-Based Misinformation Detection Strategies 2 Mar 2025 · 0 repositories · arXiv:2503.00724
-
Variance reduction in output from generative AI 2 Mar 2025 · 0 repositories · arXiv:2503.01033
-
Wavelet-Driven Masked Image Modeling: A Path to Efficient Visual Representation 2 Mar 2025 · 0 repositories · arXiv:2503.00782
-
Pseudo-Knowledge Graph: Meta-Path Guided Retrieval and In-Graph Text for RAG-Equipped LLM 1 Mar 2025 · 0 repositories · arXiv:2503.00309
-
BGM2Pose: Active 3D Human Pose Estimation with Non-Stationary Sounds 1 Mar 2025 · 0 repositories · arXiv:2503.00389
-
PodAgent: A Comprehensive Framework for Podcast Generation 1 Mar 2025 · 1 repository · arXiv:2503.00455
-
BadJudge: Backdoor Vulnerabilities of LLM-as-a-Judge 1 Mar 2025 · 0 repositories · arXiv:2503.00596
-
3D Shape Completion using Multi-Resolution Spectral Encoding 1 Mar 2025 · 1 repository
-
Advances in Anti-Deception Jamming Strategies for Radar Systems: A Survey 1 Mar 2025 · 0 repositories · arXiv:2503.00285
-
Approaching the Limits to EFL Writing Enhancement with AI-generated Text and Diverse Learners 1 Mar 2025 · 0 repositories · arXiv:2503.00367
-
Artificially Generated Visual Scanpath Improves Multi-label Thoracic Disease Classification in Chest X-Ray Images 1 Mar 2025 · 1 repository · arXiv:2503.00657
-
CL-MoE: Enhancing Multimodal Large Language Model with Dual Momentum Mixture-of-Experts for Continual Visual Question Answering 1 Mar 2025 · 0 repositories · arXiv:2503.00413
-
Cross-Attention Fusion of MRI and Jacobian Maps for Alzheimer's Disease Diagnosis 1 Mar 2025 · 0 repositories · arXiv:2503.00586
-
Hierarchical Multi-Stage BERT Fusion Framework with Dual Attention for Enhanced Cyberbullying Detection in Social Media 1 Mar 2025 · 0 repositories · arXiv:2503.00342
-
Learning Conditional Average Treatment Effects in Regression Discontinuity Designs using Bayesian Additive Regression Trees 1 Mar 2025 · 0 repositories · arXiv:2503.00326
-
MIRROR: Multi-Modal Pathological Self-Supervised Representation Learning via Modality Alignment and Retention 1 Mar 2025 · 1 repository · arXiv:2503.00374
-
Never too Prim to Swim: An LLM-Enhanced RL-based Adaptive S-Surface Controller for AUVs under Extreme Sea Conditions 1 Mar 2025 · 0 repositories · arXiv:2503.00527
-
Psychological Counseling Ability of Large Language Models 1 Mar 2025 · 0 repositories · arXiv:2503.07627
-
Qilin: A Multimodal Information Retrieval Dataset with APP-level User Sessions 1 Mar 2025 · 1 repository · arXiv:2503.00501
-
SegImgNet: Segmentation-Guided Dual-Branch Network for Retinal Disease Diagnoses 1 Mar 2025 · 1 repository · arXiv:2503.00267
-
Sentence-level Reward Model can Generalize Better for Aligning LLM from Human Preference 1 Mar 2025 · 0 repositories · arXiv:2503.04793
-
Streaming Video Question-Answering with In-context Video KV-Cache Retrieval 1 Mar 2025 · 1 repository · arXiv:2503.00540Syntology official (archive's flag): 2 ran · 2 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified; every one of the 2 samples that ran constructed an object rather than computing a result (of 6 harvested samples) · 6 pointer-only (licence)
-
TSDW: A Tri-Stream Dynamic Weight Network for Cloth-Changing Person Re-Identification 1 Mar 2025 · 0 repositories · arXiv:2503.00477
-
Two-stream Beats One-stream: Asymmetric Siamese Network for Efficient Visual Tracking 1 Mar 2025 · 1 repository · arXiv:2503.00516
-
U-NIAH: Unified RAG and LLM Evaluation for Long Context Needle-In-A-Haystack 1 Mar 2025 · 1 repository · arXiv:2503.00353
-
UL-UNAS: Ultra-Lightweight U-Nets for Real-Time Speech Enhancement via Network Architecture Search 1 Mar 2025 · 1 repository · arXiv:2503.00340