Methods › General › Attention Mechanisms › Attention › Papers, page 56
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 56 of 316: papers 5,501 to 5,600 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Traits of a Leader: User Influence Level Prediction through Sociolinguistic Modeling 5 Jan 2025 · 0 repositories · arXiv:2501.04046
-
Unified Guidance for Geometry-Conditioned Molecular Generation 5 Jan 2025 · 0 repositories · arXiv:2501.02526
-
Watch Video, Catch Keyword: Context-aware Keyword Attention for Moment Retrieval and Highlight Detection 5 Jan 2025 · 1 repository · arXiv:2501.02504Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 8 where Syntology's instrument failed) · 2 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
Bridge the Inference Gaps of Neural Processes via Expectation Maximization 4 Jan 2025 · 1 repository · arXiv:2501.03264
-
Context Aware Lemmatization and Morphological Tagging Method in Turkish 4 Jan 2025 · 0 repositories · arXiv:2501.02361
-
CorrFill: Enhancing Faithfulness in Reference-based Inpainting with Correspondence Guidance in Diffusion Models 4 Jan 2025 · 0 repositories · arXiv:2501.02355
-
Deep Learning-Driven Segmentation of Ischemic Stroke Lesions Using Multi-Channel MRI 4 Jan 2025 · 0 repositories · arXiv:2501.02287
-
Examining the Robustness of Homogeneity Bias to Hyperparameter Adjustments in GPT-4 4 Jan 2025 · 0 repositories · arXiv:2501.02211
-
Exploring the Capabilities and Limitations of Large Language Models for Radiation Oncology Decision Support 4 Jan 2025 · 0 repositories · arXiv:2501.02346
-
Graph-Aware Isomorphic Attention for Adaptive Dynamics in Transformers 4 Jan 2025 · 1 repository · arXiv:2501.02393
-
Guiding Medical Vision-Language Models with Explicit Visual Prompts: Framework Design and Comprehensive Exploration of Prompt Variations 4 Jan 2025 · 0 repositories · arXiv:2501.02385
-
Knowledge Graph Retrieval-Augmented Generation for LLM-based Recommendation 4 Jan 2025 · 0 repositories · arXiv:2501.02226
-
Learning Evolution via Optimization Knowledge Adaptation 4 Jan 2025 · 0 repositories · arXiv:2501.02200
-
LLM Content Moderation and User Satisfaction: Evidence from Response Refusals in Chatbot Arena 4 Jan 2025 · 0 repositories · arXiv:2501.03266
-
TDM: Temporally-Consistent Diffusion Model for All-in-One Real-World Video Restoration 4 Jan 2025 · 0 repositories · arXiv:2501.02269
-
The Application of Large Language Models in Recommendation Systems 4 Jan 2025 · 0 repositories · arXiv:2501.02178
-
UAVs Meet LLMs: Overviews and Perspectives Toward Agentic Low-Altitude Mobility 4 Jan 2025 · 1 repository · arXiv:2501.02341
-
V2X-DGPE: Addressing Domain Gaps and Pose Errors for Robust Collaborative 3D Object Detection 4 Jan 2025 · 1 repository · arXiv:2501.02363
-
A Separable Self-attention Inspired by the State Space Model for Computer Vision 3 Jan 2025 · 1 repository · arXiv:2501.02040
-
A Survey on Large Language Models with some Insights on their Capabilities and Limitations 3 Jan 2025 · 0 repositories · arXiv:2501.04040
-
Age-Based Device Selection and Transmit Power Optimization in Over-the-Air Federated Learning 3 Jan 2025 · 0 repositories · arXiv:2501.01828
-
AgentRefine: Enhancing Agent Generalization through Refinement Tuning 3 Jan 2025 · 0 repositories · arXiv:2501.01702
-
AI-Powered Cow Detection in Complex Farm Environments 3 Jan 2025 · 0 repositories · arXiv:2501.02080
-
Applying Text Mining to Analyze Human Question Asking in Creativity Research 3 Jan 2025 · 1 repository · arXiv:2501.02090
-
ArtCrafter: Text-Image Aligning Style Transfer via Embedding Reframing 3 Jan 2025 · 0 repositories · arXiv:2501.02064
-
BARTPredict: Empowering IoT Security with LLM-Driven Cyber Threat Prediction 3 Jan 2025 · 0 repositories · arXiv:2501.01664
-
BERT4MIMO: A Foundation Model using BERT Architecture for Massive MIMO Channel State Information Prediction 3 Jan 2025 · 1 repository · arXiv:2501.01802
-
Catch Causal Signals from Edges for Label Imbalance in Graph Classification 3 Jan 2025 · 1 repository · arXiv:2501.01707
-
Classifier-Guided Captioning Across Modalities 3 Jan 2025 · 0 repositories · arXiv:2501.03183
-
Compressed Domain Prior-Guided Video Super-Resolution for Cloud Gaming Content 3 Jan 2025 · 0 repositories · arXiv:2501.01773
-
CoT-based Synthesizer: Enhancing LLM Performance through Answer Synthesis 3 Jan 2025 · 1 repository · arXiv:2501.01668
-
DreamMask: Boosting Open-vocabulary Panoptic Segmentation with Synthetic Data 3 Jan 2025 · 0 repositories · arXiv:2501.02048
-
End-to-End Long Document Summarization using Gradient Caching 3 Jan 2025 · 0 repositories · arXiv:2501.01805
-
GoBERT: Gene Ontology Graph Informed BERT for Universal Gene Function Prediction 3 Jan 2025 · 0 repositories · arXiv:2501.01930
-
HSTforU: anomaly detection in aerial and ground-based videos with hierarchical spatio-temporal transformer for U-net 3 Jan 2025 · 1 repository
-
IAM: Enhancing RGB-D Instance Segmentation with New Benchmarks 3 Jan 2025 · 3 repositories · arXiv:2501.01685
-
IGAF: Incremental Guided Attention Fusion for Depth Super-Resolution 3 Jan 2025 · 0 repositories · arXiv:2501.01723
-
Improving Transducer-Based Spoken Language Understanding with Self-Conditioned CTC and Knowledge Transfer 3 Jan 2025 · 0 repositories · arXiv:2501.01936
-
LLMs & Legal Aid: Understanding Legal Needs Exhibited Through User Queries 3 Jan 2025 · 0 repositories · arXiv:2501.01711
-
MADGEN: Mass-Spec attends to De Novo Molecular generation 3 Jan 2025 · 1 repository · arXiv:2501.01950Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
MIRAGE: Exploring How Large Language Models Perform in Complex Social Interactive Environments 3 Jan 2025 · 1 repository · arXiv:2501.01652
-
Mitigating Hallucination for Large Vision Language Model by Inter-Modality Correlation Calibration Decoding 3 Jan 2025 · 1 repository · arXiv:2501.01926
-
PersonaAI: Leveraging Retrieval-Augmented Generation and Personalized Context for AI-Driven Digital Avatars 3 Jan 2025 · 0 repositories · arXiv:2503.15489
-
Quantitative Gait Analysis from Single RGB Videos Using a Dual-Input Transformer-Based Network 3 Jan 2025 · 1 repository · arXiv:2501.01689
-
Relaxation-assisted reverse annealing on nonnegative/binary matrix factorization 3 Jan 2025 · 0 repositories · arXiv:2501.02114
-
Spot Risks Before Speaking! Unraveling Safety Attention Heads in Large Vision-Language Models 3 Jan 2025 · 0 repositories · arXiv:2501.02029
-
TCPFormer: Learning Temporal Correlation with Implicit Pose Proxy for 3D Human Pose Estimation 3 Jan 2025 · 1 repository · arXiv:2501.01770
-
Towards Hard and Soft Shadow Removal via Dual-Branch Separation Network and Vision Transformer 3 Jan 2025 · 0 repositories · arXiv:2501.01864
-
Transformer-Driven Inverse Problem Transform for Fast Blind Hyperspectral Image Dehazing 3 Jan 2025 · 0 repositories · arXiv:2501.01924
-
Turning Logic Against Itself : Probing Model Defenses Through Contrastive Questions 3 Jan 2025 · 1 repository · arXiv:2501.01872Syntology official: no sample here; runs from other or unrecorded repositories · 9 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 10 harvested samples) · 1 pointer-only (licence)
-
VidFormer: A novel end-to-end framework fused by 3DCNN and Transformer for Video-based Remote Physiological Measurement 3 Jan 2025 · 0 repositories · arXiv:2501.01691
-
Virgo: A Preliminary Exploration on Reproducing o1-like MLLM 3 Jan 2025 · 2 repositories · arXiv:2501.01904
-
3D-LLaVA: Towards Generalist 3D LMMs with Omni Superpoint Transformer 2 Jan 2025 · 0 repositories · arXiv:2501.01163
-
A Unified Hyperparameter Optimization Pipeline for Transformer-Based Time Series Forecasting Models 2 Jan 2025 · 1 repository · arXiv:2501.01394
-
An Efficient Attention Mechanism for Sequential Recommendation Tasks: HydraRec 2 Jan 2025 · 0 repositories · arXiv:2501.01242
-
BeliN: A Novel Corpus for Bengali Religious News Headline Generation using Contextual Feature Fusion 2 Jan 2025 · 1 repository · arXiv:2501.01069
-
Detail Matters: Mamba-Inspired Joint Unfolding Network for Snapshot Spectral Compressive Imaging 2 Jan 2025 · 1 repository · arXiv:2501.01262
-
Disambiguation of Chinese Polyphones in an End-to-End Framework with Semantic Features Extracted by Pre-trained BERT 2 Jan 2025 · 0 repositories · arXiv:2501.01102
-
Does a Large Language Model Really Speak in Human-Like Language? 2 Jan 2025 · 0 repositories · arXiv:2501.01273
-
Dynamic Attention-Guided Context Decoding for Mitigating Context Faithfulness Hallucinations in Large Language Models 2 Jan 2025 · 0 repositories · arXiv:2501.01059
-
EHCTNet: Enhanced Hybrid of CNN and Transformer Network for Remote Sensing Image Change Detection 2 Jan 2025 · 0 repositories · arXiv:2501.01238
-
FlashInfer: Efficient and Customizable Attention Engine for LLM Inference Serving 2 Jan 2025 · 1 repository · arXiv:2501.01005Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 11 unverified (of 16 harvested samples)
-
Graph Generative Pre-trained Transformer 2 Jan 2025 · 0 repositories · arXiv:2501.01073
-
Hadamard Attention Recurrent Transformer: A Strong Baseline for Stereo Matching Transformer 2 Jan 2025 · 1 repository · arXiv:2501.01023
-
Harnessing Multi-Agent LLMs for Complex Engineering Problem-Solving: A Framework for Senior Design Projects 2 Jan 2025 · 0 repositories · arXiv:2501.01205
-
KANS: Knowledge Discovery Graph Attention Network for Soft Sensing in Multivariate Industrial Processes 2 Jan 2025 · 0 repositories · arXiv:2501.02015
-
Large Language Models for Mental Health Diagnostic Assessments: Exploring The Potential of Large Language Models for Assisting with Mental Health Diagnostic Assessments -- The Depression and Anxiety Case 2 Jan 2025 · 0 repositories · arXiv:2501.01305
-
Learning Spectral Methods by Transformers 2 Jan 2025 · 0 repositories · arXiv:2501.01312
-
Leverage Cross-Attention for End-to-End Open-Vocabulary Panoptic Reconstruction 2 Jan 2025 · 0 repositories · arXiv:2501.01119
-
Long-range Brain Graph Transformer 2 Jan 2025 · 1 repository · arXiv:2501.01100Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Missing Data as Augmentation in the Earth Observation Domain: A Multi-View Learning Approach 2 Jan 2025 · 1 repository · arXiv:2501.01132
-
MSWA: Refining Local Attention with Multi-ScaleWindow Attention 2 Jan 2025 · 0 repositories · arXiv:2501.01039
-
Multi-Head Explainer: A General Framework to Improve Explainability in CNNs and Transformers 2 Jan 2025 · 0 repositories · arXiv:2501.01311
-
Multi-Modal Video Feature Extraction for Popularity Prediction 2 Jan 2025 · 0 repositories · arXiv:2501.01422
-
Multi-Task Semantic Communication With Graph Attention-Based Feature Correlation Extraction 2 Jan 2025 · 0 repositories · arXiv:2501.02006
-
Nested Attention: Semantic-aware Attention Values for Concept Personalization 2 Jan 2025 · 0 repositories · arXiv:2501.01407
-
nnY-Net: Swin-NeXt with Cross-Attention for 3D Medical Images Segmentation 2 Jan 2025 · 0 repositories · arXiv:2501.01406
-
Operator Learning for Reconstructing Flow Fields from Sparse Measurements: an Energy Transformer Approach 2 Jan 2025 · 0 repositories · arXiv:2501.08339
-
Predicting the Performance of Black-box LLMs through Self-Queries 2 Jan 2025 · 1 repository · arXiv:2501.01558Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 2 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
RealDiffFusionNet: Neural Controlled Differential Equation Informed Multi-Head Attention Fusion Networks for Disease Progression Modeling Using Real-World Data 2 Jan 2025 · 0 repositories · arXiv:2501.02025
-
Reconstruction vs. Generation: Taming Optimization Dilemma in Latent Diffusion Models 2 Jan 2025 · 2 repositories · arXiv:2501.01423
-
RingFormer: A Neural Vocoder with Ring Attention and Convolution-Augmented Transformer 2 Jan 2025 · 1 repository · arXiv:2501.01182
-
ROME: Robust Model Ensembling for Semantic Communication Against Semantic Jamming Attacks 2 Jan 2025 · 0 repositories · arXiv:2501.01172
-
ScarNet: A Novel Foundation Model for Automated Myocardial Scar Quantification from LGE in Cardiac MRI 2 Jan 2025 · 0 repositories · arXiv:2501.01372
-
SeedVR: Seeding Infinity in Diffusion Transformer Towards Generic Video Restoration 2 Jan 2025 · 0 repositories · arXiv:2501.01320
-
TART: Token-based Architecture Transformer for Neural Network Performance Prediction 2 Jan 2025 · 1 repository · arXiv:2501.02007
-
TED: Turn Emphasis with Dialogue Feature Attention for Emotion Recognition in Conversation 2 Jan 2025 · 0 repositories · arXiv:2501.01123
-
Test-time Controllable Image Generation by Explicit Spatial Constraint Enforcement 2 Jan 2025 · 0 repositories · arXiv:2501.01368
-
Toward Inclusive Educational AI: Auditing Frontier LLMs through a Multiplexity Lens 2 Jan 2025 · 0 repositories · arXiv:2501.03259
-
Towards Adversarially Robust Deep Metric Learning 2 Jan 2025 · 0 repositories · arXiv:2501.01025
-
Weakly Supervised Learning on Large Graphs 2 Jan 2025 · 0 repositories · arXiv:2501.02021
-
3D-MVP: 3D Multiview Pretraining for Manipulation 1 Jan 2025 · 0 repositories
-
A Polarization-Aided Transformer for Image Deblurring via Motion Vector Decomposition 1 Jan 2025 · 0 repositories
-
A Universal Scale-Adaptive Deformable Transformer for Image Restoration across Diverse Artifacts 1 Jan 2025 · 1 repository
-
A4A: Adapter for Adapter Transfer via All-for-All Mapping for Cross-Architecture Models 1 Jan 2025 · 0 repositories
-
ABBSPO: Adaptive Bounding Box Scaling and Symmetric Prior based Orientation Prediction for Detecting Aerial Image Objects 1 Jan 2025 · 0 repositories
-
ABC-Former: Auxiliary Bimodal Cross-domain Transformer with Interactive Channel Attention for White Balance 1 Jan 2025 · 1 repository
-
ACL: Activating Capability of Linear Attention for Image Restoration 1 Jan 2025 · 0 repositories
-
Activating Sparse Part Concepts for 3D Class Incremental Learning 1 Jan 2025 · 0 repositories
-
Adaptive Part Learning for Fine-Grained Generalized Category Discovery: A Plug-and-Play Enhancement 1 Jan 2025 · 0 repositories