Methods › General › Attention Mechanisms › Attention › Papers, page 74
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 74 of 316: papers 7,301 to 7,400 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
LiTformer: Efficient Modeling and Analysis of High-Speed Link Transmitters Using Non-Autoregressive Transformer 18 Nov 2024 · 0 repositories · arXiv:2411.11699
-
Mechanism and Emergence of Stacked Attention Heads in Multi-Layer Transformers 18 Nov 2024 · 0 repositories · arXiv:2411.12118
-
Multi-Hyperbolic Space-based Heterogeneous Graph Attention Network 18 Nov 2024 · 0 repositories · arXiv:2411.11283
-
PerfCodeGen: Improving Performance of LLM Generated Code with Execution Feedback 18 Nov 2024 · 1 repository · arXiv:2412.03578Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
Popular LLMs Amplify Race and Gender Disparities in Human Mobility 18 Nov 2024 · 0 repositories · arXiv:2411.14469
-
SeqProFT: Applying LoRA Finetuning for Sequence-only Protein Property Predictions 18 Nov 2024 · 0 repositories · arXiv:2411.11530
-
ST-Tree with Interpretability for Multivariate Time Series Classification 18 Nov 2024 · 0 repositories · arXiv:2411.11620
-
Suicide Risk Assessment on Social Media with Semi-Supervised Learning 18 Nov 2024 · 0 repositories · arXiv:2411.12767
-
Superpixel-informed Implicit Neural Representation for Multi-Dimensional Data 18 Nov 2024 · 0 repositories · arXiv:2411.11356
-
Enhancing LLM Reasoning with Reward-guided Tree Search 18 Nov 2024 · 2 repositories · arXiv:2411.11694
-
The ADUULM-360 Dataset -- A Multi-Modal Dataset for Depth Estimation in Adverse Weather 18 Nov 2024 · 0 repositories · arXiv:2411.11455
-
TimeFormer: Capturing Temporal Relationships of Deformable 3D Gaussians for Robust Reconstruction 18 Nov 2024 · 1 repository · arXiv:2411.11941
-
Towards a Practical Ethics of Generative AI in Creative Production Processes 18 Nov 2024 · 0 repositories · arXiv:2412.03579
-
TrojanRobot: Physical-World Backdoor Attacks Against VLM-based Robotic Manipulation 18 Nov 2024 · 0 repositories · arXiv:2411.11683
-
Uncovering the role of semantic and acoustic cues in normal and dichotic listening 18 Nov 2024 · 0 repositories · arXiv:2411.11308
-
Understanding Student Sentiment on Mental Health Support in Colleges Using Large Language Models 18 Nov 2024 · 0 repositories · arXiv:2412.04326
-
Unveiling the Inflexibility of Adaptive Embedding in Traffic Forecasting 18 Nov 2024 · 1 repository · arXiv:2411.11448
-
VersaTune: An Efficient Data Composition Framework for Training Multi-Capability LLMs 18 Nov 2024 · 1 repository · arXiv:2411.11266
-
Video-to-Task Learning via Motion-Guided Attention for Few-Shot Action Recognition 18 Nov 2024 · 0 repositories · arXiv:2411.11335
-
Different Horses for Different Courses: Comparing Bias Mitigation Algorithms in ML 17 Nov 2024 · 0 repositories · arXiv:2411.11101
-
Direct and Explicit 3D Generation from a Single Image 17 Nov 2024 · 0 repositories · arXiv:2411.10947
-
Exploiting VLM Localizability and Semantics for Open Vocabulary Action Detection 17 Nov 2024 · 1 repository · arXiv:2411.10922
-
Freqformer: Frequency-Domain Transformer for 3-D Visualization and Quantification of Human Retinal Circulation 17 Nov 2024 · 0 repositories · arXiv:2411.11189
-
IVE: Enhanced Probabilistic Forecasting of Intraday Volume Ratio with Transformers 17 Nov 2024 · 0 repositories · arXiv:2411.10956
-
Knowledge-enhanced Transformer for Multivariate Long Sequence Time-series Forecasting 17 Nov 2024 · 0 repositories · arXiv:2411.11046
-
SageAttention2: Efficient Attention with Thorough Outlier Smoothing and Per-thread INT4 Quantization 17 Nov 2024 · 2 repositories · arXiv:2411.10958Syntology official: harvested, nothing ran · 0 ran · 6 unverified (of 6 harvested samples) · 3 pointer-only (licence)
-
Skeleton-Guided Spatial-Temporal Feature Learning for Video-Based Visible-Infrared Person Re-Identification 17 Nov 2024 · 0 repositories · arXiv:2411.11069
-
A Novel Approach to Eliminating Hallucinations in Large Language Model-Assisted Causal Discovery 16 Nov 2024 · 0 repositories · arXiv:2411.12759
-
A Wearable Gait Monitoring System for 17 Gait Parameters Based on Computer Vision 16 Nov 2024 · 0 repositories · arXiv:2411.10739
-
AllRestorer: All-in-One Transformer for Image Restoration under Composite Degradations 16 Nov 2024 · 0 repositories · arXiv:2411.10708
-
Attention-based U-Net Method for Autonomous Lane Detection 16 Nov 2024 · 0 repositories · arXiv:2411.10902
-
Bag of Design Choices for Inference of High-Resolution Masked Generative Transformer 16 Nov 2024 · 1 repository · arXiv:2411.10781
-
Constructing accurate machine-learned potentials and performing highly efficient atomistic simulations to predict structural and thermal properties 16 Nov 2024 · 0 repositories · arXiv:2411.10911
-
Deep BI-RADS Network for Improved Cancer Detection from Mammograms 16 Nov 2024 · 0 repositories · arXiv:2411.10894
-
Explainable DNN-based Beamformer with Postfilter 16 Nov 2024 · 1 repository · arXiv:2411.10854
-
FIAS: Feature Imbalance-Aware Medical Image Segmentation with Dynamic Fusion and Mixing Attention 16 Nov 2024 · 0 repositories · arXiv:2411.10881
-
Infrared-Assisted Single-Stage Framework for Joint Restoration and Fusion of Visible and Infrared Images under Hazy Conditions 16 Nov 2024 · 0 repositories · arXiv:2411.12586
-
IntentGPT: Few-shot Intent Discovery with Large Language Models 16 Nov 2024 · 0 repositories · arXiv:2411.10670
-
MetaLA: Unified Optimal Linear Approximation to Softmax Attention Map 16 Nov 2024 · 1 repository · arXiv:2411.10741Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
MpoxVLM: A Vision-Language Model for Diagnosing Skin Lesions from Mpox Virus Infection 16 Nov 2024 · 1 repository · arXiv:2411.10888
-
One-Layer Transformer Provably Learns One-Nearest Neighbor In Context 16 Nov 2024 · 0 repositories · arXiv:2411.10830
-
SPDFusion: An Infrared and Visible Image Fusion Network Based on a Non-Euclidean Representation of Riemannian Manifolds 16 Nov 2024 · 0 repositories · arXiv:2411.10679
-
A Low-Resolution Image is Worth 1x1 Words: Enabling Fine Image Super-Resolution with Transformers and TaylorShift 15 Nov 2024 · 0 repositories · arXiv:2411.10231
-
A Multi-Scale Spatial-Temporal Network for Wireless Video Transmission 15 Nov 2024 · 0 repositories · arXiv:2411.09936
-
Boundary Attention Constrained Zero-Shot Layout-To-Image Generation 15 Nov 2024 · 0 repositories · arXiv:2411.10495
-
Building 6G Radio Foundation Models with Transformer Architectures 15 Nov 2024 · 0 repositories · arXiv:2411.09996
-
CMATH: Cross-Modality Augmented Transformer with Hierarchical Variational Distillation for Multimodal Emotion Recognition in Conversation 15 Nov 2024 · 0 repositories · arXiv:2411.10060
-
ColorEdit: Training-free Image-Guided Color editing with diffusion model 15 Nov 2024 · 0 repositories · arXiv:2411.10232
-
CoSAM: Self-Correcting SAM for Domain Generalization in 2D Medical Image Segmentation 15 Nov 2024 · 0 repositories · arXiv:2411.10136
-
DaYu: Data-Driven Model for Geostationary Satellite Observed Cloud Images Forecasting 15 Nov 2024 · 0 repositories · arXiv:2411.10144
-
Debias-CLR: A Contrastive Learning Based Debiasing Method for Algorithmic Fairness in Healthcare Applications 15 Nov 2024 · 0 repositories · arXiv:2411.10544
-
Detecting the Adversarially-Learned Injection Attacks via Knowledge Graphs 15 Nov 2024 · 0 repositories
-
DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization 15 Nov 2024 · 0 repositories · arXiv:2411.10193
-
Does Prompt Formatting Have Any Impact on LLM Performance? 15 Nov 2024 · 0 repositories · arXiv:2411.10541
-
EchoMimicV2: Towards Striking, Simplified, and Semi-Body Human Animation 15 Nov 2024 · 1 repository · arXiv:2411.10061Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples)
-
Evidential Federated Learning for Skin Lesion Image Classification 15 Nov 2024 · 0 repositories · arXiv:2411.10071
-
FitDiT: Advancing the Authentic Garment Details for High-fidelity Virtual Try-on 15 Nov 2024 · 2 repositories · arXiv:2411.10499Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 5 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Hysteresis Activation Function for Efficient Inference 15 Nov 2024 · 1 repository · arXiv:2411.10573Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Information Extraction from Clinical Notes: Are We Ready to Switch to Large Language Models? 15 Nov 2024 · 1 repository · arXiv:2411.10020
-
KAT to KANs: A Review of Kolmogorov-Arnold Networks and the Neural Leap Forward 15 Nov 2024 · 0 repositories · arXiv:2411.10622
-
Lateral Movement Detection via Time-aware Subgraph Classification on Authentication Logs 15 Nov 2024 · 0 repositories · arXiv:2411.10279
-
LoRA-LiteE: A Computationally Efficient Framework for Chatbot Preference-Tuning 15 Nov 2024 · 0 repositories · arXiv:2411.09947
-
MARS: Unleashing the Power of Variance Reduction for Training Large Models 15 Nov 2024 · 2 repositories · arXiv:2411.10438Syntology official (archive's flag): 2 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 1 pointer-only (licence)
-
Memorization in Attention-only Transformers 15 Nov 2024 · 1 repository · arXiv:2411.10115
-
Mitigating Parameter Degeneracy using Joint Conditional Diffusion Model for WECC Composite Load Model in Power Systems 15 Nov 2024 · 0 repositories · arXiv:2411.10431
-
Morpho-Aware Global Attention for Image Matting 15 Nov 2024 · 0 repositories · arXiv:2411.10251
-
"On the goals of linguistic theory": Revisiting Chomskyan theories in the era of AI 15 Nov 2024 · 0 repositories · arXiv:2411.10533
-
Probabilistic Prior Driven Attention Mechanism Based on Diffusion Model for Imaging Through Atmospheric Turbulence 15 Nov 2024 · 0 repositories · arXiv:2411.10321
-
Prompting and Fine-tuning Large Language Models for Automated Code Review Comment Generation 15 Nov 2024 · 0 repositories · arXiv:2411.10129
-
Repurposing Stable Diffusion Attention for Training-Free Unsupervised Interactive Segmentation 15 Nov 2024 · 0 repositories · arXiv:2411.10411
-
RETR: Multi-View Radar Detection Transformer for Indoor Perception 15 Nov 2024 · 1 repository · arXiv:2411.10293Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 2 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 13 harvested samples) · 13 pointer-only (licence)
-
P² Law: Scaling Law for Post-Training After Model Pruning 15 Nov 2024 · 0 repositories · arXiv:2411.10272
-
Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs 15 Nov 2024 · 0 repositories · arXiv:2411.09968
-
SmoothCache: A Universal Inference Acceleration Technique for Diffusion Transformers 15 Nov 2024 · 1 repository · arXiv:2411.10510
-
SoftLMs: Efficient Adaptive Low-Rank Approximation of Language Models using Soft-Thresholding Mechanism 15 Nov 2024 · 0 repositories · arXiv:2411.10543
-
Take Package as Language: Anomaly Detection Using Transformer 15 Nov 2024 · 0 repositories · arXiv:2412.04473
-
ULTra: Unveiling Latent Token Interpretability in Transformer Based Understanding 15 Nov 2024 · 0 repositories · arXiv:2411.12589
-
Vision Eagle Attention: a new lens for advancing image classification 15 Nov 2024 · 1 repository · arXiv:2411.10564
-
Visual question answering based evaluation metrics for text-to-image generation 15 Nov 2024 · 0 repositories · arXiv:2411.10183
-
WavChat: A Survey of Spoken Dialogue Models 15 Nov 2024 · 1 repository · arXiv:2411.13577
-
A Centralized-Distributed Transfer Model for Cross-Domain Recommendation Based on Multi-Source Heterogeneous Transfer Learning 14 Nov 2024 · 0 repositories · arXiv:2411.09286
-
A Strategic Topology on Information Structures 14 Nov 2024 · 0 repositories · arXiv:2411.09149
-
Adopting RAG for LLM-Aided Future Vehicle Design 14 Nov 2024 · 0 repositories · arXiv:2411.09590
-
AI-driven inverse design of materials: Past, present and future 14 Nov 2024 · 0 repositories · arXiv:2411.09429
-
An Explainable Attention Model for Cervical Precancer Risk Classification using Colposcopic Images 14 Nov 2024 · 0 repositories · arXiv:2411.09469
-
Assessing the Performance of the DINOv2 Self-supervised Learning Vision Transformer Model for the Segmentation of the Left Atrium from MRI Images 14 Nov 2024 · 0 repositories · arXiv:2411.09598
-
Automating Autograding: Large Language Models as Test Suite Generators for Introductory Programming 14 Nov 2024 · 0 repositories · arXiv:2411.09261
-
BabyLM Challenge: Exploring the Effect of Variation Sets on Language Model Training Efficiency 14 Nov 2024 · 0 repositories · arXiv:2411.09587
-
Beyond Static Tools: Evaluating Large Language Models for Cryptographic Misuse Detection 14 Nov 2024 · 0 repositories · arXiv:2411.09772
-
Comprehensive and Practical Evaluation of Retrieval-Augmented Generation Systems for Medical Question Answering 14 Nov 2024 · 0 repositories · arXiv:2411.09213
-
Deep Learning for Fetal Inflammatory Response Diagnosis in the Umbilical Cord 14 Nov 2024 · 0 repositories · arXiv:2411.09767
-
DSCformer: A Dual-Branch Network Integrating Enhanced Dynamic Snake Convolution and SegFormer for Crack Segmentation 14 Nov 2024 · 0 repositories · arXiv:2411.09371
-
DT-JRD: Deep Transformer based Just Recognizable Difference Prediction Model for Video Coding for Machines 14 Nov 2024 · 0 repositories · arXiv:2411.09308
-
Evaluating Gender Bias in Large Language Models 14 Nov 2024 · 0 repositories · arXiv:2411.09826
-
GRAINRec: Graph and Attention Integrated Approach for Real-Time Session-Based Item Recommendations 14 Nov 2024 · 0 repositories · arXiv:2411.09152
-
Harnessing Vision Foundation Models for High-Performance, Training-Free Open Vocabulary Segmentation 14 Nov 2024 · 1 repository · arXiv:2411.09219
-
HateGPT: Unleashing GPT-3.5 Turbo to Combat Hate Speech on X 14 Nov 2024 · 0 repositories · arXiv:2411.09214
-
Heuristical Comparison of Vision Transformers Against Convolutional Neural Networks for Semantic Segmentation on Remote Sensing Imagery 14 Nov 2024 · 1 repository · arXiv:2411.09101
-
Initial Nugget Evaluation Results for the TREC 2024 RAG Track with the AutoNuggetizer Framework 14 Nov 2024 · 2 repositories · arXiv:2411.09607
-
Learning Parameter Sharing with Tensor Decompositions and Sparsity 14 Nov 2024 · 1 repository · arXiv:2411.09816