Methods › General › Attention Mechanisms › Attention › Papers, page 111
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 111 of 316: papers 11,001 to 11,100 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
JARViS: Detecting Actions in Video Using Unified Actor-Scene Context Relation Modeling 7 Aug 2024 · 0 repositories · arXiv:2408.03612
-
Query3D: LLM-Powered Open-Vocabulary Scene Segmentation with Language Embedded 3D Gaussian 7 Aug 2024 · 1 repository · arXiv:2408.03516
-
MaxMind: A Memory Loop Network to Enhance Software Productivity based on Large Language Models 7 Aug 2024 · 0 repositories · arXiv:2408.03841
-
NACL: A General and Effective KV Cache Eviction Framework for LLMs at Inference Time 7 Aug 2024 · 2 repositories · arXiv:2408.03675Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
PackMamba: Efficient Processing of Variable-Length Sequences in Mamba training 7 Aug 2024 · 0 repositories · arXiv:2408.03865
-
PaveCap: The First Multimodal Framework for Comprehensive Pavement Condition Assessment with Dense Captioning and PCI Estimation 7 Aug 2024 · 1 repository · arXiv:2408.04110
-
Pick of the Bunch: Detecting Infrared Small Targets Beyond Hit-Miss Trade-Offs via Selective Rank-Aware Attention 7 Aug 2024 · 1 repository · arXiv:2408.03717Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
PRISM: PRogressive dependency maxImization for Scale-invariant image Matching 7 Aug 2024 · 0 repositories · arXiv:2408.03598
-
RailTrack-DaViT: A Vision Transformer-Based Approach for Automated Railway Track Defect Detection 7 Aug 2024 · 1 repository
-
Retrieval Augmentation via User Interest Clustering 7 Aug 2024 · 0 repositories · arXiv:2408.03886
-
Path-SAM2: Transfer SAM2 for digital pathology semantic segmentation 7 Aug 2024 · 1 repository · arXiv:2408.03651
-
SocFedGPT: Federated GPT-based Adaptive Content Filtering System Leveraging User Interactions in Social Networks 7 Aug 2024 · 0 repositories · arXiv:2408.05243
-
Soft-Hard Attention U-Net Model and Benchmark Dataset for Multiscale Image Shadow Removal 7 Aug 2024 · 0 repositories · arXiv:2408.03734
-
Surgformer: Surgical Transformer with Hierarchical Temporal Attention for Surgical Phase Recognition 7 Aug 2024 · 1 repository · arXiv:2408.03867
-
SwinShadow: Shifted Window for Ambiguous Adjacent Shadow Detection 7 Aug 2024 · 1 repository · arXiv:2408.03521
-
TALE: Training-free Cross-domain Image Composition via Adaptive Latent Manipulation and Energy-guided Optimization 7 Aug 2024 · 0 repositories · arXiv:2408.03637
-
Time is Not Enough: Time-Frequency based Explanation for Time-Series Black-Box Models 7 Aug 2024 · 1 repository · arXiv:2408.03636
-
Tree Attention: Topology-aware Decoding for Long-Context Attention on GPU clusters 7 Aug 2024 · 1 repository · arXiv:2408.04093Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Trustworthy Image Semantic Communication with GenAI: Explainablity, Controllability, and Efficiency 7 Aug 2024 · 0 repositories · arXiv:2408.03806
-
FDC: Fast KV Dimensionality Compression for Efficient LLM Inference 7 Aug 2024 · 0 repositories · arXiv:2408.04107
-
VizECGNet: Visual ECG Image Network for Cardiovascular Diseases Classification with Multi-Modal Training and Knowledge Distillation 6 Aug 2024 · 0 repositories · arXiv:2408.02888
-
Intermediate direct preference optimization 6 Aug 2024 · 0 repositories · arXiv:2408.02923
-
Data Poisoning in LLMs: Jailbreak-Tuning and Scaling Laws 6 Aug 2024 · 2 repositories · arXiv:2408.02946Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
EC-Guide: A Comprehensive E-Commerce Guide for Instruction Tuning and Quantization 6 Aug 2024 · 1 repository · arXiv:2408.02970
-
Empathy Level Alignment via Reinforcement Learning for Empathetic Response Generation 6 Aug 2024 · 1 repository · arXiv:2408.02976
-
Nighttime Pedestrian Detection Based on Fore-Background Contrast Learning 6 Aug 2024 · 0 repositories · arXiv:2408.03030
-
Training-Free Condition Video Diffusion Models for single frame Spatial-Semantic Echocardiogram Synthesis 6 Aug 2024 · 1 repository · arXiv:2408.03035
-
Analysis of Argument Structure Constructions in a Deep Recurrent Language Model 6 Aug 2024 · 0 repositories · arXiv:2408.03062
-
SCOPE: A Synthetic Multi-Modal Dataset for Collective Perception Including Physical-Correct Weather Conditions 6 Aug 2024 · 0 repositories · arXiv:2408.03065
-
QADQN: Quantum Attention Deep Q-Network for Financial Market Prediction 6 Aug 2024 · 0 repositories · arXiv:2408.03088
-
Topic Modeling with Fine-tuning LLMs and Bag of Sentences 6 Aug 2024 · 1 repository · arXiv:2408.03099Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Evaluating the Translation Performance of Large Language Models Based on Euas-20 6 Aug 2024 · 0 repositories · arXiv:2408.03119
-
Leveraging Entity Information for Cross-Modality Correlation Learning: The Entity-Guided Multimodal Summarization 6 Aug 2024 · 1 repository · arXiv:2408.03149
-
Iterative CT Reconstruction via Latent Variable Optimization of Shallow Diffusion Models 6 Aug 2024 · 0 repositories · arXiv:2408.03156
-
Leveraging Parameter Efficient Training Methods for Low Resource Text Classification: A Case Study in Marathi 6 Aug 2024 · 0 repositories · arXiv:2408.03172
-
SGSR: Structure-Guided Multi-Contrast MRI Super-Resolution via Spatio-Frequency Co-Query Attention 6 Aug 2024 · 0 repositories · arXiv:2408.03194
-
Learning to Learn without Forgetting using Attention 6 Aug 2024 · 1 repository · arXiv:2408.03219
-
LAC-Net: Linear-Fusion Attention-Guided Convolutional Network for Accurate Robotic Grasping Under the Occlusion 6 Aug 2024 · 0 repositories · arXiv:2408.03238
-
ReSyncer: Rewiring Style-based Generator for Unified Audio-Visually Synced Facial Performer 6 Aug 2024 · 0 repositories · arXiv:2408.03284
-
DopQ-ViT: Towards Distribution-Friendly and Outlier-Aware Post-Training Quantization for Vision Transformers 6 Aug 2024 · 0 repositories · arXiv:2408.03291
-
MDT-A2G: Exploring Masked Diffusion Transformers for Co-Speech Gesture Generation 6 Aug 2024 · 0 repositories · arXiv:2408.03312
-
Advancing EEG-Based Gaze Prediction Using Depthwise Separable Convolution and Enhanced Pre-Processing 6 Aug 2024 · 1 repository · arXiv:2408.03480
-
Can LLMs Serve As Time Series Anomaly Detectors? 6 Aug 2024 · 0 repositories · arXiv:2408.03475
-
FDA Jamming Against Airborne Phased-MIMO Radar-Part I: Matched Filtering and Spatial Filtering 6 Aug 2024 · 0 repositories · arXiv:2408.03050
-
FLASH: Federated Learning-Based LLMs for Advanced Query Processing in Social Networks through RAG 6 Aug 2024 · 0 repositories · arXiv:2408.05242
-
LLM-Aided Compilation for Tensor Accelerators 6 Aug 2024 · 0 repositories · arXiv:2408.03408
-
LLM-based MOFs Synthesis Condition Extraction using Few-Shot Demonstrations 6 Aug 2024 · 0 repositories · arXiv:2408.04665
-
Set2Seq Transformer: Learning Permutation Aware Set Representations of Artistic Sequences 6 Aug 2024 · 0 repositories · arXiv:2408.03404
-
Static IR Drop Prediction with Attention U-Net and Saliency-Based Explainability 6 Aug 2024 · 0 repositories · arXiv:2408.03292
-
TF-Locoformer: Transformer with Local Modeling by Convolution for Speech Separation and Enhancement 6 Aug 2024 · 1 repository · arXiv:2408.03440
-
The Use of Large Language Models (LLM) for Cyber Threat Intelligence (CTI) in Cybercrime Forums 6 Aug 2024 · 0 repositories · arXiv:2408.03354
-
TrafficGPT: An LLM Approach for Open-Set Encrypted Traffic Classification 6 Aug 2024 · 1 repository
-
Training LLMs to Recognize Hedges in Spontaneous Narratives 6 Aug 2024 · 1 repository · arXiv:2408.03319
-
ULLME: A Unified Framework for Large Language Model Embeddings with Generation-Augmented Learning 6 Aug 2024 · 1 repository · arXiv:2408.03402Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
AssemAI: Interpretable Image-Based Anomaly Detection for Manufacturing Pipelines 5 Aug 2024 · 1 repository · arXiv:2408.02181
-
Is Large Language Model Good at Database Knob Tuning? A Comprehensive Experimental Evaluation 5 Aug 2024 · 0 repositories · arXiv:2408.02213
-
Cross-modulated Attention Transformer for RGBT Tracking 5 Aug 2024 · 0 repositories · arXiv:2408.02222
-
Do Large Language Models Speak All Languages Equally? A Comparative Study in Low-Resource Settings 5 Aug 2024 · 0 repositories · arXiv:2408.02237
-
COM Kitchens: An Unedited Overhead-view Video Dataset as a Vision-Language Benchmark 5 Aug 2024 · 1 repository · arXiv:2408.02272
-
DRFormer: Multi-Scale Transformer Utilizing Diverse Receptive Fields for Long Time-Series Forecasting 5 Aug 2024 · 1 repository · arXiv:2408.02279
-
Spin glass model of in-context learning 5 Aug 2024 · 0 repositories · arXiv:2408.02288
-
The NPU-ASLP System Description for Visual Speech Recognition in CNVSRC 2024 5 Aug 2024 · 1 repository · arXiv:2408.02369
-
Why Are My Prompts Leaked? Unraveling Prompt Extraction Threats in Customized Large Language Models 5 Aug 2024 · 1 repository · arXiv:2408.02416
-
Enhancing Heterogeneous Knowledge Graph Completion with a Novel GAT-based Approach 5 Aug 2024 · 0 repositories · arXiv:2408.02456
-
An investigation into the causes of race bias in AI-based cine CMR segmentation 5 Aug 2024 · 0 repositories · arXiv:2408.02462
-
RAG Foundry: A Framework for Enhancing LLMs for Retrieval Augmented Generation 5 Aug 2024 · 2 repositories · arXiv:2408.02545
-
On Using Quasirandom Sequences in Machine Learning for Model Weight Initialization 5 Aug 2024 · 1 repository · arXiv:2408.02654
-
PSNE: Efficient Spectral Sparsification Algorithms for Scaling Network Embedding 5 Aug 2024 · 0 repositories · arXiv:2408.02705
-
A Novel Hybrid Approach for Tornado Prediction in the United States: Kalman-Convolutional BiLSTM with Multi-Head Attention 5 Aug 2024 · 0 repositories · arXiv:2408.02751
-
Dimensionality Reduction and Nearest Neighbors for Improving Out-of-Distribution Detection in Medical Image Segmentation 5 Aug 2024 · 1 repository · arXiv:2408.02761
-
From Recognition to Prediction: Leveraging Sequence Reasoning for Action Anticipation 5 Aug 2024 · 1 repository · arXiv:2408.02769
-
LR-Net: A Lightweight and Robust Network for Infrared Small Target Detection 5 Aug 2024 · 0 repositories · arXiv:2408.02780
-
GazeXplain: Learning to Predict Natural Language Explanations of Visual Scanpaths 5 Aug 2024 · 0 repositories · arXiv:2408.02788
-
Heterogeneous graph attention network improves cancer multiomics integration 5 Aug 2024 · 1 repository · arXiv:2408.02845
-
Wiping out the limitations of Large Language Models -- A Taxonomy for Retrieval Augmented Generation 5 Aug 2024 · 0 repositories · arXiv:2408.02854
-
AppAgent v2: Advanced Agent for Flexible Mobile Interactions 5 Aug 2024 · 0 repositories · arXiv:2408.11824
-
LLM Agents Improve Semantic Code Search 5 Aug 2024 · 0 repositories · arXiv:2408.11058
-
Modelling Visual Semantics via Image Captioning to extract Enhanced Multi-Level Cross-Modal Semantic Incongruity Representation with Attention for Multimodal Sarcasm Detection 5 Aug 2024 · 0 repositories · arXiv:2408.02595
-
SEAS: Self-Evolving Adversarial Safety Optimization for Large Language Models 5 Aug 2024 · 1 repository · arXiv:2408.02632
-
Self-Taught Evaluators 5 Aug 2024 · 0 repositories · arXiv:2408.02666
-
The Mechanics of Conceptual Interpretation in GPT Models: Interpretative Insights 5 Aug 2024 · 0 repositories · arXiv:2408.11827
-
Toward Attention-based TinyML: A Heterogeneous Accelerated Architecture and Automated Deployment Flow 5 Aug 2024 · 1 repository · arXiv:2408.02473
-
VyAnG-Net: A Novel Multi-Modal Sarcasm Recognition Model by Uncovering Visual, Acoustic and Glossary Features 5 Aug 2024 · 0 repositories · arXiv:2408.10246
-
XMainframe: A Large Language Model for Mainframe Modernization 5 Aug 2024 · 1 repository · arXiv:2408.04660
-
Cross-layer Attention Sharing for Large Language Models 4 Aug 2024 · 0 repositories · arXiv:2408.01890
-
CAF-YOLO: A Robust Framework for Multi-Scale Lesion Detection in Biomedical Imagery 4 Aug 2024 · 1 repository · arXiv:2408.01897
-
Advancing H&E-to-IHC Stain Translation in Breast Cancer: A Multi-Magnification and Attention-Based Approach 4 Aug 2024 · 0 repositories · arXiv:2408.01929
-
DiReCT: Diagnostic Reasoning for Clinical Notes via Large Language Models 4 Aug 2024 · 1 repository · arXiv:2408.01933Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 1 violated, 8 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 13 harvested samples) · 13 pointer-only (licence)
-
Defining and Evaluating Decision and Composite Risk in Language Models Applied to Natural Language Inference 4 Aug 2024 · 0 repositories · arXiv:2408.01935
-
CACE-Net: Co-guidance Attention and Contrastive Enhancement for Effective Audio-Visual Event Localization 4 Aug 2024 · 1 repository · arXiv:2408.01952Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples) · 1 pointer-only (licence)
-
ML-EAT: A Multilevel Embedding Association Test for Interpretable and Transparent Social Science 4 Aug 2024 · 1 repository · arXiv:2408.01966
-
AdaCBM: An Adaptive Concept Bottleneck Model for Explainable and Accurate Diagnosis 4 Aug 2024 · 1 repository · arXiv:2408.02001
-
Mini-Monkey: Alleviating the Semantic Sawtooth Effect for Lightweight MLLMs via Complementary Image Pyramid 4 Aug 2024 · 1 repository · arXiv:2408.02034Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Pixel-Level Domain Adaptation: A New Perspective for Enhancing Weakly Supervised Semantic Segmentation 4 Aug 2024 · 1 repository · arXiv:2408.02039
-
MedSyn: LLM-based Synthetic Medical Text Generation Framework 4 Aug 2024 · 1 repository · arXiv:2408.02056
-
KAN-RCBEVDepth: A multi-modal fusion algorithm in object detection for autonomous driving 4 Aug 2024 · 0 repositories · arXiv:2408.02088
-
Effective Demonstration Annotation for In-Context Learning via Language Model-Based Determinantal Point Process 4 Aug 2024 · 0 repositories · arXiv:2408.02103
-
Leveraging Large Language Models with Chain-of-Thought and Prompt Engineering for Traffic Crash Severity Analysis and Inference 4 Aug 2024 · 0 repositories · arXiv:2408.04652
-
Mamba-Spike: Enhancing the Mamba Architecture with a Spiking Front-End for Efficient Temporal Data Processing 4 Aug 2024 · 1 repository · arXiv:2408.11823Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Advancing Mental Health Pre-Screening: A New Custom GPT for Psychological Distress Assessment 3 Aug 2024 · 0 repositories · arXiv:2408.01614