Methods › General › Attention Mechanisms › Attention › Papers, page 93
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 93 of 316: papers 9,201 to 9,300 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Towards a Deeper Understanding of Transformer for Residential Non-intrusive Load Monitoring 2 Oct 2024 · 0 repositories · arXiv:2410.03758
-
Towards Dynamic Graph Neural Networks with Provably High-Order Expressive Power 2 Oct 2024 · 0 repositories · arXiv:2410.01367
-
Tracking objects that change in appearance with phase synchrony 2 Oct 2024 · 0 repositories · arXiv:2410.02094
-
UlcerGPT: A Multimodal Approach Leveraging Large Language and Vision Models for Diabetic Foot Ulcer Image Transcription 2 Oct 2024 · 0 repositories · arXiv:2410.01989
-
VectorGraphNET: Graph Attention Networks for Accurate Segmentation of Complex Technical Drawings 2 Oct 2024 · 0 repositories · arXiv:2410.01336
-
Why context matters in VQA and Reasoning: Semantic interventions for VLM input modalities 2 Oct 2024 · 0 repositories · arXiv:2410.01690
-
Addition is All You Need for Energy-efficient Language Models 1 Oct 2024 · 0 repositories · arXiv:2410.00907
-
Advanced Arabic Alphabet Sign Language Recognition Using Transfer Learning and Transformer Models 1 Oct 2024 · 0 repositories · arXiv:2410.00681
-
Advancing RVFL networks: Robust classification with the HawkEye loss function 1 Oct 2024 · 1 repository · arXiv:2410.00510
-
Unleashing the Unseen: Harnessing Benign Datasets for Jailbreaking Large Language Models 1 Oct 2024 · 1 repository · arXiv:2410.00451
-
AI Persuasion, Bayesian Attribution, and Career Concerns of Doctors 1 Oct 2024 · 0 repositories · arXiv:2410.01114
-
AlignSum: Data Pyramid Hierarchical Fine-tuning for Aligning with Human Summarization Preference 1 Oct 2024 · 1 repository · arXiv:2410.00409Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Explainable AI for Fraud Detection: An Attention-Based Ensemble of CNNs, GNNs, and A Confidence-Driven Gating Mechanism 1 Oct 2024 · 0 repositories · arXiv:2410.09069
-
BabelBench: An Omni Benchmark for Code-Driven Analysis of Multimodal and Multistructured Data 1 Oct 2024 · 1 repository · arXiv:2410.00773
-
Creative and Context-Aware Translation of East Asian Idioms with GPT-4 1 Oct 2024 · 1 repository · arXiv:2410.00988
-
Data-driven Framework for Forward and Inverse Problems in Guided Waves-Based Structural Health Monitoring Under Varying Environmental and Operating Conditions 1 Oct 2024 · 0 repositories · arXiv:2410.01127
-
Decoding Hate: Exploring Language Models' Reactions to Hate Speech 1 Oct 2024 · 0 repositories · arXiv:2410.00775
-
Deep Multimodal Fusion for Semantic Segmentation of Remote Sensing Earth Observation Data 1 Oct 2024 · 0 repositories · arXiv:2410.00469
-
Domain Aware Multi-Task Pretraining of 3D Swin Transformer for T1-weighted Brain MRI 1 Oct 2024 · 1 repository · arXiv:2410.00410
-
End-to-End Speech Recognition with Pre-trained Masked Language Model 1 Oct 2024 · 1 repository · arXiv:2410.00528
-
Exploring the Learning Capabilities of Language Models using LEVERWORLDS 1 Oct 2024 · 0 repositories · arXiv:2410.00519
-
Pediatric Wrist Fracture Detection Using Feature Context Excitation Modules in X-ray Images 1 Oct 2024 · 1 repository · arXiv:2410.01031
-
GLMHA A Guided Low-rank Multi-Head Self-Attention for Efficient Image Restoration and Spectral Reconstruction 1 Oct 2024 · 0 repositories · arXiv:2410.00380
-
GSPR: Multimodal Place Recognition Using 3D Gaussian Splatting for Autonomous Driving 1 Oct 2024 · 1 repository · arXiv:2410.00299
-
Insight: A Multi-Modal Diagnostic Pipeline using LLMs for Ocular Surface Disease Diagnosis 1 Oct 2024 · 0 repositories · arXiv:2410.00292
-
Language Enhanced Model for Eye (LEME): An Open-Source Ophthalmology-Specific Large Language Model 1 Oct 2024 · 0 repositories · arXiv:2410.03740
-
Learning Adaptive Hydrodynamic Models Using Neural ODEs in Complex Conditions 1 Oct 2024 · 0 repositories · arXiv:2410.00490
-
MAP: Unleashing Hybrid Mamba-Transformer Vision Backbone's Potential with Masked Autoregressive Pretraining 1 Oct 2024 · 0 repositories · arXiv:2410.00871
-
Multi-Scale Temporal Transformer For Speech Emotion Recognition 1 Oct 2024 · 0 repositories · arXiv:2410.00390
-
nGPT: Normalized Transformer with Representation Learning on the Hypersphere 1 Oct 2024 · 0 repositories · arXiv:2410.01131
-
Optimizing and Evaluating Enterprise Retrieval-Augmented Generation (RAG): A Content Design Perspective 1 Oct 2024 · 1 repository · arXiv:2410.12812
-
PclGPT: A Large Language Model for Patronizing and Condescending Language Detection 1 Oct 2024 · 1 repository · arXiv:2410.00361
-
Quantifying reliance on external information over parametric knowledge during Retrieval Augmented Generation (RAG) using mechanistic analysis 1 Oct 2024 · 0 repositories · arXiv:2410.00857
-
RATIONALYST: Pre-training Process-Supervision for Improving Reasoning 1 Oct 2024 · 1 repository · arXiv:2410.01044
-
Replacing Paths with Connection-Biased Attention for Knowledge Graph Completion 1 Oct 2024 · 1 repository · arXiv:2410.00876
-
Revisiting the Role of Texture in 3D Person Re-identification 1 Oct 2024 · 0 repositories · arXiv:2410.00348
-
Robust Traffic Forecasting against Spatial Shift over Years 1 Oct 2024 · 1 repository · arXiv:2410.00373Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 1 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 13 harvested samples)
-
Scene Graph Disentanglement and Composition for Generalizable Complex Image Generation 1 Oct 2024 · 0 repositories · arXiv:2410.00447
-
SCINet: Spatial and Contrast Interactive Super-Resolution Assisted Infrared UAV Target Detection 1 Oct 2024 · 2 repositories
-
Simplified priors for Object-Centric Learning 1 Oct 2024 · 0 repositories · arXiv:2410.00728
-
Sparse Attention Decomposition Applied to Circuit Tracing 1 Oct 2024 · 1 repository · arXiv:2410.00340
-
Spatial Action Unit Cues for Interpretable Deep Facial Expression Recognition 1 Oct 2024 · 1 repository · arXiv:2410.01848Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
STGformer: Efficient Spatiotemporal Graph Transformer for Traffic Forecasting 1 Oct 2024 · 1 repository · arXiv:2410.00385
-
Graph-Based Representation Learning of Neuronal Dynamics and Behavior 1 Oct 2024 · 1 repository · arXiv:2410.00665
-
TFCT-I2P: Three stream fusion network with color aware transformer for image-to-point cloud registration 1 Oct 2024 · 1 repository · arXiv:2410.00360
-
TPN: Transferable Proto-Learning Network towards Few-shot Document-Level Relation Extraction 1 Oct 2024 · 1 repository · arXiv:2410.00412
-
TransResNet: Integrating the Strengths of ViTs and CNNs for High Resolution Medical Image Segmentation via Feature Grafting 1 Oct 2024 · 1 repository · arXiv:2410.00986
-
Y-CA-Net: A Convolutional Attention Based Network for Volumetric Medical Image Segmentation 1 Oct 2024 · 0 repositories · arXiv:2410.01003
-
A Looming Replication Crisis in Evaluating Behavior in Language Models? Evidence and Solutions 30 Sep 2024 · 0 repositories · arXiv:2409.20303
-
A Methodology for Explainable Large Language Models with Integrated Gradients and Linguistic Analysis in Text Classification 30 Sep 2024 · 0 repositories · arXiv:2410.00250
-
ACE: All-round Creator and Editor Following Instructions via Diffusion Transformer 30 Sep 2024 · 0 repositories · arXiv:2410.00086
-
Adapting LLMs for the Medical Domain in Portuguese: A Study on Fine-Tuning and Model Evaluation 30 Sep 2024 · 0 repositories · arXiv:2410.00163
-
ASQuery: A Query-based Model for Action Segmentation 30 Sep 2024 · 1 repository
-
BSharedRAG: Backbone Shared Retrieval-Augmented Generation for the E-commerce Domain 30 Sep 2024 · 0 repositories · arXiv:2409.20075
-
CBAM-SwinT-BL: Small Rail Surface Defect Detection Method Based on Swin Transformer with Block Level CBAM Enhancement 30 Sep 2024 · 0 repositories · arXiv:2409.20113
-
Characterizing and Efficiently Accelerating Multimodal Generation Model Inference 30 Sep 2024 · 0 repositories · arXiv:2410.00215
-
CliMB: An AI-enabled Partner for Clinical Predictive Modeling 30 Sep 2024 · 1 repository · arXiv:2410.03736Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 10 harvested samples)
-
CLR-GAN: Improving GANs Stability and Quality via Consistent Latent Representation and Reconstruction 30 Sep 2024 · 1 repository
-
Contrastive Token Learning with Similarity Decay for Repetition Suppression in Machine Translation 30 Sep 2024 · 0 repositories · arXiv:2409.19877
-
Depression detection in social media posts using transformer-based models and auxiliary features 30 Sep 2024 · 0 repositories · arXiv:2409.20048
-
Enhancing Romanian Offensive Language Detection through Knowledge Distillation, Multi-Task Learning, and Data Augmentation 30 Sep 2024 · 0 repositories · arXiv:2409.20498
-
Evaluating the fairness of task-adaptive pretraining on unlabeled test data before few-shot text classification 30 Sep 2024 · 1 repository · arXiv:2410.00179
-
FreeMask: Rethinking the Importance of Attention Masks for Zero-Shot Video Editing 30 Sep 2024 · 0 repositories · arXiv:2409.20500
-
GTransPDM: A Graph-embedded Transformer with Positional Decoupling for Pedestrian Crossing Intention Prediction 30 Sep 2024 · 0 repositories · arXiv:2409.20223
-
HELPD: Mitigating Hallucination of LVLMs by Hierarchical Feedback Learning with Vision-enhanced Penalty Decoding 30 Sep 2024 · 1 repository · arXiv:2409.20429
-
ImmersePro: End-to-End Stereo Video Synthesis Via Implicit Disparity Learning 30 Sep 2024 · 1 repository · arXiv:2410.00262
-
Ingest-And-Ground: Dispelling Hallucinations from Continually-Pretrained LLMs with RAG 30 Sep 2024 · 0 repositories · arXiv:2410.02825
-
Is Preference Alignment Always the Best Option to Enhance LLM-Based Translation? An Empirical Analysis 30 Sep 2024 · 0 repositories · arXiv:2409.20059
-
KV-Compress: Paged KV-Cache Compression with Variable Compression Rates per Attention Head 30 Sep 2024 · 1 repository · arXiv:2410.00161Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Learning Multimodal Latent Generative Models with Energy-Based Prior 30 Sep 2024 · 1 repository · arXiv:2409.19862
-
Maia-2: A Unified Model for Human-AI Alignment in Chess 30 Sep 2024 · 2 repositories · arXiv:2409.20553Syntology official (archive's flag): 4 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 10 unverified (of 18 harvested samples) · 11 pointer-only (licence)
-
MaskMamba: A Hybrid Mamba-Transformer Model for Masked Image Generation 30 Sep 2024 · 0 repositories · arXiv:2409.19937
-
Mechanism Design with Endogenous Perception 30 Sep 2024 · 0 repositories · arXiv:2409.19853
-
Modelando procesos cognitivos de la lectura natural con GPT-2 30 Sep 2024 · 0 repositories · arXiv:2409.20174
-
Numerically Robust Fixed-Point Smoothing Without State Augmentation 30 Sep 2024 · 1 repository · arXiv:2409.20004
-
Exploring Social Media Image Categorization Using Large Models with Different Adaptation Methods: A Case Study on Cultural Nature's Contributions to People 30 Sep 2024 · 0 repositories · arXiv:2410.00275
-
On The Planning Abilities of OpenAI's o1 Models: Feasibility, Optimality, and Generalizability 30 Sep 2024 · 2 repositories · arXiv:2409.19924
-
QAEncoder: Towards Aligned Representation Learning in Question Answering System 30 Sep 2024 · 1 repository · arXiv:2409.20434
-
Social Conjuring: Multi-User Runtime Collaboration with AI in Building Virtual 3D Worlds 30 Sep 2024 · 0 repositories · arXiv:2410.00274
-
SWIM: Short-Window CNN Integrated with Mamba for EEG-Based Auditory Spatial Attention Decoding 30 Sep 2024 · 1 repository · arXiv:2409.19884
-
Systemic Risk Asymptotics in a Renewal Model with Multiple Business Lines and Heterogeneous Claims 30 Sep 2024 · 0 repositories · arXiv:2410.00158
-
T-KAER: Transparency-enhanced Knowledge-Augmented Entity Resolution Framework 30 Sep 2024 · 1 repository · arXiv:2410.00218
-
The age of spiritual machines: Language quietus induces synthetic altered states of consciousness in artificial intelligence 30 Sep 2024 · 0 repositories · arXiv:2410.00257
-
Towards Open-Vocabulary Semantic Segmentation Without Semantic Labels 30 Sep 2024 · 0 repositories · arXiv:2409.19846
-
Whole-Graph Representation Learning For the Classification of Signed Networks 30 Sep 2024 · 1 repository · arXiv:2409.20073
-
2D-TPE: Two-Dimensional Positional Encoding Enhances Table Understanding for Large Language Models 29 Sep 2024 · 1 repository · arXiv:2409.19700
-
A multimodal LLM for the non-invasive decoding of spoken text from brain recordings 29 Sep 2024 · 0 repositories · arXiv:2409.19710
-
Abstractive Summarization of Low resourced Nepali language using Multilingual Transformers 29 Sep 2024 · 0 repositories · arXiv:2409.19566
-
Adversarial Examples for DNA Classification 29 Sep 2024 · 0 repositories · arXiv:2409.19788
-
Black-Box Segmentation of Electronic Medical Records 29 Sep 2024 · 0 repositories · arXiv:2409.19796
-
Can Models Learn Skill Composition from Examples? 29 Sep 2024 · 0 repositories · arXiv:2409.19808
-
Causal Deciphering and Inpainting in Spatio-Temporal Dynamics via Diffusion Model 29 Sep 2024 · 0 repositories · arXiv:2409.19608
-
CRScore: Grounding Automated Evaluation of Code Review Comments in Code Claims and Smells 29 Sep 2024 · 0 repositories · arXiv:2409.19801
-
Differentially Private Bilevel Optimization 29 Sep 2024 · 0 repositories · arXiv:2409.19800
-
DIIT: A Domain-Invariant Information Transfer Method for Industrial Cross-Domain Recommendation 29 Sep 2024 · 0 repositories · arXiv:2410.10835
-
Discerning the Chaos: Detecting Adversarial Perturbations while Disentangling Intentional from Unintentional Noises 29 Sep 2024 · 0 repositories · arXiv:2409.19619
-
Does RAG Introduce Unfairness in LLMs? Evaluating Fairness in Retrieval-Augmented Generation Systems 29 Sep 2024 · 1 repository · arXiv:2409.19804Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
Dual-Attention Frequency Fusion at Multi-Scale for Joint Segmentation and Deformable Medical Image Registration 29 Sep 2024 · 0 repositories · arXiv:2409.19658
-
Federated Learning from Vision-Language Foundation Models: Theoretical Analysis and Method 29 Sep 2024 · 1 repository · arXiv:2409.19610Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 4 where Syntology's instrument failed) · 3 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
Flipped Classroom: Aligning Teacher Attention with Student in Generalized Category Discovery 29 Sep 2024 · 0 repositories · arXiv:2409.19659