Methods › General › Attention Mechanisms › Attention › Papers, page 26
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 26 of 316: papers 2,501 to 2,600 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
LaViC: Adapting Large Vision-Language Models to Visually-Aware Conversational Recommendation 30 Mar 2025 · 1 repository · arXiv:2503.23312
-
Measuring Online Hate on 4chan using Pre-trained Deep Learning Models 30 Mar 2025 · 0 repositories · arXiv:2504.00045
-
Mixture of Routers 30 Mar 2025 · 0 repositories · arXiv:2503.23362
-
MoCha: Towards Movie-Grade Talking Character Synthesis 30 Mar 2025 · 0 repositories · arXiv:2503.23307
-
Multi-Dimensional AGV Path Planning in 3D Warehouses Using Ant Colony Optimization and Advanced Neural Networks 30 Mar 2025 · 0 repositories · arXiv:2504.01985
-
Multi-Stakeholder Disaster Insights from Social Media Using Large Language Models 30 Mar 2025 · 0 repositories · arXiv:2504.00046
-
Object Isolated Attention for Consistent Story Visualization 30 Mar 2025 · 0 repositories · arXiv:2503.23353
-
PromptDistill: Query-based Selective Token Retention in Intermediate Layers for Efficient Large Language Model Inference 30 Mar 2025 · 1 repository · arXiv:2503.23274
-
Question-Aware Knowledge Graph Prompting for Enhancing Large Language Models 30 Mar 2025 · 1 repository · arXiv:2503.23523
-
RARE: Retrieval-Augmented Reasoning Modeling 30 Mar 2025 · 1 repository · arXiv:2503.23513Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 0 honoured, 0 violated, 13 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 14 harvested samples)
-
Redundant feature screening method for human activity recognition based on attention purification mechanism 30 Mar 2025 · 0 repositories · arXiv:2503.23537
-
ReferDINO-Plus: 2nd Solution for 4th PVUW MeViS Challenge at CVPR 2025 30 Mar 2025 · 1 repository · arXiv:2503.23509
-
SCORE: Story Coherence and Retrieval Enhancement for AI Narratives 30 Mar 2025 · 0 repositories · arXiv:2503.23512
-
Semantic-Spatial Feature Fusion with Dynamic Graph Refinement for Remote Sensing Image Captioning 30 Mar 2025 · 0 repositories · arXiv:2503.23453
-
SketchVideo: Sketch-based Video Generation and Editing 30 Mar 2025 · 0 repositories · arXiv:2503.23284
-
Towards Physically Plausible Video Generation via VLM Planning 30 Mar 2025 · 0 repositories · arXiv:2503.23368
-
Understanding Visual Saliency of Outlier Items in Product Search 30 Mar 2025 · 0 repositories · arXiv:2503.23596
-
A Training-free LLM Framework with Interaction between Contextually Related Subtasks in Solving Complex Tasks 29 Mar 2025 · 0 repositories · arXiv:2503.23053
-
Efficient Adaptation For Remote Sensing Visual Grounding 29 Mar 2025 · 0 repositories · arXiv:2503.23083
-
Enhancing Knowledge Graph Completion with Entity Neighborhood and Relation Context 29 Mar 2025 · 0 repositories · arXiv:2503.23205
-
Large Self-Supervised Models Bridge the Gap in Domain Adaptive Object Detection 29 Mar 2025 · 1 repository · arXiv:2503.23220
-
Memory-Aware and Uncertainty-Guided Retrieval for Multi-Hop Question Answering 29 Mar 2025 · 0 repositories · arXiv:2503.23095
-
MHTS: Multi-Hop Tree Structure Framework for Generating Difficulty-Controllable QA Datasets for RAG Evaluation 29 Mar 2025 · 0 repositories · arXiv:2504.08756
-
Multimodal machine learning with large language embedding model for polymer property prediction 29 Mar 2025 · 1 repository · arXiv:2503.22962
-
ShiftLIC: Lightweight Learned Image Compression with Spatial-Channel Shift Operations 29 Mar 2025 · 1 repository · arXiv:2503.23052
-
The geomagnetic storm and Kp prediction using Wasserstein transformer 29 Mar 2025 · 0 repositories · arXiv:2503.23102
-
The realization of tones in spontaneous spoken Taiwan Mandarin: a corpus-based survey and theory-driven computational modeling 29 Mar 2025 · 0 repositories · arXiv:2503.23163
-
TRA: Better Length Generalisation with Threshold Relative Attention 29 Mar 2025 · 0 repositories · arXiv:2503.23174
-
UP-dROM : Uncertainty-Aware and Parametrised dynamic Reduced-Order Model, application to unsteady flows 29 Mar 2025 · 0 repositories · arXiv:2503.23236
-
Z-SASLM: Zero-Shot Style-Aligned SLI Blending Latent Manipulation 29 Mar 2025 · 1 repository · arXiv:2503.23234
-
A Refined Analysis of Massive Activations in LLMs 28 Mar 2025 · 1 repository · arXiv:2503.22329
-
A Survey of Circuit Foundation Model: Foundation AI Models for VLSI Circuit Design and EDA 28 Mar 2025 · 0 repositories · arXiv:2504.03711
-
An Advanced Ensemble Deep Learning Framework for Stock Price Prediction Using VAE, Transformer, and LSTM Model 28 Mar 2025 · 0 repositories · arXiv:2503.22192
-
AnnoPage Dataset: Dataset of Non-Textual Elements in Documents with Fine-Grained Categorization 28 Mar 2025 · 0 repositories · arXiv:2503.22526
-
Autonomous AI for Multi-Pathology Detection in Chest X-Rays: A Multi-Site Study in the Indian Healthcare System 28 Mar 2025 · 0 repositories · arXiv:2504.00022
-
Bridging the Dimensional Chasm: Uncover Layer-wise Dimensional Reduction in Transformers through Token Correlation 28 Mar 2025 · 0 repositories · arXiv:2503.22547
-
Camera Model Identification with SPAIR-Swin and Entropy based Non-Homogeneous Patches 28 Mar 2025 · 0 repositories · arXiv:2503.22120
-
Correlation-Attention Masked Temporal Transformer for User Identity Linkage Using Heterogeneous Mobility Data 28 Mar 2025 · 1 repository · arXiv:2504.01979
-
DeepOFormer: Deep Operator Learning with Domain-informed Features for Fatigue Life Prediction 28 Mar 2025 · 0 repositories · arXiv:2503.22475
-
DiTFastAttnV2: Head-wise Attention Compression for Multi-Modality Diffusion Transformers 28 Mar 2025 · 0 repositories · arXiv:2503.22796
-
DREMnet: An Interpretable Denoising Framework for Semi-Airborne Transient Electromagnetic Signal 28 Mar 2025 · 0 repositories · arXiv:2503.22223
-
DynaGraph: Interpretable Multi-Label Prediction from EHRs via Dynamic Graph Learning and Contrastive Augmentation 28 Mar 2025 · 0 repositories · arXiv:2503.22257
-
EdgeInfinite: A Memory-Efficient Infinite-Context Transformer for Edge Devices 28 Mar 2025 · 0 repositories · arXiv:2503.22196
-
Efficient Building Roof Type Classification: A Domain-Specific Self-Supervised Approach 28 Mar 2025 · 0 repositories · arXiv:2503.22251
-
Endo-TTAP: Robust Endoscopic Tissue Tracking via Multi-Facet Guided Attention and Hybrid Flow-point Supervision 28 Mar 2025 · 0 repositories · arXiv:2503.22394
-
Energy-efficient UAV movement and user-UAV association in multi-UAV networks 28 Mar 2025 · 0 repositories · arXiv:2503.22489
-
Enhancing Dance-to-Music Generation via Negative Conditioning Latent Diffusion Model 28 Mar 2025 · 0 repositories · arXiv:2503.22138
-
FLAM: Foundation Model-Based Body Stabilization for Humanoid Locomotion and Manipulation 28 Mar 2025 · 0 repositories · arXiv:2503.22249
-
Follow Your Motion: A Generic Temporal Consistency Portrait Editing Framework with Trajectory Guidance 28 Mar 2025 · 0 repositories · arXiv:2503.22225
-
Historical Ink: Exploring Large Language Models for Irony Detection in 19th-Century Spanish 28 Mar 2025 · 1 repository · arXiv:2503.22585
-
How Well Can Vison-Language Models Understand Humans' Intention? An Open-ended Theory of Mind Question Evaluation Benchmark 28 Mar 2025 · 0 repositories · arXiv:2503.22093
-
Hyperspectral Adapter for Object Tracking based on Hyperspectral Video 28 Mar 2025 · 0 repositories · arXiv:2503.22199
-
Integrating Artificial Intelligence with Human Expertise: An In-depth Analysis of ChatGPT's Capabilities in Generating Metamorphic Relations 28 Mar 2025 · 0 repositories · arXiv:2503.22141
-
Intrinsic Image Decomposition for Robust Self-supervised Monocular Depth Estimation on Reflective Surfaces 28 Mar 2025 · 0 repositories · arXiv:2503.22209
-
Leveraging LLMs for Predicting Unknown Diagnoses from Clinical Notes 28 Mar 2025 · 0 repositories · arXiv:2503.22092
-
MCRB for Parameter Estimation from One-Bit Quantized and Oversampled Measurements 28 Mar 2025 · 0 repositories · arXiv:2503.22860
-
Multimodal Machine Learning for Real Estate Appraisal: A Comprehensive Survey 28 Mar 2025 · 0 repositories · arXiv:2503.22119
-
SafeCast: Risk-Responsive Motion Forecasting for Autonomous Vehicles 28 Mar 2025 · 0 repositories · arXiv:2503.22541
-
Scenario Dreamer: Vectorized Latent Diffusion for Generating Driving Simulation Environments 28 Mar 2025 · 0 repositories · arXiv:2503.22496
-
Segment Any Motion in Videos 28 Mar 2025 · 0 repositories · arXiv:2503.22268
-
Spatial Transport Optimization by Repositioning Attention Map for Training-Free Text-to-Image Synthesis 28 Mar 2025 · 0 repositories · arXiv:2503.22168
-
Understanding Inequality of LLM Fact-Checking over Geographic Regions with Agent and Retrieval models 28 Mar 2025 · 0 repositories · arXiv:2503.22877
-
AdaMHF: Adaptive Multimodal Hierarchical Fusion for Survival Prediction 27 Mar 2025 · 1 repository · arXiv:2503.21124
-
AGILE: A Diffusion-Based Attention-Guided Image and Label Translation for Efficient Cross-Domain Plant Trait Identification 27 Mar 2025 · 1 repository · arXiv:2503.22019
-
An evaluation of LLMs and Google Translate for translation of selected Indian languages via sentiment and semantic analyses 27 Mar 2025 · 0 repositories · arXiv:2503.21393
-
An improved EfficientNetV2 for garbage classification 27 Mar 2025 · 0 repositories · arXiv:2503.21208
-
As easy as PIE: understanding when pruning causes language models to disagree 27 Mar 2025 · 1 repository · arXiv:2503.21714
-
Benchmarking Deep Learning-Based Methods for Irradiance Nowcasting with Sky Images 27 Mar 2025 · 0 repositories · arXiv:2503.21966
-
CMD-HAR: Cross-Modal Disentanglement for Wearable Human Activity Recognition 27 Mar 2025 · 0 repositories · arXiv:2503.21843
-
Collab: Controlled Decoding using Mixture of Agents for LLM Alignment 27 Mar 2025 · 0 repositories · arXiv:2503.21720
-
Combining Graph Attention Networks and Distributed Optimization for Multi-Robot Mixed-Integer Convex Programming 27 Mar 2025 · 0 repositories · arXiv:2503.21548
-
DSU-Net:An Improved U-Net Model Based on DINOv2 and SAM2 with Multi-scale Cross-model Feature Enhancement 27 Mar 2025 · 1 repository · arXiv:2503.21187
-
Dual-Splitting Conformal Prediction for Multi-Step Time Series Forecasting 27 Mar 2025 · 0 repositories · arXiv:2503.21251
-
Dynamic Asset Pricing Theory for Life Contingent Risks 27 Mar 2025 · 0 repositories · arXiv:2503.21256
-
Exploring the Evolution of Physics Cognition in Video Generation: A Survey 27 Mar 2025 · 1 repository · arXiv:2503.21765
-
Fine-Grained Behavior and Lane Constraints Guided Trajectory Prediction Method 27 Mar 2025 · 0 repositories · arXiv:2503.21477
-
From Individual to Group: Developing a Context-Aware Multi-Criteria Group Recommender System 27 Mar 2025 · 0 repositories · arXiv:2503.22752
-
HSLiNets: Evaluating Band Ordering Strategies in Hyperspectral and LiDAR Fusion 27 Mar 2025 · 1 repository · arXiv:2503.21072
-
Hybrid Emotion Recognition: Enhancing Customer Interactions Through Acoustic and Textual Analysis 27 Mar 2025 · 0 repositories · arXiv:2503.21927
-
HybridoNet-Adapt: A Domain-Adapted Framework for Accurate Lithium-Ion Battery RUL Prediction 27 Mar 2025 · 0 repositories · arXiv:2503.21392
-
HyperGraphRAG: Retrieval-Augmented Generation with Hypergraph-Structured Knowledge Representation 27 Mar 2025 · 1 repository · arXiv:2503.21322Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Integrating Travel Behavior Forecasting and Generative Modeling for Predicting Future Urban Mobility and Spatial Transformations 27 Mar 2025 · 0 repositories · arXiv:2503.21158
-
Learn by Reasoning: Analogical Weight Generation for Few-Shot Class-Incremental Learning 27 Mar 2025 · 0 repositories · arXiv:2503.21258
-
Local Normalization Distortion and the Thermodynamic Formalism of Decoding Strategies for Large Language Models 27 Mar 2025 · 0 repositories · arXiv:2503.21929
-
LOCATEdit: Graph Laplacian Optimized Cross Attention for Localized Text-Guided Image Editing 27 Mar 2025 · 1 repository · arXiv:2503.21541
-
MemInsight: Autonomous Memory Augmentation for LLM Agents 27 Mar 2025 · 0 repositories · arXiv:2503.21760
-
Molecular Quantum Transformer 27 Mar 2025 · 0 repositories · arXiv:2503.21686
-
NeuroLIP: Interpretable and Fair Cross-Modal Alignment of fMRI and Phenotypic Text 27 Mar 2025 · 0 repositories · arXiv:2503.21964
-
OccRobNet : Occlusion Robust Network for Accurate 3D Interacting Hand-Object Pose Estimation 27 Mar 2025 · 0 repositories · arXiv:2503.21723
-
OminiAdapt: Learning Cross-Task Invariance for Robust and Environment-Aware Robotic Manipulation 27 Mar 2025 · 0 repositories · arXiv:2503.21257
-
ProHOC: Probabilistic Hierarchical Out-of-Distribution Classification via Multi-Depth Networks 27 Mar 2025 · 1 repository · arXiv:2503.21397Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Q-MambaIR: Accurate Quantized Mamba for Efficient Image Restoration 27 Mar 2025 · 0 repositories · arXiv:2503.21970
-
Real-Time Evaluation Models for RAG: Who Detects Hallucinations Best? 27 Mar 2025 · 0 repositories · arXiv:2503.21157
-
ReaRAG: Knowledge-guided Reasoning Enhances Factuality of Large Reasoning Models with Iterative Retrieval Augmented Generation 27 Mar 2025 · 1 repository · arXiv:2503.21729
-
ReCoM: Realistic Co-Speech Motion Generation with Recurrent Embedded Transformer 27 Mar 2025 · 0 repositories · arXiv:2503.21847
-
Recurrent Feature Mining and Keypoint Mixup Padding for Category-Agnostic Pose Estimation 27 Mar 2025 · 1 repository · arXiv:2503.21140
-
Reinforced Model Merging 27 Mar 2025 · 1 repository · arXiv:2503.21272
-
Retinal Fundus Multi-Disease Image Classification using Hybrid CNN-Transformer-Ensemble Architectures 27 Mar 2025 · 1 repository · arXiv:2503.21465
-
ThinkEdit: Interpretable Weight Editing to Mitigate Overly Short Thinking in Reasoning Models 27 Mar 2025 · 1 repository · arXiv:2503.22048
-
Using large language models to produce literature reviews: Usages and systematic biases of microphysics parametrizations in 2699 publications 27 Mar 2025 · 0 repositories · arXiv:2503.21352