Methods › General › Attention Mechanisms › Attention › Papers, page 69
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 69 of 316: papers 6,801 to 6,900 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
AlignFormer: Modality Matching Can Achieve Better Zero-shot Instruction-Following Speech-LLM 2 Dec 2024 · 0 repositories · arXiv:2412.01145
-
Automated Extraction of Acronym-Expansion Pairs from Scientific Papers 2 Dec 2024 · 0 repositories · arXiv:2412.01093
-
Automated Toll Management System Using RFID and Image Processing 2 Dec 2024 · 0 repositories · arXiv:2412.01728
-
Class Distance Weighted Cross Entropy Loss for Classification of Disease Severity 2 Dec 2024 · 0 repositories · arXiv:2412.01246
-
Beyond Text-Visual Attention: Exploiting Visual Cues for Effective Token Pruning in VLMs 2 Dec 2024 · 2 repositories · arXiv:2412.01818Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 1 violated, 4 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples) · 1 pointer-only (licence)
-
Convolutional Transformer Neural Collaborative Filtering 2 Dec 2024 · 0 repositories · arXiv:2412.01376
-
CopyrightShield: Spatial Similarity Guided Backdoor Defense against Copyright Infringement in Diffusion Models 2 Dec 2024 · 0 repositories · arXiv:2412.01528
-
CPA: Camera-pose-awareness Diffusion Transformer for Video Generation 2 Dec 2024 · 0 repositories · arXiv:2412.01429
-
Cross-Modal Visual Relocalization in Prior LiDAR Maps Utilizing Intensity Textures 2 Dec 2024 · 0 repositories · arXiv:2412.01299
-
CSP-AIT-Net: A contrastive learning-enhanced spatiotemporal graph attention framework for short-term metro OD flow prediction with asynchronous inflow tracking 2 Dec 2024 · 0 repositories · arXiv:2412.01419
-
Data Uncertainty-Aware Learning for Multimodal Aspect-based Sentiment Analysis 2 Dec 2024 · 0 repositories · arXiv:2412.01249
-
Deep Learning Based Near-Field User Localization with Beam Squint in Wideband XL-MIMO Systems 2 Dec 2024 · 0 repositories · arXiv:2412.01029
-
Dual-Branch Graph Transformer Network for 3D Human Mesh Reconstruction from Video 2 Dec 2024 · 1 repository · arXiv:2412.01179
-
Efficient Semantic Communication Through Transformer-Aided Compression 2 Dec 2024 · 0 repositories · arXiv:2412.01817
-
Enhancing Crop Segmentation in Satellite Image Time Series with Transformer Networks 2 Dec 2024 · 0 repositories · arXiv:2412.01944
-
Epipolar Attention Field Transformers for Bird's Eye View Semantic Segmentation 2 Dec 2024 · 0 repositories · arXiv:2412.01595
-
FedPAW: Federated Learning with Personalized Aggregation Weights for Urban Vehicle Speed Prediction 2 Dec 2024 · 1 repository · arXiv:2412.01281
-
FGATT: A Robust Framework for Wireless Data Imputation Using Fuzzy Graph Attention Networks and Transformer Encoders 2 Dec 2024 · 0 repositories · arXiv:2412.01979
-
GETAE: Graph information Enhanced deep neural NeTwork ensemble ArchitecturE for fake news detection 2 Dec 2024 · 1 repository · arXiv:2412.01825
-
Global Average Feature Augmentation for Robust Semantic Segmentation with Transformers 2 Dec 2024 · 0 repositories · arXiv:2412.01941
-
HandOS: 3D Hand Reconstruction in One Stage 2 Dec 2024 · 0 repositories · arXiv:2412.01537
-
High-Throughput Detection of Risk Factors to Sudden Cardiac Arrest in Youth Athletes: A Smartwatch-Based Screening Platform 2 Dec 2024 · 0 repositories · arXiv:2412.12118
-
Identifying Reliable Predictions in Detection Transformers 2 Dec 2024 · 0 repositories · arXiv:2412.01782
-
InstantSwap: Fast Customized Concept Swapping across Sharp Shape Differences 2 Dec 2024 · 1 repository · arXiv:2412.01197
-
Learning Adaptive Lighting via Channel-Aware Guidance 2 Dec 2024 · 0 repositories · arXiv:2412.01493
-
Linear stimulus reconstruction works on the KU Leuven audiovisual, gaze-controlled auditory attention decoding dataset 2 Dec 2024 · 0 repositories · arXiv:2412.01401
-
LSceneLLM: Enhancing Large 3D Scene Understanding Using Adaptive Visual Preferences 2 Dec 2024 · 1 repository · arXiv:2412.01292Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 6 where Syntology's instrument failed) · 1 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
MBA-RAG: a Bandit Approach for Adaptive Retrieval-Augmented Generation through Question Complexity 2 Dec 2024 · 1 repository · arXiv:2412.01572Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
MFTF: Mask-free Training-free Object Level Layout Control Diffusion Model 2 Dec 2024 · 1 repository · arXiv:2412.01284
-
Multi-Agent Deep Reinforcement Learning for Distributed and Autonomous Platoon Coordination via Speed-regulation over Large-scale Transportation Networks 2 Dec 2024 · 0 repositories · arXiv:2412.01075
-
Multimodal Fusion Learning with Dual Attention for Medical Imaging 2 Dec 2024 · 1 repository · arXiv:2412.01248
-
MuSiCNet: A Gradual Coarse-to-Fine Framework for Irregularly Sampled Multivariate Time Series Analysis 2 Dec 2024 · 0 repositories · arXiv:2412.01063
-
Mutli-View 3D Reconstruction using Knowledge Distillation 2 Dec 2024 · 1 repository · arXiv:2412.02039
-
Neuron Abandoning Attention Flow: Visual Explanation of Dynamics inside CNN Models 2 Dec 2024 · 0 repositories · arXiv:2412.01202
-
NYT-Connections: A Deceptively Simple Text Classification Task that Stumps System-1 Thinkers 2 Dec 2024 · 0 repositories · arXiv:2412.01621
-
PainterNet: Adaptive Image Inpainting with Actual-Token Attention and Diverse Mask Control 2 Dec 2024 · 0 repositories · arXiv:2412.01223
-
Phaseformer: Phase-based Attention Mechanism for Underwater Image Restoration and Beyond 2 Dec 2024 · 1 repository · arXiv:2412.01456
-
PKRD-CoT: A Unified Chain-of-thought Prompting for Multi-Modal Large Language Models in Autonomous Driving 2 Dec 2024 · 0 repositories · arXiv:2412.02025
-
R-Bot: An LLM-based Query Rewrite System 2 Dec 2024 · 0 repositories · arXiv:2412.01661
-
ReHub: Linear Complexity Graph Transformers with Adaptive Hub-Spoke Reassignment 2 Dec 2024 · 0 repositories · arXiv:2412.01519
-
Research on Cervical Cancer p16/Ki-67 Immunohistochemical Dual-Staining Image Recognition Algorithm Based on YOLO 2 Dec 2024 · 0 repositories · arXiv:2412.01372
-
SEAL: Semantic Attention Learning for Long Video Representation 2 Dec 2024 · 0 repositories · arXiv:2412.01798
-
SiTSE: Sinhala Text Simplification Dataset and Evaluation 2 Dec 2024 · 1 repository · arXiv:2412.01293
-
Su-RoBERTa: A Semi-supervised Approach to Predicting Suicide Risk through Social Media using Base Language Models 2 Dec 2024 · 0 repositories · arXiv:2412.01353
-
Swin Transformer with Enhanced Dropout and Layer-wise Unfreezing for Facial Expression Recognition in Mental Health Detection 2 Dec 2024 · 1 repository
-
The Promise and Peril of Generative AI: Evidence from GPT-4 as Sell-Side Analysts 2 Dec 2024 · 0 repositories · arXiv:2412.01069
-
Tokenizing 3D Molecule Structure with Quantized Spherical Coordinates 2 Dec 2024 · 0 repositories · arXiv:2412.01564
-
VideoLights: Feature Refinement and Cross-Task Alignment Transformer for Joint Video Highlight Detection and Moment Retrieval 2 Dec 2024 · 1 repository · arXiv:2412.01558
-
A Comprehensive Guide to Explainable AI: From Classical Models to LLMs 1 Dec 2024 · 1 repository · arXiv:2412.00800
-
Advanced Video Inpainting Using Optical Flow-Guided Efficient Diffusion 1 Dec 2024 · 1 repository · arXiv:2412.00857
-
AniMer: Animal Pose and Shape Estimation Using Family Aware Transformer 1 Dec 2024 · 0 repositories · arXiv:2412.00837
-
Categorical Keypoint Positional Embedding for Robust Animal Re-Identification 1 Dec 2024 · 0 repositories · arXiv:2412.00818
-
Decision Transformer vs. Decision Mamba: Analysing the Complexity of Sequential Decision Making in Atari Games 1 Dec 2024 · 1 repository · arXiv:2412.00725
-
Deep Learning for Longitudinal Gross Tumor Volume Segmentation in MRI-Guided Adaptive Radiotherapy for Head and Neck Cancer 1 Dec 2024 · 1 repository · arXiv:2412.00663
-
DSSRNN: Decomposition-Enhanced State-Space Recurrent Neural Network for Time-Series Analysis 1 Dec 2024 · 1 repository · arXiv:2412.00994
-
DyMO: Training-Free Diffusion Model Alignment with Dynamic Multi-Objective Scheduling 1 Dec 2024 · 0 repositories · arXiv:2412.00759
-
EDTformer: An Efficient Decoder Transformer for Visual Place Recognition 1 Dec 2024 · 1 repository · arXiv:2412.00784Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples)
-
Effects of time aggregation, product aggregation, and seasonality in measuring bullwhip ratio 1 Dec 2024 · 0 repositories · arXiv:2412.00716
-
EventGPT: Event Stream Understanding with Multimodal Large Language Models 1 Dec 2024 · 0 repositories · arXiv:2412.00832
-
Learning to Forget using Hypernetworks 1 Dec 2024 · 0 repositories · arXiv:2412.00761
-
Lightweight Contenders: Navigating Semi-Supervised Text Mining through Peer Collaboration and Self Transcendence 1 Dec 2024 · 1 repository · arXiv:2412.00883
-
Long text outline generation: Chinese text outline based on unsupervised framework and large language mode 1 Dec 2024 · 0 repositories · arXiv:2412.00810
-
CONCERTO: Complex Query Execution Mechanism-Aware Learned Cost Estimation 1 Dec 2024 · 0 repositories · arXiv:2412.00749
-
MIMIC: Multimodal Islamophobic Meme Identification and Classification 1 Dec 2024 · 1 repository · arXiv:2412.00681
-
Motion-Aware Optical Camera Communication with Event Cameras 1 Dec 2024 · 2 repositories · arXiv:2412.00816
-
Precise Facial Landmark Detection by Dynamic Semantic Aggregation Transformer 1 Dec 2024 · 1 repository · arXiv:2412.00740
-
TGTOD: A Global Temporal Graph Transformer for Outlier Detection at Scale 1 Dec 2024 · 1 repository · arXiv:2412.00984
-
Visual Modality Prompt for Adapting Vision-Language Object Detectors 1 Dec 2024 · 1 repository · arXiv:2412.00622
-
A Self-Explainable Heterogeneous GNN for Relational Deep Learning 30 Nov 2024 · 1 repository · arXiv:2412.00521
-
Accelerating Multimodal Large Language Models by Searching Optimal Vision Token Reduction 30 Nov 2024 · 0 repositories · arXiv:2412.00556
-
BGM: Background Mixup for X-ray Prohibited Items Detection 30 Nov 2024 · 0 repositories · arXiv:2412.00460
-
CDEMapper: Enhancing NIH Common Data Element Normalization using Large Language Models 30 Nov 2024 · 0 repositories · arXiv:2412.00491
-
Cognitive Biases in Large Language Models: A Survey and Mitigation Experiments 30 Nov 2024 · 0 repositories · arXiv:2412.00323
-
Does Self-Attention Need Separate Weights in Transformers? 30 Nov 2024 · 0 repositories · arXiv:2412.00359
-
Dynamic Token Selection for Aerial-Ground Person Re-Identification 30 Nov 2024 · 0 repositories · arXiv:2412.00433
-
Empowering the Deaf and Hard of Hearing Community: Enhancing Video Captions Using Large Language Models 30 Nov 2024 · 0 repositories · arXiv:2412.00342
-
Fairness at Every Intersection: Uncovering and Mitigating Intersectional Biases in Multimodal Clinical Predictions 30 Nov 2024 · 0 repositories · arXiv:2412.00606
-
Forma mentis networks predict creativity ratings of short texts via interpretable artificial intelligence in human and GPT-simulated raters 30 Nov 2024 · 0 repositories · arXiv:2412.00530
-
Homeostasis and Sparsity in Transformer 30 Nov 2024 · 0 repositories · arXiv:2412.00503
-
HSLiNets: Hyperspectral Image and LiDAR Data Fusion Using Efficient Dual Non-Linear Feature Learning Networks 30 Nov 2024 · 0 repositories · arXiv:2412.00302
-
Multi-scale Feature Enhancement in Multi-task Learning for Medical Image Analysis 30 Nov 2024 · 0 repositories · arXiv:2412.00351
-
Pruned Convolutional Attention Network Based Wideband Spectrum Sensing with Sub-Nyquist Sampling 30 Nov 2024 · 1 repository · arXiv:2412.00562
-
Signal Processing over Time-Varying Graphs: A Systematic Review 30 Nov 2024 · 0 repositories · arXiv:2412.00462
-
Towards Fault Tolerance in Multi-Agent Reinforcement Learning 30 Nov 2024 · 1 repository · arXiv:2412.00534
-
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings 29 Nov 2024 · 1 repository · arXiv:2411.19628
-
Advanced System Integration: Analyzing OpenAPI Chunking for Retrieval-Augmented Generation 29 Nov 2024 · 0 repositories · arXiv:2411.19804
-
BatchLLM: Optimizing Large Batched LLM Inference with Global Prefix Sharing and Throughput-oriented Token Batching 29 Nov 2024 · 0 repositories · arXiv:2412.03594
-
Dynamic ETF Portfolio Optimization Using enhanced Transformer-Based Models for Covariance and Semi-Covariance Prediction(Work in Progress) 29 Nov 2024 · 0 repositories · arXiv:2411.19649
-
Dynamic Neural Curiosity Enhances Learning Flexibility for Autonomous Goal Discovery 29 Nov 2024 · 1 repository · arXiv:2412.00152
-
Excretion Detection in Pigsties Using Convolutional and Transformerbased Deep Neural Networks 29 Nov 2024 · 0 repositories · arXiv:2412.00256
-
Forecasting Foreign Exchange Market Prices Using Technical Indicators with Deep Learning and Attention Mechanism 29 Nov 2024 · 0 repositories · arXiv:2411.19763
-
Generating a Low-code Complete Workflow via Task Decomposition and RAG 29 Nov 2024 · 0 repositories · arXiv:2412.00239
-
Graph Neural Networks for Heart Failure Prediction on an EHR-Based Patient Similarity Graph 29 Nov 2024 · 1 repository · arXiv:2411.19742
-
HVAC-DPT: A Decision Pretrained Transformer for HVAC Control 29 Nov 2024 · 0 repositories · arXiv:2411.19746
-
Hyperspectral Images Efficient Spatial and Spectral non-Linear Model with Bidirectional Feature Learning 29 Nov 2024 · 0 repositories · arXiv:2412.00283
-
Interleaved-Modal Chain-of-Thought 29 Nov 2024 · 0 repositories · arXiv:2411.19488
-
Know Your RAG: Dataset Taxonomy and Generation Strategies for Evaluating RAG Systems 29 Nov 2024 · 0 repositories · arXiv:2411.19710
-
Knowledge Management for Automobile Failure Analysis Using Graph RAG 29 Nov 2024 · 0 repositories · arXiv:2411.19539
-
KV Shifting Attention Enhances Language Modeling 29 Nov 2024 · 1 repository · arXiv:2411.19574
-
LLM Teacher-Student Framework for Text Classification With No Manually Annotated Data: A Case Study in IPTC News Topic Classification 29 Nov 2024 · 1 repository · arXiv:2411.19638