Methods › General › Attention Mechanisms › Attention › Papers, page 21
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 21 of 316: papers 2,001 to 2,100 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Attention GhostUNet++: Enhanced Segmentation of Adipose Tissue and Liver in CT Images 14 Apr 2025 · 1 repository · arXiv:2504.11491
-
Beyond Chains of Thought: Benchmarking Latent-Space Reasoning Abilities in Large Language Models 14 Apr 2025 · 0 repositories · arXiv:2504.10615
-
Building Trustworthy Multimodal AI: A Review of Fairness, Transparency, and Ethics in Vision-Language Tasks 14 Apr 2025 · 0 repositories · arXiv:2504.13199
-
Can LLMs handle WebShell detection? Overcoming Detection Challenges with Behavioral Function-Aware Framework 14 Apr 2025 · 0 repositories · arXiv:2504.13811
-
COUNTS: Benchmarking Object Detectors and Multimodal Large Language Models under Distribution Shifts 14 Apr 2025 · 0 repositories · arXiv:2504.10158
-
Differentially Private 2D Human Pose Estimation 14 Apr 2025 · 0 repositories · arXiv:2504.10190
-
DiffMOD: Progressive Diffusion Point Denoising for Moving Object Detection in Remote Sensing 14 Apr 2025 · 0 repositories · arXiv:2504.10278
-
DioR: Adaptive Cognitive Detection and Contextual Retrieval Optimization for Dynamic Retrieval-Augmented Generation 14 Apr 2025 · 0 repositories · arXiv:2504.10198
-
DTFSal: Audio-Visual Dynamic Token Fusion for Video Saliency Prediction 14 Apr 2025 · 0 repositories · arXiv:2504.10070
-
EMAFusion: A Self-Optimizing System for Seamless LLM Selection and Integration 14 Apr 2025 · 0 repositories · arXiv:2504.10681
-
Enhancing Image Restoration through Learning Context-Rich and Detail-Accurate Features 14 Apr 2025 · 1 repository · arXiv:2504.10558
-
Epistemic Uncertainty-aware Recommendation Systems via Bayesian Deep Ensemble Learning 14 Apr 2025 · 0 repositories · arXiv:2504.10753
-
Global and Local Mamba Network for Multi-Modality Medical Image Super-Resolution 14 Apr 2025 · 0 repositories · arXiv:2504.10105
-
Hallucination Detection in LLMs via Topological Divergence on Attention Graphs 14 Apr 2025 · 0 repositories · arXiv:2504.10063
-
Hierarchical and Step-Layer-Wise Tuning of Attention Specialty for Multi-Instance Synthesis in Diffusion Transformers 14 Apr 2025 · 0 repositories · arXiv:2504.10148
-
Integrating Vision and Location with Transformers: A Multimodal Deep Learning Framework for Medical Wound Analysis 14 Apr 2025 · 0 repositories · arXiv:2504.10452
-
KeepKV: Eliminating Output Perturbation in KV Cache Compression for Efficient LLMs Inference 14 Apr 2025 · 0 repositories · arXiv:2504.09936
-
Keyword Extraction, and Aspect Classification in Sinhala, English, and Code-Mixed Content 14 Apr 2025 · 0 repositories · arXiv:2504.10679
-
Learning to Beamform for Cooperative Localization and Communication: A Link Heterogeneous GNN-Based Approach 14 Apr 2025 · 0 repositories · arXiv:2504.10060
-
LLM Can be a Dangerous Persuader: Empirical Study of Persuasion Safety in Large Language Models 14 Apr 2025 · 0 repositories · arXiv:2504.10430
-
MiMu: Mitigating Multiple Shortcut Learning Behavior of Transformers 14 Apr 2025 · 0 repositories · arXiv:2504.10551
-
MMKB-RAG: A Multi-Modal Knowledge-Based Retrieval-Augmented Generation Framework 14 Apr 2025 · 0 repositories · arXiv:2504.10074
-
Multi-Object Grounding via Hierarchical Contrastive Siamese Transformers 14 Apr 2025 · 0 repositories · arXiv:2504.10048
-
Multimodal Long Video Modeling Based on Temporal Dynamic Context 14 Apr 2025 · 1 repository · arXiv:2504.10443
-
OVERLORD: Ultimate Scaling of DataLoader for Multi-Source Large Foundation Model Training 14 Apr 2025 · 0 repositories · arXiv:2504.09844
-
Paging Dr. GPT: Extracting Information from Clinical Notes to Enhance Patient Predictions 14 Apr 2025 · 0 repositories · arXiv:2504.12338
-
Pay Attention to What and Where? Interpretable Feature Extractor in Vision-based Deep Reinforcement Learning 14 Apr 2025 · 1 repository · arXiv:2504.10071
-
RAKG:Document-level Retrieval Augmented Knowledge Graph Construction 14 Apr 2025 · 1 repository · arXiv:2504.09823
-
Relative Illumination Fields: Learning Medium and Light Independent Underwater Scenes 14 Apr 2025 · 0 repositories · arXiv:2504.10024
-
Self-Controlled Dynamic Expansion Model for Continual Learning 14 Apr 2025 · 0 repositories · arXiv:2504.10561
-
Separate to Collaborate: Dual-Stream Diffusion Model for Coordinated Piano Hand Motion Synthesis 14 Apr 2025 · 0 repositories · arXiv:2504.09885
-
Siamese Network with Dual Attention for EEG-Driven Social Learning: Bridging the Human-Robot Gap in Long-Tail Autonomous Driving 14 Apr 2025 · 0 repositories · arXiv:2504.10296
-
ST-Booster: An Iterative SpatioTemporal Perception Booster for Vision-and-Language Navigation in Continuous Environments 14 Apr 2025 · 0 repositories · arXiv:2504.09843
-
TAMP: Token-Adaptive Layerwise Pruning in Multimodal Large Language Models 14 Apr 2025 · 1 repository · arXiv:2504.09897
-
The Scalability of Simplicity: Empirical Analysis of Vision-Language Learning with a Single Transformer 14 Apr 2025 · 1 repository · arXiv:2504.10462Syntology official (archive's flag): 2 ran · 2 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 7 unverified; every one of the 2 samples that ran constructed an object rather than computing a result (of 9 harvested samples)
-
Toward Aligning Human and Robot Actions via Multi-Modal Demonstration Learning 14 Apr 2025 · 1 repository · arXiv:2504.11493
-
Understanding and Optimizing Multi-Stage AI Inference Pipelines 14 Apr 2025 · 0 repositories · arXiv:2504.09775
-
UP-Person: Unified Parameter-Efficient Transfer Learning for Text-based Person Retrieval 14 Apr 2025 · 1 repository · arXiv:2504.10084
-
VDocRAG: Retrieval-Augmented Generation over Visually-Rich Documents 14 Apr 2025 · 0 repositories · arXiv:2504.09795
-
Weight-of-Thought Reasoning: Exploring Neural Network Weights for Enhanced LLM Reasoning 14 Apr 2025 · 1 repository · arXiv:2504.10646
-
Will AI shape the way we speak? The emerging sociolinguistic influence of synthetic voices 14 Apr 2025 · 0 repositories · arXiv:2504.10650
-
XY-Cut++: Advanced Layout Ordering via Hierarchical Mask Mechanism on a Novel Benchmark 14 Apr 2025 · 1 repository · arXiv:2504.10258
-
Automatic Detection of Intro and Credits in Video using CLIP and Multihead Attention 13 Apr 2025 · 0 repositories · arXiv:2504.09738
-
ClinicalGPT-R1: Pushing reasoning capability of generalist disease diagnosis with large language model 13 Apr 2025 · 1 repository · arXiv:2504.09421
-
ControlNET: A Firewall for RAG-based LLM System 13 Apr 2025 · 0 repositories · arXiv:2504.09593
-
DiTSE: High-Fidelity Generative Speech Enhancement via Latent Diffusion Transformers 13 Apr 2025 · 0 repositories · arXiv:2504.09381
-
Enhanced Filterless Multi-Color VLC via QCT 13 Apr 2025 · 0 repositories · arXiv:2504.09743
-
Ensemble-Enhanced Graph Autoencoder with GAT and Transformer-Based Encoders for Robust Fault Diagnosis 13 Apr 2025 · 0 repositories · arXiv:2504.09427
-
Federated Prototype Graph Learning 13 Apr 2025 · 0 repositories · arXiv:2504.09493
-
HalluShift: Measuring Distribution Shifts towards Hallucination Detection in LLMs 13 Apr 2025 · 1 repository · arXiv:2504.09482Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
HD-RAG: Retrieval-Augmented Generation for Hybrid Documents Containing Text and Hierarchical Tables 13 Apr 2025 · 0 repositories · arXiv:2504.09554
-
HM-RAG: Hierarchical Multi-Agent Multimodal Retrieval Augmented Generation 13 Apr 2025 · 1 repository · arXiv:2504.12330
-
Imaging Transformer for MRI Denoising: a Scalable Model Architecture that enables SNR << 1 Imaging 13 Apr 2025 · 0 repositories · arXiv:2504.10534
-
Integrating Large Language Models for Automated Structural Analysis 13 Apr 2025 · 0 repositories · arXiv:2504.09754
-
Iterative Self-Training for Code Generation via Reinforced Re-Ranking 13 Apr 2025 · 0 repositories · arXiv:2504.09643
-
Mixture-of-Shape-Experts (MoSE): End-to-End Shape Dictionary Framework to Prompt SAM for Generalizable Medical Segmentation 13 Apr 2025 · 0 repositories · arXiv:2504.09601
-
Ordinary Least Squares as an Attention Mechanism 13 Apr 2025 · 0 repositories · arXiv:2504.09663
-
Question Tokens Deserve More Attention: Enhancing Large Language Models without Training through Step-by-Step Reading and Question Attention Recalibration 13 Apr 2025 · 0 repositories · arXiv:2504.09402
-
Sparse Deformable Mamba for Hyperspectral Image Classification 13 Apr 2025 · 0 repositories · arXiv:2504.09446
-
Spatially Directional Dual-Attention GAT for Spatial Fluoride Health Risk Modeling 13 Apr 2025 · 0 repositories · arXiv:2504.09416
-
Trajectory-guided Motion Perception for Facial Expression Quality Assessment in Neurological Disorders 13 Apr 2025 · 1 repository · arXiv:2504.09530
-
Uncertainty Guided Refinement for Fine-Grained Salient Object Detection 13 Apr 2025 · 1 repository · arXiv:2504.09666
-
Accurate Diagnosis of Respiratory Viruses Using an Explainable Machine Learning with Mid-Infrared Biomolecular Fingerprinting of Nasopharyngeal Secretions 12 Apr 2025 · 0 repositories · arXiv:2504.09211
-
AMNet: An Acoustic Model Network for Enhanced Mandarin Speech Synthesis 12 Apr 2025 · 0 repositories · arXiv:2504.09225
-
Beyond Degradation Conditions: All-in-One Image Restoration via HOG Transformers 12 Apr 2025 · 1 repository · arXiv:2504.09377
-
EchoMask: Speech-Queried Attention-based Mask Modeling for Holistic Co-Speech Motion Generation 12 Apr 2025 · 0 repositories · arXiv:2504.09209
-
Flux Already Knows -- Activating Subject-Driven Image Generation without Training 12 Apr 2025 · 1 repository · arXiv:2504.11478
-
Head-Aware KV Cache Compression for Efficient Visual Autoregressive Modeling 12 Apr 2025 · 0 repositories · arXiv:2504.09261
-
HeteRAG: A Heterogeneous Retrieval-augmented Generation Framework with Decoupled Knowledge Representations 12 Apr 2025 · 0 repositories · arXiv:2504.10529
-
Large Language Models as Particle Swarm Optimizers 12 Apr 2025 · 0 repositories · arXiv:2504.09247
-
Learning Occlusion-Robust Vision Transformers for Real-Time UAV Tracking 12 Apr 2025 · 1 repository · arXiv:2504.09228Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample)
-
Local Temporal Feature Enhanced Transformer with ROI-rank Based Masking for Diagnosis of ADHD 12 Apr 2025 · 0 repositories · arXiv:2504.11474
-
Lumos: Efficient Performance Modeling and Estimation for Large-scale LLM Training 12 Apr 2025 · 0 repositories · arXiv:2504.09307
-
MoE-Lens: Towards the Hardware Limit of High-Throughput MoE LLM Serving Under Resource Constraints 12 Apr 2025 · 0 repositories · arXiv:2504.09345
-
Multi-Modal Brain Tumor Segmentation via 3D Multi-Scale Self-attention and Cross-attention 12 Apr 2025 · 0 repositories · arXiv:2504.09088
-
Multi-scale Activation, Refinement, and Aggregation: Exploring Diverse Cues for Fine-Grained Bird Recognition 12 Apr 2025 · 0 repositories · arXiv:2504.09215
-
NetTAG: A Multimodal RTL-and-Layout-Aligned Netlist Foundation Model via Text-Attributed Graph 12 Apr 2025 · 1 repository · arXiv:2504.09260
-
PathSeqSAM: Sequential Modeling for Pathology Image Segmentation with SAM2 12 Apr 2025 · 0 repositories · arXiv:2504.10526
-
Pneuma: Leveraging LLMs for Tabular Data Representation and Retrieval in an End-to-End System 12 Apr 2025 · 1 repository · arXiv:2504.09207
-
Semantic Commit: Helping Users Update Intent Specifications for AI Memory at Scale 12 Apr 2025 · 0 repositories · arXiv:2504.09283
-
Stability Control of Metastable States as a Unified Mechanism for Flexible Temporal Modulation in Cognitive Processing 12 Apr 2025 · 0 repositories · arXiv:2504.09080
-
Adaptive Additive Parameter Updates of Vision Transformers for Few-Shot Continual Learning 11 Apr 2025 · 0 repositories · arXiv:2504.08982
-
Adopting Large Language Models to Automated System Integration 11 Apr 2025 · 0 repositories · arXiv:2504.08490
-
Annealed Mean Field Descent Is Highly Effective for Quadratic Unconstrained Binary Optimization 11 Apr 2025 · 0 repositories · arXiv:2504.08315
-
Application of machine learning models to predict the relationship between air pollution, ecosystem degradation, and health disparities and lung cancer in Vietnam 11 Apr 2025 · 0 repositories · arXiv:2504.08651
-
Artifact detection and localization in single-channel mobile EEG for sleep research using deep learning and attention mechanisms 11 Apr 2025 · 0 repositories · arXiv:2504.08469
-
Detecting Credit Card Fraud via Heterogeneous Graph Neural Networks with Graph Attention 11 Apr 2025 · 0 repositories · arXiv:2504.08183
-
Distributed Kalman Filter with Ultimately Accurate Fused Measurement Covariance 11 Apr 2025 · 0 repositories · arXiv:2504.08302
-
DreamFuse: Adaptive Image Fusion with Diffusion Transformer 11 Apr 2025 · 0 repositories · arXiv:2504.08291
-
DrivAer Transformer: A high-precision and fast prediction method for vehicle aerodynamic drag coefficient based on the DrivAerNet++ dataset 11 Apr 2025 · 0 repositories · arXiv:2504.08217
-
Examining GPT's Capability to Generate and Map Course Concepts and Their Relationship 11 Apr 2025 · 0 repositories · arXiv:2504.08856
-
HyperCore: The Core Framework for Building Hyperbolic Foundation Models with Comprehensive Modules 11 Apr 2025 · 1 repository · arXiv:2504.08912Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Hypergraph Vision Transformers: Images are More than Nodes, More than Edges 11 Apr 2025 · 0 repositories · arXiv:2504.08710
-
Integrated ensemble of BERT- and features-based models for authorship attribution in Japanese literary works 11 Apr 2025 · 0 repositories · arXiv:2504.08527
-
Jupiter: Fast and Resource-Efficient Collaborative Inference of Generative LLMs on Edge Devices 11 Apr 2025 · 0 repositories · arXiv:2504.08242
-
Learning from Elders: Making an LLM-powered Chatbot for Retirement Communities more Accessible through User-centered Design 11 Apr 2025 · 0 repositories · arXiv:2504.08985
-
LLM for Comparative Narrative Analysis 11 Apr 2025 · 0 repositories · arXiv:2504.08211
-
LLMTaxo: Leveraging Large Language Models for Constructing Taxonomy of Factual Claims from Social Media 11 Apr 2025 · 0 repositories · arXiv:2504.12325
-
Long Context In-Context Compression by Getting to the Gist of Gisting 11 Apr 2025 · 0 repositories · arXiv:2504.08934
-
Millions of States: Designing a Scalable MoE Architecture with RWKV-7 Meta-learner 11 Apr 2025 · 0 repositories · arXiv:2504.08247