Methods › General › Attention Mechanisms › Attention › Papers, page 34
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 34 of 316: papers 3,301 to 3,400 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
FEDS: Feature and Entropy-Based Distillation Strategy for Efficient Learned Image Compression 9 Mar 2025 · 0 repositories · arXiv:2503.06399
-
Fine-Grained Alignment and Noise Refinement for Compositional Text-to-Image Generation 9 Mar 2025 · 1 repository · arXiv:2503.06506
-
Global-Aware Monocular Semantic Scene Completion with State Space Models 9 Mar 2025 · 0 repositories · arXiv:2503.06569
-
GroMo: Plant Growth Modeling with Multiview Images 9 Mar 2025 · 1 repository · arXiv:2503.06608
-
Heterogeneous bimodal attention fusion for speech emotion recognition 9 Mar 2025 · 0 repositories · arXiv:2503.06405
-
Human Cognition Inspired RAG with Knowledge Graph for Complex Problem Solving 9 Mar 2025 · 0 repositories · arXiv:2503.06567
-
Intelligent Control of Merging Car-following and Lane-Changing Behavior 9 Mar 2025 · 0 repositories · arXiv:2503.06572
-
LSA: Latent Style Augmentation Towards Stain-Agnostic Cervical Cancer Screening 9 Mar 2025 · 0 repositories · arXiv:2503.06563
-
Multimodal Emotion Recognition and Sentiment Analysis in Multi-Party Conversation Contexts 9 Mar 2025 · 0 repositories · arXiv:2503.06805
-
Pre-Training Meta-Rule Selection Policy for Visual Generative Abductive Learning 9 Mar 2025 · 1 repository · arXiv:2503.06427
-
Probabilistic Shielding for Safe Reinforcement Learning 9 Mar 2025 · 0 repositories · arXiv:2503.07671
-
SAQ-SAM: Semantically-Aligned Quantization for Segment Anything Model 9 Mar 2025 · 0 repositories · arXiv:2503.06515
-
Seeing Delta Parameters as JPEG Images: Data-Free Delta Compression with Discrete Cosine Transform 9 Mar 2025 · 0 repositories · arXiv:2503.06676
-
SKG-LLM: Developing a Mathematical Model for Stroke Knowledge Graph Construction Using Large Language Models 9 Mar 2025 · 0 repositories · arXiv:2503.06475
-
Small Vision-Language Models: A Survey on Compact Architectures and Techniques 9 Mar 2025 · 0 repositories · arXiv:2503.10665
-
UniGenX: Unified Generation of Sequence and Structure with Autoregressive Diffusion 9 Mar 2025 · 0 repositories · arXiv:2503.06687
-
A Noise-Robust Turn-Taking System for Real-World Dialogue Robots: A Field Experiment 8 Mar 2025 · 1 repository · arXiv:2503.06241
-
ACAM-KD: Adaptive and Cooperative Attention Masking for Knowledge Distillation 8 Mar 2025 · 0 repositories · arXiv:2503.06307
-
AF-KAN: Activation Function-Based Kolmogorov-Arnold Networks for Efficient Representation Learning 8 Mar 2025 · 1 repository · arXiv:2503.06112
-
Analyzing the Role of Permutation Invariance in Linear Mode Connectivity 8 Mar 2025 · 0 repositories · arXiv:2503.06001
-
Attention on the Wires (AttWire): A Foundation Model for Detecting Devices and Catheters in X-ray Fluoroscopic Images 8 Mar 2025 · 1 repository · arXiv:2503.06190
-
Bimodal Connection Attention Fusion for Speech Emotion Recognition 8 Mar 2025 · 0 repositories · arXiv:2503.05858
-
Breaking Free from MMI: A New Frontier in Rationalization by Probing Input Utilization 8 Mar 2025 · 1 repository · arXiv:2503.06202Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 1 honoured, 1 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
Constructions are Revealed in Word Distributions 8 Mar 2025 · 1 repository · arXiv:2503.06048
-
Disrupting Model Merging: A Parameter-Level Defense Without Sacrificing Accuracy 8 Mar 2025 · 0 repositories · arXiv:2503.07661
-
End-to-End Action Segmentation Transformer 8 Mar 2025 · 0 repositories · arXiv:2503.06316
-
End-to-End HOI Reconstruction Transformer with Graph-based Encoding 8 Mar 2025 · 0 repositories · arXiv:2503.06012
-
Evaluating Discourse Cohesion in Pre-trained Language Models 8 Mar 2025 · 0 repositories · arXiv:2503.06137
-
Feature Fusion Attention Network with CycleGAN for Image Dehazing, De-Snowing and De-Raining 8 Mar 2025 · 0 repositories · arXiv:2503.06107
-
Fine-Grained Bias Detection in LLM: Enhancing detection mechanisms for nuanced biases 8 Mar 2025 · 0 repositories · arXiv:2503.06054
-
Fish2Mesh Transformer: 3D Human Mesh Recovery from Egocentric Vision 8 Mar 2025 · 0 repositories · arXiv:2503.06089
-
Get In Video: Add Anything You Want to the Video 8 Mar 2025 · 0 repositories · arXiv:2503.06268
-
Improving SAM for Camouflaged Object Detection via Dual Stream Adapters 8 Mar 2025 · 0 repositories · arXiv:2503.06042
-
Interpretable High-order Knowledge Graph Neural Network for Predicting Synthetic Lethality in Human Cancers 8 Mar 2025 · 0 repositories · arXiv:2503.06052
-
Lightweight Software Kernels and Hardware Extensions for Efficient Sparse Deep Neural Networks on Microcontrollers 8 Mar 2025 · 0 repositories · arXiv:2503.06183
-
LimTopic: LLM-based Topic Modeling and Text Summarization for Analyzing Scientific Articles limitations 8 Mar 2025 · 1 repository · arXiv:2503.10658
-
MARRO: Multi-headed Attention for Rhetorical Role Labeling in Legal Documents 8 Mar 2025 · 0 repositories · arXiv:2503.10659
-
MoEMoE: Question Guided Dense and Scalable Sparse Mixture-of-Expert for Multi-source Multi-modal Answering 8 Mar 2025 · 0 repositories · arXiv:2503.06296
-
MSConv: Multiplicative and Subtractive Convolution for Face Recognition 8 Mar 2025 · 0 repositories · arXiv:2503.06187
-
NeuraLoc: Visual Localization in Neural Implicit Map with Dual Complementary Features 8 Mar 2025 · 0 repositories · arXiv:2503.06117
-
Object-Centric World Model for Language-Guided Manipulation 8 Mar 2025 · 0 repositories · arXiv:2503.06170
-
Optimizing Generative AI's Accuracy and Transparency in Inductive Thematic Analysis: A Human-AI Comparison 8 Mar 2025 · 0 repositories · arXiv:2503.16485
-
Poisoned-MRAG: Knowledge Poisoning Attacks to Multimodal Retrieval Augmented Generation 8 Mar 2025 · 0 repositories · arXiv:2503.06254
-
Rethinking Lanes and Points in Complex Scenarios for Monocular 3D Lane Detection 8 Mar 2025 · 0 repositories · arXiv:2503.06237
-
RGB-Phase Speckle: Cross-Scene Stereo 3D Reconstruction via Wrapped Pre-Normalization 8 Mar 2025 · 0 repositories · arXiv:2503.06125
-
X2I: Seamless Integration of Multimodal Understanding into Diffusion Transformer via Attention Distillation 8 Mar 2025 · 1 repository · arXiv:2503.06134Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Your Large Vision-Language Model Only Needs A Few Attention Heads For Visual Grounding 8 Mar 2025 · 0 repositories · arXiv:2503.06287
-
A Hybrid Model/Data-Driven Solution to Channel, Position and Orientation Tracking in mmWave Vehicular Systems 7 Mar 2025 · 0 repositories · arXiv:2503.05091
-
A Real-time Multimodal Transformer Neural Network-powered Wildfire Forecasting System 7 Mar 2025 · 0 repositories · arXiv:2503.05971
-
A Survey on Sparse Autoencoders: Interpreting the Internal Mechanisms of Large Language Models 7 Mar 2025 · 0 repositories · arXiv:2503.05613
-
BARK: A Fully Bayesian Tree Kernel for Black-box Optimization 7 Mar 2025 · 0 repositories · arXiv:2503.05574
-
CASP: Compression of Large Multimodal Models Based on Attention Sparsity 7 Mar 2025 · 1 repository · arXiv:2503.05936
-
ColFigPhotoAttnNet: Reliable Finger Photo Presentation Attack Detection Leveraging Window-Attention on Color Spaces 7 Mar 2025 · 1 repository · arXiv:2503.05247
-
CoMoGaussian: Continuous Motion-Aware Gaussian Splatting from Motion-Blurred Images 7 Mar 2025 · 1 repository · arXiv:2503.05332
-
Deep Frequency Attention Networks for Single Snapshot Sparse Array Interpolation 7 Mar 2025 · 0 repositories · arXiv:2503.05486
-
Energy-Free Sensing and Context Recognition Using Photovoltaic Cells 7 Mar 2025 · 1 repository · arXiv:2503.05406
-
Evaluating Large Language Models in Code Generation: INFINITE Methodology for Defining the Inference Index 7 Mar 2025 · 0 repositories · arXiv:2503.05852
-
Simulating and Analysing Human Survey Responses with Large Language Models: A Case Study in Energy Stated Preference 7 Mar 2025 · 0 repositories · arXiv:2503.10652
-
Evaluating open-source Large Language Models for automated fact-checking 7 Mar 2025 · 0 repositories · arXiv:2503.05565
-
Explaining the Unexplainable: A Systematic Review of Explainable AI in Finance 7 Mar 2025 · 0 repositories · arXiv:2503.05966
-
Exploring FMCW Radars and Feature Maps for Activity Recognition: A Benchmark Study 7 Mar 2025 · 0 repositories · arXiv:2503.05629
-
FastMap: Fast Queries Initialization Based Vectorized HD Map Reconstruction Framework 7 Mar 2025 · 1 repository · arXiv:2503.05492
-
FMCHS: Advancing Traditional Chinese Medicine Herb Recommendation with Fusion of Multiscale Correlations of Herbs and Symptoms 7 Mar 2025 · 0 repositories · arXiv:2503.05167
-
FMT:A Multimodal Pneumonia Detection Model Based on Stacking MOE Framework 7 Mar 2025 · 0 repositories · arXiv:2503.05626
-
GaussianCAD: Robust Self-Supervised CAD Reconstruction from Three Orthographic Views Using 3D Gaussian Splatting 7 Mar 2025 · 0 repositories · arXiv:2503.05161
-
Language modelling techniques for analysing the impact of human genetic variation 7 Mar 2025 · 0 repositories · arXiv:2503.10655
-
Leveraging Approximate Caching for Faster Retrieval-Augmented Generation 7 Mar 2025 · 0 repositories · arXiv:2503.05530
-
Leveraging Semantic Type Dependencies for Clinical Named Entity Recognition 7 Mar 2025 · 0 repositories · arXiv:2503.05373
-
Lightweight Hypercomplex MRI Reconstruction: A Generalized Kronecker-Parameterized Approach 7 Mar 2025 · 0 repositories · arXiv:2503.05063
-
Look Before You Leap: Using Serialized State Machine for Language Conditioned Robotic Manipulation 7 Mar 2025 · 0 repositories · arXiv:2503.05114
-
MagicInfinite: Generating Infinite Talking Videos with Your Words and Voice 7 Mar 2025 · 0 repositories · arXiv:2503.05978
-
MastermindEval: A Simple But Scalable Reasoning Benchmark 7 Mar 2025 · 1 repository · arXiv:2503.05891Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
MedCAM-OsteoCls: Medical Context Aware Multimodal Classification of Knee Osteoarthritis 7 Mar 2025 · 1 repository
-
Mol-CADiff: Causality-Aware Autoregressive Diffusion for Molecule Generation 7 Mar 2025 · 0 repositories · arXiv:2503.05499
-
MPTSNet: Integrating Multiscale Periodic Local Patterns and Global Dependencies for Multivariate Time Series Classification 7 Mar 2025 · 1 repository · arXiv:2503.05582
-
Personalized Federated Learning via Learning Dynamic Graphs 7 Mar 2025 · 0 repositories · arXiv:2503.05474
-
Pi-GPS: Enhancing Geometry Problem Solving by Unleashing the Power of Diagrammatic Information 7 Mar 2025 · 0 repositories · arXiv:2503.05543Syntology 7 ran (of which 1 constructed an object rather than computing a result; 5 with no instrument failure: 1 honoured, 3 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
Quantifying the Robustness of Retrieval-Augmented Language Models Against Spurious Features in Grounding Data 7 Mar 2025 · 0 repositories · arXiv:2503.05587
-
R1-Searcher: Incentivizing the Search Capability in LLMs via Reinforcement Learning 7 Mar 2025 · 5 repositories · arXiv:2503.05592Syntology 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
S2S-Arena, Evaluating Speech2Speech Protocols on Instruction Following with Paralinguistic Information 7 Mar 2025 · 0 repositories · arXiv:2503.05085
-
Slim attention: cut your context memory in half without loss of accuracy -- K-cache is all you need for MHA 7 Mar 2025 · 1 repository · arXiv:2503.05840
-
SplatPose: Geometry-Aware 6-DoF Pose Estimation from Single RGB Image via 3D Gaussian Splatting 7 Mar 2025 · 0 repositories · arXiv:2503.05174
-
Task-oriented Uncertainty Collaborative Learning for Label-Efficient Brain Tumor Segmentation 7 Mar 2025 · 1 repository · arXiv:2503.05682
-
Tractable Representations for Convergent Approximation of Distributional HJB Equations 7 Mar 2025 · 0 repositories · arXiv:2503.05563
-
Zero-shot Medical Event Prediction Using a Generative Pre-trained Transformer on Electronic Health Records 7 Mar 2025 · 0 repositories · arXiv:2503.05893
-
A Generalist Cross-Domain Molecular Learning Framework for Structure-Based Drug Discovery 6 Mar 2025 · 0 repositories · arXiv:2503.04362
-
A Unified Framework with Novel Metrics for Evaluating the Effectiveness of XAI Techniques in LLMs 6 Mar 2025 · 0 repositories · arXiv:2503.05050
-
Beyond RAG: Task-Aware KV Cache Compression for Comprehensive Knowledge Reasoning 6 Mar 2025 · 0 repositories · arXiv:2503.04973
-
BicliqueEncoder: An Efficient Method for Link Prediction in Bipartite Networks using Formal Concept Analysis and Transformer Encoder 6 Mar 2025 · 0 repositories · arXiv:2503.07645
-
Can We Optimize Deep RL Policy Weights as Trajectory Modeling? 6 Mar 2025 · 0 repositories · arXiv:2503.04074
-
Chart-HQA: A Benchmark for Hypothetical Question Answering in Charts 6 Mar 2025 · 0 repositories · arXiv:2503.04095
-
Collapse of Dense Retrievers: Short, Early, and Literal Biases Outranking Factual Evidence 6 Mar 2025 · 0 repositories · arXiv:2503.05037
-
Compositional Causal Reasoning Evaluation in Language Models 6 Mar 2025 · 0 repositories · arXiv:2503.04556
-
Conformal forecasting for surgical instrument trajectory 6 Mar 2025 · 0 repositories · arXiv:2503.04191
-
DB-Explore: Automated Database Exploration and Instruction Synthesis for Text-to-SQL 6 Mar 2025 · 0 repositories · arXiv:2503.04959
-
Early Detection of Mental Health Issues Using Social Media Posts 6 Mar 2025 · 0 repositories · arXiv:2503.07653
-
Frequency-Based Alignment of EEG and Audio Signals Using Contrastive Learning and SincNet for Auditory Attention Detection 6 Mar 2025 · 1 repository · arXiv:2503.04156
-
Gate-Shift-Pose: Enhancing Action Recognition in Sports with Skeleton Information 6 Mar 2025 · 1 repository · arXiv:2503.04470
-
GBT-SAM: Adapting a Foundational Deep Learning Model for Generalizable Brain Tumor Segmentation via Efficient Integration of Multi-Parametric MRI Data 6 Mar 2025 · 1 repository · arXiv:2503.04325
-
Hedging with Sparse Reward Reinforcement Learning 6 Mar 2025 · 0 repositories · arXiv:2503.04218