Methods › General › Attention Mechanisms › Attention › Papers, page 35
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 35 of 316: papers 3,401 to 3,500 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
High-Precision Transformer-Based Visual Servoing for Humanoid Robots in Aligning Tiny Objects 6 Mar 2025 · 0 repositories · arXiv:2503.04862
-
HILGEN: Hierarchically-Informed Data Generation for Biomedical NER Using Knowledgebases and Large Language Models 6 Mar 2025 · 0 repositories · arXiv:2503.04930
-
HybridNorm: Towards Stable and Efficient Transformer Training via Hybrid Normalization 6 Mar 2025 · 1 repository · arXiv:2503.04598
-
In-depth Analysis of Graph-based RAG in a Unified Framework 6 Mar 2025 · 0 repositories · arXiv:2503.04338
-
Incentivizing Multi-Tenant Split Federated Learning for Foundation Models at the Network Edge 6 Mar 2025 · 0 repositories · arXiv:2503.04971
-
Interpretable Transformation and Analysis of Timelines through Learning via Surprisability 6 Mar 2025 · 0 repositories · arXiv:2503.04502
-
Joint Masked Reconstruction and Contrastive Learning for Mining Interactions Between Proteins 6 Mar 2025 · 1 repository · arXiv:2503.04650
-
Layer-Specific Scaling of Positional Encodings for Superior Long-Context Modeling 6 Mar 2025 · 0 repositories · arXiv:2503.04355
-
Learning Transformer-based World Models with Contrastive Predictive Coding 6 Mar 2025 · 0 repositories · arXiv:2503.04416
-
Learning Wideband User Scheduling and Hybrid Precoding with Graph Neural Networks 6 Mar 2025 · 0 repositories · arXiv:2503.04233
-
LEDiT: Your Length-Extrapolatable Diffusion Transformer without Positional Encoding 6 Mar 2025 · 0 repositories · arXiv:2503.04344
-
Leveraging Large Language Models to Address Data Scarcity in Machine Learning: Applications in Graphene Synthesis 6 Mar 2025 · 1 repository · arXiv:2503.04870
-
Multi-modal Summarization in Model-Based Engineering: Automotive Software Development Case Study 6 Mar 2025 · 0 repositories · arXiv:2503.04506
-
Scale-Invariant Adversarial Attack against Arbitrary-scale Super-resolution 6 Mar 2025 · 0 repositories · arXiv:2503.04385
-
Toward Lightweight and Fast Decoders for Diffusion Models in Image and Video Generation 6 Mar 2025 · 1 repository · arXiv:2503.04871
-
Towards Autonomous Reinforcement Learning for Real-World Robotic Manipulation with Large Language Models 6 Mar 2025 · 0 repositories · arXiv:2503.04280
-
A Multimodal Framework for Topic Propagation Classification in Social Networks 5 Mar 2025 · 0 repositories · arXiv:2503.03112
-
Addressing Overprescribing Challenges: Fine-Tuning Large Language Models for Medication Recommendation Tasks 5 Mar 2025 · 1 repository · arXiv:2503.03687Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples)
-
Afford-X: Generalizable and Slim Affordance Reasoning for Task-oriented Manipulation 5 Mar 2025 · 0 repositories · arXiv:2503.03556
-
AHCPTQ: Accurate and Hardware-Compatible Post-Training Quantization for Segment Anything Model 5 Mar 2025 · 0 repositories · arXiv:2503.03088
-
All-atom Diffusion Transformers: Unified generative modelling of molecules and materials 5 Mar 2025 · 1 repository · arXiv:2503.03965Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
An Aspect Extraction Framework using Different Embedding Types, Learning Models, and Dependency Structure 5 Mar 2025 · 1 repository · arXiv:2503.03512
-
Analogical Reasoning Inside Large Language Models: Concept Vectors and the Limits of Abstraction 5 Mar 2025 · 1 repository · arXiv:2503.03666
-
BANet: Bilateral Aggregation Network for Mobile Stereo Matching 5 Mar 2025 · 1 repository · arXiv:2503.03259Syntology official (archive's flag): 8 ran · 8 ran (of which 4 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 11 harvested samples)
-
Can Frontier LLMs Replace Annotators in Biomedical Text Mining? Analyzing Challenges and Exploring Solutions 5 Mar 2025 · 1 repository · arXiv:2503.03261
-
Conformal Transformations for Symmetric Power Transformers 5 Mar 2025 · 0 repositories · arXiv:2503.03269
-
DA-STGCN: 4D Trajectory Prediction Based on Spatiotemporal Feature Extraction 5 Mar 2025 · 0 repositories · arXiv:2503.04823
-
Deictic Codes, Demonstratives, and Reference: A Step Toward Solving the Grounding Problem 5 Mar 2025 · 0 repositories · arXiv:2503.03495
-
DTU-Net: A Multi-Scale Dilated Transformer Network for Nonlinear Hyperspectral Unmixing 5 Mar 2025 · 0 repositories · arXiv:2503.03465
-
DualDiff+: Dual-Branch Diffusion for High-Fidelity Video Generation with Reward Guidance 5 Mar 2025 · 1 repository · arXiv:2503.03689
-
Intermediate-Task Transfer Learning: Leveraging Sarcasm Detection for Stance Detection 5 Mar 2025 · 0 repositories · arXiv:2503.03172
-
Introduction to Artificial Consciousness: History, Current Trends and Ethical Challenges 5 Mar 2025 · 0 repositories · arXiv:2503.05823
-
Knowledge Augmentation in Federation: Rethinking What Collaborative Learning Can Bring Back to Decentralized Data 5 Mar 2025 · 0 repositories · arXiv:2503.03140
-
Learning to Reduce Search Space for Generalizable Neural Routing Solver 5 Mar 2025 · 0 repositories · arXiv:2503.03137
-
Large language models in finance : what is financial sentiment? 5 Mar 2025 · 0 repositories · arXiv:2503.03612
-
MA-LoT: Multi-Agent Lean-based Long Chain-of-Thought Reasoning enhances Formal Theorem Proving 5 Mar 2025 · 1 repository · arXiv:2503.03205
-
Multi-View Depth Consistent Image Generation Using Generative AI Models: Application on Architectural Design of University Buildings 5 Mar 2025 · 0 repositories · arXiv:2503.03068
-
On the Relation Between Speech Quality and Quantized Latent Representations of Neural Codecs 5 Mar 2025 · 0 repositories · arXiv:2503.03304
-
Partial Convolution Meets Visual Attention 5 Mar 2025 · 0 repositories · arXiv:2503.03148
-
PathRWKV: Enabling Whole Slide Prediction with Recurrent-Transformer 5 Mar 2025 · 0 repositories · arXiv:2503.03199
-
Personalized Federated Fine-tuning for Heterogeneous Data: An Automatic Rank Learning Approach via Two-Level LoRA 5 Mar 2025 · 0 repositories · arXiv:2503.03920
-
Petri Timo 5 Mar 2025 · 0 repositories
-
PowerAttention: Exponentially Scaling of Receptive Fields for Effective Sparse Attention 5 Mar 2025 · 0 repositories · arXiv:2503.03588
-
Pretrained LLMs as Real-Time Controllers for Robot Operated Serial Production Line 5 Mar 2025 · 0 repositories · arXiv:2503.03889
-
Qieemo: Speech Is All You Need in the Emotion Recognition in Conversations 5 Mar 2025 · 0 repositories · arXiv:2503.22687
-
RiskAgent: Autonomous Medical AI Copilot for Generalist Risk Prediction 5 Mar 2025 · 0 repositories · arXiv:2503.03802
-
RGB-Thermal Infrared Fusion for Robust Depth Estimation in Complex Environments 5 Mar 2025 · 0 repositories · arXiv:2503.04821
-
RVAFM: Re-parameterizing Vertical Attention Fusion Module for Handwritten Paragraph Text Recognition 5 Mar 2025 · 0 repositories · arXiv:2503.03104
-
Sarcasm Detection as a Catalyst: Improving Stance Detection with Cross-Target Capabilities 5 Mar 2025 · 0 repositories · arXiv:2503.03787
-
ScaleFusionNet: Transformer-Guided Multi-Scale Feature Fusion for Skin Lesion Segmentation 5 Mar 2025 · 1 repository · arXiv:2503.03327
-
See What You Are Told: Visual Attention Sink in Large Multimodal Models 5 Mar 2025 · 0 repositories · arXiv:2503.03321
-
The Box is in the Pen: Evaluating Commonsense Reasoning in Neural Machine Translation 5 Mar 2025 · 1 repository · arXiv:2503.03308
-
The Signed Two-Space Proximity Model for Learning Representations in Protein-Protein Interaction Networks 5 Mar 2025 · 0 repositories · arXiv:2503.03904
-
A Joint Visual Compression and Perception Framework for Neuralmorphic Spiking Camera 4 Mar 2025 · 0 repositories · arXiv:2503.02725
-
A Transformer Model for Predicting Chemical Reaction Products from Generic Templates 4 Mar 2025 · 0 repositories · arXiv:2503.05810
-
Adapting Decoder-Based Language Models for Diverse Encoder Downstream Tasks 4 Mar 2025 · 0 repositories · arXiv:2503.02656
-
Attention Bootstrapping for Multi-Modal Test-Time Adaptation 4 Mar 2025 · 0 repositories · arXiv:2503.02221
-
BdSLW401: Transformer-Based Word-Level Bangla Sign Language Recognition Using Relative Quantization Encoding (RQE) 4 Mar 2025 · 0 repositories · arXiv:2503.02360
-
BHViT: Binarized Hybrid Vision Transformer 4 Mar 2025 · 1 repository · arXiv:2503.02394Syntology official (archive's flag): 16 ran · 16 ran (of which 13 constructed an object rather than computing a result; 15 with no instrument failure: 0 honoured, 0 violated, 15 with no contract checked; 1 where Syntology's instrument failed) · 13 unverified (of 29 harvested samples)
-
Boltzmann Attention Sampling for Image Analysis with Small Objects 4 Mar 2025 · 0 repositories · arXiv:2503.02841
-
Controllable Motion Generation via Diffusion Modal Coupling 4 Mar 2025 · 1 repository · arXiv:2503.02353
-
CoServe: Efficient Collaboration-of-Experts (CoE) Model Inference with Limited Memory 4 Mar 2025 · 0 repositories · arXiv:2503.02354
-
CrystalFramer: Rethinking the Role of Frames for SE(3)-Invariant Crystal Structure Modeling 4 Mar 2025 · 0 repositories · arXiv:2503.02209
-
Developing a PET/CT Foundation Model for Cross-Modal Anatomical and Functional Imaging 4 Mar 2025 · 0 repositories · arXiv:2503.02824
-
Disentangled Knowledge Tracing for Alleviating Cognitive Bias 4 Mar 2025 · 2 repositories · arXiv:2503.02539
-
Effectively Steer LLM To Follow Preference via Building Confident Directions 4 Mar 2025 · 0 repositories · arXiv:2503.02989
-
LREA: Low-Rank Efficient Attention on Modeling Long-Term User Behaviors for CTR Prediction 4 Mar 2025 · 0 repositories · arXiv:2503.02542
-
Exploring Token-Level Augmentation in Vision Transformer for Semi-Supervised Semantic Segmentation 4 Mar 2025 · 1 repository · arXiv:2503.02459
-
Extrapolating the long-term seasonal component of electricity prices for forecasting in the day-ahead market 4 Mar 2025 · 0 repositories · arXiv:2503.02518
-
Fair Play in the Fast Lane: Integrating Sportsmanship into Autonomous Racing Systems 4 Mar 2025 · 0 repositories · arXiv:2503.03774
-
FourierNAT: A Fourier-Mixing-Based Non-Autoregressive Transformer for Parallel Sequence Generation 4 Mar 2025 · 0 repositories · arXiv:2503.07630
-
Graph Transformer with Disease Subgraph Positional Encoding for Improved Comorbidity Prediction 4 Mar 2025 · 1 repository · arXiv:2503.03046
-
Haste Makes Waste: Evaluating Planning Abilities of LLMs for Efficient and Feasible Multitasking with Time Constraints Between Actions 4 Mar 2025 · 1 repository · arXiv:2503.02238
-
Interpretable Few-Shot Retinal Disease Diagnosis with Concept-Guided Prompting of Vision-Language Models 4 Mar 2025 · 0 repositories · arXiv:2503.02917
-
JPDS-NN: Reinforcement Learning-Based Dynamic Task Allocation for Agricultural Vehicle Routing Optimization 4 Mar 2025 · 0 repositories · arXiv:2503.02369
-
LADM: Long-context Training Data Selection with Attention-based Dependency Measurement for LLMs 4 Mar 2025 · 0 repositories · arXiv:2503.02502
-
Learning Precoding in Multi-user Multi-antenna Systems: Transformer or Graph Transformer? 4 Mar 2025 · 0 repositories · arXiv:2503.02998
-
LLaVE: Large Language and Vision Embedding Models with Hardness-Weighted Contrastive Learning 4 Mar 2025 · 0 repositories · arXiv:2503.04812
-
LLM Misalignment via Adversarial RLHF Platforms 4 Mar 2025 · 0 repositories · arXiv:2503.03039
-
Multilingualism, Transnationality, and K-pop in the Online #StopAsianHate Movement 4 Mar 2025 · 1 repository · arXiv:2503.02707
-
Network Traffic Classification Using Machine Learning, Transformer, and Large Language Models 4 Mar 2025 · 0 repositories · arXiv:2503.02141
-
NodeNAS: Node-Specific Graph Neural Architecture Search for Out-of-Distribution Generalization 4 Mar 2025 · 0 repositories · arXiv:2503.02448
-
Numerical methods for two-dimensional G-heat equation 4 Mar 2025 · 0 repositories · arXiv:2503.02395
-
Optimizing open-domain question answering with graph-based retrieval augmented generation 4 Mar 2025 · 0 repositories · arXiv:2503.02922
-
PanguIR Technical Report for NTCIR-18 AEOLLM Task 4 Mar 2025 · 0 repositories · arXiv:2503.04809
-
PennyLang: Pioneering LLM-Based Quantum Code Generation with a Novel PennyLane-Centric Dataset 4 Mar 2025 · 0 repositories · arXiv:2503.02497
-
Q-Filters: Leveraging QK Geometry for Efficient KV Cache Compression 4 Mar 2025 · 1 repository · arXiv:2503.02812
-
RACNN: Residual Attention Convolutional Neural Network for Near-Field Channel Estimation in 6G Wireless Communications 4 Mar 2025 · 1 repository · arXiv:2503.02299
-
Resource-Efficient Affordance Grounding with Complementary Depth and Semantic Prompts 4 Mar 2025 · 1 repository · arXiv:2503.02600
-
Seeing is Understanding: Unlocking Causal Attention into Modality-Mutual Attention for Multimodal LLMs 4 Mar 2025 · 1 repository · arXiv:2503.02597Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 2 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 6 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
Sparse Meets Dense: Unified Generative Recommendations with Cascaded Sparse-Dense Representations 4 Mar 2025 · 0 repositories · arXiv:2503.02453
-
STAA-SNN: Spatial-Temporal Attention Aggregator for Spiking Neural Networks 4 Mar 2025 · 0 repositories · arXiv:2503.02689
-
Tabby: Tabular Data Synthesis with Language Models 4 Mar 2025 · 0 repositories · arXiv:2503.02152
-
Target Return Optimizer for Multi-Game Decision Transformer 4 Mar 2025 · 0 repositories · arXiv:2503.02311
-
TeTRA-VPR: A Ternary Transformer Approach for Compact Visual Place Recognition 4 Mar 2025 · 0 repositories · arXiv:2503.02511
-
Towards Robust Multi-UAV Collaboration: MARL with Noise-Resilient Communication and Attention Mechanisms 4 Mar 2025 · 1 repository · arXiv:2503.02913
-
Union of Experts: Adapting Hierarchical Routing to Equivalently Decomposed Transformer 4 Mar 2025 · 1 repository · arXiv:2503.02495
-
Use Me Wisely: AI-Driven Assessment for LLM Prompting Skills Development 4 Mar 2025 · 0 repositories · arXiv:2503.02532
-
Weak-to-Strong Generalization Even in Random Feature Networks, Provably 4 Mar 2025 · 0 repositories · arXiv:2503.02877
-
Wikipedia in the Era of LLMs: Evolution and Risks 4 Mar 2025 · 1 repository · arXiv:2503.02879