Methods › General › Attention Mechanisms › Attention › Papers, page 41
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 41 of 316: papers 4,001 to 4,100 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Modeling Behavior Change for Multi-model At-Risk Students Early Prediction (extended version) 19 Feb 2025 · 0 repositories · arXiv:2503.05734
-
ModSkill: Physical Character Skill Modularization 19 Feb 2025 · 0 repositories · arXiv:2502.14140
-
MoM: Linear Sequence Modeling with Mixture-of-Memories 19 Feb 2025 · 2 repositories · arXiv:2502.13685Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 4 harvested samples)
-
MuDAF: Long-Context Multi-Document Attention Focusing through Contrastive Learning on Attention Heads 19 Feb 2025 · 1 repository · arXiv:2502.13963Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 14 harvested samples) · 7 pointer-only (licence)
-
PitVQA++: Vector Matrix-Low-Rank Adaptation for Open-Ended Visual Question Answering in Pituitary Surgery 19 Feb 2025 · 1 repository · arXiv:2502.14149
-
PLDR-LLMs Learn A Generalizable Tensor Operator That Can Replace Its Own Deep Neural Net At Inference 19 Feb 2025 · 1 repository · arXiv:2502.13502
-
Qwen2.5-VL Technical Report 19 Feb 2025 · 4 repositories · arXiv:2502.13923Syntology 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
RAG-Gym: Optimizing Reasoning and Search Agents with Process Supervision 19 Feb 2025 · 0 repositories · arXiv:2502.13957
-
RAPTOR: Refined Approach for Product Table Object Recognition 19 Feb 2025 · 0 repositories · arXiv:2502.14918
-
Rectified Lagrangian for Out-of-Distribution Detection in Modern Hopfield Networks 19 Feb 2025 · 0 repositories · arXiv:2502.14003
-
Reproducing NevIR: Negation in Neural Information Retrieval 19 Feb 2025 · 2 repositories · arXiv:2502.13506
-
RGAR: Recurrence Generation-augmented Retrieval for Factual-aware Medical Question Answering 19 Feb 2025 · 0 repositories · arXiv:2502.13361
-
RocketKV: Accelerating Long-Context LLM Inference via Two-Stage KV Cache Compression 19 Feb 2025 · 0 repositories · arXiv:2502.14051Syntology 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Spiking Point Transformer for Point Cloud Classification 19 Feb 2025 · 1 repository · arXiv:2502.15811
-
STaR-SQL: Self-Taught Reasoner for Text-to-SQL 19 Feb 2025 · 0 repositories · arXiv:2502.13550
-
The Risk-Neutral Equivalent Pricing of Model-Uncertainty 19 Feb 2025 · 0 repositories · arXiv:2502.13744
-
Token Adaptation via Side Graph Convolution for Temporally and Spatially Efficient Fine-tuning of 3D Point Cloud Transformers 19 Feb 2025 · 1 repository · arXiv:2502.14142
-
Toward Robust Non-Transferable Learning: A Survey and Benchmark 19 Feb 2025 · 1 repository · arXiv:2502.13593
-
TrustRAG: An Information Assistant with Retrieval Augmented Generation 19 Feb 2025 · 1 repository · arXiv:2502.13719
-
UNGT: Ultrasound Nasogastric Tube Dataset for Medical Image Analysis 19 Feb 2025 · 0 repositories · arXiv:2502.14915
-
Universal Semantic Embeddings of Chemical Elements for Enhanced Materials Inference and Discovery 19 Feb 2025 · 0 repositories · arXiv:2502.14912
-
What are Models Thinking about? Understanding Large Language Model Hallucinations "Psychology" through Model Inner State Analysis 19 Feb 2025 · 0 repositories · arXiv:2502.13490
-
Where's the Bug? Attention Probing for Scalable Fault Localization 19 Feb 2025 · 0 repositories · arXiv:2502.13966
-
A²ATS: Retrieval-Based KV Cache Reduction via Windowed Rotary Position Embedding and Query-Aware Vector Quantization 18 Feb 2025 · 0 repositories · arXiv:2502.12665
-
A Survey of Sim-to-Real Methods in RL: Progress, Prospects and Challenges with Foundation Models 18 Feb 2025 · 0 repositories · arXiv:2502.13187
-
An Attention-Assisted Multi-Modal Data Fusion Model for Real-Time Estimation of Underwater Sound Velocity 18 Feb 2025 · 0 repositories · arXiv:2502.12817
-
An LLM-Powered Agent for Physiological Data Analysis: A Case Study on PPG-based Heart Rate Estimation 18 Feb 2025 · 0 repositories · arXiv:2502.12836
-
BaKlaVa -- Budgeted Allocation of KV cache for Long-context Inference 18 Feb 2025 · 0 repositories · arXiv:2502.13176
-
DAMamba: Vision State Space Model with Dynamic Adaptive Scan 18 Feb 2025 · 1 repository · arXiv:2502.12627
-
DeepResonance: Enhancing Multimodal Music Understanding via Music-centric Multi-way Instruction Tuning 18 Feb 2025 · 0 repositories · arXiv:2502.12623Syntology 7 ran (of which 4 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 3 where Syntology's instrument failed) · 5 unverified (of 12 harvested samples)
-
Efficient UAV Coverage in Large Convex Quadrilateral Areas with Elliptical Footprints 18 Feb 2025 · 0 repositories · arXiv:2502.13032
-
Enhancing Audio-Visual Spiking Neural Networks through Semantic-Alignment and Cross-Modal Residual Learning 18 Feb 2025 · 1 repository · arXiv:2502.12488
-
Generative AI Enabled Robust Data Augmentation for Wireless Sensing in ISAC Networks 18 Feb 2025 · 0 repositories · arXiv:2502.12622
-
Gesture-Aware Zero-Shot Speech Recognition for Patients with Language Disorders 18 Feb 2025 · 0 repositories · arXiv:2502.13983
-
HeadInfer: Memory-Efficient LLM Inference by Head-wise Offloading 18 Feb 2025 · 1 repository · arXiv:2502.12574
-
HopRAG: Multi-Hop Reasoning for Logic-Aware Retrieval-Augmented Generation 18 Feb 2025 · 0 repositories · arXiv:2502.12442
-
Hybrid Frequency Transmission for Upload Latency Minimization of IoT Devices in HSR Scenario Aided by Intelligent Reflecting Surfaces 18 Feb 2025 · 0 repositories · arXiv:2502.12642
-
Improving Clinical Question Answering with Multi-Task Learning: A Joint Approach for Answer Extraction and Medical Categorization 18 Feb 2025 · 0 repositories · arXiv:2502.13108
-
Infinite Retrieval: Attention Enhanced LLMs in Long-Context Processing 18 Feb 2025 · 2 repositories · arXiv:2502.12962
-
Instance-Level Moving Object Segmentation from a Single Image with Events 18 Feb 2025 · 0 repositories · arXiv:2502.12975
-
Context-Aware Lifelong Sequential Modeling for Online Click-Through Rate Prediction 18 Feb 2025 · 0 repositories · arXiv:2502.12634
-
Label Drop for Multi-Aspect Relation Modeling in Universal Information Extraction 18 Feb 2025 · 1 repository · arXiv:2502.12614
-
Language Barriers: Evaluating Cross-Lingual Performance of CNN and Transformer Architectures for Speech Quality Estimation 18 Feb 2025 · 0 repositories · arXiv:2502.13004
-
Language Models are Few-Shot Graders 18 Feb 2025 · 0 repositories · arXiv:2502.13337
-
LMN: A Tool for Generating Machine Enforceable Policies from Natural Language Access Control Rules using LLMs 18 Feb 2025 · 0 repositories · arXiv:2502.12460
-
LocalEscaper: A Weakly-supervised Framework with Regional Reconstruction for Scalable Neural TSP Solvers 18 Feb 2025 · 0 repositories · arXiv:2502.12484
-
MALT Diffusion: Memory-Augmented Latent Transformers for Any-Length Video Generation 18 Feb 2025 · 0 repositories · arXiv:2502.12632
-
Masking the Gaps: An Imputation-Free Approach to Time Series Modeling with Missing Data 18 Feb 2025 · 0 repositories · arXiv:2502.15785
-
MatterChat: A Multi-Modal LLM for Material Science 18 Feb 2025 · 0 repositories · arXiv:2502.13107
-
MindLLM: A Subject-Agnostic and Versatile Model for fMRI-to-Text Decoding 18 Feb 2025 · 0 repositories · arXiv:2502.15786
-
MoBA: Mixture of Block Attention for Long-Context LLMs 18 Feb 2025 · 1 repository · arXiv:2502.13189Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
MVCNet: Multi-View Contrastive Network for Motor Imagery Classification 18 Feb 2025 · 1 repository · arXiv:2502.17482
-
Multimodal Mamba: Decoder-only Multimodal State Space Model via Quadratic to Linear Distillation 18 Feb 2025 · 1 repository · arXiv:2502.13145
-
Multimodal Sleep Stage and Sleep Apnea Classification Using Vision Transformer: A Multitask Explainable Learning Approach 18 Feb 2025 · 0 repositories · arXiv:2502.17486
-
Myna: Masking-Based Contrastive Learning of Musical Representations 18 Feb 2025 · 1 repository · arXiv:2502.12511
-
Natural Language Generation from Visual Sequences: Challenges and Future Directions 18 Feb 2025 · 0 repositories · arXiv:2502.13034
-
Neural Attention Search 18 Feb 2025 · 0 repositories · arXiv:2502.13251
-
Oreo: A Plug-in Context Reconstructor to Enhance Retrieval-Augmented Generation 18 Feb 2025 · 0 repositories · arXiv:2502.13019
-
PathRAG: Pruning Graph-based Retrieval Augmented Generation with Relational Paths 18 Feb 2025 · 1 repository · arXiv:2502.14902Syntology official: no sample here; runs from other or unrecorded repositories · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Range Retrieval with Graph-Based Indices 18 Feb 2025 · 0 repositories · arXiv:2502.13245
-
RingFormer: Rethinking Recurrent Transformer with Adaptive Level Signals 18 Feb 2025 · 0 repositories · arXiv:2502.13181
-
Self-Supervised Transformers as Iterative Solution Improvers for Constraint Satisfaction 18 Feb 2025 · 0 repositories · arXiv:2502.15794
-
SparAMX: Accelerating Compressed LLMs Token Generation on AMX-powered CPUs 18 Feb 2025 · 1 repository · arXiv:2502.12444
-
Spherical Dense Text-to-Image Synthesis 18 Feb 2025 · 0 repositories · arXiv:2502.12691
-
Towards an automated workflow in materials science for combining multi-modal simulative and experimental information using data mining and large language models 18 Feb 2025 · 0 repositories · arXiv:2502.14904
-
Tuning Algorithmic and Architectural Hyperparameters in Graph-Based Semi-Supervised Learning with Provable Guarantees 18 Feb 2025 · 0 repositories · arXiv:2502.12937
-
When Segmentation Meets Hyperspectral Image: New Paradigm for Hyperspectral Image Classification 18 Feb 2025 · 1 repository · arXiv:2502.12541
-
Graph Neural Network-based Spectral Filtering Mechanism for Imbalance Classification in Network Digital Twin 17 Feb 2025 · 0 repositories · arXiv:2502.11505
-
A Survey on Bridging EEG Signals and Generative AI: From Image and Text to Beyond 17 Feb 2025 · 0 repositories · arXiv:2502.12048
-
AAKT: Enhancing Knowledge Tracing with Alternate Autoregressive Modeling 17 Feb 2025 · 1 repository · arXiv:2502.11817
-
Fate: Fast Edge Inference of Mixture-of-Experts Models via Cross-Layer Gate 17 Feb 2025 · 1 repository · arXiv:2502.12224
-
AdaSplash: Adaptive Sparse Flash Attention 17 Feb 2025 · 1 repository · arXiv:2502.12082Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
AI-generated Text Detection with a GLTR-based Approach 17 Feb 2025 · 0 repositories · arXiv:2502.12064
-
APB: Accelerating Distributed Long-Context Inference by Passing Compressed Context Blocks across GPUs 17 Feb 2025 · 1 repository · arXiv:2502.12085
-
Beyond Cortisol! Physiological Indicators of Welfare for Dogs: Deficits, Misunderstandings and Opportunities 17 Feb 2025 · 0 repositories · arXiv:2502.11384
-
Biases in Edge Language Models: Detection, Analysis, and Mitigation 17 Feb 2025 · 0 repositories · arXiv:2502.11349
-
Can LLMs Simulate Social Media Engagement? A Study on Action-Guided Response Generation 17 Feb 2025 · 0 repositories · arXiv:2502.12073
-
CMQCIC-Bench: A Chinese Benchmark for Evaluating Large Language Models in Medical Quality Control Indicator Calculation 17 Feb 2025 · 0 repositories · arXiv:2502.11703
-
Control-CLIP: Decoupling Category and Style Guidance in CLIP for Specific-Domain Generation 17 Feb 2025 · 0 repositories · arXiv:2502.11532
-
Deep Spatio-Temporal Neural Network for Air Quality Reanalysis 17 Feb 2025 · 1 repository · arXiv:2502.11941
-
Dictionary-Learning-Based Data Pruning for System Identification 17 Feb 2025 · 0 repositories · arXiv:2502.11484
-
DiSCo: Device-Server Collaborative LLM-Based Text Streaming Services 17 Feb 2025 · 0 repositories · arXiv:2502.11417
-
Diversity-Oriented Data Augmentation with Large Language Models 17 Feb 2025 · 0 repositories · arXiv:2502.11671
-
Does RAG Really Perform Bad For Long-Context Processing? 17 Feb 2025 · 0 repositories · arXiv:2502.11444
-
Exploring Translation Mechanism of Large Language Models 17 Feb 2025 · 0 repositories · arXiv:2502.11806
-
Fast or Better? Balancing Accuracy and Cost in Retrieval-Augmented Generation with Flexible User Control 17 Feb 2025 · 1 repository · arXiv:2502.12145
-
FineFilter: A Fine-grained Noise Filtering Mechanism for Retrieval-Augmented Large Language Models 17 Feb 2025 · 0 repositories · arXiv:2502.11811
-
From Gaming to Research: GTA V for Synthetic Data Generation for Robotics and Navigations 17 Feb 2025 · 0 repositories · arXiv:2502.12303
-
Generative Multi-Agent Collaboration in Embodied AI: A Systematic Review 17 Feb 2025 · 0 repositories · arXiv:2502.11518
-
GLTW: Joint Improved Graph Transformer and LLM via Three-Word Language for Knowledge Graph Completion 17 Feb 2025 · 0 repositories · arXiv:2502.11471
-
Hierarchical Graph Topic Modeling with Topic Tree-based Transformer 17 Feb 2025 · 0 repositories · arXiv:2502.11345
-
Hyperspherical Energy Transformer with Recurrent Depth 17 Feb 2025 · 0 repositories · arXiv:2502.11646
-
If Attention Serves as a Cognitive Model of Human Memory Retrieval, What is the Plausible Memory Representation? 17 Feb 2025 · 0 repositories · arXiv:2502.11469Syntology 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 1 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 8 harvested samples) · 1 pointer-only (licence)
-
Leader and Follower: Interactive Motion Generation under Trajectory Constraints 17 Feb 2025 · 0 repositories · arXiv:2502.11563
-
Linear Diffusion Networks 17 Feb 2025 · 1 repository · arXiv:2502.12381
-
LMFCA-Net: A Lightweight Model for Multi-Channel Speech Enhancement with Efficient Narrow-Band and Cross-Band Attention 17 Feb 2025 · 0 repositories · arXiv:2502.11462
-
Low-Rank Thinning 17 Feb 2025 · 1 repository · arXiv:2502.12063
-
MaskGWM: A Generalizable Driving World Model with Video Mask Reconstruction 17 Feb 2025 · 1 repository · arXiv:2502.11663Syntology official (archive's flag): 3 ran · 4 ran (of which 1 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 8 harvested samples) · 1 pointer-only (licence)
-
MuSC: Improving Complex Instruction Following with Multi-granularity Self-Contrastive Training 17 Feb 2025 · 1 repository · arXiv:2502.11541
-
OCT Data is All You Need: How Vision Transformers with and without Pre-training Benefit Imaging 17 Feb 2025 · 0 repositories · arXiv:2502.12379