Methods › General › Attention Mechanisms › Attention › Papers, page 5
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 5 of 316: papers 401 to 500 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Measuring Participant Contributions in Decentralized Federated Learning 29 May 2025 · 0 repositories · arXiv:2505.23246
-
MMGT: Motion Mask Guided Two-Stage Network for Co-Speech Gesture Video Generation 29 May 2025 · 1 repository · arXiv:2505.23120
-
MOVi: Training-free Text-conditioned Multi-Object Video Generation 29 May 2025 · 0 repositories · arXiv:2505.22980
-
Multi-Group Proportional Representation for Text-to-Image Models 29 May 2025 · 0 repositories · arXiv:2505.24023
-
Multi-Sourced Compositional Generalization in Visual Question Answering 29 May 2025 · 1 repository · arXiv:2505.23045
-
Neural Interpretable PDEs: Harmonizing Fourier Insights with Attention for Scalable and Interpretable Physics Discovery 29 May 2025 · 1 repository · arXiv:2505.23106
-
On the Validity of Head Motion Patterns as Generalisable Depression Biomarkers 29 May 2025 · 0 repositories · arXiv:2505.23427
-
PAN-Crafter: Learning Modality-Consistent Alignment for PAN-Sharpening 29 May 2025 · 0 repositories · arXiv:2505.23367
-
Parameter-Free Bio-Inspired Channel Attention for Enhanced Cardiac MRI Reconstruction 29 May 2025 · 0 repositories · arXiv:2505.23872
-
Patient Domain Supervised Contrastive Learning for Lung Sound Classification Using Mobile Phone 29 May 2025 · 0 repositories · arXiv:2505.23132
-
Probing Association Biases in LLM Moderation Over-Sensitivity 29 May 2025 · 0 repositories · arXiv:2505.23914
-
Query Routing for Retrieval-Augmented Language Models 29 May 2025 · 0 repositories · arXiv:2505.23052
-
Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information 29 May 2025 · 1 repository · arXiv:2505.23558
-
Reducing Latency in LLM-Based Natural Language Commands Processing for Robot Navigation 29 May 2025 · 0 repositories · arXiv:2506.00075
-
Rethinking Regularization Methods for Knowledge Graph Completion 29 May 2025 · 0 repositories · arXiv:2505.23442
-
SafeCOMM: What about Safety Alignment in Fine-Tuned Telecom Large Language Models? 29 May 2025 · 0 repositories · arXiv:2506.00062
-
Semantics-Aware Human Motion Generation from Audio Instructions 29 May 2025 · 0 repositories · arXiv:2505.23465
-
Sentinel: Attention Probing of Proxy Models for LLM Context Compression with an Understanding Perspective 29 May 2025 · 1 repository · arXiv:2505.23277
-
Table-R1: Inference-Time Scaling for Table Reasoning 29 May 2025 · 1 repository · arXiv:2505.23621
-
The Warmup Dilemma: How Learning Rate Strategies Impact Speech-to-Text Model Convergence 29 May 2025 · 1 repository · arXiv:2505.23420
-
Threading the Needle: Reweaving Chain-of-Thought Reasoning to Explain Human Label Variation 29 May 2025 · 1 repository · arXiv:2505.23368Syntology official (archive's flag): 2 ran · 2 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition 29 May 2025 · 0 repositories · arXiv:2505.24059
-
Towards Robust Overlapping Speech Detection: A Speaker-Aware Progressive Approach Using WavLM 29 May 2025 · 0 repositories · arXiv:2505.23207
-
VF-Eval: Evaluating Multimodal LLMs for Generating Feedback on AIGC Videos 29 May 2025 · 1 repository · arXiv:2505.23693
-
VITON-DRR: Details Retention Virtual Try-on via Non-rigid Registration 29 May 2025 · 1 repository · arXiv:2505.23439
-
Zero-to-Hero: Zero-Shot Initialization Empowering Reference-Based Video Appearance Editing 29 May 2025 · 1 repository · arXiv:2505.23134
-
ZPressor: Bottleneck-Aware Compression for Scalable Feed-Forward 3DGS 29 May 2025 · 1 repository · arXiv:2505.23734Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
Agent-UniRAG: A Trainable Open-Source LLM Agent Framework for Unified Retrieval-Augmented Generation Systems 28 May 2025 · 0 repositories · arXiv:2505.22571
-
Are classical deep neural networks weakly adversarially robust? 28 May 2025 · 0 repositories · arXiv:2506.02016
-
Attention-Enhanced Prompt Decision Transformers for UAV-Assisted Communications with AoI 28 May 2025 · 0 repositories · arXiv:2505.22170
-
Bayesian Attention Mechanism: A Probabilistic Framework for Positional Encoding and Context Length Extrapolation 28 May 2025 · 1 repository · arXiv:2505.22842Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Breaking the Cloak! Unveiling Chinese Cloaked Toxicity with Homophone Graph and Toxic Lexicon 28 May 2025 · 0 repositories · arXiv:2505.22184
-
Climate Finance Bench 28 May 2025 · 1 repository · arXiv:2505.22752
-
Contextual Memory Intelligence -- A Foundational Paradigm for Human-AI Collaboration and Reflective Generative AI Systems 28 May 2025 · 0 repositories · arXiv:2506.05370
-
Cross-modal RAG: Sub-dimensional Retrieval-Augmented Text-to-Image Generation 28 May 2025 · 1 repository · arXiv:2505.21956
-
Curse of High Dimensionality Issue in Transformer for Long-context Modeling 28 May 2025 · 1 repository · arXiv:2505.22107Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 5 where Syntology's instrument failed) · 3 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
HiDream-I1: A High-Efficient Image Generative Foundation Model with Sparse Diffusion Transformer 28 May 2025 · 2 repositories · arXiv:2505.22705Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Hierarchical Material Recognition from Local Appearance 28 May 2025 · 0 repositories · arXiv:2505.22911
-
HydraNet: Momentum-Driven State Space Duality for Multi-Granularity Tennis Tournaments Analysis 28 May 2025 · 1 repository · arXiv:2505.21882
-
Improving QA Efficiency with DistilBERT: Fine-Tuning and Inference on mobile Intel CPUs 28 May 2025 · 0 repositories · arXiv:2505.22937
-
Judging LLMs on a Simplex 28 May 2025 · 0 repositories · arXiv:2505.21972
-
Large Language Models for Depression Recognition in Spoken Language Integrating Psychological Knowledge 28 May 2025 · 1 repository · arXiv:2505.22863
-
Mitigating Audiovisual Mismatch in Visual-Guide Audio Captioning 28 May 2025 · 0 repositories · arXiv:2505.22045
-
Multi-MLLM Knowledge Distillation for Out-of-Context News Detection 28 May 2025 · 0 repositories · arXiv:2505.22517
-
MultiFormer: A Multi-Person Pose Estimation System Based on CSI and Attention Mechanism 28 May 2025 · 0 repositories · arXiv:2505.22555
-
Mustafar: Promoting Unstructured Sparsity for KV Cache Pruning in LLM Inference 28 May 2025 · 1 repository · arXiv:2505.22913Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
NGPU-LM: GPU-Accelerated N-Gram Language Model for Context-Biasing in Greedy ASR Decoding 28 May 2025 · 0 repositories · arXiv:2505.22857
-
Online Fair Division for Personalized 2-Value Instances 28 May 2025 · 0 repositories · arXiv:2505.22174
-
Physics-Informed Distillation of Diffusion Models for PDE-Constrained Generation 28 May 2025 · 0 repositories · arXiv:2505.22391
-
PS4PRO: Pixel-to-pixel Supervision for Photorealistic Rendering and Optimization 28 May 2025 · 0 repositories · arXiv:2505.22616
-
RAGPPI: RAG Benchmark for Protein-Protein Interactions in Drug Discovery 28 May 2025 · 1 repository · arXiv:2505.23823
-
Re-ttention: Ultra Sparse Visual Generation via Attention Statistical Reshape 28 May 2025 · 1 repository · arXiv:2505.22918
-
Say What You Mean: Natural Language Access Control with Large Language Models for Internet of Things 28 May 2025 · 0 repositories · arXiv:2505.23835
-
SHTOcc: Effective 3D Occupancy Prediction with Sparse Head and Tail Voxels 28 May 2025 · 0 repositories · arXiv:2505.22461
-
SkewRoute: Training-Free LLM Routing for Knowledge Graph Retrieval-Augmented Generation via Score Skewness of Retrieved Context 28 May 2025 · 0 repositories · arXiv:2505.23841
-
SlimLLM: Accurate Structured Pruning for Large Language Models 28 May 2025 · 0 repositories · arXiv:2505.22689
-
SplitLoRA: Balancing Stability and Plasticity in Continual Learning Through Gradient Space Splitting 28 May 2025 · 0 repositories · arXiv:2505.22370
-
Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models 28 May 2025 · 0 repositories · arXiv:2505.22271
-
Triple Attention Transformer Architecture for Time-Dependent Concrete Creep Prediction 28 May 2025 · 0 repositories · arXiv:2506.04243
-
UP-SLAM: Adaptively Structured Gaussian SLAM with Uncertainty Prediction in Dynamic Environments 28 May 2025 · 0 repositories · arXiv:2505.22335
-
Update Your Transformer to the Latest Release: Re-Basin of Task Vectors 28 May 2025 · 1 repository · arXiv:2505.22697Syntology official (archive's flag): 1 ran · 5 ran (of which 1 constructed an object rather than computing a result; 5 with no instrument failure: 1 honoured, 3 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning 28 May 2025 · 1 repository · arXiv:2505.22019Syntology official: no sample here; runs from other or unrecorded repositories · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
When Does Neuroevolution Outcompete Reinforcement Learning in Transfer Learning Tasks? 28 May 2025 · 1 repository · arXiv:2505.22696
-
A domain adaptation neural network for digital twin-supported fault diagnosis 27 May 2025 · 1 repository · arXiv:2505.21046
-
AgriFM: A Multi-source Temporal Remote Sensing Foundation Model for Crop Mapping 27 May 2025 · 1 repository · arXiv:2505.21357
-
BacktrackAgent: Enhancing GUI Agent with Error Detection and Backtracking Mechanism 27 May 2025 · 0 repositories · arXiv:2505.20660
-
Beyond 1D: Vision Transformers and Multichannel Signal Images for PPG-to-ECG Reconstruction 27 May 2025 · 0 repositories · arXiv:2505.21767
-
CNN-Based Channel Map Estimation for Movable Antenna Systems 27 May 2025 · 0 repositories · arXiv:2505.21001
-
Continuous-Time Attention: PDE-Guided Mechanisms for Long-Sequence Transformers 27 May 2025 · 0 repositories · arXiv:2505.20666
-
DeepConvContext: A Multi-Scale Approach to Timeseries Classification in Human Activity Recognition 27 May 2025 · 1 repository · arXiv:2505.20894
-
Diagnosing and Resolving Cloud Platform Instability with Multi-modal RAG LLMs 27 May 2025 · 0 repositories · arXiv:2505.21419
-
DiMoSR: Feature Modulation via Multi-Branch Dilated Convolutions for Efficient Image Super-Resolution 27 May 2025 · 1 repository · arXiv:2505.21262
-
Don't Think Longer, Think Wisely: Optimizing Thinking Dynamics for Large Reasoning Models 27 May 2025 · 0 repositories · arXiv:2505.21765
-
Efficient Leaf Disease Classification and Segmentation using Midpoint Normalization Technique and Attention Mechanism 27 May 2025 · 0 repositories · arXiv:2505.21316
-
Emotion-aware Dual Cross-Attentive Neural Network with Label Fusion for Stance Detection in Misinformative Social Media Content 27 May 2025 · 1 repository · arXiv:2505.23812
-
Explainability of Large Language Models using SMILE: Statistical Model-agnostic Interpretability with Local Explanations 27 May 2025 · 1 repository · arXiv:2505.21657
-
FastFace: Tuning Identity Preservation in Distilled Diffusion via Guidance and Attention 27 May 2025 · 1 repository · arXiv:2505.21144
-
From prosthetic memory to prosthetic denial: Auditing whether large language models are prone to mass atrocity denialism 27 May 2025 · 0 repositories · arXiv:2505.21753
-
HAD: Hybrid Architecture Distillation Outperforms Teacher in Genomic Sequence Modeling 27 May 2025 · 0 repositories · arXiv:2505.20836
-
Hardware-Efficient Attention for Fast Decoding 27 May 2025 · 2 repositories · arXiv:2505.21487
-
HTMNet: A Hybrid Network with Transformer-Mamba Bottleneck Multimodal Fusion for Transparent and Reflective Objects Depth Completion 27 May 2025 · 0 repositories · arXiv:2505.20904
-
Long Context Scaling: Divide and Conquer via Multi-Agent Question-driven Collaboration 27 May 2025 · 0 repositories · arXiv:2505.20625
-
MARS-Bench: A Multi-turn Athletic Real-world Scenario Benchmark for Dialogue Evaluation 27 May 2025 · 0 repositories · arXiv:2505.23810
-
Minute-Long Videos with Dual Parallelisms 27 May 2025 · 1 repository · arXiv:2505.21070
-
Mitigating Hallucination in Large Vision-Language Models via Adaptive Attention Calibration 27 May 2025 · 0 repositories · arXiv:2505.21472
-
MoE-Gyro: Self-Supervised Over-Range Reconstruction and Denoising for MEMS Gyroscopes 27 May 2025 · 0 repositories · arXiv:2506.06318
-
MoPFormer: Motion-Primitive Transformer for Wearable-Sensor Activity Recognition 27 May 2025 · 0 repositories · arXiv:2505.20744
-
Music's Multimodal Complexity in AVQA: Why We Need More than General Multimodal LLMs 27 May 2025 · 0 repositories · arXiv:2505.20638
-
Object-Centric Action-Enhanced Representations for Robot Visuo-Motor Policy Learning 27 May 2025 · 0 repositories · arXiv:2505.20962
-
Pause Tokens Strictly Increase the Expressivity of Constant-Depth Transformers 27 May 2025 · 0 repositories · arXiv:2505.21024
-
Plug-and-Play Co-Occurring Face Attention for Robust Audio-Visual Speaker Extraction 27 May 2025 · 0 repositories · arXiv:2505.20635
-
Position is Power: System Prompts as a Mechanism of Bias in Large Language Models (LLMs) 27 May 2025 · 0 repositories · arXiv:2505.21091
-
Privacy-Preserving Chest X-ray Report Generation via Multimodal Federated Learning with ViT and GPT-2 27 May 2025 · 0 repositories · arXiv:2505.21715
-
SageAttention2++: A More Efficient Implementation of SageAttention2 27 May 2025 · 2 repositories · arXiv:2505.21136Syntology official: harvested, nothing ran · 0 ran · 5 unverified (of 5 harvested samples) · 3 pointer-only (licence)
-
SOSBENCH: Benchmarking Safety Alignment on Scientific Knowledge 27 May 2025 · 0 repositories · arXiv:2505.21605
-
SpecExtend: A Drop-in Enhancement for Speculative Decoding of Long Sequences 27 May 2025 · 1 repository · arXiv:2505.20776
-
tenSVD algorithm for compression 27 May 2025 · 0 repositories · arXiv:2505.21686
-
Time-Series Learning for Proactive Fault Prediction in Distributed Systems with Deep Neural Structures 27 May 2025 · 0 repositories · arXiv:2505.20705
-
Towards Robust Assessment of Pathological Voices via Combined Low-Level Descriptors and Foundation Model Representations 27 May 2025 · 0 repositories · arXiv:2505.21356
-
Unpaired Image-to-Image Translation for Segmentation and Signal Unmixing 27 May 2025 · 0 repositories · arXiv:2505.20746