Methods › General › Attention Mechanisms › Attention › Papers, page 7
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 7 of 316: papers 601 to 700 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Freqformer: Image-Demoiréing Transformer via Efficient Frequency Decomposition 25 May 2025 · 1 repository · arXiv:2505.19120
-
GhostPrompt: Jailbreaking Text-to-image Generative Models based on Dynamic Optimization 25 May 2025 · 0 repositories · arXiv:2505.18979
-
Graph-Based Operator Learning from Limited Data on Irregular Domains 25 May 2025 · 0 repositories · arXiv:2505.18923
-
Hermes@DravidianLangTech 2025: Sentiment Analysis of Dravidian Languages using XLM-RoBERTa 25 May 2025 · 1 repository
-
Hypercube-RAG: Hypercube-Based Retrieval-Augmented Generation for In-domain Scientific Question-Answering 25 May 2025 · 1 repository · arXiv:2505.19288
-
Investigating Pedagogical Teacher and Student LLM Agents: Genetic Adaptation Meets Retrieval Augmented Generation Across Learning Style 25 May 2025 · 0 repositories · arXiv:2505.19173
-
JEDI: The Force of Jensen-Shannon Divergence in Disentangling Diffusion Models 25 May 2025 · 0 repositories · arXiv:2505.19166
-
Kernel Space Diffusion Model for Efficient Remote Sensing Pansharpening 25 May 2025 · 0 repositories · arXiv:2505.18991
-
Optimized Text Embedding Models and Benchmarks for Amharic Passage Retrieval 25 May 2025 · 1 repository · arXiv:2505.19356
-
POQD: Performance-Oriented Query Decomposer for Multi-vector retrieval 25 May 2025 · 1 repository · arXiv:2505.19189Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 3 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
PosePilot: An Edge-AI Solution for Posture Correction in Physical Exercises 25 May 2025 · 0 repositories · arXiv:2505.19186
-
Rethinking Metrics and Benchmarks of Video Anomaly Detection 25 May 2025 · 0 repositories · arXiv:2505.19022
-
Retrieval-Augmented Generation for Service Discovery: Chunking Strategies and Benchmarking 25 May 2025 · 0 repositories · arXiv:2505.19310
-
SATORI-R1: Incentivizing Multimodal Reasoning with Spatial Grounding and Verifiable Rewards 25 May 2025 · 1 repository · arXiv:2505.19094
-
SETransformer: A Hybrid Attention-Based Architecture for Robust Human Activity Recognition 25 May 2025 · 0 repositories · arXiv:2505.19369
-
Sparse-to-Dense: A Free Lunch for Lossless Acceleration of Video Understanding in LLMs 25 May 2025 · 0 repositories · arXiv:2505.19155
-
System-1.5 Reasoning: Traversal in Language and Latent Spaces with Dynamic Shortcuts 25 May 2025 · 0 repositories · arXiv:2505.18962
-
Veta-GS: View-dependent deformable 3D Gaussian Splatting for thermal infrared Novel-view Synthesis 25 May 2025 · 0 repositories · arXiv:2505.19138
-
A Survey of LLM × DATA 24 May 2025 · 2 repositories · arXiv:2505.18458
-
ALPS: Attention Localization and Pruning Strategy for Efficient Alignment of Large Language Models 24 May 2025 · 0 repositories · arXiv:2505.18799
-
Benchmarking Poisoning Attacks against Retrieval-Augmented Generation 24 May 2025 · 0 repositories · arXiv:2505.18543
-
BRIT: Bidirectional Retrieval over Unified Image-Text Graph 24 May 2025 · 0 repositories · arXiv:2505.18450
-
CiRL: Open-Source Environments for Reinforcement Learning in Circular Economy and Net Zero 24 May 2025 · 1 repository · arXiv:2505.21536
-
Distribution-Aware Mobility-Assisted Decentralized Federated Learning 24 May 2025 · 0 repositories · arXiv:2505.18866
-
Doc-CoB: Enhancing Multi-Modal Document Understanding with Visual Chain-of-Boxes Reasoning 24 May 2025 · 0 repositories · arXiv:2505.18603
-
Efficient and Workload-Aware LLM Serving via Runtime Layer Swapping and KV Cache Resizing 24 May 2025 · 0 repositories · arXiv:2506.02006
-
Enhancing Efficiency and Exploration in Reinforcement Learning for LLMs 24 May 2025 · 1 repository · arXiv:2505.18573
-
Evaluating the Usefulness of Non-Diagnostic Speech Data for Developing Parkinson's Disease Classifiers 24 May 2025 · 1 repository · arXiv:2505.18722
-
EvdCLIP: Improving Vision-Language Retrieval with Entity Visual Descriptions from Large Language Models 24 May 2025 · 0 repositories · arXiv:2505.18594
-
Federated Retrieval-Augmented Generation: A Systematic Mapping Study 24 May 2025 · 0 repositories · arXiv:2505.18906
-
Focus on What Matters: Enhancing Medical Vision-Language Models with Automatic Attention Alignment Tuning 24 May 2025 · 0 repositories · arXiv:2505.18503
-
From Reddit to Generative AI: Evaluating Large Language Models for Anxiety Support Fine-tuned on Social Media Data 24 May 2025 · 0 repositories · arXiv:2505.18464
-
GainRAG: Preference Alignment in Retrieval-Augmented Generation through Gain Signal Synthesis 24 May 2025 · 1 repository · arXiv:2505.18710
-
High-order Equivariant Flow Matching for Density Functional Theory Hamiltonian Prediction 24 May 2025 · 0 repositories · arXiv:2505.18817
-
How Does Sequence Modeling Architecture Influence Base Capabilities of Pre-trained Language Models? Exploring Key Architecture Design Principles to Avoid Base Capabilities Degradation 24 May 2025 · 0 repositories · arXiv:2505.18522
-
HyperFake: Hyperspectral Reconstruction and Attention-Guided Analysis for Advanced Deepfake Detection 24 May 2025 · 0 repositories · arXiv:2505.18587
-
Is Attention Required for Transformer Inference? Explore Function-preserving Attention Replacement 24 May 2025 · 0 repositories · arXiv:2505.21535
-
LLMs for Supply Chain Management 24 May 2025 · 0 repositories · arXiv:2505.18597
-
Localizing Knowledge in Diffusion Transformers 24 May 2025 · 0 repositories · arXiv:2505.18832
-
Lookahead Q-Cache: Achieving More Consistent KV Cache Eviction via Pseudo Query 24 May 2025 · 0 repositories · arXiv:2505.20334
-
Manifold-aware Representation Learning for Degradation-agnostic Image Restoration 24 May 2025 · 0 repositories · arXiv:2505.18679
-
MonarchAttention: Zero-Shot Conversion to Fast, Hardware-Aware Structured Attention 24 May 2025 · 1 repository · arXiv:2505.18698Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
MSLAU-Net: A Hybird CNN-Transformer Network for Medical Image Segmentation 24 May 2025 · 1 repository · arXiv:2505.18823
-
Partition Generative Modeling: Masked Modeling Without Masks 24 May 2025 · 1 repository · arXiv:2505.18883Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 1 honoured, 0 violated, 10 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 14 harvested samples) · 12 pointer-only (licence)
-
Predictive Performance of Deep Quantum Data Re-uploading Models 24 May 2025 · 0 repositories · arXiv:2505.20337
-
Pruning for Performance: Efficient Idiom and Metaphor Classification in Low-Resource Konkani Using mBERT 24 May 2025 · 0 repositories · arXiv:2506.02005
-
Reinforcement Fine-Tuning Powers Reasoning Capability of Multimodal Large Language Models 24 May 2025 · 1 repository · arXiv:2505.18536
-
Removal of Hallucination on Hallucination: Debate-Augmented RAG 24 May 2025 · 1 repository · arXiv:2505.18581Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 4 harvested samples)
-
Rethinking Causal Mask Attention for Vision-Language Inference 24 May 2025 · 0 repositories · arXiv:2505.18605
-
Security Concerns for Large Language Models: A Survey 24 May 2025 · 0 repositories · arXiv:2505.18889
-
Smart Energy Guardian: A Hybrid Deep Learning Model for Detecting Fraudulent PV Generation 24 May 2025 · 0 repositories · arXiv:2505.18755
-
Sparse VideoGen2: Accelerate Video Generation with Sparse Attention via Semantic-Aware Permutation 24 May 2025 · 0 repositories · arXiv:2505.18875
-
Strong Membership Inference Attacks on Massive Datasets and (Moderately) Large Language Models 24 May 2025 · 0 repositories · arXiv:2505.18773
-
SW-ViT: A Spatio-Temporal Vision Transformer Network with Post Denoiser for Sequential Multi-Push Ultrasound Shear Wave Elastography 24 May 2025 · 0 repositories · arXiv:2505.18865
-
The Silent Saboteur: Imperceptible Adversarial Attacks against Black-Box Retrieval-Augmented Generation Systems 24 May 2025 · 0 repositories · arXiv:2505.18583
-
ToDRE: Visual Token Pruning via Diversity and Task Awareness for Efficient Large Vision-Language Models 24 May 2025 · 0 repositories · arXiv:2505.18757
-
TrajMoE: Spatially-Aware Mixture of Experts for Unified Human Mobility Modeling 24 May 2025 · 0 repositories · arXiv:2505.18670
-
Unifying Attention Heads and Task Vectors via Hidden State Geometry in In-Context Learning 24 May 2025 · 0 repositories · arXiv:2505.18752
-
Unleashing Diffusion Transformers for Visual Correspondence by Modulating Massive Activations 24 May 2025 · 0 repositories · arXiv:2505.18584
-
VORTA: Efficient Video Diffusion via Routing Sparse Attention 24 May 2025 · 1 repository · arXiv:2505.18809Syntology official (archive's flag): 2 ran · 2 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
A Robust PPO-optimized Tabular Transformer Framework for Intrusion Detection in Industrial IoT Systems 23 May 2025 · 1 repository · arXiv:2505.18234
-
Anatomy-Guided Multitask Learning for MRI-Based Classification of Placenta Accreta Spectrum and its Subtypes 23 May 2025 · 0 repositories · arXiv:2505.17484
-
ATMM-SAGA: Alternating Training for Multi-Module with Score-Aware Gated Attention SASV system 23 May 2025 · 0 repositories · arXiv:2505.18273
-
BOTM: Echocardiography Segmentation via Bi-directional Optimal Token Matching 23 May 2025 · 0 repositories · arXiv:2505.18052
-
CENet: Context Enhancement Network for Medical Image Segmentation 23 May 2025 · 1 repository · arXiv:2505.18423
-
COLORA: Efficient Fine-Tuning for Convolutional Models with a Study Case on Optical Coherence Tomography Image Classification 23 May 2025 · 0 repositories · arXiv:2505.18315
-
ComfyMind: Toward General-Purpose Generation via Tree-Based Planning and Reactive Feedback 23 May 2025 · 2 repositories · arXiv:2505.17908Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 10 harvested samples) · 2 pointer-only (licence)
-
Contrastive Distillation of Emotion Knowledge from LLMs for Zero-Shot Emotion Recognition 23 May 2025 · 1 repository · arXiv:2505.18040
-
DECT-based Space-Squeeze Method for Multi-Class Classification of Metastatic Lymph Nodes in Breast Cancer 23 May 2025 · 1 repository · arXiv:2505.17528
-
DetailFusion: A Dual-branch Framework with Detail Enhancement for Composed Image Retrieval 23 May 2025 · 0 repositories · arXiv:2505.17796
-
Direct3D-S2: Gigascale 3D Generation Made Easy with Spatial Sparse Attention 23 May 2025 · 1 repository · arXiv:2505.17412Syntology 6 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 10 harvested samples)
-
Dual Attention Residual U-Net for Accurate Brain Ultrasound Segmentation in IVH Detection 23 May 2025 · 1 repository · arXiv:2505.17683
-
EVM-Fusion: An Explainable Vision Mamba Architecture with Neural Algorithmic Fusion 23 May 2025 · 0 repositories · arXiv:2505.17367
-
Explainable Anatomy-Guided AI for Prostate MRI: Foundation Models and In Silico Clinical Trials for Virtual Biopsy-based Risk Assessment 23 May 2025 · 0 repositories · arXiv:2505.17971
-
FinRAGBench-V: A Benchmark for Multimodal RAG with Visual Citation in the Financial Domain 23 May 2025 · 0 repositories · arXiv:2505.17471
-
FlashForge: Ultra-Efficient Prefix-Aware Attention for LLM Decoding 23 May 2025 · 0 repositories · arXiv:2505.17694
-
FreqU-FNet: Frequency-Aware U-Net for Imbalanced Medical Image Segmentation 23 May 2025 · 0 repositories · arXiv:2505.17544
-
Gaming Tool Preferences in Agentic LLMs 23 May 2025 · 1 repository · arXiv:2505.18135
-
Hybrid Mamba-Transformer Decoder for Error-Correcting Codes 23 May 2025 · 0 repositories · arXiv:2505.17834
-
Is It Bad to Work All the Time? Cross-Cultural Evaluation of Social Norm Biases in GPT-4 23 May 2025 · 0 repositories · arXiv:2505.18322
-
LLM assisted web application functional requirements generation: A case study of four popular LLMs over a Mess Management System 23 May 2025 · 0 repositories · arXiv:2505.18019
-
Model Editing with Graph-Based External Memory 23 May 2025 · 0 repositories · arXiv:2505.18343
-
Multi-Scale Probabilistic Generation Theory: A Hierarchical Framework for Interpreting Large Language Models 23 May 2025 · 0 repositories · arXiv:2505.18244
-
Object-level Cross-view Geo-localization with Location Enhancement and Multi-Head Cross Attention 23 May 2025 · 1 repository · arXiv:2505.17911
-
One Model Transfer to All: On Robust Jailbreak Prompts Generation against LLMs 23 May 2025 · 1 repository · arXiv:2505.17598
-
Online Statistical Inference of Constrained Stochastic Optimization via Random Scaling 23 May 2025 · 0 repositories · arXiv:2505.18327
-
QwenLong-CPRS: Towards ∞-LLMs with Dynamic Context Optimization 23 May 2025 · 0 repositories · arXiv:2505.18092
-
ReqBrain: Task-Specific Instruction Tuning of LLMs for AI-Assisted Requirements Generation 23 May 2025 · 0 repositories · arXiv:2505.17632
-
Resolving Conflicting Evidence in Automated Fact-Checking: A Study on Retrieval-Augmented LLMs 23 May 2025 · 1 repository · arXiv:2505.17762
-
ShIOEnv: A CLI Behavior-Capturing Environment Enabling Grammar-Guided Command Synthesis for Dataset Curation 23 May 2025 · 1 repository · arXiv:2505.18374
-
Slot-MLLM: Object-Centric Visual Tokenization for Multimodal LLM 23 May 2025 · 0 repositories · arXiv:2505.17726
-
Small Models, Smarter Learning: The Power of Joint Task Training 23 May 2025 · 0 repositories · arXiv:2505.18369
-
SpikeGen: Generative Framework for Visual Spike Stream Processing 23 May 2025 · 0 repositories · arXiv:2505.18049
-
Superplatforms Have to Attack AI Agents 23 May 2025 · 0 repositories · arXiv:2505.17861
-
Token Reduction Should Go Beyond Efficiency in Generative Models -- From Vision, Language to Multimodality 23 May 2025 · 1 repository · arXiv:2505.18227
-
TopoPoint: Enhance Topology Reasoning via Endpoint Detection in Autonomous Driving 23 May 2025 · 2 repositories · arXiv:2505.17771
-
Towards more transferable adversarial attack in black-box manner 23 May 2025 · 0 repositories · arXiv:2505.18097
-
Universal Biological Sequence Reranking for Improved De Novo Peptide Sequencing 23 May 2025 · 1 repository · arXiv:2505.17552
-
VEAttack: Downstream-agnostic Vision Encoder Attack against Large Vision Language Models 23 May 2025 · 1 repository · arXiv:2505.17440
-
A Multi-Head Attention Soft Random Forest for Interpretable Patient No-Show Prediction 22 May 2025 · 0 repositories · arXiv:2505.17344