Methods › General › Attention Mechanisms › Attention › Papers, page 4
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 4 of 316: papers 301 to 400 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Attention-Aided MMSE for OFDM Channel Estimation: Learning Linear Filters with Attention 31 May 2025 · 0 repositories · arXiv:2506.00452
-
Blockchain-Enabled Privacy-Preserving Second-Order Federated Edge Learning in Personalized Healthcare 31 May 2025 · 0 repositories · arXiv:2506.00416
-
Channel-Imposed Fusion: A Simple yet Effective Method for Medical Time Series Classification 31 May 2025 · 0 repositories · arXiv:2506.00337
-
Evaluating Robot Policies in a World Model 31 May 2025 · 0 repositories · arXiv:2506.00613
-
FinBERT2: A Specialized Bidirectional Encoder for Bridging the Gap in Finance-Specific Deployment of Large Language Models 31 May 2025 · 0 repositories · arXiv:2506.06335
-
Machine vs Machine: Using AI to Tackle Generative AI Threats in Assessment 31 May 2025 · 0 repositories · arXiv:2506.02046
-
Multi-Objective Neural Network Assisted Design Optimization of Soft Fin-Ray Grippers for Enhanced Grasping Performance 31 May 2025 · 0 repositories · arXiv:2506.00494
-
Position: Olfaction Standardization is Essential for the Advancement of Embodied Artificial Intelligence 31 May 2025 · 0 repositories · arXiv:2506.00398
-
Power-of-Two (PoT) Weights in Large Language Models (LLMs) 31 May 2025 · 0 repositories · arXiv:2506.00315
-
Towards Graph-Based Privacy-Preserving Federated Learning: ModelNet -- A ResNet-based Model Classification Dataset 31 May 2025 · 0 repositories · arXiv:2506.00476
-
Translate With Care: Addressing Gender Bias, Neutrality, and Reasoning in Large Language Model Translations 31 May 2025 · 1 repository · arXiv:2506.00748
-
Using Diffusion Ensembles to Estimate Uncertainty for End-to-End Autonomous Driving 31 May 2025 · 0 repositories · arXiv:2506.00560
-
Adaptive LoRA Merge with Parameter Pruning for Low-Resource Generation 30 May 2025 · 1 repository · arXiv:2505.24174
-
Adversarial Threat Vectors and Risk Mitigation for Retrieval-Augmented Generation Systems 30 May 2025 · 0 repositories · arXiv:2506.00281
-
Agent-X: Evaluating Deep Multimodal Reasoning in Vision-Centric Agentic Tasks 30 May 2025 · 1 repository · arXiv:2505.24876
-
Bayesian Data Sketching for Varying Coefficient Regression Models 30 May 2025 · 0 repositories · arXiv:2506.00270
-
Cloud Optical Thickness Retrievals Using Angle Invariant Attention Based Deep Learning Models 30 May 2025 · 0 repositories · arXiv:2505.24638
-
ClueAnchor: Clue-Anchored Knowledge Reasoning Exploration and Optimization for Retrieval-Augmented Generation 30 May 2025 · 1 repository · arXiv:2505.24388
-
Cross-Attention Speculative Decoding 30 May 2025 · 0 repositories · arXiv:2505.24544
-
D2AF: A Dual-Driven Annotation and Filtering Framework for Visual Grounding 30 May 2025 · 0 repositories · arXiv:2505.24372
-
Decoding Knowledge Attribution in Mixture-of-Experts: A Framework of Basic-Refinement Collaboration and Efficiency Analysis 30 May 2025 · 0 repositories · arXiv:2505.24593
-
Deformable Attention Mechanisms Applied to Object Detection, case of Remote Sensing 30 May 2025 · 0 repositories · arXiv:2505.24489
-
Efficient Text Encoders for Labor Market Analysis 30 May 2025 · 0 repositories · arXiv:2505.24640
-
Explainable Depression Detection using Masked Hard Instance Mining 30 May 2025 · 0 repositories · arXiv:2505.24609
-
From Hallucinations to Jailbreaks: Rethinking the Vulnerability of Large Foundation Models 30 May 2025 · 0 repositories · arXiv:2505.24232
-
HELM: Hyperbolic Large Language Models via Mixture-of-Curvature Experts 30 May 2025 · 1 repository · arXiv:2505.24722Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 7 harvested samples)
-
Interactive Video Generation via Domain Adaptation 30 May 2025 · 0 repositories · arXiv:2505.24253
-
Interpretable phenotyping of Heart Failure patients with Dutch discharge letters 30 May 2025 · 0 repositories · arXiv:2505.24619
-
Interpreting Large Text-to-Image Diffusion Models with Dictionary Learning 30 May 2025 · 1 repository · arXiv:2505.24360
-
Large Language Models are Locally Linear Mappings 30 May 2025 · 1 repository · arXiv:2505.24293
-
Leveraging Intermediate Features of Vision Transformer for Face Anti-Spoofing 30 May 2025 · 0 repositories · arXiv:2505.24402
-
Lightweight Relational Embedding in Task-Interpolated Few-Shot Networks for Enhanced Gastrointestinal Disease Classification 30 May 2025 · 0 repositories · arXiv:2505.24792
-
LPASS: Linear Probes as Stepping Stones for vulnerability detection using compressed LLMs 30 May 2025 · 0 repositories · arXiv:2505.24451
-
Mamba Knockout for Unraveling Factual Information Flow 30 May 2025 · 1 repository · arXiv:2505.24244Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 1 pointer-only (licence)
-
Mastering Massive Multi-Task Reinforcement Learning via Mixture-of-Expert Decision Transformer 30 May 2025 · 1 repository · arXiv:2505.24378
-
Model-Guided Network with Cluster-Based Operators for Spatio-Spectral Super-Resolution 30 May 2025 · 1 repository · arXiv:2505.24605
-
MOFGPT: Generative Design of Metal-Organic Frameworks using Language Models 30 May 2025 · 1 repository · arXiv:2506.00198
-
PCIE_Pose Solution for EgoExo4D Pose and Proficiency Estimation Challenge 30 May 2025 · 0 repositories · arXiv:2505.24411
-
PersianMedQA: Language-Centric Evaluation of LLMs in the Persian Medical Domain 30 May 2025 · 0 repositories · arXiv:2506.00250
-
RealDrive: Retrieval-Augmented Driving with Diffusion Models 30 May 2025 · 0 repositories · arXiv:2505.24808
-
ReCalKV: Low-Rank KV Cache Compression via Head Reordering and Offline Calibration 30 May 2025 · 1 repository · arXiv:2505.24357
-
S3CE-Net: Spike-guided Spatiotemporal Semantic Coupling and Expansion Network for Long Sequence Event Re-Identification 30 May 2025 · 1 repository · arXiv:2505.24401
-
SALE : Low-bit Estimation for Efficient Sparse Attention in Long-context LLM Prefilling 30 May 2025 · 1 repository · arXiv:2505.24179
-
SPPSFormer: High-quality Superpoint-based Transformer for Roof Plane Instance Segmentation from Point Clouds 30 May 2025 · 0 repositories · arXiv:2505.24475
-
STAR-Net: An Interpretable Model-Aided Network for Remote Sensing Image Denoising 30 May 2025 · 1 repository · arXiv:2505.24327
-
The Hype Index: an NLP-driven Measure of Market News Attention 30 May 2025 · 0 repositories · arXiv:2506.06329
-
Transformers Are Universally Consistent 30 May 2025 · 0 repositories · arXiv:2505.24531
-
Two failure modes of deep transformers and how to avoid them: a unified theory of signal propagation at initialisation 30 May 2025 · 0 repositories · arXiv:2505.24333
-
UniGeo: Taming Video Diffusion for Unified Consistent Geometry Estimation 30 May 2025 · 0 repositories · arXiv:2505.24521
-
Visual Embodied Brain: Let Multimodal Large Language Models See, Think, and Control in Spaces 30 May 2025 · 0 repositories · arXiv:2506.00123
-
When GPT Spills the Tea: Comprehensive Assessment of Knowledge File Leakage in GPTs 30 May 2025 · 0 repositories · arXiv:2506.00197
-
Deep Learning-Based Breast Cancer Detection in Mammography: A Multi-Center Validation Study in Thai Population 29 May 2025 · 0 repositories · arXiv:2506.03177
-
Accelerated Training of Federated Learning via Second-Order Methods 29 May 2025 · 0 repositories · arXiv:2505.23588
-
Adversarial Semantic and Label Perturbation Attack for Pedestrian Attribute Recognition 29 May 2025 · 2 repositories · arXiv:2505.23313
-
AnchorAttention: Difference-Aware Sparse Attention with Stripe Granularity 29 May 2025 · 1 repository · arXiv:2505.23520
-
Argus: Vision-Centric Reasoning with Grounded Chain-of-Thought 29 May 2025 · 0 repositories · arXiv:2505.23766
-
ATLAS: Learning to Optimally Memorize the Context at Test Time 29 May 2025 · 0 repositories · arXiv:2505.23735
-
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds 29 May 2025 · 0 repositories · arXiv:2505.23673
-
Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time 29 May 2025 · 0 repositories · arXiv:2505.23729
-
Bridging Geometric and Semantic Foundation Models for Generalized Monocular Depth Estimation 29 May 2025 · 0 repositories · arXiv:2505.23400
-
Bridging the Gap Between Semantic and User Preference Spaces for Multi-modal Music Representation Learning 29 May 2025 · 0 repositories · arXiv:2505.23298
-
CF-DETR: Coarse-to-Fine Transformer for Real-Time Object Detection 29 May 2025 · 0 repositories · arXiv:2505.23317
-
Characterizing the Expressivity of Transformer Language Models 29 May 2025 · 0 repositories · arXiv:2505.23623
-
CLaC at SemEval-2025 Task 6: A Multi-Architecture Approach for Corporate Environmental Promise Verification 29 May 2025 · 0 repositories · arXiv:2505.23538
-
CLIP-AE: CLIP-assisted Cross-view Audio-Visual Enhancement for Unsupervised Temporal Action Localization 29 May 2025 · 0 repositories · arXiv:2505.23524
-
Cora: Correspondence-aware image editing using few step diffusion 29 May 2025 · 1 repository · arXiv:2505.23907
-
Critical Batch Size Revisited: A Simple Empirical Approach to Large-Batch Language Model Training 29 May 2025 · 0 repositories · arXiv:2505.23971
-
DA-VPT: Semantic-Guided Visual Prompt Tuning for Vision Transformers 29 May 2025 · 1 repository · arXiv:2505.23694Syntology official (archive's flag): 4 ran · 4 ran (of which 4 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; every one of the 4 samples that ran constructed an object rather than computing a result (of 5 harvested samples)
-
Data-efficient Meta-models for Evaluation of Context-based Questions and Answers in LLMs 29 May 2025 · 0 repositories · arXiv:2505.23299
-
DATD3: Depthwise Attention Twin Delayed Deep Deterministic Policy Gradient For Model Free Reinforcement Learning Under Output Feedback Control 29 May 2025 · 0 repositories · arXiv:2505.23857
-
Daunce: Data Attribution through Uncertainty Estimation 29 May 2025 · 0 repositories · arXiv:2505.23223
-
Decom-Renorm-Merge: Model Merging on the Right Space Improves Multitasking 29 May 2025 · 0 repositories · arXiv:2505.23117
-
Deep Modeling and Optimization of Medical Image Classification 29 May 2025 · 1 repository · arXiv:2505.23040
-
Differential Gated Self-Attention 29 May 2025 · 0 repositories · arXiv:2505.24054
-
Dimension-Reduction Attack! Video Generative Models are Experts on Controllable Image Synthesis 29 May 2025 · 0 repositories · arXiv:2505.23325
-
DINO-R1: Incentivizing Reasoning Capability in Vision Foundation Models 29 May 2025 · 0 repositories · arXiv:2505.24025
-
Does Machine Unlearning Truly Remove Model Knowledge? A Framework for Auditing Unlearning in LLMs 29 May 2025 · 0 repositories · arXiv:2505.23270
-
Enhancing LLM-Based Code Generation with Complexity Metrics: A Feedback-Driven Approach 29 May 2025 · 0 repositories · arXiv:2505.23953
-
Equivariant Spherical Transformer for Efficient Molecular Modeling 29 May 2025 · 0 repositories · arXiv:2505.23086
-
Evaluating AI capabilities in detecting conspiracy theories on YouTube 29 May 2025 · 1 repository · arXiv:2505.23570
-
From Images to Signals: Are Large Vision Models Useful for Time Series Analysis? 29 May 2025 · 0 repositories · arXiv:2505.24030
-
Generating Fit Check Videos with a Handheld Camera 29 May 2025 · 0 repositories · arXiv:2505.23886
-
Graph Positional Autoencoders as Self-supervised Learners 29 May 2025 · 0 repositories · arXiv:2505.23345
-
Grounded Reinforcement Learning for Visual Reasoning 29 May 2025 · 1 repository · arXiv:2505.23678
-
Grower-in-the-Loop Interactive Reinforcement Learning for Greenhouse Climate Control 29 May 2025 · 0 repositories · arXiv:2505.23355
-
How Does Response Length Affect Long-Form Factuality 29 May 2025 · 1 repository · arXiv:2505.23295
-
HyperPointFormer: Multimodal Fusion in 3D Space with Dual-Branch Cross-Attention Transformers 29 May 2025 · 1 repository · arXiv:2505.23206
-
Identity resolution of software metadata using Large Language Models 29 May 2025 · 0 repositories · arXiv:2505.23500
-
Improving Time Series Forecasting via Instance-aware Post-hoc Revision 29 May 2025 · 0 repositories · arXiv:2505.23583Syntology 3 ran (of which 3 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; every one of the 3 samples that ran constructed an object rather than computing a result (of 4 harvested samples)
-
Interspeech 2025 URGENT Speech Enhancement Challenge 29 May 2025 · 0 repositories · arXiv:2505.23212
-
KVzip: Query-Agnostic KV Cache Compression with Context Reconstruction 29 May 2025 · 1 repository · arXiv:2505.23416Syntology official (archive's flag): 4 ran · 9 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 5 where Syntology's instrument failed) · 3 unverified (of 12 harvested samples) · 1 pointer-only (licence)
-
LayerPeeler: Autoregressive Peeling for Layer-wise Image Vectorization 29 May 2025 · 0 repositories · arXiv:2505.23740
-
Learning to Regulate: A New Event-Level Dataset of Capital Control Measures 29 May 2025 · 0 repositories · arXiv:2505.23025
-
LeMoRe: Learn More Details for Lightweight Semantic Segmentation 29 May 2025 · 1 repository · arXiv:2505.23093
-
Let's Reason Formally: Natural-Formal Hybrid Reasoning Enhances LLM's Math Capability 29 May 2025 · 0 repositories · arXiv:2505.23703
-
LoLA: Low-Rank Linear Attention With Sparse Caching 29 May 2025 · 0 repositories · arXiv:2505.23666
-
MangoLeafViT: Leveraging Lightweight Vision Transformer with Runtime Augmentation for Efficient Mango Leaf Disease Classification 29 May 2025 · 0 repositories · arXiv:2505.23961
-
Matryoshka Model Learning for Improved Elastic Student Models 29 May 2025 · 0 repositories · arXiv:2505.23337
-
MCFNet: A Multimodal Collaborative Fusion Network for Fine-Grained Semantic Classification 29 May 2025 · 0 repositories · arXiv:2505.23365
-
MCP Safety Training: Learning to Refuse Falsely Benign MCP Exploits using Improved Preference Alignment 29 May 2025 · 0 repositories · arXiv:2505.23634