Methods › General › Attention Mechanisms › Attention › Papers, page 79
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 79 of 316: papers 7,801 to 7,900 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Internship Report: Benchmark of Deep Learning-based Imaging PPG in Automotive Domain 1 Nov 2024 · 0 repositories · arXiv:2411.00919
-
LAM-YOLO: Drones-based Small Object Detection on Lighting-Occlusion Attention Mechanism YOLO 1 Nov 2024 · 0 repositories · arXiv:2411.00485
-
LLM-Ref: Enhancing Reference Handling in Technical Writing with Large Language Models 1 Nov 2024 · 0 repositories · arXiv:2411.00294
-
LLMs: A Game-Changer for Software Engineers? 1 Nov 2024 · 0 repositories · arXiv:2411.00932
-
Machine Learning-Accelerated Multi-Objective Design of Fractured Geothermal Systems 1 Nov 2024 · 1 repository · arXiv:2411.00504
-
Multiple Information Prompt Learning for Cloth-Changing Person Re-Identification 1 Nov 2024 · 0 repositories · arXiv:2411.00330
-
MV-Adapter: Enhancing Underwater Instance Segmentation via Adaptive Channel Attention 1 Nov 2024 · 0 repositories · arXiv:2411.00472
-
Provenance: A Light-weight Fact-checker for Retrieval Augmented LLM Generation Output 1 Nov 2024 · 0 repositories · arXiv:2411.01022
-
Rationale-Guided Retrieval Augmented Generation for Medical Question Answering 1 Nov 2024 · 1 repository · arXiv:2411.00300Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Self-Evolved Reward Learning for LLMs 1 Nov 2024 · 1 repository · arXiv:2411.00418Syntology 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample)
-
STAA: Spatio-Temporal Attention Attribution for Real-Time Interpreting Transformer-based Video Models 1 Nov 2024 · 1 repository · arXiv:2411.00630
-
Target-Guided Adversarial Point Cloud Transformer Towards Recognition Against Real-world Corruptions 1 Nov 2024 · 1 repository · arXiv:2411.00462Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Towards High-fidelity Head Blending with Chroma Keying for Industrial Applications 1 Nov 2024 · 0 repositories · arXiv:2411.00652
-
Towards Multi-Source Retrieval-Augmented Generation via Synergizing Reasoning and Preference-Driven Retrieval 1 Nov 2024 · 0 repositories · arXiv:2411.00689
-
Tumor Location-weighted MRI-Report Contrastive Learning: A Framework for Improving the Explainability of Pediatric Brain Tumor Diagnosis 1 Nov 2024 · 0 repositories · arXiv:2411.00609
-
ZIM: Zero-Shot Image Matting for Anything 1 Nov 2024 · 1 repository · arXiv:2411.00626
-
A Multiphysics Analysis and Investigation of Soft Magnetics Effect on IPMSM: Case Study Dynamometer 31 Oct 2024 · 0 repositories · arXiv:2410.24172
-
Ada-MSHyper: Adaptive Multi-Scale Hypergraph Transformer for Time Series Forecasting 31 Oct 2024 · 1 repository · arXiv:2410.23992Syntology official: harvested, nothing ran · 0 ran · 4 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Aerial Flood Scene Classification Using Fine-Tuned Attention-based Architecture for Flood-Prone Countries in South Asia 31 Oct 2024 · 0 repositories · arXiv:2411.00169
-
Analyzing & Reducing the Need for Learning Rate Warmup in GPT Training 31 Oct 2024 · 0 repositories · arXiv:2410.23922
-
RAM: Replace Attention with MLP for Efficient Multivariate Time Series Forecasting 31 Oct 2024 · 0 repositories · arXiv:2410.24023
-
Attention is All You Need to Optimize Wind Farm Operations and Maintenance 31 Oct 2024 · 0 repositories · arXiv:2410.24052
-
Automating Quantum Software Maintenance: Flakiness Detection and Root Cause Analysis 31 Oct 2024 · 0 repositories · arXiv:2410.23578
-
Benchmark Data Repositories for Better Benchmarking 31 Oct 2024 · 0 repositories · arXiv:2410.24100
-
Beyond Label Attention: Transparency in Language Models for Automated Medical Coding via Dictionary Learning 31 Oct 2024 · 0 repositories · arXiv:2411.00173
-
Commonsense Knowledge Editing Based on Free-Text in LLMs 31 Oct 2024 · 1 repository · arXiv:2410.23844
-
Context-Aware Token Selection and Packing for Enhanced Vision Transformer 31 Oct 2024 · 0 repositories · arXiv:2410.23608
-
DC-Spin: A Speaker-invariant Speech Tokenizer for Spoken Language Models 31 Oct 2024 · 0 repositories · arXiv:2410.24177
-
Deep Convolutional Neural Networks on Multiclass Classification of Three-Dimensional Brain Images for Parkinson's Disease Stage Prediction 31 Oct 2024 · 0 repositories · arXiv:2410.23649
-
Deep Learning in Long-Short Stock Portfolio Allocation: An Empirical Study 31 Oct 2024 · 0 repositories · arXiv:2411.13555
-
Deep Learning with HM-VGG: AI Strategies for Multi-modal Image Analysis 31 Oct 2024 · 0 repositories · arXiv:2410.24046
-
DELTA: Dense Efficient Long-range 3D Tracking for any video 31 Oct 2024 · 0 repositories · arXiv:2410.24211
-
Desert Camels and Oil Sheikhs: Arab-Centric Red Teaming of Frontier LLMs 31 Oct 2024 · 0 repositories · arXiv:2410.24049
-
DIP: Diffusion Learning of Inconsistency Pattern for General DeepFake Detection 31 Oct 2024 · 0 repositories · arXiv:2410.23663
-
EDT: An Efficient Diffusion Transformer Framework Inspired by Human-like Sketching 31 Oct 2024 · 1 repository · arXiv:2410.23788Syntology official (archive's flag): 10 ran · 12 ran (of which 7 constructed an object rather than computing a result; 9 with no instrument failure: 1 honoured, 0 violated, 8 with no contract checked; 3 where Syntology's instrument failed) · 3 unverified (of 15 harvested samples) · 2 pointer-only (licence)
-
Enhancing Brain Tumor Classification Using TrAdaBoost and Multi-Classifier Deep Learning Approaches 31 Oct 2024 · 0 repositories · arXiv:2411.00875
-
Enhancing Chess Reinforcement Learning with Graph Representation 31 Oct 2024 · 1 repository · arXiv:2410.23753
-
Global bifurcation in a virus, defective genomes, satellite RNAs tripartite system: breakdown of a coexistence quasi-neutral curve 31 Oct 2024 · 0 repositories · arXiv:2411.00070
-
Handwriting Recognition in Historical Documents with Multimodal LLM 31 Oct 2024 · 0 repositories · arXiv:2410.24034
-
Human Action Recognition (HAR) Using Skeleton-based Spatial Temporal Relative Transformer Network: ST-RTR 31 Oct 2024 · 0 repositories · arXiv:2410.23806
-
In-Context LoRA for Diffusion Transformers 31 Oct 2024 · 1 repository · arXiv:2410.23775
-
IO Transformer: Evaluating SwinV2-Based Reward Models for Computer Vision 31 Oct 2024 · 0 repositories · arXiv:2411.00252
-
JEMA: A Joint Embedding Framework for Scalable Co-Learning with Multimodal Alignment 31 Oct 2024 · 0 repositories · arXiv:2410.23988
-
JudgeRank: Leveraging Large Language Models for Reasoning-Intensive Reranking 31 Oct 2024 · 0 repositories · arXiv:2411.00142
-
Language-Assisted Skeleton Action Understanding for Skeleton-Based Temporal Action Segmentation 31 Oct 2024 · 1 repository
-
Large Language Models for Patient Comments Multi-Label Classification 31 Oct 2024 · 0 repositories · arXiv:2410.23528
-
LEAF: Learning and Evaluation Augmented by Fact-Checking to Improve Factualness in Large Language Models 31 Oct 2024 · 0 repositories · arXiv:2410.23526
-
Learning Low-Dimensional Strain Models of Soft Robots by Looking at the Evolution of Their Shape with Application to Model-Based Control 31 Oct 2024 · 0 repositories · arXiv:2411.00138
-
LLM4Mat-Bench: Benchmarking Large Language Models for Materials Property Prediction 31 Oct 2024 · 1 repository · arXiv:2411.00177Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Localization, balance and affinity: a stronger multifaceted collaborative salient object detector in remote sensing images 31 Oct 2024 · 0 repositories · arXiv:2410.23991
-
LSEAttention is All You Need for Time Series Forecasting 31 Oct 2024 · 0 repositories · arXiv:2410.23749
-
MLLA-UNet: Mamba-like Linear Attention in an Efficient U-Shape Model for Medical Image Segmentation 31 Oct 2024 · 1 repository · arXiv:2410.23738
-
NIMBA: Towards Robust and Principled Processing of Point Clouds With SSMs 31 Oct 2024 · 0 repositories · arXiv:2411.00151
-
On Learning Multi-Modal Forgery Representation for Diffusion Generated Video Detection 31 Oct 2024 · 1 repository · arXiv:2410.23623Syntology official (archive's flag): 20 ran · 20 ran (of which 0 constructed an object rather than computing a result; 19 with no instrument failure: 0 honoured, 1 violated, 18 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (of 24 harvested samples)
-
On Positional Bias of Faithfulness for Long-form Summarization 31 Oct 2024 · 1 repository · arXiv:2410.23609
-
One Sample Fits All: Approximating All Probabilistic Values Simultaneously and Efficiently 31 Oct 2024 · 1 repository · arXiv:2410.23808Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Phrase Decoupling Cross-Modal Hierarchical Matching and Progressive Position Correction for Visual Grounding 31 Oct 2024 · 1 repository · arXiv:2410.23570
-
Reasons and Solutions for the Decline in Model Performance after Editing 31 Oct 2024 · 1 repository · arXiv:2410.23843Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Reinforcement Learning Gradients as Vitamin for Online Finetuning Decision Transformers 31 Oct 2024 · 1 repository · arXiv:2410.24108Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
ResiDual Transformer Alignment with Spectral Decomposition 31 Oct 2024 · 0 repositories · arXiv:2411.00246
-
Responsible Retrieval Augmented Generation for Climate Decision Making from Documents 31 Oct 2024 · 0 repositories · arXiv:2410.23902
-
RSL-SQL: Robust Schema Linking in Text-to-SQL Generation 31 Oct 2024 · 1 repository · arXiv:2411.00073Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
SelfCodeAlign: Self-Alignment for Code Generation 31 Oct 2024 · 2 repositories · arXiv:2410.24198Syntology official (archive's flag): 9 ran · 30 ran (of which 3 constructed an object rather than computing a result; 22 with no instrument failure: 1 honoured, 0 violated, 21 with no contract checked; 8 where Syntology's instrument failed) · 7 unverified (of 37 harvested samples)
-
The Importance of Being Scalable: Improving the Speed and Accuracy of Neural Network Interatomic Potentials Across Chemical Domains 31 Oct 2024 · 1 repository · arXiv:2410.24169Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 8 harvested samples)
-
Towards Generative Ray Path Sampling for Faster Point-to-Point Ray Tracing 31 Oct 2024 · 1 repository · arXiv:2410.23773
-
ViT-LCA: A Neuromorphic Approach for Vision Transformers 31 Oct 2024 · 0 repositories · arXiv:2411.00140
-
Weight decay induces low-rank attention layers 31 Oct 2024 · 0 repositories · arXiv:2410.23819
-
A Comprehensive Study on Quantization Techniques for Large Language Models 30 Oct 2024 · 0 repositories · arXiv:2411.02530
-
A Transformer Model for Segmentation, Classification, and Caller Identification of Marmoset Vocalization 30 Oct 2024 · 0 repositories · arXiv:2410.23279
-
Dataset Awareness is not Enough: Implementing Sample-level Tail Encouragement in Long-tailed Self-supervised Learning 30 Oct 2024 · 0 repositories · arXiv:2410.22883
-
An Individual Identity-Driven Framework for Animal Re-Identification 30 Oct 2024 · 1 repository · arXiv:2410.22927
-
Backdoor Attack Against Vision Transformers via Attention Gradient-Based Image Erosion 30 Oct 2024 · 0 repositories · arXiv:2410.22678
-
CLIPErase: Efficient Unlearning of Visual-Textual Associations in CLIP 30 Oct 2024 · 0 repositories · arXiv:2410.23330
-
CORAL: Benchmarking Multi-turn Conversational Retrieval-Augmentation Generation 30 Oct 2024 · 1 repository · arXiv:2410.23090
-
Danoliteracy of Generative, Large Language Models 30 Oct 2024 · 0 repositories · arXiv:2410.22839
-
DisenTS: Disentangled Channel Evolving Pattern Modeling for Multivariate Time Series Forecasting 30 Oct 2024 · 0 repositories · arXiv:2410.22981
-
Don't Just Pay Attention, PLANT It: Transfer L2R Models to Fine-tune Attention in Extreme Multi-Label Text Classification 30 Oct 2024 · 0 repositories · arXiv:2410.23066
-
EchoFM: Foundation Model for Generalizable Echocardiogram Analysis 30 Oct 2024 · 1 repository · arXiv:2410.23413Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 1 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
Eliciting Critical Reasoning in Retrieval-Augmented Language Models via Contrastive Explanations 30 Oct 2024 · 0 repositories · arXiv:2410.22874
-
Emergence of Human-Like Attention in Self-Supervised Vision Transformers: an eye-tracking study 30 Oct 2024 · 1 repository · arXiv:2410.22768
-
Emergence of meta-stable clustering in mean-field transformer models 30 Oct 2024 · 0 repositories · arXiv:2410.23228
-
Emotional RAG: Enhancing Role-Playing Agents through Emotional Retrieval 30 Oct 2024 · 1 repository · arXiv:2410.23041
-
Epipolar-Free 3D Gaussian Splatting for Generalizable Novel View Synthesis 30 Oct 2024 · 0 repositories · arXiv:2410.22817
-
EvoCodeBench: An Evolving Code Generation Benchmark with Domain-Specific Evaluations 30 Oct 2024 · 0 repositories · arXiv:2410.22821Syntology 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 3 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
FilterViT and DropoutViT 30 Oct 2024 · 0 repositories · arXiv:2410.22709
-
FlexTSF: A Universal Forecasting Model for Time Series with Variable Regularities 30 Oct 2024 · 1 repository · arXiv:2410.23160
-
HelloMeme: Integrating Spatial Knitting Attentions to Embed High-Level and Fidelity-Rich Conditions in Diffusion Models 30 Oct 2024 · 1 repository · arXiv:2410.22901
-
High-Fidelity Document Stain Removal via A Large-Scale Real-World Dataset and A Memory-Augmented Transformer 30 Oct 2024 · 1 repository · arXiv:2410.22922
-
Higher-order Cross-structural Embedding Model for Time Series Analysis 30 Oct 2024 · 0 repositories · arXiv:2410.22984
-
HijackRAG: Hijacking Attacks against Retrieval-Augmented Large Language Models 30 Oct 2024 · 0 repositories · arXiv:2410.22832
-
The Sample Complexity of Learning Lipschitz Operators with respect to Gaussian Measures 30 Oct 2024 · 0 repositories · arXiv:2410.23440
-
Learning to Achieve Goals with Belief State Transformers 30 Oct 2024 · 0 repositories · arXiv:2410.23506
-
Lina-Speech: Gated Linear Attention is a Fast and Parameter-Efficient Learner for text-to-speech synthesis 30 Oct 2024 · 1 repository · arXiv:2410.23320
-
LoFLAT: Local Feature Matching using Focused Linear Attention Transformer 30 Oct 2024 · 0 repositories · arXiv:2410.22710
-
MoLE: Enhancing Human-centric Text-to-image Diffusion via Mixture of Low-rank Experts 30 Oct 2024 · 0 repositories · arXiv:2410.23332
-
Neural Attention Field: Emerging Point Relevance in 3D Scenes for One-Shot Dexterous Grasping 30 Oct 2024 · 0 repositories · arXiv:2410.23039
-
NMformer: A Transformer for Noisy Modulation Classification in Wireless Communication 30 Oct 2024 · 1 repository · arXiv:2411.02428
-
ProTransformer: Robustify Transformers via Plug-and-Play Paradigm 30 Oct 2024 · 1 repository · arXiv:2410.23182
-
Retrieval-Augmented Generation with Estimation of Source Reliability 30 Oct 2024 · 0 repositories · arXiv:2410.22954
-
Return Augmented Decision Transformer for Off-Dynamics Reinforcement Learning 30 Oct 2024 · 0 repositories · arXiv:2410.23450