Methods › General › Attention Mechanisms › Attention › Papers, page 98
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 98 of 316: papers 9,701 to 9,800 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Recommendation with Generative Models 18 Sep 2024 · 0 repositories · arXiv:2409.15173
-
ReFu: Recursive Fusion for Exemplar-Free 3D Class-Incremental Learning 18 Sep 2024 · 0 repositories · arXiv:2409.12326
-
Reinforcement Learning as an Improvement Heuristic for Real-World Production Scheduling 18 Sep 2024 · 0 repositories · arXiv:2409.11933
-
Secure Control Systems for Autonomous Quadrotors against Cyber-Attacks 18 Sep 2024 · 1 repository · arXiv:2409.11897
-
SIM-OFE: Structure Information Mining and Object-aware Feature Enhancement for Fine-Grained Visual Categorization 18 Sep 2024 · 0 repositories
-
TART: An Open-Source Tool-Augmented Framework for Explainable Table-based Reasoning 18 Sep 2024 · 1 repository · arXiv:2409.11724
-
Unsupervised Feature Orthogonalization for Learning Distortion-Invariant Representations 18 Sep 2024 · 1 repository · arXiv:2409.12276
-
User Subgrouping in Scalable Cell-Free Massive MIMO Multicasting Systems 18 Sep 2024 · 0 repositories · arXiv:2409.11871
-
VERA: Validation and Enhancement for Retrieval Augmented systems 18 Sep 2024 · 0 repositories · arXiv:2409.15364
-
WiLoR: End-to-end 3D Hand Localization and Reconstruction in-the-wild 18 Sep 2024 · 1 repository · arXiv:2409.12259
-
Exploring the Trade-Offs: Quantization Methods, Task Difficulty, and Model Size in Large Language Models From Edge to Giant 17 Sep 2024 · 1 repository · arXiv:2409.11055
-
A Unified Framework to Classify Business Activities into International Standard Industrial Classification through Large Language Models for Circular Economy 17 Sep 2024 · 0 repositories · arXiv:2409.18988
-
Adaptive Large Language Models By Layerwise Attention Shortcuts 17 Sep 2024 · 0 repositories · arXiv:2409.10870
-
American Sign Language to Text Translation using Transformer and Seq2Seq with LSTM 17 Sep 2024 · 0 repositories · arXiv:2409.10874
-
Attention-Seeker: Dynamic Self-Attention Scoring for Unsupervised Keyphrase Extraction 17 Sep 2024 · 1 repository · arXiv:2409.10907
-
Beyond LoRA: Exploring Efficient Fine-Tuning Techniques for Time Series Foundational Models 17 Sep 2024 · 0 repositories · arXiv:2409.11302
-
Chain-of-Thought Prompting for Speech Translation 17 Sep 2024 · 0 repositories · arXiv:2409.11538
-
Contrasformer: A Brain Network Contrastive Transformer for Neurodegenerative Condition Identification 17 Sep 2024 · 1 repository · arXiv:2409.10944
-
Cross-lingual transfer of multilingual models on low resource African Languages 17 Sep 2024 · 2 repositories · arXiv:2409.10965
-
D2Vformer: A Flexible Time Series Prediction Model Based on Time Position Embedding 17 Sep 2024 · 0 repositories · arXiv:2409.11024
-
Guess What I Think: Streamlined EEG-to-Image Generation with Latent Diffusion Models 17 Sep 2024 · 1 repository · arXiv:2410.02780Syntology official (archive's flag): 2 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Harnessing AI data-driven global weather models for climate attribution: An analysis of the 2017 Oroville Dam extreme atmospheric river 17 Sep 2024 · 1 repository · arXiv:2409.11605
-
HEARTS: A Holistic Framework for Explainable, Sustainable and Robust Text Stereotype Detection 17 Sep 2024 · 1 repository · arXiv:2409.11579
-
High-Order Evolving Graphs for Enhanced Representation of Traffic Dynamics 17 Sep 2024 · 0 repositories · arXiv:2409.11206
-
A Hybrid Multi-Factor Network with Dynamic Sequence Modeling for Early Warning of Intraoperative Hypotension 17 Sep 2024 · 1 repository · arXiv:2409.11064
-
Implicit Reasoning in Deep Time Series Forecasting 17 Sep 2024 · 0 repositories · arXiv:2409.10840
-
Investigating Context-Faithfulness in Large Language Models: The Roles of Memory Strength and Evidence Style 17 Sep 2024 · 0 repositories · arXiv:2409.10955
-
Large Language Models are Good Multi-lingual Learners : When LLMs Meet Cross-lingual Prompts 17 Sep 2024 · 1 repository · arXiv:2409.11056
-
Learning variant product relationship and variation attributes from e-commerce website structures 17 Sep 2024 · 0 repositories · arXiv:2410.02779
-
Less is More: A Simple yet Effective Token Reduction Method for Efficient Multi-modal LLMs 17 Sep 2024 · 1 repository · arXiv:2409.10994Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 1 violated, 2 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 8 harvested samples) · 1 pointer-only (licence)
-
Linear Recency Bias During Training Improves Transformers' Fit to Reading Times 17 Sep 2024 · 0 repositories · arXiv:2409.11250
-
LOLA -- An Open-Source Massively Multilingual Large Language Model 17 Sep 2024 · 1 repository · arXiv:2409.11272
-
Measuring and Enhancing Trustworthiness of LLMs in RAG through Grounded Attributions and Learning to Refuse 17 Sep 2024 · 1 repository · arXiv:2409.11242Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Multi-Cohort Framework with Cohort-Aware Attention and Adversarial Mutual-Information Minimization for Whole Slide Image Classification 17 Sep 2024 · 0 repositories · arXiv:2409.11119
-
Multi-frequency Electrical Impedance Tomography Reconstruction with Multi-Branch Attention Image Prior 17 Sep 2024 · 0 repositories · arXiv:2409.10794
-
Multimodal Attention-Enhanced Feature Fusion-based Weekly Supervised Anomaly Violence Detection 17 Sep 2024 · 0 repositories · arXiv:2409.11223
-
Multimodality Adaptive Transformer and Mutual Learning for Unsupervised Domain Adaptation Vehicle Re-Identification 17 Sep 2024 · 0 repositories
-
Norm of Mean Contextualized Embeddings Determines their Variance 17 Sep 2024 · 1 repository · arXiv:2409.11253
-
P-RAG: Progressive Retrieval Augmented Generation For Planning on Embodied Everyday Task 17 Sep 2024 · 0 repositories · arXiv:2409.11279
-
Retinal Vessel Segmentation with Deep Graph and Capsule Reasoning 17 Sep 2024 · 0 repositories · arXiv:2409.11508
-
Semformer: Transformer Language Models with Semantic Planning 17 Sep 2024 · 0 repositories · arXiv:2409.11143
-
Shaking the Fake: Detecting Deepfake Videos in Real Time via Active Probes 17 Sep 2024 · 0 repositories · arXiv:2409.10889
-
ShapeAug++: More Realistic Shape Augmentation for Event Data 17 Sep 2024 · 0 repositories · arXiv:2409.11075
-
SkinMamba: A Precision Skin Lesion Segmentation Architecture with Cross-Scale Global State Modeling and Frequency Boundary Guidance 17 Sep 2024 · 1 repository · arXiv:2409.10890
-
Small Language Models can Outperform Humans in Short Creative Writing: A Study Comparing SLMs with Humans and LLMs 17 Sep 2024 · 1 repository · arXiv:2409.11547
-
Sparks of Artificial General Intelligence(AGI) in Semiconductor Material Science: Early Explorations into the Next Frontier of Generative AI-Assisted Electron Micrograph Analysis 17 Sep 2024 · 0 repositories · arXiv:2409.12244
-
THaMES: An End-to-End Tool for Hallucination Mitigation and Evaluation in Large Language Models 17 Sep 2024 · 1 repository · arXiv:2409.11353Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Towards Fair RAG: On the Impact of Fair Ranking in Retrieval-Augmented Generation 17 Sep 2024 · 1 repository · arXiv:2409.11598Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Unleashing the Potential of Mamba: Boosting a LiDAR 3D Sparse Detector by Using Cross-Model Knowledge Distillation 17 Sep 2024 · 0 repositories · arXiv:2409.11018
-
A Green Multi-Attribute Client Selection for Over-The-Air Federated Learning: A Grey-Wolf-Optimizer Approach 16 Sep 2024 · 0 repositories · arXiv:2409.11442
-
A Riemannian Approach to Ground Metric Learning for Optimal Transport 16 Sep 2024 · 0 repositories · arXiv:2409.10085
-
Are Deep Learning Models Robust to Partial Object Occlusion in Visual Recognition Tasks? 16 Sep 2024 · 0 repositories · arXiv:2409.10775
-
ASMA: An Adaptive Safety Margin Algorithm for Vision-Language Drone Navigation via Scene-Aware Control Barrier Functions 16 Sep 2024 · 0 repositories · arXiv:2409.10283
-
AttnMod: Attention-Based New Art Styles 16 Sep 2024 · 0 repositories · arXiv:2409.10028
-
BAFNet: Bilateral Attention Fusion Network for Lightweight Semantic Segmentation of Urban Remote Sensing Images 16 Sep 2024 · 0 repositories · arXiv:2409.10269
-
beeFormer: Bridging the Gap Between Semantic and Interaction Similarity in Recommender Systems 16 Sep 2024 · 1 repository · arXiv:2409.10309
-
Benchmarking Large Language Model Uncertainty for Prompt Optimization 16 Sep 2024 · 1 repository · arXiv:2409.10044
-
Causal Discovery in Recommender Systems: Example and Discussion 16 Sep 2024 · 0 repositories · arXiv:2409.10271
-
Causal Language Modeling Can Elicit Search and Reasoning Capabilities on Logic Puzzles 16 Sep 2024 · 1 repository · arXiv:2409.10502Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified (of 11 harvested samples)
-
CoMamba: Real-time Cooperative Perception Unlocked with State Space Models 16 Sep 2024 · 1 repository · arXiv:2409.10699
-
Context-Conditioned Spatio-Temporal Predictive Learning for Reliable V2V Channel Prediction 16 Sep 2024 · 0 repositories · arXiv:2409.09978
-
Deep Graph Anomaly Detection: A Survey and New Perspectives 16 Sep 2024 · 1 repository · arXiv:2409.09957
-
Deep Learning tools to support deforestation monitoring in the Ivory Coast using SAR and Optical satellite imagery 16 Sep 2024 · 0 repositories · arXiv:2409.11186
-
DILA: Dictionary Label Attention for Mechanistic Interpretability in High-dimensional Multi-label Medical Coding Prediction 16 Sep 2024 · 0 repositories · arXiv:2409.10504
-
Exploring Fine-tuned Generative Models for Keyphrase Selection: A Case Study for Russian 16 Sep 2024 · 0 repositories · arXiv:2409.10640
-
Fit and Prune: Fast and Training-free Visual Token Pruning for Multi-modal Large Language Models 16 Sep 2024 · 1 repository · arXiv:2409.10197Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 1 violated, 2 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
Flash STU: Fast Spectral Transform Units 16 Sep 2024 · 1 repository · arXiv:2409.10489
-
Garment Attribute Manipulation with Multi-level Attention 16 Sep 2024 · 0 repositories · arXiv:2409.10206
-
GPT takes the SAT: Tracing changes in Test Difficulty and Math Performance of Students 16 Sep 2024 · 0 repositories · arXiv:2409.10750
-
Improving Multi-candidate Speculative Decoding 16 Sep 2024 · 1 repository · arXiv:2409.10644Syntology official (archive's flag): 2 ran · 2 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Kolmogorov-Arnold Networks in Low-Data Regimes: A Comparative Study with Multilayer Perceptrons 16 Sep 2024 · 0 repositories · arXiv:2409.10463
-
Kolmogorov-Arnold Transformer 16 Sep 2024 · 1 repository · arXiv:2409.10594Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
Lab-AI: Using Retrieval Augmentation to Enhance Language Models for Personalized Lab Test Interpretation in Clinical Medicine 16 Sep 2024 · 0 repositories · arXiv:2409.18986
-
Learning large softmax mixtures with warm start EM 16 Sep 2024 · 0 repositories · arXiv:2409.09903
-
LLM-DER:A Named Entity Recognition Method Based on Large Language Models for Chinese Coal Chemical Domain 16 Sep 2024 · 0 repositories · arXiv:2409.10077
-
LLMs for clinical risk prediction 16 Sep 2024 · 0 repositories · arXiv:2409.10191
-
MGSA: Multi-Granularity Graph Structure Attention for Knowledge Graph-to-Text Generation 16 Sep 2024 · 0 repositories · arXiv:2409.10294
-
MindGuard: Towards Accessible and Sitgma-free Mental Health First Aid via Edge LLM 16 Sep 2024 · 0 repositories · arXiv:2409.10064
-
Mitigating Partial Observability in Adaptive Traffic Signal Control with Transformers 16 Sep 2024 · 0 repositories · arXiv:2409.10693
-
Model Tells Itself Where to Attend: Faithfulness Meets Automatic Attention Steering 16 Sep 2024 · 0 repositories · arXiv:2409.10790
-
NARX Transformer: A Dynamic Model for Leveraging Multicycle Data in Long-Term Battery State of Health Estimation 16 Sep 2024 · 2 repositories
-
Personalized Speech Emotion Recognition in Human-Robot Interaction using Vision Transformers 16 Sep 2024 · 0 repositories · arXiv:2409.10687
-
Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models 16 Sep 2024 · 1 repository · arXiv:2409.10695Syntology 18 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 4 honoured, 1 violated, 4 with no contract checked; 9 where Syntology's instrument failed) · 3 unverified (of 21 harvested samples) · 21 pointer-only (licence)
-
Predicting Punctuation in Ancient Chinese Texts: A Multi-Layered LSTM and Attention-Based Approach 16 Sep 2024 · 0 repositories · arXiv:2409.10783
-
Recurrent Graph Transformer Network for Multiple Fault Localization in Naval Shipboard Systems 16 Sep 2024 · 0 repositories · arXiv:2409.10792
-
RetrievalAttention: Accelerating Long-Context LLM Inference via Vector Retrieval 16 Sep 2024 · 1 repository · arXiv:2409.10516
-
SelECT-SQL: Self-correcting ensemble Chain-of-Thought for Text-to-SQL 16 Sep 2024 · 1 repository · arXiv:2409.10007
-
Self-Attention Limits Working Memory Capacity of Transformer-Based Models 16 Sep 2024 · 0 repositories · arXiv:2409.10715
-
Self-Supervised Syllable Discovery Based on Speaker-Disentangled HuBERT 16 Sep 2024 · 1 repository · arXiv:2409.10103
-
SFR-RAG: Towards Contextually Faithful LLMs 16 Sep 2024 · 0 repositories · arXiv:2409.09916
-
Signed Graph Autoencoder for Explainable and Polarization-Aware Network Embeddings 16 Sep 2024 · 0 repositories · arXiv:2409.10452
-
TCDformer-based Momentum Transfer Model for Long-term Sports Prediction 16 Sep 2024 · 0 repositories · arXiv:2409.10176
-
The Importance of Causality in Decision Making: A Perspective on Recommender Systems 16 Sep 2024 · 0 repositories · arXiv:2410.01822
-
Trustworthiness in Retrieval-Augmented Generation Systems: A Survey 16 Sep 2024 · 1 repository · arXiv:2409.10102Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Uncertainty-Guided Appearance-Motion Association Network for Out-of-Distribution Action Detection 16 Sep 2024 · 1 repository · arXiv:2409.09953
-
XLM for Autonomous Driving Systems: A Comprehensive Review 16 Sep 2024 · 0 repositories · arXiv:2409.10484
-
A Comprehensive Methodological Survey of Human Activity Recognition Across Divers Data Modalities 15 Sep 2024 · 0 repositories · arXiv:2409.09678
-
Detection Made Easy: Potentials of Large Language Models for Solidity Vulnerabilities 15 Sep 2024 · 0 repositories · arXiv:2409.10574
-
Enhancing Weakly-Supervised Object Detection on Static Images through (Hallucinated) Motion 15 Sep 2024 · 0 repositories · arXiv:2409.09616
-
Entity-Aware Self-Attention and Contextualized GCN for Enhanced Relation Extraction in Long Sentences 15 Sep 2024 · 0 repositories · arXiv:2409.13755