Methods › General › Attention Mechanisms › Attention › Papers, page 91
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 91 of 316: papers 9,001 to 9,100 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Large Language Model Inference Acceleration: A Comprehensive Hardware Perspective 6 Oct 2024 · 1 repository · arXiv:2410.04466
-
Large Language Models for Knowledge-Free Network Management: Feasibility Study and Opportunities 6 Oct 2024 · 0 repositories · arXiv:2410.17259
-
MindScope: Exploring cognitive biases in large language models through Multi-Agent Systems 6 Oct 2024 · 1 repository · arXiv:2410.04452
-
ProtocoLLM: Automatic Evaluation Framework of LLMs on Domain-Specific Scientific Protocol Formulation Tasks 6 Oct 2024 · 0 repositories · arXiv:2410.04601
-
Putting Gale & Shapley to Work: Guaranteeing Stability Through Learning 6 Oct 2024 · 0 repositories · arXiv:2410.04376
-
Social Choice for Heterogeneous Fairness in Recommendation 6 Oct 2024 · 0 repositories · arXiv:2410.04551
-
TimeBridge: Non-Stationarity Matters for Long-term Time Series Forecasting 6 Oct 2024 · 1 repository · arXiv:2410.04442Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 10 harvested samples) · 3 pointer-only (licence)
-
Video Summarization Techniques: A Comprehensive Review 6 Oct 2024 · 0 repositories · arXiv:2410.04449
-
VISTA: A Visual and Textual Attention Dataset for Interpreting Multimodal Models 6 Oct 2024 · 0 repositories · arXiv:2410.04609
-
Wrong-of-Thought: An Integrated Reasoning Framework with Multi-Perspective Verification and Wrong Information 6 Oct 2024 · 1 repository · arXiv:2410.04463Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Applying Hybrid Graph Neural Networks to Strengthen Credit Risk Analysis 5 Oct 2024 · 0 repositories · arXiv:2410.04283
-
Assessing the Performance of Human-Capable LLMs -- Are LLMs Coming for Your Job? 5 Oct 2024 · 0 repositories · arXiv:2410.16285
-
Beyond Language: Applying MLX Transformers to Engineering Physics 5 Oct 2024 · 1 repository · arXiv:2410.04167
-
Can the Variation of Model Weights be used as a Criterion for Self-Paced Multilingual NMT? 5 Oct 2024 · 0 repositories · arXiv:2410.04147
-
Correlation-Aware Select and Merge Attention for Efficient Fine-Tuning and Context Length Extension 5 Oct 2024 · 0 repositories · arXiv:2410.04211
-
Cross Resolution Encoding-Decoding For Detection Transformers 5 Oct 2024 · 1 repository · arXiv:2410.04088
-
DB-SAM: Delving into High Quality Universal Medical Image Segmentation 5 Oct 2024 · 1 repository · arXiv:2410.04172
-
Deep Transfer Learning Based Peer Review Aggregation and Meta-review Generation for Scientific Articles 5 Oct 2024 · 0 repositories · arXiv:2410.04202
-
DeFoG: Discrete Flow Matching for Graph Generation 5 Oct 2024 · 1 repository · arXiv:2410.04263Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 1 honoured, 0 violated, 10 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 15 harvested samples) · 1 pointer-only (licence)
-
Designing Concise ConvNets with Columnar Stages 5 Oct 2024 · 0 repositories · arXiv:2410.04089
-
ECon: On the Detection and Resolution of Evidence Conflicts 5 Oct 2024 · 1 repository · arXiv:2410.04068
-
Efficient Large-Scale Urban Parking Prediction: Graph Coarsening Based on Real-Time Parking Service Capability 5 Oct 2024 · 0 repositories · arXiv:2410.04022
-
Equivariant Neural Functional Networks for Transformers 5 Oct 2024 · 0 repositories · arXiv:2410.04209Syntology 4 ran (of which 3 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples)
-
Fundamental Limitations on Subquadratic Alternatives to Transformers 5 Oct 2024 · 0 repositories · arXiv:2410.04271
-
Gamified crowd-sourcing of high-quality data for visual fine-tuning 5 Oct 2024 · 0 repositories · arXiv:2410.04038
-
Is Score Matching Suitable for Estimating Point Processes? 5 Oct 2024 · 1 repository · arXiv:2410.04037
-
Metadata-based Data Exploration with Retrieval-Augmented Generation for Large Language Models 5 Oct 2024 · 0 repositories · arXiv:2410.04231
-
Multimodal Large Language Models for Inverse Molecular Design with Retrosynthetic Planning 5 Oct 2024 · 1 repository · arXiv:2410.04223Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples) · 1 pointer-only (licence)
-
On the Sample Complexity of a Policy Gradient Algorithm with Occupancy Approximation for General Utility Reinforcement Learning 5 Oct 2024 · 0 repositories · arXiv:2410.04108
-
Optimizing Medical Image Segmentation with Advanced Decoder Design 5 Oct 2024 · 1 repository · arXiv:2410.04128
-
Persona Knowledge-Aligned Prompt Tuning Method for Online Debate 5 Oct 2024 · 1 repository · arXiv:2410.04239
-
PsFuture: A Pseudo-Future-based Zero-Shot Adaptive Policy for Simultaneous Machine Translation 5 Oct 2024 · 0 repositories · arXiv:2410.04075
-
Self-Supervised Anomaly Detection in the Wild: Favor Joint Embeddings Methods 5 Oct 2024 · 0 repositories · arXiv:2410.04289
-
Sinc Kolmogorov-Arnold Network and Its Applications on Physics-informed Neural Networks 5 Oct 2024 · 1 repository · arXiv:2410.04096
-
Take It Easy: Label-Adaptive Self-Rationalization for Fact Verification and Explanation Generation 5 Oct 2024 · 1 repository · arXiv:2410.04002Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Unveiling the Impact of Local Homophily on GNN Fairness: In-Depth Analysis and New Benchmarks 5 Oct 2024 · 0 repositories · arXiv:2410.04287
-
WAVE-UNET: Wavelength based Image Reconstruction method using attention UNET for OCT images 5 Oct 2024 · 0 repositories · arXiv:2410.04123
-
Action Selection Learning for Multi-label Multi-view Action Recognition 4 Oct 2024 · 1 repository · arXiv:2410.03302
-
Adaptive Masking Enhances Visual Grounding 4 Oct 2024 · 0 repositories · arXiv:2410.03161
-
An X-Ray Is Worth 15 Features: Sparse Autoencoders for Interpretable Radiology Report Generation 4 Oct 2024 · 0 repositories · arXiv:2410.03334
-
Audio-Agent: Leveraging LLMs For Audio Generation, Editing and Composition 4 Oct 2024 · 0 repositories · arXiv:2410.03335
-
Auto-GDA: Automatic Domain Adaptation for Efficient Grounding Verification in Retrieval Augmented Generation 4 Oct 2024 · 0 repositories · arXiv:2410.03461
-
Autoregressive Action Sequence Learning for Robotic Manipulation 4 Oct 2024 · 1 repository · arXiv:2410.03132Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 0 honoured, 1 violated, 12 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 14 harvested samples) · 14 pointer-only (licence)
-
Autoregressive Moving-average Attention Mechanism for Time Series Forecasting 4 Oct 2024 · 1 repository · arXiv:2410.03159Syntology official (archive's flag): 6 ran · 6 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 5 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
Benchmarking the Fidelity and Utility of Synthetic Relational Data 4 Oct 2024 · 1 repository · arXiv:2410.03411
-
Beyond Film Subtitles: Is YouTube the Best Approximation of Spoken Vocabulary? 4 Oct 2024 · 1 repository · arXiv:2410.03240
-
Can Mamba Always Enjoy the "Free Lunch"? 4 Oct 2024 · 0 repositories · arXiv:2410.03810
-
Crafting Narrative Closures: Zero-Shot Learning with SSM Mamba for Short Story Ending Generation 4 Oct 2024 · 0 repositories · arXiv:2410.10848
-
Cross-lingual Transfer for Automatic Question Generation by Learning Interrogative Structures in Target Languages 4 Oct 2024 · 0 repositories · arXiv:2410.03197
-
Detecting Machine-Generated Long-Form Content with Latent-Space Variables 4 Oct 2024 · 0 repositories · arXiv:2410.03856
-
DOTS: Learning to Reason Dynamically in LLMs via Optimal Reasoning Trajectories Search 4 Oct 2024 · 1 repository · arXiv:2410.03864Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 6 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Dynamic Diffusion Transformer 4 Oct 2024 · 2 repositories · arXiv:2410.03456Syntology official (archive's flag): 16 ran · 16 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 5 honoured, 0 violated, 5 with no contract checked; 6 where Syntology's instrument failed) · 3 unverified (of 19 harvested samples) · 19 pointer-only (licence)
-
Error Correction Code Transformer: From Non-Unified to Unified 4 Oct 2024 · 0 repositories · arXiv:2410.03364
-
Explaining the (Not So) Obvious: Simple and Fast Explanation of STAN, a Next Point of Interest Recommendation System 4 Oct 2024 · 1 repository · arXiv:2410.03841
-
From Epilepsy Seizures Classification to Detection: A Deep Learning-based Approach for Raw EEG Signals 4 Oct 2024 · 0 repositories · arXiv:2410.03385
-
How Discrete and Continuous Diffusion Meet: Comprehensive Analysis of Discrete Diffusion Models via a Stochastic Integral Framework 4 Oct 2024 · 0 repositories · arXiv:2410.03601
-
How Language Models Prioritize Contextual Grammatical Cues? 4 Oct 2024 · 1 repository · arXiv:2410.03447
-
Learning Semantic Structure through First-Order-Logic Translation 4 Oct 2024 · 0 repositories · arXiv:2410.03203
-
Learning to Balance: Diverse Normalization for Cloth-Changing Person Re-Identification 4 Oct 2024 · 0 repositories · arXiv:2410.03977
-
Linear Transformer Topological Masking with Graph Random Features 4 Oct 2024 · 0 repositories · arXiv:2410.03462
-
Local Attention Mechanism: Boosting the Transformer Architecture for Long-Sequence Time Series Forecasting 4 Oct 2024 · 1 repository · arXiv:2410.03805
-
LoRC: Low-Rank Compression for LLMs KV Cache with a Progressive Compression Strategy 4 Oct 2024 · 0 repositories · arXiv:2410.03111
-
MARE: Multi-Aspect Rationale Extractor on Unsupervised Rationale Extraction 4 Oct 2024 · 0 repositories · arXiv:2410.03531Syntology 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
MELODI: Exploring Memory Compression for Long Contexts 4 Oct 2024 · 1 repository · arXiv:2410.03156
-
Metadata Matters for Time Series: Informative Forecasting with Transformers 4 Oct 2024 · 0 repositories · arXiv:2410.03806
-
Not All Diffusion Model Activations Have Been Evaluated as Discriminative Features 4 Oct 2024 · 1 repository · arXiv:2410.03558Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples)
-
Predictive Coding for Decision Transformer 4 Oct 2024 · 1 repository · arXiv:2410.03408
-
PRF: Parallel Resonate and Fire Neuron for Long Sequence Learning in Spiking Neural Networks 4 Oct 2024 · 0 repositories · arXiv:2410.03530
-
SAG: Style-Aligned Article Generation via Model Collaboration 4 Oct 2024 · 0 repositories · arXiv:2410.03137
-
SDA-GRIN for Adaptive Spatial-Temporal Multivariate Time Series Imputation 4 Oct 2024 · 1 repository · arXiv:2410.03954
-
Selective Transformer for Hyperspectral Image Classification 4 Oct 2024 · 0 repositories · arXiv:2410.03171
-
Steering Large Language Models between Code Execution and Textual Reasoning 4 Oct 2024 · 1 repository · arXiv:2410.03524
-
Still Not Quite There! Evaluating Large Language Models for Comorbid Mental Health Diagnosis 4 Oct 2024 · 0 repositories · arXiv:2410.03908
-
Structured List-Grounded Question Answering 4 Oct 2024 · 0 repositories · arXiv:2410.03950
-
SwiftKV: Fast Prefill-Optimized Inference with Knowledge-Preserving Model Transformation 4 Oct 2024 · 2 repositories · arXiv:2410.03960
-
Towards Linguistically-Aware and Language-Independent Tokenization for Large Language Models (LLMs) 4 Oct 2024 · 0 repositories · arXiv:2410.03568
-
TrustEMG-Net: Using Representation-Masking Transformer with U-Net for Surface Electromyography Enhancement 4 Oct 2024 · 1 repository · arXiv:2410.03843
-
UNComp: Uncertainty-Aware Long-Context Compressor for Efficient Large Language Model Inference 4 Oct 2024 · 0 repositories · arXiv:2410.03090
-
Using Prompts to Guide Large Language Models in Imitating a Real Person's Language Style 4 Oct 2024 · 0 repositories · arXiv:2410.03848
-
Variational Language Concepts for Interpreting Foundation Language Models 4 Oct 2024 · 1 repository · arXiv:2410.03964
-
Vulnerability Detection via Topological Analysis of Attention Maps 4 Oct 2024 · 1 repository · arXiv:2410.03470
-
Ward: Provable RAG Dataset Inference via LLM Watermarks 4 Oct 2024 · 0 repositories · arXiv:2410.03537Syntology 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
A Comprehensive Survey of Mamba Architectures for Medical Image Analysis: Classification, Segmentation, Restoration and Beyond 3 Oct 2024 · 1 repository · arXiv:2410.02362
-
A Comprehensive Survey of Retrieval-Augmented Generation (RAG): Evolution, Current Landscape and Future Directions 3 Oct 2024 · 0 repositories · arXiv:2410.12837
-
A Novel Method for Accurate & Real-time Food Classification: The Synergistic Integration of EfficientNetB7, CBAM, Transfer Learning, and Data Augmentation 3 Oct 2024 · 0 repositories · arXiv:2410.02304
-
Adaptive Inference-Time Compute: LLMs Can Predict if They Can Do Better, Even Mid-Generation 3 Oct 2024 · 0 repositories · arXiv:2410.02725
-
AlphaIntegrator: Transformer Action Search for Symbolic Integration Proofs 3 Oct 2024 · 0 repositories · arXiv:2410.02666
-
Attention in Large Language Models Yields Efficient Zero-Shot Re-Rankers 3 Oct 2024 · 0 repositories · arXiv:2410.02642
-
BrainTransformers: SNN-LLM 3 Oct 2024 · 0 repositories · arXiv:2410.14687
-
Can LLMs Reliably Simulate Human Learner Actions? A Simulation Authoring Framework for Open-Ended Learning Environments 3 Oct 2024 · 1 repository · arXiv:2410.02110
-
CAX: Cellular Automata Accelerated in JAX 3 Oct 2024 · 1 repository · arXiv:2410.02651
-
Coal Mining Question Answering with LLMs 3 Oct 2024 · 0 repositories · arXiv:2410.02959
-
CodeJudge: Evaluating Code Generation with Large Language Models 3 Oct 2024 · 1 repository · arXiv:2410.02184Syntology official (archive's flag): 13 ran · 15 ran (of which 0 constructed an object rather than computing a result; 14 with no instrument failure: 2 honoured, 0 violated, 12 with no contract checked; 1 where Syntology's instrument failed) · 8 unverified (of 23 harvested samples) · 2 pointer-only (licence)
-
CoLLAP: Contrastive Long-form Language-Audio Pretraining with Musical Temporal Structure Augmentation 3 Oct 2024 · 0 repositories · arXiv:2410.02271
-
Adversarial Decoding: Generating Readable Documents for Adversarial Objectives 3 Oct 2024 · 1 repository · arXiv:2410.02163
-
Cross-Domain Comparative Analysis of Digital Twins and Universalised Solutions 3 Oct 2024 · 0 repositories · arXiv:2410.02358
-
Deconstructing Recurrence, Attention, and Gating: Investigating the transferability of Transformers and Gated Recurrent Neural Networks in forecasting of dynamical systems 3 Oct 2024 · 0 repositories · arXiv:2410.02654
-
Defining Knowledge: Bridging Epistemology and Large Language Models 3 Oct 2024 · 0 repositories · arXiv:2410.02499
-
Differentiation and Specialization of Attention Heads via the Refined Local Learning Coefficient 3 Oct 2024 · 0 repositories · arXiv:2410.02984
-
Domain-Specific Retrieval-Augmented Generation Using Vector Stores, Knowledge Graphs, and Tensor Factorization 3 Oct 2024 · 0 repositories · arXiv:2410.02721