Methods › General › Attention Mechanisms › Attention › Papers, page 127
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 127 of 316: papers 12,601 to 12,700 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
SUM: Saliency Unification through Mamba for Visual Attention Modeling 25 Jun 2024 · 1 repository · arXiv:2406.17815Syntology official (archive's flag): 14 ran · 14 ran (of which 0 constructed an object rather than computing a result; 14 with no instrument failure: 0 honoured, 0 violated, 14 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 14 harvested samples)
-
Task-Agnostic Federated Learning 25 Jun 2024 · 0 repositories · arXiv:2406.17235
-
Temporal-Channel Modeling in Multi-head Self-Attention for Synthetic Speech Detection 25 Jun 2024 · 1 repository · arXiv:2406.17376
-
This Paper Had the Smartest Reviewers -- Flattery Detection Utilising an Audio-Textual Transformer-Based Approach 25 Jun 2024 · 1 repository · arXiv:2406.17667
-
Towards Open-set Camera 3D Object Detection 25 Jun 2024 · 0 repositories · arXiv:2406.17297
-
Towards Optimal Trade-offs in Knowledge Distillation for CNNs and Vision Transformers at the Edge 25 Jun 2024 · 0 repositories · arXiv:2407.12808
-
Transformer-based Named Entity Recognition with Combined Data Representation 25 Jun 2024 · 0 repositories · arXiv:2406.17474
-
Improving ovarian cancer segmentation accuracy with transformers through AI-guided labeling 25 Jun 2024 · 0 repositories · arXiv:2406.17666
-
Transformer Normalisation Layers and the Independence of Semantic Subspaces 25 Jun 2024 · 0 repositories · arXiv:2406.17837
-
TRIP: Trainable Region-of-Interest Prediction for Hardware-Efficient Neuromorphic Processing on Event-based Vision 25 Jun 2024 · 1 repository · arXiv:2406.17483
-
Univariate Skeleton Prediction in Multivariate Systems Using Transformers 25 Jun 2024 · 1 repository · arXiv:2406.17834
-
Unlocking Continual Learning Abilities in Language Models 25 Jun 2024 · 1 repository · arXiv:2406.17245Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples)
-
Using joint angles based on the international biomechanical standards for human action recognition and related tasks 25 Jun 2024 · 0 repositories · arXiv:2406.17443
-
Understanding Language Model Circuits through Knowledge Editing 25 Jun 2024 · 0 repositories · arXiv:2406.17241
-
A Dual-Channel Particle Swarm Optimization Algorithm Based on Adaptive Balance Search 24 Jun 2024 · 0 repositories · arXiv:2406.16500
-
A Wiener Process Perspective on Local Intrinsic Dimension Estimation Methods 24 Jun 2024 · 0 repositories · arXiv:2406.17125
-
Accelerating Phase Field Simulations Through a Hybrid Adaptive Fourier Neural Operator with U-Net Backbone 24 Jun 2024 · 0 repositories · arXiv:2406.17119
-
Anomaly Detection of Tabular Data Using LLMs 24 Jun 2024 · 0 repositories · arXiv:2406.16308
-
Artistic-style text detector and a new Movie-Poster dataset 24 Jun 2024 · 1 repository · arXiv:2406.16307
-
Attention Instruction: Amplifying Attention in the Middle via Prompting 24 Jun 2024 · 1 repository · arXiv:2406.17095
-
BrainMAE: A Region-aware Self-supervised Learning Framework for Brain Signals 24 Jun 2024 · 0 repositories · arXiv:2406.17086
-
Building on Efficient Foundations: Effectively Training LLMs with Structured Feedforward Layers 24 Jun 2024 · 1 repository · arXiv:2406.16450
-
CausalFormer: An Interpretable Transformer for Temporal Causal Discovery 24 Jun 2024 · 1 repository · arXiv:2406.16708Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 1 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
Classification of Geological Borehole Descriptions Using a Domain Adapted Large Language Model 24 Jun 2024 · 0 repositories · arXiv:2407.10991
-
CLIMATELI: Evaluating Entity Linking on Climate Change Data 24 Jun 2024 · 0 repositories · arXiv:2406.16732
-
DaLPSR: Leverage Degradation-Aligned Language Prompt for Real-World Image Super-Resolution 24 Jun 2024 · 1 repository · arXiv:2406.16477
-
Demystifying the Effect of Receptive Field Size in U-Net Models for Medical Image Segmentation 24 Jun 2024 · 1 repository · arXiv:2406.16701
-
Diff3Dformer: Leveraging Slice Sequence Diffusion for Enhanced 3D CT Classification with Transformer Networks 24 Jun 2024 · 0 repositories · arXiv:2406.17173
-
DreamBench++: A Human-Aligned Benchmark for Personalized Image Generation 24 Jun 2024 · 1 repository · arXiv:2406.16855Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
DWARF: Disease-weighted network for attention map refinement 24 Jun 2024 · 0 repositories · arXiv:2406.17032
-
Evaluation of Instruction-Following Ability for Large Language Models on Story-Ending Generation 24 Jun 2024 · 0 repositories · arXiv:2406.16356
-
Evaluation of Language Models in the Medical Context Under Resource-Constrained Settings 24 Jun 2024 · 1 repository · arXiv:2406.16611
-
Exploring Cross-Domain Few-Shot Classification via Frequency-Aware Prompting 24 Jun 2024 · 1 repository · arXiv:2406.16422Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Exploring Factual Entailment with NLI: A News Media Study 24 Jun 2024 · 0 repositories · arXiv:2406.16842
-
FASTC: A Fast Attentional Framework for Semantic Traversability Classification Using Point Cloud 24 Jun 2024 · 1 repository · arXiv:2406.16564
-
Feature Fusion for Human Activity Recognition using Parameter-Optimized Multi-Stage Graph Convolutional Network and Transformer Models 24 Jun 2024 · 0 repositories · arXiv:2406.16638
-
Finding Transformer Circuits with Edge Pruning 24 Jun 2024 · 1 repository · arXiv:2406.16778Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 3 honoured, 0 violated, 0 with no contract checked; 5 where Syntology's instrument failed) · 0 unverified (of 8 harvested samples)
-
FreeTraj: Tuning-Free Trajectory Control in Video Diffusion Models 24 Jun 2024 · 1 repository · arXiv:2406.16863Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 2 honoured, 1 violated, 7 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 15 harvested samples) · 4 pointer-only (licence)
-
From Decoding to Meta-Generation: Inference-time Algorithms for Large Language Models 24 Jun 2024 · 0 repositories · arXiv:2406.16838
-
GeoMFormer: A General Architecture for Geometric Molecular Representation Learning 24 Jun 2024 · 1 repository · arXiv:2406.16853
-
GMT: Guided Mask Transformer for Leaf Instance Segmentation 24 Jun 2024 · 1 repository · arXiv:2406.17109
-
Hacking a surrogate model approach to XAI 24 Jun 2024 · 0 repositories · arXiv:2406.16626
-
Large Language Models in Student Assessment: Comparing ChatGPT and Human Graders 24 Jun 2024 · 0 repositories · arXiv:2406.16510
-
Lesion-Aware Cross-Phase Attention Network for Renal Tumor Subtype Classification on Multi-Phase CT Scans 24 Jun 2024 · 0 repositories · arXiv:2406.16322
-
Make Graph Neural Networks Great Again: A Generic Integration Paradigm of Topology-Free Patterns for Traffic Speed Prediction 24 Jun 2024 · 1 repository · arXiv:2406.16992
-
METRIK: Measurement-Efficient Randomized Controlled Trials using Transformers with Input Masking 24 Jun 2024 · 0 repositories · arXiv:2406.16351
-
Minimax Optimality in Contextual Dynamic Pricing with General Valuation Models 24 Jun 2024 · 0 repositories · arXiv:2406.17184
-
modeLing: A Novel Dataset for Testing Linguistic Reasoning in Language Models 24 Jun 2024 · 0 repositories · arXiv:2406.17038
-
Multi-LogiEval: Towards Evaluating Multi-Step Logical Reasoning Ability of Large Language Models 24 Jun 2024 · 1 repository · arXiv:2406.17169
-
Multi-Modal Vision Transformers for Crop Mapping from Satellite Image Time Series 24 Jun 2024 · 0 repositories · arXiv:2406.16513
-
On the Role of Long-tail Knowledge in Retrieval Augmented Large Language Models 24 Jun 2024 · 0 repositories · arXiv:2406.16367
-
OTCE: Hybrid SSM and Attention with Cross Domain Mixture of Experts to construct Observer-Thinker-Conceiver-Expresser 24 Jun 2024 · 1 repository · arXiv:2406.16495
-
Panza: Design and Analysis of a Fully-Local Personalized Text Writing Assistant 24 Jun 2024 · 1 repository · arXiv:2407.10994
-
PlagBench: Exploring the Duality of Large Language Models in Plagiarism Generation and Detection 24 Jun 2024 · 0 repositories · arXiv:2406.16288
-
Priorformer: A UGC-VQA Method with content and distortion priors 24 Jun 2024 · 0 repositories · arXiv:2406.16297
-
Ragnarök: A Reusable RAG Framework and Baselines for TREC 2024 Retrieval-Augmented Generation Track 24 Jun 2024 · 2 repositories · arXiv:2406.16828Syntology official (archive's flag): 9 ran · 19 ran (of which 0 constructed an object rather than computing a result; 19 with no instrument failure: 0 honoured, 0 violated, 19 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 23 harvested samples)
-
Repairing Catastrophic-Neglect in Text-to-Image Diffusion Models via Attention-Guided Feature Enhancement 24 Jun 2024 · 1 repository · arXiv:2406.16272
-
Scaling Laws for Linear Complexity Language Models 24 Jun 2024 · 1 repository · arXiv:2406.16690
-
ShadowLLM: Predictor-based Contextual Sparsity for Large Language Models 24 Jun 2024 · 1 repository · arXiv:2406.16635Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Sparser is Faster and Less is More: Efficient Sparse Attention for Long-Range Transformers 24 Jun 2024 · 0 repositories · arXiv:2406.16747
-
The GPT-WritingPrompts Dataset: A Comparative Analysis of Character Portrayal in Short Stories 24 Jun 2024 · 1 repository · arXiv:2406.16767
-
The Progression of Transformers from Language to Vision to MOT: A Literature Review on Multi-Object Tracking with Transformers 24 Jun 2024 · 0 repositories · arXiv:2406.16784
-
Theory on Mixture-of-Experts in Continual Learning 24 Jun 2024 · 0 repositories · arXiv:2406.16437
-
Towards Better Graph-based Cross-document Relation Extraction via Non-bridge Entity Enhancement and Prediction Debiasing 24 Jun 2024 · 1 repository · arXiv:2406.16529
-
Training-Free Exponential Context Extension via Cascading KV Cache 24 Jun 2024 · 1 repository · arXiv:2406.17808Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
MixTex: Unambiguous Recognition Should Not Rely Solely on Real Data 24 Jun 2024 · 1 repository · arXiv:2406.17148
-
UNO Arena for Evaluating Sequential Decision-Making Capability of Large Language Models 24 Jun 2024 · 0 repositories · arXiv:2406.16382
-
USDC: A Dataset of User Stance and Dogmatism in Long Conversations 24 Jun 2024 · 0 repositories · arXiv:2406.16833
-
Venturing into Uncharted Waters: The Navigation Compass from Transformer to Mamba 24 Jun 2024 · 0 repositories · arXiv:2406.16722
-
Video-Infinity: Distributed Long Video Generation 24 Jun 2024 · 0 repositories · arXiv:2406.16260
-
Vision Mamba-based autonomous crack segmentation on concrete, asphalt, and masonry surfaces 24 Jun 2024 · 0 repositories · arXiv:2406.16518
-
A First Running Time Analysis of the Strength Pareto Evolutionary Algorithm 2 (SPEA2) 23 Jun 2024 · 0 repositories · arXiv:2406.16116
-
A Mechanism for Optimizing Media Recommender Systems 23 Jun 2024 · 0 repositories · arXiv:2406.16212
-
Breaking the Frame: Visual Place Recognition by Overlap Prediction 23 Jun 2024 · 1 repository · arXiv:2406.16204Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
DV-3DLane: End-to-end Multi-modal 3D Lane Detection with Dual-view Representation 23 Jun 2024 · 1 repository · arXiv:2406.16072
-
EditFollower: Tunable Car Following Models for Customizable Adaptive Cruise Control Systems 23 Jun 2024 · 0 repositories · arXiv:2407.02516
-
Enhancing Commentary Strategies for Imperfect Information Card Games: A Study of Large Language Models in Guandan Commentary 23 Jun 2024 · 1 repository · arXiv:2406.17807
-
Evaluating Ensemble Methods for News Recommender Systems 23 Jun 2024 · 0 repositories · arXiv:2406.16106
-
Evaluating the Effectiveness of the Foundational Models for Q&A Classification in Mental Health care 23 Jun 2024 · 0 repositories · arXiv:2406.15966
-
Found in the Middle: Calibrating Positional Attention Bias Improves Long Context Utilization 23 Jun 2024 · 0 repositories · arXiv:2406.16008
-
GraphEval2000: Benchmarking and Improving Large Language Models on Graph Datasets 23 Jun 2024 · 0 repositories · arXiv:2406.16176
-
Intensity Confusion Matters: An Intensity-Distance Guided Loss for Bronchus Segmentation 23 Jun 2024 · 1 repository · arXiv:2406.16150
-
Learning Accurate and Enriched Features for Stereo Image Super-Resolution 23 Jun 2024 · 1 repository · arXiv:2406.16001
-
Multi-Scale Temporal Difference Transformer for Video-Text Retrieval 23 Jun 2024 · 0 repositories · arXiv:2406.16111
-
Wound Tissue Segmentation in Diabetic Foot Ulcer Images Using Deep Learning: A Pilot Study 23 Jun 2024 · 1 repository · arXiv:2406.16012
-
A multi-speaker multi-lingual voice cloning system based on vits2 for limmits 2024 challenge 22 Jun 2024 · 0 repositories · arXiv:2406.17801
-
Are Language Models Actually Useful for Time Series Forecasting? 22 Jun 2024 · 3 repositories · arXiv:2406.16964Syntology official (archive's flag): 2 ran · 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 2 pointer-only (licence)
-
Beyond Individual Facts: Investigating Categorical Knowledge Locality of Taxonomy and Meronomy Concepts in GPT Models 22 Jun 2024 · 0 repositories · arXiv:2406.15940
-
Can LLMs Generate Visualizations with Dataless Prompts? 22 Jun 2024 · 0 repositories · arXiv:2406.17805
-
Decentralized Transformers with Centralized Aggregation are Sample-Efficient Multi-Agent World Models 22 Jun 2024 · 1 repository · arXiv:2406.15836
-
Enhancing Solar Driver Forecasting with Multivariate Transformers 22 Jun 2024 · 1 repository · arXiv:2406.15847
-
Fair Clustering: Critique, Caveats, and Future Directions 22 Jun 2024 · 0 repositories · arXiv:2406.15960
-
Fast Tree-Field Integrators: From Low Displacement Rank to Topological Transformers 22 Jun 2024 · 1 repository · arXiv:2406.15881Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 5 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Injectivity of ReLU-layers: Tools from Frame Theory 22 Jun 2024 · 1 repository · arXiv:2406.15856
-
Integrating Attentional Factors and Spacing in Logistic Knowledge Tracing Models to Explore the Impact of Training Sequences on Category Learning 22 Jun 2024 · 0 repositories · arXiv:2407.15020
-
Ladder: A Model-Agnostic Framework Boosting LLM-based Machine Translation to the Next Level 22 Jun 2024 · 3 repositories · arXiv:2406.15741Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 6 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
LaneSegNet Design Study 22 Jun 2024 · 0 repositories · arXiv:2406.15946
-
MVOC: a training-free multiple video object composition method with diffusion models 22 Jun 2024 · 1 repository · arXiv:2406.15829
-
Remaining useful life prediction of rolling bearings based on refined composite multi-scale attention entropy and dispersion entropy 22 Jun 2024 · 0 repositories · arXiv:2406.16967
-
Unveiling Entity-Level Unlearning for Large Language Models: A Comprehensive Analysis 22 Jun 2024 · 0 repositories · arXiv:2406.15796