Methods › General › Attention Mechanisms › Attention › Papers, page 123
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 123 of 316: papers 12,201 to 12,300 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
HCS-TNAS: Hybrid Constraint-driven Semi-supervised Transformer-NAS for Ultrasound Image Segmentation 5 Jul 2024 · 0 repositories · arXiv:2407.04203
-
Improving ensemble extreme precipitation forecasts using generative artificial intelligence 5 Jul 2024 · 0 repositories · arXiv:2407.04882
-
Improving Knowledge Distillation in Transfer Learning with Layer-wise Learning Rates 5 Jul 2024 · 0 repositories · arXiv:2407.04871
-
LaRa: Efficient Large-Baseline Radiance Fields 5 Jul 2024 · 1 repository · arXiv:2407.04699Syntology 14 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 2 honoured, 0 violated, 11 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 14 harvested samples) · 2 pointer-only (licence)
-
LayerShuffle: Enhancing Robustness in Vision Transformers by Randomizing Layer Execution Order 5 Jul 2024 · 1 repository · arXiv:2407.04513
-
Learning to (Learn at Test Time): RNNs with Expressive Hidden States 5 Jul 2024 · 3 repositories · arXiv:2407.04620Syntology official (archive's flag): 2 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
Leveraging Graph Structures to Detect Hallucinations in Large Language Models 5 Jul 2024 · 1 repository · arXiv:2407.04485
-
Looking into Black Box Code Language Models 5 Jul 2024 · 0 repositories · arXiv:2407.04868
-
Multi-modal Masked Siamese Network Improves Chest X-Ray Representation Learning 5 Jul 2024 · 3 repositories · arXiv:2407.04449
-
PartCraft: Crafting Creative Objects by Parts 5 Jul 2024 · 1 repository · arXiv:2407.04604Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 5 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Robust Decision Transformer: Tackling Data Corruption in Offline RL via Sequence Modeling 5 Jul 2024 · 0 repositories · arXiv:2407.04285
-
Robust Multimodal Learning via Representation Decoupling 5 Jul 2024 · 0 repositories · arXiv:2407.04458
-
Segmenting Medical Images: From UNet to Res-UNet and nnUNet 5 Jul 2024 · 0 repositories · arXiv:2407.04353
-
Self-Supervised Representation Learning for Adversarial Attack Detection 5 Jul 2024 · 0 repositories · arXiv:2407.04382
-
Strengthening Structural Inductive Biases by Pre-training to Perform Syntactic Transformations 5 Jul 2024 · 1 repository · arXiv:2407.04543
-
Towards Stable 3D Object Detection 5 Jul 2024 · 0 repositories · arXiv:2407.04305
-
Using LLMs to label medical papers according to the CIViC evidence model 5 Jul 2024 · 1 repository · arXiv:2407.04466
-
VCD-Texture: Variance Alignment based 3D-2D Co-Denoising for Text-Guided Texturing 5 Jul 2024 · 0 repositories · arXiv:2407.04461
-
Spatiotemporal Forecasting of Traffic Flow using Wavelet-based Temporal Attention 5 Jul 2024 · 1 repository · arXiv:2407.04440
-
XLSR-Transducer: Streaming ASR for Self-Supervised Pretrained Models 5 Jul 2024 · 0 repositories · arXiv:2407.04439
-
YourMT3+: Multi-instrument Music Transcription with Enhanced Transformer Architectures and Cross-dataset Stem Augmentation 5 Jul 2024 · 1 repository · arXiv:2407.04822
-
A Computer Vision Approach to Estimate the Localized Sea State 4 Jul 2024 · 0 repositories · arXiv:2407.03755
-
A Critical Assessment of Interpretable and Explainable Machine Learning for Intrusion Detection 4 Jul 2024 · 0 repositories · arXiv:2407.04009
-
A Fully Parameter-Free Second-Order Algorithm for Convex-Concave Minimax Problems with Optimal Iteration Complexity 4 Jul 2024 · 0 repositories · arXiv:2407.03571
-
A Systematic Survey and Critical Review on Evaluating Large Language Models: Challenges, Limitations, and Recommendations 4 Jul 2024 · 1 repository · arXiv:2407.04069Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
ADAPT: Multimodal Learning for Detecting Physiological Changes under Missing Modalities 4 Jul 2024 · 1 repository · arXiv:2407.03836
-
Adaptive Step-size Perception Unfolding Network with Non-local Hybrid Attention for Hyperspectral Image Reconstruction 4 Jul 2024 · 0 repositories · arXiv:2407.04024
-
Attention Normalization Impacts Cardinality Generalization in Slot Attention 4 Jul 2024 · 1 repository · arXiv:2407.04170
-
Convolutional vs Large Language Models for Software Log Classification in Edge-Deployable Cellular Network Testing 4 Jul 2024 · 0 repositories · arXiv:2407.03759
-
CRiM-GS: Continuous Rigid Motion-Aware Gaussian Splatting from Motion-Blurred Images 4 Jul 2024 · 0 repositories · arXiv:2407.03923
-
DASS: Distilled Audio State Space Models Are Stronger and More Duration-Scalable Learners 4 Jul 2024 · 1 repository · arXiv:2407.04082Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Deep Content Understanding Toward Entity and Aspect Target Sentiment Analysis on Foundation Models 4 Jul 2024 · 1 repository · arXiv:2407.04050
-
Detection and Multi-Parameter Estimation for NLOS Targets: An IRS-assisted Framework 4 Jul 2024 · 0 repositories · arXiv:2407.03902
-
Diverse and Fine-Grained Instruction-Following Ability Exploration with Synthetic Data 4 Jul 2024 · 0 repositories · arXiv:2407.03942
-
DSLR: Document Refinement with Sentence-Level Re-ranking and Reconstruction to Enhance Retrieval-Augmented Generation 4 Jul 2024 · 0 repositories · arXiv:2407.03627
-
Evaluating Language Model Context Windows: A "Working Memory" Test and Inference-time Correction 4 Jul 2024 · 1 repository · arXiv:2407.03651
-
Feelings about Bodies: Emotions on Diet and Fitness Forums Reveal Gendered Stereotypes and Body Image Concerns 4 Jul 2024 · 0 repositories · arXiv:2407.03551
-
FIPGNet:Pyramid grafting network with feature interaction strategies 4 Jul 2024 · 0 repositories · arXiv:2407.04085
-
From Data to Commonsense Reasoning: The Use of Large Language Models for Explainable AI 4 Jul 2024 · 0 repositories · arXiv:2407.03778
-
Generalizing Graph Transformers Across Diverse Graphs and Tasks via Pre-Training on Industrial-Scale Data 4 Jul 2024 · 0 repositories · arXiv:2407.03953
-
GPT-4 vs. Human Translators: A Comprehensive Evaluation of Translation Quality Across Languages, Domains, and Expertise Levels 4 Jul 2024 · 0 repositories · arXiv:2407.03658
-
QET: Enhancing Quantized LLM Parameters and KV cache Compression through Element Substitution and Residual Clustering 4 Jul 2024 · 0 repositories · arXiv:2407.03637
-
Heterogeneous Hypergraph Embedding for Recommendation Systems 4 Jul 2024 · 1 repository · arXiv:2407.03665
-
HYBRINFOX at CheckThat! 2024 -- Task 1: Enhancing Language Models with Structured Information for Check-Worthiness Estimation 4 Jul 2024 · 0 repositories · arXiv:2407.03850
-
HYBRINFOX at CheckThat! 2024 -- Task 2: Enriching BERT Models with the Expert System VAGO for Subjectivity Detection 4 Jul 2024 · 0 repositories · arXiv:2407.03770
-
Learning Video Temporal Dynamics with Cross-Modal Attention for Robust Audio-Visual Speech Recognition 4 Jul 2024 · 1 repository · arXiv:2407.03563Syntology official (archive's flag): 12 ran · 12 ran (of which 2 constructed an object rather than computing a result; 8 with no instrument failure: 1 honoured, 0 violated, 7 with no contract checked; 4 where Syntology's instrument failed) · 1 unverified (of 13 harvested samples) · 8 pointer-only (licence)
-
Looking for Tiny Defects via Forward-Backward Feature Transfer 4 Jul 2024 · 0 repositories · arXiv:2407.04092
-
MAPO: Boosting Large Language Model Performance with Model-Adaptive Prompt Optimization 4 Jul 2024 · 0 repositories · arXiv:2407.04118
-
Mitigating Low-Frequency Bias: Feature Recalibration and Frequency Attention Regularization for Adversarial Robustness 4 Jul 2024 · 0 repositories · arXiv:2407.04016
-
MRIR: Integrating Multimodal Insights for Diffusion-based Realistic Image Restoration 4 Jul 2024 · 0 repositories · arXiv:2407.03635
-
NutriBench: A Dataset for Evaluating Large Language Models on Nutrition Estimation from Meal Descriptions 4 Jul 2024 · 0 repositories · arXiv:2407.12843
-
On the Benchmarking of LLMs for Open-Domain Dialogue Evaluation 4 Jul 2024 · 0 repositories · arXiv:2407.03841
-
Controllable Conversations: Planning-Based Dialogue Agent with Large Language Models 4 Jul 2024 · 1 repository · arXiv:2407.03884
-
Query-Guided Self-Supervised Summarization of Nursing Notes 4 Jul 2024 · 0 repositories · arXiv:2407.04125
-
Query-oriented Data Augmentation for Session Search 4 Jul 2024 · 0 repositories · arXiv:2407.03720
-
Question-Analysis Prompting Improves LLM Performance in Reasoning Tasks 4 Jul 2024 · 0 repositories · arXiv:2407.03624
-
Serialized Output Training by Learned Dominance 4 Jul 2024 · 0 repositories · arXiv:2407.03966
-
Slice-100K: A Multimodal Dataset for Extrusion-based 3D Printing 4 Jul 2024 · 0 repositories · arXiv:2407.04180
-
Solving Zebra Puzzles Using Constraint-Guided Multi-Agent Systems 4 Jul 2024 · 0 repositories · arXiv:2407.03956
-
Towards Automating Text Annotation: A Case Study on Semantic Proximity Annotation using GPT-4 4 Jul 2024 · 0 repositories · arXiv:2407.04130
-
A Comparative Study of DSL Code Generation: Fine-Tuning vs. Optimized Retrieval Augmentation 3 Jul 2024 · 0 repositories · arXiv:2407.02742
-
A Unified Framework for 3D Scene Understanding 3 Jul 2024 · 1 repository · arXiv:2407.03263Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples)
-
AgentInstruct: Toward Generative Teaching with Agentic Flows 3 Jul 2024 · 0 repositories · arXiv:2407.03502
-
AMEX: Android Multi-annotation Expo Dataset for Mobile GUI Agents 3 Jul 2024 · 0 repositories · arXiv:2407.17490
-
Attention Incorporated Network for Sharing Low-rank, Image and K-space Information during MR Image Reconstruction to Achieve Single Breath-hold Cardiac Cine Imaging 3 Jul 2024 · 1 repository · arXiv:2407.03034
-
Gradient descent with generalized Newton's method 3 Jul 2024 · 1 repository · arXiv:2407.02772Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
CATT: Character-based Arabic Tashkeel Transformer 3 Jul 2024 · 1 repository · arXiv:2407.03236
-
Codec-ASR: Training Performant Automatic Speech Recognition Systems with Discrete Speech Representations 3 Jul 2024 · 0 repositories · arXiv:2407.03495
-
Collective Attention in Human-AI Teams 3 Jul 2024 · 0 repositories · arXiv:2407.17489
-
Croppable Knowledge Graph Embedding 3 Jul 2024 · 0 repositories · arXiv:2407.02779
-
DACB-Net: Dual Attention Guided Compact Bilinear Convolution Neural Network for Skin Disease Classification 3 Jul 2024 · 0 repositories · arXiv:2407.03439
-
Differential Encoding for Improved Representation Learning over Graphs 3 Jul 2024 · 0 repositories · arXiv:2407.02758
-
Fine-Grained Scene Image Classification with Modality-Agnostic Adapter 3 Jul 2024 · 1 repository · arXiv:2407.02769
-
Fisher-aware Quantization for DETR Detectors with Critical-category Objectives 3 Jul 2024 · 0 repositories · arXiv:2407.03442
-
Graph and Skipped Transformer: Exploiting Spatial and Temporal Modeling Capacities for Efficient 3D Human Pose Estimation 3 Jul 2024 · 0 repositories · arXiv:2407.02990
-
Human-like Linguistic Biases in Neural Speech Models: Phonetic Categorization and Phonotactic Constraints in Wav2Vec2.0 3 Jul 2024 · 1 repository · arXiv:2407.03005
-
Improving LLM Abilities in Idiomatic Translation 3 Jul 2024 · 0 repositories · arXiv:2407.03518
-
ISWSST: Index-space-wave State Superposition Transformers for Multispectral Remotely Sensed Imagery Semantic Segmentation 3 Jul 2024 · 0 repositories · arXiv:2407.03033
-
JailbreakHunter: A Visual Analytics Approach for Jailbreak Prompts Discovery from Large-Scale Human-LLM Conversational Datasets 3 Jul 2024 · 0 repositories · arXiv:2407.03045
-
LANE: Logic Alignment of Non-tuning Large Language Models and Online Recommendation Systems for Explainable Reason Generation 3 Jul 2024 · 0 repositories · arXiv:2407.02833
-
Large Language Models as Evaluators for Scientific Synthesis 3 Jul 2024 · 0 repositories · arXiv:2407.02977
-
Learning Positional Attention for Sequential Recommendation 3 Jul 2024 · 1 repository · arXiv:2407.02793
-
Learning to Reduce: Towards Improving Performance of Large Language Models on Structured Data 3 Jul 2024 · 0 repositories · arXiv:2407.02750
-
LMBF-Net: A Lightweight Multipath Bidirectional Focal Attention Network for Multifeatures Segmentation 3 Jul 2024 · 0 repositories · arXiv:2407.02871
-
M5: A Whole Genome Bacterial Encoder at Single Nucleotide Resolution 3 Jul 2024 · 0 repositories · arXiv:2407.03392
-
MLKD-BERT: Multi-level Knowledge Distillation for Pre-trained Language Models 3 Jul 2024 · 0 repositories · arXiv:2407.02775
-
Motion meets Attention: Video Motion Prompts 3 Jul 2024 · 1 repository · arXiv:2407.03179Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 1 honoured, 0 violated, 7 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples) · 7 pointer-only (licence)
-
Exploiting Precision Mapping and Component-Specific Feature Enhancement for Breast Cancer Segmentation and Identification 3 Jul 2024 · 0 repositories · arXiv:2407.02844
-
MVGT: A Multi-view Graph Transformer Based on Spatial Relations for EEG Emotion Recognition 3 Jul 2024 · 0 repositories · arXiv:2407.03131
-
ObfuscaTune: Obfuscated Offsite Fine-tuning and Inference of Proprietary LLMs on Private Datasets 3 Jul 2024 · 0 repositories · arXiv:2407.02960
-
On Large Language Models in National Security Applications 3 Jul 2024 · 1 repository · arXiv:2407.03453
-
OSPC: Artificial VLM Features for Hateful Meme Detection 3 Jul 2024 · 0 repositories · arXiv:2407.12836
-
RDBE: Reasoning Distillation-Based Evaluation Enhances Automatic Essay Scoring 3 Jul 2024 · 0 repositories · arXiv:2407.13781
-
Regurgitative Training: The Value of Real Data in Training Large Language Models 3 Jul 2024 · 0 repositories · arXiv:2407.12835
-
Relating CNN-Transformer Fusion Network for Change Detection 3 Jul 2024 · 1 repository · arXiv:2407.03178
-
SegVG: Transferring Object Bounding Box to Segmentation for Visual Grounding 3 Jul 2024 · 1 repository · arXiv:2407.03200Syntology official (archive's flag): 20 ran · 20 ran (of which 15 constructed an object rather than computing a result; 15 with no instrument failure: 0 honoured, 0 violated, 15 with no contract checked; 5 where Syntology's instrument failed) · 11 unverified (of 31 harvested samples) · 31 pointer-only (licence)
-
Self-supervised Vision Transformer are Scalable Generative Models for Domain Generalization 3 Jul 2024 · 1 repository · arXiv:2407.02900
-
SemioLLM: Assessing Large Language Models for Semiological Analysis in Epilepsy Research 3 Jul 2024 · 0 repositories · arXiv:2407.03004
-
Generative AI Enables EEG Super-Resolution via Spatio-Temporal Adaptive Diffusion Learning 3 Jul 2024 · 0 repositories · arXiv:2407.03089
-
Style Alignment based Dynamic Observation Method for UAV-View Geo-localization 3 Jul 2024 · 0 repositories · arXiv:2407.02832