Methods › General › Attention Mechanisms › Attention › Papers, page 24
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 24 of 316: papers 2,301 to 2,400 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Metamorphic Testing for Fairness Evaluation in Large Language Models: Identifying Intersectional Bias in LLaMA and GPT 4 Apr 2025 · 0 repositories · arXiv:2504.07982
-
Model Reveals What to Cache: Profiling-Based Feature Reuse for Video Diffusion Models 4 Apr 2025 · 1 repository · arXiv:2504.03140Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample) · 1 pointer-only (licence)
-
Multi-encoder nnU-Net outperforms Transformer models with self-supervised pretraining 4 Apr 2025 · 0 repositories · arXiv:2504.03474
-
Multi-Granularity Vision Fastformer with Fusion Mechanism for Skin Lesion Segmentation 4 Apr 2025 · 0 repositories · arXiv:2504.03108
-
Multilingual Retrieval-Augmented Generation for Knowledge-Intensive Task 4 Apr 2025 · 0 repositories · arXiv:2504.03616
-
NAACL2025 Tutorial: Adaptation of Large Language Models 4 Apr 2025 · 0 repositories · arXiv:2504.03931
-
Nemotron-H: A Family of Accurate and Efficient Hybrid Mamba-Transformer Models 4 Apr 2025 · 0 repositories · arXiv:2504.03624
-
Practical Poisoning Attacks against Retrieval-Augmented Generation 4 Apr 2025 · 0 repositories · arXiv:2504.03957
-
Rotation Invariance in Floor Plan Digitization using Zernike Moments 4 Apr 2025 · 0 repositories · arXiv:2504.03241
-
Structured Extraction of Process Structure Properties Relationships in Materials Science 4 Apr 2025 · 0 repositories · arXiv:2504.03979
-
TQD-Track: Temporal Query Denoising for 3D Multi-Object Tracking 4 Apr 2025 · 0 repositories · arXiv:2504.03258
-
VISTA-OCR: Towards generative and interactive end to end OCR models 4 Apr 2025 · 0 repositories · arXiv:2504.03621
-
ZFusion: An Effective Fuser of Camera and 4D Radar for 3D Object Perception in Autonomous Driving 4 Apr 2025 · 0 repositories · arXiv:2504.03438
-
A Framework for Situating Innovations, Opportunities, and Challenges in Advancing Vertical Systems with Large AI Models 3 Apr 2025 · 0 repositories · arXiv:2504.02793
-
A Sensorimotor Vision Transformer 3 Apr 2025 · 0 repositories · arXiv:2504.02536
-
AC-LoRA: Auto Component LoRA for Personalized Artistic Style Image Generation 3 Apr 2025 · 0 repositories · arXiv:2504.02231
-
AD-GPT: Large Language Models in Alzheimer's Disease 3 Apr 2025 · 0 repositories · arXiv:2504.03071
-
Adapting Large Language Models for Multi-Domain Retrieval-Augmented-Generation 3 Apr 2025 · 0 repositories · arXiv:2504.02411
-
Attention-Aware Multi-View Pedestrian Tracking 3 Apr 2025 · 0 repositories · arXiv:2504.03047
-
Beyond Conventional Transformers: The Medical X-ray Attention (MXA) Block for Improved Multi-Label Diagnosis Using Knowledge Distillation 3 Apr 2025 · 1 repository · arXiv:2504.02277
-
Cognitive Memory in Large Language Models 3 Apr 2025 · 0 repositories · arXiv:2504.02441
-
CoLa -- Learning to Interactively Collaborate with Large LMs 3 Apr 2025 · 0 repositories · arXiv:2504.02965
-
Computing High-dimensional Confidence Sets for Arbitrary Distributions 3 Apr 2025 · 0 repositories · arXiv:2504.02723
-
Deep Reinforcement Learning via Object-Centric Attention 3 Apr 2025 · 1 repository · arXiv:2504.03024
-
F-ViTA: Foundation Model Guided Visible to Thermal Translation 3 Apr 2025 · 1 repository · arXiv:2504.02801
-
FT-Transformer: Resilient and Reliable Transformer with End-to-End Fault Tolerant Attention 3 Apr 2025 · 0 repositories · arXiv:2504.02211
-
GPTAQ: Efficient Finetuning-Free Quantization for Asymmetric Calibration 3 Apr 2025 · 2 repositories · arXiv:2504.02692Syntology official: harvested, nothing ran · 0 ran · 2 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Graph Attention for Heterogeneous Graphs with Positional Encoding 3 Apr 2025 · 1 repository · arXiv:2504.02938
-
Graphs are everywhere -- Psst! In Music Recommendation too 3 Apr 2025 · 0 repositories · arXiv:2504.02598
-
Group-based Distinctive Image Captioning with Memory Difference Encoding and Attention 3 Apr 2025 · 0 repositories · arXiv:2504.02496
-
HGFormer: Topology-Aware Vision Transformer with HyperGraph Learning 3 Apr 2025 · 0 repositories · arXiv:2504.02440
-
HQViT: Hybrid Quantum Vision Transformer for Image Classification 3 Apr 2025 · 0 repositories · arXiv:2504.02730
-
Hummus: A Dataset of Humorous Multimodal Metaphor Use 3 Apr 2025 · 1 repository · arXiv:2504.02983
-
HyperRAG: Enhancing Quality-Efficiency Tradeoffs in Retrieval-Augmented Generation with Reranker KV-Cache Reuse 3 Apr 2025 · 0 repositories · arXiv:2504.02921
-
Hyperspectral Remote Sensing Images Salient Object Detection: The First Benchmark Dataset and Baseline 3 Apr 2025 · 1 repository · arXiv:2504.02416
-
LearNAT: Learning NL2SQL with AST-guided Task Decomposition for Large Language Models 3 Apr 2025 · 0 repositories · arXiv:2504.02327
-
Learning Audio-guided Video Representation with Gated Attention for Video-Text Retrieval 3 Apr 2025 · 0 repositories · arXiv:2504.02397
-
Localized Definitions and Distributed Reasoning: A Proof-of-Concept Mechanistic Interpretability Study via Activation Patching 3 Apr 2025 · 1 repository · arXiv:2504.02976
-
MegaScale-Infer: Serving Mixture-of-Experts at Scale with Disaggregated Expert Parallelism 3 Apr 2025 · 0 repositories · arXiv:2504.02263
-
MMTL-UniAD: A Unified Framework for Multimodal and Multi-Task Learning in Assistive Driving Perception 3 Apr 2025 · 1 repository · arXiv:2504.02264Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Multi-Head Adaptive Graph Convolution Network for Sparse Point Cloud-Based Human Activity Recognition 3 Apr 2025 · 1 repository · arXiv:2504.02778
-
On Vanishing Variance in Transformer Length Generalization 3 Apr 2025 · 0 repositories · arXiv:2504.02827
-
QID: Efficient Query-Informed ViTs in Data-Scarce Regimes for OCR-free Visual Document Understanding 3 Apr 2025 · 0 repositories · arXiv:2504.02971
-
Secure Generalization through Stochastic Bidirectional Parameter Updates Using Dual-Gradient Mechanism 3 Apr 2025 · 0 repositories · arXiv:2504.02213
-
Semiconductor Wafer Map Defect Classification with Tiny Vision Transformers 3 Apr 2025 · 0 repositories · arXiv:2504.02494
-
SLACK: Attacking LiDAR-based SLAM with Adversarial Point Injections 3 Apr 2025 · 0 repositories · arXiv:2504.03089
-
Spline-based Transformers 3 Apr 2025 · 0 repositories · arXiv:2504.02797
-
Task as Context Prompting for Accurate Medical Symptom Coding Using Large Language Models 3 Apr 2025 · 1 repository · arXiv:2504.03051
-
Towards Computation- and Communication-efficient Computational Pathology 3 Apr 2025 · 0 repositories · arXiv:2504.02628
-
Why do LLMs attend to the first token? 3 Apr 2025 · 1 repository · arXiv:2504.02732Syntology 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples)
-
WonderTurbo: Generating Interactive 3D World in 0.72 Seconds 3 Apr 2025 · 0 repositories · arXiv:2504.02261
-
A Prefixed Patch Time Series Transformer for Two-Point Boundary Value Problems in Three-Body Problems 2 Apr 2025 · 0 repositories · arXiv:2504.01464
-
A thorough benchmark of automatic text classification: From traditional approaches to large language models 2 Apr 2025 · 1 repository · arXiv:2504.01930
-
Analysis of an Idealized Stochastic Polyak Method and its Application to Black-Box Model Distillation 2 Apr 2025 · 0 repositories · arXiv:2504.01898
-
Attention Mamba: Time Series Modeling with Adaptive Pooling Acceleration and Receptive Field Enhancements 2 Apr 2025 · 0 repositories · arXiv:2504.02013
-
BioAtt: Anatomical Prior Driven Low-Dose CT Denoising 2 Apr 2025 · 0 repositories · arXiv:2504.01662
-
Biomedical Question Answering via Multi-Level Summarization on a Local Knowledge Graph 2 Apr 2025 · 0 repositories · arXiv:2504.01309
-
BOLDSimNet: Examining Brain Network Similarity between Task and Resting-State fMRI 2 Apr 2025 · 0 repositories · arXiv:2504.01274
-
Breaking BERT: Gradient Attack on Twitter Sentiment Analysis for Targeted Misclassification 2 Apr 2025 · 1 repository · arXiv:2504.01345
-
Chain of Correction for Full-text Speech Recognition with Large Language Models 2 Apr 2025 · 0 repositories · arXiv:2504.01519
-
Coarse-to-Fine Semantic Communication Systems for Text Transmission 2 Apr 2025 · 0 repositories · arXiv:2504.01442
-
Context-Aware Toxicity Detection in Multiplayer Games: Integrating Domain-Adaptive Pretraining and Match Metadata 2 Apr 2025 · 1 repository · arXiv:2504.01534
-
CoRAG: Collaborative Retrieval-Augmented Generation 2 Apr 2025 · 0 repositories · arXiv:2504.01883
-
Decoding Covert Speech from EEG Using a Functional Areas Spatio-Temporal Transformer 2 Apr 2025 · 1 repository · arXiv:2504.03762
-
Deep Representation Learning for Unsupervised Clustering of Myocardial Fiber Trajectories in Cardiac Diffusion Tensor Imaging 2 Apr 2025 · 0 repositories · arXiv:2504.01953
-
Dual-stream Transformer-GCN Model with Contextualized Representations Learning for Monocular 3D Human Pose Estimation 2 Apr 2025 · 1 repository · arXiv:2504.01764
-
Efficient Model Selection for Time Series Forecasting via LLMs 2 Apr 2025 · 0 repositories · arXiv:2504.02119
-
Enhancing Traffic Sign Recognition On The Performance Based On Yolov8 2 Apr 2025 · 0 repositories · arXiv:2504.02884
-
GaussianLSS -- Toward Real-world BEV Perception: Depth Uncertainty Estimation via Gaussian Splatting 2 Apr 2025 · 0 repositories · arXiv:2504.01957
-
Geometric Reasoning in the Embedding Space 2 Apr 2025 · 0 repositories · arXiv:2504.02018
-
GeoRAG: A Question-Answering Approach from a Geographical Perspective 2 Apr 2025 · 0 repositories · arXiv:2504.01458
-
GPT Adoption and the Impact of Disclosure Policies 2 Apr 2025 · 0 repositories · arXiv:2504.01566
-
GTR: Graph-Table-RAG for Cross-Table Question Answering 2 Apr 2025 · 0 repositories · arXiv:2504.01346
-
InvFussion: Bridging Supervised and Zero-shot Diffusion for Inverse Problems 2 Apr 2025 · 1 repository · arXiv:2504.01689
-
LARGE: Legal Retrieval Augmented Generation Evaluation Tool 2 Apr 2025 · 1 repository · arXiv:2504.01840
-
On Model Protection in Federated Learning against Eavesdropping Attacks 2 Apr 2025 · 0 repositories · arXiv:2504.02114
-
OnRL-RAG: Real-Time Personalized Mental Health Dialogue System 2 Apr 2025 · 0 repositories · arXiv:2504.02894
-
Overlap-Aware Feature Learning for Robust Unsupervised Domain Adaptation for 3D Semantic Segmentation 2 Apr 2025 · 0 repositories · arXiv:2504.01668
-
PiCo: Jailbreaking Multimodal Large Language Models via Pictorial Code Contextualization 2 Apr 2025 · 0 repositories · arXiv:2504.01444
-
Prompt-Guided Attention Head Selection for Focus-Oriented Image Retrieval 2 Apr 2025 · 0 repositories · arXiv:2504.01348
-
Prompting Medical Vision-Language Models to Mitigate Diagnosis Bias by Generating Realistic Dermoscopic Images 2 Apr 2025 · 1 repository · arXiv:2504.01838
-
ProtoGuard-guided PROPEL: Class-Aware Prototype Enhancement and Progressive Labeling for Incremental 3D Point Cloud Segmentation 2 Apr 2025 · 0 repositories · arXiv:2504.01648
-
Quattro: Transformer-Accelerated Iterative Linear Quadratic Regulator Framework for Fast Trajectory Optimization 2 Apr 2025 · 1 repository · arXiv:2504.01806
-
Revisiting Funnel Transformers for Modern LLM Architectures with Comprehensive Ablations in Training and Inference Configurations 2 Apr 2025 · 0 repositories · arXiv:2504.02877
-
Robust Channel Estimation for Optical Wireless Communications Using Neural Network 2 Apr 2025 · 1 repository · arXiv:2504.02134
-
Scaling Test-Time Inference with Policy-Optimized, Dynamic Retrieval-Augmented Generation via KV Caching and Decoding 2 Apr 2025 · 0 repositories · arXiv:2504.01281
-
Strategize Globally, Adapt Locally: A Multi-Turn Red Teaming Agent with Dual-Level Learning 2 Apr 2025 · 0 repositories · arXiv:2504.01278
-
Testing Low-Resource Language Support in LLMs Using Language Proficiency Exams: the Case of Luxembourgish 2 Apr 2025 · 0 repositories · arXiv:2504.01667
-
Time-to-event prediction for grouped variables using Exclusive Lasso 2 Apr 2025 · 0 repositories · arXiv:2504.01520
-
Towards Interpretable Soft Prompts 2 Apr 2025 · 1 repository · arXiv:2504.02144
-
UniViTAR: Unified Vision Transformer with Native Resolution 2 Apr 2025 · 0 repositories · arXiv:2504.01792
-
A Unified Virtual Mixture-of-Experts Framework:Enhanced Inference and Hallucination Mitigation in Single-Model System 1 Apr 2025 · 0 repositories · arXiv:2504.03739
-
Accelerating Causal Network Discovery of Alzheimer Disease Biomarkers via Scientific Literature-based Retrieval Augmented Generation 1 Apr 2025 · 0 repositories · arXiv:2504.08768
-
Adaptive Low Light Enhancement via Joint Global-Local Illumination Adjustment 1 Apr 2025 · 0 repositories · arXiv:2504.00400
-
Alleviating Performance Disparity in Adversarial Spatiotemporal Graph Learning Under Zero-Inflated Distribution 1 Apr 2025 · 0 repositories · arXiv:2504.00721
-
Attention in Diffusion Model: A Survey 1 Apr 2025 · 0 repositories · arXiv:2504.03738
-
Automated Factual Benchmarking for In-Car Conversational Systems using Large Language Models 1 Apr 2025 · 0 repositories · arXiv:2504.01248
-
CamoSAM2: Motion-Appearance Induced Auto-Refining Prompts for Video Camouflaged Object Detection 1 Apr 2025 · 0 repositories · arXiv:2504.00375
-
CellVTA: Enhancing Vision Foundation Models for Accurate Cell Segmentation and Classification 1 Apr 2025 · 1 repository · arXiv:2504.00784
-
Collaborative LLM Numerical Reasoning with Local Data Protection 1 Apr 2025 · 0 repositories · arXiv:2504.00299