Methods › General › Attention Mechanisms › Attention › Papers, page 30
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 30 of 316: papers 2,901 to 3,000 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Enhancing LLM Generation with Knowledge Hypergraph for Evidence-Based Medicine 18 Mar 2025 · 0 repositories · arXiv:2503.16530
-
Fast Autoregressive Video Generation with Diagonal Decoding 18 Mar 2025 · 0 repositories · arXiv:2503.14070
-
Fibonacci-Net: A Lightweight CNN model for Automatic Brain Tumor Classification 18 Mar 2025 · 0 repositories · arXiv:2503.13928
-
Good/Evil Reputation Judgment of Celebrities by LLMs via Retrieval Augmented Generation 18 Mar 2025 · 0 repositories · arXiv:2503.14382
-
Gricean Norms as a Basis for Effective Collaboration 18 Mar 2025 · 1 repository · arXiv:2503.14484
-
Growing a Twig to Accelerate Large Vision-Language Models 18 Mar 2025 · 0 repositories · arXiv:2503.14075
-
Identifying and Mitigating Position Bias of Multi-image Vision-Language Models 18 Mar 2025 · 0 repositories · arXiv:2503.13792
-
Inference-Time Intervention in Large Language Models for Reliable Requirement Verification 18 Mar 2025 · 1 repository · arXiv:2503.14130
-
Intra and Inter Parser-Prompted Transformers for Effective Image Restoration 18 Mar 2025 · 1 repository · arXiv:2503.14037
-
Involution and BSConv Multi-Depth Distillation Network for Lightweight Image Super-Resolution 18 Mar 2025 · 0 repositories · arXiv:2503.14779
-
JuDGE: Benchmarking Judgment Document Generation for Chinese Legal System 18 Mar 2025 · 1 repository · arXiv:2503.14258
-
Beyond Single Pass, Looping Through Time: KG-IRAG with Iterative Knowledge Retrieval 18 Mar 2025 · 0 repositories · arXiv:2503.14234
-
Large Language Models for Virtual Human Gesture Selection 18 Mar 2025 · 0 repositories · arXiv:2503.14408
-
Learning Shape-Independent Transformation via Spherical Representations for Category-Level Object Pose Estimation 18 Mar 2025 · 0 repositories · arXiv:2503.13926
-
MagicComp: Training-free Dual-Phase Refinement for Compositional Video Generation 18 Mar 2025 · 0 repositories · arXiv:2503.14428
-
Mapping Urban Villages in China: Progress and Challenges 18 Mar 2025 · 1 repository · arXiv:2503.14195
-
MDocAgent: A Multi-Modal Multi-Agent Framework for Document Understanding 18 Mar 2025 · 1 repository · arXiv:2503.13964Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 1 pointer-only (licence)
-
MoK-RAG: Mixture of Knowledge Paths Enhanced Retrieval-Augmented Generation for Embodied AI Environments 18 Mar 2025 · 0 repositories · arXiv:2503.13882
-
Multi-task Learning for Identification of Porcelain in Song and Yuan Dynasties 18 Mar 2025 · 0 repositories · arXiv:2503.14231
-
Multi-user Wireless Image Semantic Transmission over MIMO Multiple Access Channels 18 Mar 2025 · 0 repositories · arXiv:2504.07969
-
Multimodal Feature-Driven Deep Learning for the Prediction of Duck Body Dimensions and Weight 18 Mar 2025 · 0 repositories · arXiv:2503.14001
-
PENCIL: Long Thoughts with Short Memory 18 Mar 2025 · 1 repository · arXiv:2503.14337Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Predicting Human Choice Between Textually Described Lotteries 18 Mar 2025 · 0 repositories · arXiv:2503.14004
-
RAGO: Systematic Performance Optimization for Retrieval-Augmented Generation Serving 18 Mar 2025 · 1 repository · arXiv:2503.14649
-
Scale-Aware Contrastive Reverse Distillation for Unsupervised Medical Anomaly Detection 18 Mar 2025 · 1 repository · arXiv:2503.13828Syntology official (archive's flag): 6 ran · 8 ran (of which 3 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 0 violated, 3 with no contract checked; 4 where Syntology's instrument failed) · 7 unverified (of 15 harvested samples) · 15 pointer-only (licence)
-
SCJD: Sparse Correlation and Joint Distillation for Efficient 3D Human Pose Estimation 18 Mar 2025 · 0 repositories · arXiv:2503.14097
-
SMILE: a Scale-aware Multiple Instance Learning Method for Multicenter STAS Lung Cancer Histopathology Diagnosis 18 Mar 2025 · 0 repositories · arXiv:2503.13799
-
State Space Model Meets Transformer: A New Paradigm for 3D Object Detection 18 Mar 2025 · 1 repository · arXiv:2503.14493Syntology 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 1 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 10 harvested samples) · 2 pointer-only (licence)
-
Text-Guided Image Invariant Feature Learning for Robust Image Watermarking 18 Mar 2025 · 0 repositories · arXiv:2503.13805
-
Theoretical Foundation of Flow-Based Time Series Generation: Provable Approximation, Generalization, and Efficiency 18 Mar 2025 · 0 repositories · arXiv:2503.14076
-
Tiled Flash Linear Attention: More Efficient Linear RNN and xLSTM Kernels 18 Mar 2025 · 1 repository · arXiv:2503.14376
-
Unique Hard Attention: A Tale of Two Sides 18 Mar 2025 · 0 repositories · arXiv:2503.14615
-
Where do Large Vision-Language Models Look at when Answering Questions? 18 Mar 2025 · 1 repository · arXiv:2503.13891
-
XOXO: Stealthy Cross-Origin Context Poisoning Attacks against AI Coding Assistants 18 Mar 2025 · 0 repositories · arXiv:2503.14281
-
YOLO-LLTS: Real-Time Low-Light Traffic Sign Detection via Prior-Guided Enhancement and Multi-Branch Feature Interaction 18 Mar 2025 · 0 repositories · arXiv:2503.13883
-
A Reinforcement Learning-Driven Transformer GAN for Molecular Generation 17 Mar 2025 · 0 repositories · arXiv:2503.12796
-
A super-resolution reconstruction method for lightweight building images based on an expanding feature modulation network 17 Mar 2025 · 0 repositories · arXiv:2503.13179
-
A Survey on Transformer Context Extension: Approaches and Evaluation 17 Mar 2025 · 0 repositories · arXiv:2503.13299
-
ACT360: An Efficient 360-Degree Action Detection and Summarization Framework for Mission-Critical Training and Debriefing 17 Mar 2025 · 0 repositories · arXiv:2503.12852
-
Advancing Chronic Tuberculosis Diagnostics Using Vision-Language Models: A Multi modal Framework for Precision Analysis 17 Mar 2025 · 0 repositories · arXiv:2503.14536
-
An interpretable approach to automating the assessment of biofouling in video footage 17 Mar 2025 · 1 repository · arXiv:2503.12875
-
Are LLMs (Really) Ideological? An IRT-based Analysis and Alignment Tool for Perceived Socio-Economic Bias in LLMs 17 Mar 2025 · 0 repositories · arXiv:2503.13149
-
C2D-ISR: Optimizing Attention-based Image Super-resolution from Continuous to Discrete Scales 17 Mar 2025 · 0 repositories · arXiv:2503.13740
-
Can Language Models Follow Multiple Turns of Entangled Instructions? 17 Mar 2025 · 1 repository · arXiv:2503.13222
-
ClearSight: Visual Signal Enhancement for Object Hallucination Mitigation in Multimodal Large language Models 17 Mar 2025 · 1 repository · arXiv:2503.13107Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Computation Mechanism Behind LLM Position Generalization 17 Mar 2025 · 0 repositories · arXiv:2503.13305
-
DTGBrepGen: A Novel B-rep Generative Model through Decoupling Topology and Geometry 17 Mar 2025 · 1 repository · arXiv:2503.13110Syntology official (archive's flag): 5 ran · 5 ran (of which 3 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
Dynamic Derivation and Elimination: Audio Visual Segmentation with Enhanced Audio Semantics 17 Mar 2025 · 1 repository · arXiv:2503.12840
-
Explainable Dual-Attention Tabular Transformer for Soil Electrical Resistivity Prediction: A Decision Support Framework for High-Voltage Substation Construction 17 Mar 2025 · 0 repositories · arXiv:2504.02834
-
Exploring 3D Activity Reasoning and Planning: From Implicit Human Intentions to Route-Aware Planning 17 Mar 2025 · 0 repositories · arXiv:2503.12974
-
Feature Extraction and Analysis for GPT-Generated Text 17 Mar 2025 · 0 repositories · arXiv:2503.13687
-
Generative AI for Software Architecture. Applications, Trends, Challenges, and Future Directions 17 Mar 2025 · 0 repositories · arXiv:2503.13310
-
GFSNetwork: Differentiable Feature Selection via Gumbel-Sigmoid Relaxation 17 Mar 2025 · 1 repository · arXiv:2503.13304
-
HICD: Hallucination-Inducing via Attention Dispersion for Contrastive Decoding to Mitigate Hallucinations in Large Language Models 17 Mar 2025 · 1 repository · arXiv:2503.12908
-
High-Resolution Range-Doppler Imaging from One-Bit PMCW Radar via Generative Adversarial Networks 17 Mar 2025 · 0 repositories · arXiv:2503.12841
-
Humanoid Policy ~ Human Policy 17 Mar 2025 · 0 repositories · arXiv:2503.13441
-
In-Context Linear Regression Demystified: Training Dynamics and Mechanistic Interpretability of Multi-Head Softmax Attention 17 Mar 2025 · 1 repository · arXiv:2503.12734
-
Intra-neuronal attention within language models Relationships between activation and semantics 17 Mar 2025 · 0 repositories · arXiv:2503.12992
-
KVShare: An LLM Service System with Efficient and Effective Multi-Tenant KV Cache Reuse 17 Mar 2025 · 0 repositories · arXiv:2503.16525
-
Let Synthetic Data Shine: Domain Reassembly and Soft-Fusion for Single Domain Generalization 17 Mar 2025 · 0 repositories · arXiv:2503.13617
-
MaTVLM: Hybrid Mamba-Transformer for Efficient Vision-Language Modeling 17 Mar 2025 · 1 repository · arXiv:2503.13440Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 4 where Syntology's instrument failed) · 1 unverified (of 10 harvested samples)
-
MES-RAG: Bringing Multi-modal, Entity-Storage, and Secure Enhancements to RAG 17 Mar 2025 · 1 repository · arXiv:2503.13563
-
Mitigating Visual Forgetting via Take-along Visual Conditioning for Multi-modal Long CoT Reasoning 17 Mar 2025 · 0 repositories · arXiv:2503.13360
-
OKRA: an Explainable, Heterogeneous, Multi-Stakeholder Job Recommender System 17 Mar 2025 · 0 repositories · arXiv:2504.07108
-
OSCAR: Online Soft Compression And Reranking 17 Mar 2025 · 0 repositories · arXiv:2504.07109
-
OSLO-IC: On-the-Sphere Learned Omnidirectional Image Compression with Attention Modules and Spatial Context 17 Mar 2025 · 0 repositories · arXiv:2503.13119
-
PAUSE: Low-Latency and Privacy-Aware Active User Selection for Federated Learning 17 Mar 2025 · 1 repository · arXiv:2503.13173
-
Privacy-Aware RAG: Secure and Isolated Knowledge Retrieval 17 Mar 2025 · 0 repositories · arXiv:2503.15548
-
Robust Audio-Visual Segmentation via Audio-Guided Visual Convergent Alignment 17 Mar 2025 · 0 repositories · arXiv:2503.12847
-
SeisRDT: Latent Diffusion Model Based On Representation Learning For Seismic Data Interpolation And Reconstruction 17 Mar 2025 · 0 repositories · arXiv:2503.21791
-
Synchronous vs Asynchronous Reinforcement Learning in a Real World Robot 17 Mar 2025 · 0 repositories · arXiv:2503.14554
-
Towards Scalable Foundation Model for Multi-modal and Hyperspectral Geospatial Data 17 Mar 2025 · 0 repositories · arXiv:2503.12843
-
Unlock Pose Diversity: Accurate and Efficient Implicit Keypoint-based Spatiotemporal Diffusion for Audio-driven Talking Portrait 17 Mar 2025 · 1 repository · arXiv:2503.12963
-
VeriContaminated: Assessing LLM-Driven Verilog Coding for Data Contamination 17 Mar 2025 · 0 repositories · arXiv:2503.13572
-
Atlas: Multi-Scale Attention Improves Long Context Image Modeling 16 Mar 2025 · 1 repository · arXiv:2503.12355
-
Fourier-Based 3D Multistage Transformer for Aberration Correction in Multicellular Specimens 16 Mar 2025 · 2 repositories · arXiv:2503.12593
-
Fragile Mastery: Are Domain-Specific Trade-Offs Undermining On-Device Language Models? 16 Mar 2025 · 0 repositories · arXiv:2503.22698
-
GCBLANE: A graph-enhanced convolutional BiLSTM attention network for improved transcription factor binding site prediction 16 Mar 2025 · 1 repository · arXiv:2503.12377
-
GraphEval: A Lightweight Graph-Based LLM Framework for Idea Evaluation 16 Mar 2025 · 0 repositories · arXiv:2503.12600
-
GS-I³: Gaussian Splatting for Surface Reconstruction from Illumination-Inconsistent Images 16 Mar 2025 · 1 repository · arXiv:2503.12335
-
HyperKAN: Hypergraph Representation Learning with Kolmogorov-Arnold Networks 16 Mar 2025 · 0 repositories · arXiv:2503.12365
-
MambaIC: State Space Models for High-Performance Learned Image Compression 16 Mar 2025 · 1 repository · arXiv:2503.12461
-
MAVEN: Multi-modal Attention for Valence-Arousal Emotion Network 16 Mar 2025 · 1 repository · arXiv:2503.12623
-
Modality-Composable Diffusion Policy via Inference-Time Distribution-level Composition 16 Mar 2025 · 1 repository · arXiv:2503.12466
-
MSCMHMST: A traffic flow prediction model based on Transformer 16 Mar 2025 · 0 repositories · arXiv:2503.13540
-
SAM2-ELNet: Label Enhancement and Automatic Annotation for Remote Sensing Segmentation 16 Mar 2025 · 0 repositories · arXiv:2503.12404
-
Semantic Matters: Multimodal Features for Affective Analysis 16 Mar 2025 · 0 repositories · arXiv:2504.11460
-
State Fourier Diffusion Language Model (SFDLM): A Scalable, Novel Iterative Approach to Language Modeling 16 Mar 2025 · 0 repositories · arXiv:2503.17382
-
TuneNSearch: a hybrid transfer learning and local search approach for solving vehicle routing problems 16 Mar 2025 · 0 repositories · arXiv:2503.12662
-
PA-CFL: Privacy-Adaptive Clustered Federated Learning for Transformer-Based Sales Forecasting on Heterogeneous Retail Data 15 Mar 2025 · 0 repositories · arXiv:2503.12220
-
Applications of Large Language Model Reasoning in Feature Generation 15 Mar 2025 · 0 repositories · arXiv:2503.11989
-
Att-Adapter: A Robust and Precise Domain-Specific Multi-Attributes T2I Diffusion Adapter via Conditional Variational Autoencoder 15 Mar 2025 · 0 repositories · arXiv:2503.11937
-
Changing Base Without Losing Pace: A GPU-Efficient Alternative to MatMul in DNNs 15 Mar 2025 · 0 repositories · arXiv:2503.12211
-
Cognitive Activation and Chaotic Dynamics in Large Language Models: A Quasi-Lyapunov Analysis of Reasoning Mechanisms 15 Mar 2025 · 0 repositories · arXiv:2503.13530
-
Design of an Expression Recognition Solution Employing the Global Channel-Spatial Attention Mechanism 15 Mar 2025 · 0 repositories · arXiv:2503.11935
-
EHNet: An Efficient Hybrid Network for Crowd Counting and Localization 15 Mar 2025 · 0 repositories · arXiv:2503.12061
-
Fast Critical Clearing Time Calculation for Power Systems with Synchronous and Asynchronous Generation 15 Mar 2025 · 0 repositories · arXiv:2503.12132
-
From Laboratory to Real World: A New Benchmark Towards Privacy-Preserved Visible-Infrared Person Re-Identification 15 Mar 2025 · 0 repositories · arXiv:2503.12232
-
United we stand, Divided we fall: Handling Weak Complementary Relationships for Audio-Visual Emotion Recognition in Valence-Arousal Space 15 Mar 2025 · 0 repositories · arXiv:2503.12261
-
Integrating Chain-of-Thought and Retrieval Augmented Generation Enhances Rare Disease Diagnosis from Clinical Notes 15 Mar 2025 · 0 repositories · arXiv:2503.12286