Methods › General › Attention Modules › Multi-Head Attention › Papers, page 7
Multi-Head Attention
Papers archive 2025-07-28
archive papers tagged: 24,855 · with a code link: 11,214 · where Syntology ran a sample: 3,454 (2,916 with a run with no instrument failure, 538 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,454 of 24,855 tagged: 2,916 with a run with no instrument failure, 538 where every run was a failure of Syntology's instrument)
Page 7 of 249: papers 601 to 700 of 24,855, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
LLM-e Guess: Can LLMs Capabilities Advance Without Hardware Progress? 7 May 2025 · 1 repository · arXiv:2505.04075
-
M2Rec: Multi-scale Mamba for Efficient Sequential Recommendation 7 May 2025 · 0 repositories · arXiv:2505.04445
-
ORBIT-2: Scaling Exascale Vision Foundation Models for Weather and Climate Downscaling 7 May 2025 · 0 repositories · arXiv:2505.04802
-
Osiris: A Lightweight Open-Source Hallucination Detection System 7 May 2025 · 0 repositories · arXiv:2505.04844
-
Personalized Risks and Regulatory Strategies of Large Language Models in Digital Advertising 7 May 2025 · 0 repositories · arXiv:2505.04665
-
Pose Estimation for Intra-cardiac Echocardiography Catheter via AI-Based Anatomical Understanding 7 May 2025 · 0 repositories · arXiv:2505.07851
-
Red Teaming the Mind of the Machine: A Systematic Evaluation of Prompt Injection and Jailbreak Vulnerabilities in LLMs 7 May 2025 · 0 repositories · arXiv:2505.04806
-
Retrieval Augmented Generation Evaluation for Health Documents 7 May 2025 · 0 repositories · arXiv:2505.04680
-
SwinLip: An Efficient Visual Speech Encoder for Lip Reading Using Swin Transformer 7 May 2025 · 0 repositories · arXiv:2505.04394
-
Theoretical Guarantees for LT-TTD: A Unified Transformer-based Architecture for Two-Level Ranking Systems 7 May 2025 · 0 repositories · arXiv:2505.04434
-
Towards Effectively Leveraging Execution Traces for Program Repair with Code LLMs 7 May 2025 · 0 repositories · arXiv:2505.04441
-
A Comparative Analysis of Ethical and Safety Gaps in LLMs using Relative Danger Coefficient 6 May 2025 · 0 repositories · arXiv:2505.04654
-
A Reasoning-Focused Legal Retrieval Benchmark 6 May 2025 · 0 repositories · arXiv:2505.03970
-
A review of DNA restriction-free overlapping sequence cloning techniques for synthetic biology 6 May 2025 · 0 repositories · arXiv:2505.03681
-
An Analysis of Hyper-Parameter Optimization Methods for Retrieval Augmented Generation 6 May 2025 · 0 repositories · arXiv:2505.03452
-
Attonsecond Streaking Phase Retrieval Via Deep Learning Methods 6 May 2025 · 0 repositories · arXiv:2505.06275
-
Evaluation of LLMs on Long-tail Entity Linking in Historical Documents 6 May 2025 · 0 repositories · arXiv:2505.03473
-
From Word to Sentence: A Large-Scale Multi-Instance Dataset for Open-Set Aerial Detection 6 May 2025 · 0 repositories · arXiv:2505.03334
-
Hesitation is defeat? Connecting Linguistic and Predictive Uncertainty 6 May 2025 · 0 repositories · arXiv:2505.03910
-
Image Recognition with Online Lightweight Vision Transformer: A Survey 6 May 2025 · 0 repositories · arXiv:2505.03113
-
IndicSQuAD: A Comprehensive Multilingual Question Answering Dataset for Indic Languages 6 May 2025 · 1 repository · arXiv:2505.03688
-
MergeGuard: Efficient Thwarting of Trojan Attacks in Machine Learning Models 6 May 2025 · 1 repository · arXiv:2505.04015
-
Physics-inspired Energy Transition Neural Network for Sequence Learning 6 May 2025 · 0 repositories · arXiv:2505.03281
-
Rethinking Boundary Detection in Deep Learning-Based Medical Image Segmentation 6 May 2025 · 1 repository · arXiv:2505.04652
-
Transformers for Learning on Noisy and Task-Level Manifolds: Approximation and Generalization Insights 6 May 2025 · 0 repositories · arXiv:2505.03205
-
Advancing Email Spam Detection: Leveraging Zero-Shot Learning and Large Language Models 5 May 2025 · 1 repository · arXiv:2505.02362
-
Automatic Proficiency Assessment in L2 English Learners 5 May 2025 · 0 repositories · arXiv:2505.02615
-
Deep learning of personalized priors from past MRI scans enables fast, quality-enhanced point-of-care MRI with low-cost systems 5 May 2025 · 0 repositories · arXiv:2505.02470
-
Direct Retrieval-augmented Optimization: Synergizing Knowledge Selection and Language Models 5 May 2025 · 1 repository · arXiv:2505.03075
-
Knowing You Don't Know: Learning When to Continue Search in Multi-round RAG through Self-Practicing 5 May 2025 · 1 repository · arXiv:2505.02811
-
Large Language Model Partitioning for Low-Latency Inference at the Edge 5 May 2025 · 0 repositories · arXiv:2505.02533
-
Less is More: Efficient Weight Farcasting with 1-Layer Neural Network 5 May 2025 · 0 repositories · arXiv:2505.02714
-
LLM4FTS: Enhancing Large Language Models for Financial Time Series Prediction 5 May 2025 · 0 repositories · arXiv:2505.02880
-
Low-Loss Space in Neural Networks is Continuous and Fully Connected 5 May 2025 · 0 repositories · arXiv:2505.02604Syntology 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 13 harvested samples)
-
Memorization or Interpolation ? Detecting LLM Memorization through Input Perturbation Analysis 5 May 2025 · 0 repositories · arXiv:2505.03019
-
Rapid yet accurate Tile-circuit and device modeling for Analog In-Memory Computing 5 May 2025 · 0 repositories · arXiv:2506.00004
-
SCFormer: Structured Channel-wise Transformer with Cumulative Historical State for Multivariate Time Series Forecasting 5 May 2025 · 1 repository · arXiv:2505.02655
-
SymbioticRAG: Enhancing Document Intelligence Through Human-LLM Symbiotic Collaboration 5 May 2025 · 0 repositories · arXiv:2505.02418
-
T2S: High-resolution Time Series Generation with Text-to-Series Diffusion Models 5 May 2025 · 1 repository · arXiv:2505.02417
-
Voila: Voice-Language Foundation Models for Real-Time Autonomous Interaction and Voice Role-Play 5 May 2025 · 1 repository · arXiv:2505.02707Syntology official (archive's flag): 7 ran · 7 ran (of which 6 constructed an object rather than computing a result; 7 with no instrument failure: 1 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 11 harvested samples)
-
A New HOPE: Domain-agnostic Automatic Evaluation of Text Chunking 4 May 2025 · 0 repositories · arXiv:2505.02171
-
Adversarial Cooperative Rationalization: The Risk of Spurious Correlations in Even Clean Datasets 4 May 2025 · 1 repository · arXiv:2505.02118Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample) · 1 pointer-only (licence)
-
CASA: CNN Autoencoder-based Score Attention for Efficient Multivariate Long-term Time-series Forecasting 4 May 2025 · 1 repository · arXiv:2505.02011
-
DualReal: Adaptive Joint Training for Lossless Identity-Motion Fusion in Video Customization 4 May 2025 · 0 repositories · arXiv:2505.02192
-
Exploring new Approaches for Information Retrieval through Natural Language Processing 4 May 2025 · 0 repositories · arXiv:2505.02199
-
Learning Local Causal World Models with State Space Models and Attention 4 May 2025 · 0 repositories · arXiv:2505.02074
-
LLM-OptiRA: LLM-Driven Optimization of Resource Allocation for Non-Convex Problems in Wireless Communications 4 May 2025 · 1 repository · arXiv:2505.02091
-
Local Herb Identification Using Transfer Learning: A CNN-Powered Mobile Application for Nepalese Flora 4 May 2025 · 0 repositories · arXiv:2505.02147
-
Real-time Spatial Retrieval Augmented Generation for Urban Environments 4 May 2025 · 0 repositories · arXiv:2505.02271
-
Retrieval-augmented in-context learning for multimodal large language models in disease classification 4 May 2025 · 0 repositories · arXiv:2505.02087
-
SEval-Ex: A Statement-Level Framework for Explainable Summarization Evaluation 4 May 2025 · 0 repositories · arXiv:2505.02235
-
Positional Attention for Efficient BERT-Based Named Entity Recognition 3 May 2025 · 0 repositories · arXiv:2505.01868
-
Securing 5G and Beyond-Enabled UAV Networks: Resilience Through Multiagent Learning and Transformers Detection 3 May 2025 · 0 repositories · arXiv:2505.01885
-
Semantic Intelligence: Integrating GPT-4 with A Planning in Low-Cost Robotics 3 May 2025 · 0 repositories · arXiv:2505.01931
-
Toward Onboard AI-Enabled Solutions to Space Object Detection for Space Sustainability 3 May 2025 · 0 repositories · arXiv:2505.01650
-
3D Human Pose Estimation via Spatial Graph Order Attention and Temporal Body Aware Transformer 2 May 2025 · 1 repository · arXiv:2505.01003
-
A Character-based Diffusion Embedding Algorithm for Enhancing the Generation Quality of Generative Linguistic Steganographic Texts 2 May 2025 · 0 repositories · arXiv:2505.00977
-
A Domain Adaptation of Large Language Models for Classifying Mechanical Assembly Components 2 May 2025 · 0 repositories · arXiv:2505.01627
-
A Self-Supervised Transformer for Unusable Shared Bike Detection 2 May 2025 · 0 repositories · arXiv:2505.00932
-
A Transformer-based Neural Architecture Search Method 2 May 2025 · 1 repository · arXiv:2505.01314
-
Asset Pricing in Pre-trained Transformer 2 May 2025 · 0 repositories · arXiv:2505.01575
-
CHORUS: Zero-shot Hierarchical Retrieval and Orchestration for Generating Linear Programming Code 2 May 2025 · 0 repositories · arXiv:2505.01485
-
Compact Recurrent Transformer with Persistent Memory 2 May 2025 · 0 repositories · arXiv:2505.00929
-
Enhancing SPARQL Query Rewriting for Complex Ontology Alignments 2 May 2025 · 0 repositories · arXiv:2505.01309
-
FalconWing: An Open-Source Platform for Ultra-Light Fixed-Wing Aircraft Research 2 May 2025 · 0 repositories · arXiv:2505.01383
-
FreCT: Frequency-augmented Convolutional Transformer for Robust Time Series Anomaly Detection 2 May 2025 · 0 repositories · arXiv:2505.00941
-
Good News for Script Kiddies? Evaluating Large Language Models for Automated Exploit Generation 2 May 2025 · 0 repositories · arXiv:2505.01065
-
Multimodal Transformers are Hierarchical Modal-wise Heterogeneous Graphs 2 May 2025 · 0 repositories · arXiv:2505.01068Syntology 0 ran · 2 unverified (of 2 harvested samples)
-
NeuroLoc: Encoding Navigation Cells for 6-DOF Camera Localization 2 May 2025 · 0 repositories · arXiv:2505.01113
-
On the effectiveness of Large Language Models in the mechanical design domain 2 May 2025 · 1 repository · arXiv:2505.01559
-
Retrieval-Augmented Generation in Biomedicine: A Survey of Technologies, Datasets, and Clinical Applications 2 May 2025 · 0 repositories · arXiv:2505.01146
-
Token-free Models for Sarcasm Detection 2 May 2025 · 0 repositories · arXiv:2505.01006
-
Zero-Shot Document-Level Biomedical Relation Extraction via Scenario-based Prompt Design in Two-Stage with LLM 2 May 2025 · 0 repositories · arXiv:2505.01077
-
DARTer: Dynamic Adaptive Representation Tracker for Nighttime UAV Tracking 1 May 2025 · 0 repositories · arXiv:2505.00752
-
A Time-Series Data Augmentation Model through Diffusion and Transformer Integration 1 May 2025 · 0 repositories · arXiv:2505.03790
-
CSE-SFP: Enabling Unsupervised Sentence Representation Learning via a Single Forward Pass 1 May 2025 · 1 repository · arXiv:2505.00389
-
Directly Forecasting Belief for Reinforcement Learning with Delays 1 May 2025 · 1 repository · arXiv:2505.00546
-
Efficient Recommendation with Millions of Items by Dynamic Pruning of Sub-Item Embeddings 1 May 2025 · 0 repositories · arXiv:2505.00560
-
Enhancing Tropical Cyclone Path Forecasting with an Improved Transformer Network 1 May 2025 · 0 repositories · arXiv:2505.00495
-
EnronQA: Towards Personalized RAG over Private Documents 1 May 2025 · 0 repositories · arXiv:2505.00263
-
Gateformer: Advancing Multivariate Time Series Forecasting through Temporal and Variate-Wise Attention with Gated Representations 1 May 2025 · 1 repository · arXiv:2505.00307Syntology official (archive's flag): 2 ran · 2 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; every one of the 2 samples that ran constructed an object rather than computing a result (of 2 harvested samples) · 2 pointer-only (licence)
-
Open-Source LLM-Driven Federated Transformer for Predictive IoV Management 1 May 2025 · 0 repositories · arXiv:2505.00651
-
Patchwork: A Unified Framework for RAG Serving 1 May 2025 · 0 repositories · arXiv:2505.07833
-
Pixel3DMM: Versatile Screen-Space Priors for Single-Image 3D Face Reconstruction 1 May 2025 · 0 repositories · arXiv:2505.00615
-
Unlocking the Potential of Linear Networks for Irregular Multivariate Time Series Forecasting 1 May 2025 · 0 repositories · arXiv:2505.00590
-
Consistency-aware Fake Videos Detection on Short Video Platforms 30 Apr 2025 · 1 repository · arXiv:2504.21495
-
DOPE: Dual Object Perception-Enhancement Network for Vision-and-Language Navigation 30 Apr 2025 · 0 repositories · arXiv:2505.00743
-
Enhancing Security and Strengthening Defenses in Automated Short-Answer Grading Systems 30 Apr 2025 · 0 repositories · arXiv:2505.00061
-
LLM-Empowered Embodied Agent for Memory-Augmented Task Planning in Household Robotics 30 Apr 2025 · 1 repository · arXiv:2504.21716Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
MatMMFuse: Multi-Modal Fusion model for Material Property Prediction 30 Apr 2025 · 1 repository · arXiv:2505.04634Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
Talk Before You Retrieve: Agent-Led Discussions for Better RAG in Medical QA 30 Apr 2025 · 1 repository · arXiv:2504.21252
-
Traceback of Poisoning Attacks to Retrieval-Augmented Generation 30 Apr 2025 · 0 repositories · arXiv:2504.21668
-
Advance Fake Video Detection via Vision Transformers 29 Apr 2025 · 0 repositories · arXiv:2504.20669
-
AlignDiT: Multimodal Aligned Diffusion Transformer for Synchronized Speech Generation 29 Apr 2025 · 0 repositories · arXiv:2504.20629
-
ARCS: Agentic Retrieval-Augmented Code Synthesis with Iterative Refinement 29 Apr 2025 · 0 repositories · arXiv:2504.20434
-
BrightCookies at SemEval-2025 Task 9: Exploring Data Augmentation for Food Hazard Classification 29 Apr 2025 · 1 repository · arXiv:2504.20703
-
CBM-RAG: Demonstrating Enhanced Interpretability in Radiology Report Generation with Multi-Agent RAG and Concept Bottleneck Models 29 Apr 2025 · 1 repository · arXiv:2504.20898
-
DB-GNN: Dual-Branch Graph Neural Network with Multi-Level Contrastive Learning for Jointly Identifying Within- and Cross-Frequency Coupled Brain Networks 29 Apr 2025 · 0 repositories · arXiv:2504.20744
-
Efficient LLMs with AMP: Attention Heads and MLP Pruning 29 Apr 2025 · 2 repositories · arXiv:2504.21174
-
Geolocating Earth Imagery from ISS: Integrating Machine Learning with Astronaut Photography for Enhanced Geographic Mapping 29 Apr 2025 · 1 repository · arXiv:2504.21194