Methods › General › Attention Modules › Multi-Head Attention › Papers, page 85
Multi-Head Attention
Papers archive 2025-07-28
archive papers tagged: 24,855 · with a code link: 11,214 · where Syntology ran a sample: 3,454 (2,916 with a run with no instrument failure, 538 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,454 of 24,855 tagged: 2,916 with a run with no instrument failure, 538 where every run was a failure of Syntology's instrument)
Page 85 of 249: papers 8,401 to 8,500 of 24,855, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
AI Insights: A Case Study on Utilizing ChatGPT Intelligence for Research Paper Analysis 5 Mar 2024 · 0 repositories · arXiv:2403.03293
-
An Empirical Study of LLM-as-a-Judge for LLM Evaluation: Fine-tuned Judge Model is not a General Substitute for GPT-4 5 Mar 2024 · 1 repository · arXiv:2403.02839Syntology official (archive's flag): 15 ran · 15 ran (of which 0 constructed an object rather than computing a result; 15 with no instrument failure: 0 honoured, 0 violated, 15 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 18 harvested samples) · 18 pointer-only (licence)
-
ARNN: Attentive Recurrent Neural Network for Multi-channel EEG Signals to Identify Epileptic Seizures 5 Mar 2024 · 1 repository · arXiv:2403.03276
-
AttentionStitch: How Attention Solves the Speech Editing Problem 5 Mar 2024 · 0 repositories · arXiv:2403.04804
-
Behavior Generation with Latent Actions 5 Mar 2024 · 2 repositories · arXiv:2403.03181Syntology official (archive's flag): 7 ran · 8 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 1 honoured, 2 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 6 unverified (of 14 harvested samples) · 2 pointer-only (licence)
-
CLEVR-POC: Reasoning-Intensive Visual Question Answering in Partially Observable Environments 5 Mar 2024 · 0 repositories · arXiv:2403.03203
-
Drug Resistance Predictions Based on a Directed Flag Transformer 5 Mar 2024 · 0 repositories · arXiv:2403.02603
-
Emerging Synergies Between Large Language Models and Machine Learning in Ecommerce Recommendations 5 Mar 2024 · 0 repositories · arXiv:2403.02760
-
Enhancing Weakly Supervised 3D Medical Image Segmentation through Probabilistic-aware Learning 5 Mar 2024 · 1 repository · arXiv:2403.02566
-
Evaluating and Optimizing Educational Content with Large Language Model Judgments 5 Mar 2024 · 1 repository · arXiv:2403.02795Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Evolution Transformer: In-Context Evolutionary Optimization 5 Mar 2024 · 1 repository · arXiv:2403.02985
-
Exploring Naive Approaches to Tell Apart LLMs Productions from Human-written Text 5 Mar 2024 · 1 repository
-
FAR: Flexible, Accurate and Robust 6DoF Relative Camera Pose Estimation 5 Mar 2024 · 0 repositories · arXiv:2403.03221
-
How Well Can Transformers Emulate In-context Newton's Method? 5 Mar 2024 · 0 repositories · arXiv:2403.03183
-
Improving Event Definition Following For Zero-Shot Event Detection 5 Mar 2024 · 0 repositories · arXiv:2403.02586
-
InjecAgent: Benchmarking Indirect Prompt Injections in Tool-Integrated Large Language Model Agents 5 Mar 2024 · 2 repositories · arXiv:2403.02691Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
InjectTST: A Transformer Method of Injecting Global Information into Independent Channels for Long Time Series Forecasting 5 Mar 2024 · 0 repositories · arXiv:2403.02814
-
Interactive Continual Learning: Fast and Slow Thinking 5 Mar 2024 · 1 repository · arXiv:2403.02628Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
JMI at SemEval 2024 Task 3: Two-step approach for multimodal ECAC using in-context learning with GPT and instruction-tuned Llama models 5 Mar 2024 · 1 repository · arXiv:2403.04798Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 7 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
Knowledge Graphs as Context Sources for LLM-Based Explanations of Learning Recommendations 5 Mar 2024 · 0 repositories · arXiv:2403.03008
-
Language Guided Exploration for RL Agents in Text Environments 5 Mar 2024 · 0 repositories · arXiv:2403.03141
-
Learning without Exact Guidance: Updating Large-scale High-resolution Land Cover Maps from Low-resolution Historical Labels 5 Mar 2024 · 3 repositories · arXiv:2403.02746Syntology official (archive's flag): 2 ran · 5 ran (of which 2 constructed an object rather than computing a result; 5 with no instrument failure: 1 honoured, 1 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 1 pointer-only (licence)
-
MathScale: Scaling Instruction Tuning for Mathematical Reasoning 5 Mar 2024 · 1 repository · arXiv:2403.02884
-
MiKASA: Multi-Key-Anchor & Scene-Aware Transformer for 3D Visual Grounding 5 Mar 2024 · 1 repository · arXiv:2403.03077Syntology official (archive's flag): 11 ran · 11 ran (of which 10 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 16 harvested samples) · 16 pointer-only (licence)
-
PARADISE: Evaluating Implicit Planning Skills of Language Models with Procedural Warnings and Tips Dataset 5 Mar 2024 · 1 repository · arXiv:2403.03167
-
MeanCache: User-Centric Semantic Caching for LLM Web Services 5 Mar 2024 · 0 repositories · arXiv:2403.02694
-
Scope of Large Language Models for Mining Emerging Opinions in Online Health Discourse 5 Mar 2024 · 0 repositories · arXiv:2403.03336
-
SNIFFER: Multimodal Large Language Model for Explainable Out-of-Context Misinformation Detection 5 Mar 2024 · 0 repositories · arXiv:2403.03170
-
Towards Democratized Flood Risk Management: An Advanced AI Assistant Enabled by GPT-4 for Enhanced Interpretability and Public Engagement 5 Mar 2024 · 2 repositories · arXiv:2403.03188
-
Towards Training A Chinese Large Language Model for Anesthesiology 5 Mar 2024 · 0 repositories · arXiv:2403.02742
-
Zero-Shot Cross-Lingual Document-Level Event Causality Identification with Heterogeneous Graph Contrastive Transfer Learning 5 Mar 2024 · 0 repositories · arXiv:2403.02893
-
A Spatio-temporal Aligned SUNet Model for Low-light Video Enhancement 4 Mar 2024 · 0 repositories · arXiv:2403.02408
-
adaptNMT: an open-source, language-agnostic development environment for Neural Machine Translation 4 Mar 2024 · 0 repositories · arXiv:2403.02367
-
GCAN: Generative Counterfactual Attention-guided Network for Explainable Cognitive Decline Diagnostics based on fMRI Functional Connectivity 4 Mar 2024 · 2 repositories · arXiv:2403.01758
-
Automated Generation of Multiple-Choice Cloze Questions for Assessing English Vocabulary Using GPT-turbo 3.5 4 Mar 2024 · 0 repositories · arXiv:2403.02078
-
Brand Visibility in Packaging: A Deep Learning Approach for Logo Detection, Saliency-Map Prediction, and Logo Placement Analysis 4 Mar 2024 · 1 repository · arXiv:2403.02336
-
Can LLMs Generate Architectural Design Decisions? -An Exploratory Empirical study 4 Mar 2024 · 0 repositories · arXiv:2403.01709
-
COLA: Cross-city Mobility Transformer for Human Trajectory Simulation 4 Mar 2024 · 1 repository · arXiv:2403.01801Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
Differentially Private Synthetic Data via Foundation Model APIs 2: Text 4 Mar 2024 · 2 repositories · arXiv:2403.01749Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 1 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 11 harvested samples)
-
EEE-QA: Exploring Effective and Efficient Question-Answer Representations 4 Mar 2024 · 1 repository · arXiv:2403.02176
-
How does Architecture Influence the Base Capabilities of Pre-trained Language Models? A Case Study Based on FFN-Wider and MoE Transformers 4 Mar 2024 · 0 repositories · arXiv:2403.02436
-
Human Evaluation of English--Irish Transformer-Based NMT 4 Mar 2024 · 0 repositories · arXiv:2403.02366
-
Hypertext Entity Extraction in Webpage 4 Mar 2024 · 0 repositories · arXiv:2403.01698
-
Key-Point-Driven Data Synthesis with its Enhancement on Mathematical Reasoning 4 Mar 2024 · 0 repositories · arXiv:2403.02333
-
Lightweight Object Detection: A Study Based on YOLOv7 Integrated with ShuffleNetv2 and Vision Transformer 4 Mar 2024 · 0 repositories · arXiv:2403.01736
-
LLM-Oriented Retrieval Tuner 4 Mar 2024 · 0 repositories · arXiv:2403.01999
-
NiNformer: A Network in Network Transformer with Token Mixing Generated Gating Function 4 Mar 2024 · 1 repository · arXiv:2403.02411
-
NoteLLM: A Retrievable Large Language Model for Note Recommendation 4 Mar 2024 · 0 repositories · arXiv:2403.01744
-
PHAnToM: Persona-based Prompting Has An Effect on Theory-of-Mind Reasoning in Large Language Models 4 Mar 2024 · 0 repositories · arXiv:2403.02246
-
Predicting Learning Performance with Large Language Models: A Study in Adult Literacy 4 Mar 2024 · 0 repositories · arXiv:2403.14668
-
ProTrix: Building Models for Planning and Reasoning over Tables with Sentence Context 4 Mar 2024 · 1 repository · arXiv:2403.02177
-
SciAssess: Benchmarking LLM Proficiency in Scientific Literature Analysis 4 Mar 2024 · 1 repository · arXiv:2403.01976Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Training-Free Pretrained Model Merging 4 Mar 2024 · 1 repository · arXiv:2403.01753Syntology official (archive's flag): 3 ran · 3 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Transformer for Times Series: an Application to the S&P500 4 Mar 2024 · 0 repositories · arXiv:2403.02523
-
Transformers for Low-Resource Languages:Is Féidir Linn! 4 Mar 2024 · 0 repositories · arXiv:2403.01985
-
Using LLMs for the Extraction and Normalization of Product Attribute Values 4 Mar 2024 · 1 repository · arXiv:2403.02130
-
Vanilla Transformers are Transfer Capability Teachers 4 Mar 2024 · 0 repositories · arXiv:2403.01994
-
VariErr NLI: Separating Annotation Error from Human Label Variation 4 Mar 2024 · 0 repositories · arXiv:2403.01931Syntology 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Vision-RWKV: Efficient and Scalable Visual Perception with RWKV-Like Architectures 4 Mar 2024 · 1 repository · arXiv:2403.02308Syntology official (archive's flag): 15 ran · 15 ran (of which 0 constructed an object rather than computing a result; 15 with no instrument failure: 0 honoured, 0 violated, 15 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 19 harvested samples) · 5 pointer-only (licence)
-
Align-to-Distill: Trainable Attention Alignment for Knowledge Distillation in Neural Machine Translation 3 Mar 2024 · 1 repository · arXiv:2403.01479Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
ConvTimeNet: A Deep Hierarchical Fully Convolutional Model for Multivariate Time Series Analysis 3 Mar 2024 · 3 repositories · arXiv:2403.01493Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Enhancing Neural Machine Translation of Low-Resource Languages: Corpus Development, Human Evaluation and Explainable AI Architectures 3 Mar 2024 · 0 repositories · arXiv:2403.01580
-
Enhancing Retinal Vascular Structure Segmentation in Images With a Novel Design Two-Path Interactive Fusion Module Model 3 Mar 2024 · 1 repository · arXiv:2403.01362
-
Fine Tuning vs. Retrieval Augmented Generation for Less Popular Knowledge 3 Mar 2024 · 1 repository · arXiv:2403.01432Syntology official (archive's flag): 15 ran · 15 ran (of which 0 constructed an object rather than computing a result; 15 with no instrument failure: 0 honoured, 0 violated, 15 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 19 harvested samples) · 19 pointer-only (licence)
-
Learning A Physical-aware Diffusion Model Based on Transformer for Underwater Image Enhancement 3 Mar 2024 · 1 repository · arXiv:2403.01497Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 1 honoured, 2 violated, 6 with no contract checked; 3 where Syntology's instrument failed) · 3 unverified (of 15 harvested samples) · 5 pointer-only (licence)
-
LUM-ViT: Learnable Under-sampling Mask Vision Transformer for Bandwidth Limited Optical Signal Acquisition 3 Mar 2024 · 1 repository · arXiv:2403.01412Syntology official (archive's flag): 5 ran · 5 ran (of which 3 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 1 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
MovieLLM: Enhancing Long Video Understanding with AI-Generated Movies 3 Mar 2024 · 0 repositories · arXiv:2403.01422
-
Multi-level Product Category Prediction through Text Classification 3 Mar 2024 · 1 repository · arXiv:2403.01638
-
SERVAL: Synergy Learning between Vertical Models and LLMs towards Oracle-Level Zero-shot Medical Prediction 3 Mar 2024 · 0 repositories · arXiv:2403.01570Syntology 6 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 5 where Syntology's instrument failed) · 2 unverified (of 8 harvested samples)
-
Transformers for Supervised Online Continual Learning 3 Mar 2024 · 0 repositories · arXiv:2403.01554
-
You Need to Pay Better Attention: Rethinking the Mathematics of Attention Mechanism 3 Mar 2024 · 0 repositories · arXiv:2403.01643
-
Analysis of Privacy Leakage in Federated Large Language Models 2 Mar 2024 · 1 repository · arXiv:2403.04784Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
AutoDefense: Multi-Agent LLM Defense against Jailbreak Attacks 2 Mar 2024 · 1 repository · arXiv:2403.04783Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Evaluating Large Language Models as Virtual Annotators for Time-series Physical Sensing Data 2 Mar 2024 · 0 repositories · arXiv:2403.01133
-
Improving the Validity of Automatically Generated Feedback via Reinforcement Learning 2 Mar 2024 · 1 repository · arXiv:2403.01304Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 13 harvested samples) · 13 pointer-only (licence)
-
LAB: Large-Scale Alignment for ChatBots 2 Mar 2024 · 1 repository · arXiv:2403.01081Syntology 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples)
-
Large Language Multimodal Models for 5-Year Chronic Disease Cohort Prediction Using EHR Data 2 Mar 2024 · 0 repositories · arXiv:2403.04785
-
LLaMoCo: Instruction Tuning of Large Language Models for Optimization Code Generation 2 Mar 2024 · 0 repositories · arXiv:2403.01131
-
LM4OPT: Unveiling the Potential of Large Language Models in Formulating Mathematical Optimization Problems 2 Mar 2024 · 0 repositories · arXiv:2403.01342
-
Machine Translation in the Covid domain: an English-Irish case study for LoResMT 2021 2 Mar 2024 · 0 repositories · arXiv:2403.01196
-
OpenGraph: Towards Open Graph Foundation Models 2 Mar 2024 · 1 repository · arXiv:2403.01121Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 5 harvested samples)
-
RAGged Edges: The Double-Edged Sword of Retrieval-Augmented Chatbots 2 Mar 2024 · 0 repositories · arXiv:2403.01193
-
Reading Subtext: Evaluating Large Language Models on Short Story Summarization with Writers 2 Mar 2024 · 2 repositories · arXiv:2403.01061
-
VBART: The Turkish LLM 2 Mar 2024 · 0 repositories · arXiv:2403.01308
-
Comparing large language models and human programmers for generating programming code 1 Mar 2024 · 0 repositories · arXiv:2403.00894
-
ATP: Enabling Fast LLM Serving via Attention on Top Principal Keys 1 Mar 2024 · 0 repositories · arXiv:2403.02352
-
Crimson: Empowering Strategic Reasoning in Cybersecurity through Large Language Models 1 Mar 2024 · 0 repositories · arXiv:2403.00878
-
DAMSDet: Dynamic Adaptive Multispectral Detection Transformer with Competitive Query Selection and Adaptive Feature Fusion 1 Mar 2024 · 2 repositories · arXiv:2403.00326Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 1 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples) · 1 pointer-only (licence)
-
Deformable One-shot Face Stylization via DINO Semantic Guidance 1 Mar 2024 · 1 repository · arXiv:2403.00459
-
DFIN-SQL: Integrating Focused Schema with DIN-SQL for Superior Accuracy in Large-Scale Databases 1 Mar 2024 · 0 repositories · arXiv:2403.00872
-
Dual-domain strip attention for image restoration 1 Mar 2024 · 1 repository
-
Efficient Adapter Tuning of Pre-trained Speech Models for Automatic Speaker Verification 1 Mar 2024 · 0 repositories · arXiv:2403.00293
-
Event-Triggered Robust Cooperative Output Regulation for a Class of Linear Multi-Agent Systems with an Unknown Exosystem 1 Mar 2024 · 0 repositories · arXiv:2403.00645
-
Gender Bias in Large Language Models across Multiple Languages 1 Mar 2024 · 0 repositories · arXiv:2403.00277
-
Large Language Models for Simultaneous Named Entity Extraction and Spelling Correction 1 Mar 2024 · 0 repositories · arXiv:2403.00528
-
Merging Text Transformer Models from Different Initializations 1 Mar 2024 · 1 repository · arXiv:2403.00986Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Multi-modal Heart Failure Risk Estimation based on Short ECG and Sampled Long-Term HRV 1 Mar 2024 · 0 repositories · arXiv:2403.15408
-
SoftTiger: A Clinical Foundation Model for Healthcare Workflows 1 Mar 2024 · 1 repository · arXiv:2403.00868
-
Surveying the Dead Minds: Historical-Psychological Text Analysis with Contextualized Construct Representation (CCR) for Classical Chinese 1 Mar 2024 · 0 repositories · arXiv:2403.00509
-
Task Indicating Transformer for Task-conditional Dense Predictions 1 Mar 2024 · 1 repository · arXiv:2403.00327