Methods › General › Attention Modules › Multi-Head Attention › Papers, page 13
Multi-Head Attention
Papers archive 2025-07-28
archive papers tagged: 24,855 · with a code link: 11,214 · where Syntology ran a sample: 3,454 (2,916 with a run with no instrument failure, 538 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,454 of 24,855 tagged: 2,916 with a run with no instrument failure, 538 where every run was a failure of Syntology's instrument)
Page 13 of 249: papers 1,201 to 1,300 of 24,855, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
GPT Meets Graphs and KAN Splines: Testing Novel Frameworks on Multitask Fine-Tuned GPT-2 with LoRA 25 Mar 2025 · 0 repositories · arXiv:2504.10490
-
iNatAg: Multi-Class Classification Models Enabled by a Large-Scale Benchmark Dataset with 4.7M Images of 2,959 Crop and Weed Species 25 Mar 2025 · 1 repository · arXiv:2503.20068
-
M²CD: A Unified MultiModal Framework for Optical-SAR Change Detection with Mixture of Experts and Self-Distillation 25 Mar 2025 · 0 repositories · arXiv:2503.19406
-
Machine-assisted writing evaluation: Exploring pre-trained language models in analyzing argumentative moves 25 Mar 2025 · 0 repositories · arXiv:2503.19279
-
Mask²DiT: Dual Mask-based Diffusion Transformer for Multi-Scene Long Video Generation 25 Mar 2025 · 0 repositories · arXiv:2503.19881
-
RoboFlamingo-Plus: Fusion of Depth and RGB Perception with Vision-Language Models for Enhanced Robotic Manipulation 25 Mar 2025 · 0 repositories · arXiv:2503.19510
-
Scaling Down Text Encoders of Text-to-Image Diffusion Models 25 Mar 2025 · 1 repository · arXiv:2503.19897
-
SCI-IDEA: Context-Aware Scientific Ideation Using Token and Sentence Embeddings 25 Mar 2025 · 0 repositories · arXiv:2503.19257
-
Social Network User Profiling for Anomaly Detection Based on Graph Neural Networks 25 Mar 2025 · 0 repositories · arXiv:2503.19380
-
Surg-3M: A Dataset and Foundation Model for Perception in Surgical Settings 25 Mar 2025 · 1 repository · arXiv:2503.19740
-
Taxonomy Inference for Tabular Data Using Large Language Models 25 Mar 2025 · 0 repositories · arXiv:2503.21810
-
TRIDIS: A Comprehensive Medieval and Early Modern Corpus for HTR and NER 25 Mar 2025 · 0 repositories · arXiv:2503.22714
-
VGAT: A Cancer Survival Analysis Framework Transitioning from Generative Visual Question Answering to Genomic Reconstruction 25 Mar 2025 · 1 repository · arXiv:2503.19367
-
Analyzing Islamophobic Discourse Using Semi-Coded Terms and LLMs 24 Mar 2025 · 0 repositories · arXiv:2503.18273
-
Chirp Localization via Fine-Tuned Transformer Model: A Proof-of-Concept Study 24 Mar 2025 · 0 repositories · arXiv:2503.22713
-
Coeff-Tuning: A Graph Filter Subspace View for Tuning Attention-Based Large Models 24 Mar 2025 · 1 repository · arXiv:2503.18337Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Construction Identification and Disambiguation Using BERT: A Case Study of NPN 24 Mar 2025 · 0 repositories · arXiv:2503.18751
-
Context-Enhanced Memory-Refined Transformer for Online Action Detection 24 Mar 2025 · 1 repository · arXiv:2503.18359
-
Deep learning-based identification of precipitation clouds from all-sky camera data for observatory safety 24 Mar 2025 · 0 repositories · arXiv:2503.18670
-
Distil-xLSTM: Learning Attention Mechanisms through Recurrent Structures 24 Mar 2025 · 0 repositories · arXiv:2503.18565
-
Enhancing Recommender Systems Using Textual Embeddings from Pre-trained Language Models 24 Mar 2025 · 0 repositories · arXiv:2504.08746
-
Exploring the Integration of Key-Value Attention Into Pure and Hybrid Transformers for Semantic Segmentation 24 Mar 2025 · 0 repositories · arXiv:2503.18862
-
Exploring Training and Inference Scaling Laws in Generative Retrieval 24 Mar 2025 · 1 repository · arXiv:2503.18941
-
Global-Local Tree Search in VLMs for 3D Indoor Scene Generation 24 Mar 2025 · 1 repository · arXiv:2503.18476Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
How to Capture and Study Conversations Between Research Participants and ChatGPT: GPT for Researchers (g4r.org) 24 Mar 2025 · 0 repositories · arXiv:2503.18303
-
Image-to-Text for Medical Reports Using Adaptive Co-Attention and Triple-LSTM Module 24 Mar 2025 · 0 repositories · arXiv:2503.18297
-
Improving RAG for Personalization with Author Features and Contrastive Examples 24 Mar 2025 · 1 repository · arXiv:2504.08745
-
LGI-DETR: Local-Global Interaction for UAV Object Detection 24 Mar 2025 · 0 repositories · arXiv:2503.18785
-
LoTUS: Large-Scale Machine Unlearning with a Taste of Uncertainty 24 Mar 2025 · 1 repository · arXiv:2503.18314
-
Predicting the Road Ahead: A Knowledge Graph based Foundation Model for Scene Understanding in Autonomous Driving 24 Mar 2025 · 0 repositories · arXiv:2503.18730
-
REALM: A Dataset of Real-World LLM Use Cases 24 Mar 2025 · 0 repositories · arXiv:2503.18792
-
SPMTrack: Spatio-Temporal Parameter-Efficient Fine-Tuning with Mixture of Experts for Scalable Visual Tracking 24 Mar 2025 · 1 repository · arXiv:2503.18338Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Synthetic Function Demonstrations Improve Generation in Low-Resource Programming Languages 24 Mar 2025 · 0 repositories · arXiv:2503.18760
-
U-REPA: Aligning Diffusion U-Nets to ViTs 24 Mar 2025 · 1 repository · arXiv:2503.18414
-
Video-XL-Pro: Reconstructive Token Compression for Extremely Long Video Understanding 24 Mar 2025 · 0 repositories · arXiv:2503.18478
-
Your ViT is Secretly an Image Segmentation Model 24 Mar 2025 · 1 repository · arXiv:2503.19108Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
ZeroLM: Data-Free Transformer Architecture Search for Language Models 24 Mar 2025 · 0 repositories · arXiv:2503.18646
-
Adaptive Rank Allocation: Speeding Up Modern Transformers with RaNA Adapters 23 Mar 2025 · 1 repository · arXiv:2503.18216Syntology official (archive's flag): 3 ran · 3 ran (of which 1 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
End-to-End Implicit Neural Representations for Classification 23 Mar 2025 · 1 repository · arXiv:2503.18123
-
ExpertRAG: Efficient RAG with Mixture of Experts -- Optimizing Context Retrieval for Adaptive LLM Responses 23 Mar 2025 · 0 repositories · arXiv:2504.08744
-
Investigating Recent Large Language Models for Vietnamese Machine Reading Comprehension 23 Mar 2025 · 0 repositories · arXiv:2503.18062
-
LakotaBERT: A Transformer-based Model for Low Resource Lakota Language 23 Mar 2025 · 0 repositories · arXiv:2503.18212
-
PathoHR: Breast Cancer Survival Prediction on High-Resolution Pathological Images 23 Mar 2025 · 1 repository · arXiv:2503.17970
-
Payload-Aware Intrusion Detection with CMAE and Large Language Models 23 Mar 2025 · 0 repositories · arXiv:2503.20798
-
Retrieval Augmented Generation and Understanding in Vision: A Survey and New Outlook 23 Mar 2025 · 1 repository · arXiv:2503.18016
-
SymmCompletion: High-Fidelity and High-Consistency Point Cloud Completion with Symmetry Guidance 23 Mar 2025 · 1 repository · arXiv:2503.18007
-
A Modular Dataset to Demonstrate LLM Abstraction Capability 22 Mar 2025 · 0 repositories · arXiv:2503.17645
-
Automated diagnosis of lung diseases using vision transformer: a comparative study on chest x-ray classification 22 Mar 2025 · 0 repositories · arXiv:2503.18973
-
Bandwidth Reservation for Time-Critical Vehicular Applications: A Multi-Operator Environment 22 Mar 2025 · 0 repositories · arXiv:2503.17756
-
EMPLACE: Self-Supervised Urban Scene Change Detection 22 Mar 2025 · 1 repository · arXiv:2503.17716Syntology official (archive's flag): 4 ran · 4 ran (of which 4 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified; every one of the 4 samples that ran constructed an object rather than computing a result (of 7 harvested samples) · 7 pointer-only (licence)
-
Enhancing Arabic Automated Essay Scoring with Synthetic Data and Error Injection 22 Mar 2025 · 0 repositories · arXiv:2503.17739
-
Hierarchy-Aware and Channel-Adaptive Semantic Communication for Bandwidth-Limited Data Fusion 22 Mar 2025 · 0 repositories · arXiv:2503.17777
-
Serial Low-rank Adaptation of Vision Transformer 22 Mar 2025 · 0 repositories · arXiv:2503.17750
-
TDRI: Two-Phase Dialogue Refinement and Co-Adaptation for Interactive Image Generation 22 Mar 2025 · 0 repositories · arXiv:2503.17669
-
Assessing the Reliability and Validity of GPT-4 in Annotating Emotion Appraisal Ratings 21 Mar 2025 · 0 repositories · arXiv:2503.16883
-
Autonomous Radiotherapy Treatment Planning Using DOLA: A Privacy-Preserving, LLM-Based Optimization Agent 21 Mar 2025 · 0 repositories · arXiv:2503.17553
-
CoKe: Customizable Fine-Grained Story Evaluation via Chain-of-Keyword Rationalization 21 Mar 2025 · 0 repositories · arXiv:2503.17136
-
Efficient Intent-Based Filtering for Multi-Party Conversations Using Knowledge Distillation from LLMs 21 Mar 2025 · 0 repositories · arXiv:2503.17336
-
Feature-Based Dual Visual Feature Extraction Model for Compound Multimodal Emotion Recognition 21 Mar 2025 · 1 repository · arXiv:2503.17453
-
Federated Cross-Domain Click-Through Rate Prediction With Large Language Model Augmentation 21 Mar 2025 · 0 repositories · arXiv:2503.16875
-
Meme Similarity and Emotion Detection using Multimodal Analysis 21 Mar 2025 · 0 repositories · arXiv:2503.17493
-
SaudiCulture: A Benchmark for Evaluating Large Language Models Cultural Competence within Saudi Arabia 21 Mar 2025 · 0 repositories · arXiv:2503.17485
-
Vision Transformer Based Semantic Communications for Next Generation Wireless Networks 21 Mar 2025 · 0 repositories · arXiv:2503.17275
-
When Words Outperform Vision: VLMs Can Self-Improve Via Text-Only Training For Human-Centered Decision Making 21 Mar 2025 · 0 repositories · arXiv:2503.16965Syntology 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Zero-Shot Styled Text Image Generation, but Make It Autoregressive 21 Mar 2025 · 0 repositories · arXiv:2503.17074
-
Binarized Mamba-Transformer for Lightweight Quad Bayer HybridEVS Demosaicing 20 Mar 2025 · 1 repository · arXiv:2503.16134
-
Design and Implementation of an FPGA-Based Hardware Accelerator for Transformer 20 Mar 2025 · 1 repository · arXiv:2503.16731
-
DocVideoQA: Towards Comprehensive Understanding of Document-Centric Videos through Question Answering 20 Mar 2025 · 0 repositories · arXiv:2503.15887
-
Don't Fight Hallucinations, Use Them: Estimating Image Realism using NLI over Atomic Facts 20 Mar 2025 · 1 repository · arXiv:2503.15948
-
EDiT: Efficient Diffusion Transformers with Linear Compressed Attention 20 Mar 2025 · 0 repositories · arXiv:2503.16726
-
Financial Analysis: Intelligent Financial Data Analysis System Based on LLM-RAG 20 Mar 2025 · 0 repositories · arXiv:2504.06279
-
FreeFlux: Understanding and Exploiting Layer-Specific Roles in RoPE-Based MMDiT for Versatile Image Editing 20 Mar 2025 · 0 repositories · arXiv:2503.16153
-
GraPLUS: Graph-based Placement Using Semantics for Image Composition 20 Mar 2025 · 0 repositories · arXiv:2503.15761
-
Hyperspectral Imaging for Identifying Foreign Objects on Pork Belly 20 Mar 2025 · 0 repositories · arXiv:2503.16086
-
iFlame: Interleaving Full and Linear Attention for Efficient Mesh Generation 20 Mar 2025 · 0 repositories · arXiv:2503.16653
-
InhibiDistilbert: Knowledge Distillation for a ReLU and Addition-based Transformer 20 Mar 2025 · 0 repositories · arXiv:2503.15983
-
Iterative Optimal Attention and Local Model for Single Image Rain Streak Removal 20 Mar 2025 · 1 repository · arXiv:2503.16165
-
Deep learning framework for action prediction reveals multi-timescale locomotor control 20 Mar 2025 · 0 repositories · arXiv:2503.16340
-
Parameters vs. Context: Fine-Grained Control of Knowledge Reliance in Language Models 20 Mar 2025 · 1 repository · arXiv:2503.15888
-
PromptHash: Affinity-Prompted Collaborative Cross-Modal Learning for Adaptive Hashing Retrieval 20 Mar 2025 · 0 repositories · arXiv:2503.16064
-
SenseExpo: Efficient Autonomous Exploration with Prediction Information from Lightweight Neural Networks 20 Mar 2025 · 0 repositories · arXiv:2503.16000
-
SpiLiFormer: Enhancing Spiking Transformers with Lateral Inhibition 20 Mar 2025 · 0 repositories · arXiv:2503.15986
-
The Lighthouse of Language: Enhancing LLM Agents via Critique-Guided Improvement 20 Mar 2025 · 0 repositories · arXiv:2503.16024
-
Towards Lighter and Robust Evaluation for Retrieval Augmented Generation 20 Mar 2025 · 1 repository · arXiv:2503.16161Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Transformer-based Wireless Symbol Detection Over Fading Channels 20 Mar 2025 · 0 repositories · arXiv:2503.16594
-
Tuning LLMs by RAG Principles: Towards LLM-native Memory 20 Mar 2025 · 1 repository · arXiv:2503.16071
-
Typed-RAG: Type-aware Multi-Aspect Decomposition for Non-Factoid Question Answering 20 Mar 2025 · 1 repository · arXiv:2503.15879
-
Unify and Triumph: Polyglot, Diverse, and Self-Consistent Generation of Unit Tests with LLMs 20 Mar 2025 · 0 repositories · arXiv:2503.16144
-
UniHDSA: A Unified Relation Prediction Approach for Hierarchical Document Structure Analysis 20 Mar 2025 · 1 repository · arXiv:2503.15893
-
XAttention: Block Sparse Attention with Antidiagonal Scoring 20 Mar 2025 · 1 repository · arXiv:2503.16428
-
A Novel Channel Boosted Residual CNN-Transformer with Regional-Boundary Learning for Breast Cancer Detection 19 Mar 2025 · 0 repositories · arXiv:2503.15008
-
ChatGPT or A Silent Everywhere Helper: A Survey of Large Language Models 19 Mar 2025 · 0 repositories · arXiv:2503.17403
-
ELTEX: A Framework for Domain-Driven Synthetic Data Generation 19 Mar 2025 · 1 repository · arXiv:2503.15055
-
Enhancing Code LLM Training with Programmer Attention 19 Mar 2025 · 0 repositories · arXiv:2503.14936
-
Enhancing Pancreatic Cancer Staging with Large Language Models: The Role of Retrieval-Augmented Generation 19 Mar 2025 · 0 repositories · arXiv:2503.15664
-
Bias Evaluation and Mitigation in Retrieval-Augmented Medical Question-Answering Systems 19 Mar 2025 · 0 repositories · arXiv:2503.15454
-
FP4DiT: Towards Effective Floating Point Quantization for Diffusion Transformers 19 Mar 2025 · 1 repository · arXiv:2503.15465Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
GenM³: Generative Pretrained Multi-path Motion Model for Text Conditional Human Motion Generation 19 Mar 2025 · 0 repositories · arXiv:2503.14919
-
Optimizing Retrieval Strategies for Financial Question Answering Documents in Retrieval-Augmented Generation Systems 19 Mar 2025 · 1 repository · arXiv:2503.15191Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
RAG-based User Profiling for Precision Planning in Mixed-precision Over-the-Air Federated Learning 19 Mar 2025 · 0 repositories · arXiv:2503.15569