Methods › General › Attention Modules › Multi-Head Attention › Papers, page 18
Multi-Head Attention
Papers archive 2025-07-28
archive papers tagged: 24,855 · with a code link: 11,214 · where Syntology ran a sample: 3,454 (2,916 with a run with no instrument failure, 538 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,454 of 24,855 tagged: 2,916 with a run with no instrument failure, 538 where every run was a failure of Syntology's instrument)
Page 18 of 249: papers 1,701 to 1,800 of 24,855, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
CalibRefine: Deep Learning-Based Online Automatic Targetless LiDAR-Camera Calibration with Iterative and Attention-Driven Post-Refinement 24 Feb 2025 · 1 repository · arXiv:2502.17648
-
CipherPrune: Efficient and Scalable Private Transformer Inference 24 Feb 2025 · 1 repository · arXiv:2502.16782Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Dimitra: Audio-driven Diffusion model for Expressive Talking Head Generation 24 Feb 2025 · 0 repositories · arXiv:2502.17198
-
Disentangling Visual Transformers: Patch-level Interpretability for Image Classification 24 Feb 2025 · 0 repositories · arXiv:2502.17196
-
ENACT-Heart -- ENsemble-based Assessment Using CNN and Transformer on Heart Sounds 24 Feb 2025 · 0 repositories · arXiv:2502.16914
-
Enhancing Image Matting in Real-World Scenes with Mask-Guided Iterative Refinement 24 Feb 2025 · 0 repositories · arXiv:2502.17093
-
Evaluating the Effect of Retrieval Augmentation on Social Biases 24 Feb 2025 · 0 repositories · arXiv:2502.17611
-
Functional BART with Shape Priors: A Bayesian Tree Approach to Constrained Functional Regression 24 Feb 2025 · 0 repositories · arXiv:2502.16888
-
GaussianFlowOcc: Sparse and Weakly Supervised Occupancy Estimation using Gaussian Splatting and Temporal Flow 24 Feb 2025 · 0 repositories · arXiv:2502.17288
-
LettuceDetect: A Hallucination Detection Framework for RAG Applications 24 Feb 2025 · 2 repositories · arXiv:2502.17125
-
LLM Inference Acceleration via Efficient Operation Fusion 24 Feb 2025 · 0 repositories · arXiv:2502.17728
-
Logic Haystacks: Probing LLMs Long-Context Logical Reasoning (Without Easily Identifiable Unrelated Padding) 24 Feb 2025 · 0 repositories · arXiv:2502.17169
-
MaxGlaViT: A novel lightweight vision transformer-based approach for early diagnosis of glaucoma stages from fundus images 24 Feb 2025 · 1 repository · arXiv:2502.17154
-
MDN: Mamba-Driven Dualstream Network For Medical Hyperspectral Image Segmentation 24 Feb 2025 · 0 repositories · arXiv:2502.17255
-
MEMERAG: A Multilingual End-to-End Meta-Evaluation Benchmark for Retrieval Augmented Generation 24 Feb 2025 · 1 repository · arXiv:2502.17163
-
Mitigating Bias in RAG: Controlling the Embedder 24 Feb 2025 · 1 repository · arXiv:2502.17390
-
Mutual Reinforcement of LLM Dialogue Synthesis and Summarization Capabilities for Few-Shot Dialogue Summarization 24 Feb 2025 · 0 repositories · arXiv:2502.17328
-
Towards Typologically Aware Rescoring to Mitigate Unfaithfulness in Lower-Resource Languages 24 Feb 2025 · 0 repositories · arXiv:2502.17664
-
Unraveling the geometry of visual relational reasoning 24 Feb 2025 · 1 repository · arXiv:2502.17382
-
A Fine-Tuning Approach for T5 Using Knowledge Graphs to Address Complex Tasks 23 Feb 2025 · 0 repositories · arXiv:2502.16484
-
A Split-Window Transformer for Multi-Model Sequence Spammer Detection using Multi-Model Variational Autoencoder 23 Feb 2025 · 0 repositories · arXiv:2502.16483
-
D2S-FLOW: Automated Parameter Extraction from Datasheets for SPICE Model Generation Using Large Language Models 23 Feb 2025 · 0 repositories · arXiv:2502.16540
-
AeroReformer: Aerial Referring Transformer for UAV-based Referring Image Segmentation 23 Feb 2025 · 1 repository · arXiv:2502.16680
-
Co-MTP: A Cooperative Trajectory Prediction Framework with Multi-Temporal Fusion for Autonomous Driving 23 Feb 2025 · 1 repository · arXiv:2502.16589
-
Code Summarization Beyond Function Level 23 Feb 2025 · 1 repository · arXiv:2502.16704
-
Dynamic LLM Routing and Selection based on User Preferences: Balancing Performance, Cost, and Ethics 23 Feb 2025 · 0 repositories · arXiv:2502.16696
-
GS-TransUNet: Integrated 2D Gaussian Splatting and Transformer UNet for Accurate Skin Lesion Analysis 23 Feb 2025 · 1 repository · arXiv:2502.16748
-
Layer-Wise Evolution of Representations in Fine-Tuned Transformers: Insights from Sparse AutoEncoders 23 Feb 2025 · 0 repositories · arXiv:2502.16722
-
Optimizing Retrieval-Augmented Generation of Medical Content for Spaced Repetition Learning 23 Feb 2025 · 0 repositories · arXiv:2503.01859
-
PMAT: Optimizing Action Generation Order in Multi-Agent Reinforcement Learning 23 Feb 2025 · 1 repository · arXiv:2502.16496
-
Reasoning about Affordances: Causal and Compositional Reasoning in LLMs 23 Feb 2025 · 0 repositories · arXiv:2502.16606
-
Retrieval-Augmented Visual Question Answering via Built-in Autoregressive Search Engines 23 Feb 2025 · 0 repositories · arXiv:2502.16641
-
Visual-RAG: Benchmarking Text-to-Image Retrieval Augmented Generation for Visual Knowledge Intensive Queries 23 Feb 2025 · 1 repository · arXiv:2502.16636
-
VPNeXt -- Rethinking Dense Decoding for Plain Vision Transformer 23 Feb 2025 · 0 repositories · arXiv:2502.16654
-
An End-to-End Homomorphically Encrypted Neural Network 22 Feb 2025 · 0 repositories · arXiv:2502.16176
-
Enhancing LLMs for Identifying and Prioritizing Important Medical Jargons from Electronic Health Record Notes Utilizing Data Augmentation 22 Feb 2025 · 0 repositories · arXiv:2502.16022
-
Iterative Auto-Annotation for Scientific Named Entity Recognition Using BERT-Based Models 22 Feb 2025 · 0 repositories · arXiv:2502.16312
-
RAG-Enhanced Collaborative LLM Agents for Drug Discovery 22 Feb 2025 · 0 repositories · arXiv:2502.17506
-
Uncertainty-Aware Fusion: An Ensemble Framework for Mitigating Hallucinations in Large Language Models 22 Feb 2025 · 0 repositories · arXiv:2503.05757
-
Vision Transformer Accelerator ASIC for Real-Time, Low-Power Sleep Staging 22 Feb 2025 · 0 repositories · arXiv:2502.16334
-
Worse than Zero-shot? A Fact-Checking Dataset for Evaluating the Robustness of RAG Against Misleading Retrievals 22 Feb 2025 · 0 repositories · arXiv:2502.16101
-
A Close Look at Decomposition-based XAI-Methods for Transformer Language Models 21 Feb 2025 · 2 repositories · arXiv:2502.15886
-
Auto-Bench: An Automated Benchmark for Scientific Discovery in LLMs 21 Feb 2025 · 0 repositories · arXiv:2502.15224
-
AutoMedPrompt: A New Framework for Optimizing LLM Medical Prompts Using Textual Gradients 21 Feb 2025 · 0 repositories · arXiv:2502.15944
-
BP-GPT: Auditory Neural Decoding Using fMRI-prompted LLM 21 Feb 2025 · 1 repository · arXiv:2502.15172Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Chain-of-Rank: Enhancing Large Language Models for Domain-Specific RAG in Edge Device 21 Feb 2025 · 0 repositories · arXiv:2502.15134
-
Comparative Analysis of Large Language Models for Context-Aware Code Completion using SAFIM Framework 21 Feb 2025 · 0 repositories · arXiv:2502.15243
-
Corrections Meet Explanations: A Unified Framework for Explainable Grammatical Error Correction 21 Feb 2025 · 0 repositories · arXiv:2502.15261
-
Cross-Format Retrieval-Augmented Generation in XR with LLMs for Context-Aware Maintenance Assistance 21 Feb 2025 · 0 repositories · arXiv:2502.15604
-
Empowering LLMs with Logical Reasoning: A Comprehensive Survey 21 Feb 2025 · 0 repositories · arXiv:2502.15652
-
Enhancing Domain-Specific Retrieval-Augmented Generation: Synthetic Data Generation and Evaluation using Reasoning Models 21 Feb 2025 · 1 repository · arXiv:2502.15854
-
Extraction multi-étiquettes de relations en utilisant des couches de Transformer 21 Feb 2025 · 0 repositories · arXiv:2502.15619
-
GNN-Coder: Boosting Semantic Code Retrieval with Combined GNNs and Transformer 21 Feb 2025 · 0 repositories · arXiv:2502.15202
-
Jeffrey's update rule as a minimizer of Kullback-Leibler divergence 21 Feb 2025 · 0 repositories · arXiv:2502.15504
-
Lightweight yet Efficient: An External Attentive Graph Convolutional Network with Positional Prompts for Sequential Recommendation 21 Feb 2025 · 1 repository · arXiv:2502.15331
-
Mantis: Lightweight Calibrated Foundation Model for User-Friendly Time Series Classification 21 Feb 2025 · 1 repository · arXiv:2502.15637Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples)
-
Optimizing Pre-Training Data Mixtures with Mixtures of Data Expert Models 21 Feb 2025 · 0 repositories · arXiv:2502.15950
-
Retrieval-Augmented Speech Recognition Approach for Domain Challenges 21 Feb 2025 · 0 repositories · arXiv:2502.15264
-
Robust Bias Detection in MLMs and its Application to Human Trait Ratings 21 Feb 2025 · 1 repository · arXiv:2502.15600
-
SentiFormer: Metadata Enhanced Transformer for Image Sentiment Analysis 21 Feb 2025 · 1 repository · arXiv:2502.15322
-
Single-pass Detection of Jailbreaking Input in Large Language Models 21 Feb 2025 · 0 repositories · arXiv:2502.15435
-
Soybean pod and seed counting in both outdoor fields and indoor laboratories using unions of deep neural networks 21 Feb 2025 · 0 repositories · arXiv:2502.15286
-
Tokenization is Sensitive to Language Variation 21 Feb 2025 · 0 repositories · arXiv:2502.15343
-
TransMamba: Fast Universal Architecture Adaption from Transformers to Mamba 21 Feb 2025 · 0 repositories · arXiv:2502.15130
-
TurboFuzzLLM: Turbocharging Mutation-based Fuzzing for Effectively Jailbreaking Large Language Models in Practice 21 Feb 2025 · 1 repository · arXiv:2502.18504
-
Utilizing Sequential Information of General Lab-test Results and Diagnoses History for Differential Diagnosis of Dementia 21 Feb 2025 · 0 repositories · arXiv:2502.15317
-
A Socratic RAG Approach to Connect Natural Language Queries on Research Topics with Knowledge Organization Systems 20 Feb 2025 · 0 repositories · arXiv:2502.15005
-
Argument-Based Comparative Question Answering Evaluation Benchmark 20 Feb 2025 · 0 repositories · arXiv:2502.14476
-
Bridging Text and Vision: A Multi-View Text-Vision Registration Approach for Cross-Modal Place Recognition 20 Feb 2025 · 1 repository · arXiv:2502.14195
-
DeepRTL: Bridging Verilog Understanding and Generation with a Unified Representation Model 20 Feb 2025 · 0 repositories · arXiv:2502.15832
-
Do LLMs Consider Security? An Empirical Study on Responses to Programming Questions 20 Feb 2025 · 0 repositories · arXiv:2502.14202
-
Entropy-UID: A Method for Optimizing Information Density 20 Feb 2025 · 0 repositories · arXiv:2502.14366
-
FIND: Fine-grained Information Density Guided Adaptive Retrieval-Augmented Generation for Disease Diagnosis 20 Feb 2025 · 0 repositories · arXiv:2502.14614
-
Forecasting Local Ionospheric Parameters Using Transformers 20 Feb 2025 · 1 repository · arXiv:2502.15093
-
From Knowledge Generation to Knowledge Verification: Examining the BioMedical Generative Capabilities of ChatGPT 20 Feb 2025 · 0 repositories · arXiv:2502.14714
-
From RAG to Memory: Non-Parametric Continual Learning for Large Language Models 20 Feb 2025 · 1 repository · arXiv:2502.14802Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
Full-Step-DPO: Self-Supervised Preference Optimization with Step-wise Rewards for Mathematical Reasoning 20 Feb 2025 · 0 repositories · arXiv:2502.14356
-
Hallucination Detection in Large Language Models with Metamorphic Relations 20 Feb 2025 · 0 repositories · arXiv:2502.15844
-
Hardware-Friendly Static Quantization Method for Video Diffusion Transformers 20 Feb 2025 · 0 repositories · arXiv:2502.15077
-
Is Relevance Propagated from Retriever to Generator in RAG? 20 Feb 2025 · 0 repositories · arXiv:2502.15025
-
KITAB-Bench: A Comprehensive Multi-Domain Benchmark for Arabic OCR and Document Understanding 20 Feb 2025 · 0 repositories · arXiv:2502.14949
-
Mechanistic Understanding of Language Models in Syntactic Code Completion 20 Feb 2025 · 0 repositories · arXiv:2502.18499
-
Multiscale Byte Language Models -- A Hierarchical Architecture for Causal Million-Length Sequence Modeling 20 Feb 2025 · 1 repository · arXiv:2502.14553
-
On the Influence of Context Size and Model Choice in Retrieval-Augmented Generation Systems 20 Feb 2025 · 1 repository · arXiv:2502.14759
-
PaperHelper: Knowledge-Based LLM QA Paper Reading Assistant 20 Feb 2025 · 0 repositories · arXiv:2502.14271
-
Predicting Fetal Birthweight from High Dimensional Data using Advanced Machine Learning 20 Feb 2025 · 0 repositories · arXiv:2502.14270
-
QUAD-LLM-MLTC: Large Language Models Ensemble Learning for Healthcare Text Multi-Label Classification 20 Feb 2025 · 0 repositories · arXiv:2502.14189
-
RelaCtrl: Relevance-Guided Efficient Control for Diffusion Transformers 20 Feb 2025 · 0 repositories · arXiv:2502.14377
-
Tabular Embeddings for Tables with Bi-Dimensional Hierarchical Metadata and Nesting 20 Feb 2025 · 0 repositories · arXiv:2502.15819
-
Towards Economical Inference: Enabling DeepSeek's Multi-Head Latent Attention in Any Transformer-based LLMs 20 Feb 2025 · 1 repository · arXiv:2502.14837Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 15 harvested samples)
-
WavRAG: Audio-Integrated Retrieval Augmented Generation for Spoken Dialogue Models 20 Feb 2025 · 0 repositories · arXiv:2502.14727
-
Adapting Large Language Models for Time Series Modeling via a Novel Parameter-efficient Adaptation Method 19 Feb 2025 · 0 repositories · arXiv:2502.13725
-
Are Large Language Models In-Context Graph Learners? 19 Feb 2025 · 0 repositories · arXiv:2502.13562
-
Building Age Estimation: A New Multi-Modal Benchmark Dataset and Community Challenge 19 Feb 2025 · 1 repository · arXiv:2502.13818
-
Capturing Rich Behavior Representations: A Dynamic Action Semantic-Aware Graph Transformer for Video Captioning 19 Feb 2025 · 0 repositories · arXiv:2502.13754
-
DH-RAG: A Dynamic Historical Context-Powered Retrieval-Augmented Generation Method for Multi-Turn Dialogue 19 Feb 2025 · 0 repositories · arXiv:2502.13847
-
Extracting Social Connections from Finnish Karelian Refugee Interviews Using LLMs 19 Feb 2025 · 0 repositories · arXiv:2502.13566
-
FairKV: Balancing Per-Head KV Cache for Fast Multi-GPU Inference 19 Feb 2025 · 0 repositories · arXiv:2502.15804
-
FlexTok: Resampling Images into 1D Token Sequences of Flexible Length 19 Feb 2025 · 0 repositories · arXiv:2502.13967
-
From Correctness to Comprehension: AI Agents for Personalized Error Diagnosis in Education 19 Feb 2025 · 0 repositories · arXiv:2502.13789