Methods › General › Attention Modules › Multi-Head Attention › Papers, page 12
Multi-Head Attention
Papers archive 2025-07-28
archive papers tagged: 24,855 · with a code link: 11,214 · where Syntology ran a sample: 3,454 (2,916 with a run with no instrument failure, 538 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,454 of 24,855 tagged: 2,916 with a run with no instrument failure, 538 where every run was a failure of Syntology's instrument)
Page 12 of 249: papers 1,101 to 1,200 of 24,855, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Coarse-to-Fine Learning for Multi-Pipette Localisation in Robot-Assisted In Vivo Patch-Clamp 31 Mar 2025 · 0 repositories · arXiv:2504.01044
-
Comparing representations of long clinical texts for the task of patient note-identification 31 Mar 2025 · 0 repositories · arXiv:2503.24006
-
Conformal uncertainty quantification to evaluate predictive fairness of foundation AI model for skin lesion classes across patient demographics 31 Mar 2025 · 0 repositories · arXiv:2503.23819
-
CrossFormer: Cross-Segment Semantic Fusion for Document Segmentation 31 Mar 2025 · 0 repositories · arXiv:2503.23671
-
Does "Reasoning" with Large Language Models Improve Recognizing, Generating, and Reframing Unhelpful Thoughts? 31 Mar 2025 · 0 repositories · arXiv:2504.00163
-
Easi3R: Estimating Disentangled Motion from DUSt3R Without Training 31 Mar 2025 · 1 repository · arXiv:2503.24391
-
Enhancing Large Language Models (LLMs) for Telecommunications using Knowledge Graphs and Retrieval-Augmented Generation 31 Mar 2025 · 0 repositories · arXiv:2503.24245
-
Foundation Models For Seismic Data Processing: An Extensive Review 31 Mar 2025 · 1 repository · arXiv:2503.24166
-
Frequency-Aware Attention-LSTM for PM_(2.5) Time Series Forecasting 31 Mar 2025 · 0 repositories · arXiv:2503.24043
-
Graph Transformer-Based Flood Susceptibility Mapping: Application to the French Riviera and Railway Infrastructure Under Climate Change 31 Mar 2025 · 0 repositories · arXiv:2504.03727
-
JudgeLRM: Large Reasoning Models as a Judge 31 Mar 2025 · 0 repositories · arXiv:2504.00050
-
Large Language Models Pass the Turing Test 31 Mar 2025 · 0 repositories · arXiv:2503.23674
-
LLM4FS: Leveraging Large Language Models for Feature Selection and How to Improve It 31 Mar 2025 · 0 repositories · arXiv:2503.24157
-
NeuRaLaTeX: A machine learning library written in pure LaTeX 31 Mar 2025 · 0 repositories · arXiv:2503.24187
-
Rubric Is All You Need: Enhancing LLM-based Code Evaluation With Question-Specific Rubrics 31 Mar 2025 · 0 repositories · arXiv:2503.23989
-
Synthetic News Generation for Fake News Classification 31 Mar 2025 · 0 repositories · arXiv:2503.24206
-
Text Chunking for Document Classification for Urban System Management using Large Language Models 31 Mar 2025 · 1 repository · arXiv:2504.00274
-
TransMamba: Flexibly Switching between Transformer and Mamba 31 Mar 2025 · 0 repositories · arXiv:2503.24067
-
UltraRAG: A Modular and Automated Toolkit for Adaptive Retrieval-Augmented Generation 31 Mar 2025 · 1 repository · arXiv:2504.08761
-
A Lightweight Image Super-Resolution Transformer Trained on Low-Resolution Images Only 30 Mar 2025 · 1 repository · arXiv:2503.23265
-
Advancing Sentiment Analysis in Tamil-English Code-Mixed Texts: Challenges and Transformer-Based Solutions 30 Mar 2025 · 0 repositories · arXiv:2503.23295
-
Beyond Detection: Designing AI-Resilient Assessments with Automated Feedback Tool to Foster Critical Thinking 30 Mar 2025 · 0 repositories · arXiv:2503.23622
-
CADFormer: Fine-Grained Cross-modal Alignment and Decoding Transformer for Referring Remote Sensing Image Segmentation 30 Mar 2025 · 0 repositories · arXiv:2503.23456
-
Exploring GPT-4 for Robotic Agent Strategy with Real-Time State Feedback and a Reactive Behaviour Framework 30 Mar 2025 · 0 repositories · arXiv:2503.23601
-
FeRG-LLM : Feature Engineering by Reason Generation Large Language Models 30 Mar 2025 · 0 repositories · arXiv:2503.23371
-
HiPART: Hierarchical Pose AutoRegressive Transformer for Occluded 3D Human Pose Estimation 30 Mar 2025 · 0 repositories · arXiv:2503.23331
-
Hyper-RAG: Combating LLM Hallucinations using Hypergraph-Driven Retrieval-Augmented Generation 30 Mar 2025 · 0 repositories · arXiv:2504.08758
-
JavisDiT: Joint Audio-Video Diffusion Transformer with Hierarchical Spatio-Temporal Prior Synchronization 30 Mar 2025 · 0 repositories · arXiv:2503.23377
-
Large Language Models Are Better Logical Fallacy Reasoners with Counterargument, Explanation, and Goal-Aware Prompt Formulation 30 Mar 2025 · 1 repository · arXiv:2503.23363
-
LaViC: Adapting Large Vision-Language Models to Visually-Aware Conversational Recommendation 30 Mar 2025 · 1 repository · arXiv:2503.23312
-
Measuring Online Hate on 4chan using Pre-trained Deep Learning Models 30 Mar 2025 · 0 repositories · arXiv:2504.00045
-
Multi-Stakeholder Disaster Insights from Social Media Using Large Language Models 30 Mar 2025 · 0 repositories · arXiv:2504.00046
-
Object Isolated Attention for Consistent Story Visualization 30 Mar 2025 · 0 repositories · arXiv:2503.23353
-
RARE: Retrieval-Augmented Reasoning Modeling 30 Mar 2025 · 1 repository · arXiv:2503.23513Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 0 honoured, 0 violated, 13 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 14 harvested samples)
-
SCORE: Story Coherence and Retrieval Enhancement for AI Narratives 30 Mar 2025 · 0 repositories · arXiv:2503.23512
-
A Training-free LLM Framework with Interaction between Contextually Related Subtasks in Solving Complex Tasks 29 Mar 2025 · 0 repositories · arXiv:2503.23053
-
Efficient Adaptation For Remote Sensing Visual Grounding 29 Mar 2025 · 0 repositories · arXiv:2503.23083
-
Enhancing Knowledge Graph Completion with Entity Neighborhood and Relation Context 29 Mar 2025 · 0 repositories · arXiv:2503.23205
-
Large Self-Supervised Models Bridge the Gap in Domain Adaptive Object Detection 29 Mar 2025 · 1 repository · arXiv:2503.23220
-
MHTS: Multi-Hop Tree Structure Framework for Generating Difficulty-Controllable QA Datasets for RAG Evaluation 29 Mar 2025 · 0 repositories · arXiv:2504.08756
-
Multimodal machine learning with large language embedding model for polymer property prediction 29 Mar 2025 · 1 repository · arXiv:2503.22962
-
The geomagnetic storm and Kp prediction using Wasserstein transformer 29 Mar 2025 · 0 repositories · arXiv:2503.23102
-
The realization of tones in spontaneous spoken Taiwan Mandarin: a corpus-based survey and theory-driven computational modeling 29 Mar 2025 · 0 repositories · arXiv:2503.23163
-
Z-SASLM: Zero-Shot Style-Aligned SLI Blending Latent Manipulation 29 Mar 2025 · 1 repository · arXiv:2503.23234
-
An Advanced Ensemble Deep Learning Framework for Stock Price Prediction Using VAE, Transformer, and LSTM Model 28 Mar 2025 · 0 repositories · arXiv:2503.22192
-
AnnoPage Dataset: Dataset of Non-Textual Elements in Documents with Fine-Grained Categorization 28 Mar 2025 · 0 repositories · arXiv:2503.22526
-
Bridging the Dimensional Chasm: Uncover Layer-wise Dimensional Reduction in Transformers through Token Correlation 28 Mar 2025 · 0 repositories · arXiv:2503.22547
-
Camera Model Identification with SPAIR-Swin and Entropy based Non-Homogeneous Patches 28 Mar 2025 · 0 repositories · arXiv:2503.22120
-
Correlation-Attention Masked Temporal Transformer for User Identity Linkage Using Heterogeneous Mobility Data 28 Mar 2025 · 1 repository · arXiv:2504.01979
-
DeepOFormer: Deep Operator Learning with Domain-informed Features for Fatigue Life Prediction 28 Mar 2025 · 0 repositories · arXiv:2503.22475
-
DREMnet: An Interpretable Denoising Framework for Semi-Airborne Transient Electromagnetic Signal 28 Mar 2025 · 0 repositories · arXiv:2503.22223
-
EdgeInfinite: A Memory-Efficient Infinite-Context Transformer for Edge Devices 28 Mar 2025 · 0 repositories · arXiv:2503.22196
-
Historical Ink: Exploring Large Language Models for Irony Detection in 19th-Century Spanish 28 Mar 2025 · 1 repository · arXiv:2503.22585
-
How Well Can Vison-Language Models Understand Humans' Intention? An Open-ended Theory of Mind Question Evaluation Benchmark 28 Mar 2025 · 0 repositories · arXiv:2503.22093
-
Integrating Artificial Intelligence with Human Expertise: An In-depth Analysis of ChatGPT's Capabilities in Generating Metamorphic Relations 28 Mar 2025 · 0 repositories · arXiv:2503.22141
-
Leveraging LLMs for Predicting Unknown Diagnoses from Clinical Notes 28 Mar 2025 · 0 repositories · arXiv:2503.22092
-
Scenario Dreamer: Vectorized Latent Diffusion for Generating Driving Simulation Environments 28 Mar 2025 · 0 repositories · arXiv:2503.22496
-
Understanding Inequality of LLM Fact-Checking over Geographic Regions with Agent and Retrieval models 28 Mar 2025 · 0 repositories · arXiv:2503.22877
-
An evaluation of LLMs and Google Translate for translation of selected Indian languages via sentiment and semantic analyses 27 Mar 2025 · 0 repositories · arXiv:2503.21393
-
As easy as PIE: understanding when pruning causes language models to disagree 27 Mar 2025 · 1 repository · arXiv:2503.21714
-
Collab: Controlled Decoding using Mixture of Agents for LLM Alignment 27 Mar 2025 · 0 repositories · arXiv:2503.21720
-
From Individual to Group: Developing a Context-Aware Multi-Criteria Group Recommender System 27 Mar 2025 · 0 repositories · arXiv:2503.22752
-
Hybrid Emotion Recognition: Enhancing Customer Interactions Through Acoustic and Textual Analysis 27 Mar 2025 · 0 repositories · arXiv:2503.21927
-
HyperGraphRAG: Retrieval-Augmented Generation with Hypergraph-Structured Knowledge Representation 27 Mar 2025 · 1 repository · arXiv:2503.21322Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Integrating Travel Behavior Forecasting and Generative Modeling for Predicting Future Urban Mobility and Spatial Transformations 27 Mar 2025 · 0 repositories · arXiv:2503.21158
-
MemInsight: Autonomous Memory Augmentation for LLM Agents 27 Mar 2025 · 0 repositories · arXiv:2503.21760
-
Molecular Quantum Transformer 27 Mar 2025 · 0 repositories · arXiv:2503.21686
-
Real-Time Evaluation Models for RAG: Who Detects Hallucinations Best? 27 Mar 2025 · 0 repositories · arXiv:2503.21157
-
ReaRAG: Knowledge-guided Reasoning Enhances Factuality of Large Reasoning Models with Iterative Retrieval Augmented Generation 27 Mar 2025 · 1 repository · arXiv:2503.21729
-
ReCoM: Realistic Co-Speech Motion Generation with Recurrent Embedded Transformer 27 Mar 2025 · 0 repositories · arXiv:2503.21847
-
Retinal Fundus Multi-Disease Image Classification using Hybrid CNN-Transformer-Ensemble Architectures 27 Mar 2025 · 1 repository · arXiv:2503.21465
-
Using large language models to produce literature reviews: Usages and systematic biases of microphysics parametrizations in 2699 publications 27 Mar 2025 · 0 repositories · arXiv:2503.21352
-
VALLR: Visual ASR Language Model for Lip Reading 27 Mar 2025 · 0 repositories · arXiv:2503.21408
-
Vision Language Models versus Machine Learning Models Performance on Polyp Detection and Classification in Colonoscopy Images 27 Mar 2025 · 1 repository · arXiv:2503.21840
-
A Survey of Multimodal Retrieval-Augmented Generation 26 Mar 2025 · 0 repositories · arXiv:2504.08748
-
Advancements in Natural Language Processing: Exploring Transformer-Based Architectures for Text Understanding 26 Mar 2025 · 0 repositories · arXiv:2503.20227
-
Advancing Vulnerability Classification with BERT: A Multi-Objective Learning Model 26 Mar 2025 · 0 repositories · arXiv:2503.20831
-
Can We Make Code Green? Understanding Trade-Offs in LLMs vs. Human Code Optimizations 26 Mar 2025 · 0 repositories · arXiv:2503.20126
-
CNN+Transformer Based Anomaly Traffic Detection in UAV Networks for Emergency Rescue 26 Mar 2025 · 0 repositories · arXiv:2503.20355
-
Novel Deep Neural OFDM Receiver Architectures for LLR Estimation 26 Mar 2025 · 1 repository · arXiv:2503.20500
-
Devil is in the Uniformity: Exploring Diverse Learners within Transformer for Image Restoration 26 Mar 2025 · 1 repository · arXiv:2503.20174Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
In vitro 2 In vivo : Bidirectional and High-Precision Generation of In Vitro and In Vivo Neuronal Spike Data 26 Mar 2025 · 0 repositories · arXiv:2503.20841
-
ITA-MDT: Image-Timestep-Adaptive Masked Diffusion Transformer Framework for Image-Based Virtual Try-On 26 Mar 2025 · 0 repositories · arXiv:2503.20418
-
Iterative Prompting with Persuasion Skills in Jailbreaking Large Language Models 26 Mar 2025 · 0 repositories · arXiv:2503.20320
-
MCTS-RAG: Enhancing Retrieval-Augmented Generation with Monte Carlo Tree Search 26 Mar 2025 · 1 repository · arXiv:2503.20757
-
MVFNet: Multipurpose Video Forensics Network using Multiple Forms of Forensic Evidence 26 Mar 2025 · 0 repositories · arXiv:2503.20991
-
Patients Speak, AI Listens: LLM-based Analysis of Online Reviews Uncovers Key Drivers for Urgent Care Satisfaction 26 Mar 2025 · 0 repositories · arXiv:2503.20981
-
Progressive Focused Transformer for Single Image Super-Resolution 26 Mar 2025 · 1 repository · arXiv:2503.20337Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 6 where Syntology's instrument failed) · 1 unverified (of 10 harvested samples) · 3 pointer-only (licence)
-
RALLRec+: Retrieval Augmented Large Language Model Recommendation with Reasoning 26 Mar 2025 · 1 repository · arXiv:2503.20430
-
RSRWKV: A Linear-Complexity 2D Attention Mechanism for Efficient Remote Sensing Vision Task 26 Mar 2025 · 0 repositories · arXiv:2503.20382
-
RGL: A Graph-Centric, Modular Framework for Efficient Retrieval-Augmented Generation on Graphs 25 Mar 2025 · 1 repository · arXiv:2503.19314
-
A novel forecasting framework combining virtual samples and enhanced Transformer models for tourism demand forecasting 25 Mar 2025 · 0 repositories · arXiv:2503.19423
-
BiblioPage: A Dataset of Scanned Title Pages for Bibliographic Metadata Extraction 25 Mar 2025 · 1 repository · arXiv:2503.19658
-
BugCraft: End-to-End Crash Bug Reproduction Using LLM Agents in Minecraft 25 Mar 2025 · 0 repositories · arXiv:2503.20036
-
CausalRAG: Integrating Causal Graphs into Retrieval-Augmented Generation 25 Mar 2025 · 0 repositories · arXiv:2503.19878
-
Context-Aware Semantic Segmentation: Enhancing Pixel-Level Understanding with Large Language Models for Advanced Vision Applications 25 Mar 2025 · 0 repositories · arXiv:2503.19276
-
Dita: Scaling Diffusion Transformer for Generalist Vision-Language-Action Policy 25 Mar 2025 · 1 repository · arXiv:2503.19757Syntology 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Enabling Rapid Shared Human-AI Mental Model Alignment via the After-Action Review 25 Mar 2025 · 1 repository · arXiv:2503.19607
-
Face Spoofing Detection using Deep Learning 25 Mar 2025 · 1 repository · arXiv:2503.19223
-
Fundamental Limits of Perfect Concept Erasure 25 Mar 2025 · 1 repository · arXiv:2503.20098