Methods › General › Attention Modules › Multi-Head Attention › Papers, page 54
Multi-Head Attention
Papers archive 2025-07-28
archive papers tagged: 24,855 · with a code link: 11,214 · where Syntology ran a sample: 3,454 (2,916 with a run with no instrument failure, 538 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,454 of 24,855 tagged: 2,916 with a run with no instrument failure, 538 where every run was a failure of Syntology's instrument)
Page 54 of 249: papers 5,301 to 5,400 of 24,855, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Could ChatGPT get an Engineering Degree? Evaluating Higher Education Vulnerability to AI Assistants 7 Aug 2024 · 0 repositories · arXiv:2408.11841
-
Early Prediction of Causes (not Effects) in Healthcare by Long-Term Clinical Time Series Forecasting 7 Aug 2024 · 1 repository · arXiv:2408.03816
-
VulScribeR: Exploring RAG-based Vulnerability Augmentation with LLMs 7 Aug 2024 · 1 repository · arXiv:2408.04125
-
FMiFood: Multi-modal Contrastive Learning for Food Image Classification 7 Aug 2024 · 0 repositories · arXiv:2408.03922
-
No-Reference Image Quality Assessment with Global-Local Progressive Integration and Semantic-Aligned Quality Transfer 7 Aug 2024 · 1 repository · arXiv:2408.03885
-
Image-to-LaTeX Converter for Mathematical Formulas and Text 7 Aug 2024 · 1 repository · arXiv:2408.04015
-
Inter-Series Transformer: Attending to Products in Time Series Forecasting 7 Aug 2024 · 0 repositories · arXiv:2408.03872
-
Is Child-Directed Speech Effective Training Data for Language Models? 7 Aug 2024 · 1 repository · arXiv:2408.03617Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 1 violated, 4 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples) · 2 pointer-only (licence)
-
JARViS: Detecting Actions in Video Using Unified Actor-Scene Context Relation Modeling 7 Aug 2024 · 0 repositories · arXiv:2408.03612
-
Query3D: LLM-Powered Open-Vocabulary Scene Segmentation with Language Embedded 3D Gaussian 7 Aug 2024 · 1 repository · arXiv:2408.03516
-
MaxMind: A Memory Loop Network to Enhance Software Productivity based on Large Language Models 7 Aug 2024 · 0 repositories · arXiv:2408.03841
-
PackMamba: Efficient Processing of Variable-Length Sequences in Mamba training 7 Aug 2024 · 0 repositories · arXiv:2408.03865
-
PaveCap: The First Multimodal Framework for Comprehensive Pavement Condition Assessment with Dense Captioning and PCI Estimation 7 Aug 2024 · 1 repository · arXiv:2408.04110
-
RailTrack-DaViT: A Vision Transformer-Based Approach for Automated Railway Track Defect Detection 7 Aug 2024 · 1 repository
-
SocFedGPT: Federated GPT-based Adaptive Content Filtering System Leveraging User Interactions in Social Networks 7 Aug 2024 · 0 repositories · arXiv:2408.05243
-
Surgformer: Surgical Transformer with Hierarchical Temporal Attention for Surgical Phase Recognition 7 Aug 2024 · 1 repository · arXiv:2408.03867
-
SwinShadow: Shifted Window for Ambiguous Adjacent Shadow Detection 7 Aug 2024 · 1 repository · arXiv:2408.03521
-
Intermediate direct preference optimization 6 Aug 2024 · 0 repositories · arXiv:2408.02923
-
Data Poisoning in LLMs: Jailbreak-Tuning and Scaling Laws 6 Aug 2024 · 2 repositories · arXiv:2408.02946Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Empathy Level Alignment via Reinforcement Learning for Empathetic Response Generation 6 Aug 2024 · 1 repository · arXiv:2408.02976
-
Analysis of Argument Structure Constructions in a Deep Recurrent Language Model 6 Aug 2024 · 0 repositories · arXiv:2408.03062
-
Topic Modeling with Fine-tuning LLMs and Bag of Sentences 6 Aug 2024 · 1 repository · arXiv:2408.03099Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Evaluating the Translation Performance of Large Language Models Based on Euas-20 6 Aug 2024 · 0 repositories · arXiv:2408.03119
-
Leveraging Entity Information for Cross-Modality Correlation Learning: The Entity-Guided Multimodal Summarization 6 Aug 2024 · 1 repository · arXiv:2408.03149
-
Leveraging Parameter Efficient Training Methods for Low Resource Text Classification: A Case Study in Marathi 6 Aug 2024 · 0 repositories · arXiv:2408.03172
-
ReSyncer: Rewiring Style-based Generator for Unified Audio-Visually Synced Facial Performer 6 Aug 2024 · 0 repositories · arXiv:2408.03284
-
MDT-A2G: Exploring Masked Diffusion Transformers for Co-Speech Gesture Generation 6 Aug 2024 · 0 repositories · arXiv:2408.03312
-
Advancing EEG-Based Gaze Prediction Using Depthwise Separable Convolution and Enhanced Pre-Processing 6 Aug 2024 · 1 repository · arXiv:2408.03480
-
Can LLMs Serve As Time Series Anomaly Detectors? 6 Aug 2024 · 0 repositories · arXiv:2408.03475
-
FLASH: Federated Learning-Based LLMs for Advanced Query Processing in Social Networks through RAG 6 Aug 2024 · 0 repositories · arXiv:2408.05242
-
LLM-Aided Compilation for Tensor Accelerators 6 Aug 2024 · 0 repositories · arXiv:2408.03408
-
LLM-based MOFs Synthesis Condition Extraction using Few-Shot Demonstrations 6 Aug 2024 · 0 repositories · arXiv:2408.04665
-
Set2Seq Transformer: Learning Permutation Aware Set Representations of Artistic Sequences 6 Aug 2024 · 0 repositories · arXiv:2408.03404
-
TF-Locoformer: Transformer with Local Modeling by Convolution for Speech Separation and Enhancement 6 Aug 2024 · 1 repository · arXiv:2408.03440
-
The Use of Large Language Models (LLM) for Cyber Threat Intelligence (CTI) in Cybercrime Forums 6 Aug 2024 · 0 repositories · arXiv:2408.03354
-
TrafficGPT: An LLM Approach for Open-Set Encrypted Traffic Classification 6 Aug 2024 · 1 repository
-
Training LLMs to Recognize Hedges in Spontaneous Narratives 6 Aug 2024 · 1 repository · arXiv:2408.03319
-
AssemAI: Interpretable Image-Based Anomaly Detection for Manufacturing Pipelines 5 Aug 2024 · 1 repository · arXiv:2408.02181
-
Is Large Language Model Good at Database Knob Tuning? A Comprehensive Experimental Evaluation 5 Aug 2024 · 0 repositories · arXiv:2408.02213
-
Cross-modulated Attention Transformer for RGBT Tracking 5 Aug 2024 · 0 repositories · arXiv:2408.02222
-
Do Large Language Models Speak All Languages Equally? A Comparative Study in Low-Resource Settings 5 Aug 2024 · 0 repositories · arXiv:2408.02237
-
DRFormer: Multi-Scale Transformer Utilizing Diverse Receptive Fields for Long Time-Series Forecasting 5 Aug 2024 · 1 repository · arXiv:2408.02279
-
The NPU-ASLP System Description for Visual Speech Recognition in CNVSRC 2024 5 Aug 2024 · 1 repository · arXiv:2408.02369
-
Why Are My Prompts Leaked? Unraveling Prompt Extraction Threats in Customized Large Language Models 5 Aug 2024 · 1 repository · arXiv:2408.02416
-
RAG Foundry: A Framework for Enhancing LLMs for Retrieval Augmented Generation 5 Aug 2024 · 2 repositories · arXiv:2408.02545
-
On Using Quasirandom Sequences in Machine Learning for Model Weight Initialization 5 Aug 2024 · 1 repository · arXiv:2408.02654
-
A Novel Hybrid Approach for Tornado Prediction in the United States: Kalman-Convolutional BiLSTM with Multi-Head Attention 5 Aug 2024 · 0 repositories · arXiv:2408.02751
-
Dimensionality Reduction and Nearest Neighbors for Improving Out-of-Distribution Detection in Medical Image Segmentation 5 Aug 2024 · 1 repository · arXiv:2408.02761
-
Wiping out the limitations of Large Language Models -- A Taxonomy for Retrieval Augmented Generation 5 Aug 2024 · 0 repositories · arXiv:2408.02854
-
AppAgent v2: Advanced Agent for Flexible Mobile Interactions 5 Aug 2024 · 0 repositories · arXiv:2408.11824
-
LLM Agents Improve Semantic Code Search 5 Aug 2024 · 0 repositories · arXiv:2408.11058
-
SEAS: Self-Evolving Adversarial Safety Optimization for Large Language Models 5 Aug 2024 · 1 repository · arXiv:2408.02632
-
Self-Taught Evaluators 5 Aug 2024 · 0 repositories · arXiv:2408.02666
-
The Mechanics of Conceptual Interpretation in GPT Models: Interpretative Insights 5 Aug 2024 · 0 repositories · arXiv:2408.11827
-
Toward Attention-based TinyML: A Heterogeneous Accelerated Architecture and Automated Deployment Flow 5 Aug 2024 · 1 repository · arXiv:2408.02473
-
XMainframe: A Large Language Model for Mainframe Modernization 5 Aug 2024 · 1 repository · arXiv:2408.04660
-
DiReCT: Diagnostic Reasoning for Clinical Notes via Large Language Models 4 Aug 2024 · 1 repository · arXiv:2408.01933Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 1 violated, 8 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 13 harvested samples) · 13 pointer-only (licence)
-
Defining and Evaluating Decision and Composite Risk in Language Models Applied to Natural Language Inference 4 Aug 2024 · 0 repositories · arXiv:2408.01935
-
ML-EAT: A Multilevel Embedding Association Test for Interpretable and Transparent Social Science 4 Aug 2024 · 1 repository · arXiv:2408.01966
-
AdaCBM: An Adaptive Concept Bottleneck Model for Explainable and Accurate Diagnosis 4 Aug 2024 · 1 repository · arXiv:2408.02001
-
MedSyn: LLM-based Synthetic Medical Text Generation Framework 4 Aug 2024 · 1 repository · arXiv:2408.02056
-
KAN-RCBEVDepth: A multi-modal fusion algorithm in object detection for autonomous driving 4 Aug 2024 · 0 repositories · arXiv:2408.02088
-
Effective Demonstration Annotation for In-Context Learning via Language Model-Based Determinantal Point Process 4 Aug 2024 · 0 repositories · arXiv:2408.02103
-
Leveraging Large Language Models with Chain-of-Thought and Prompt Engineering for Traffic Crash Severity Analysis and Inference 4 Aug 2024 · 0 repositories · arXiv:2408.04652
-
Advancing Mental Health Pre-Screening: A New Custom GPT for Psychological Distress Assessment 3 Aug 2024 · 0 repositories · arXiv:2408.01614
-
Self-Emotion Blended Dialogue Generation in Social Simulation Agents 3 Aug 2024 · 0 repositories · arXiv:2408.01633
-
Stimulating Imagination: Towards General-purpose Object Rearrangement 3 Aug 2024 · 0 repositories · arXiv:2408.01655
-
A Novel Evaluation Framework for Image2Text Generation 3 Aug 2024 · 0 repositories · arXiv:2408.01723
-
LAM3D: Leveraging Attention for Monocular 3D Object Detection 3 Aug 2024 · 0 repositories · arXiv:2408.01739
-
Indexing and Visualization of Climate Change Narratives Using BERT and Causal Extraction 3 Aug 2024 · 0 repositories · arXiv:2408.01745
-
GLDiTalker: Speech-Driven 3D Facial Animation with Graph Latent Diffusion Transformer 3 Aug 2024 · 0 repositories · arXiv:2408.01826
-
Tracking Emotional Dynamics in Chat Conversations: A Hybrid Approach using DistilBERT and Emoji Sentiment Analysis 3 Aug 2024 · 0 repositories · arXiv:2408.01838
-
Efficient Solutions For An Intriguing Failure of LLMs: Long Context Window Does Not Mean LLMs Can Analyze Long Sequences Flawlessly 3 Aug 2024 · 0 repositories · arXiv:2408.01866
-
MALADE: Orchestration of LLM-powered Agents with Retrieval Augmented Generation for Pharmacovigilance 3 Aug 2024 · 1 repository · arXiv:2408.01869
-
Building Trust in Mental Health Chatbots: Safety Metrics and LLM-Based Evaluation Tools 3 Aug 2024 · 0 repositories · arXiv:2408.04650
-
Distinguishing Chatbot from Human 3 Aug 2024 · 0 repositories · arXiv:2408.04647
-
JambaTalk: Speech-Driven 3D Talking Head Generation Based on Hybrid Transformer-Mamba Language Model 3 Aug 2024 · 0 repositories · arXiv:2408.01627
-
Cross-domain Named Entity Recognition via Graph Matching 2 Aug 2024 · 0 repositories · arXiv:2408.00981
-
POA: Pre-training Once for Models of All Sizes 2 Aug 2024 · 1 repository · arXiv:2408.01031
-
MambaST: A Plug-and-Play Cross-Spectral Spatial-Temporal Fuser for Efficient Pedestrian Detection 2 Aug 2024 · 1 repository · arXiv:2408.01037
-
Privacy-Preserving Split Learning with Vision Transformers using Patch-Wise Random and Noisy CutMix 2 Aug 2024 · 0 repositories · arXiv:2408.01040
-
LLM as Runtime Error Handler: A Promising Pathway to Adaptive Self-Healing of Software Systems 2 Aug 2024 · 0 repositories · arXiv:2408.01055
-
Leveraging Encoder-only Large Language Models for Mobile App Review Feature Extraction 2 Aug 2024 · 1 repository · arXiv:2408.01063
-
BioRAG: A RAG-LLM Framework for Biological Question Reasoning 2 Aug 2024 · 0 repositories · arXiv:2408.01107
-
An Efficient and Effective Transformer Decoder-Based Framework for Multi-Task Visual Grounding 2 Aug 2024 · 1 repository · arXiv:2408.01120Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 15 harvested samples) · 15 pointer-only (licence)
-
A Survey of Mamba 2 Aug 2024 · 0 repositories · arXiv:2408.01129
-
Rethinking Pre-Trained Feature Extractor Selection in Multiple Instance Learning for Whole Slide Image Classification 2 Aug 2024 · 1 repository · arXiv:2408.01167
-
Nested Music Transformer: Sequentially Decoding Compound Tokens in Symbolic Music and Audio Generation 2 Aug 2024 · 1 repository · arXiv:2408.01180
-
High-Throughput Phenotyping of Clinical Text Using Large Language Models 2 Aug 2024 · 0 repositories · arXiv:2408.01214
-
Multi-head Spatial-Spectral Mamba for Hyperspectral Image Classification 2 Aug 2024 · 1 repository · arXiv:2408.01224
-
HeteroMorpheus: Universal Control Based on Morphological Heterogeneity Modeling 2 Aug 2024 · 1 repository · arXiv:2408.01230
-
WaveMamba: Spatial-Spectral Wavelet Mamba for Hyperspectral Image Classification 2 Aug 2024 · 0 repositories · arXiv:2408.01231
-
RAGEval: Scenario Specific RAG Evaluation Dataset Generation Framework 2 Aug 2024 · 1 repository · arXiv:2408.01262
-
Underwater Object Detection Enhancement via Channel Stabilization 2 Aug 2024 · 1 repository · arXiv:2408.01293
-
Transformers are Universal In-context Learners 2 Aug 2024 · 0 repositories · arXiv:2408.01367
-
Spatial and Spatial-Spectral Morphological Mamba for Hyperspectral Image Classification 2 Aug 2024 · 2 repositories · arXiv:2408.01372
-
NOLO: Navigate Only Look Once 2 Aug 2024 · 0 repositories · arXiv:2408.01384
-
Pre-trained Language Models Improve the Few-shot Prompt Ability of Decision Transformer 2 Aug 2024 · 0 repositories · arXiv:2408.01402
-
THOR2: Topological Analysis for 3D Shape and Color-Based Human-Inspired Object Recognition in Unseen Environments 2 Aug 2024 · 1 repository · arXiv:2408.01579
-
Evaluating the Impact of Advanced LLM Techniques on AI-Lecture Tutors for a Robotics Course 2 Aug 2024 · 0 repositories · arXiv:2408.04645