Methods › General › Attention Modules › Multi-Head Attention › Papers, page 66
Multi-Head Attention
Papers archive 2025-07-28
archive papers tagged: 24,855 · with a code link: 11,214 · where Syntology ran a sample: 3,454 (2,916 with a run with no instrument failure, 538 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,454 of 24,855 tagged: 2,916 with a run with no instrument failure, 538 where every run was a failure of Syntology's instrument)
Page 66 of 249: papers 6,501 to 6,600 of 24,855, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Data-Efficient Learning with Neural Programs 10 Jun 2024 · 1 repository · arXiv:2406.06246Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples)
-
Diving into Underwater: Segment Anything Model Guided Underwater Salient Instance Segmentation and A Large-scale Dataset 10 Jun 2024 · 1 repository · arXiv:2406.06039
-
Emotion-Aware Speech Self-Supervised Representation Learning with Intensity Knowledge 10 Jun 2024 · 0 repositories · arXiv:2406.06646
-
Husky: A Unified, Open-Source Language Agent for Multi-Step Reasoning 10 Jun 2024 · 1 repository · arXiv:2406.06469Syntology official (archive's flag): 5 ran · 6 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 2 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
In-Context Learning and Fine-Tuning GPT for Argument Mining 10 Jun 2024 · 1 repository · arXiv:2406.06699
-
Learning Physical Simulation with Message Passing Transformer 10 Jun 2024 · 0 repositories · arXiv:2406.06060
-
Leveraging Large Language Models for Knowledge-free Weak Supervision in Clinical Natural Language Processing 10 Jun 2024 · 0 repositories · arXiv:2406.06723
-
LLM-dCache: Improving Tool-Augmented LLMs with GPT-Driven Localized Data Caching 10 Jun 2024 · 0 repositories · arXiv:2406.06799
-
PointABM:Integrating Bidirectional State Space Model with Multi-Head Self-Attention for Point Cloud Analysis 10 Jun 2024 · 0 repositories · arXiv:2406.06069
-
SecureNet: A Comparative Study of DeBERTa and Large Language Models for Phishing Detection 10 Jun 2024 · 0 repositories · arXiv:2406.06663
-
Separate and Reconstruct: Asymmetric Encoder-Decoder for Speech Separation 10 Jun 2024 · 1 repository · arXiv:2406.05983
-
Symmetric Dot-Product Attention for Efficient Training of BERT Language Models 10 Jun 2024 · 0 repositories · arXiv:2406.06366
-
The Impact of Quantization on Retrieval-Augmented Generation: An Analysis of Small LLMs 10 Jun 2024 · 0 repositories · arXiv:2406.10251
-
UMBRELA: UMbrela is the (Open-Source Reproduction of the) Bing RELevance Assessor 10 Jun 2024 · 1 repository · arXiv:2406.06519
-
A Knowledge-Component-Based Methodology for Evaluating AI Assistants 9 Jun 2024 · 0 repositories · arXiv:2406.05603
-
Are Large Language Models Actually Good at Text Style Transfer? 9 Jun 2024 · 1 repository · arXiv:2406.05885
-
Attention as a Hypernetwork 9 Jun 2024 · 1 repository · arXiv:2406.05816Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
CAMS: Convolution and Attention-Free Mamba-based Cardiac Image Segmentation 9 Jun 2024 · 1 repository · arXiv:2406.05786
-
DomainRAG: A Chinese Benchmark for Evaluating Domain-specific Retrieval-Augmented Generation 9 Jun 2024 · 2 repositories · arXiv:2406.05654Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
Exploring the Efficacy of Large Language Models (GPT-4) in Binary Reverse Engineering 9 Jun 2024 · 0 repositories · arXiv:2406.06637
-
GCtx-UNet: Efficient Network for Medical Image Segmentation 9 Jun 2024 · 1 repository · arXiv:2406.05891
-
Hidden Holes: topological aspects of language models 9 Jun 2024 · 0 repositories · arXiv:2406.05798
-
Large Language Models Memorize Sensor Datasets! Implications on Human Activity Recognition Research 9 Jun 2024 · 0 repositories · arXiv:2406.05900
-
Machine Against the RAG: Jamming Retrieval-Augmented Generation with Blocker Documents 9 Jun 2024 · 0 repositories · arXiv:2406.05870
-
MedREQAL: Examining Medical Knowledge Recall of Large Language Models via Question Answering 9 Jun 2024 · 0 repositories · arXiv:2406.05845
-
OD-DETR: Online Distillation for Stabilizing Training of Detection Transformer 9 Jun 2024 · 0 repositories · arXiv:2406.05791
-
RE-RAG: Improving Open-Domain QA Performance and Interpretability with Relevance Estimator in Retrieval-Augmented Generation 9 Jun 2024 · 1 repository · arXiv:2406.05794
-
SinkLoRA: Enhanced Efficiency and Chat Capabilities for Long-Context Large Language Models 9 Jun 2024 · 1 repository · arXiv:2406.05678Syntology official: no sample here; runs from other or unrecorded repositories · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Smiles2Dock: an open large-scale multi-task dataset for ML-based molecular docking 9 Jun 2024 · 1 repository · arXiv:2406.05738
-
Text2VP: Generative AI for Visual Programming and Parametric Modeling 9 Jun 2024 · 0 repositories · arXiv:2407.07732
-
Vision Mamba: Cutting-Edge Classification of Alzheimer's Disease with 3D MRI Scans 9 Jun 2024 · 0 repositories · arXiv:2406.05757
-
1st Place Winner of the 2024 Pixel-level Video Understanding in the Wild (CVPR'24 PVUW) Challenge in Video Panoptic Segmentation and Best Long Video Consistency of Video Semantic Segmentation 8 Jun 2024 · 0 repositories · arXiv:2406.05352
-
A Fine-tuning Dataset and Benchmark for Large Language Models for Protein Understanding 8 Jun 2024 · 1 repository · arXiv:2406.05540Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples)
-
Advancing Semantic Textual Similarity Modeling: A Regression Framework with Translated ReLU and Smooth K2 Loss 8 Jun 2024 · 2 repositories · arXiv:2406.05326
-
Automata Extraction from Transformers 8 Jun 2024 · 1 repository · arXiv:2406.05564
-
Benchmarking Neural Decoding Backbones towards Enhanced On-edge iBCI Applications 8 Jun 2024 · 0 repositories · arXiv:2406.06626
-
Concept Formation and Alignment in Language Models: Bridging Statistical Patterns in Latent Space to Concept Taxonomy 8 Jun 2024 · 0 repositories · arXiv:2406.05315
-
Critical Phase Transition in Large Language Models 8 Jun 2024 · 0 repositories · arXiv:2406.05335
-
Do LLMs Recognize me, When I is not me: Assessment of LLMs Understanding of Turkish Indexical Pronouns in Indexical Shift Contexts 8 Jun 2024 · 0 repositories · arXiv:2406.05569
-
G-Transformer: Counterfactual Outcome Prediction under Dynamic and Time-varying Treatment Regimes 8 Jun 2024 · 0 repositories · arXiv:2406.05504
-
MaTableGPT: GPT-based Table Data Extractor from Materials Science Literature 8 Jun 2024 · 0 repositories · arXiv:2406.05431
-
SelfDefend: LLMs Can Defend Themselves against Jailbreaking in a Practical Manner 8 Jun 2024 · 0 repositories · arXiv:2406.05498
-
Teaching-Assistant-in-the-Loop: Improving Knowledge Distillation from Imperfect Teacher Models in Low-Budget Scenarios 8 Jun 2024 · 0 repositories · arXiv:2406.05322
-
Toward Reliable Ad-hoc Scientific Information Extraction: A Case Study on Two Materials Datasets 8 Jun 2024 · 1 repository · arXiv:2406.05348
-
Transformer Conformal Prediction for Time Series 8 Jun 2024 · 1 repository · arXiv:2406.05332Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
U-Net Ensemble for Enhanced Semantic Segmentation in Remote Sensing Imagery 8 Jun 2024 · 0 repositories
-
VP-LLM: Text-Driven 3D Volume Completion with Large Language Models through Patchification 8 Jun 2024 · 0 repositories · arXiv:2406.05543
-
Are Large Language Models More Empathetic than Humans? 7 Jun 2024 · 0 repositories · arXiv:2406.05063
-
BAMO at SemEval-2024 Task 9: BRAINTEASER: A Novel Task Defying Common Sense 7 Jun 2024 · 1 repository · arXiv:2406.04947
-
BERTs are Generative In-Context Learners 7 Jun 2024 · 1 repository · arXiv:2406.04823Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 1 violated, 11 with no contract checked; 1 where Syntology's instrument failed) · 13 unverified (of 26 harvested samples)
-
Conti-Fuse: A Novel Continuous Decomposition-based Fusion Framework for Infrared and Visible Images 7 Jun 2024 · 0 repositories · arXiv:2406.04689
-
Corpus Poisoning via Approximate Greedy Gradient Descent 7 Jun 2024 · 1 repository · arXiv:2406.05087
-
CRAG -- Comprehensive RAG Benchmark 7 Jun 2024 · 2 repositories · arXiv:2406.04744Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
CTSyn: A Foundational Model for Cross Tabular Data Generation 7 Jun 2024 · 0 repositories · arXiv:2406.04619
-
DiNeR: a Large Realistic Dataset for Evaluating Compositional Generalization 7 Jun 2024 · 1 repository · arXiv:2406.04669
-
Diving Deep into the Motion Representation of Video-Text Models 7 Jun 2024 · 1 repository · arXiv:2406.05075
-
GameBench: Evaluating Strategic Reasoning Abilities of LLM Agents 7 Jun 2024 · 2 repositories · arXiv:2406.06613Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Hints-In-Browser: Benchmarking Language Models for Programming Feedback Generation 7 Jun 2024 · 0 repositories · arXiv:2406.05053
-
DALD: Improving Logits-based Detector without Logits from Black-box LLMs 7 Jun 2024 · 1 repository · arXiv:2406.05232
-
Large Generative Graph Models 7 Jun 2024 · 0 repositories · arXiv:2406.05109
-
LlavaGuard: An Open VLM-based Framework for Safeguarding Vision Datasets and Models 7 Jun 2024 · 1 repository · arXiv:2406.05113Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
LLMs Are Not Intelligent Thinkers: Introducing Mathematical Topic Tree Benchmark for Comprehensive Evaluation of LLMs 7 Jun 2024 · 1 repository · arXiv:2406.05194
-
Logic Synthesis with Generative Deep Neural Networks 7 Jun 2024 · 0 repositories · arXiv:2406.04699
-
Low-Resource Cross-Lingual Summarization through Few-Shot Learning with Large Language Models 7 Jun 2024 · 0 repositories · arXiv:2406.04630
-
Mixture-of-Agents Enhances Large Language Model Capabilities 7 Jun 2024 · 3 repositories · arXiv:2406.04692Syntology 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 3 pointer-only (licence)
-
Multi-Head RAG: Solving Multi-Aspect Problems with LLMs 7 Jun 2024 · 2 repositories · arXiv:2406.05085
-
Multiplane Prior Guided Few-Shot Aerial Scene Rendering 7 Jun 2024 · 0 repositories · arXiv:2406.04961
-
On Subjective Uncertainty Quantification and Calibration in Natural Language Generation 7 Jun 2024 · 1 repository · arXiv:2406.05213Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 0 violated, 12 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 14 harvested samples)
-
Pretraining Decision Transformers with Reward Prediction for In-Context Multi-task Structured Bandit Learning 7 Jun 2024 · 0 repositories · arXiv:2406.05064
-
REP: Resource-Efficient Prompting for Rehearsal-Free Continual Learning 7 Jun 2024 · 0 repositories · arXiv:2406.04772
-
Towards Interpretable Deep Local Learning with Successive Gradient Reconciliation 7 Jun 2024 · 0 repositories · arXiv:2406.05222
-
UniTST: Effectively Modeling Inter-Series and Intra-Series Dependencies for Multivariate Time Series Forecasting 7 Jun 2024 · 0 repositories · arXiv:2406.04975
-
VTrans: Accelerating Transformer Compression with Variational Information Bottleneck based Pruning 7 Jun 2024 · 0 repositories · arXiv:2406.05276
-
A Survey on Medical Large Language Models: Technology, Application, Trustworthiness, and Future Directions 6 Jun 2024 · 0 repositories · arXiv:2406.03712
-
ABEX: Data Augmentation for Low-Resource NLU via Expanding Abstract Descriptions 6 Jun 2024 · 1 repository · arXiv:2406.04286Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
Are Graphs and GCNs necessary for short-term metro ridership forecasting? 6 Jun 2024 · 1 repository
-
Benchmark Data Contamination of Large Language Models: A Survey 6 Jun 2024 · 0 repositories · arXiv:2406.04244
-
Characterizing Similarities and Divergences in Conversational Tones in Humans and LLMs by Sampling with People 6 Jun 2024 · 1 repository · arXiv:2406.04278
-
Credit Card Fraud Detection Using Advanced Transformer Model 6 Jun 2024 · 0 repositories · arXiv:2406.03733
-
Cross-variable Linear Integrated ENhanced Transformer for Photovoltaic power forecasting 6 Jun 2024 · 0 repositories · arXiv:2406.03808
-
Decoder-only Streaming Transformer for Simultaneous Translation 6 Jun 2024 · 1 repository · arXiv:2406.03878Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
DeepStack: Deeply Stacking Visual Tokens is Surprisingly Simple and Effective for LMMs 6 Jun 2024 · 0 repositories · arXiv:2406.04334
-
Do Language Models Understand Morality? Towards a Robust Detection of Moral Content 6 Jun 2024 · 1 repository · arXiv:2406.04143
-
Empirical Guidelines for Deploying LLMs onto Resource-constrained Edge Devices 6 Jun 2024 · 0 repositories · arXiv:2406.03777
-
Enhancing In-Context Learning Performance with just SVD-Based Weight Pruning: A Theoretical Perspective 6 Jun 2024 · 1 repository · arXiv:2406.03768Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Exploring the Latest LLMs for Leaderboard Extraction 6 Jun 2024 · 0 repositories · arXiv:2406.04383
-
Generalization-Enhanced Code Vulnerability Detection via Multi-Task Instruction Fine-Tuning 6 Jun 2024 · 1 repository · arXiv:2406.03718
-
GLINT-RU: Gated Lightweight Intelligent Recurrent Units for Sequential Recommender Systems 6 Jun 2024 · 0 repositories · arXiv:2406.10244
-
HORAE: A Domain-Agnostic Language for Automated Service Regulation 6 Jun 2024 · 1 repository · arXiv:2406.06600
-
LLMEmbed: Rethinking Lightweight LLM's Genuine Function in Text Classification 6 Jun 2024 · 1 repository · arXiv:2406.03725
-
NATURAL PLAN: Benchmarking LLMs on Natural Language Planning 6 Jun 2024 · 0 repositories · arXiv:2406.04520
-
Proactive Detection of Physical Inter-rule Vulnerabilities in IoT Services Using a Deep Learning Approach 6 Jun 2024 · 0 repositories · arXiv:2406.03836
-
ReDistill: Residual Encoded Distillation for Peak Memory Reduction 6 Jun 2024 · 0 repositories · arXiv:2406.03744
-
RoboCoder: Robotic Learning from Basic Skills to General Tasks with Large Language Models 6 Jun 2024 · 0 repositories · arXiv:2406.03757
-
Scaling and evaluating sparse autoencoders 6 Jun 2024 · 5 repositories · arXiv:2406.04093Syntology official (archive's flag): 1 ran · 7 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 10 harvested samples) · 4 pointer-only (licence)
-
Simplified and Generalized Masked Diffusion for Discrete Data 6 Jun 2024 · 1 repository · arXiv:2406.04329Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 0 violated, 1 with no contract checked; 4 where Syntology's instrument failed) · 15 unverified (of 21 harvested samples)
-
Tool-Planner: Task Planning with Clusters across Multiple Tools 6 Jun 2024 · 1 repository · arXiv:2406.03807Syntology official (archive's flag): 7 ran · 7 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 6 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Tox-BART: Leveraging Toxicity Attributes for Explanation Generation of Implicit Hate Speech 6 Jun 2024 · 1 repository · arXiv:2406.03953
-
Transformers need glasses! Information over-squashing in language tasks 6 Jun 2024 · 0 repositories · arXiv:2406.04267
-
TwinS: Revisiting Non-Stationarity in Multivariate Time Series Forecasting 6 Jun 2024 · 0 repositories · arXiv:2406.03710