Methods › General › Attention Modules › Multi-Head Attention › Papers, page 88
Multi-Head Attention
Papers archive 2025-07-28
archive papers tagged: 24,855 · with a code link: 11,214 · where Syntology ran a sample: 3,454 (2,916 with a run with no instrument failure, 538 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,454 of 24,855 tagged: 2,916 with a run with no instrument failure, 538 where every run was a failure of Syntology's instrument)
Page 88 of 249: papers 8,701 to 8,800 of 24,855, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
EffLoc: Lightweight Vision Transformer for Efficient 6-DOF Camera Relocalization 21 Feb 2024 · 0 repositories · arXiv:2402.13537
-
Event-aware Video Corpus Moment Retrieval 21 Feb 2024 · 0 repositories · arXiv:2402.13566
-
Exploring ChatGPT and its Impact on Society 21 Feb 2024 · 0 repositories · arXiv:2403.14643
-
EyeTrans: Merging Human and Machine Attention for Neural Code Summarization 21 Feb 2024 · 1 repository · arXiv:2402.14096
-
FanOutQA: A Multi-Hop, Multi-Document Question Answering Benchmark for Large Language Models 21 Feb 2024 · 1 repository · arXiv:2402.14116Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Green AI: A Preliminary Empirical Study on Energy Consumption in DL Models Across Different Runtime Infrastructures 21 Feb 2024 · 0 repositories · arXiv:2402.13640
-
Hallucinations or Attention Misdirection? The Path to Strategic Value Extraction in Business Using Large Language Models 21 Feb 2024 · 0 repositories · arXiv:2402.14002
-
Improving Language Understanding from Screenshots 21 Feb 2024 · 1 repository · arXiv:2402.14073Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Knowledge Graph Enhanced Large Language Model Editing 21 Feb 2024 · 0 repositories · arXiv:2402.13593
-
Kuaiji: the First Chinese Accounting Large Language Model 21 Feb 2024 · 0 repositories · arXiv:2402.13866
-
Large Language Models for Data Annotation and Synthesis: A Survey 21 Feb 2024 · 1 repository · arXiv:2402.13446
-
A Comprehensive Study of Jailbreak Attack versus Defense for Large Language Models 21 Feb 2024 · 1 repository · arXiv:2402.13457Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
MSTAR: Multi-Scale Backbone Architecture Search for Timeseries Classification 21 Feb 2024 · 0 repositories · arXiv:2402.13822
-
Multi-scale Spatio-temporal Transformer-based Imbalanced Longitudinal Learning for Glaucoma Forecasting from Irregular Time Series Images 21 Feb 2024 · 0 repositories · arXiv:2402.13475
-
OMGEval: An Open Multilingual Generative Evaluation Benchmark for Large Language Models 21 Feb 2024 · 1 repository · arXiv:2402.13524
-
AlgoFormer: An Efficient Transformer Framework with Algorithmic Structures 21 Feb 2024 · 0 repositories · arXiv:2402.13572
-
PCA-Bench: Evaluating Multimodal Large Language Models in Perception-Cognition-Action Chain 21 Feb 2024 · 1 repository · arXiv:2402.15527Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
SYNFAC-EDIT: Synthetic Imitation Edit Feedback for Factual Alignment in Clinical Summarization 21 Feb 2024 · 1 repository · arXiv:2402.13919
-
Test-Driven Development for Code Generation 21 Feb 2024 · 0 repositories · arXiv:2402.13521
-
Towards Building Multilingual Language Model for Medicine 21 Feb 2024 · 1 repository · arXiv:2402.13963Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
TransGOP: Transformer-Based Gaze Object Prediction 21 Feb 2024 · 1 repository · arXiv:2402.13578Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 1 violated, 4 with no contract checked; 5 where Syntology's instrument failed) · 2 unverified (of 12 harvested samples) · 3 pointer-only (licence)
-
UniGraph: Learning a Unified Cross-Domain Foundation Model for Text-Attributed Graphs 21 Feb 2024 · 1 repository · arXiv:2402.13630Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
What's in a Name? Auditing Large Language Models for Race and Gender Bias 21 Feb 2024 · 1 repository · arXiv:2402.14875
-
WinoViz: Probing Visual Properties of Objects Under Different States 21 Feb 2024 · 0 repositories · arXiv:2402.13584
-
A Survey on Knowledge Distillation of Large Language Models 20 Feb 2024 · 1 repository · arXiv:2402.13116
-
Advancing GenAI Assisted Programming--A Comparative Study on Prompt Efficiency and Code Quality Between GPT-4 and GLM-4 20 Feb 2024 · 0 repositories · arXiv:2402.12782
-
AgentMD: Empowering Language Agents for Risk Prediction with Large-Scale Clinical Tool Learning 20 Feb 2024 · 0 repositories · arXiv:2402.13225
-
Are ELECTRA's Sentence Embeddings Beyond Repair? The Case of Semantic Textual Similarity 20 Feb 2024 · 1 repository · arXiv:2402.13130
-
ASCEND: Accurate yet Efficient End-to-End Stochastic Computing Acceleration of Vision Transformer 20 Feb 2024 · 0 repositories · arXiv:2402.12820
-
Benchmarking Retrieval-Augmented Generation for Medicine 20 Feb 2024 · 2 repositories · arXiv:2402.13178Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Can GNN be Good Adapter for LLMs? 20 Feb 2024 · 2 repositories · arXiv:2402.12984Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
Can Large Language Models be Used to Provide Psychological Counselling? An Analysis of GPT-4-Generated Responses Using Role-play Dialogues 20 Feb 2024 · 0 repositories · arXiv:2402.12738
-
Cell Graph Transformer for Nuclei Classification 20 Feb 2024 · 1 repository · arXiv:2402.12946Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
ChatEL: Entity Linking with Chatbots 20 Feb 2024 · 1 repository · arXiv:2402.14858
-
Conditional Logical Message Passing Transformer for Complex Query Answering 20 Feb 2024 · 1 repository · arXiv:2402.12954
-
DINOBot: Robot Manipulation via Retrieval and Alignment with Vision Foundation Models 20 Feb 2024 · 0 repositories · arXiv:2402.13181
-
An Equivariant Pretrained Transformer for Unified 3D Molecular Representation Learning 20 Feb 2024 · 0 repositories · arXiv:2402.12714
-
EvoGrad: A Dynamic Take on the Winograd Schema Challenge with Human Adversaries 20 Feb 2024 · 0 repositories · arXiv:2402.13372
-
Exploring the Impact of Table-to-Text Methods on Augmenting LLM-based Question Answering with Domain Hybrid Data 20 Feb 2024 · 0 repositories · arXiv:2402.12869
-
HumanEval on Latest GPT Models -- 2024 20 Feb 2024 · 1 repository · arXiv:2402.14852Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
Is the System Message Really Important to Jailbreaks in Large Language Models? 20 Feb 2024 · 0 repositories · arXiv:2402.14857
-
KetGPT -- Dataset Augmentation of Quantum Circuits using Transformers 20 Feb 2024 · 0 repositories · arXiv:2402.13352
-
Me LLaMA: Foundation Large Language Models for Medical Applications 20 Feb 2024 · 1 repository · arXiv:2402.12749Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
MoELoRA: Contrastive Learning Guided Mixture of Experts on Parameter-Efficient Fine-Tuning for Large Language Models 20 Feb 2024 · 1 repository · arXiv:2402.12851
-
NL2Formula: Generating Spreadsheet Formulas from Natural Language Queries 20 Feb 2024 · 0 repositories · arXiv:2402.14853
-
OLViT: Multi-Modal State Tracking via Attention-Based Embeddings for Video-Grounded Dialog 20 Feb 2024 · 0 repositories · arXiv:2402.13146
-
OPDAI at SemEval-2024 Task 6: Small LLMs can Accelerate Hallucination Detection with Weakly Supervised Data 20 Feb 2024 · 0 repositories · arXiv:2402.12913
-
PRECISE Framework: GPT-based Text For Improved Readability, Reliability, and Understandability of Radiology Reports For Patient-Centered Care 20 Feb 2024 · 0 repositories · arXiv:2403.00788
-
PromptKD: Distilling Student-Friendly Knowledge for Generative Language Models via Prompt Tuning 20 Feb 2024 · 1 repository · arXiv:2402.12842Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 10 harvested samples)
-
Quantum Embedding with Transformer for High-dimensional Data 20 Feb 2024 · 0 repositories · arXiv:2402.12704
-
Reflect-RL: Two-Player Online RL Fine-Tuning for LMs 20 Feb 2024 · 1 repository · arXiv:2402.12621Syntology official (archive's flag): 2 ran · 2 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (of 6 harvested samples)
-
RhythmFormer: Extracting Patterned rPPG Signals based on Periodic Sparse Attention 20 Feb 2024 · 1 repository · arXiv:2402.12788
-
R³: "This is My SQL, Are You With Me?" A Consensus-Based Multi-Agent System for Text-to-SQL Tasks 20 Feb 2024 · 0 repositories · arXiv:2402.14851
-
FinBen: A Holistic Financial Benchmark for Large Language Models 20 Feb 2024 · 2 repositories · arXiv:2402.12659Syntology community repositories only · 8 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 12 harvested samples) · 7 pointer-only (licence)
-
The Impact of Demonstrations on Multilingual In-Context Learning: A Multidimensional Analysis 20 Feb 2024 · 1 repository · arXiv:2402.12976
-
TofuEval: Evaluating Hallucinations of LLMs on Topic-Focused Dialogue Summarization 20 Feb 2024 · 1 repository · arXiv:2402.13249
-
Tree-Planted Transformers: Unidirectional Transformer Language Models with Implicit Syntactic Supervision 20 Feb 2024 · 0 repositories · arXiv:2402.12691
-
UMBCLU at SemEval-2024 Task 1A and 1C: Semantic Textual Relatedness with and without machine translation 20 Feb 2024 · 1 repository · arXiv:2402.12730
-
A Critical Evaluation of AI Feedback for Aligning Large Language Models 19 Feb 2024 · 1 repository · arXiv:2402.12366Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples) · 1 pointer-only (licence)
-
A novel molecule generative model of VAE combined with Transformer for unseen structure generation 19 Feb 2024 · 0 repositories · arXiv:2402.11950
-
A synthetic data approach for domain generalization of NLI models 19 Feb 2024 · 0 repositories · arXiv:2402.12368
-
Acquiring Clean Language Models from Backdoor Poisoned Datasets by Downscaling Frequency Space 19 Feb 2024 · 1 repository · arXiv:2402.12026Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
AnaloBench: Benchmarking the Identification of Abstract and Long-context Analogies 19 Feb 2024 · 2 repositories · arXiv:2402.12370Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Analysis of Multidomain Abstractive Summarization Using Salience Allocation 19 Feb 2024 · 0 repositories · arXiv:2402.11955
-
ArtPrompt: ASCII Art-based Jailbreak Attacks against Aligned LLMs 19 Feb 2024 · 1 repository · arXiv:2402.11753Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 15 harvested samples)
-
Ask Optimal Questions: Aligning Large Language Models with Retriever's Preference in Conversational Search 19 Feb 2024 · 0 repositories · arXiv:2402.11827
-
Asynchronous and Segmented Bidirectional Encoding for NMT 19 Feb 2024 · 0 repositories · arXiv:2402.14849
-
CodeArt: Better Code Models by Attention Regularization When Symbols Are Lacking 19 Feb 2024 · 1 repository · arXiv:2402.11842
-
Creating a Fine Grained Entity Type Taxonomy Using LLMs 19 Feb 2024 · 0 repositories · arXiv:2402.12557
-
DeepCode AI Fix: Fixing Security Vulnerabilities with Large Language Models 19 Feb 2024 · 0 repositories · arXiv:2402.13291
-
Surprising Efficacy of Fine-Tuned Transformers for Fact-Checking over Larger Language Models 19 Feb 2024 · 0 repositories · arXiv:2402.12147
-
Evaluation of ChatGPT's Smart Contract Auditing Capabilities Based on Chain of Thought 19 Feb 2024 · 0 repositories · arXiv:2402.12023
-
FeB4RAG: Evaluating Federated Search in the Context of Retrieval Augmented Generation 19 Feb 2024 · 0 repositories · arXiv:2402.11891
-
FiT: Flexible Vision Transformer for Diffusion Model 19 Feb 2024 · 2 repositories · arXiv:2402.12376Syntology official (archive's flag): 15 ran · 15 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 1 honoured, 0 violated, 10 with no contract checked; 4 where Syntology's instrument failed) · 2 unverified (of 17 harvested samples) · 3 pointer-only (licence)
-
Graph-Based Retriever Captures the Long Tail of Biomedical Knowledge 19 Feb 2024 · 0 repositories · arXiv:2402.12352
-
GTBench: Uncovering the Strategic Reasoning Limitations of LLMs via Game-Theoretic Evaluations 19 Feb 2024 · 2 repositories · arXiv:2402.12348
-
Head-wise Shareable Attention for Large Language Models 19 Feb 2024 · 2 repositories · arXiv:2402.11819
-
IMBUE: Improving Interpersonal Effectiveness through Simulation and Just-in-time Feedback with Human-Language Model Interaction 19 Feb 2024 · 0 repositories · arXiv:2402.12556
-
Is Open-Source There Yet? A Comparative Study on Commercial and Open-Source LLMs in Their Ability to Label Chest X-Ray Reports 19 Feb 2024 · 0 repositories · arXiv:2402.12298
-
KARL: Knowledge-Aware Retrieval and Representations aid Retention and Learning in Students 19 Feb 2024 · 0 repositories · arXiv:2402.12291
-
Key ingredients for effective zero-shot cross-lingual knowledge transfer in generative tasks 19 Feb 2024 · 0 repositories · arXiv:2402.12279
-
Language Model Adaptation to Specialized Domains through Selective Masking based on Genre and Topical Characteristics 19 Feb 2024 · 1 repository · arXiv:2402.12036
-
Locality-Sensitive Hashing-Based Efficient Point Transformer with Applications in High-Energy Physics 19 Feb 2024 · 1 repository · arXiv:2402.12535Syntology official (archive's flag): 12 ran · 12 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 11 where Syntology's instrument failed) · 3 unverified (of 15 harvested samples)
-
Mafin: Enhancing Black-Box Embeddings with Model Augmented Fine-Tuning 19 Feb 2024 · 0 repositories · arXiv:2402.12177
-
Enabling Weak LLMs to Judge Response Reliability via Meta Ranking 19 Feb 2024 · 0 repositories · arXiv:2402.12146
-
Cofca: A Step-Wise Counterfactual Multi-hop QA benchmark 19 Feb 2024 · 0 repositories · arXiv:2402.11924
-
Ontology Enhanced Claim Detection 19 Feb 2024 · 0 repositories · arXiv:2402.12282
-
Perceiving Longer Sequences With Bi-Directional Cross-Attention Transformers 19 Feb 2024 · 2 repositories · arXiv:2402.12138Syntology official (archive's flag): 16 ran · 17 ran (of which 11 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 6 where Syntology's instrument failed) · 11 unverified (of 28 harvested samples) · 27 pointer-only (licence)
-
Query-Based Adversarial Prompt Generation 19 Feb 2024 · 2 repositories · arXiv:2402.12329
-
Robust CLIP: Unsupervised Adversarial Fine-Tuning of Vision Embeddings for Robust Large Vision-Language Models 19 Feb 2024 · 1 repository · arXiv:2402.12336Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 6 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples)
-
Rock Classification Based on Residual Networks 19 Feb 2024 · 0 repositories · arXiv:2402.11831
-
Shallow Synthesis of Knowledge in GPT-Generated Texts: A Case Study in Automatic Related Work Composition 19 Feb 2024 · 0 repositories · arXiv:2402.12255
-
SPML: A DSL for Defending Language Models Against Prompt Attacks 19 Feb 2024 · 0 repositories · arXiv:2402.11755
-
Standardize: Aligning Language Models with Expert-Defined Standards for Content Generation 19 Feb 2024 · 1 repository · arXiv:2402.12593
-
Stealing the Invisible: Unveiling Pre-Trained CNN Models through Adversarial Examples and Timing Side-Channels 19 Feb 2024 · 0 repositories · arXiv:2402.11953
-
Stick to your Role! Stability of Personal Values Expressed in Large Language Models 19 Feb 2024 · 0 repositories · arXiv:2402.14846
-
CiMNet: Towards Joint Optimization for DNN Architecture and Configuration for Compute-In-Memory Hardware 19 Feb 2024 · 0 repositories · arXiv:2402.11780
-
What Evidence Do Language Models Find Convincing? 19 Feb 2024 · 1 repository · arXiv:2402.11782Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified (of 14 harvested samples)
-
Your Large Language Model is Secretly a Fairness Proponent and You Should Prompt it Like One 19 Feb 2024 · 0 repositories · arXiv:2402.12150
-
A Curious Case of Searching for the Correlation between Training Data and Adversarial Robustness of Transformer Textual Models 18 Feb 2024 · 1 repository · arXiv:2402.11469