Methods › General › Attention Modules › Multi-Head Attention › Papers, page 82
Multi-Head Attention
Papers archive 2025-07-28
archive papers tagged: 24,855 · with a code link: 11,214 · where Syntology ran a sample: 3,454 (2,916 with a run with no instrument failure, 538 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,454 of 24,855 tagged: 2,916 with a run with no instrument failure, 538 where every run was a failure of Syntology's instrument)
Page 82 of 249: papers 8,101 to 8,200 of 24,855, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
LDTR: Transformer-based Lane Detection with Anchor-chain Representation 21 Mar 2024 · 0 repositories · arXiv:2403.14354
-
Learning with SASQuaTCh: a Novel Variational Quantum Transformer Architecture with Kernel-Based Self-Attention 21 Mar 2024 · 0 repositories · arXiv:2403.14753
-
LLM-based Extraction of Contradictions from Patents 21 Mar 2024 · 0 repositories · arXiv:2403.14258
-
OTSeg: Multi-prompt Sinkhorn Attention for Zero-Shot Semantic Segmentation 21 Mar 2024 · 1 repository · arXiv:2403.14183Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 1 honoured, 0 violated, 6 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
PSALM: Pixelwise SegmentAtion with Large Multi-Modal Model 21 Mar 2024 · 1 repository · arXiv:2403.14598Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 7 harvested samples) · 2 pointer-only (licence)
-
ReAct Meets ActRe: When Language Agents Enjoy Training Data Autonomy 21 Mar 2024 · 0 repositories · arXiv:2403.14589
-
S2LIC: Learned Image Compression with the SwinV2 Block, Adaptive Channel-wise and Global-inter Attention Context 21 Mar 2024 · 1 repository · arXiv:2403.14471
-
Speech-Aware Neural Diarization with Encoder-Decoder Attractor Guided by Attention Constraints 21 Mar 2024 · 0 repositories · arXiv:2403.14268
-
SpikeGraphormer: A High-Performance Graph Transformer with Spiking Graph Attention 21 Mar 2024 · 1 repository · arXiv:2403.15480Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 1 pointer-only (licence)
-
SpikingResformer: Bridging ResNet and Vision Transformer in Spiking Neural Networks 21 Mar 2024 · 2 repositories · arXiv:2403.14302Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Token Transformation Matters: Towards Faithful Post-hoc Explanation for Vision Transformer 21 Mar 2024 · 0 repositories · arXiv:2403.14552
-
Toward Multi-class Anomaly Detection: Exploring Class-aware Unified Model against Inter-class Interference 21 Mar 2024 · 0 repositories · arXiv:2403.14213
-
Unsupervised Audio-Visual Segmentation with Modality Alignment 21 Mar 2024 · 0 repositories · arXiv:2403.14203
-
VURF: A General-purpose Reasoning and Self-refinement Framework for Video Understanding 21 Mar 2024 · 1 repository · arXiv:2403.14743
-
AMP: Autoregressive Motion Prediction Revisited with Next Token Prediction for Autonomous Driving 20 Mar 2024 · 0 repositories · arXiv:2403.13331
-
AUD-TGN: Advancing Action Unit Detection with Temporal Convolution and GPT-2 in Wild Audiovisual Contexts 20 Mar 2024 · 0 repositories · arXiv:2403.13678
-
Ax-to-Grind Urdu: Benchmark Dataset for Urdu Fake News Detection 20 Mar 2024 · 1 repository · arXiv:2403.14037
-
DiffImpute: Tabular Data Imputation With Denoising Diffusion Probabilistic Model 20 Mar 2024 · 0 repositories · arXiv:2403.13863
-
Efficient argument classification with compact language models and ChatGPT-4 refinements 20 Mar 2024 · 0 repositories · arXiv:2403.15473
-
Facilitating Pornographic Text Detection for Open-Domain Dialogue Systems via Knowledge Distillation of Large Language Models 20 Mar 2024 · 1 repository · arXiv:2403.13250
-
High-confidence pseudo-labels for domain adaptation in COVID-19 detection 20 Mar 2024 · 0 repositories · arXiv:2403.13509
-
Incentivizing News Consumption on Social Media Platforms Using Large Language Models and Realistic Bot Accounts 20 Mar 2024 · 1 repository · arXiv:2403.13362
-
Motion Generation from Fine-grained Textual Descriptions 20 Mar 2024 · 1 repository · arXiv:2403.13518
-
MTP: Advancing Remote Sensing Foundation Model via Multi-Task Pretraining 20 Mar 2024 · 2 repositories · arXiv:2403.13430Syntology official (archive's flag): 1 ran · 5 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 1 violated, 1 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
Natural Language as Policies: Reasoning for Coordinate-Level Embodied Control with LLMs 20 Mar 2024 · 0 repositories · arXiv:2403.13801
-
Open Access NAO (OAN): a ROS2-based software framework for HRI applications with the NAO robot 20 Mar 2024 · 0 repositories · arXiv:2403.13960
-
PARAMANU-AYN: Pretrain from scratch or Continual Pretraining of LLMs for Legal Domain Adaptation? 20 Mar 2024 · 0 repositories · arXiv:2403.13681
-
Portrait4D-v2: Pseudo Multi-View Data Creates Better 4D Head Synthesizer 20 Mar 2024 · 0 repositories · arXiv:2403.13570
-
Retina Vision Transformer (RetinaViT): Introducing Scaled Patches into Vision Transformers 20 Mar 2024 · 1 repository · arXiv:2403.13677
-
Rotary Position Embedding for Vision Transformer 20 Mar 2024 · 2 repositories · arXiv:2403.13298Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
T-Pixel2Mesh: Combining Global and Local Transformer for 3D Mesh Generation from a Single Image 20 Mar 2024 · 0 repositories · arXiv:2403.13663
-
VL-Mamba: Exploring State Space Models for Multimodal Learning 20 Mar 2024 · 0 repositories · arXiv:2403.13600
-
A Comparison of Deep Learning Architectures for Spacecraft Anomaly Detection 19 Mar 2024 · 0 repositories · arXiv:2403.12864
-
Automated Data Curation for Robust Language Model Fine-Tuning 19 Mar 2024 · 0 repositories · arXiv:2403.12776
-
Automatic Information Extraction From Employment Tribunal Judgements Using Large Language Models 19 Mar 2024 · 0 repositories · arXiv:2403.12936
-
Automatic Summarization of Doctor-Patient Encounter Dialogues Using Large Language Model through Prompt Tuning 19 Mar 2024 · 0 repositories · arXiv:2403.13089
-
Can AI Outperform Human Experts in Creating Social Media Creatives? 19 Mar 2024 · 0 repositories · arXiv:2404.00018
-
DeblurDiNAT: A Compact Model with Exceptional Generalization and Visual Fidelity on Unseen Domains 19 Mar 2024 · 1 repository · arXiv:2403.13163
-
Diffusion-Driven Self-Supervised Learning for Shape Reconstruction and Pose Estimation 19 Mar 2024 · 1 repository · arXiv:2403.12728
-
Emotion Recognition Using Transformers with Masked Learning 19 Mar 2024 · 1 repository · arXiv:2403.13731
-
Efficient Encoder-Decoder Transformer Decoding for Decomposable Tasks 19 Mar 2024 · 1 repository · arXiv:2403.13112
-
Fine-Tuning Pre-trained Language Models to Detect In-Game Trash Talks 19 Mar 2024 · 0 repositories · arXiv:2403.15458
-
FlowerFormer: Empowering Neural Architecture Encoding using a Flow-aware Graph Transformer 19 Mar 2024 · 1 repository · arXiv:2403.12821Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
GraphERE: Jointly Multiple Event-Event Relation Extraction via Graph-Enhanced Event Embeddings 19 Mar 2024 · 0 repositories · arXiv:2403.12523
-
Improved EATFormer: A Vision Transformer for Medical Image Classification 19 Mar 2024 · 0 repositories · arXiv:2403.13167
-
End-to-End Neuro-Symbolic Reinforcement Learning with Textual Explanations 19 Mar 2024 · 1 repository · arXiv:2403.12451Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; the one sample that ran constructed an object rather than computing a result (of 2 harvested samples) · 2 pointer-only (licence)
-
Instructing Large Language Models to Identify and Ignore Irrelevant Conditions 19 Mar 2024 · 1 repository · arXiv:2403.12744
-
LHMKE: A Large-scale Holistic Multi-subject Knowledge Evaluation Benchmark for Chinese Large Language Models 19 Mar 2024 · 0 repositories · arXiv:2403.12601
-
LLMLingua-2: Data Distillation for Efficient and Faithful Task-Agnostic Prompt Compression 19 Mar 2024 · 1 repository · arXiv:2403.12968
-
Multimodal Fusion Method with Spatiotemporal Sequences and Relationship Learning for Valence-Arousal Estimation 19 Mar 2024 · 0 repositories · arXiv:2403.12425
-
Pipelined Biomedical Event Extraction Rivaling Joint Learning 19 Mar 2024 · 0 repositories · arXiv:2403.12386
-
Pragmatic Competence Evaluation of Large Language Models for the Korean Language 19 Mar 2024 · 1 repository · arXiv:2403.12675
-
RankPrompt: Step-by-Step Comparisons Make Language Models Better Reasoners 19 Mar 2024 · 0 repositories · arXiv:2403.12373
-
SEVEN: Pruning Transformer Model by Reserving Sentinels 19 Mar 2024 · 1 repository · arXiv:2403.12688
-
Simple Hack for Transformers against Heavy Long-Text Classification on a Time- and Memory-Limited GPU Service 19 Mar 2024 · 0 repositories · arXiv:2403.12563
-
Quantifying uncertainty in lung cancer segmentation with foundation models applied to mixed-domain datasets 19 Mar 2024 · 0 repositories · arXiv:2403.13113
-
TT-BLIP: Enhancing Fake News Detection Using BLIP and Tri-Transformer 19 Mar 2024 · 0 repositories · arXiv:2403.12481
-
Benchmarking Badminton Action Recognition with a New Fine-Grained Dataset 19 Mar 2024 · 0 repositories · arXiv:2403.12385
-
VL-ICL Bench: The Devil in the Details of Multimodal In-Context Learning 19 Mar 2024 · 1 repository · arXiv:2403.13164Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 10 harvested samples) · 2 pointer-only (licence)
-
A Disease Labeler for Chinese Chest X-Ray Report Generation 18 Mar 2024 · 0 repositories · arXiv:2404.16852
-
Boosting Continuous Emotion Recognition with Self-Pretraining using Masked Autoencoders, Temporal Convolutional Networks, and Transformers 18 Mar 2024 · 0 repositories · arXiv:2403.11440
-
CICLe: Conformal In-Context Learning for Largescale Multi-Class Food Risk Classification 18 Mar 2024 · 1 repository · arXiv:2403.11904
-
Compositional learning of functions in humans and machines 18 Mar 2024 · 0 repositories · arXiv:2403.12201
-
Construction of Hyper-Relational Knowledge Graphs Using Pre-Trained Large Language Models 18 Mar 2024 · 0 repositories · arXiv:2403.11786
-
Continual Forgetting for Pre-trained Vision Models 18 Mar 2024 · 3 repositories · arXiv:2403.11530Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Counting-Stars: A Multi-evidence, Position-aware, and Scalable Benchmark for Evaluating Long-Context Large Language Models 18 Mar 2024 · 1 repository · arXiv:2403.11802
-
Crystalformer: Infinitely Connected Attention for Periodic Structure Encoding 18 Mar 2024 · 0 repositories · arXiv:2403.11686
-
Demystifying the Physics of Deep Reinforcement Learning-Based Autonomous Vehicle Decision-Making 18 Mar 2024 · 0 repositories · arXiv:2403.11432
-
EasyJailbreak: A Unified Framework for Jailbreaking Large Language Models 18 Mar 2024 · 1 repository · arXiv:2403.12171Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Embedded Named Entity Recognition using Probing Classifiers 18 Mar 2024 · 2 repositories · arXiv:2403.11747Syntology official (archive's flag): 4 ran · 4 ran (of which 2 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Embracing the Generative AI Revolution: Advancing Tertiary Education in Cybersecurity with GPT 18 Mar 2024 · 0 repositories · arXiv:2403.11402
-
Enhancing Taiwanese Hokkien Dual Translation by Exploring and Standardizing of Four Writing Systems 18 Mar 2024 · 1 repository · arXiv:2403.12024
-
Ensuring Safe and High-Quality Outputs: A Guideline Library Approach for Language Models 18 Mar 2024 · 1 repository · arXiv:2403.11838
-
EnvGen: Generating and Adapting Environments via LLMs for Training Embodied Agents 18 Mar 2024 · 0 repositories · arXiv:2403.12014
-
Evaluating Named Entity Recognition: A comparative analysis of mono- and multilingual transformer models on a novel Brazilian corporate earnings call transcripts dataset 18 Mar 2024 · 2 repositories · arXiv:2403.12212
-
GPT-4 as Evaluator: Evaluating Large Language Models on Pest Management in Agriculture 18 Mar 2024 · 0 repositories · arXiv:2403.11858
-
HateCOT: An Explanation-Enhanced Dataset for Generalizable Offensive Speech Detection via Large Language Models 18 Mar 2024 · 1 repository · arXiv:2403.11456
-
HIRI-ViT: Scaling Vision Transformer with High Resolution Inputs 18 Mar 2024 · 0 repositories · arXiv:2403.11999
-
How Far Are We on the Decision-Making of LLMs? Evaluating LLMs' Gaming Ability in Multi-Agent Environments 18 Mar 2024 · 1 repository · arXiv:2403.11807Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 1 honoured, 2 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Meta-Prompting for Automating Zero-shot Visual Recognition with LLMs 18 Mar 2024 · 1 repository · arXiv:2403.11755Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples) · 3 pointer-only (licence)
-
Metaphor Understanding Challenge Dataset for LLMs 18 Mar 2024 · 0 repositories · arXiv:2403.11810
-
Narrative Feature or Structured Feature? A Study of Large Language Models to Identify Cancer Patients at Risk of Heart Failure 18 Mar 2024 · 1 repository · arXiv:2403.11425
-
Graph Neural Network for Neutrino Physics Event Reconstruction 18 Mar 2024 · 0 repositories · arXiv:2403.11872
-
ReGenNet: Towards Human Action-Reaction Synthesis 18 Mar 2024 · 1 repository · arXiv:2403.11882Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 1 honoured, 0 violated, 4 with no contract checked; 3 where Syntology's instrument failed) · 3 unverified (of 11 harvested samples) · 6 pointer-only (licence)
-
Semantic-Enhanced Representation Learning for Road Networks with Temporal Dynamics 18 Mar 2024 · 0 repositories · arXiv:2403.11495
-
SETA: Semantic-Aware Token Augmentation for Domain Generalization 18 Mar 2024 · 1 repository · arXiv:2403.11792
-
Leveraging Large Language Models to Detect npm Malicious Packages 18 Mar 2024 · 0 repositories · arXiv:2403.12196
-
Towards Understanding the Relationship between In-context Learning and Compositional Generalization 18 Mar 2024 · 0 repositories · arXiv:2403.11834
-
Adaptive Semantic-Enhanced Denoising Diffusion Probabilistic Model for Remote Sensing Image Super-Resolution 17 Mar 2024 · 1 repository · arXiv:2403.11078
-
Aligning Uncertainty: Leveraging LLMs to Analyze Uncertainty Transfer in Text Summarization 17 Mar 2024 · 0 repositories
-
Correcting misinformation on social media with a large language model 17 Mar 2024 · 1 repository · arXiv:2403.11169
-
Data is all you need: Finetuning LLMs for Chip Design via an Automated design-data augmentation framework 17 Mar 2024 · 1 repository · arXiv:2403.11202Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
FlowMind: Automatic Workflow Generation with LLMs 17 Mar 2024 · 0 repositories · arXiv:2404.13050
-
Forging the Forger: An Attempt to Improve Authorship Verification via Data Augmentation 17 Mar 2024 · 0 repositories · arXiv:2403.11265
-
From Pixels to Predictions: Spectrogram and Vision Transformer for Better Time Series Forecasting 17 Mar 2024 · 0 repositories · arXiv:2403.11047
-
HumSum: A Personalized Lecture Summarization Tool for Humanities Students Using LLMs 17 Mar 2024 · 0 repositories
-
Is Mamba Effective for Time Series Forecasting? 17 Mar 2024 · 1 repository · arXiv:2403.11144
-
JORA: JAX Tensor-Parallel LoRA Library for Retrieval Augmented Fine-Tuning 17 Mar 2024 · 1 repository · arXiv:2403.11366
-
Mixture-of-Prompt-Experts for Multi-modal Semantic Understanding 17 Mar 2024 · 0 repositories · arXiv:2403.11311
-
Reasoning in Transformers - Mitigating Spurious Correlations and Reasoning Shortcuts 17 Mar 2024 · 0 repositories · arXiv:2403.11314