Methods › General › Attention Modules › Multi-Head Attention › Papers, page 126
Multi-Head Attention
Papers archive 2025-07-28
archive papers tagged: 24,855 · with a code link: 11,214 · where Syntology ran a sample: 3,454 (2,916 with a run with no instrument failure, 538 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,454 of 24,855 tagged: 2,916 with a run with no instrument failure, 538 where every run was a failure of Syntology's instrument)
Page 126 of 249: papers 12,501 to 12,600 of 24,855, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Is ChatGPT a Good Personality Recognizer? A Preliminary Study 8 Jul 2023 · 0 repositories · arXiv:2307.03952
-
VS-TransGRU: A Novel Transformer-GRU-based Framework Enhanced by Visual-Semantic Fusion for Egocentric Action Anticipation 8 Jul 2023 · 0 repositories · arXiv:2307.03918
-
Distilling Self-Supervised Vision Transformers for Weakly-Supervised Few-Shot Classification & Segmentation 7 Jul 2023 · 0 repositories · arXiv:2307.03407
-
DWReCO at CheckThat! 2023: Enhancing Subjectivity Detection through Style-based Data Sampling 7 Jul 2023 · 1 repository · arXiv:2307.03550
-
Exploring and Characterizing Large Language Models For Embedded System Development and Debugging 7 Jul 2023 · 0 repositories · arXiv:2307.03817
-
Goal-Conditioned Predictive Coding for Offline Reinforcement Learning 7 Jul 2023 · 0 repositories · arXiv:2307.03406
-
HoughLaneNet: Lane Detection with Deep Hough Transform and Dynamic Convolution 7 Jul 2023 · 0 repositories · arXiv:2307.03494
-
How does AI chat change search behaviors? 7 Jul 2023 · 0 repositories · arXiv:2307.03826
-
inTformer: A Time-Embedded Attention-Based Transformer for Crash Likelihood Prediction at Intersections Using Connected Vehicle Data 7 Jul 2023 · 0 repositories · arXiv:2307.03854
-
Large Language Models as Batteries-Included Zero-Shot ESCO Skills Matchers 7 Jul 2023 · 0 repositories · arXiv:2307.03539
-
Non-iterative Coarse-to-fine Transformer Networks for Joint Affine and Deformable Image Registration 7 Jul 2023 · 1 repository · arXiv:2307.03421
-
RADAR: Robust AI-Text Detection via Adversarial Learning 7 Jul 2023 · 0 repositories · arXiv:2307.03838
-
Teaching Arithmetic to Small Transformers 7 Jul 2023 · 1 repository · arXiv:2307.03381Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Text Simplification of Scientific Texts for Non-Expert Readers 7 Jul 2023 · 0 repositories · arXiv:2307.03569
-
TRAQ: Trustworthy Retrieval Augmented Question Answering via Conformal Prediction 7 Jul 2023 · 1 repository · arXiv:2307.04642
-
Unsupervised 3D out-of-distribution detection with latent diffusion models 7 Jul 2023 · 1 repository · arXiv:2307.03777
-
Weakly-supervised Contrastive Learning for Unsupervised Object Discovery 7 Jul 2023 · 1 repository · arXiv:2307.03376
-
When Do Transformers Shine in RL? Decoupling Memory from Credit Assignment 7 Jul 2023 · 2 repositories · arXiv:2307.03864Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
A Novel Site-Agnostic Multimodal Deep Learning Model to Identify Pro-Eating Disorder Content on Social Media 6 Jul 2023 · 0 repositories · arXiv:2307.06775
-
ACDNet: Attention-guided Collaborative Decision Network for Effective Medication Recommendation 6 Jul 2023 · 0 repositories · arXiv:2307.03332
-
Art Authentication with Vision Transformers 6 Jul 2023 · 0 repositories · arXiv:2307.03039
-
Can ChatGPT's Responses Boost Traditional Natural Language Processing? 6 Jul 2023 · 1 repository · arXiv:2307.04648
-
Contrast Is All You Need 6 Jul 2023 · 0 repositories · arXiv:2307.02882
-
Cross-Spatial Pixel Integration and Cross-Stage Feature Fusion Based Transformer Network for Remote Sensing Image Super-Resolution 6 Jul 2023 · 0 repositories · arXiv:2307.02974
-
Focused Transformer: Contrastive Training for Context Scaling 6 Jul 2023 · 1 repository · arXiv:2307.03170Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 4 where Syntology's instrument failed) · 5 unverified (of 14 harvested samples)
-
Improving Retrieval-Augmented Large Language Models via Data Importance Learning 6 Jul 2023 · 1 repository · arXiv:2307.03027Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 6 harvested samples)
-
Large Language Models Empowered Autonomous Edge AI for Connected Intelligence 6 Jul 2023 · 0 repositories · arXiv:2307.02779
-
LEA: Improving Sentence Similarity Robustness to Typos Using Lexical Attention Bias 6 Jul 2023 · 1 repository · arXiv:2307.02912
-
Structure Guided Multi-modal Pre-trained Transformer for Knowledge Graph Reasoning 6 Jul 2023 · 0 repositories · arXiv:2307.03591
-
Text Alignment Is An Efficient Unified Model for Massive NLP Tasks 6 Jul 2023 · 1 repository · arXiv:2307.02729
-
Track Mix Generation on Music Streaming Services using Transformers 6 Jul 2023 · 0 repositories · arXiv:2307.03045
-
UIT-Saviors at MEDVQA-GI 2023: Improving Multimodal Learning with Image Enhancement for Gastrointestinal Visual Question Answering 6 Jul 2023 · 0 repositories · arXiv:2307.02783
-
Vision Language Transformers: A Survey 6 Jul 2023 · 0 repositories · arXiv:2307.03254
-
Building Cooperative Embodied Agents Modularly with Large Language Models 5 Jul 2023 · 2 repositories · arXiv:2307.02485
-
CAME: Confidence-guided Adaptive Memory Efficient Optimization 5 Jul 2023 · 2 repositories · arXiv:2307.02047Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; the one sample that ran constructed an object rather than computing a result (of 2 harvested samples) · 1 pointer-only (licence)
-
Comparative Analysis of GPT-4 and Human Graders in Evaluating Praise Given to Students in Synthetic Dialogues 5 Jul 2023 · 0 repositories · arXiv:2307.02018
-
Elastic Decision Transformer 5 Jul 2023 · 0 repositories · arXiv:2307.02484
-
Emoji Prediction in Tweets using BERT 5 Jul 2023 · 1 repository · arXiv:2307.02054
-
Evaluating the Effectiveness of Large Language Models in Representing Textual Descriptions of Geometry and Spatial Relations 5 Jul 2023 · 0 repositories · arXiv:2307.03678
-
Exploring Continual Learning for Code Generation Models 5 Jul 2023 · 0 repositories · arXiv:2307.02435
-
External Reasoning: Towards Multi-Large-Language-Models Interchangeable Assistance with Human Feedback 5 Jul 2023 · 1 repository · arXiv:2307.12057
-
Hoodwinked: Deception and Cooperation in a Text-Based Game for Language Models 5 Jul 2023 · 1 repository · arXiv:2308.01404
-
Improving Automatic Parallel Training via Balanced Memory Workload Optimization 5 Jul 2023 · 1 repository · arXiv:2307.02031
-
Jailbroken: How Does LLM Safety Training Fail? 5 Jul 2023 · 1 repository · arXiv:2307.02483
-
Leveraging Denoised Abstract Meaning Representation for Grammatical Error Correction 5 Jul 2023 · 0 repositories · arXiv:2307.02127
-
LongNet: Scaling Transformers to 1,000,000,000 Tokens 5 Jul 2023 · 3 repositories · arXiv:2307.02486
-
MAE-DFER: Efficient Masked Autoencoder for Self-supervised Dynamic Facial Expression Recognition 5 Jul 2023 · 1 repository · arXiv:2307.02227Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 1 honoured, 0 violated, 11 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 14 harvested samples) · 7 pointer-only (licence)
-
Make A Long Image Short: Adaptive Token Length for Vision Transformers 5 Jul 2023 · 0 repositories · arXiv:2307.02092
-
Multi-Scale Prototypical Transformer for Whole Slide Image Classification 5 Jul 2023 · 0 repositories · arXiv:2307.02308
-
Multilingual Controllable Transformer-Based Lexical Simplification 5 Jul 2023 · 1 repository · arXiv:2307.02120
-
Named Entity Inclusion in Abstractive Text Summarization 5 Jul 2023 · 0 repositories · arXiv:2307.02570
-
Deductive Additivity for Planning of Natural Language Proofs 5 Jul 2023 · 1 repository · arXiv:2307.02472
-
Open-Source LLMs for Text Annotation: A Practical Guide for Model Setting and Fine-Tuning 5 Jul 2023 · 0 repositories · arXiv:2307.02179
-
Sumformer: Universal Approximation for Efficient Transformers 5 Jul 2023 · 0 repositories · arXiv:2307.02301
-
Task-Specific Alignment and Multiple Level Transformer for Few-Shot Action Recognition 5 Jul 2023 · 1 repository · arXiv:2307.01985
-
The FormAI Dataset: Generative AI in Software Security Through the Lens of Formal Verification 5 Jul 2023 · 0 repositories · arXiv:2307.02192
-
Deep Attention Q-Network for Personalized Treatment Recommendation 4 Jul 2023 · 1 repository · arXiv:2307.01519
-
Deep Features for Contactless Fingerprint Presentation Attack Detection: Can They Be Generalized? 4 Jul 2023 · 0 repositories · arXiv:2307.01845
-
DiT-3D: Exploring Plain Diffusion Transformers for 3D Shape Generation 4 Jul 2023 · 1 repository · arXiv:2307.01831Syntology 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
EdgeFace: Efficient Face Recognition Model for Edge Devices 4 Jul 2023 · 3 repositories · arXiv:2307.01838Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Embodied Task Planning with Large Language Models 4 Jul 2023 · 1 repository · arXiv:2307.01848Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Exploring Transformers for On-Line Handwritten Signature Verification 4 Jul 2023 · 0 repositories · arXiv:2307.01663
-
H-DenseFormer: An Efficient Hybrid Densely Connected Transformer for Multimodal Tumor Segmentation 4 Jul 2023 · 1 repository · arXiv:2307.01486
-
KDSTM: Neural Semi-supervised Topic Modeling with Knowledge Distillation 4 Jul 2023 · 0 repositories · arXiv:2307.01878Syntology 6 ran (of which 1 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 8 harvested samples)
-
Knowledge Graph for NLG in the context of conversational agents 4 Jul 2023 · 0 repositories · arXiv:2307.01548
-
Last layer state space model for representation learning and uncertainty quantification 4 Jul 2023 · 0 repositories · arXiv:2307.01566
-
MaskBEV: Joint Object Detection and Footprint Completion for Bird's-eye View 3D Point Clouds 4 Jul 2023 · 1 repository · arXiv:2307.01864
-
Pretraining is All You Need: A Multi-Atlas Enhanced Transformer Framework for Autism Spectrum Disorder Classification 4 Jul 2023 · 1 repository · arXiv:2307.01759
-
SageFormer: Series-Aware Framework for Long-term Multivariate Time Series Forecasting 4 Jul 2023 · 1 repository · arXiv:2307.01616Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 0 violated, 3 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples) · 2 pointer-only (licence)
-
SelfFed: Self-supervised Federated Learning for Data Heterogeneity and Label Scarcity in IoMT 4 Jul 2023 · 0 repositories · arXiv:2307.01514
-
Spike-driven Transformer 4 Jul 2023 · 1 repository · arXiv:2307.01694Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 1 pointer-only (licence)
-
Transformed Protoform Reconstruction 4 Jul 2023 · 1 repository · arXiv:2307.01896
-
ALBERTI, a Multilingual Domain Specific Language Model for Poetry Analysis 3 Jul 2023 · 0 repositories · arXiv:2307.01387
-
Beyond the Snapshot: Brain Tokenized Graph Transformer for Longitudinal Brain Functional Connectome Embedding 3 Jul 2023 · 1 repository · arXiv:2307.00858
-
End-To-End Prediction of Knee Osteoarthritis Progression With Multi-Modal Transformers 3 Jul 2023 · 1 repository · arXiv:2307.00873
-
Evaluating Shutdown Avoidance of Language Models in Textual Scenarios 3 Jul 2023 · 1 repository · arXiv:2307.00787
-
Guided Patch-Grouping Wavelet Transformer with Spatial Congruence for Ultra-High Resolution Segmentation 3 Jul 2023 · 0 repositories · arXiv:2307.00711
-
Implicit Memory Transformer for Computationally Efficient Simultaneous Speech Translation 3 Jul 2023 · 1 repository · arXiv:2307.01381
-
Improving Language Plasticity via Pretraining with Active Forgetting 3 Jul 2023 · 1 repository · arXiv:2307.01163
-
Interpretability and Transparency-Driven Detection and Transformation of Textual Adversarial Examples (IT-DT) 3 Jul 2023 · 0 repositories · arXiv:2307.01225
-
Iterative Zero-Shot LLM Prompting for Knowledge Graph Construction 3 Jul 2023 · 0 repositories · arXiv:2307.01128
-
Population Age Group Sensitivity for COVID-19 Infections with Deep Learning 3 Jul 2023 · 0 repositories · arXiv:2307.00751
-
Shiftable Context: Addressing Training-Inference Context Mismatch in Simultaneous Speech Translation 3 Jul 2023 · 1 repository · arXiv:2307.01377
-
Trainable Transformer in Transformer 3 Jul 2023 · 1 repository · arXiv:2307.01189Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
VOLTA: Improving Generative Diversity by Variational Mutual Information Maximizing Autoencoder 3 Jul 2023 · 0 repositories · arXiv:2307.00852
-
MedCPT: Contrastive Pre-trained Transformers with Large-scale PubMed Search Logs for Zero-shot Biomedical Information Retrieval 2 Jul 2023 · 2 repositories · arXiv:2307.00589
-
ClipSitu: Effectively Leveraging CLIP for Conditional Predictions in Situation Recognition 2 Jul 2023 · 1 repository · arXiv:2307.00586
-
Conformer LLMs -- Convolution Augmented Large Language Models 2 Jul 2023 · 0 repositories · arXiv:2307.00461
-
Bidirectional Correlation-Driven Inter-Frame Interaction Transformer for Referring Video Object Segmentation 2 Jul 2023 · 0 repositories · arXiv:2307.00536
-
TensorGPT: Efficient Compression of Large Language Models based on Tensor-Train Decomposition 2 Jul 2023 · 0 repositories · arXiv:2307.00526
-
AutoST: Training-free Neural Architecture Search for Spiking Transformers 1 Jul 2023 · 1 repository · arXiv:2307.00293
-
Effective Matching of Patients to Clinical Trials using Entity Extraction and Neural Re-ranking 1 Jul 2023 · 0 repositories · arXiv:2307.00381
-
Rearrangement Planning for General Part Assembly 1 Jul 2023 · 0 repositories · arXiv:2307.00206
-
How far is Language Model from 100% Few-shot Named Entity Recognition in Medical Domain 1 Jul 2023 · 1 repository · arXiv:2307.00186
-
Learning Content-enhanced Mask Transformer for Domain Generalized Urban-Scene Segmentation 1 Jul 2023 · 1 repository · arXiv:2307.00371
-
More for Less: Compact Convolutional Transformers Enable Robust Medical Image Classification with Limited Data 1 Jul 2023 · 0 repositories · arXiv:2307.00213
-
PM-DETR: Domain Adaptive Prompt Memory for Object Detection with Transformers 1 Jul 2023 · 0 repositories · arXiv:2307.00313
-
Spatial-Temporal Graph Enhanced DETR Towards Multi-Frame 3D Object Detection 1 Jul 2023 · 1 repository · arXiv:2307.00347Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples)
-
Act3D: 3D Feature Field Transformers for Multi-Task Robotic Manipulation 30 Jun 2023 · 2 repositories · arXiv:2306.17817
-
Harnessing LLMs in Curricular Design: Using GPT-4 to Support Authoring of Learning Objectives 30 Jun 2023 · 0 repositories · arXiv:2306.17459