Methods › General › Attention Modules › Multi-Head Attention › Papers, page 140
Multi-Head Attention
Papers archive 2025-07-28
archive papers tagged: 24,855 · with a code link: 11,214 · where Syntology ran a sample: 3,454 (2,916 with a run with no instrument failure, 538 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,454 of 24,855 tagged: 2,916 with a run with no instrument failure, 538 where every run was a failure of Syntology's instrument)
Page 140 of 249: papers 13,901 to 14,000 of 24,855, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
SikuGPT: A Generative Pre-trained Model for Intelligent Information Processing of Ancient Texts from the Perspective of Digital Humanities 16 Apr 2023 · 1 repository · arXiv:2304.07778
-
Towards Better Instruction Following Language Models for Chinese: Investigating the Impact of Training Data and Evaluation 16 Apr 2023 · 2 repositories · arXiv:2304.07854
-
TransFusionOdom: Interpretable Transformer-based LiDAR-Inertial Fusion Odometry Estimation 16 Apr 2023 · 1 repository · arXiv:2304.07728
-
A CTC Alignment-based Non-autoregressive Transformer for End-to-end Automatic Speech Recognition 15 Apr 2023 · 0 repositories · arXiv:2304.07611
-
Align-DETR: Enhancing End-to-end Object Detection with Aligned Loss 15 Apr 2023 · 1 repository · arXiv:2304.07527Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Can ChatGPT Forecast Stock Price Movements? Return Predictability and Large Language Models 15 Apr 2023 · 0 repositories · arXiv:2304.07619
-
MA-ViT: Modality-Agnostic Vision Transformers for Face Anti-Spoofing 15 Apr 2023 · 0 repositories · arXiv:2304.07549
-
A Unified HDR Imaging Method with Pixel and Patch Level 14 Apr 2023 · 0 repositories · arXiv:2304.06943
-
API-Bank: A Comprehensive Benchmark for Tool-Augmented LLMs 14 Apr 2023 · 2 repositories · arXiv:2304.08244Syntology 2 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; every one of the 2 samples that ran constructed an object rather than computing a result (of 2 harvested samples)
-
CAD-RADS scoring of coronary CT angiography with Multi-Axis Vision Transformer: a clinically-inspired deep learning pipeline 14 Apr 2023 · 1 repository · arXiv:2304.07277
-
ChatGPT: Applications, Opportunities, and Threats 14 Apr 2023 · 0 repositories · arXiv:2304.09103
-
DETR with Additional Global Aggregation for Cross-domain Weakly Supervised Object Detection 14 Apr 2023 · 0 repositories · arXiv:2304.07082
-
MedAlpaca -- An Open-Source Collection of Medical Conversational AI Models and Training Data 14 Apr 2023 · 1 repository · arXiv:2304.08247Syntology 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 13 harvested samples) · 1 pointer-only (licence)
-
On Data Sampling Strategies for Training Neural Network Speech Separation Models 14 Apr 2023 · 0 repositories · arXiv:2304.07142
-
SimpLex: a lexical text simplification architecture 14 Apr 2023 · 1 repository · arXiv:2304.07002
-
Stochastic Code Generation 14 Apr 2023 · 0 repositories · arXiv:2304.08243
-
Very high resolution canopy height maps from RGB imagery using self-supervised vision transformer and convolutional decoder trained on Aerial Lidar 14 Apr 2023 · 1 repository · arXiv:2304.07213
-
Swin3D: A Pretrained Transformer Backbone for 3D Indoor Scene Understanding 14 Apr 2023 · 2 repositories · arXiv:2304.06906Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Uncovering the Inner Workings of STEGO for Safe Unsupervised Semantic Segmentation 14 Apr 2023 · 1 repository · arXiv:2304.07314
-
AGIEval: A Human-Centric Benchmark for Evaluating Foundation Models 13 Apr 2023 · 3 repositories · arXiv:2304.06364Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Automated Mapping of CVE Vulnerability Records to MITRE CWE Weaknesses 13 Apr 2023 · 0 repositories · arXiv:2304.11130
-
ChatGPT cites the most-cited articles and journals, relying solely on Google Scholar's citation counts. As a result, AI may amplify the Matthew Effect in environmental science 13 Apr 2023 · 0 repositories · arXiv:2304.06794
-
DDT: Dual-branch Deformable Transformer for Image Denoising 13 Apr 2023 · 1 repository · arXiv:2304.06346
-
Dynamic Mobile-Former: Strengthening Dynamic Convolution with Attention and Residual Connection in Kernel Space 13 Apr 2023 · 1 repository · arXiv:2304.07254
-
DynaMITe: Dynamic Query Bootstrapping for Multi-object Interactive Segmentation Transformer 13 Apr 2023 · 0 repositories · arXiv:2304.06668
-
Evaluation of Social Biases in Recent Large Pre-Trained Models 13 Apr 2023 · 0 repositories · arXiv:2304.06861
-
EWT: Efficient Wavelet-Transformer for Single Image Denoising 13 Apr 2023 · 0 repositories · arXiv:2304.06274
-
Modeling Dense Multimodal Interactions Between Biological Pathways and Histology for Survival Prediction 13 Apr 2023 · 2 repositories · arXiv:2304.06819Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
PGTask: Introducing the Task of Profile Generation from Dialogues 13 Apr 2023 · 1 repository · arXiv:2304.06634
-
RSIR Transformer: Hierarchical Vision Transformer using Random Sampling Windows and Important Region Windows 13 Apr 2023 · 0 repositories · arXiv:2304.06250
-
Shall We Pretrain Autoregressive Language Models with Retrieval? A Comprehensive Study 13 Apr 2023 · 1 repository · arXiv:2304.06762
-
Sign Language Translation from Instructional Videos 13 Apr 2023 · 1 repository · arXiv:2304.06371
-
TransHP: Image Classification with Hierarchical Prompting 13 Apr 2023 · 1 repository · arXiv:2304.06385Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample) · 1 pointer-only (licence)
-
VISION DIFFMASK: Faithful Interpretation of Vision Transformers with Differentiable Patch Masking 13 Apr 2023 · 1 repository · arXiv:2304.06391
-
What does CLIP know about a red circle? Visual prompt engineering for VLMs 13 Apr 2023 · 0 repositories · arXiv:2304.06712
-
An Improved Heart Disease Prediction Using Stacked Ensemble Method 12 Apr 2023 · 0 repositories · arXiv:2304.06015
-
Detection of Fake Generated Scientific Abstracts 12 Apr 2023 · 1 repository · arXiv:2304.06148
-
Distilling Token-Pruned Pose Transformer for 2D Human Pose Estimation 12 Apr 2023 · 0 repositories · arXiv:2304.05548
-
DUFormer: Solving Power Line Detection Task in Aerial Images using Semantic Segmentation 12 Apr 2023 · 0 repositories · arXiv:2304.05821
-
Evaluation of ChatGPT Model for Vulnerability Detection 12 Apr 2023 · 0 repositories · arXiv:2304.07232
-
Galactic ChitChat: Using Large Language Models to Converse with Astronomy Literature 12 Apr 2023 · 0 repositories · arXiv:2304.05406
-
Localizing Model Behavior with Path Patching 12 Apr 2023 · 1 repository · arXiv:2304.05969Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 13 harvested samples)
-
MED-VT++: Unifying Multimodal Learning with a Multiscale Encoder-Decoder Video Transformer 12 Apr 2023 · 0 repositories · arXiv:2304.05930
-
Multi-scale Geometry-aware Transformer for 3D Point Cloud Classification 12 Apr 2023 · 0 repositories · arXiv:2304.05694
-
PATMAT: Person Aware Tuning of Mask-Aware Transformer for Face Inpainting 12 Apr 2023 · 2 repositories · arXiv:2304.06107
-
Real-time Trajectory-based Social Group Detection 12 Apr 2023 · 1 repository · arXiv:2304.05678
-
RECLIP: Resource-efficient CLIP by Training with Small Images 12 Apr 2023 · 0 repositories · arXiv:2304.06028
-
Towards Evaluating Explanations of Vision Transformers for Medical Imaging 12 Apr 2023 · 1 repository · arXiv:2304.06133
-
A Billion-scale Foundation Model for Remote Sensing Images 11 Apr 2023 · 0 repositories · arXiv:2304.05215
-
Approximating Online Human Evaluation of Social Chatbots with Prompting 11 Apr 2023 · 0 repositories · arXiv:2304.05253
-
Bayesian Optimization of Catalysis With In-Context Learning 11 Apr 2023 · 2 repositories · arXiv:2304.05341
-
chatClimate: Grounding Conversational AI in Climate Science 11 Apr 2023 · 0 repositories · arXiv:2304.05510
-
ChemCrow: Augmenting large-language models with chemistry tools 11 Apr 2023 · 3 repositories · arXiv:2304.05376Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 1 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Data-Efficient Image Quality Assessment with Attention-Panel Decoder 11 Apr 2023 · 1 repository · arXiv:2304.04952Syntology official (archive's flag): 2 ran · 2 ran (of which 1 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Distinguishing ChatGPT(-3.5, -4)-generated and human-written papers through Japanese stylometric analysis 11 Apr 2023 · 0 repositories · arXiv:2304.05534
-
Exploring the Use of Foundation Models for Named Entity Recognition and Lemmatization Tasks in Slavic Languages 11 Apr 2023 · 0 repositories · arXiv:2304.05336
-
MC-ViViT: Multi-branch Classifier-ViViT to detect Mild Cognitive Impairment in older adults using facial videos 11 Apr 2023 · 0 repositories · arXiv:2304.05292
-
Multi-Graph Convolution Network for Pose Forecasting 11 Apr 2023 · 0 repositories · arXiv:2304.04956
-
Multi-step Jailbreaking Privacy Attacks on ChatGPT 11 Apr 2023 · 1 repository · arXiv:2304.05197Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 5 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples)
-
Panoramic Image-to-Image Translation 11 Apr 2023 · 0 repositories · arXiv:2304.04960
-
Sim-T: Simplify the Transformer Network by Multiplexing Technique for Speech Recognition 11 Apr 2023 · 0 repositories · arXiv:2304.04991
-
Towards preserving word order importance through Forced Invalidation 11 Apr 2023 · 1 repository · arXiv:2304.05221
-
Training Large Language Models Efficiently with Sparsity and Dataflow 11 Apr 2023 · 0 repositories · arXiv:2304.05511
-
Video Event Restoration Based on Keyframes for Video Anomaly Detection 11 Apr 2023 · 0 repositories · arXiv:2304.05112
-
Weakly Supervised Intracranial Hemorrhage Segmentation using Head-Wise Gradient-Infused Self-Attention Maps from a Swin Transformer in Categorical Learning 11 Apr 2023 · 1 repository · arXiv:2304.04902
-
Automated Reading Passage Generation with OpenAI's Large Language Model 10 Apr 2023 · 0 repositories · arXiv:2304.04616
-
Detection Transformer with Stable Matching 10 Apr 2023 · 2 repositories · arXiv:2304.04742Syntology official (archive's flag): 1 ran · 7 ran (of which 2 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples) · 1 pointer-only (licence)
-
Feature Representation Learning with Adaptive Displacement Generation and Transformer Fusion for Micro-Expression Recognition 10 Apr 2023 · 0 repositories · arXiv:2304.04420
-
High Dynamic Range Imaging with Context-aware Transformer 10 Apr 2023 · 0 repositories · arXiv:2304.04416
-
HST-MRF: Heterogeneous Swin Transformer with Multi-Receptive Field for Medical Image Segmentation 10 Apr 2023 · 0 repositories · arXiv:2304.04614
-
Incorporating Structured Sentences with Time-enhanced BERT for Fully-inductive Temporal Relation Prediction 10 Apr 2023 · 0 repositories · arXiv:2304.04717
-
Is ChatGPT a Good Sentiment Analyzer? A Preliminary Study 10 Apr 2023 · 1 repository · arXiv:2304.04339Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Multilingual Machine Translation with Large Language Models: Empirical Results and Analysis 10 Apr 2023 · 2 repositories · arXiv:2304.04675
-
On the Possibilities of AI-Generated Text Detection 10 Apr 2023 · 0 repositories · arXiv:2304.04736
-
Two Steps Forward and One Behind: Rethinking Time Series Forecasting with Deep Learning 10 Apr 2023 · 0 repositories · arXiv:2304.04553
-
Use the Detection Transformer as a Data Augmenter 10 Apr 2023 · 1 repository · arXiv:2304.04554
-
Are Large Language Models Ready for Healthcare? A Comparative Study on Clinical Language Understanding 9 Apr 2023 · 1 repository · arXiv:2304.05368
-
Learning to Tokenize for Generative Retrieval 9 Apr 2023 · 1 repository · arXiv:2304.04171
-
Slide-Transformer: Hierarchical Vision Transformer with Local Self-Attention 9 Apr 2023 · 1 repository · arXiv:2304.04237Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample) · 1 pointer-only (licence)
-
Sparse Dense Fusion for 3D Object Detection 9 Apr 2023 · 0 repositories · arXiv:2304.04179
-
Transformer Utilization in Medical Image Segmentation Networks 9 Apr 2023 · 0 repositories · arXiv:2304.04225
-
Factify 2: A Multimodal Fake News and Satire News Dataset 8 Apr 2023 · 1 repository · arXiv:2304.03897
-
FlexMoE: Scaling Large-scale Sparse Pre-trained Model Training via Dynamic Device Placement 8 Apr 2023 · 0 repositories · arXiv:2304.03946
-
GPT4Rec: A Generative Framework for Personalized Recommendation and User Interests Interpretation 8 Apr 2023 · 0 repositories · arXiv:2304.03879
-
Interpretable Multi Labeled Bengali Toxic Comments Classification using Deep Learning 8 Apr 2023 · 1 repository · arXiv:2304.04087
-
Multi-class Categorization of Reasons behind Mental Disturbance in Long Texts 8 Apr 2023 · 0 repositories · arXiv:2304.04118
-
Surrogate Lagrangian Relaxation: A Path To Retrain-free Deep Neural Network Pruning 8 Apr 2023 · 0 repositories · arXiv:2304.04120
-
tmn at SemEval-2023 Task 9: Multilingual Tweet Intimacy Detection using XLM-T, Google Translate, and Ensemble Learning 8 Apr 2023 · 1 repository · arXiv:2304.04054
-
A Cross-Scale Hierarchical Transformer with Correspondence-Augmented Attention for inferring Bird's-Eye-View Semantic Segmentation 7 Apr 2023 · 0 repositories · arXiv:2304.03650
-
Cleansing Jewel: A Neural Spelling Correction Model Built On Google OCR-ed Tibetan Manuscripts 7 Apr 2023 · 0 repositories · arXiv:2304.03427
-
Evaluating the Logical Reasoning Ability of ChatGPT and GPT-4 7 Apr 2023 · 1 repository · arXiv:2304.03439
-
Hierarchical Catalogue Generation for Literature Review: A Benchmark 7 Apr 2023 · 1 repository · arXiv:2304.03512
-
PSLT: A Light-weight Vision Transformer with Ladder Self-Attention and Progressive Shift 7 Apr 2023 · 0 repositories · arXiv:2304.03481
-
SparseFormer: Sparse Visual Recognition via Limited Latent Tokens 7 Apr 2023 · 1 repository · arXiv:2304.03768
-
All Keypoints You Need: Detecting Arbitrary Keypoints on the Body of Triple, High, and Long Jump Athletes 6 Apr 2023 · 1 repository · arXiv:2304.02939
-
Can Large Language Models Play Text Games Well? Current State-of-the-Art and Open Questions 6 Apr 2023 · 0 repositories · arXiv:2304.02868
-
ChatGPT-Crawler: Find out if ChatGPT really knows what it's talking about 6 Apr 2023 · 0 repositories · arXiv:2304.03325
-
Continual Detection Transformer for Incremental Object Detection 6 Apr 2023 · 0 repositories · arXiv:2304.03110
-
Deep Learning for Opinion Mining and Topic Classification of Course Reviews 6 Apr 2023 · 0 repositories · arXiv:2304.03394
-
DeLiRa: Self-Supervised Depth, Light, and Radiance Fields 6 Apr 2023 · 0 repositories · arXiv:2304.02797