Methods › General › Attention Modules › Multi-Head Attention › Papers, page 98
Multi-Head Attention
Papers archive 2025-07-28
archive papers tagged: 24,855 · with a code link: 11,214 · where Syntology ran a sample: 3,454 (2,916 with a run with no instrument failure, 538 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,454 of 24,855 tagged: 2,916 with a run with no instrument failure, 538 where every run was a failure of Syntology's instrument)
Page 98 of 249: papers 9,701 to 9,800 of 24,855, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
UFC-Net: Unrolling Fixed-point Continuous Network for Deep Compressive Sensing 1 Jan 2024 · 0 repositories
-
Uncertainty-aware Action Decoupling Transformer for Action Anticipation 1 Jan 2024 · 0 repositories
-
Unlocking the Potential of Pre-trained Vision Transformers for Few-Shot Semantic Segmentation through Relationship Descriptors 1 Jan 2024 · 1 repository
-
Video Harmonization with Triplet Spatio-Temporal Variation Patterns 1 Jan 2024 · 1 repository
-
A Two-stream Hybrid CNN-Transformer Network for Skeleton-based Human Interaction Recognition 31 Dec 2023 · 0 repositories · arXiv:2401.00409
-
An Analysis of Embedding Layers and Similarity Scores using Siamese Neural Networks 31 Dec 2023 · 0 repositories · arXiv:2401.00582
-
Analyzing Local Representations of Self-supervised Vision Transformers 31 Dec 2023 · 0 repositories · arXiv:2401.00463
-
EMAGE: Towards Unified Holistic Co-Speech Gesture Generation via Expressive Masked Audio Gesture Modeling 31 Dec 2023 · 1 repository · arXiv:2401.00374Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 2 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Generative Model-Driven Synthetic Training Image Generation: An Approach to Cognition in Rail Defect Detection 31 Dec 2023 · 1 repository · arXiv:2401.00393
-
RAGTruth: A Hallucination Corpus for Developing Trustworthy Retrieval-Augmented Language Models 31 Dec 2023 · 3 repositories · arXiv:2401.00396Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples)
-
SVFAP: Self-supervised Video Facial Affect Perceiver 31 Dec 2023 · 1 repository · arXiv:2401.00416
-
Advancing TTP Analysis: Harnessing the Power of Large Language Models with Retrieval Augmented Generation 30 Dec 2023 · 1 repository · arXiv:2401.00280
-
HybridGait: A Benchmark for Spatial-Temporal Cloth-Changing Gait Recognition with Hybrid Explorations 30 Dec 2023 · 1 repository · arXiv:2401.00271
-
Image Super-resolution Reconstruction Network based on Enhanced Swin Transformer via Alternating Aggregation of Local-Global Features 30 Dec 2023 · 0 repositories · arXiv:2401.00241
-
L3Cube-MahaSocialNER: A Social Media based Marathi NER Dataset and BERT models 30 Dec 2023 · 1 repository · arXiv:2401.00170
-
Red Teaming for Large Language Models At Scale: Tackling Hallucinations on Mathematics Tasks 30 Dec 2023 · 1 repository · arXiv:2401.00290
-
Trace and Edit Relation Associations in GPT 30 Dec 2023 · 0 repositories · arXiv:2401.02976
-
Why is the User Interface a Dark Pattern? : Explainable Auto-Detection and its Analysis 30 Dec 2023 · 1 repository · arXiv:2401.04119
-
A Fully Automated Pipeline Using Swin Transformers for Deep Learning-Based Blood Segmentation on Head CT Scans After Aneurysmal Subarachnoid Hemorrhage 29 Dec 2023 · 0 repositories · arXiv:2312.17553
-
Action-Item-Driven Summarization of Long Meeting Transcripts 29 Dec 2023 · 1 repository · arXiv:2312.17581
-
Adaptive Control Strategy for Quadruped Robots in Actuator Degradation Scenarios 29 Dec 2023 · 1 repository · arXiv:2312.17606
-
Efficacy of Utilizing Large Language Models to Detect Public Threat Posted Online 29 Dec 2023 · 0 repositories · arXiv:2401.02974
-
Enhancing Quantitative Reasoning Skills of Large Language Models through Dimension Perception 29 Dec 2023 · 0 repositories · arXiv:2312.17532
-
Gemini in Reasoning: Unveiling Commonsense in Multimodal Large Language Models 29 Dec 2023 · 1 repository · arXiv:2312.17661
-
Jatmo: Prompt Injection Defense by Task-Specific Finetuning 29 Dec 2023 · 1 repository · arXiv:2312.17673Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 8 unverified (of 18 harvested samples) · 18 pointer-only (licence)
-
MosaicBERT: A Bidirectional Encoder Optimized for Fast Pretraining 29 Dec 2023 · 1 repository · arXiv:2312.17482
-
Multiscale Vision Transformers meet Bipartite Matching for efficient single-stage Action Localization 29 Dec 2023 · 1 repository · arXiv:2312.17686
-
TuPy-E: detecting hate speech in Brazilian Portuguese social media with a novel dataset and comprehensive analysis of models 29 Dec 2023 · 1 repository · arXiv:2312.17704
-
XAI for In-hospital Mortality Prediction via Multimodal ICU Data 29 Dec 2023 · 1 repository · arXiv:2312.17624
-
MR-GSM8K: A Meta-Reasoning Benchmark for Large Language Model Evaluation 28 Dec 2023 · 2 repositories · arXiv:2312.17080
-
Evaluating the Performance of Large Language Models for Spanish Language in Undergraduate Admissions Exams 28 Dec 2023 · 0 repositories · arXiv:2312.16845
-
Geometry-Biased Transformer for Robust Multi-View 3D Human Pose Reconstruction 28 Dec 2023 · 0 repositories · arXiv:2312.17106
-
Enhancing Open-Domain Task-Solving Capability of LLMs via Autonomous Tool Integration from GitHub 28 Dec 2023 · 1 repository · arXiv:2312.17294
-
Language Model as an Annotator: Unsupervised Context-aware Quality Phrase Generation 28 Dec 2023 · 0 repositories · arXiv:2312.17349
-
LLM4Causal: Democratized Causal Tools for Everyone via Large Language Model 28 Dec 2023 · 0 repositories · arXiv:2312.17122
-
Learning Multi-axis Representation in Frequency Domain for Medical Image Segmentation 28 Dec 2023 · 1 repository · arXiv:2312.17030
-
Learning Vision from Models Rivals Learning Vision from Data 28 Dec 2023 · 2 repositories · arXiv:2312.17742Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 1 pointer-only (licence)
-
Length Extrapolation of Transformers: A Survey from the Perspective of Positional Encoding 28 Dec 2023 · 0 repositories · arXiv:2312.17044
-
Replication-proof Bandit Mechanism Design with Bayesian Agents 28 Dec 2023 · 0 repositories · arXiv:2312.16896
-
ROI-Aware Multiscale Cross-Attention Vision Transformer for Pest Image Identification 28 Dec 2023 · 0 repositories · arXiv:2312.16914
-
SentinelLMs: Encrypted Input Adaptation and Fine-tuning of Language Models for Private and Secure Inference 28 Dec 2023 · 1 repository · arXiv:2312.17342
-
The LLM Surgeon 28 Dec 2023 · 1 repository · arXiv:2312.17244Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
A New Perspective on Speaker Verification: Joint Modeling with DFSMN and Transformer 28 Dec 2023 · 1 repository · arXiv:2312.16826
-
Weed mapping in multispectral drone imagery using lightweight vision transformers 28 Dec 2023 · 1 repository
-
A Non-Uniform Low-Light Image Enhancement Method with Multi-Scale Attention Transformer and Luminance Consistency Loss 27 Dec 2023 · 1 repository · arXiv:2312.16498
-
Group Multi-View Transformer for 3D Shape Analysis with Spatial Encoding 27 Dec 2023 · 1 repository · arXiv:2312.16477
-
Learn From Orientation Prior for Radiograph Super-Resolution: Orientation Operator Transformer 27 Dec 2023 · 0 repositories · arXiv:2312.16455
-
PanGu-π: Enhancing Language Model Architectures via Nonlinearity Compensation 27 Dec 2023 · 0 repositories · arXiv:2312.17276
-
RefineNet: Enhancing Text-to-Image Conversion with High-Resolution and Detail Accuracy through Hierarchical Transformers and Progressive Refinement 27 Dec 2023 · 0 repositories · arXiv:2312.17274
-
Relationship between auditory and semantic entrainment using Deep Neural Networks (DNN) 27 Dec 2023 · 0 repositories · arXiv:2312.16599
-
scRNA-seq Data Clustering by Cluster-aware Iterative Contrastive Learning 27 Dec 2023 · 1 repository · arXiv:2312.16600
-
Spatial-Related Sensors Matters: 3D Human Motion Reconstruction Assisted with Textual Semantics 27 Dec 2023 · 0 repositories · arXiv:2401.05412
-
Attention-aware Social Graph Transformer Networks for Stochastic Trajectory Prediction 26 Dec 2023 · 0 repositories · arXiv:2312.15881
-
C2T-Net: Channel-Aware Cross-Fused Transformer-Style Networks for Pedestrian Attribute Recognition 26 Dec 2023 · 1 repository
-
ChartBench: A Benchmark for Complex Visual Reasoning in Charts 26 Dec 2023 · 0 repositories · arXiv:2312.15915
-
Graph Context Transformation Learning for Progressive Correspondence Pruning 26 Dec 2023 · 1 repository · arXiv:2312.15971
-
Heterogeneous Encoders Scaling In The Transformer For Neural Machine Translation 26 Dec 2023 · 0 repositories · arXiv:2312.15872
-
LangSplat: 3D Language Gaussian Splatting 26 Dec 2023 · 1 repository · arXiv:2312.16084Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Modality-Collaborative Transformer with Hybrid Feature Reconstruction for Robust Emotion Recognition 26 Dec 2023 · 1 repository · arXiv:2312.15848
-
PDiT: Interleaving Perception and Decision-making Transformers for Deep Reinforcement Learning 26 Dec 2023 · 2 repositories · arXiv:2312.15863
-
Principled Instructions Are All You Need for Questioning LLaMA-1/2, GPT-3.5/4 26 Dec 2023 · 2 repositories · arXiv:2312.16171Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
RoleEval: A Bilingual Role Evaluation Benchmark for Large Language Models 26 Dec 2023 · 1 repository · arXiv:2312.16132
-
Scaling Down, LiTting Up: Efficient Zero-Shot Listwise Reranking with Seq2seq Encoder-Decoder Models 26 Dec 2023 · 2 repositories · arXiv:2312.16098
-
SecQA: A Concise Question-Answering Dataset for Evaluating Large Language Models in Computer Security 26 Dec 2023 · 1 repository · arXiv:2312.15838Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 1 pointer-only (licence)
-
Task Contamination: Language Models May Not Be Few-Shot Anymore 26 Dec 2023 · 0 repositories · arXiv:2312.16337
-
Compositional Generalization in Spoken Language Understanding 25 Dec 2023 · 0 repositories · arXiv:2312.15815
-
Deep Structure and Attention Aware Subspace Clustering 25 Dec 2023 · 1 repository · arXiv:2312.15577
-
ESGReveal: An LLM-based approach for extracting structured data from ESG reports 25 Dec 2023 · 0 repositories · arXiv:2312.17264
-
IQAGPT: Image Quality Assessment with Vision-language and ChatGPT Models 25 Dec 2023 · 0 repositories · arXiv:2312.15663
-
Lifting by Image -- Leveraging Image Cues for Accurate 3D Human Pose Estimation 25 Dec 2023 · 0 repositories · arXiv:2312.15636
-
Nighttime Person Re-Identification via Collaborative Enhancement Network with Multi-domain Learning 25 Dec 2023 · 1 repository · arXiv:2312.16246
-
Partial Fine-Tuning: A Successor to Full Fine-Tuning for Vision Transformers 25 Dec 2023 · 0 repositories · arXiv:2312.15681
-
Proximal Gradient Descent Unfolding Dense-spatial Spectral-attention Transformer for Compressive Spectral Imaging 25 Dec 2023 · 0 repositories · arXiv:2312.16237
-
UniRef++: Segment Every Reference Object in Spatial and Temporal Spaces 25 Dec 2023 · 2 repositories · arXiv:2312.15715Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 1 honoured, 0 violated, 4 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples) · 2 pointer-only (licence)
-
Word length-aware text spotting: Enhancing detection and recognition in dense text image 25 Dec 2023 · 0 repositories · arXiv:2312.15690
-
DEAP: Design Space Exploration for DNN Accelerator Parallelism 24 Dec 2023 · 0 repositories · arXiv:2312.15388
-
Deformable Audio Transformer for Audio Event Detection 24 Dec 2023 · 0 repositories · arXiv:2312.16228
-
Diffusion-EXR: Controllable Review Generation for Explainable Recommendation via Diffusion Models 24 Dec 2023 · 0 repositories · arXiv:2312.15490
-
Fairness-Aware Structured Pruning in Transformers 24 Dec 2023 · 1 repository · arXiv:2312.15398Syntology official: harvested, nothing ran · 0 ran · 2 unverified (of 2 harvested samples)
-
Multi-level biomedical NER through multi-granularity embeddings and enhanced labeling 24 Dec 2023 · 0 repositories · arXiv:2312.15550
-
PointCT: Point Central Transformer Network for Weakly-supervised Point Cloud Semantic Segmentation 24 Dec 2023 · 1 repository
-
Do LLM Agents Exhibit Social Behavior? 23 Dec 2023 · 0 repositories · arXiv:2312.15198
-
Enhancing User Intent Capture in Session-Based Recommendation with Attribute Patterns 23 Dec 2023 · 1 repository · arXiv:2312.16199Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
GestaltMML: Enhancing Rare Genetic Disease Diagnosis through Multimodal Machine Learning Combining Facial Images and Clinical Texts 23 Dec 2023 · 2 repositories · arXiv:2312.15320
-
Narrowing the semantic gaps in U-Net with learnable skip connections: The case of medical image segmentation 23 Dec 2023 · 3 repositories · arXiv:2312.15182
-
Paralinguistics-Enhanced Large Language Modeling of Spoken Dialogue 23 Dec 2023 · 0 repositories · arXiv:2312.15316
-
Understanding the Potential of FPGA-Based Spatial Acceleration for Large Language Model Inference 23 Dec 2023 · 1 repository · arXiv:2312.15159
-
Context Enhanced Transformer for Single Image Object Detection 22 Dec 2023 · 0 repositories · arXiv:2312.14492
-
Efficacy of Machine-Generated Instructions 22 Dec 2023 · 0 repositories · arXiv:2312.14423
-
Personalized Large Language Model Assistant with Evolving Conditional Memory 22 Dec 2023 · 0 repositories · arXiv:2312.17257
-
FM-OV3D: Foundation Model-based Cross-modal Knowledge Blending for Open-Vocabulary 3D Detection 22 Dec 2023 · 0 repositories · arXiv:2312.14465
-
Generative Pretraining at Scale: Transformer-Based Encoding of Transactional Behavior for Fraud Detection 22 Dec 2023 · 0 repositories · arXiv:2312.14406
-
Global Occlusion-Aware Transformer for Robust Stereo Matching 22 Dec 2023 · 1 repository · arXiv:2312.14650
-
Large Language Model (LLM) Bias Index -- LLMBI 22 Dec 2023 · 0 repositories · arXiv:2312.14769
-
MMGPL: Multimodal Medical Data Analysis with Graph Prompt Learning 22 Dec 2023 · 0 repositories · arXiv:2312.14574
-
Numerical Reasoning for Financial Reports 22 Dec 2023 · 1 repository · arXiv:2312.14870
-
QUAR-VLA: Vision-Language-Action Model for Quadruped Robots 22 Dec 2023 · 0 repositories · arXiv:2312.14457
-
Refining GPT-3 Embeddings with a Siamese Structure for Technical Post Duplicate Detection 22 Dec 2023 · 1 repository · arXiv:2312.15068
-
SCUNet++: Swin-UNet and CNN Bottleneck Hybrid Architecture with Multi-Fusion Dense Skip Connection for Pulmonary Embolism CT Image Segmentation 22 Dec 2023 · 1 repository · arXiv:2312.14705
-
Spatiotemporal-Linear: Towards Universal Multivariate Time Series Forecasting 22 Dec 2023 · 0 repositories · arXiv:2312.14869