Methods › General › Attention Modules › Multi-Head Attention › Papers, page 94
Multi-Head Attention
Papers archive 2025-07-28
archive papers tagged: 24,855 · with a code link: 11,214 · where Syntology ran a sample: 3,454 (2,916 with a run with no instrument failure, 538 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,454 of 24,855 tagged: 2,916 with a run with no instrument failure, 538 where every run was a failure of Syntology's instrument)
Page 94 of 249: papers 9,301 to 9,400 of 24,855, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
ZS4C: Zero-Shot Synthesis of Compilable Code for Incomplete Code Snippets using LLMs 25 Jan 2024 · 0 repositories · arXiv:2401.14279
-
A Unified Approach to Emotion Detection and Task-Oriented Dialogue Modeling 24 Jan 2024 · 1 repository · arXiv:2401.13789
-
Automated Root Causing of Cloud Incidents using In-Context Learning with GPT-4 24 Jan 2024 · 0 repositories · arXiv:2401.13810
-
Can GPT-3.5 Generate and Code Discharge Summaries? 24 Jan 2024 · 1 repository · arXiv:2401.13512Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Fine-Grained Stateful Knowledge Exploration: A Novel Paradigm for Integrating Knowledge Graphs with Large Language Models 24 Jan 2024 · 1 repository · arXiv:2401.13444
-
ConTextual: Evaluating Context-Sensitive Text-Rich Visual Reasoning in Large Multimodal Models 24 Jan 2024 · 1 repository · arXiv:2401.13311Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Discovering Mathematical Formulas from Data via GPT-guided Monte Carlo Tree Search 24 Jan 2024 · 0 repositories · arXiv:2401.14424
-
Evaluation of General Large Language Models in Contextually Assessing Semantic Concepts Extracted from Adult Critical Care Electronic Health Record Notes 24 Jan 2024 · 0 repositories · arXiv:2401.13588
-
Graph Guided Question Answer Generation for Procedural Question-Answering 24 Jan 2024 · 0 repositories · arXiv:2401.13594
-
How Good is ChatGPT at Face Biometrics? A First Look into Recognition, Soft Biometrics, and Explainability 24 Jan 2024 · 1 repository · arXiv:2401.13641
-
Inadequacy of common stochastic neural networks for reliable clinical decision support 24 Jan 2024 · 0 repositories · arXiv:2401.13657
-
Graph Diffusion Transformers for Multi-Conditional Molecular Generation 24 Jan 2024 · 1 repository · arXiv:2401.13858Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 1 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
Language-Guided World Models: A Model-Based Approach to AI Control 24 Jan 2024 · 0 repositories · arXiv:2402.01695
-
Learning Representations for Clustering via Partial Information Discrimination and Cross-Level Interaction 24 Jan 2024 · 1 repository · arXiv:2401.13503
-
Proactive Emotion Tracker: AI-Driven Continuous Mood and Emotion Monitoring 24 Jan 2024 · 0 repositories · arXiv:2401.13722
-
Research about the Ability of LLM in the Tamper-Detection Area 24 Jan 2024 · 0 repositories · arXiv:2401.13504
-
LPNL: Scalable Link Prediction with Large Language Models 24 Jan 2024 · 0 repositories · arXiv:2401.13227
-
SegMamba: Long-range Sequential Modeling Mamba For 3D Medical Image Segmentation 24 Jan 2024 · 1 repository · arXiv:2401.13560Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Segment Any Cell: A SAM-based Auto-prompting Fine-tuning Framework for Nuclei Segmentation 24 Jan 2024 · 0 repositories · arXiv:2401.13220
-
TAT-LLM: A Specialized Language Model for Discrete Reasoning over Tabular and Textual Data 24 Jan 2024 · 0 repositories · arXiv:2401.13223
-
ARGS: Alignment as Reward-Guided Search 23 Jan 2024 · 1 repository · arXiv:2402.01694Syntology official (archive's flag): 5 ran · 5 ran (of which 1 constructed an object rather than computing a result; 4 with no instrument failure: 3 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Contrastive Learning in Distilled Models 23 Jan 2024 · 1 repository · arXiv:2401.12472
-
Convolutional Initialization for Data-Efficient Vision Transformers 23 Jan 2024 · 1 repository · arXiv:2401.12511
-
DatUS^2: Data-driven Unsupervised Semantic Segmentation with Pre-trained Self-supervised Vision Transformer 23 Jan 2024 · 1 repository · arXiv:2401.12820
-
Detecting and recognizing characters in Greek papyri with YOLOv8, DeiT and SimCLR 23 Jan 2024 · 0 repositories · arXiv:2401.12513
-
EL-VIT: Probing Vision Transformer with Interactive Visualization 23 Jan 2024 · 0 repositories · arXiv:2401.12666
-
Exploration and Improvement of Nerf-based 3D Scene Editing Techniques 23 Jan 2024 · 0 repositories · arXiv:2401.12456
-
Fast Adversarial Training against Textual Adversarial Attacks 23 Jan 2024 · 0 repositories · arXiv:2401.12461
-
KAM-CoT: Knowledge Augmented Multimodal Chain-of-Thoughts Reasoning 23 Jan 2024 · 0 repositories · arXiv:2401.12863
-
MAST: Video Polyp Segmentation with a Mixture-Attention Siamese Transformer 23 Jan 2024 · 1 repository · arXiv:2401.12439
-
Meta-Prompting: Enhancing Language Models with Task-Agnostic Scaffolding 23 Jan 2024 · 1 repository · arXiv:2401.12954Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples) · 4 pointer-only (licence)
-
On the Efficacy of Text-Based Input Modalities for Action Anticipation 23 Jan 2024 · 0 repositories · arXiv:2401.12972
-
Quality of Answers of Generative Large Language Models vs Peer Patients for Interpreting Lab Test Results for Lay Patients: Evaluation Study 23 Jan 2024 · 0 repositories · arXiv:2402.01693
-
Revolutionizing Retrieval-Augmented Generation with Enhanced PDF Structure Recognition 23 Jan 2024 · 0 repositories · arXiv:2401.12599
-
TroVE: Inducing Verifiable and Efficient Toolboxes for Solving Programmatic Tasks 23 Jan 2024 · 1 repository · arXiv:2401.12869Syntology official (archive's flag): 16 ran · 16 ran (of which 0 constructed an object rather than computing a result; 16 with no instrument failure: 0 honoured, 0 violated, 16 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 19 harvested samples) · 19 pointer-only (licence)
-
Empowering Communication: Speech Technology for Indian and Western Accents through AI-powered Speech Synthesis 22 Jan 2024 · 0 repositories · arXiv:2401.11771
-
APT: Adaptive Pruning and Tuning Pretrained Language Models for Efficient Training and Inference 22 Jan 2024 · 1 repository · arXiv:2401.12200Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 1 pointer-only (licence)
-
BETA: Binarized Energy-Efficient Transformer Accelerator at the Edge 22 Jan 2024 · 0 repositories · arXiv:2401.11851
-
Codebook-enabled Generative End-to-end Semantic Communication Powered by Transformer 22 Jan 2024 · 0 repositories · arXiv:2402.16868
-
Enhancing In-context Learning via Linear Probe Calibration 22 Jan 2024 · 1 repository · arXiv:2401.12406Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Evaluation of QCNN-LSTM for Disability Forecasting in Multiple Sclerosis Using Sequential Multisequence MRI 22 Jan 2024 · 0 repositories · arXiv:2401.12132
-
Friends Across Time: Multi-Scale Action Segmentation Transformer for Surgical Phase Recognition 22 Jan 2024 · 0 repositories · arXiv:2401.11644
-
Investigating Large Language Models for Financial Causality Detection in Multilingual Setup 22 Jan 2024 · 0 repositories
-
Keep Decoding Parallel with Effective Knowledge Distillation from Language Models to End-to-end Speech Recognisers 22 Jan 2024 · 0 repositories · arXiv:2401.11700
-
LKFormer: Large Kernel Transformer for Infrared Image Super-Resolution 22 Jan 2024 · 1 repository · arXiv:2401.11859
-
MsSVT++: Mixed-scale Sparse Voxel Transformer with Center Voting for 3D Object Detection 22 Jan 2024 · 0 repositories · arXiv:2401.11718
-
OnDev-LCT: On-Device Lightweight Convolutional Transformers towards federated learning 22 Jan 2024 · 0 repositories · arXiv:2401.11652
-
P2DT: Mitigating Forgetting in task-incremental Learning with progressive prompt Decision Transformer 22 Jan 2024 · 0 repositories · arXiv:2401.11666
-
ATFusion: An Alternate Cross-Attention Transformer Network for Infrared and Visible Image Fusion 22 Jan 2024 · 0 repositories · arXiv:2401.11675
-
Revolutionizing Finance with LLMs: An Overview of Applications and Insights 22 Jan 2024 · 0 repositories · arXiv:2401.11641
-
Speak It Out: Solving Symbol-Related Problems with Symbol-to-Language Conversion for Language Models 22 Jan 2024 · 1 repository · arXiv:2401.11725
-
SuperCLUE-Math6: Graded Multi-Step Math Reasoning Benchmark for LLMs in Chinese 22 Jan 2024 · 1 repository · arXiv:2401.11819
-
Parsimony or Capability? Decomposition Delivers Both in Long-term Time Series Forecasting 22 Jan 2024 · 0 repositories · arXiv:2401.11929
-
The Right Model for the Job: An Evaluation of Legal Multi-Label Classification Baselines 22 Jan 2024 · 0 repositories · arXiv:2401.11852
-
Zero-Space Cost Fault Tolerance for Transformer-based Language Models on ReRAM 22 Jan 2024 · 0 repositories · arXiv:2401.11664
-
Adversarial Augmentation Training Makes Action Recognition Models More Robust to Realistic Video Distribution Shifts 21 Jan 2024 · 1 repository · arXiv:2401.11406
-
AttentionLego: An Open-Source Building Block For Spatially-Scalable Large Language Model Accelerator With Processing-In-Memory Technology 21 Jan 2024 · 0 repositories · arXiv:2401.11459
-
CheX-GPT: Harnessing Large Language Models for Enhanced Chest X-ray Report Labeling 21 Jan 2024 · 2 repositories · arXiv:2401.11505
-
Confidence Preservation Property in Knowledge Distillation Abstractions 21 Jan 2024 · 0 repositories · arXiv:2401.11365
-
Enhancing Recommendation Diversity by Re-ranking with Large Language Models 21 Jan 2024 · 0 repositories · arXiv:2401.11506
-
Epilepsy Seizure Detection and Prediction using an Approximate Spiking Convolutional Transformer 21 Jan 2024 · 0 repositories · arXiv:2402.09424
-
Finding a Needle in the Adversarial Haystack: A Targeted Paraphrasing Approach For Uncovering Edge Cases with Minimal Distribution Distortion 21 Jan 2024 · 1 repository · arXiv:2401.11373
-
Freely Long-Thinking Transformer (FraiLT) 21 Jan 2024 · 0 repositories · arXiv:2401.11626
-
Language Models as Hierarchy Encoders 21 Jan 2024 · 1 repository · arXiv:2401.11374Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples)
-
LLMRA: Multi-modal Large Language Model based Restoration Assistant 21 Jan 2024 · 0 repositories · arXiv:2401.11401
-
ProLex: A Benchmark for Language Proficiency-oriented Lexical Substitution 21 Jan 2024 · 1 repository · arXiv:2401.11356
-
Scalable High-Resolution Pixel-Space Image Synthesis with Hourglass Diffusion Transformers 21 Jan 2024 · 1 repository · arXiv:2401.11605
-
SEBERTNets: Sequence Enhanced BERT Networks for Event Entity Extraction Tasks Oriented to the Finance Field 21 Jan 2024 · 1 repository · arXiv:2401.11408
-
Training microrobots to swim by a large language model 21 Jan 2024 · 0 repositories · arXiv:2402.00044
-
A Novel Benchmark for Few-Shot Semantic Segmentation in the Era of Foundation Models 20 Jan 2024 · 1 repository · arXiv:2401.11311
-
BadChain: Backdoor Chain-of-Thought Prompting for Large Language Models 20 Jan 2024 · 1 repository · arXiv:2401.12242Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 0 honoured, 0 violated, 13 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 15 harvested samples)
-
DengueNet: Dengue Prediction using Spatiotemporal Satellite Imagery for Resource-Limited Countries 20 Jan 2024 · 1 repository · arXiv:2401.11114
-
Drop your Decoder: Pre-training with Bag-of-Word Prediction for Dense Passage Retrieval 20 Jan 2024 · 3 repositories · arXiv:2401.11248Syntology official (archive's flag): 1 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Cross-Task Affinity Learning for Multitask Dense Scene Predictions 20 Jan 2024 · 1 repository · arXiv:2401.11124
-
Enhancing Large Language Models for Clinical Decision Support by Incorporating Clinical Practice Guidelines 20 Jan 2024 · 0 repositories · arXiv:2401.11120
-
Evaluating and Enhancing Large Language Models Performance in Domain-specific Medicine: Osteoarthritis Management with DocOA 20 Jan 2024 · 0 repositories · arXiv:2401.12998
-
Density Adaptive Attention is All You Need: Robust Parameter-Efficient Fine-Tuning Across Multiple Modalities 20 Jan 2024 · 2 repositories · arXiv:2401.11143Syntology community repositories only · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Inducing High Energy-Latency of Large Vision-Language Models with Verbose Images 20 Jan 2024 · 1 repository · arXiv:2401.11170Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 5 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
LMUFormer: Low Complexity Yet Powerful Spiking Model With Legendre Memory Units 20 Jan 2024 · 1 repository · arXiv:2402.04882Syntology official (archive's flag): 9 ran · 9 ran (of which 1 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 12 harvested samples)
-
LRP-QViT: Mixed-Precision Vision Transformer Quantization via Layer-wise Relevance Propagation 20 Jan 2024 · 0 repositories · arXiv:2401.11243
-
Prompt-RAG: Pioneering Vector Embedding-Free Retrieval-Augmented Generation in Niche Domains, Exemplified by Korean Medicine 20 Jan 2024 · 0 repositories · arXiv:2401.11246
-
Uncertainty-aware Bridge based Mobile-Former Network for Event-based Pattern Recognition 20 Jan 2024 · 1 repository · arXiv:2401.11123
-
Unfair TOS: An Automated Approach using Customized BERT 20 Jan 2024 · 0 repositories · arXiv:2401.11207
-
AAT: Adapting Audio Transformer for Various Acoustics Recognition Tasks 19 Jan 2024 · 1 repository · arXiv:2401.10544
-
Attentive Fusion: A Transformer-based Approach to Multimodal Hate Speech Detection 19 Jan 2024 · 2 repositories · arXiv:2401.10653
-
DeepRLI: A Multi-objective Framework for Universal Protein--Ligand Interaction Prediction 19 Jan 2024 · 1 repository · arXiv:2401.10806
-
FinLLMs: A Framework for Financial Reasoning Dataset Generation with Large Language Models 19 Jan 2024 · 0 repositories · arXiv:2401.10744
-
LangBridge: Multilingual Reasoning Without Multilingual Supervision 19 Jan 2024 · 1 repository · arXiv:2401.10695Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 7 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
M2ORT: Many-To-One Regression Transformer for Spatial Transcriptomics Prediction from Histopathology Images 19 Jan 2024 · 1 repository · arXiv:2401.10608Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
MDGNN: Multi-Relational Dynamic Graph Neural Network for Comprehensive and Dynamic Stock Investment Prediction 19 Jan 2024 · 0 repositories · arXiv:2402.06633
-
Mementos: A Comprehensive Benchmark for Multimodal Large Language Model Reasoning over Image Sequences 19 Jan 2024 · 1 repository · arXiv:2401.10529
-
Mining experimental data from Materials Science literature with Large Language Models: an evaluation study 19 Jan 2024 · 1 repository · arXiv:2401.11052
-
PhotoBot: Reference-Guided Interactive Photography via Natural Language 19 Jan 2024 · 0 repositories · arXiv:2401.11061
-
Reinforcement learning for question answering in programming domain using public community scoring as a human feedback 19 Jan 2024 · 0 repositories · arXiv:2401.10882
-
Speech Swin-Transformer: Exploring a Hierarchical Transformer with Shifted Windows for Speech Emotion Recognition 19 Jan 2024 · 0 repositories · arXiv:2401.10536
-
Understanding Video Transformers via Universal Concept Discovery 19 Jan 2024 · 0 repositories · arXiv:2401.10831
-
An Empirical Study on the Impact of Positional Encoding in Transformer-based Monaural Speech Enhancement 18 Jan 2024 · 0 repositories · arXiv:2401.09686
-
Beyond Traditional Benchmarks: Analyzing Behaviors of Open LLMs on Data-to-Text Generation 18 Jan 2024 · 0 repositories · arXiv:2401.10186
-
BlenDA: Domain Adaptive Object Detection through diffusion-based blending 18 Jan 2024 · 1 repository · arXiv:2401.09921
-
ChatQA: Surpassing GPT-4 on Conversational QA and RAG 18 Jan 2024 · 0 repositories · arXiv:2401.10225