Methods › General › Attention Modules › Multi-Head Attention › Papers, page 48
Multi-Head Attention
Papers archive 2025-07-28
archive papers tagged: 24,855 · with a code link: 11,214 · where Syntology ran a sample: 3,454 (2,916 with a run with no instrument failure, 538 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,454 of 24,855 tagged: 2,916 with a run with no instrument failure, 538 where every run was a failure of Syntology's instrument)
Page 48 of 249: papers 4,701 to 4,800 of 24,855, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Active Learning to Guide Labeling Efforts for Question Difficulty Estimation 14 Sep 2024 · 1 repository · arXiv:2409.09258
-
An empirical evaluation of using ChatGPT to summarize disputes for recommending similar labor and employment cases in Chinese 14 Sep 2024 · 0 repositories · arXiv:2409.09280
-
Autoregressive + Chain of Thought = Recurrent: Recurrence's Role in Language Models' Computability and a Revisit of Recurrent Transformer 14 Sep 2024 · 0 repositories · arXiv:2409.09239
-
Block-Attention for Efficient RAG 14 Sep 2024 · 1 repository · arXiv:2409.15355Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 6 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
Comparing Retrieval-Augmentation and Parameter-Efficient Fine-Tuning for Privacy-Preserving Personalization of Large Language Models 14 Sep 2024 · 1 repository · arXiv:2409.09510
-
Investigation of Hierarchical Spectral Vision Transformer Architecture for Classification of Hyperspectral Imagery 14 Sep 2024 · 0 repositories · arXiv:2409.09244
-
Keeping Humans in the Loop: Human-Centered Automated Annotation with Generative AI 14 Sep 2024 · 0 repositories · arXiv:2409.09467
-
LLM-Powered Ensemble Learning for Paper Source Tracing: A GPU-Free Approach 14 Sep 2024 · 1 repository · arXiv:2409.09383
-
Multi-Microphone and Multi-Modal Emotion Recognition in Reverberant Environment 14 Sep 2024 · 0 repositories · arXiv:2409.09545
-
Planning Transformer: Long-Horizon Offline Reinforcement Learning with Planning Tokens 14 Sep 2024 · 0 repositories · arXiv:2409.09513
-
SEA-ViT: Sea Surface Currents Forecasting Using Vision Transformer and GRU-Based Spatio-Temporal Covariance Modeling 14 Sep 2024 · 1 repository · arXiv:2409.16313
-
SEE: Semantically Aligned EEG-to-Text Translation 14 Sep 2024 · 0 repositories · arXiv:2409.16312
-
Tran-GCN: A Transformer-Enhanced Graph Convolutional Network for Person Re-Identification in Monitoring Videos 14 Sep 2024 · 0 repositories · arXiv:2409.09391
-
VSFormer: Mining Correlations in Flexible View Set for Multi-view 3D Shape Understanding 14 Sep 2024 · 1 repository · arXiv:2409.09254
-
A RAG Approach for Generating Competency Questions in Ontology Engineering 13 Sep 2024 · 0 repositories · arXiv:2409.08820
-
ChangeChat: An Interactive Model for Remote Sensing Change Analysis via Multimodal Instruction Tuning 13 Sep 2024 · 1 repository · arXiv:2409.08582Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
DomURLs_BERT: Pre-trained BERT-based Model for Malicious Domains and URLs Detection and Classification 13 Sep 2024 · 1 repository · arXiv:2409.09143
-
Exploring Information Retrieval Landscapes: An Investigation of a Novel Evaluation Techniques and Comparative Document Splitting Methods 13 Sep 2024 · 1 repository · arXiv:2409.08479
-
HTR-VT: Handwritten Text Recognition with Vision Transformer 13 Sep 2024 · 2 repositories · arXiv:2409.08573
-
Integration of Mamba and Transformer -- MAT for Long-Short Range Time Series Forecasting with Application to Weather Dynamics 13 Sep 2024 · 0 repositories · arXiv:2409.08530
-
KodeXv0.1: A Family of State-of-the-Art Financial Large Language Models 13 Sep 2024 · 0 repositories · arXiv:2409.13749
-
LMAC-TD: Producing Time Domain Explanations for Audio Classifiers 13 Sep 2024 · 0 repositories · arXiv:2409.08655
-
Optimizing Ingredient Substitution Using Large Language Models to Enhance Phytochemical Content in Recipes 13 Sep 2024 · 0 repositories · arXiv:2409.08792
-
Pathfinder for Low-altitude Aircraft with Binary Neural Network 13 Sep 2024 · 1 repository · arXiv:2409.08824
-
Phikon-v2, A large and public feature extractor for biomarker prediction 13 Sep 2024 · 0 repositories · arXiv:2409.09173
-
PSTNet: Enhanced Polyp Segmentation with Multi-scale Alignment and Frequency Domain Integration 13 Sep 2024 · 0 repositories · arXiv:2409.08501
-
SkinFormer: Learning Statistical Texture Representation with Transformer for Skin Lesion Segmentation 13 Sep 2024 · 1 repository · arXiv:2409.08652
-
TabKANet: Tabular Data Modeling with Kolmogorov-Arnold Network and Transformer 13 Sep 2024 · 2 repositories · arXiv:2409.08806
-
Transformer with Controlled Attention for Synchronous Motion Captioning 13 Sep 2024 · 1 repository · arXiv:2409.09177
-
Winning Solution For Meta KDD Cup' 24 13 Sep 2024 · 0 repositories · arXiv:2410.00005
-
xTED: Cross-Domain Adaptation via Diffusion-Based Trajectory Editing 13 Sep 2024 · 1 repository · arXiv:2409.08687
-
AD-Lite Net: A Lightweight and Concatenated CNN Model for Alzheimer's Detection from MRI Images 12 Sep 2024 · 0 repositories · arXiv:2409.08170
-
AudioBERT: Audio Knowledge Augmented Language Model 12 Sep 2024 · 1 repository · arXiv:2409.08199
-
Collaborative Automatic Modulation Classification via Deep Edge Inference for Hierarchical Cognitive Radio Networks 12 Sep 2024 · 0 repositories · arXiv:2409.07946
-
Depth Matters: Exploring Deep Interactions of RGB-D for Semantic Segmentation in Traffic Scenes 12 Sep 2024 · 0 repositories · arXiv:2409.07995
-
Enhanced Online Grooming Detection Employing Context Determination and Message-Level Analysis 12 Sep 2024 · 0 repositories · arXiv:2409.07958
-
Experimenting with Legal AI Solutions: The Case of Question-Answering for Access to Justice 12 Sep 2024 · 0 repositories · arXiv:2409.07713
-
Fine-tuning Large Language Models for Entity Matching 12 Sep 2024 · 1 repository · arXiv:2409.08185
-
Generated Data with Fake Privacy: Hidden Dangers of Fine-tuning Large Language Models on Generated Data 12 Sep 2024 · 0 repositories · arXiv:2409.11423
-
GRE^2-MDCL: Graph Representation Embedding Enhanced via Multidimensional Contrastive Learning 12 Sep 2024 · 0 repositories · arXiv:2409.07725
-
HiRT: Enhancing Robotic Control with Hierarchical Robot Transformers 12 Sep 2024 · 0 repositories · arXiv:2410.05273
-
Lagrange Duality and Compound Multi-Attention Transformer for Semi-Supervised Medical Image Segmentation 12 Sep 2024 · 1 repository · arXiv:2409.07793
-
Model Ensemble for Brain Tumor Segmentation in Magnetic Resonance Imaging 12 Sep 2024 · 1 repository · arXiv:2409.08232
-
OmniQuery: Contextually Augmenting Captured Multimodal Memory to Enable Personal Question Answering 12 Sep 2024 · 0 repositories · arXiv:2409.08250
-
On the Vulnerability of Applying Retrieval-Augmented Generation within Knowledge-Intensive Application Domains 12 Sep 2024 · 0 repositories · arXiv:2409.17275
-
Online vs Offline: A Comparative Study of First-Party and Third-Party Evaluations of Social Chatbots 12 Sep 2024 · 0 repositories · arXiv:2409.07823
-
Q-value Regularized Decision ConvFormer for Offline Reinforcement Learning 12 Sep 2024 · 0 repositories · arXiv:2409.08062
-
SDformer: Efficient End-to-End Transformer for Depth Completion 12 Sep 2024 · 1 repository · arXiv:2409.08159
-
SoloAudio: Target Sound Extraction with Language-oriented Audio Diffusion Transformer 12 Sep 2024 · 1 repository · arXiv:2409.08425
-
Stable Language Model Pre-training by Reducing Embedding Variability 12 Sep 2024 · 0 repositories · arXiv:2409.07787
-
Unleashing Worms and Extracting Data: Escalating the Outcome of Attacks against RAG-based Inference in Scale and Severity Using Jailbreaking 12 Sep 2024 · 1 repository · arXiv:2409.08045
-
How Effectively Do LLMs Extract Feature-Sentiment Pairs from App Reviews? 11 Sep 2024 · 1 repository · arXiv:2409.07162
-
A Novel Mathematical Framework for Objective Characterization of Ideas 11 Sep 2024 · 0 repositories · arXiv:2409.07578
-
ART: Artifact Removal Transformer for Reconstructing Noise-Free Multichannel Electroencephalographic Signals 11 Sep 2024 · 0 repositories · arXiv:2409.07326
-
Attention Down-Sampling Transformer, Relative Ranking and Self-Consistency for Blind Image Quality Assessment 11 Sep 2024 · 1 repository · arXiv:2409.07115
-
Bio-Eng-LMM AI Assist chatbot: A Comprehensive Tool for Research and Education 11 Sep 2024 · 1 repository · arXiv:2409.07110
-
Can We Count on LLMs? The Fixed-Effect Fallacy and Claims of GPT-4 Capabilities 11 Sep 2024 · 0 repositories · arXiv:2409.07638
-
Cross-Dialect Text-To-Speech in Pitch-Accent Language Incorporating Multi-Dialect Phoneme-Level BERT 11 Sep 2024 · 0 repositories · arXiv:2409.07265
-
CWT-Net: Super-resolution of Histopathology Images Using a Cross-scale Wavelet-based Transformer 11 Sep 2024 · 0 repositories · arXiv:2409.07092
-
Foundation Models Boost Low-Level Perceptual Similarity Metrics 11 Sep 2024 · 1 repository · arXiv:2409.07650
-
Integrating SPARQL and LLMs for Question Answering over Scholarly Data Sources 11 Sep 2024 · 0 repositories · arXiv:2409.18969
-
Intrapartum Ultrasound Image Segmentation of Pubic Symphysis and Fetal Head Using Dual Student-Teacher Framework with CNN-ViT Collaborative Learning 11 Sep 2024 · 1 repository · arXiv:2409.06928
-
Mamba for Scalable and Efficient Personalized Recommendations 11 Sep 2024 · 0 repositories · arXiv:2409.17165
-
Multimodal Emotion Recognition with Vision-language Prompting and Modality Dropout 11 Sep 2024 · 0 repositories · arXiv:2409.07078
-
SimulBench: Evaluating Language Models with Creative Simulation Tasks 11 Sep 2024 · 0 repositories · arXiv:2409.07641
-
SSR-Speech: Towards Stable, Safe and Robust Zero-shot Text-based Speech Editing and Synthesis 11 Sep 2024 · 1 repository · arXiv:2409.07556Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 11 harvested samples)
-
Swin-LiteMedSAM: A Lightweight Box-Based Segment Anything Model for Large-Scale Medical Image Datasets 11 Sep 2024 · 1 repository · arXiv:2409.07172Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Token Turing Machines are Efficient Vision Models 11 Sep 2024 · 1 repository · arXiv:2409.07613Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 1 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 14 harvested samples) · 1 pointer-only (licence)
-
Towards Fairer Health Recommendations: finding informative unbiased samples via Word Sense Disambiguation 11 Sep 2024 · 0 repositories · arXiv:2409.07424
-
VMAS: Video-to-Music Generation via Semantic Alignment in Web Music Videos 11 Sep 2024 · 0 repositories · arXiv:2409.07450
-
Weather-Informed Probabilistic Forecasting and Scenario Generation in Power Systems 11 Sep 2024 · 0 repositories · arXiv:2409.07637
-
Mapping Biomedical Ontology Terms to IDs: Effect of Domain Prevalence on Prediction Accuracy 11 Sep 2024 · 0 repositories · arXiv:2409.13746
-
KAG: Boosting LLMs in Professional Domains via Knowledge Augmented Generation 10 Sep 2024 · 1 repository · arXiv:2409.13731Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
A Dataset for Evaluating LLM-based Evaluation Functions for Research Question Extraction Task 10 Sep 2024 · 0 repositories · arXiv:2409.06883
-
A Practical Gated Recurrent Transformer Network Incorporating Multiple Fusions for Video Denoising 10 Sep 2024 · 0 repositories · arXiv:2409.06603
-
Accelerating Large Language Model Pretraining via LFR Pedagogy: Learn, Focus, and Review 10 Sep 2024 · 0 repositories · arXiv:2409.06131
-
Adaptive Transformer Modelling of Density Function for Nonparametric Survival Analysis 10 Sep 2024 · 1 repository · arXiv:2409.06209Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 1 honoured, 0 violated, 5 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
AgileIR: Memory-Efficient Group Shifted Windows Attention for Agile Image Restoration 10 Sep 2024 · 0 repositories · arXiv:2409.06206
-
Can Large Language Models Unlock Novel Scientific Research Ideas? 10 Sep 2024 · 1 repository · arXiv:2409.06185Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 9 unverified (of 11 harvested samples)
-
Connecting Concept Convexity and Human-Machine Alignment in Deep Neural Networks 10 Sep 2024 · 0 repositories · arXiv:2409.06362
-
DA-MoE: Towards Dynamic Expert Allocation for Mixture-of-Experts Models 10 Sep 2024 · 0 repositories · arXiv:2409.06669
-
Generative AI for Requirements Engineering: A Systematic Literature Review 10 Sep 2024 · 0 repositories · arXiv:2409.06741
-
GroUSE: A Benchmark to Evaluate Evaluators in Grounded Question Answering 10 Sep 2024 · 1 repository · arXiv:2409.06595
-
Knowledge Distillation via Query Selection for Detection Transformer 10 Sep 2024 · 0 repositories · arXiv:2409.06443
-
Lightweight single-image super-resolution network based on dual paths 10 Sep 2024 · 0 repositories · arXiv:2409.06590
-
Static for Dynamic: Towards a Deeper Understanding of Dynamic Facial Expressions Using Static Expression Data 10 Sep 2024 · 1 repository · arXiv:2409.06154
-
What is the Role of Small Models in the LLM Era: A Survey 10 Sep 2024 · 1 repository · arXiv:2409.06857
-
Classification performance and reproducibility of GPT-4 omni for information extraction from veterinary electronic health records 9 Sep 2024 · 1 repository · arXiv:2409.13727
-
Rule Extrapolation in Language Models: A Study of Compositional Generalization on OOD Prompts 9 Sep 2024 · 1 repository · arXiv:2409.13728
-
A Small Claims Court for the NLP: Judging Legal Text Classification Strategies With Small Datasets 9 Sep 2024 · 0 repositories · arXiv:2409.05972
-
AbGPT: De Novo Antibody Design via Generative Language Modeling 9 Sep 2024 · 1 repository · arXiv:2409.06090
-
Application Specific Compression of Deep Learning Models 9 Sep 2024 · 1 repository · arXiv:2409.05368
-
Assessing SPARQL capabilities of Large Language Models 9 Sep 2024 · 2 repositories · arXiv:2409.05925
-
Deep Generative Model for Mechanical System Configuration Design 9 Sep 2024 · 0 repositories · arXiv:2409.06016
-
DriveScape: Towards High-Resolution Controllable Multi-View Driving Video Generation 9 Sep 2024 · 0 repositories · arXiv:2409.05463
-
DSDFormer: An Innovative Transformer-Mamba Framework for Robust High-Precision Driver Distraction Identification 9 Sep 2024 · 0 repositories · arXiv:2409.05587
-
Elsevier Arena: Human Evaluation of Chemistry/Biology/Health Foundational Large Language Models 9 Sep 2024 · 0 repositories · arXiv:2409.05486
-
Ethereum Fraud Detection via Joint Transaction Language Model and Graph Representation Learning 9 Sep 2024 · 0 repositories · arXiv:2409.07494
-
Exploring Rich Subjective Quality Information for Image Quality Assessment in the Wild 9 Sep 2024 · 0 repositories · arXiv:2409.05540
-
FairHome: A Fair Housing and Fair Lending Dataset 9 Sep 2024 · 0 repositories · arXiv:2409.05990