Methods › General › Attention Modules › Multi-Head Attention › Papers, page 71
Multi-Head Attention
Papers archive 2025-07-28
archive papers tagged: 24,855 · with a code link: 11,214 · where Syntology ran a sample: 3,454 (2,916 with a run with no instrument failure, 538 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,454 of 24,855 tagged: 2,916 with a run with no instrument failure, 538 where every run was a failure of Syntology's instrument)
Page 71 of 249: papers 7,001 to 7,100 of 24,855, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
A Transformer variant for multi-step forecasting of water level and hydrometeorological sensitivity analysis based on explainable artificial intelligence technology 22 May 2024 · 0 repositories · arXiv:2405.13646
-
A General Graph Spectral Wavelet Convolution via Chebyshev Order Decomposition 22 May 2024 · 1 repository · arXiv:2405.13806Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
Affine-based Deformable Attention and Selective Fusion for Semi-dense Matching 22 May 2024 · 0 repositories · arXiv:2405.13874
-
Automated Evaluation of Retrieval-Augmented Language Models with Task-Specific Exam Generation 22 May 2024 · 1 repository · arXiv:2405.13622
-
Automatically Identifying Local and Global Circuits with Linear Computation Graphs 22 May 2024 · 0 repositories · arXiv:2405.13868
-
CViT: Continuous Vision Transformer for Operator Learning 22 May 2024 · 2 repositories · arXiv:2405.13998Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
Comparative Analysis of Hyperspectral Image Reconstruction Using Deep Learning for Agricultural and Biological Applications 22 May 2024 · 0 repositories · arXiv:2405.13331
-
Discrete Cosine Transform Based Decorrelated Attention for Vision Transformers 22 May 2024 · 0 repositories · arXiv:2405.13901
-
Evaluating Large Language Models with Human Feedback: Establishing a Swedish Benchmark 22 May 2024 · 1 repository · arXiv:2405.14006
-
FlashRAG: A Modular Toolkit for Efficient Retrieval-Augmented Generation Research 22 May 2024 · 1 repository · arXiv:2405.13576
-
From CNNs to Transformers in Multimodal Human Action Recognition: A Survey 22 May 2024 · 0 repositories · arXiv:2405.15813
-
High Performance P300 Spellers Using GPT2 Word Prediction With Cross-Subject Training 22 May 2024 · 0 repositories · arXiv:2405.13329
-
KU-DMIS at EHRSQL 2024:Generating SQL query via question templatization in EHR 22 May 2024 · 0 repositories · arXiv:2406.00014
-
Leveraging 2D Information for Long-term Time Series Forecasting with Vanilla Transformers 22 May 2024 · 1 repository · arXiv:2405.13810
-
LookHere: Vision Transformers with Directed Attention Generalize and Extrapolate 22 May 2024 · 1 repository · arXiv:2405.13985Syntology official (archive's flag): 8 ran · 8 ran (of which 5 constructed an object rather than computing a result; 7 with no instrument failure: 2 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (of 12 harvested samples)
-
Unlocking the Power of Patch: Patch-Based MLP for Long-Term Time Series Forecasting 22 May 2024 · 0 repositories · arXiv:2405.13575
-
Semantic Equitable Clustering: A Simple and Effective Strategy for Clustering Vision Tokens 22 May 2024 · 0 repositories · arXiv:2405.13337
-
Task-agnostic Decision Transformer for Multi-type Agent Control with Federated Split Training 22 May 2024 · 0 repositories · arXiv:2405.13445
-
Text Prompting for Multi-Concept Video Customization by Autoregressive Generation 22 May 2024 · 0 repositories · arXiv:2405.13951
-
TOPA: Extending Large Language Models for Video Understanding via Text-Only Pre-Alignment 22 May 2024 · 1 repository · arXiv:2405.13911Syntology official (archive's flag): 3 ran · 6 ran (of which 3 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 1 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 9 harvested samples) · 3 pointer-only (licence)
-
TrojanRAG: Retrieval-Augmented Generation Can Be Backdoor Driver in Large Language Models 22 May 2024 · 1 repository · arXiv:2405.13401
-
Unleashing the Power of Unlabeled Data: A Self-supervised Learning Framework for Cyber Attack Detection in Smart Grids 22 May 2024 · 0 repositories · arXiv:2405.13965
-
Unsupervised Pre-training with Language-Vision Prompts for Low-Data Instance Segmentation 22 May 2024 · 1 repository · arXiv:2405.13388
-
Why Not Transform Chat Large Language Models to Non-English? 22 May 2024 · 1 repository · arXiv:2405.13923
-
WordGame: Efficient & Effective LLM Jailbreak via Simultaneous Obfuscation in Query and Response 22 May 2024 · 0 repositories · arXiv:2405.14023
-
A Masked Semi-Supervised Learning Approach for Otago Micro Labels Recognition 21 May 2024 · 0 repositories · arXiv:2405.12711
-
BIMM: Brain Inspired Masked Modeling for Video Representation Learning 21 May 2024 · 1 repository · arXiv:2405.12757
-
BiomedParse: a biomedical foundation model for image parsing of everything everywhere all at once 21 May 2024 · 0 repositories · arXiv:2405.12971
-
Enhancing Transformer-based models for Long Sequence Time Series Forecasting via Structured Matrix 21 May 2024 · 1 repository · arXiv:2405.12462
-
Computational Tradeoffs in Image Synthesis: Diffusion, Masked-Token, and Next-Token Prediction 21 May 2024 · 0 repositories · arXiv:2405.13218
-
GASE: Graph Attention Sampling with Edges Fusion for Solving Vehicle Routing Problems 21 May 2024 · 0 repositories · arXiv:2405.12475
-
Generative AI in Cybersecurity: A Comprehensive Review of LLM Applications and Vulnerabilities 21 May 2024 · 0 repositories · arXiv:2405.12750
-
Global-Local Detail Guided Transformer for Sea Ice Recognition in Optical Remote Sensing Images 21 May 2024 · 0 repositories · arXiv:2405.13197
-
GPT-4 Jailbreaks Itself with Near-Perfect Success Using Self-Explanation 21 May 2024 · 0 repositories · arXiv:2405.13077
-
How Reliable AI Chatbots are for Disease Prediction from Patient Complaints? 21 May 2024 · 0 repositories · arXiv:2405.13219
-
Investigating Persuasion Techniques in Arabic: An Empirical Study Leveraging Large Language Models 21 May 2024 · 0 repositories · arXiv:2405.12884
-
Is Dataset Quality Still a Concern in Diagnosis Using Large Foundation Model? 21 May 2024 · 0 repositories · arXiv:2405.12584
-
Mamba in Speech: Towards an Alternative to Self-Attention 21 May 2024 · 1 repository · arXiv:2405.12609
-
Mitigating Overconfidence in Out-of-Distribution Detection by Capturing Extreme Activations 21 May 2024 · 1 repository · arXiv:2405.12658
-
Global-local Fourier Neural Operator for Accelerating Coronal Magnetic Field Model 21 May 2024 · 1 repository · arXiv:2405.12754Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
PathOCL: Path-Based Prompt Augmentation for OCL Generation with GPT-4 21 May 2024 · 0 repositories · arXiv:2405.12450
-
Pseudo Channel: Time Embedding for Motor Imagery Decoding 21 May 2024 · 0 repositories · arXiv:2405.15812
-
Quantifying Semantic Emergence in Language Models 21 May 2024 · 1 repository · arXiv:2405.12617Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 6 harvested samples)
-
Self-Supervised Modality-Agnostic Pre-Training of Swin Transformers 21 May 2024 · 1 repository · arXiv:2405.12781
-
System Safety Monitoring of Learned Components Using Temporal Metric Forecasting 21 May 2024 · 0 repositories · arXiv:2405.13254
-
The 2nd FutureDial Challenge: Dialog Systems with Retrieval Augmented Generation (FutureDial-RAG) 21 May 2024 · 1 repository · arXiv:2405.13084
-
Time Matters: Enhancing Pre-trained News Recommendation Models with Robust User Dwell Time Injection 21 May 2024 · 0 repositories · arXiv:2405.12486
-
Transformer in Touch: A Survey 21 May 2024 · 0 repositories · arXiv:2405.12779
-
A review on the use of large language models as virtual tutors 20 May 2024 · 0 repositories · arXiv:2405.11983
-
Asymptotic theory of in-context learning by linear attention 20 May 2024 · 1 repository · arXiv:2405.11751Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Can AI Relate: Testing Large Language Model Response for Mental Health Support 20 May 2024 · 1 repository · arXiv:2405.12021Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 10 harvested samples)
-
CReMa: Crisis Response through Computational Identification and Matching of Cross-Lingual Requests and Offers Shared on Social Media 20 May 2024 · 0 repositories · arXiv:2405.11897
-
CT-Eval: Benchmarking Chinese Text-to-Table Performance in Large Language Models 20 May 2024 · 0 repositories · arXiv:2405.12174
-
Degree of Irrationality: Sentiment and Implied Volatility Surface 20 May 2024 · 0 repositories · arXiv:2405.11730
-
Efficiency optimization of large-scale language models based on deep learning in natural language processing tasks 20 May 2024 · 0 repositories · arXiv:2405.11704
-
Evaluating and Modeling Social Intelligence: A Comparative Study of Human and AI Capabilities 20 May 2024 · 1 repository · arXiv:2405.11841
-
Fennec: Fine-grained Language Model Evaluation and Correction Extended through Branching and Bridging 20 May 2024 · 1 repository · arXiv:2405.12163
-
Is Mamba Compatible with Trajectory Optimization in Offline Reinforcement Learning? 20 May 2024 · 1 repository · arXiv:2405.12094Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (of 7 harvested samples) · 1 pointer-only (licence)
-
Large-Scale Multi-Center CT and MRI Segmentation of Pancreas with Deep Learning 20 May 2024 · 1 repository · arXiv:2405.12367
-
Metacognitive Capabilities of LLMs: An Exploration in Mathematical Problem Solving 20 May 2024 · 0 repositories · arXiv:2405.12205
-
Question-Based Retrieval using Atomic Units for Enterprise RAG 20 May 2024 · 0 repositories · arXiv:2405.12363
-
SSAMBA: Self-Supervised Audio Representation Learning with Mamba State Space Model 20 May 2024 · 1 repository · arXiv:2405.11831Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 1 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 8 harvested samples)
-
A Method on Searching Better Activation Functions 19 May 2024 · 0 repositories · arXiv:2405.12954
-
A Multi-Perspective Analysis of Memorization in Large Language Models 19 May 2024 · 0 repositories · arXiv:2405.11577
-
ColorFoil: Investigating Color Blindness in Large Vision and Language Models 19 May 2024 · 1 repository · arXiv:2405.11685
-
DaVinci at SemEval-2024 Task 9: Few-shot prompting GPT-3.5 for Unconventional Reasoning 19 May 2024 · 0 repositories · arXiv:2405.11559
-
Du-IN: Discrete units-guided mask modeling for decoding speech from Intracranial Neural signals 19 May 2024 · 1 repository · arXiv:2405.11459
-
Human-Centered LLM-Agent User Interface: A Position Paper 19 May 2024 · 1 repository · arXiv:2405.13050
-
Hummer: Towards Limited Competitive Preference Dataset 19 May 2024 · 0 repositories · arXiv:2405.11647
-
Hybrid CNN-Transformer Architecture for Efficient Large-Scale Video Snapshot Compressive Imaging 19 May 2024 · 1 repository
-
Large Language Models Can Infer Personality from Free-Form User Interactions 19 May 2024 · 0 repositories · arXiv:2405.13052
-
MHPP: Exploring the Capabilities and Limitations of Language Models Beyond Basic Code Generation 19 May 2024 · 1 repository · arXiv:2405.11430
-
NetMamba: Efficient Network Traffic Classification via Pre-training Unidirectional Mamba 19 May 2024 · 1 repository · arXiv:2405.11449
-
Review of deep learning models for crypto price prediction: implementation and evaluation 19 May 2024 · 2 repositories · arXiv:2405.11431
-
Track Anything Rapter(TAR) 19 May 2024 · 1 repository · arXiv:2405.11655
-
VCformer: Variable Correlation Transformer with Inherent Lagged Correlation for Multivariate Time Series Forecasting 19 May 2024 · 1 repository · arXiv:2405.11470Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Your Transformer is Secretly Linear 19 May 2024 · 1 repository · arXiv:2405.12250Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Zero-Shot Stance Detection using Contextual Data Generation with LLMs 19 May 2024 · 1 repository · arXiv:2405.11637Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
A Dual Power Grid Cascading Failure Model for the Vulnerability Analysis 18 May 2024 · 0 repositories · arXiv:2405.11311
-
Automating PTSD Diagnostics in Clinical Interviews: Leveraging Large Language Models for Trauma Assessments 18 May 2024 · 0 repositories · arXiv:2405.11178
-
Can Public LLMs be used for Self-Diagnosis of Medical Conditions ? 18 May 2024 · 0 repositories · arXiv:2405.11407
-
Cross-Language Assessment of Mathematical Capability of ChatGPT 18 May 2024 · 0 repositories · arXiv:2405.11264
-
Exploring speech style spaces with language models: Emotional TTS without emotion labels 18 May 2024 · 0 repositories · arXiv:2405.11413
-
Towards SAR Automatic Target Recognition MultiCategory SAR Image Classification Based on Light Weight Vision Transformer 18 May 2024 · 0 repositories · arXiv:2407.06128
-
A Hybrid Deep Learning Framework for Stock Price Prediction Considering the Investor Sentiment of Online Forum Enhanced by Popularity 17 May 2024 · 0 repositories · arXiv:2405.10584
-
ActiveLLM: Large Language Model-based Active Learning for Textual Few-Shot Scenarios 17 May 2024 · 0 repositories · arXiv:2405.10808
-
Are Large Language Models Moral Hypocrites? A Study Based on Moral Foundations 17 May 2024 · 0 repositories · arXiv:2405.11100
-
Benchmarking Large Language Models on CFLUE -- A Chinese Financial Language Understanding Evaluation Dataset 17 May 2024 · 2 repositories · arXiv:2405.10542
-
DINO as a von Mises-Fisher mixture model 17 May 2024 · 0 repositories · arXiv:2405.10939
-
Empowering Prior to Court Legal Analysis: A Transparent and Accessible Dataset for Defensive Statement Classification and Interpretation 17 May 2024 · 0 repositories · arXiv:2405.10702
-
Enhancing Dialogue State Tracking Models through LLM-backed User-Agents Simulation 17 May 2024 · 0 repositories · arXiv:2405.13037
-
Enhancing the analysis of murine neonatal ultrasonic vocalizations: Development, evaluation, and application of different mathematical models 17 May 2024 · 1 repository · arXiv:2405.12957
-
Evaluation of large language model performance on the Biomedical Language Understanding and Reasoning Benchmark 17 May 2024 · 0 repositories
-
Hi-GMAE: Hierarchical Graph Masked Autoencoders 17 May 2024 · 1 repository · arXiv:2405.10642
-
Know in AdVance: Linear-Complexity Forecasting of Ad Campaign Performance with Evolving User Interest 17 May 2024 · 1 repository · arXiv:2405.10681
-
Language Models can Evaluate Themselves via Probability Discrepancy 17 May 2024 · 1 repository · arXiv:2405.10516
-
Language Models can Exploit Cross-Task In-context Learning for Data-Scarce Novel Tasks 17 May 2024 · 1 repository · arXiv:2405.10548Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples)
-
Large Language Models in Wireless Application Design: In-Context Learning-enhanced Automatic Network Intrusion Detection 17 May 2024 · 0 repositories · arXiv:2405.11002
-
NeuroAssist: Enhancing Cognitive-Computer Synergy with Adaptive AI and Advanced Neural Decoding for Efficient EEG Signal Classification 17 May 2024 · 0 repositories · arXiv:2406.01600
-
Observational Scaling Laws and the Predictability of Language Model Performance 17 May 2024 · 1 repository · arXiv:2405.10938Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 10 harvested samples)