Methods › General › Attention Modules › Multi-Head Attention › Papers, page 42
Multi-Head Attention
Papers archive 2025-07-28
archive papers tagged: 24,855 · with a code link: 11,214 · where Syntology ran a sample: 3,454 (2,916 with a run with no instrument failure, 538 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,454 of 24,855 tagged: 2,916 with a run with no instrument failure, 538 where every run was a failure of Syntology's instrument)
Page 42 of 249: papers 4,101 to 4,200 of 24,855, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Automatic Speech Recognition with BERT and CTC Transformers: A Review 12 Oct 2024 · 0 repositories · arXiv:2410.09456
-
Diabetic retinopathy image classification method based on GreenBen data augmentation 12 Oct 2024 · 0 repositories · arXiv:2410.09444
-
EG-SpikeFormer: Eye-Gaze Guided Transformer on Spiking Neural Networks for Medical Image Analysis 12 Oct 2024 · 0 repositories · arXiv:2410.09674
-
Extended Japanese Commonsense Morality Dataset with Masked Token and Label Enhancement 12 Oct 2024 · 0 repositories · arXiv:2410.09564
-
GPTON: Generative Pre-trained Transformers enhanced with Ontology Narration for accurate annotation of biological data 12 Oct 2024 · 0 repositories · arXiv:2410.10899
-
Improving 3D Finger Traits Recognition via Generalizable Neural Rendering 12 Oct 2024 · 0 repositories · arXiv:2410.09582
-
\llinstruct: An Instruction-tuned model for English Language Proficiency Assessments 12 Oct 2024 · 0 repositories · arXiv:2410.09314
-
Looped ReLU MLPs May Be All You Need as Practical Programmable Computers 12 Oct 2024 · 0 repositories · arXiv:2410.09375
-
Token Pruning using a Lightweight Background Aware Vision Transformer 12 Oct 2024 · 0 repositories · arXiv:2410.09324
-
Toward General Instruction-Following Alignment for Retrieval-Augmented Generation 12 Oct 2024 · 1 repository · arXiv:2410.09584Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
A Methodology for Evaluating RAG Systems: A Case Study On Configuration Dependency Validation 11 Oct 2024 · 1 repository · arXiv:2410.08801
-
Rethinking Gradient-Based Methods: Multi-Property Materials Design Beyond Differentiable Targets 11 Oct 2024 · 1 repository · arXiv:2410.08562
-
AttnGCG: Enhancing Jailbreaking Attacks on LLMs with Attention Manipulation 11 Oct 2024 · 1 repository · arXiv:2410.09040
-
CoTCoNet: An Optimized Coupled Transformer-Convolutional Network with an Adaptive Graph Reconstruction for Leukemia Detection 11 Oct 2024 · 0 repositories · arXiv:2410.08797
-
DAT: Dialogue-Aware Transformer with Modality-Group Fusion for Human Engagement Estimation 11 Oct 2024 · 1 repository · arXiv:2410.08470
-
DeBiFormer: Vision Transformer with Deformable Agent Bi-level Routing Attention 11 Oct 2024 · 1 repository · arXiv:2410.08582
-
Developing a Pragmatic Benchmark for Assessing Korean Legal Language Understanding in Large Language Models 11 Oct 2024 · 1 repository · arXiv:2410.08731
-
Efficiently Scanning and Resampling Spatio-Temporal Tasks with Irregular Observations 11 Oct 2024 · 0 repositories · arXiv:2410.08681
-
Encoding Agent Trajectories as Representations with Sequence Transformers 11 Oct 2024 · 0 repositories · arXiv:2410.09204
-
Enhancing Long Context Performance in LLMs Through Inner Loop Query Mechanism 11 Oct 2024 · 0 repositories · arXiv:2410.12859
-
Extra Global Attention Designation Using Keyword Detection in Sparse Transformer Architectures 11 Oct 2024 · 0 repositories · arXiv:2410.08971
-
Fine-Tuning In-House Large Language Models to Infer Differential Diagnosis from Radiology Reports 11 Oct 2024 · 0 repositories · arXiv:2410.09234
-
HorGait: A Hybrid Model for Accurate Gait Recognition in LiDAR Point Cloud Planar Projections 11 Oct 2024 · 0 repositories · arXiv:2410.08454
-
Humanity in AI: Detecting the Personality of Large Language Models 11 Oct 2024 · 0 repositories · arXiv:2410.08545
-
Hypothesis-only Biases in Large Language Model-Elicited Natural Language Inference 11 Oct 2024 · 0 repositories · arXiv:2410.08996
-
JAILJUDGE: A Comprehensive Jailbreak Judge Benchmark with Multi-Agent Enhanced Explanation Evaluation Framework 11 Oct 2024 · 1 repository · arXiv:2410.12855Syntology 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples)
-
L3Cube-MahaSum: A Comprehensive Dataset and BART Models for Abstractive Text Summarization in Marathi 11 Oct 2024 · 1 repository · arXiv:2410.09184
-
Large Language Models for Medical OSCE Assessment: A Novel Approach to Transcript Analysis 11 Oct 2024 · 0 repositories · arXiv:2410.12858
-
Long Range Named Entity Recognition for Marathi Documents 11 Oct 2024 · 0 repositories · arXiv:2410.09192
-
Observing the Southern US Culture of Honor Using Large-Scale Social Media Analysis 11 Oct 2024 · 0 repositories · arXiv:2410.13887
-
Optimized Biomedical Question-Answering Services with LLM and Multi-BERT Integration 11 Oct 2024 · 0 repositories · arXiv:2410.12856
-
oRetrieval Augmented Generation for 10 Large Language Models and its Generalizability in Assessing Medical Fitness 11 Oct 2024 · 0 repositories · arXiv:2410.08431
-
pLDDT-Predictor: High-speed Protein Screening Using Transformer and ESM2 11 Oct 2024 · 1 repository · arXiv:2410.21283
-
Retriever-and-Memory: Towards Adaptive Note-Enhanced Retrieval-Augmented Generation 11 Oct 2024 · 1 repository · arXiv:2410.08821Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
Scaling Gaussian Processes for Learning Curve Prediction via Latent Kronecker Structure 11 Oct 2024 · 0 repositories · arXiv:2410.09239
-
SocialGaze: Improving the Integration of Human Social Norms in Large Language Models 11 Oct 2024 · 1 repository · arXiv:2410.08698
-
StructRAG: Boosting Knowledge Intensive Reasoning of LLMs via Inference-time Hybrid Information Structurization 11 Oct 2024 · 1 repository · arXiv:2410.08815
-
SuperCorrect: Supervising and Correcting Language Models with Error-Driven Insights 11 Oct 2024 · 2 repositories · arXiv:2410.09008Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Synth-SONAR: Sonar Image Synthesis with Enhanced Diversity and Realism via Dual Diffusion Models and GPT Prompting 11 Oct 2024 · 1 repository · arXiv:2410.08612
-
ViT3D Alignment of LLaMA3: 3D Medical Image Report Generation 11 Oct 2024 · 0 repositories · arXiv:2410.08588
-
Adam Exploits ℓ_∞-geometry of Loss Landscape via Coordinate-wise Adaptivity 10 Oct 2024 · 1 repository · arXiv:2410.08198Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 10 harvested samples)
-
Benchmarking Agentic Workflow Generation 10 Oct 2024 · 1 repository · arXiv:2410.07869Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
Can Looped Transformers Learn to Implement Multi-step Gradient Descent for In-context Learning? 10 Oct 2024 · 0 repositories · arXiv:2410.08292
-
DICE: Discrete Inversion Enabling Controllable Editing for Multinomial Diffusion and Masked Generative Models 10 Oct 2024 · 0 repositories · arXiv:2410.08207
-
Diversity of Thought Elicits Stronger Reasoning Capabilities in Multi-Agent Debate Frameworks 10 Oct 2024 · 0 repositories · arXiv:2410.12853
-
Do You Know What You Are Talking About? Characterizing Query-Knowledge Relevance For Reliable Retrieval Augmented Generation 10 Oct 2024 · 0 repositories · arXiv:2410.08320
-
Explainability of Deep Neural Networks for Brain Tumor Detection 10 Oct 2024 · 1 repository · arXiv:2410.07613
-
Fine-Tuning Language Models for Ethical Ambiguity: A Comparative Study of Alignment with Human Responses 10 Oct 2024 · 0 repositories · arXiv:2410.07826
-
FLIER: Few-shot Language Image Models Embedded with Latent Representations 10 Oct 2024 · 0 repositories · arXiv:2410.07648
-
IceDiff: High Resolution and High-Quality Sea Ice Forecasting with Generative Diffusion Prior 10 Oct 2024 · 0 repositories · arXiv:2410.09111
-
News Reporter: A Multi-lingual LLM Framework for Broadcast T.V News 10 Oct 2024 · 0 repositories · arXiv:2410.07520
-
No Free Lunch: Retrieval-Augmented Generation Undermines Fairness in LLMs, Even for Vigilant Users 10 Oct 2024 · 0 repositories · arXiv:2410.07589
-
Offline Inverse Constrained Reinforcement Learning for Safe-Critical Decision Making in Healthcare 10 Oct 2024 · 0 repositories · arXiv:2410.07525
-
PLaMo-100B: A Ground-Up Language Model Designed for Japanese Proficiency 10 Oct 2024 · 0 repositories · arXiv:2410.07563
-
Pretraining Graph Transformers with Atom-in-a-Molecule Quantum Properties for Improved ADMET Modeling 10 Oct 2024 · 1 repository · arXiv:2410.08024
-
Privately Learning from Graphs with Applications in Fine-tuning Large Language Models 10 Oct 2024 · 1 repository · arXiv:2410.08299
-
Prompt Engineering a Schizophrenia Chatbot: Utilizing a Multi-Agent Approach for Enhanced Compliance with Prompt Instructions 10 Oct 2024 · 0 repositories · arXiv:2410.12848
-
RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation 10 Oct 2024 · 1 repository · arXiv:2410.07864Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples)
-
Reducing the Cost of Dropout in Flash-Attention by Hiding RNG with GEMM 10 Oct 2024 · 0 repositories · arXiv:2410.07531
-
Rescriber: Smaller-LLM-Powered User-Led Data Minimization for LLM-Based Chatbots 10 Oct 2024 · 0 repositories · arXiv:2410.11876
-
Robust AI-Generated Text Detection by Restricted Embeddings 10 Oct 2024 · 1 repository · arXiv:2410.08113Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
SeMv-3D: Towards Concurrency of Semantic and Multi-view Consistency in General Text-to-3D Generation 10 Oct 2024 · 0 repositories · arXiv:2410.07658
-
SNN-PAR: Energy Efficient Pedestrian Attribute Recognition via Spiking Neural Networks 10 Oct 2024 · 1 repository · arXiv:2410.07857
-
SPA: 3D Spatial-Awareness Enables Effective Embodied Representation 10 Oct 2024 · 1 repository · arXiv:2410.08208Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples)
-
Teaching-Inspired Integrated Prompting Framework: A Novel Approach for Enhancing Reasoning in Large Language Models 10 Oct 2024 · 1 repository · arXiv:2410.08068Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
The Rise of AI-Generated Content in Wikipedia 10 Oct 2024 · 1 repository · arXiv:2410.08044Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
Theoretical limits of descending ℓ₀ sparse-regression ML algorithms 10 Oct 2024 · 0 repositories · arXiv:2410.07651
-
Think Beyond Size: Adaptive Prompting for More Effective Reasoning 10 Oct 2024 · 0 repositories · arXiv:2410.08130
-
Thought2Text: Text Generation from EEG Signal using Large Language Models (LLMs) 10 Oct 2024 · 1 repository · arXiv:2410.07507Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
TurboRAG: Accelerating Retrieval-Augmented Generation with Precomputed KV Caches for Chunked Text 10 Oct 2024 · 1 repository · arXiv:2410.07590
-
VibeCheck: Discover and Quantify Qualitative Differences in Large Language Models 10 Oct 2024 · 1 repository · arXiv:2410.12851Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
A Two-Model Approach for Humour Style Recognition 9 Oct 2024 · 1 repository · arXiv:2410.12842
-
Adaptive High-Frequency Transformer for Diverse Wildlife Re-Identification 9 Oct 2024 · 1 repository · arXiv:2410.06977
-
Astute RAG: Overcoming Imperfect Retrieval Augmentation and Knowledge Conflicts for Large Language Models 9 Oct 2024 · 0 repositories · arXiv:2410.07176
-
AutoFeedback: An LLM-based Framework for Efficient and Accurate API Request Generation 9 Oct 2024 · 0 repositories · arXiv:2410.06943
-
Bridge the Points: Graph-based Few-shot Segment Anything Semantically 9 Oct 2024 · 1 repository · arXiv:2410.06964Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples)
-
Can Transformers Reason Logically? A Study in SAT Solving 9 Oct 2024 · 0 repositories · arXiv:2410.07432
-
Capturing Bias Diversity in LLMs 9 Oct 2024 · 0 repositories · arXiv:2410.12839
-
Cluster-wise Graph Transformer with Dual-granularity Kernelized Attention 9 Oct 2024 · 1 repository · arXiv:2410.06746Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 2 honoured, 0 violated, 8 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 14 harvested samples) · 14 pointer-only (licence)
-
Detecting Bias and Enhancing Diagnostic Accuracy in Large Language Models for Healthcare 9 Oct 2024 · 0 repositories · arXiv:2410.06566
-
Do great minds think alike? Investigating Human-AI Complementarity in Question Answering with CAIMIRA 9 Oct 2024 · 0 repositories · arXiv:2410.06524
-
Efficient training strategies for natural sounding speech synthesis and speaker adaptation based on FastPitch 9 Oct 2024 · 0 repositories · arXiv:2410.06787
-
ETA: Evaluating Then Aligning Safety of Vision Language Models at Inference Time 9 Oct 2024 · 1 repository · arXiv:2410.06625Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 1 violated, 5 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
F5-TTS: A Fairytaler that Fakes Fluent and Faithful Speech with Flow Matching 9 Oct 2024 · 1 repository · arXiv:2410.06885Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples)
-
Generative Model for Less-Resourced Language with 1 billion parameters 9 Oct 2024 · 0 repositories · arXiv:2410.06898
-
Improving Data Efficiency via Curating LLM-Driven Rating Systems 9 Oct 2024 · 0 repositories · arXiv:2410.10877
-
Instructional Segment Embedding: Improving LLM Safety with Instruction Hierarchy 9 Oct 2024 · 0 repositories · arXiv:2410.09102
-
Investigating Cost-Efficiency of LLM-Generated Training Data for Conversational Semantic Frame Analysis 9 Oct 2024 · 0 repositories · arXiv:2410.06550
-
Large Language Models as Code Executors: An Exploratory Study 9 Oct 2024 · 0 repositories · arXiv:2410.06667
-
LLM Self-Correction with DeCRIM: Decompose, Critique, and Refine for Enhanced Following of Instructions with Multiple Constraints 9 Oct 2024 · 0 repositories · arXiv:2410.06458
-
MaD-Scientist: AI-based Scientist solving Convection-Diffusion-Reaction Equations Using Massive PINN-Based Prior Data 9 Oct 2024 · 0 repositories · arXiv:2410.06442
-
MatMamba: A Matryoshka State Space Model 9 Oct 2024 · 1 repository · arXiv:2410.06718
-
Mental Disorders Detection in the Era of Large Language Models 9 Oct 2024 · 0 repositories · arXiv:2410.07129
-
MentalArena: Self-play Training of Language Models for Diagnosis and Treatment of Mental Health Disorders 9 Oct 2024 · 1 repository · arXiv:2410.06845
-
NetDiff: Deep Graph Denoising Diffusion for Ad Hoc Network Topology Generation 9 Oct 2024 · 0 repositories · arXiv:2410.08238
-
Optimizing Transformer based on high-performance optimizer for predicting employment sentiment in American social media content 9 Oct 2024 · 0 repositories · arXiv:2410.10874
-
Pair-VPR: Place-Aware Pre-training and Contrastive Pair Classification for Visual Place Recognition with Vision Transformers 9 Oct 2024 · 1 repository · arXiv:2410.06614
-
QuadMamba: Learning Quadtree-based Selective Scan for Visual State Space Model 9 Oct 2024 · 1 repository · arXiv:2410.06806
-
Retrieval-Augmented Decision Transformer: External Memory for In-context RL 9 Oct 2024 · 1 repository · arXiv:2410.07071Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples)
-
SAGE: Scalable Ground Truth Evaluations for Large Sparse Autoencoders 9 Oct 2024 · 0 repositories · arXiv:2410.07456