Methods › General › Attention Mechanisms › Attention › Papers, page 88
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 88 of 316: papers 8,701 to 8,800 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
M2M-Gen: A Multimodal Framework for Automated Background Music Generation in Japanese Manga Using Large Language Models 13 Oct 2024 · 0 repositories · arXiv:2410.09928
-
Meta-Reinforcement Learning with Universal Policy Adaptation: Provable Near-Optimality under All-task Optimum Comparator 13 Oct 2024 · 0 repositories · arXiv:2410.09728
-
Online Multi-modal Root Cause Analysis 13 Oct 2024 · 0 repositories · arXiv:2410.10021
-
Single Ground Truth Is Not Enough: Add Linguistic Variability to Aspect-based Sentiment Analysis Evaluation 13 Oct 2024 · 0 repositories · arXiv:2410.09807
-
STA-Unet: Rethink the semantic redundant for Medical Imaging Segmentation 13 Oct 2024 · 1 repository · arXiv:2410.11578
-
TextMaster: Universal Controllable Text Edit 13 Oct 2024 · 0 repositories · arXiv:2410.09879
-
Automatic Speech Recognition with BERT and CTC Transformers: A Review 12 Oct 2024 · 0 repositories · arXiv:2410.09456
-
Beyond Exact Match: Semantically Reassessing Event Extraction by Large Language Models 12 Oct 2024 · 0 repositories · arXiv:2410.09418
-
Bi-temporal Gaussian Feature Dependency Guided Change Detection in Remote Sensing Images 12 Oct 2024 · 0 repositories · arXiv:2410.09539
-
Bridging Text and Image for Artist Style Transfer via Contrastive Learning 12 Oct 2024 · 0 repositories · arXiv:2410.09566
-
CollabEdit: Towards Non-destructive Collaborative Knowledge Editing 12 Oct 2024 · 1 repository · arXiv:2410.09508Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 9 harvested samples) · 6 pointer-only (licence)
-
Diabetic retinopathy image classification method based on GreenBen data augmentation 12 Oct 2024 · 0 repositories · arXiv:2410.09444
-
EG-SpikeFormer: Eye-Gaze Guided Transformer on Spiking Neural Networks for Medical Image Analysis 12 Oct 2024 · 0 repositories · arXiv:2410.09674
-
EmbodiedCity: A Benchmark Platform for Embodied Agent in Real-world City Environment 12 Oct 2024 · 0 repositories · arXiv:2410.09604
-
Emphasis Rendering for Conversational Text-to-Speech with Multi-modal Multi-scale Context Modeling 12 Oct 2024 · 0 repositories · arXiv:2410.09524
-
Extended Japanese Commonsense Morality Dataset with Masked Token and Label Enhancement 12 Oct 2024 · 0 repositories · arXiv:2410.09564
-
Fine-grained Attention I/O Complexity: Comprehensive Analysis for Backward Passes 12 Oct 2024 · 0 repositories · arXiv:2410.09397
-
GPTON: Generative Pre-trained Transformers enhanced with Ontology Narration for accurate annotation of biological data 12 Oct 2024 · 0 repositories · arXiv:2410.10899
-
Hey AI Can You Grade My Essay?: Automatic Essay Grading 12 Oct 2024 · 0 repositories · arXiv:2410.09319
-
Improving 3D Finger Traits Recognition via Generalizable Neural Rendering 12 Oct 2024 · 0 repositories · arXiv:2410.09582
-
\llinstruct: An Instruction-tuned model for English Language Proficiency Assessments 12 Oct 2024 · 0 repositories · arXiv:2410.09314
-
Looped ReLU MLPs May Be All You Need as Practical Programmable Computers 12 Oct 2024 · 0 repositories · arXiv:2410.09375
-
ReLU's Revival: On the Entropic Overload in Normalization-Free Large Language Models 12 Oct 2024 · 1 repository · arXiv:2410.09637Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Scaled and Inter-token Relation Enhanced Transformer for Sample-restricted Residential NILM 12 Oct 2024 · 0 repositories · arXiv:2410.12861
-
SimBrainNet: Evaluating Brain Network Similarity for Attention Disorders 12 Oct 2024 · 0 repositories · arXiv:2410.09422
-
Token Pruning using a Lightweight Background Aware Vision Transformer 12 Oct 2024 · 0 repositories · arXiv:2410.09324
-
Toward General Instruction-Following Alignment for Retrieval-Augmented Generation 12 Oct 2024 · 1 repository · arXiv:2410.09584Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Training Dynamics of Transformers to Recognize Word Co-occurrence via Gradient Flow Analysis 12 Oct 2024 · 0 repositories · arXiv:2410.09605
-
Unraveling Movie Genres through Cross-Attention Fusion of Bi-Modal Synergy of Poster 12 Oct 2024 · 0 repositories · arXiv:2410.19764
-
Accelerated Distributed Stochastic Non-Convex Optimization over Time-Varying Directed Networks 11 Oct 2024 · 0 repositories · arXiv:2410.08508
-
A Methodology for Evaluating RAG Systems: A Case Study On Configuration Dependency Validation 11 Oct 2024 · 1 repository · arXiv:2410.08801
-
A Social Context-aware Graph-based Multimodal Attentive Learning Framework for Disaster Content Classification during Emergencies 11 Oct 2024 · 0 repositories · arXiv:2410.08814
-
Rethinking Gradient-Based Methods: Multi-Property Materials Design Beyond Differentiable Targets 11 Oct 2024 · 1 repository · arXiv:2410.08562
-
AttnGCG: Enhancing Jailbreaking Attacks on LLMs with Attention Manipulation 11 Oct 2024 · 1 repository · arXiv:2410.09040
-
Convolutional Neural Network Design and Evaluation for Real-Time Multivariate Time Series Fault Detection in Spacecraft Attitude Sensors 11 Oct 2024 · 0 repositories · arXiv:2410.09126
-
CoTCoNet: An Optimized Coupled Transformer-Convolutional Network with an Adaptive Graph Reconstruction for Leukemia Detection 11 Oct 2024 · 0 repositories · arXiv:2410.08797
-
Cross-Modal Bidirectional Interaction Model for Referring Remote Sensing Image Segmentation 11 Oct 2024 · 1 repository · arXiv:2410.08613
-
DAT: Dialogue-Aware Transformer with Modality-Group Fusion for Human Engagement Estimation 11 Oct 2024 · 1 repository · arXiv:2410.08470
-
DeBiFormer: Vision Transformer with Deformable Agent Bi-level Routing Attention 11 Oct 2024 · 1 repository · arXiv:2410.08582
-
Developing a Pragmatic Benchmark for Assessing Korean Legal Language Understanding in Large Language Models 11 Oct 2024 · 1 repository · arXiv:2410.08731
-
Efficiently Scanning and Resampling Spatio-Temporal Tasks with Irregular Observations 11 Oct 2024 · 0 repositories · arXiv:2410.08681
-
Encoding Agent Trajectories as Representations with Sequence Transformers 11 Oct 2024 · 0 repositories · arXiv:2410.09204
-
Enhancing Long Context Performance in LLMs Through Inner Loop Query Mechanism 11 Oct 2024 · 0 repositories · arXiv:2410.12859
-
Extra Global Attention Designation Using Keyword Detection in Sparse Transformer Architectures 11 Oct 2024 · 0 repositories · arXiv:2410.08971
-
Fine-Tuning In-House Large Language Models to Infer Differential Diagnosis from Radiology Reports 11 Oct 2024 · 0 repositories · arXiv:2410.09234
-
HorGait: A Hybrid Model for Accurate Gait Recognition in LiDAR Point Cloud Planar Projections 11 Oct 2024 · 0 repositories · arXiv:2410.08454
-
Humanity in AI: Detecting the Personality of Large Language Models 11 Oct 2024 · 0 repositories · arXiv:2410.08545
-
Hypothesis-only Biases in Large Language Model-Elicited Natural Language Inference 11 Oct 2024 · 0 repositories · arXiv:2410.08996
-
JAILJUDGE: A Comprehensive Jailbreak Judge Benchmark with Multi-Agent Enhanced Explanation Evaluation Framework 11 Oct 2024 · 1 repository · arXiv:2410.12855Syntology 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples)
-
L3Cube-MahaSum: A Comprehensive Dataset and BART Models for Abstractive Text Summarization in Marathi 11 Oct 2024 · 1 repository · arXiv:2410.09184
-
Large Language Models for Medical OSCE Assessment: A Novel Approach to Transcript Analysis 11 Oct 2024 · 0 repositories · arXiv:2410.12858
-
Learning General Representation of 12-Lead Electrocardiogram with a Joint-Embedding Predictive Architecture 11 Oct 2024 · 1 repository · arXiv:2410.08559Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Learning Interaction-aware 3D Gaussian Splatting for One-shot Hand Avatars 11 Oct 2024 · 1 repository · arXiv:2410.08840Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
Long Range Named Entity Recognition for Marathi Documents 11 Oct 2024 · 0 repositories · arXiv:2410.09192
-
Low-complexity Attention-based Unsupervised Anomalous Sound Detection exploiting Separable Convolutions and Angular Loss 11 Oct 2024 · 1 repository · arXiv:2410.08919
-
Maximizing the Potential of Synthetic Data: Insights from Random Matrix Theory 11 Oct 2024 · 0 repositories · arXiv:2410.08942
-
MeshGS: Adaptive Mesh-Aligned Gaussian Splatting for High-Quality Rendering 11 Oct 2024 · 0 repositories · arXiv:2410.08941
-
Multi-modal Fusion based Q-distribution Prediction for Controlled Nuclear Fusion 11 Oct 2024 · 0 repositories · arXiv:2410.08879
-
Multi-Source Temporal Attention Network for Precipitation Nowcasting 11 Oct 2024 · 0 repositories · arXiv:2410.08641
-
NoVo: Norm Voting off Hallucinations with Attention Heads in Large Language Models 11 Oct 2024 · 1 repository · arXiv:2410.08970Syntology 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Observing the Southern US Culture of Honor Using Large-Scale Social Media Analysis 11 Oct 2024 · 0 repositories · arXiv:2410.13887
-
On the token distance modeling ability of higher RoPE attention dimension 11 Oct 2024 · 0 repositories · arXiv:2410.08703
-
Optimized Biomedical Question-Answering Services with LLM and Multi-BERT Integration 11 Oct 2024 · 0 repositories · arXiv:2410.12856
-
oRetrieval Augmented Generation for 10 Large Language Models and its Generalizability in Assessing Medical Fitness 11 Oct 2024 · 0 repositories · arXiv:2410.08431
-
pLDDT-Predictor: High-speed Protein Screening Using Transformer and ESM2 11 Oct 2024 · 1 repository · arXiv:2410.21283
-
Quality Prediction of AI Generated Images and Videos: Emerging Trends and Opportunities 11 Oct 2024 · 0 repositories · arXiv:2410.08534
-
Retriever-and-Memory: Towards Adaptive Note-Enhanced Retrieval-Augmented Generation 11 Oct 2024 · 1 repository · arXiv:2410.08821Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
Scaling Gaussian Processes for Learning Curve Prediction via Latent Kronecker Structure 11 Oct 2024 · 0 repositories · arXiv:2410.09239
-
Score Neural Operator: A Generative Model for Learning and Generalizing Across Multiple Probability Distributions 11 Oct 2024 · 0 repositories · arXiv:2410.08549
-
SmartPretrain: Model-Agnostic and Dataset-Agnostic Representation Learning for Motion Prediction 11 Oct 2024 · 1 repository · arXiv:2410.08669Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
SocialGaze: Improving the Integration of Human Social Norms in Large Language Models 11 Oct 2024 · 1 repository · arXiv:2410.08698
-
StructRAG: Boosting Knowledge Intensive Reasoning of LLMs via Inference-time Hybrid Information Structurization 11 Oct 2024 · 1 repository · arXiv:2410.08815
-
SuperCorrect: Supervising and Correcting Language Models with Error-Driven Insights 11 Oct 2024 · 2 repositories · arXiv:2410.09008Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Synth-SONAR: Sonar Image Synthesis with Enhanced Diversity and Realism via Dual Diffusion Models and GPT Prompting 11 Oct 2024 · 1 repository · arXiv:2410.08612
-
ViT3D Alignment of LLaMA3: 3D Medical Image Report Generation 11 Oct 2024 · 0 repositories · arXiv:2410.08588
-
VOVTrack: Exploring the Potentiality in Videos for Open-Vocabulary Object Tracking 11 Oct 2024 · 0 repositories · arXiv:2410.08529
-
ZipVL: Efficient Large Vision-Language Models with Dynamic Token Sparsification 11 Oct 2024 · 0 repositories · arXiv:2410.08584
-
A Target-Aware Analysis of Data Augmentation for Hate Speech Detection 10 Oct 2024 · 0 repositories · arXiv:2410.08053
-
Adam Exploits ℓ_∞-geometry of Loss Landscape via Coordinate-wise Adaptivity 10 Oct 2024 · 1 repository · arXiv:2410.08198Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 10 harvested samples)
-
BA-Net: Bridge Attention in Deep Neural Networks 10 Oct 2024 · 0 repositories · arXiv:2410.07860
-
Benchmarking Agentic Workflow Generation 10 Oct 2024 · 1 repository · arXiv:2410.07869Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
Benign Overfitting in Single-Head Attention 10 Oct 2024 · 0 repositories · arXiv:2410.07746
-
Can Looped Transformers Learn to Implement Multi-step Gradient Descent for In-context Learning? 10 Oct 2024 · 0 repositories · arXiv:2410.08292
-
Deconstructing equivariant representations in molecular systems 10 Oct 2024 · 1 repository · arXiv:2410.08131Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 8 harvested samples)
-
Deep Learning for Generalised Planning with Background Knowledge 10 Oct 2024 · 0 repositories · arXiv:2410.07923
-
DICE: Discrete Inversion Enabling Controllable Editing for Multinomial Diffusion and Masked Generative Models 10 Oct 2024 · 0 repositories · arXiv:2410.08207
-
Diversity of Thought Elicits Stronger Reasoning Capabilities in Multi-Agent Debate Frameworks 10 Oct 2024 · 0 repositories · arXiv:2410.12853
-
Do You Know What You Are Talking About? Characterizing Query-Knowledge Relevance For Reliable Retrieval Augmented Generation 10 Oct 2024 · 0 repositories · arXiv:2410.08320
-
Emerging Pixel Grounding in Large Multimodal Models Without Grounding Supervision 10 Oct 2024 · 0 repositories · arXiv:2410.08209
-
Explainability of Deep Neural Networks for Brain Tumor Detection 10 Oct 2024 · 1 repository · arXiv:2410.07613
-
Federated Graph Learning for Cross-Domain Recommendation 10 Oct 2024 · 0 repositories · arXiv:2410.08249
-
Fine-detailed Neural Indoor Scene Reconstruction using multi-level importance sampling and multi-view consistency 10 Oct 2024 · 0 repositories · arXiv:2410.07597
-
Fine-Tuning Language Models for Ethical Ambiguity: A Comparative Study of Alignment with Human Responses 10 Oct 2024 · 0 repositories · arXiv:2410.07826
-
FLIER: Few-shot Language Image Models Embedded with Latent Representations 10 Oct 2024 · 0 repositories · arXiv:2410.07648
-
Frequency-Temporal Attention Network for Remote Sensing Imagery Change Detection 10 Oct 2024 · 1 repository
-
Full-Rank No More: Low-Rank Weight Training for Modern Speech Recognition Models 10 Oct 2024 · 0 repositories · arXiv:2410.07771
-
HeightFormer: A Semantic Alignment Monocular 3D Object Detection Method from Roadside Perspective 10 Oct 2024 · 0 repositories · arXiv:2410.07758
-
Heterogeneous Graph Auto-Encoder for CreditCard Fraud Detection 10 Oct 2024 · 0 repositories · arXiv:2410.08121
-
IceDiff: High Resolution and High-Quality Sea Ice Forecasting with Generative Diffusion Prior 10 Oct 2024 · 0 repositories · arXiv:2410.09111
-
Mind the Gap: a Spectral Analysis of Rank Collapse and Signal Propagation in Attention Layers 10 Oct 2024 · 0 repositories · arXiv:2410.07799