Methods › General › Attention Mechanisms › Attention › Papers, page 12
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 12 of 316: papers 1,101 to 1,200 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Artificial Intelligence Bias on English Language Learners in Automatic Scoring 15 May 2025 · 0 repositories · arXiv:2505.10643
-
Advancing Multiple Instance Learning with Continual Learning for Whole Slide Imaging 15 May 2025 · 0 repositories · arXiv:2505.10649
-
Seasonal Forecasting of Pan-Arctic Sea Ice with State Space Model 15 May 2025 · 1 repository · arXiv:2505.10665
-
GaussianFormer3D: Multi-Modal Gaussian-based Semantic Occupancy Prediction with 3D Deformable Attention 15 May 2025 · 0 repositories · arXiv:2505.10685
-
A Modular Approach for Clinical SLMs Driven by Synthetic Data with Pre-Instruction Tuning, Model Merging, and Clinical-Tasks Alignment 15 May 2025 · 0 repositories · arXiv:2505.10717
-
IMAGE-ALCHEMY: Advancing subject fidelity in personalised text-to-image generation 15 May 2025 · 0 repositories · arXiv:2505.10743
-
Advances in Radiance Field for Dynamic Scene: From Neural Field to Gaussian Field 15 May 2025 · 1 repository · arXiv:2505.10049
-
AI Agents vs. Agentic AI: A Conceptual Taxonomy, Applications and Challenges 15 May 2025 · 0 repositories · arXiv:2505.10468
-
All You Need Is Synthetic Task Augmentation 15 May 2025 · 0 repositories · arXiv:2505.10120
-
Are Sparse Autoencoders Useful for Java Function Bug Detection? 15 May 2025 · 1 repository · arXiv:2505.10375
-
Assessing Collective Reasoning in Multi-Agent LLMs via Hidden Profile Tasks 15 May 2025 · 0 repositories · arXiv:2505.11556
-
Automating Security Audit Using Large Language Model based Agent: An Exploration Experiment 15 May 2025 · 0 repositories · arXiv:2505.10732
-
Avocado Price Prediction Using a Hybrid Deep Learning Model: TCN-MLP-Attention Architecture 15 May 2025 · 0 repositories · arXiv:2505.09907
-
CAFE: Retrieval Head-based Coarse-to-Fine Information Seeking to Enhance Multi-Document QA Capability 15 May 2025 · 0 repositories · arXiv:2505.10063
-
CL-RAG: Bridging the Gap in Retrieval-Augmented Generation with Curriculum Learning 15 May 2025 · 0 repositories · arXiv:2505.10493
-
Comparing LLM Text Annotation Skills: A Study on Human Rights Violations in Social Media Data 15 May 2025 · 1 repository · arXiv:2505.10260
-
ComplexFormer: Disruptively Advancing Transformer Inference Ability via Head-Specific Complex Vector Attention 15 May 2025 · 1 repository · arXiv:2505.10222
-
Defending the Edge: Representative-Attention for Mitigating Backdoor Attacks in Federated Learning 15 May 2025 · 0 repositories · arXiv:2505.10297
-
Does Scaling Law Apply in Time Series Forecasting? 15 May 2025 · 0 repositories · arXiv:2505.10172
-
Exploring Implicit Visual Misunderstandings in Multimodal Large Language Models through Attention Analysis 15 May 2025 · 1 repository · arXiv:2505.10541
-
Hierarchical Document Refinement for Long-context Retrieval-augmented Generation 15 May 2025 · 1 repository · arXiv:2505.10413
-
ILIF: Temporal Inhibitory Leaky Integrate-and-Fire Neuron for Overactivation in Spiking Neural Networks 15 May 2025 · 1 repository · arXiv:2505.10371
-
Leveraging Graph Retrieval-Augmented Generation to Support Learners' Understanding of Knowledge Concepts in MOOCs 15 May 2025 · 0 repositories · arXiv:2505.10074
-
MASS: Multi-Agent Simulation Scaling for Portfolio Construction 15 May 2025 · 1 repository · arXiv:2505.10278
-
MSCI: Addressing CLIP's Inherent Limitations for Compositional Zero-Shot Learning 15 May 2025 · 1 repository · arXiv:2505.10289Syntology official (archive's flag): 22 ran · 23 ran (of which 13 constructed an object rather than computing a result; 19 with no instrument failure: 0 honoured, 0 violated, 19 with no contract checked; 4 where Syntology's instrument failed) · 5 unverified (of 28 harvested samples) · 28 pointer-only (licence)
-
MTVCrafter: 4D Motion Tokenization for Open-World Human Image Animation 15 May 2025 · 1 repository · arXiv:2505.10238
-
On Technique Identification and Threat-Actor Attribution using LLMs and Embedding Models 15 May 2025 · 1 repository · arXiv:2505.11547
-
One Shot Dominance: Knowledge Poisoning Attack on Retrieval-Augmented Generation Systems 15 May 2025 · 0 repositories · arXiv:2505.11548
-
Optimizing Electric Bus Charging Scheduling with Uncertainties Using Hierarchical Deep Reinforcement Learning 15 May 2025 · 0 repositories · arXiv:2505.10296
-
Pre-Act: Multi-Step Planning and Reasoning Improves Acting in LLM Agents 15 May 2025 · 0 repositories · arXiv:2505.09970
-
Private Transformer Inference in MLaaS: A Survey 15 May 2025 · 0 repositories · arXiv:2505.10315
-
Rethinking Prompt Optimizers: From Prompt Merits to Optimization 15 May 2025 · 1 repository · arXiv:2505.09930
-
SpikeVideoFormer: An Efficient Spike-Driven Video Transformer with Hamming Attention and 𝒪(T) Complexity 15 May 2025 · 1 repository · arXiv:2505.10352Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
TKFNet: Learning Texture Key Factor Driven Feature for Facial Expression Recognition 15 May 2025 · 0 repositories · arXiv:2505.09967
-
UniEval: Unified Holistic Evaluation for Unified Multimodal Understanding and Generation 15 May 2025 · 0 repositories · arXiv:2505.10483
-
VRU-CIPI: Crossing Intention Prediction at Intersections for Improving Vulnerable Road Users Safety 15 May 2025 · 0 repositories · arXiv:2505.09935
-
Aquarius: A Family of Industry-Level Video Generation Models for Marketing Scenarios 14 May 2025 · 0 repositories · arXiv:2505.10584
-
2D-3D Attention and Entropy for Pose Robust 2D Facial Recognition 14 May 2025 · 0 repositories · arXiv:2505.09073
-
A Comprehensive Analysis of Large Language Model Outputs: Similarity, Diversity, and Bias 14 May 2025 · 0 repositories · arXiv:2505.09056
-
Accelerating Machine Learning Systems via Category Theory: Applications to Spherical Attention for Gene Regulatory Networks 14 May 2025 · 0 repositories · arXiv:2505.09326
-
AdaFortiTran: An Adaptive Transformer Model for Robust OFDM Channel Estimation 14 May 2025 · 1 repository · arXiv:2505.09076
-
An Efficient deep learning model to Predict Stock Price Movement Based on Limit Order Book 14 May 2025 · 0 repositories · arXiv:2505.22678
-
Atomic Consistency Preference Optimization for Long-Form Question Answering 14 May 2025 · 1 repository · arXiv:2505.09039
-
Beyond the Known: Decision Making with Counterfactual Reasoning Decision Transformer 14 May 2025 · 1 repository · arXiv:2505.09114
-
BiECVC: Gated Diversification of Bidirectional Contexts for Learned Video Compression 14 May 2025 · 1 repository · arXiv:2505.09193
-
BLIP3-o: A Family of Fully Open Unified Multimodal Models-Architecture, Training and Dataset 14 May 2025 · 1 repository · arXiv:2505.09568Syntology official (archive's flag): 1 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
BrainNetMLP: An Efficient and Effective Baseline for Functional Brain Network Classification 14 May 2025 · 1 repository · arXiv:2505.11538
-
Customizing a Large Language Model for VHDL Design of High-Performance Microprocessors 14 May 2025 · 0 repositories · arXiv:2505.09610
-
CXMArena: Unified Dataset to benchmark performance in realistic CXM Scenarios 14 May 2025 · 1 repository · arXiv:2505.09436
-
Denoising and Alignment: Rethinking Domain Generalization for Multimodal Face Anti-Spoofing 14 May 2025 · 0 repositories · arXiv:2505.09484
-
Display Content, Display Methods and Evaluation Methods of the HCI in Explainable Recommender Systems: A Survey 14 May 2025 · 0 repositories · arXiv:2505.09065
-
Distance-aware Self-adaptive Graph Convolution for Fine-grained Hierarchical Recommendation 14 May 2025 · 1 repository · arXiv:2505.09590
-
Don't Forget your Inverse DDIM for Image Editing 14 May 2025 · 0 repositories · arXiv:2505.09571
-
FAS-LLM: Large Language Model-Based Channel Prediction for OTFS-Enabled Satellite-FAS Links 14 May 2025 · 0 repositories · arXiv:2505.09751
-
Few-Shot Learning of Visual Compositional Concepts through Probabilistic Schema Induction 14 May 2025 · 0 repositories · arXiv:2505.09859
-
HMamba: Hyperbolic Mamba for Sequential Recommendation 14 May 2025 · 0 repositories · arXiv:2505.09205
-
How Hungry is AI? Benchmarking Energy, Water, and Carbon Footprint of LLM Inference 14 May 2025 · 0 repositories · arXiv:2505.09598
-
Insights into DeepSeek-V3: Scaling Challenges and Reflections on Hardware for AI Architectures 14 May 2025 · 0 repositories · arXiv:2505.09343
-
LAS: Loss-less ANN-SNN Conversion for Fully Spike-Driven Large Language Models 14 May 2025 · 1 repository · arXiv:2505.09659
-
Llama See, Llama Do: A Mechanistic Perspective on Contextual Entrainment and Distraction in LLMs 14 May 2025 · 1 repository · arXiv:2505.09338
-
Mission Balance: Generating Under-represented Class Samples using Video Diffusion Models 14 May 2025 · 1 repository · arXiv:2505.09858
-
MoRAL: Motion-aware Multi-Frame 4D Radar and LiDAR Fusion for Robust 3D Object Detection 14 May 2025 · 0 repositories · arXiv:2505.09422
-
Multilingual Machine Translation with Quantum Encoder Decoder Attention-based Convolutional Variational Circuits 14 May 2025 · 0 repositories · arXiv:2505.09407
-
MUST: Multi-Scale Structural-Temporal Link Prediction Model for UAV Ad Hoc Networks 14 May 2025 · 1 repository · arXiv:2505.09331
-
Neural models for prediction of spatially patterned phase transitions: methods and challenges 14 May 2025 · 0 repositories · arXiv:2505.09718
-
Out-of-distribution generalisation is hard: evidence from ARC-like tasks 14 May 2025 · 0 repositories · arXiv:2505.09716
-
Q-space Guided Collaborative Attention Translation Network for Flexible Diffusion-Weighted Images Synthesis 14 May 2025 · 1 repository · arXiv:2505.09323
-
Quotient Complex Transformer (QCformer) for Perovskite Data Analysis 14 May 2025 · 0 repositories · arXiv:2505.09174
-
Sequence-Only Prediction of Binding Affinity Changes: A Robust and Interpretable Model for Antibody Engineering 14 May 2025 · 1 repository · arXiv:2505.20301
-
Sequential Treatment Effect Estimation with Unmeasured Confounders 14 May 2025 · 0 repositories · arXiv:2505.09113
-
Spec2VolCAMU-Net: A Spectrogram-to-Volume Model for EEG-to-fMRI Reconstruction based on Multi-directional Time-Frequency Convolutional Attention Encoder and Vision-Mamba U-Net 14 May 2025 · 1 repository · arXiv:2505.09521
-
Text-driven Motion Generation: Overview, Challenges and Directions 14 May 2025 · 0 repositories · arXiv:2505.09379
-
TopoDiT-3D: Topology-Aware Diffusion Transformer with Bottleneck Structure for 3D Point Cloud Generation 14 May 2025 · 1 repository · arXiv:2505.09140
-
WavReward: Spoken Dialogue Models With Generalist Reward Evaluators 14 May 2025 · 1 repository · arXiv:2505.09558
-
Zero-Shot Multi-modal Large Language Model v.s. Supervised Deep Learning: A Comparative Study on CT-Based Intracranial Hemorrhage Subtyping 14 May 2025 · 1 repository · arXiv:2505.09252
-
A Deep Learning-Driven Inhalation Injury Grading Assistant Using Bronchoscopy Images 13 May 2025 · 0 repositories · arXiv:2505.08517
-
A Head to Predict and a Head to Question: Pre-trained Uncertainty Quantification Heads for Hallucination Detection in LLM Outputs 13 May 2025 · 1 repository · arXiv:2505.08200Syntology 2 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; every one of the 2 samples that ran constructed an object rather than computing a result (of 3 harvested samples)
-
A suite of LMs comprehend puzzle statements as well as humans 13 May 2025 · 0 repositories · arXiv:2505.08996
-
AC-PKAN: Attention-Enhanced and Chebyshev Polynomial-Based Physics-Informed Kolmogorov-Arnold Networks 13 May 2025 · 0 repositories · arXiv:2505.08687
-
AC-Reason: Towards Theory-Guided Actual Causality Reasoning with Large Language Models 13 May 2025 · 1 repository · arXiv:2505.08750
-
Achieving Scalable Robot Autonomy via neurosymbolic planning using lightweight local LLM 13 May 2025 · 1 repository · arXiv:2505.08492
-
Are We Paying Attention to Her? Investigating Gender Disambiguation and Attention in Machine Translation 13 May 2025 · 1 repository · arXiv:2505.08546
-
CNN and ViT Efficiency Study on Tiny ImageNet and DermaMNIST Datasets 13 May 2025 · 0 repositories · arXiv:2505.08259
-
Constrained Edge AI Deployment: Fine-Tuning vs Distillation for LLM Compression 13 May 2025 · 0 repositories · arXiv:2505.18166
-
Controllable Image Colorization with Instance-aware Texts and Masks 13 May 2025 · 0 repositories · arXiv:2505.08705
-
DHECA-SuperGaze: Dual Head-Eye Cross-Attention and Super-Resolution for Unconstrained Gaze Estimation 13 May 2025 · 0 repositories · arXiv:2505.08426
-
Differentiable Channel Selection in Self-Attention For Person Re-Identification 13 May 2025 · 1 repository · arXiv:2505.08961
-
DyGSSM: Multi-view Dynamic Graph Embeddings with State Space Model Gradient Update 13 May 2025 · 0 repositories · arXiv:2505.09017
-
Enhancing Thyroid Cytology Diagnosis with RAG-Optimized LLMs and Pa-thology Foundation Models 13 May 2025 · 0 repositories · arXiv:2505.08590
-
Evaluating LLM Metrics Through Real-World Capabilities 13 May 2025 · 0 repositories · arXiv:2505.08253
-
Evaluating the Effectiveness of Black-Box Prompt Optimization as the Scale of LLMs Continues to Grow 13 May 2025 · 0 repositories · arXiv:2505.08303
-
EventDiff: A Unified and Efficient Diffusion Model Framework for Event-based Video Frame Interpolation 13 May 2025 · 0 repositories · arXiv:2505.08235
-
FlashMLA-ETAP: Efficient Transpose Attention Pipeline for Accelerating MLA Inference on NVIDIA H20 GPUs 13 May 2025 · 1 repository · arXiv:2506.01969
-
For GPT-4 as with Humans: Information Structure Predicts Acceptability of Long-Distance Dependencies 13 May 2025 · 0 repositories · arXiv:2505.09005
-
Generative Molecular Design with Steerable and Granular Synthesizability Control 13 May 2025 · 1 repository · arXiv:2505.08774
-
Granite-speech: open-source speech-aware LLMs with strong English ASR capabilities 13 May 2025 · 0 repositories · arXiv:2505.08699
-
Hakim: Farsi Text Embedding Model 13 May 2025 · 0 repositories · arXiv:2505.08435
-
HealthBench: Evaluating Large Language Models Towards Improved Human Health 13 May 2025 · 1 repository · arXiv:2505.08775Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
IterKey: Iterative Keyword Generation with LLMs for Enhanced Retrieval Augmented Generation 13 May 2025 · 0 repositories · arXiv:2505.08450
-
Judging the Judges: Can Large Vision-Language Models Fairly Evaluate Chart Comprehension and Reasoning? 13 May 2025 · 1 repository · arXiv:2505.08468