Methods › General › Attention Mechanisms › Attention › Papers, page 6
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 6 of 316: papers 501 to 600 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Absolute Coordinates Make Motion Generation Easy 26 May 2025 · 0 repositories · arXiv:2505.19377
-
Accelerating Flow-Matching-Based Text-to-Speech via Empirically Pruned Step Sampling 26 May 2025 · 0 repositories · arXiv:2505.19931
-
Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing 26 May 2025 · 0 repositories · arXiv:2505.19578
-
AdaTP: Attention-Debiased Token Pruning for Video Large Language Models 26 May 2025 · 0 repositories · arXiv:2505.20100
-
Aggregated Structural Representation with Large Language Models for Human-Centric Layout Generation 26 May 2025 · 0 repositories · arXiv:2505.19554
-
Align and Surpass Human Camouflaged Perception: Visual Refocus Reinforcement Fine-Tuning 26 May 2025 · 1 repository · arXiv:2505.19611
-
AmpleHate: Amplifying the Attention for Versatile Implicit Hate Detection 26 May 2025 · 1 repository · arXiv:2505.19528
-
AMQA: An Adversarial Dataset for Benchmarking Bias of LLMs in Medicine and Healthcare 26 May 2025 · 1 repository · arXiv:2505.19562
-
Anveshana: A New Benchmark Dataset for Cross-Lingual Information Retrieval On English Queries and Sanskrit Documents 26 May 2025 · 0 repositories · arXiv:2505.19494
-
Automated evaluation of children's speech fluency for low-resource languages 26 May 2025 · 0 repositories · arXiv:2505.19671
-
Balancing Computation Load and Representation Expressivity in Parallel Hybrid Neural Networks 26 May 2025 · 0 repositories · arXiv:2505.19472
-
Benchmarking Multimodal Knowledge Conflict for Large Multimodal Models 26 May 2025 · 1 repository · arXiv:2505.19509
-
Beyond Simple Concatenation: Fairly Assessing PLM Architectures for Multi-Chain Protein-Protein Interactions Prediction 26 May 2025 · 0 repositories · arXiv:2505.20036
-
Beyond Specialization: Benchmarking LLMs for Transliteration of Indian Languages 26 May 2025 · 0 repositories · arXiv:2505.19851
-
Burst Image Super-Resolution via Multi-Cross Attention Encoding and Multi-Scan State-Space Decoding 26 May 2025 · 0 repositories · arXiv:2505.19668
-
CA3D: Convolutional-Attentional 3D Nets for Efficient Video Activity Recognition on the Edge 26 May 2025 · 0 repositories · arXiv:2505.19928
-
Calibrating Pre-trained Language Classifiers on LLM-generated Noisy Labels via Iterative Refinement 26 May 2025 · 1 repository · arXiv:2505.19675
-
CardioPatternFormer: Pattern-Guided Attention for Interpretable ECG Classification with Transformer Architecture 26 May 2025 · 0 repositories · arXiv:2505.20481
-
Compliance-to-Code: Enhancing Financial Compliance Checking via Code Generation 26 May 2025 · 1 repository · arXiv:2505.19804
-
Conversational Lexicography: Querying Lexicographic Data on Knowledge Graphs with SPARQL through Natural Language 26 May 2025 · 0 repositories · arXiv:2505.19971
-
Dependency Parsing is More Parameter-Efficient with Normalization 26 May 2025 · 0 repositories · arXiv:2505.20215
-
Detection of Suicidal Risk on Social Media: A Hybrid Model 26 May 2025 · 0 repositories · arXiv:2505.23797
-
DGRAG: Distributed Graph-based Retrieval-Augmented Generation in Edge-Cloud Systems 26 May 2025 · 0 repositories · arXiv:2505.19847
-
DoctorRAG: Medical RAG Fusing Knowledge with Patient Analogy through Textual Gradients 26 May 2025 · 0 repositories · arXiv:2505.19538
-
DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation 26 May 2025 · 0 repositories · arXiv:2505.19774
-
Electrolyzers-HSI: Close-Range Multi-Scene Hyperspectral Imaging Benchmark Dataset 26 May 2025 · 0 repositories · arXiv:2505.20507
-
Emotion Classification In-Context in Spanish 26 May 2025 · 0 repositories · arXiv:2505.20571
-
Equivariant Representation Learning for Symmetry-Aware Inference with Guarantees 26 May 2025 · 0 repositories · arXiv:2505.19809
-
ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining 26 May 2025 · 0 repositories · arXiv:2505.19893
-
FlowCut: Rethinking Redundancy via Information Flow for Efficient Vision-Language Models 26 May 2025 · 1 repository · arXiv:2505.19536Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
GoLF-NRT: Integrating Global Context and Local Geometry for Few-Shot View Synthesis 26 May 2025 · 1 repository · arXiv:2505.19813Syntology official (archive's flag): 15 ran · 15 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 1 honoured, 0 violated, 8 with no contract checked; 6 where Syntology's instrument failed) · 9 unverified (of 24 harvested samples) · 24 pointer-only (licence)
-
Grokking ExPLAIND: Unifying Model, Data, and Training Attribution to Study Model Behavior 26 May 2025 · 1 repository · arXiv:2505.20076Syntology official (archive's flag): 3 ran · 3 ran (of which 3 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified; every one of the 3 samples that ran constructed an object rather than computing a result (of 5 harvested samples)
-
Hierarchical Masked Autoregressive Models with Low-Resolution Token Pivots 26 May 2025 · 1 repository · arXiv:2505.20288
-
Hierarchical Tree Search-based User Lifelong Behavior Modeling on Large Language Model 26 May 2025 · 0 repositories · arXiv:2505.19505
-
How Syntax Specialization Emerges in Language Models 26 May 2025 · 0 repositories · arXiv:2505.19548
-
Improvement Strategies for Few-Shot Learning in OCT Image Classification of Rare Retinal Diseases 26 May 2025 · 0 repositories · arXiv:2505.20149
-
Improving Speech Emotion Recognition Through Cross Modal Attention Alignment and Balanced Stacking Model 26 May 2025 · 1 repository · arXiv:2505.20007
-
In-Context Brush: Zero-shot Customized Subject Insertion with Context-Aware Latent Space Manipulation 26 May 2025 · 0 repositories · arXiv:2505.20271
-
Inference-time Alignment in Continuous Space 26 May 2025 · 1 repository · arXiv:2505.20081
-
KnowTrace: Bootstrapping Iterative Retrieval-Augmented Generation with Structured Knowledge Tracing 26 May 2025 · 1 repository · arXiv:2505.20245
-
Large Language Models' Reasoning Stalls: An Investigation into the Capabilities of Frontier Models 26 May 2025 · 0 repositories · arXiv:2505.19676
-
LeCoDe: A Benchmark Dataset for Interactive Legal Consultation Dialogue Evaluation 26 May 2025 · 0 repositories · arXiv:2505.19667
-
LlamaSeg: Image Segmentation via Autoregressive Mask Generation 26 May 2025 · 0 repositories · arXiv:2505.19422
-
Long-Context State-Space Video World Models 26 May 2025 · 0 repositories · arXiv:2505.20171
-
Lung Nodule Segmentation: Exploring Data Efficiency and Advanced Architectures 26 May 2025 · 0 repositories
-
MA-RAG: Multi-Agent Retrieval-Augmented Generation via Collaborative Chain-of-Thought Reasoning 26 May 2025 · 0 repositories · arXiv:2505.20096
-
Memory-Efficient Visual Autoregressive Modeling with Scale-Aware KV Cache Compression 26 May 2025 · 1 repository · arXiv:2505.19602Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
MetaSTNet: Multimodal Meta-learning for Cellular Traffic Conformal Prediction 26 May 2025 · 0 repositories · arXiv:2505.21553
-
Minimalist Softmax Attention Provably Learns Constrained Boolean Functions 26 May 2025 · 0 repositories · arXiv:2505.19531
-
Multi-modal brain encoding models for multi-modal stimuli 26 May 2025 · 1 repository · arXiv:2505.20027Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 3 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning 26 May 2025 · 0 repositories · arXiv:2505.19938
-
Multimodal Machine Translation with Visual Scene Graph Pruning 26 May 2025 · 0 repositories · arXiv:2505.19507
-
NeuSym-RAG: Hybrid Neural Symbolic Retrieval with Multiview Structuring for PDF Question Answering 26 May 2025 · 1 repository · arXiv:2505.19754
-
One Surrogate to Fool Them All: Universal, Transferable, and Targeted Adversarial Attacks with CLIP 26 May 2025 · 1 repository · arXiv:2505.19840
-
Pangu Light: Weight Re-Initialization for Pruning and Accelerating LLMs 26 May 2025 · 0 repositories · arXiv:2505.20155
-
PHI: Bridging Domain Shift in Long-Term Action Quality Assessment via Progressive Hierarchical Instruction 26 May 2025 · 1 repository · arXiv:2505.19972Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples)
-
Private MEV Protection RPCs: Benchmark Stud 26 May 2025 · 0 repositories · arXiv:2505.19708
-
REARANK: Reasoning Re-ranking Agent via Reinforcement Learning 26 May 2025 · 1 repository · arXiv:2505.20046Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 2 where Syntology's instrument failed) · 6 unverified (of 18 harvested samples) · 4 pointer-only (licence)
-
ReasonPlan: Unified Scene Prediction and Decision Reasoning for Closed-loop Autonomous Driving 26 May 2025 · 1 repository · arXiv:2505.20024
-
Regularized Personalization of Text-to-Image Diffusion Models without Distributional Drift 26 May 2025 · 0 repositories · arXiv:2505.19519
-
Research on feature fusion and multimodal patent text based on graph attention network 26 May 2025 · 0 repositories · arXiv:2505.20188
-
Rethinking Text-based Protein Understanding: Retrieval or LLM? 26 May 2025 · 1 repository · arXiv:2505.20354Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
Rotation-Equivariant Self-Supervised Method in Image Denoising 26 May 2025 · 1 repository · arXiv:2505.19618Syntology official (archive's flag): 8 ran · 8 ran (of which 5 constructed an object rather than computing a result; 6 with no instrument failure: 1 honoured, 0 violated, 5 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
SESaMo: Symmetry-Enforcing Stochastic Modulation for Normalizing Flows 26 May 2025 · 0 repositories · arXiv:2505.19619
-
Small Language Models: Architectures, Techniques, Evaluation, Problems and Future Adaptation 26 May 2025 · 0 repositories · arXiv:2505.19529
-
Sparse2DGS: Sparse-View Surface Reconstruction using 2D Gaussian Splatting with Dense Point Cloud 26 May 2025 · 0 repositories · arXiv:2505.19854
-
Structured Initialization for Vision Transformers 26 May 2025 · 0 repositories · arXiv:2505.19985
-
syftr: Pareto-Optimal Generative AI 26 May 2025 · 1 repository · arXiv:2505.20266
-
Synthetic Time Series Forecasting with Transformer Architectures: Extensive Simulation Benchmarks 26 May 2025 · 1 repository · arXiv:2505.20048
-
Tensorization is a powerful but underexplored tool for compression and interpretability of neural networks 26 May 2025 · 0 repositories · arXiv:2505.20132
-
The Avengers: A Simple Recipe for Uniting Smaller Language Models to Challenge Proprietary Giants 26 May 2025 · 1 repository · arXiv:2505.19797
-
The Missing Point in Vision Transformers for Universal Image Segmentation 26 May 2025 · 1 repository · arXiv:2505.19795
-
Training LLM-Based Agents with Synthetic Self-Reflected Trajectories and Partial Masking 26 May 2025 · 0 repositories · arXiv:2505.20023
-
Transformers in Protein: A Survey 26 May 2025 · 0 repositories · arXiv:2505.20098
-
Translation-Equivariance of Normalization Layers and Aliasing in Convolutional Neural Networks 26 May 2025 · 1 repository · arXiv:2505.19805
-
Uncertainty-Aware Attention Heads: Efficient Unsupervised Uncertainty Quantification for LLMs 26 May 2025 · 0 repositories · arXiv:2505.20045
-
Understanding Transformer from the Perspective of Associative Memory 26 May 2025 · 0 repositories · arXiv:2505.19488
-
Underwater Diffusion Attention Network with Contrastive Language-Image Joint Learning for Underwater Image Enhancement 26 May 2025 · 0 repositories · arXiv:2505.19895
-
VADER: A Human-Evaluated Benchmark for Vulnerability Assessment, Detection, Explanation, and Remediation 26 May 2025 · 1 repository · arXiv:2505.19395
-
VisCRA: A Visual Chain Reasoning Attack for Jailbreaking Multimodal Large Language Models 26 May 2025 · 0 repositories · arXiv:2505.19684Syntology 9 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 12 harvested samples) · 2 pointer-only (licence)
-
VSCBench: Bridging the Gap in Vision-Language Model Safety Calibration 26 May 2025 · 1 repository · arXiv:2505.20362
-
WeatherEdit: Controllable Weather Editing with 4D Gaussian Field 26 May 2025 · 1 repository · arXiv:2505.20471
-
A Smart Healthcare System for Monkeypox Skin Lesion Detection and Tracking 25 May 2025 · 0 repositories · arXiv:2505.19023
-
ADGSyn: Dual-Stream Learning for Efficient Anticancer Drug Synergy Prediction 25 May 2025 · 1 repository · arXiv:2505.19144
-
AI4Math: A Native Spanish Benchmark for University-Level Mathematical Reasoning in Large Language Models 25 May 2025 · 0 repositories · arXiv:2505.18978
-
ASPO: Adaptive Sentence-Level Preference Optimization for Fine-Grained Multimodal Reasoning 25 May 2025 · 0 repositories · arXiv:2505.19100
-
Assistant-Guided Mitigation of Teacher Preference Bias in LLM-as-a-Judge 25 May 2025 · 1 repository · arXiv:2505.19176
-
Bayesian sparse modeling for interpretable prediction of hydroxide ion conductivity in anion-conductive polymer membranes 25 May 2025 · 0 repositories · arXiv:2505.19044
-
Benchmarking Large Language Models for Cyberbullying Detection in Real-World YouTube Comments 25 May 2025 · 0 repositories · arXiv:2505.18927
-
CDPDNet: Integrating Text Guidance with Hybrid Vision Encoders for Medical Image Segmentation 25 May 2025 · 1 repository · arXiv:2505.18958
-
Communication-Efficient Multi-Device Inference Acceleration for Transformer Models 25 May 2025 · 1 repository · arXiv:2505.19342
-
Conventional Contrastive Learning Often Falls Short: Improving Dense Retrieval with Cross-Encoder Listwise Distillation and Synthetic Data 25 May 2025 · 1 repository · arXiv:2505.19274
-
CreatiDesign: A Unified Multi-Conditional Diffusion Transformer for Creative Graphic Design 25 May 2025 · 0 repositories · arXiv:2505.19114
-
Curvature Dynamic Black-box Attack: revisiting adversarial robustness via dynamic curvature estimation 25 May 2025 · 0 repositories · arXiv:2505.19194
-
Disentangled Human Body Representation Based on Unsupervised Semantic-Aware Learning 25 May 2025 · 0 repositories · arXiv:2505.19049
-
DLF: Enhancing Explicit-Implicit Interaction via Dynamic Low-Order-Aware Fusion for CTR Prediction 25 May 2025 · 1 repository · arXiv:2505.19182
-
DREAM: Drafting with Refined Target Features and Entropy-Adaptive Cross-Attention Fusion for Multimodal Speculative Decoding 25 May 2025 · 1 repository · arXiv:2505.19201
-
DriveX: Omni Scene Modeling for Learning Generalizable World Knowledge in Autonomous Driving 25 May 2025 · 0 repositories · arXiv:2505.19239
-
EventEgoHands: Event-based Egocentric 3D Hand Mesh Reconstruction 25 May 2025 · 0 repositories · arXiv:2505.19169
-
Exploring Magnitude Preservation and Rotation Modulation in Diffusion Transformers 25 May 2025 · 0 repositories · arXiv:2505.19122