Methods › General › Attention Mechanisms › Attention › Papers, page 22
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 22 of 316: papers 2,101 to 2,200 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Mimic In-Context Learning for Multimodal Tasks 11 Apr 2025 · 1 repository · arXiv:2504.08851
-
MineWorld: a Real-Time and Open-Source Interactive World Model on Minecraft 11 Apr 2025 · 0 repositories · arXiv:2504.08388
-
MixDiT: Accelerating Image Diffusion Transformer Inference with Mixed-Precision MX Quantization 11 Apr 2025 · 0 repositories · arXiv:2504.08398
-
ModernBERT or DeBERTaV3? Examining Architecture and Data Influence on Transformer Encoder Models Performance 11 Apr 2025 · 0 repositories · arXiv:2504.08716
-
MotionDreamer: One-to-Many Motion Synthesis with Localized Generative Masked Transformer 11 Apr 2025 · 0 repositories · arXiv:2504.08959
-
Muon-Accelerated Attention Distillation for Real-Time Edge Synthesis via Optimized Latent Diffusion 11 Apr 2025 · 0 repositories · arXiv:2504.08451
-
Out of Style: RAG's Fragility to Linguistic Variation 11 Apr 2025 · 1 repository · arXiv:2504.08231Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 14 harvested samples)
-
PACT: Pruning and Clustering-Based Token Reduction for Faster Visual Language Models 11 Apr 2025 · 1 repository · arXiv:2504.08966Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
Passive Underwater Acoustic Signal Separation based on Feature Decoupling Dual-path Network 11 Apr 2025 · 0 repositories · arXiv:2504.08371
-
PCA-RAG: Principal Component Analysis for Efficient Retrieval-Augmented Generation 11 Apr 2025 · 0 repositories · arXiv:2504.08386
-
PNE-SGAN: Probabilistic NDT-Enhanced Semantic Graph Attention Network for LiDAR Loop Closure Detection 11 Apr 2025 · 0 repositories · arXiv:2504.08280
-
RTLRepoCoder: Repository-Level RTL Code Completion through the Combination of Fine-Tuning and Retrieval Augmentation 11 Apr 2025 · 0 repositories · arXiv:2504.08862
-
SARFormer -- An Acquisition Parameter Aware Vision Transformer for Synthetic Aperture Radar Data 11 Apr 2025 · 0 repositories · arXiv:2504.08441
-
seeBias: A Comprehensive Tool for Assessing and Visualizing AI Fairness 11 Apr 2025 · 1 repository · arXiv:2504.08418
-
Steering CLIP's vision transformer with sparse autoencoders 11 Apr 2025 · 0 repositories · arXiv:2504.08729
-
SWAN-GPT: An Efficient and Scalable Approach for Long-Context Language Modeling 11 Apr 2025 · 0 repositories · arXiv:2504.08719
-
The Other Side of the Coin: Exploring Fairness in Retrieval-Augmented Generation 11 Apr 2025 · 1 repository · arXiv:2504.12323
-
Training-free Guidance in Text-to-Video Generation via Multimodal Planning and Structured Noise Initialization 11 Apr 2025 · 0 repositories · arXiv:2504.08641
-
Transformer Learns Optimal Variable Selection in Group-Sparse Classification 11 Apr 2025 · 0 repositories · arXiv:2504.08638
-
VLMT: Vision-Language Multimodal Transformer for Multimodal Multi-hop Question Answering 11 Apr 2025 · 0 repositories · arXiv:2504.08269
-
ZipIR: Latent Pyramid Diffusion Transformer for High-Resolution Image Restoration 11 Apr 2025 · 0 repositories · arXiv:2504.08591
-
A System for Comprehensive Assessment of RAG Frameworks 10 Apr 2025 · 1 repository · arXiv:2504.07803
-
AgentAda: Skill-Adaptive Data Analytics for Tailored Insight Discovery 10 Apr 2025 · 1 repository · arXiv:2504.07421
-
AI Coding with Few-Shot Prompting for Thematic Analysis 10 Apr 2025 · 0 repositories · arXiv:2504.07408
-
AI-Slop to AI-Polish? Aligning Language Models through Edit-Based Writing Rewards and Test-time Computation 10 Apr 2025 · 1 repository · arXiv:2504.07532Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
APSQ: Additive Partial Sum Quantization with Algorithm-Hardware Co-Design 10 Apr 2025 · 0 repositories · arXiv:2505.03748
-
AttentionDefense: Leveraging System Prompt Attention for Explainable Defense Against Novel Jailbreaks 10 Apr 2025 · 0 repositories · arXiv:2504.12321
-
Beating Transformers using Synthetic Cognition 10 Apr 2025 · 0 repositories · arXiv:2504.07619
-
Beyond Feature Importance: Feature Interactions in Predicting Post-Stroke Rigidity with Graph Explainable AI 10 Apr 2025 · 0 repositories · arXiv:2504.08150
-
Beyond LLMs: A Linguistic Approach to Causal Graph Generation from Narrative Texts 10 Apr 2025 · 0 repositories · arXiv:2504.07459
-
Breaking the Barriers: Video Vision Transformers for Word-Level Sign Language Recognition 10 Apr 2025 · 0 repositories · arXiv:2504.07792
-
Can Reasoning LLMs Enhance Clinical Document Classification? 10 Apr 2025 · 1 repository · arXiv:2504.08040
-
ChronoFormer: Time-Aware Transformer Architectures for Structured Clinical Event Modeling 10 Apr 2025 · 0 repositories · arXiv:2504.07373
-
ClimateBench-M: A Multi-Modal Climate Data Benchmark with a Simple Generative Method 10 Apr 2025 · 1 repository · arXiv:2504.07394
-
ConceptFormer: Towards Efficient Use of Knowledge-Graph Embeddings in Large Language Models 10 Apr 2025 · 0 repositories · arXiv:2504.07624
-
ContrastiveGaussian: High-Fidelity 3D Generation with Contrastive Learning and Gaussian Splatting 10 Apr 2025 · 1 repository · arXiv:2504.08100
-
Deep Learning Meets Teleconnections: Improving S2S Predictions for European Winter Weather 10 Apr 2025 · 1 repository · arXiv:2504.07625
-
DGOcc: Depth-aware Global Query-based Network for Monocular 3D Occupancy Prediction 10 Apr 2025 · 0 repositories · arXiv:2504.07524
-
Distilling Knowledge from Heterogeneous Architectures for Semantic Segmentation 10 Apr 2025 · 0 repositories · arXiv:2504.07691
-
End-to-End Facial Expression Detection in Long Videos 10 Apr 2025 · 0 repositories · arXiv:2504.07660
-
LSR-MCTS: Alleviating Long Range Dependency in Code Generation 10 Apr 2025 · 0 repositories · arXiv:2504.07433
-
Generative Artificial Intelligence for Internet of Things Computing: A Systematic Survey 10 Apr 2025 · 0 repositories · arXiv:2504.07635
-
Genetic Programming with Reinforcement Learning Trained Transformer for Real-World Dynamic Scheduling Problems 10 Apr 2025 · 0 repositories · arXiv:2504.07779
-
Has the Creativity of Large-Language Models peaked? An analysis of inter- and intra-LLM variability 10 Apr 2025 · 0 repositories · arXiv:2504.12320
-
Heart Failure Prediction using Modal Decomposition and Masked Autoencoders for Scarce Echocardiography Databases 10 Apr 2025 · 1 repository · arXiv:2504.07606
-
HoloPart: Generative 3D Part Amodal Segmentation 10 Apr 2025 · 0 repositories · arXiv:2504.07943
-
How do Large Language Models Understand Relevance? A Mechanistic Interpretability Perspective 10 Apr 2025 · 1 repository · arXiv:2504.07898
-
JEPA4Rec: Learning Effective Language Representations for Sequential Recommendation via Joint Embedding Predictive Architecture 10 Apr 2025 · 0 repositories · arXiv:2504.10512
-
Learning Object Focused Attention 10 Apr 2025 · 0 repositories · arXiv:2504.08166
-
Malware analysis assisted by AI with R2AI 10 Apr 2025 · 0 repositories · arXiv:2504.07574
-
MRD-RAG: Enhancing Medical Diagnosis with Multi-Round Retrieval-Augmented Generation 10 Apr 2025 · 1 repository · arXiv:2504.07724
-
Novel Pooling-based VGG-Lite for Pneumonia and Covid-19 Detection from Imbalanced Chest X-Ray Datasets 10 Apr 2025 · 0 repositories · arXiv:2504.07468
-
On the Practice of Deep Hierarchical Ensemble Network for Ad Conversion Rate Prediction 10 Apr 2025 · 0 repositories · arXiv:2504.08169
-
P2Object: Single Point Supervised Object Detection and Instance Segmentation 10 Apr 2025 · 1 repository · arXiv:2504.07813
-
Pangu Ultra: Pushing the Limits of Dense Large Language Models on Ascend NPUs 10 Apr 2025 · 0 repositories · arXiv:2504.07866
-
PoGO: A Scalable Proof of Useful Work via Quantized Gradient Descent and Merkle Proofs 10 Apr 2025 · 0 repositories · arXiv:2504.07540
-
RadZero: Similarity-Based Cross-Attention for Explainable Vision-Language Alignment in Radiology with Zero-Shot Multi-Task Capability 10 Apr 2025 · 0 repositories · arXiv:2504.07416Syntology 5 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Representation Meets Optimization: Training PINNs and PIKANs for Gray-Box Discovery in Systems Pharmacology 10 Apr 2025 · 0 repositories · arXiv:2504.07379
-
Revisiting Prompt Optimization with Large Reasoning Models-A Case Study on Event Extraction 10 Apr 2025 · 0 repositories · arXiv:2504.07357
-
SRVP: Strong Recollection Video Prediction Model Using Attention-Based Spatiotemporal Correlation Fusion 10 Apr 2025 · 1 repository · arXiv:2504.08012
-
Synthetic Fluency: Hallucinations, Confabulations, and the Creation of Irish Words in LLM-Generated Translations 10 Apr 2025 · 0 repositories · arXiv:2504.07680
-
The Urban Impact of AI: Modeling Feedback Loops in Next-Venue Recommendation 10 Apr 2025 · 1 repository · arXiv:2504.07911
-
ThermoStereoRT: Thermal Stereo Matching in Real Time via Knowledge Distillation and Attention-based Refinement 10 Apr 2025 · 0 repositories · arXiv:2504.07418
-
V2V3D: View-to-View Denoised 3D Reconstruction for Light-Field Microscopy 10 Apr 2025 · 0 repositories · arXiv:2504.07853
-
A Deep Single Image Rectification Approach for Pan-Tilt-Zoom Cameras 9 Apr 2025 · 0 repositories · arXiv:2504.06965
-
A new training approach for text classification in Mental Health: LatentGLoss 9 Apr 2025 · 1 repository · arXiv:2504.07245
-
A Survey of New Mid-Band/FR3 for 6G: Channel Measurement, Characterization and Modeling in Outdoor Environment 9 Apr 2025 · 0 repositories · arXiv:2504.06727
-
A Unified Agentic Framework for Evaluating Conditional Image Generation 9 Apr 2025 · 1 repository · arXiv:2504.07046
-
AMAD: AutoMasked Attention for Unsupervised Multivariate Time Series Anomaly Detection 9 Apr 2025 · 0 repositories · arXiv:2504.06643
-
Attributes-aware Visual Emotion Representation Learning 9 Apr 2025 · 0 repositories · arXiv:2504.06578
-
Benchmarking Multimodal CoT Reward Model Stepwise by Visual Program 9 Apr 2025 · 1 repository · arXiv:2504.06606
-
ColorizeDiffusion v2: Enhancing Reference-based Sketch Colorization Through Separating Utilities 9 Apr 2025 · 2 repositories · arXiv:2504.06895
-
DiffusionCom: Structure-Aware Multimodal Diffusion Model for Multimodal Knowledge Graph Completion 9 Apr 2025 · 0 repositories · arXiv:2504.06543
-
DyDiT++: Dynamic Diffusion Transformers for Efficient Visual Generation 9 Apr 2025 · 1 repository · arXiv:2504.06803
-
Endowing Embodied Agents with Spatial Reasoning Capabilities for Vision-and-Language Navigation 9 Apr 2025 · 0 repositories · arXiv:2504.08806
-
Evaluating Retrieval Augmented Generative Models for Document Queries in Transportation Safety 9 Apr 2025 · 0 repositories · arXiv:2504.07022
-
Face-LLaVA: Facial Expression and Attribute Understanding through Instruction Tuning 9 Apr 2025 · 0 repositories · arXiv:2504.07198
-
FANeRV: Frequency Separation and Augmentation based Neural Representation for Video 9 Apr 2025 · 0 repositories · arXiv:2504.06755
-
GenDoP: Auto-regressive Camera Trajectory Generation as a Director of Photography 9 Apr 2025 · 0 repositories · arXiv:2504.07083
-
Kaleidoscope: In-language Exams for Massively Multilingual Vision Evaluation 9 Apr 2025 · 1 repository · arXiv:2504.07072
-
Linguistic Interpretability of Transformer-based Language Models: a systematic review 9 Apr 2025 · 1 repository · arXiv:2504.08001
-
OPAL: Encoding Causal Understanding of Physical Systems for Robot Learning 9 Apr 2025 · 0 repositories · arXiv:2504.06538
-
Perception in Reflection 9 Apr 2025 · 0 repositories · arXiv:2504.07165
-
Poly-Vector Retrieval: Reference and Content Embeddings for Legal Documents 9 Apr 2025 · 0 repositories · arXiv:2504.10508
-
HiFlow: Training-free High-Resolution Image Generation with Flow-Aligned Guidance 8 Apr 2025 · 1 repository · arXiv:2504.06232Syntology 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 1 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples)
-
Accelerating LLM Inference Throughput via Asynchronous KV Cache Prefetching 8 Apr 2025 · 0 repositories · arXiv:2504.06319
-
Assessing how hyperparameters impact Large Language Models' sarcasm detection performance 8 Apr 2025 · 0 repositories · arXiv:2504.06166
-
Control-Oriented Modelling and Adaptive Parameter Estimation for Hybrid Wind-Wave Energy Systems 8 Apr 2025 · 0 repositories · arXiv:2504.05948
-
Defending Deep Neural Networks against Backdoor Attacks via Module Switching 8 Apr 2025 · 0 repositories · arXiv:2504.05902
-
FASR-Net: Unsupervised Shadow Removal Leveraging Inherent Frequency Priors 8 Apr 2025 · 0 repositories · arXiv:2504.05779
-
Fusing Global and Local: Transformer-CNN Synergy for Next-Gen Current Estimation 8 Apr 2025 · 0 repositories · arXiv:2504.07996
-
Gaze-Guided Learning: Avoiding Shortcut Bias in Visual Classification 8 Apr 2025 · 1 repository · arXiv:2504.05583
-
GIGA: Generalizable Sparse Image-driven Gaussian Avatars 8 Apr 2025 · 0 repositories · arXiv:2504.07144
-
Graph-based Approaches and Functionalities in Retrieval-Augmented Generation: A Comprehensive Survey 8 Apr 2025 · 0 repositories · arXiv:2504.10499
-
Hogwild! Inference: Parallel LLM Generation via Concurrent Attention 8 Apr 2025 · 1 repository · arXiv:2504.06261Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Holistic Fusion: Task- and Setup-Agnostic Robot Localization and State Estimation with Factor Graphs 8 Apr 2025 · 1 repository · arXiv:2504.06479
-
HRMedSeg: Unlocking High-resolution Medical Image segmentation via Memory-efficient Attention Modeling 8 Apr 2025 · 1 repository · arXiv:2504.06205
-
Knowledge Graph Completion with Relation-Aware Anchor Enhancement 8 Apr 2025 · 1 repository · arXiv:2504.06129
-
Large Language Models Enhanced Hyperbolic Space Recommender Systems 8 Apr 2025 · 0 repositories · arXiv:2504.05694
-
Leveraging Auto-Distillation and Generative Self-Supervised Learning in Residual Graph Transformers for Enhanced Recommender Systems 8 Apr 2025 · 0 repositories · arXiv:2504.10500