Methods › General › Attention Mechanisms › Attention › Papers, page 25
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 25 of 316: papers 2,401 to 2,500 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Deconver: A Deconvolutional Network for Medical Image Segmentation 1 Apr 2025 · 1 repository · arXiv:2504.00302
-
Detecting Financial Fraud with Hybrid Deep Learning: A Mix-of-Experts Approach to Sequential and Anomalous Patterns 1 Apr 2025 · 0 repositories · arXiv:2504.03750
-
FUSION: Frequency-guided Underwater Spatial Image recOnstructioN 1 Apr 2025 · 0 repositories · arXiv:2504.01243
-
Grade Guard: A Smart System for Short Answer Automated Grading 1 Apr 2025 · 0 repositories · arXiv:2504.01253
-
GRU-AUNet: A Domain Adaptation Framework for Contactless Fingerprint Presentation Attack Detection 1 Apr 2025 · 0 repositories · arXiv:2504.01213
-
GS_DravidianLangTech@2025: Women Targeted Abusive Texts Detection on Social Media 1 Apr 2025 · 0 repositories · arXiv:2504.02863
-
Hierarchical Attention Networks for Lossless Point Cloud Attribute Compression 1 Apr 2025 · 0 repositories · arXiv:2504.00481
-
How does Watermarking Affect Visual Language Models in Document Understanding? 1 Apr 2025 · 0 repositories · arXiv:2504.01048
-
Improved Visual-Spatial Reasoning via R1-Zero-Like Training 1 Apr 2025 · 1 repository · arXiv:2504.00883
-
Inaccuracy of an E-Dictionary and Its Influence on Chinese Language Users 1 Apr 2025 · 0 repositories · arXiv:2504.00799
-
Investigating the Capabilities and Limitations of Machine Learning for Identifying Bias in English Language Data with Information and Heritage Professionals 1 Apr 2025 · 1 repository · arXiv:2504.00860
-
Learned Image Compression with Dictionary-based Entropy Model 1 Apr 2025 · 1 repository · arXiv:2504.00496
-
LLM-Assisted Proactive Threat Intelligence for Automated Reasoning 1 Apr 2025 · 0 repositories · arXiv:2504.00428
-
MergeVQ: A Unified Framework for Visual Generation and Representation with Disentangled Token Merging and Quantization 1 Apr 2025 · 1 repository · arXiv:2504.00999
-
Misaligned Roles, Misplaced Images: Structural Input Perturbations Expose Multimodal Alignment Blind Spots 1 Apr 2025 · 0 repositories · arXiv:2504.03735
-
Multi-Token Attention 1 Apr 2025 · 0 repositories · arXiv:2504.00927
-
NeuRadar: Neural Radiance Fields for Automotive Radar Point Clouds 1 Apr 2025 · 0 repositories · arXiv:2504.00859
-
On the Robustness of Agentic Function Calling 1 Apr 2025 · 0 repositories · arXiv:2504.00914
-
ParallelFlow: Parallelizing Linear Transformers via Flow Discretization 1 Apr 2025 · 0 repositories · arXiv:2504.00492
-
Physics-informed machine learning for building performance simulation-A review of a nascent field 1 Apr 2025 · 0 repositories · arXiv:2504.00937
-
QSViT: A Methodology for Quantizing Spiking Vision Transformers 1 Apr 2025 · 0 repositories · arXiv:2504.00948
-
Repetitions are not all alike: distinct mechanisms sustain repetition in language models 1 Apr 2025 · 0 repositories · arXiv:2504.01100
-
Role and Use of Race in AI/ML Models Related to Health 1 Apr 2025 · 0 repositories · arXiv:2504.00899
-
SRLCG: Self-Rectified Large-Scale Code Generation with Multidimensional Chain-of-Thought and Dynamic Backtracking 1 Apr 2025 · 0 repositories · arXiv:2504.00532
-
SViQA: A Unified Speech-Vision Multimodal Model for Textless Visual Question Answering 1 Apr 2025 · 0 repositories · arXiv:2504.01049
-
Synthesized Annotation Guidelines are Knowledge-Lite Boosters for Clinical Information Extraction 1 Apr 2025 · 0 repositories · arXiv:2504.02871
-
Two-stage deep learning framework for the restoration of incomplete-ring PET images 1 Apr 2025 · 0 repositories · arXiv:2504.00816
-
Transformer-Based Named Entity Recognition for Automated Server Provisioning 1 Apr 2025 · 1 repository
-
WikiVideo: Article Generation from Multiple Videos 1 Apr 2025 · 1 repository · arXiv:2504.00939
-
A Comparative Study of Scanpath Models in Graph-Based Visualization 31 Mar 2025 · 0 repositories · arXiv:2503.24160
-
A Plasticity-Aware Method for Continual Self-Supervised Learning in Remote Sensing 31 Mar 2025 · 0 repositories · arXiv:2503.24088
-
A Systematic Evaluation of LLM Strategies for Mental Health Text Analysis: Fine-tuning vs. Prompt Engineering vs. RAG 31 Mar 2025 · 0 repositories · arXiv:2503.24307
-
Accelerating High-Efficiency Organic Photovoltaic Discovery via Pretrained Graph Neural Networks and Generative Reinforcement Learning 31 Mar 2025 · 0 repositories · arXiv:2503.23766
-
Adaptive Layer-skipping in Pre-trained LLMs 31 Mar 2025 · 0 repositories · arXiv:2503.23798
-
AI-Enhanced Resilience in Power Systems: Adversarial Deep Learning for Robust Short-Term Voltage Stability Assessment under Cyber-Attacks 31 Mar 2025 · 0 repositories · arXiv:2504.02859
-
AirCache: Activating Inter-modal Relevancy KV Cache Compression for Efficient Large Vision-Language Model Inference 31 Mar 2025 · 0 repositories · arXiv:2503.23956
-
An extension of linear self-attention for in-context learning 31 Mar 2025 · 0 repositories · arXiv:2503.23814
-
Dynamic Parametric Retrieval Augmented Generation for Test-time Knowledge Enhancement 31 Mar 2025 · 1 repository · arXiv:2503.23895Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
CITRAS: Covariate-Informed Transformer for Time Series Forecasting 31 Mar 2025 · 0 repositories · arXiv:2503.24007
-
Coarse-to-Fine Learning for Multi-Pipette Localisation in Robot-Assisted In Vivo Patch-Clamp 31 Mar 2025 · 0 repositories · arXiv:2504.01044
-
CoMatch: Dynamic Covisibility-Aware Transformer for Bilateral Subpixel-Level Semi-Dense Image Matching 31 Mar 2025 · 0 repositories · arXiv:2503.23925Syntology 2 ran (of which 1 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples) · 1 pointer-only (licence)
-
Comparing representations of long clinical texts for the task of patient note-identification 31 Mar 2025 · 0 repositories · arXiv:2503.24006
-
Conformal uncertainty quantification to evaluate predictive fairness of foundation AI model for skin lesion classes across patient demographics 31 Mar 2025 · 0 repositories · arXiv:2503.23819
-
Context-Independent OCR with Multimodal LLMs: Effects of Image Resolution and Visual Complexity 31 Mar 2025 · 0 repositories · arXiv:2503.23667
-
CrossFormer: Cross-Segment Semantic Fusion for Document Segmentation 31 Mar 2025 · 0 repositories · arXiv:2503.23671
-
Crossing Boundaries: Leveraging Semantic Divergences to Explore Cultural Novelty in Cooking Recipes 31 Mar 2025 · 1 repository · arXiv:2503.24027
-
Does "Reasoning" with Large Language Models Improve Recognizing, Generating, and Reframing Unhelpful Thoughts? 31 Mar 2025 · 0 repositories · arXiv:2504.00163
-
Easi3R: Estimating Disentangled Motion from DUSt3R Without Training 31 Mar 2025 · 1 repository · arXiv:2503.24391
-
Enhancing Large Language Models (LLMs) for Telecommunications using Knowledge Graphs and Retrieval-Augmented Generation 31 Mar 2025 · 0 repositories · arXiv:2503.24245
-
Enhancing Time Series Forecasting with Fuzzy Attention-Integrated Transformers 31 Mar 2025 · 1 repository · arXiv:2504.00070
-
Exploring In-Context Learning Capabilities of ChatGPT for Pathological Speech Detection 31 Mar 2025 · 0 repositories · arXiv:2503.23873
-
FineCausal: A Causal-Based Framework for Interpretable Fine-Grained Action Quality Assessment 31 Mar 2025 · 1 repository · arXiv:2503.23911
-
Foundation Models For Seismic Data Processing: An Extensive Review 31 Mar 2025 · 1 repository · arXiv:2503.24166
-
Frequency-Aware Attention-LSTM for PM_(2.5) Time Series Forecasting 31 Mar 2025 · 0 repositories · arXiv:2503.24043
-
From Colors to Classes: Emergence of Concepts in Vision Transformers 31 Mar 2025 · 1 repository · arXiv:2503.24071
-
GAL-MAD: Towards Explainable Anomaly Detection in Microservice Applications Using Graph Attention Networks 31 Mar 2025 · 0 repositories · arXiv:2504.00058
-
Graph Transformer-Based Flood Susceptibility Mapping: Application to the French Riviera and Railway Infrastructure Under Climate Change 31 Mar 2025 · 0 repositories · arXiv:2504.03727
-
Integrating Quantum-Classical Attention in Patch Transformers for Enhanced Time Series Forecasting 31 Mar 2025 · 1 repository · arXiv:2504.00068
-
JudgeLRM: Large Reasoning Models as a Judge 31 Mar 2025 · 0 repositories · arXiv:2504.00050
-
Large Language Models Pass the Turing Test 31 Mar 2025 · 0 repositories · arXiv:2503.23674
-
Learning a Single Index Model from Anisotropic Data with vanilla Stochastic Gradient Descent 31 Mar 2025 · 0 repositories · arXiv:2503.23642
-
LLM4FS: Leveraging Large Language Models for Feature Selection and How to Improve It 31 Mar 2025 · 0 repositories · arXiv:2503.24157
-
Model Hemorrhage and the Robustness Limits of Large Language Models 31 Mar 2025 · 0 repositories · arXiv:2503.23924
-
NeuRaLaTeX: A machine learning library written in pure LaTeX 31 Mar 2025 · 0 repositories · arXiv:2503.24187
-
On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices 31 Mar 2025 · 1 repository · arXiv:2503.23796
-
Order Matters: On Parameter-Efficient Image-to-Video Probing for Recognizing Nearly Symmetric Actions 31 Mar 2025 · 0 repositories · arXiv:2503.24298
-
Performance analysis of metasurface-based spatial multimode transmission for 6G wireless communications 31 Mar 2025 · 0 repositories · arXiv:2504.00137
-
PIM-LLM: A High-Throughput Hybrid PIM Architecture for 1-bit LLMs 31 Mar 2025 · 0 repositories · arXiv:2504.01994
-
PupiNet: Seamless OCT-OCTA Interconversion Through Wavelet-Driven and Multi-Scale Attention Mechanisms 31 Mar 2025 · 0 repositories · arXiv:2503.23933
-
Rubric Is All You Need: Enhancing LLM-based Code Evaluation With Question-Specific Rubrics 31 Mar 2025 · 0 repositories · arXiv:2503.23989
-
SQuat: Subspace-orthogonal KV Cache Quantization 31 Mar 2025 · 0 repositories · arXiv:2503.24358
-
Synthetic News Generation for Fake News Classification 31 Mar 2025 · 0 repositories · arXiv:2503.24206
-
Text Chunking for Document Classification for Urban System Management using Large Language Models 31 Mar 2025 · 1 repository · arXiv:2504.00274
-
Texture or Semantics? Vision-Language Models Get Lost in Font Recognition 31 Mar 2025 · 1 repository · arXiv:2503.23768
-
TransMamba: Flexibly Switching between Transformer and Mamba 31 Mar 2025 · 0 repositories · arXiv:2503.24067
-
UltraRAG: A Modular and Automated Toolkit for Adaptive Retrieval-Augmented Generation 31 Mar 2025 · 1 repository · arXiv:2504.08761
-
What the F*ck Is Artificial General Intelligence? 31 Mar 2025 · 0 repositories · arXiv:2503.23923
-
A Lightweight Image Super-Resolution Transformer Trained on Low-Resolution Images Only 30 Mar 2025 · 1 repository · arXiv:2503.23265
-
A Survey on Unlearnable Data 30 Mar 2025 · 1 repository · arXiv:2503.23536
-
Advancing Sentiment Analysis in Tamil-English Code-Mixed Texts: Challenges and Transformer-Based Solutions 30 Mar 2025 · 0 repositories · arXiv:2503.23295
-
Beyond Academic Benchmarks: Critical Analysis and Best Practices for Visual Industrial Anomaly Detection 30 Mar 2025 · 1 repository · arXiv:2503.23451
-
Beyond Detection: Designing AI-Resilient Assessments with Automated Feedback Tool to Foster Critical Thinking 30 Mar 2025 · 0 repositories · arXiv:2503.23622
-
CADFormer: Fine-Grained Cross-modal Alignment and Decoding Transformer for Referring Remote Sensing Image Segmentation 30 Mar 2025 · 0 repositories · arXiv:2503.23456
-
DiT4SR: Taming Diffusion Transformer for Real-World Image Super-Resolution 30 Mar 2025 · 0 repositories · arXiv:2503.23580
-
Efficient Dynamic Attention 3D Convolution for Hyperspectral Image Classification 30 Mar 2025 · 1 repository · arXiv:2503.23472
-
Efficient Token Compression for Vision Transformer with Spatial Information Preserved 30 Mar 2025 · 1 repository · arXiv:2503.23455Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 0 honoured, 0 violated, 13 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 14 harvested samples) · 14 pointer-only (licence)
-
Embedding Shift Dissection on CLIP: Effects of Augmentations on VLM's Representation Learning 30 Mar 2025 · 0 repositories · arXiv:2503.23495
-
Enhancing 3D Gaussian Splatting Compression via Spatial Condition-based Prediction 30 Mar 2025 · 0 repositories · arXiv:2503.23337
-
Exploring GPT-4 for Robotic Agent Strategy with Real-Time State Feedback and a Reactive Behaviour Framework 30 Mar 2025 · 0 repositories · arXiv:2503.23601
-
FeRG-LLM : Feature Engineering by Reason Generation Large Language Models 30 Mar 2025 · 0 repositories · arXiv:2503.23371
-
Focus Directions Make Your Language Models Pay More Attention to Relevant Contexts 30 Mar 2025 · 0 repositories · arXiv:2503.23306
-
GMapLatent: Geometric Mapping in Latent Space 30 Mar 2025 · 0 repositories · arXiv:2503.23407
-
Google and China's Trade 30 Mar 2025 · 0 repositories · arXiv:2503.23557
-
HiPART: Hierarchical Pose AutoRegressive Transformer for Occluded 3D Human Pose Estimation 30 Mar 2025 · 0 repositories · arXiv:2503.23331
-
Hyper-RAG: Combating LLM Hallucinations using Hypergraph-Driven Retrieval-Augmented Generation 30 Mar 2025 · 0 repositories · arXiv:2504.08758
-
Improved Ear Verification with Vision Transformers and Overlapping Patches 30 Mar 2025 · 0 repositories · arXiv:2503.23275
-
Improving underwater semantic segmentation with underwater image quality attention and muti-scale aggregation attention 30 Mar 2025 · 1 repository · arXiv:2503.23422
-
JavisDiT: Joint Audio-Video Diffusion Transformer with Hierarchical Spatio-Temporal Prior Synchronization 30 Mar 2025 · 0 repositories · arXiv:2503.23377
-
KernelDNA: Dynamic Kernel Sharing via Decoupled Naive Adapters 30 Mar 2025 · 1 repository · arXiv:2503.23379
-
Large Language Models Are Better Logical Fallacy Reasoners with Counterargument, Explanation, and Goal-Aware Prompt Formulation 30 Mar 2025 · 1 repository · arXiv:2503.23363