Methods › General › Attention Mechanisms › Attention › Papers, page 54
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 54 of 316: papers 5,301 to 5,400 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
FocusDD: Real-World Scene Infusion for Robust Dataset Distillation 11 Jan 2025 · 0 repositories · arXiv:2501.06405
-
Ladder-residual: parallelism-aware architecture for accelerating large model inference with communication overlapping 11 Jan 2025 · 1 repository · arXiv:2501.06589Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 0 violated, 3 with no contract checked; 5 where Syntology's instrument failed) · 1 unverified (of 10 harvested samples) · 1 pointer-only (licence)
-
Multi-View Factorizing and Disentangling: A Novel Framework for Incomplete Multi-View Multi-Label Classification 11 Jan 2025 · 0 repositories · arXiv:2501.06524
-
Natural Language Supervision for Low-light Image Enhancement 11 Jan 2025 · 0 repositories · arXiv:2501.06546
-
Reliable Imputed-Sample Assisted Vertical Federated Learning 11 Jan 2025 · 0 repositories · arXiv:2501.06429
-
Savanna dynamics with grazing, browsing, and migration effects 11 Jan 2025 · 0 repositories · arXiv:2501.06549
-
Tensor Product Attention Is All You Need 11 Jan 2025 · 1 repository · arXiv:2501.06425Syntology official (archive's flag): 5 ran · 8 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 1 violated, 2 with no contract checked; 5 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples) · 2 pointer-only (licence)
-
TopoFormer: Integrating Transformers and ConvLSTMs for Coastal Topography Prediction 11 Jan 2025 · 0 repositories · arXiv:2501.06494
-
Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization 11 Jan 2025 · 0 repositories · arXiv:2501.06663
-
VASparse: Towards Efficient Visual Hallucination Mitigation for Large Vision-Language Model via Visual-Aware Sparsification 11 Jan 2025 · 1 repository · arXiv:2501.06553
-
YO-CSA-T: A Real-time Badminton Tracking System Utilizing YOLO Based on Contextual and Spatial Attention 11 Jan 2025 · 0 repositories · arXiv:2501.06472
-
A Holistically Point-guided Text Framework for Weakly-Supervised Camouflaged Object Detection 10 Jan 2025 · 0 repositories · arXiv:2501.06038
-
AI-powered virtual tissues from spatial proteomics for clinical diagnostics and biomedical discovery 10 Jan 2025 · 1 repository · arXiv:2501.06039
-
An Attention-Guided Deep Learning Approach for Classifying 39 Skin Lesion Types 10 Jan 2025 · 1 repository · arXiv:2501.05991
-
Analyzing Spatio-Temporal Dynamics of Dissolved Oxygen for the River Thames using Superstatistical Methods and Machine Learning 10 Jan 2025 · 0 repositories · arXiv:2501.07599
-
Binary Event-Driven Spiking Transformer 10 Jan 2025 · 0 repositories · arXiv:2501.05904
-
Bridging Dialects: Translating Standard Bangla to Regional Variants Using Neural Models 10 Jan 2025 · 0 repositories · arXiv:2501.05749
-
CognoSpeak: an automatic, remote assessment of early cognitive decline in real-world conversational speech 10 Jan 2025 · 0 repositories · arXiv:2501.05755
-
COLOR: A compositional linear operation-based representation of protein sequences for identification of monomer contributions to properties 10 Jan 2025 · 0 repositories · arXiv:2501.06371
-
Discovery of sustainable energy materials via the machine-learned material space 10 Jan 2025 · 0 repositories · arXiv:2501.05903
-
EDNet: Edge-Optimized Small Target Detection in UAV Imagery -- Faster Context Attention, Better Feature Fusion, and Hardware Acceleration 10 Jan 2025 · 1 repository · arXiv:2501.05885
-
eKalibr: Dynamic Intrinsic Calibration for Event Cameras From First Principles of Events 10 Jan 2025 · 1 repository · arXiv:2501.05688
-
Element-wise Attention Is All You Need 10 Jan 2025 · 0 repositories · arXiv:2501.05730
-
ELFATT: Efficient Linear Fast Attention for Vision Transformers 10 Jan 2025 · 0 repositories · arXiv:2501.06098
-
Enhancing, Refining, and Fusing: Towards Robust Multi-Scale and Dense Ship Detection 10 Jan 2025 · 0 repositories · arXiv:2501.06053
-
Enhancing Unsupervised Graph Few-shot Learning via Set Functions and Optimal Transport 10 Jan 2025 · 1 repository · arXiv:2501.05635Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples)
-
EXION: Exploiting Inter- and Intra-Iteration Output Sparsity for Diffusion Models 10 Jan 2025 · 0 repositories · arXiv:2501.05680
-
Explainable Federated Bayesian Causal Inference and Its Application in Advanced Manufacturing 10 Jan 2025 · 1 repository · arXiv:2501.06077
-
Fine-tuning is Not Fine: Mitigating Backdoor Attacks in GNNs with Limited Clean Data 10 Jan 2025 · 0 repositories · arXiv:2501.05835
-
Halal or Not: Knowledge Graph Completion for Predicting Cultural Appropriateness of Daily Products 10 Jan 2025 · 1 repository · arXiv:2501.05768
-
Iconicity in Large Language Models 10 Jan 2025 · 0 repositories · arXiv:2501.05643
-
Identity-aware Feature Decoupling Learning for Clothing-change Person Re-identification 10 Jan 2025 · 0 repositories · arXiv:2501.05851
-
Merging Feed-Forward Sublayers for Compressed Transformers 10 Jan 2025 · 1 repository · arXiv:2501.06126
-
Mix-QViT: Mixed-Precision Vision Transformer Quantization Driven by Layer Importance and Quantization Sensitivity 10 Jan 2025 · 0 repositories · arXiv:2501.06357
-
Model Inversion in Split Learning for Personalized LLMs: New Insights from Information Bottleneck Theory 10 Jan 2025 · 0 repositories · arXiv:2501.05965
-
MSCViT: A Small-size ViT architecture with Multi-Scale Self-Attention Mechanism for Tiny Datasets 10 Jan 2025 · 0 repositories · arXiv:2501.06040
-
Multi-subject Open-set Personalization in Video Generation 10 Jan 2025 · 0 repositories · arXiv:2501.06187
-
Aligning Brain Activity with Advanced Transformer Models: Exploring the Role of Punctuation in Semantic Processing 10 Jan 2025 · 1 repository · arXiv:2501.06278
-
Soft regression trees: a model variant and a decomposition training algorithm 10 Jan 2025 · 0 repositories · arXiv:2501.05942
-
Swin-X2S: Reconstructing 3D Shape from 2D Biplanar X-ray with Swin Transformers 10 Jan 2025 · 0 repositories · arXiv:2501.05961
-
TTS-Transducer: End-to-End Speech Synthesis with Neural Transducer 10 Jan 2025 · 0 repositories · arXiv:2501.06320
-
Understanding How Paper Writers Use AI-Generated Captions in Figure Caption Writing 10 Jan 2025 · 0 repositories · arXiv:2501.06317
-
Using Pre-trained LLMs for Multivariate Time Series Forecasting 10 Jan 2025 · 0 repositories · arXiv:2501.06386
-
VideoRAG: Retrieval-Augmented Generation over Video Corpus 10 Jan 2025 · 1 repository · arXiv:2501.05874Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Weakly Supervised Segmentation of Hyper-Reflective Foci with Compact Convolutional Transformers and SAM2 10 Jan 2025 · 0 repositories · arXiv:2501.05933
-
1-2-1: Renaissance of Single-Network Paradigm for Virtual Try-On 9 Jan 2025 · 0 repositories · arXiv:2501.05369
-
3DIS-FLUX: simple and efficient multi-instance generation with DiT rendering 9 Jan 2025 · 1 repository · arXiv:2501.05131
-
A General Retrieval-Augmented Generation Framework for Multimodal Case-Based Reasoning Applications 9 Jan 2025 · 0 repositories · arXiv:2501.05030
-
Addressing Domain Shift via Imbalance-Aware Domain Adaptation in Embryo Development Assessment 9 Jan 2025 · 1 repository · arXiv:2501.04958
-
Analyzing Memorization in Large Language Models through the Lens of Model Attribution 9 Jan 2025 · 1 repository · arXiv:2501.05078
-
Biomedical Relation Extraction via Adaptive Document-Relation Cross-Mapping and Concept Unique Identifier 9 Jan 2025 · 0 repositories · arXiv:2501.05155
-
BRATI: Bidirectional Recurrent Attention for Time-Series Imputation 9 Jan 2025 · 0 repositories · arXiv:2501.05401
-
Compression with Global Guidance: Towards Training-free High-Resolution MLLMs Acceleration 9 Jan 2025 · 1 repository · arXiv:2501.05179Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription 9 Jan 2025 · 1 repository · arXiv:2501.05068
-
De-centering the (Traditional) User: Multistakeholder Evaluation of Recommender Systems 9 Jan 2025 · 0 repositories · arXiv:2501.05170
-
DisSim-FinBERT: Text Simplification for Core Message Extraction in Complex Financial Texts 9 Jan 2025 · 0 repositories · arXiv:2501.04959
-
DPF^*: improved Depth Potential Function for scale-invariant sulcal depth estimation 9 Jan 2025 · 1 repository · arXiv:2501.05436
-
Enhancing Plagiarism Detection in Marathi with a Weighted Ensemble of TF-IDF and BERT Embeddings for Low-Resource Language Processing 9 Jan 2025 · 1 repository · arXiv:2501.05260
-
Entangled Mean Estimation in High-Dimensions 9 Jan 2025 · 0 repositories · arXiv:2501.05425
-
FairCoder: Evaluating Social Bias of LLMs in Code Generation 9 Jan 2025 · 2 repositories · arXiv:2501.05396
-
Hyperdimensional Computing for ADHD Classification using EEG Signals 9 Jan 2025 · 0 repositories · arXiv:2501.05186
-
Interpretable deep learning illuminates multiple structures fluorescence imaging: a path toward trustworthy artificial intelligence in microscopy 9 Jan 2025 · 0 repositories · arXiv:2501.05490
-
Large language models streamline automated systematic review: A preliminary study 9 Jan 2025 · 0 repositories · arXiv:2502.15702
-
LLMQuoter: Enhancing RAG Capabilities Through Efficient Quote Extraction From Large Contexts 9 Jan 2025 · 1 repository · arXiv:2501.05554
-
LongViTU: Instruction Tuning for Long-Form Video Understanding 9 Jan 2025 · 0 repositories · arXiv:2501.05037
-
MHAFF: Multi-Head Attention Feature Fusion of CNN and Transformer for Cattle Identification 9 Jan 2025 · 0 repositories · arXiv:2501.05209
-
OpenAI ChatGPT interprets Radiological Images: GPT-4 as a Medical Doctor for a Fast Check-Up 9 Jan 2025 · 0 repositories · arXiv:2501.06269
-
Optimizing Multitask Industrial Processes with Predictive Action Guidance 9 Jan 2025 · 0 repositories · arXiv:2501.05108
-
RAG-WM: An Efficient Black-Box Watermarking Approach for Retrieval-Augmented Generation of Large Language Models 9 Jan 2025 · 0 repositories · arXiv:2501.05249
-
ReFocus: Visual Editing as a Chain of Thought for Structured Image Understanding 9 Jan 2025 · 1 repository · arXiv:2501.05452
-
Rethinking Evaluation of Sparse Autoencoders through the Representation of Polysemous Words 9 Jan 2025 · 1 repository · arXiv:2501.06254
-
SpecTf: Transformers Enable Data-Driven Imaging Spectroscopy Cloud Detection 9 Jan 2025 · 1 repository · arXiv:2501.04916
-
The dynamics of meaning through time: Assessment of Large Language Models 9 Jan 2025 · 0 repositories · arXiv:2501.05552
-
The more polypersonal the better -- a short look on space geometry of fine-tuned layers 9 Jan 2025 · 0 repositories · arXiv:2501.05503
-
UAV-VLA: Vision-Language-Action System for Large Scale Aerial Mission Generation 9 Jan 2025 · 1 repository · arXiv:2501.05014
-
Zero-1-to-G: Taming Pretrained 2D Diffusion Model for Direct 3D Generation 9 Jan 2025 · 0 repositories · arXiv:2501.05427
-
A Plug-and-Play Bregman ADMM Module for Inferring Event Branches in Temporal Point Processes 8 Jan 2025 · 1 repository · arXiv:2501.04529
-
A Survey on Algorithmic Developments in Optimal Transport Problem with Applications 8 Jan 2025 · 0 repositories · arXiv:2501.06247
-
ActPC-Geom: Towards Scalable Online Neural-Symbolic Learning via Accelerating Active Predictive Coding with Information Geometry & Diverse Cognitive Mechanisms 8 Jan 2025 · 0 repositories · arXiv:2501.04832
-
Advancing Retrieval-Augmented Generation for Persian: Development of Language Models, Comprehensive Benchmarks, and Best Practices for Optimization 8 Jan 2025 · 0 repositories · arXiv:2501.04858
-
An Interpretable ML-based Model for Predicting p-y Curves of Monopile Foundations in Sand 8 Jan 2025 · 0 repositories · arXiv:2501.06232
-
Circuit Complexity Bounds for Visual Autoregressive Model 8 Jan 2025 · 0 repositories · arXiv:2501.04299
-
Discrete Wavelet Transform-Based Capsule Network for Hyperspectral Image Classification 8 Jan 2025 · 0 repositories · arXiv:2501.04643
-
GRAPHITE: Graph-Based Interpretable Tissue Examination for Enhanced Explainability in Breast Cancer Histopathology 8 Jan 2025 · 1 repository · arXiv:2501.04206
-
HyFusion: Enhanced Reception Field Transformer for Hyperspectral Image Fusion 8 Jan 2025 · 0 repositories · arXiv:2501.04665
-
Integrating LLMs with ITS: Recent Advances, Potentials, Challenges, and Future Directions 8 Jan 2025 · 0 repositories · arXiv:2501.04437
-
Knowledge Retrieval Based on Generative AI 8 Jan 2025 · 0 repositories · arXiv:2501.04635
-
Learnable Scaled Gradient Descent for Guaranteed Robust Tensor PCA 8 Jan 2025 · 0 repositories · arXiv:2501.04565
-
LipGen: Viseme-Guided Lip Video Generation for Enhancing Visual Speech Recognition 8 Jan 2025 · 0 repositories · arXiv:2501.04204
-
Machine Learning and statistical classification of CRISPR-Cas12a diagnostic assays 8 Jan 2025 · 0 repositories · arXiv:2501.04413
-
Mapping the Edge of Chaos: Fractal-Like Boundaries in The Trainability of Decoder-Only Transformer Models 8 Jan 2025 · 1 repository · arXiv:2501.04286
-
MB-TaylorFormer V2: Improved Multi-branch Linear Transformer Expanded by Taylor Formula for Image Restoration 8 Jan 2025 · 2 repositories · arXiv:2501.04486
-
Mechanics and Design of Metastructured Auxetic Patches with Bio-inspired Materials 8 Jan 2025 · 0 repositories · arXiv:2501.06233
-
Multi-task retriever fine-tuning for domain-specific and efficient RAG 8 Jan 2025 · 0 repositories · arXiv:2501.04652
-
Online Gaussian Test-Time Adaptation of Vision-Language Models 8 Jan 2025 · 1 repository · arXiv:2501.04352Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample)
-
Planing It by Ear: Convolutional Neural Networks for Acoustic Anomaly Detection in Industrial Wood Planers 8 Jan 2025 · 1 repository · arXiv:2501.04819
-
Quantum-inspired Embeddings Projection and Similarity Metrics for Representation Learning 8 Jan 2025 · 1 repository · arXiv:2501.04591
-
Re-ranking the Context for Multimodal Retrieval Augmented Generation 8 Jan 2025 · 0 repositories · arXiv:2501.04695
-
Rethinking domain generalization in medical image segmentation: One image as one domain 8 Jan 2025 · 0 repositories · arXiv:2501.04741
-
Scaling Large Language Model Training on Frontier with Low-Bandwidth Partitioning 8 Jan 2025 · 0 repositories · arXiv:2501.04266