Methods › General › Attention Mechanisms › Attention › Papers, page 68
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 68 of 316: papers 6,701 to 6,800 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Pinco: Position-induced Consistent Adapter for Diffusion Transformer in Foreground-conditioned Inpainting 5 Dec 2024 · 0 repositories · arXiv:2412.03812
-
Reconstruction of boosted and resolved multi-Higgs-boson events with symmetry-preserving attention networks 5 Dec 2024 · 0 repositories · arXiv:2412.03819
-
SwiftEdit: Lightning Fast Text-Guided Image Editing via One-Step Diffusion 5 Dec 2024 · 0 repositories · arXiv:2412.04301
-
Text Change Detection in Multilingual Documents Using Image Comparison 5 Dec 2024 · 0 repositories · arXiv:2412.04137
-
Towards Real-Time Open-Vocabulary Video Instance Segmentation 5 Dec 2024 · 0 repositories · arXiv:2412.04434
-
TransAdapter: Vision Transformer for Feature-Centric Unsupervised Domain Adaptation 5 Dec 2024 · 1 repository · arXiv:2412.04073
-
Understanding the Excess Bond Premium 5 Dec 2024 · 0 repositories · arXiv:2412.04063
-
Uniform Discretized Integrated Gradients: An effective attribution based method for explaining large language models 5 Dec 2024 · 0 repositories · arXiv:2412.03886
-
A new Time-decay Radiomics Integrated Network (TRINet) for short-term breast cancer risk prediction 4 Dec 2024 · 0 repositories · arXiv:2412.03081
-
A Stitch in Time Saves Nine: Small VLM is a Precise Guidance for Accelerating Large VLMs 4 Dec 2024 · 1 repository · arXiv:2412.03324Syntology official: harvested, nothing ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 2 honoured, 2 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
A Water Efficiency Dataset for African Data Centers 4 Dec 2024 · 0 repositories · arXiv:2412.03716
-
Advanced Risk Prediction and Stability Assessment of Banks Using Time Series Transformer Models 4 Dec 2024 · 0 repositories · arXiv:2412.03606
-
Advancing Conversational Psychotherapy: Integrating Privacy, Dual-Memory, and Domain Expertise with Large Language Models 4 Dec 2024 · 0 repositories · arXiv:2412.02987
-
AntLM: Bridging Causal and Masked Language Models 4 Dec 2024 · 0 repositories · arXiv:2412.03275
-
ASIGN: An Anatomy-aware Spatial Imputation Graphic Network for 3D Spatial Transcriptomics 4 Dec 2024 · 1 repository · arXiv:2412.03026
-
Benchmarking Attention Mechanisms and Consistency Regularization Semi-Supervised Learning for Post-Flood Building Damage Assessment in Satellite Images 4 Dec 2024 · 0 repositories · arXiv:2412.03015
-
Beyond algorithm hyperparameters: on preprocessing hyperparameters and associated pitfalls in machine learning applications 4 Dec 2024 · 1 repository · arXiv:2412.03491
-
Beyond [cls]: Exploring the true potential of Masked Image Modeling representations 4 Dec 2024 · 1 repository · arXiv:2412.03215
-
Continual Low-Rank Scaled Dot-product Attention 4 Dec 2024 · 0 repositories · arXiv:2412.03214
-
Controlling the Mutation in Large Language Models for the Efficient Evolution of Algorithms 4 Dec 2024 · 0 repositories · arXiv:2412.03250
-
DIVE: Taming DINO for Subject-Driven Video Editing 4 Dec 2024 · 0 repositories · arXiv:2412.03347
-
Does Safety Training of LLMs Generalize to Semantically Related Natural Prompts? 4 Dec 2024 · 0 repositories · arXiv:2412.03235
-
Dynamic Consistent k-Center Clustering with Optimal Recourse 4 Dec 2024 · 0 repositories · arXiv:2412.03238
-
EMPATH: MediaPipe-Aided Ensemble Learning with Attention-Based Transformers for Accurate Recognition of Bangla Word-Level Sign Language 4 Dec 2024 · 1 repository
-
Fab-ME: A Vision State-Space and Attention-Enhanced Framework for Fabric Defect Detection 4 Dec 2024 · 0 repositories · arXiv:2412.03200
-
FANAL -- Financial Activity News Alerting Language Modeling Framework 4 Dec 2024 · 0 repositories · arXiv:2412.03527
-
FLAIR: VLM with Fine-grained Language-informed Image Representations 4 Dec 2024 · 2 repositories · arXiv:2412.03561Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 1 violated, 10 with no contract checked; 2 where Syntology's instrument failed) · 7 unverified (of 20 harvested samples) · 20 pointer-only (licence)
-
GraPix: Exploring Graph Modularity Optimization for Unsupervised Pixel Clustering 4 Dec 2024 · 1 repository
-
Higher Order Transformers: Efficient Attention Mechanism for Tensor Structured Data 4 Dec 2024 · 0 repositories · arXiv:2412.02919
-
HIIF: Hierarchical Encoding based Implicit Image Function for Continuous Super-resolution 4 Dec 2024 · 0 repositories · arXiv:2412.03748
-
Interpretable Hierarchical Attention Network for Medical Condition Identification 4 Dec 2024 · 0 repositories · arXiv:2412.03701
-
Interpreting Transformers for Jet Tagging 4 Dec 2024 · 1 repository · arXiv:2412.03673Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Is Foreground Prototype Sufficient? Few-Shot Medical Image Segmentation with Background-Fused Prototype 4 Dec 2024 · 0 repositories · arXiv:2412.02983
-
MTVNet: Mapping using Transformers for Volumes -- Network for Super-Resolution with Long-Range Interactions 4 Dec 2024 · 1 repository · arXiv:2412.03379
-
MaterialPicker: Multi-Modal Material Generation with Diffusion Transformers 4 Dec 2024 · 0 repositories · arXiv:2412.03225
-
MIDI: Multi-Instance Diffusion for Single Image to 3D Scene Generation 4 Dec 2024 · 0 repositories · arXiv:2412.03558
-
MRNet: Multifaceted Resilient Networks for Medical Image-to-Image Translation 4 Dec 2024 · 0 repositories · arXiv:2412.03039
-
Multi-Branch Mutual-Distillation Transformer for EEG-Based Seizure Subtype Classification 4 Dec 2024 · 0 repositories · arXiv:2412.15224
-
Multi-view Image Diffusion via Coordinate Noise and Fourier Attention 4 Dec 2024 · 0 repositories · arXiv:2412.03756
-
Multimodal Sentiment Analysis Based on BERT and ResNet 4 Dec 2024 · 0 repositories · arXiv:2412.03625
-
MV-Adapter: Multi-view Consistent Image Generation Made Easy 4 Dec 2024 · 1 repository · arXiv:2412.03632Syntology 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 1 pointer-only (licence)
-
Navigation World Models 4 Dec 2024 · 1 repository · arXiv:2412.03572Syntology 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Online Experimental Design With Estimation-Regret Trade-off Under Network Interference 4 Dec 2024 · 0 repositories · arXiv:2412.03727
-
Online Physics-Informed Dynamic Mode Decomposition: Theory and Applications 4 Dec 2024 · 1 repository · arXiv:2412.03609
-
ParetoFlow: Guided Flows in Multi-Objective Optimization 4 Dec 2024 · 0 repositories · arXiv:2412.03718
-
PEMF-VVTO: Point-Enhanced Video Virtual Try-on via Mask-free Paradigm 4 Dec 2024 · 0 repositories · arXiv:2412.03021
-
Point-GR: Graph Residual Point Cloud Network for 3D Object Classification and Segmentation 4 Dec 2024 · 0 repositories · arXiv:2412.03052
-
Seeing Beyond Views: Multi-View Driving Scene Video Generation with Holistic Attention 4 Dec 2024 · 0 repositories · arXiv:2412.03520
-
STDCformer: A Transformer-Based Model with a Spatial-Temporal Causal De-Confounding Strategy for Crowd Flow Prediction 4 Dec 2024 · 0 repositories · arXiv:2412.02942
-
Style3D: Attention-guided Multi-view Style Transfer for 3D Object Generation 4 Dec 2024 · 0 repositories · arXiv:2412.03571
-
Theoretical limitations of multi-layer Transformer 4 Dec 2024 · 1 repository · arXiv:2412.02975
-
Unifying KV Cache Compression for Large Language Models with LeanKV 4 Dec 2024 · 0 repositories · arXiv:2412.03131
-
3D Face Reconstruction From Radar Images 3 Dec 2024 · 0 repositories · arXiv:2412.02403
-
A Comprehensive Evaluation of Large Language Models on Aspect-Based Sentiment Analysis 3 Dec 2024 · 0 repositories · arXiv:2412.02279
-
Achieving Semantic Consistency: Contextualized Word Representations for Political Text Analysis 3 Dec 2024 · 0 repositories · arXiv:2412.04505
-
Active Negative Loss: A Robust Framework for Learning with Noisy Labels 3 Dec 2024 · 1 repository · arXiv:2412.02373
-
CAISSON: Concept-Augmented Inference Suite of Self-Organizing Neural Networks 3 Dec 2024 · 0 repositories · arXiv:2412.02835
-
Cascaded Multi-Scale Attention for Enhanced Multi-Scale Feature Extraction and Interaction with Low-Resolution Images 3 Dec 2024 · 1 repository · arXiv:2412.02197
-
Compressing KV Cache for Long-Context LLM Inference with Inter-Layer Attention Similarity 3 Dec 2024 · 0 repositories · arXiv:2412.02252
-
Controlled Spectral Uplifting for Indirect-Light-Metamerism 3 Dec 2024 · 1 repository
-
CPTQuant -- A Novel Mixed Precision Post-Training Quantization Techniques for Large Language Models 3 Dec 2024 · 0 repositories · arXiv:2412.03599
-
CubeFormer: A Simple yet Effective Baseline for Lightweight Image Super-Resolution 3 Dec 2024 · 0 repositories · arXiv:2412.02234
-
Direct Coloring for Self-Supervised Enhanced Feature Decoupling 3 Dec 2024 · 0 repositories · arXiv:2412.02109
-
DP-2Stage: Adapting Language Models as Differentially Private Tabular Data Generators 3 Dec 2024 · 1 repository · arXiv:2412.02467
-
ESA: Example Sieve Approach for Multi-Positive and Unlabeled Learning 3 Dec 2024 · 0 repositories · arXiv:2412.02240
-
FCL-ViT: Task-Aware Attention Tuning for Continual Learning 3 Dec 2024 · 0 repositories · arXiv:2412.02509
-
Flattering to Deceive: The Impact of Sycophantic Behavior on User Trust in Large Language Model 3 Dec 2024 · 0 repositories · arXiv:2412.02802
-
GQWformer: A Quantum-based Transformer for Graph Representation Learning 3 Dec 2024 · 0 repositories · arXiv:2412.02285
-
Gracefully Filtering Backdoor Samples for Generative Large Language Models without Retraining 3 Dec 2024 · 1 repository · arXiv:2412.02454
-
Graph-Powered Defense: Controller Area Network Intrusion Detection for Unmanned Aerial Vehicles 3 Dec 2024 · 0 repositories · arXiv:2412.02539
-
HumanRig: Learning Automatic Rigging for Humanoid Character in a Large Scale Dataset 3 Dec 2024 · 1 repository · arXiv:2412.02317
-
Impact of Data Snooping on Deep Learning Models for Locating Vulnerabilities in Lifted Code 3 Dec 2024 · 0 repositories · arXiv:2412.02048
-
AI-driven Inverse Design of Band-Tunable Mechanical Metastructures for Tailored Vibration Mitigation 3 Dec 2024 · 0 repositories · arXiv:2412.12122
-
Investigating the importance of social vulnerability in opioid-related mortality across the United States 3 Dec 2024 · 0 repositories · arXiv:2412.15218
-
Learning Koopman-based Stability Certificates for Unknown Nonlinear Systems 3 Dec 2024 · 1 repository · arXiv:2412.02807
-
Leveraging Large Language Models for Comparative Literature Summarization with Reflective Incremental Mechanisms 3 Dec 2024 · 0 repositories · arXiv:2412.02149
-
MAGMA: Manifold Regularization for MAEs 3 Dec 2024 · 1 repository · arXiv:2412.02871
-
MetaShadow: Object-Centered Shadow Detection, Removal, and Synthesis 3 Dec 2024 · 0 repositories · arXiv:2412.02635
-
Multi-scale and Multi-path Cascaded Convolutional Network for Semantic Segmentation of Colorectal Polyps 3 Dec 2024 · 0 repositories · arXiv:2412.02443
-
OCR Hinders RAG: Evaluating the Cascading Impact of OCR on Retrieval-Augmented Generation 3 Dec 2024 · 1 repository · arXiv:2412.02592
-
OmniCreator: Self-Supervised Unified Generation with Universal Editing 3 Dec 2024 · 0 repositories · arXiv:2412.02114
-
Optimization of Transformer heart disease prediction model based on particle swarm optimization algorithm 3 Dec 2024 · 0 repositories · arXiv:2412.02801
-
Patent-CR: A Dataset for Patent Claim Revision 3 Dec 2024 · 0 repositories · arXiv:2412.02549
-
RARE: Retrieval-Augmented Reasoning Enhancement for Large Language Models 3 Dec 2024 · 1 repository · arXiv:2412.02830
-
Revisiting the Initial Steps in Adaptive Gradient Descent Optimization 3 Dec 2024 · 0 repositories · arXiv:2412.02153
-
Scaling BERT Models for Turkish Automatic Punctuation and Capitalization Correction 3 Dec 2024 · 0 repositories · arXiv:2412.02698
-
Semantic Tokens in Retrieval Augmented Generation 3 Dec 2024 · 0 repositories · arXiv:2412.02563
-
ShadowHack: Hacking Shadows via Luminance-Color Divide and Conquer 3 Dec 2024 · 1 repository · arXiv:2412.02545
-
SNOOPI: Supercharged One-step Diffusion Distillation with Proper Guidance 3 Dec 2024 · 0 repositories · arXiv:2412.02687
-
The Asymptotic Behavior of Attention in Transformers 3 Dec 2024 · 0 repositories · arXiv:2412.02682
-
Towards Rich Emotions in 3D Avatars: A Text-to-3D Avatar Generation Benchmark 3 Dec 2024 · 1 repository · arXiv:2412.02508
-
Towards the efficacy of federated prediction for epidemics on networks 3 Dec 2024 · 1 repository · arXiv:2412.02161
-
Transformer-Based Auxiliary Loss for Face Recognition Across Age Variations 3 Dec 2024 · 0 repositories · arXiv:2412.02198
-
Trust & Safety of LLMs and LLMs in Trust & Safety 3 Dec 2024 · 0 repositories · arXiv:2412.02113
-
Can't Slow me Down: Learning Robust and Hardware-Adaptive Object Detectors against Latency Attacks for Edge Devices 3 Dec 2024 · 0 repositories · arXiv:2412.02171
-
UniForm: A Reuse Attention Mechanism Optimized for Efficient Vision Transformers on Edge Devices 3 Dec 2024 · 0 repositories · arXiv:2412.02344
-
Viewpoint Consistency in 3D Generation via Attention and CLIP Guidance 3 Dec 2024 · 0 repositories · arXiv:2412.02287
-
VR Based Emotion Recognition Using Deep Multimodal Fusion With Biosignals Across Multiple Anatomical Domains 3 Dec 2024 · 0 repositories · arXiv:2412.02283
-
A Semantic Communication System for Real-time 3D Reconstruction Tasks 2 Dec 2024 · 0 repositories · arXiv:2412.01191
-
Advancing Speech Language Models by Scaling Supervised Fine-Tuning with Over 60,000 Hours of Synthetic Speech Dialogue Data 2 Dec 2024 · 1 repository · arXiv:2412.01078