Methods › General › Attention Mechanisms › Attention › Papers, page 61
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 61 of 316: papers 6,001 to 6,100 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Fourier Position Embedding: Enhancing Attention's Periodic Extension for Length Generalization 23 Dec 2024 · 1 repository · arXiv:2412.17739
-
Free-viewpoint Human Animation with Pose-correlated Reference Selection 23 Dec 2024 · 0 repositories · arXiv:2412.17290
-
Friends-MMC: A Dataset for Multi-modal Multi-party Conversation Understanding 23 Dec 2024 · 1 repository · arXiv:2412.17295
-
Guided Real Image Dehazing using YCbCr Color Space 23 Dec 2024 · 1 repository · arXiv:2412.17496
-
HPCNeuroNet: A Neuromorphic Approach Merging SNN Temporal Dynamics with Transformer Attention for FPGA-based Particle Physics 23 Dec 2024 · 0 repositories · arXiv:2412.17571
-
LASE: Learned Adjacency Spectral Embeddings 23 Dec 2024 · 1 repository · arXiv:2412.17734
-
LayerDropBack: A Universally Applicable Approach for Accelerating Training of Deep Networks 23 Dec 2024 · 1 repository · arXiv:2412.18027
-
Learning Dynamic Local Context Representations for Infrared Small Target Detection 23 Dec 2024 · 0 repositories · arXiv:2412.17401
-
MRANet: A Modified Residual Attention Networks for Lung and Colon Cancer Classification 23 Dec 2024 · 0 repositories · arXiv:2412.17700
-
Multi-view Fuzzy Graph Attention Networks for Enhanced Graph Learning 23 Dec 2024 · 0 repositories · arXiv:2412.17271
-
Multimodal Preference Data Synthetic Alignment with Reward Model 23 Dec 2024 · 1 repository · arXiv:2412.17417
-
Personalized Large Vision-Language Models 23 Dec 2024 · 0 repositories · arXiv:2412.17610
-
Predicting Satisfied User and Machine Ratio for Compressed Images: A Unified Approach 23 Dec 2024 · 0 repositories · arXiv:2412.17477
-
QTSeg: A Query Token-Based Architecture for Efficient 2D Medical Image Segmentation 23 Dec 2024 · 1 repository · arXiv:2412.17241
-
Revisiting Multimodal Fusion for 3D Anomaly Detection from an Architectural Perspective 23 Dec 2024 · 1 repository · arXiv:2412.17297
-
STAHGNet: Modeling Hybrid-grained Heterogenous Dependency Efficiently for Traffic Prediction 23 Dec 2024 · 0 repositories · arXiv:2412.17524
-
STeInFormer: Spatial-Temporal Interaction Transformer Architecture for Remote Sensing Change Detection 23 Dec 2024 · 1 repository · arXiv:2412.17247
-
Theoretical Constraints on the Expressive Power of RoPE-based Tensor Attention Transformers 23 Dec 2024 · 0 repositories · arXiv:2412.18040
-
Token Statistics Transformer: Linear-Time Attention via Variational Rate Reduction 23 Dec 2024 · 1 repository · arXiv:2412.17810
-
Towards Unsupervised Model Selection for Domain Adaptive Object Detection 23 Dec 2024 · 1 repository · arXiv:2412.17284Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Uncertainty-Participation Context Consistency Learning for Semi-supervised Semantic Segmentation 23 Dec 2024 · 1 repository · arXiv:2412.17331
-
URoadNet: Dual Sparse Attentive U-Net for Multiscale Road Network Extraction 23 Dec 2024 · 0 repositories · arXiv:2412.17573
-
xPatch: Dual-Stream Time Series Forecasting with Exponential Seasonal-Trend Decomposition 23 Dec 2024 · 1 repository · arXiv:2412.17323
-
A Parameter-Efficient Quantum Anomaly Detection Method on a Superconducting Quantum Processor 22 Dec 2024 · 2 repositories · arXiv:2412.16867
-
A Reality Check on Context Utilisation for Retrieval-Augmented Generation 22 Dec 2024 · 1 repository · arXiv:2412.17031
-
Adapting Image-to-Video Diffusion Models for Large-Motion Frame Interpolation 22 Dec 2024 · 0 repositories · arXiv:2412.17042
-
An OpenMind for 3D medical vision self-supervised learning 22 Dec 2024 · 1 repository · arXiv:2412.17041
-
Bridging Auditory Perception and Language Comprehension through MEG-Driven Encoding Models 22 Dec 2024 · 0 repositories · arXiv:2501.03246
-
CoF: Coarse to Fine-Grained Image Understanding for Multi-modal Large Language Models 22 Dec 2024 · 1 repository · arXiv:2412.16869
-
Differentially Private Random Block Coordinate Descent 22 Dec 2024 · 0 repositories · arXiv:2412.17054
-
DR-Encoder: Encode Low-rank Gradients with Random Prior for Large Language Models Differentially Privately 22 Dec 2024 · 0 repositories · arXiv:2412.17053
-
Enhancing Supply Chain Transparency in Emerging Economies Using Online Contents and LLMs 22 Dec 2024 · 0 repositories · arXiv:2412.16922
-
FADA: Fast Diffusion Avatar Synthesis with Mixed-Supervised Multi-CFG Distillation 22 Dec 2024 · 0 repositories · arXiv:2412.16915
-
Interactive Classification Metrics: A graphical application to build robust intuition for classification model evaluation 22 Dec 2024 · 1 repository · arXiv:2412.17066
-
Multifaceted User Modeling in Recommendation: A Federated Foundation Models Approach 22 Dec 2024 · 1 repository · arXiv:2412.16969
-
NumbOD: A Spatial-Frequency Fusion Attack Against Object Detectors 22 Dec 2024 · 1 repository · arXiv:2412.16955Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
On Fusing ChatGPT and Ensemble Learning in Discon-tinuous Named Entity Recognition in Health Corpora 22 Dec 2024 · 0 repositories · arXiv:2412.16976
-
PsychAdapter: Adapting LLM Transformers to Reflect Traits, Personality and Mental Health 22 Dec 2024 · 1 repository · arXiv:2412.16882
-
Reconsidering SMT Over NMT for Closely Related Languages: A Case Study of Persian-Hindi Pair 22 Dec 2024 · 0 repositories · arXiv:2412.16877
-
Reversed Attention: On The Gradient Descent Of Attention Layers In GPT 22 Dec 2024 · 1 repository · arXiv:2412.17019
-
Robustness of Large Language Models Against Adversarial Attacks 22 Dec 2024 · 0 repositories · arXiv:2412.17011
-
Semantic Hierarchical Prompt Tuning for Parameter-Efficient Fine-Tuning 22 Dec 2024 · 1 repository · arXiv:2412.16956
-
SubstationAI: Multimodal Large Model-Based Approaches for Analyzing Substation Equipment Faults 22 Dec 2024 · 0 repositories · arXiv:2412.17077
-
Survey on Abstractive Text Summarization: Dataset, Models, and Metrics 22 Dec 2024 · 2 repositories · arXiv:2412.17165
-
TAR3D: Creating High-Quality 3D Assets via Next-Part Prediction 22 Dec 2024 · 0 repositories · arXiv:2412.16919
-
AlzheimerRAG: Multimodal Retrieval Augmented Generation for PubMed articles 21 Dec 2024 · 0 repositories · arXiv:2412.16701
-
Anchor Learning with Potential Cluster Constraints for Multi-view Clustering 21 Dec 2024 · 1 repository · arXiv:2412.16519
-
Assessing Social Alignment: Do Personality-Prompted Large Language Models Behave Like Humans? 21 Dec 2024 · 0 repositories · arXiv:2412.16772
-
Attention Entropy is a Key Factor: An Analysis of Parallel Context Encoding with Full-attention-based Pre-trained Language Models 21 Dec 2024 · 0 repositories · arXiv:2412.16545
-
Distilling Large Language Models for Efficient Clinical Information Extraction 21 Dec 2024 · 0 repositories · arXiv:2501.00031
-
Effective and Efficient Representation Learning for Flight Trajectories 21 Dec 2024 · 1 repository · arXiv:2412.16581
-
Effective Context Modeling Framework for Emotion Recognition in Conversations 21 Dec 2024 · 0 repositories · arXiv:2412.16444
-
Enhancing Contrastive Learning Inspired by the Philosophy of "The Blind Men and the Elephant" 21 Dec 2024 · 1 repository · arXiv:2412.16522
-
Enhancing Nighttime Vehicle Detection with Day-to-Night Style Transfer and Labeling-Free Augmentation 21 Dec 2024 · 0 repositories · arXiv:2412.16478
-
Evaluating the Performance of Large Language Models in Scientific Claim Detection and Classification 21 Dec 2024 · 0 repositories · arXiv:2412.16486
-
FAP-CD: Fairness-Driven Age-Friendly Community Planning via Conditional Diffusion Generation 21 Dec 2024 · 1 repository · arXiv:2412.16699
-
Flash3D: Super-scaling Point Transformers through Joint Hardware-Geometry Locality 21 Dec 2024 · 1 repository · arXiv:2412.16481
-
Follow-Your-MultiPose: Tuning-Free Multi-Character Text-to-Video Generation via Pose Guidance 21 Dec 2024 · 0 repositories · arXiv:2412.16495
-
Formal Language Knowledge Corpus for Retrieval Augmented Generation 21 Dec 2024 · 0 repositories · arXiv:2412.16689
-
From Histopathology Images to Cell Clouds: Learning Slide Representations with Hierarchical Cell Transformer 21 Dec 2024 · 0 repositories · arXiv:2412.16715
-
Identifying Cyberbullying Roles in Social Media 21 Dec 2024 · 0 repositories · arXiv:2412.16417
-
Improving FIM Code Completions via Context & Curriculum Based Learning 21 Dec 2024 · 0 repositories · arXiv:2412.16589
-
Iterative Feature Exclusion Ranking for Deep Tabular Learning 21 Dec 2024 · 1 repository · arXiv:2412.16442
-
KKANs: Kurkova-Kolmogorov-Arnold Networks and Their Learning Dynamics 21 Dec 2024 · 0 repositories · arXiv:2412.16738
-
Lillama: Large Language Models Compression via Low-Rank Feature Distillation 21 Dec 2024 · 0 repositories · arXiv:2412.16719
-
Object Detection Approaches to Identifying Hand Images with High Forensic Values 21 Dec 2024 · 0 repositories · arXiv:2412.16431
-
Paraformer: Parameterization of Sub-grid Scale Processes Using Transformers 21 Dec 2024 · 0 repositories · arXiv:2412.16763
-
Quantum-Like Contextuality in Large Language Models 21 Dec 2024 · 1 repository · arXiv:2412.16806
-
Research on Violent Text Detection System Based on BERT-fasttext Model 21 Dec 2024 · 0 repositories · arXiv:2412.16455
-
Rethinking Model Redundancy for Low-light Image Enhancement 21 Dec 2024 · 0 repositories · arXiv:2412.16459
-
Revisiting MLLMs: An In-Depth Analysis of Image Classification Abilities 21 Dec 2024 · 0 repositories · arXiv:2412.16418
-
RoomPainter: View-Integrated Diffusion for Consistent Indoor Scene Texturing 21 Dec 2024 · 0 repositories · arXiv:2412.16778
-
Semantics Prompting Data-Free Quantization for Low-Bit Vision Transformers 21 Dec 2024 · 0 repositories · arXiv:2412.16553
-
Sensitive Image Classification by Vision Transformers 21 Dec 2024 · 0 repositories · arXiv:2412.16446
-
STKDRec: Spatial-Temporal Knowledge Distillation for Takeaway Recommendation 21 Dec 2024 · 1 repository · arXiv:2412.16502
-
THeGCN: Temporal Heterophilic Graph Convolutional Network 21 Dec 2024 · 0 repositories · arXiv:2412.16435
-
TimeRAG: BOOSTING LLM Time Series Forecasting via Retrieval-Augmented Generation 21 Dec 2024 · 0 repositories · arXiv:2412.16643
-
Towards More Robust Retrieval-Augmented Generation: Evaluating RAG Under Adversarial Poisoning Attacks 21 Dec 2024 · 1 repository · arXiv:2412.16708
-
Trusted Mamba Contrastive Network for Multi-View Clustering 21 Dec 2024 · 1 repository · arXiv:2412.16487
-
UNEM: UNrolled Generalized EM for Transductive Few-Shot Learning 21 Dec 2024 · 1 repository · arXiv:2412.16739
-
VSFormer: Value and Shape-Aware Transformer with Prior-Enhanced Self-Attention for Multivariate Time Series Classification 21 Dec 2024 · 0 repositories · arXiv:2412.16515
-
A Deep Probabilistic Framework for Continuous Time Dynamic Graph Generation 20 Dec 2024 · 1 repository · arXiv:2412.15582
-
Adaptable and Precise: Enterprise-Scenario LLM Function-Calling Capability Training Pipeline 20 Dec 2024 · 0 repositories · arXiv:2412.15660
-
Adversarial Robustness through Dynamic Ensemble Learning 20 Dec 2024 · 0 repositories · arXiv:2412.16254
-
Benchmarking LLMs and SLMs for patient reported outcomes 20 Dec 2024 · 0 repositories · arXiv:2412.16291
-
Can LLMs Obfuscate Code? A Systematic Analysis of Large Language Models into Assembly Code Obfuscation 20 Dec 2024 · 0 repositories · arXiv:2412.16135
-
CCNDF: Curvature Constrained Neural Distance Fields from 3D LiDAR Sequences 20 Dec 2024 · 0 repositories · arXiv:2412.15909
-
CLEAR: Conv-Like Linearization Revs Pre-Trained Diffusion Transformers Up 20 Dec 2024 · 2 repositories · arXiv:2412.16112Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 1 violated, 8 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 13 harvested samples) · 5 pointer-only (licence)
-
CoCoGaussian: Leveraging Circle of Confusion for Gaussian Splatting from Defocused Images 20 Dec 2024 · 0 repositories · arXiv:2412.16028
-
Continual Learning with Strategic Selection and Forgetting for Network Intrusion Detection 20 Dec 2024 · 1 repository · arXiv:2412.16264
-
Decoding Linguistic Nuances in Mental Health Text Classification Using Expressive Narrative Stories 20 Dec 2024 · 0 repositories · arXiv:2412.16302
-
Demystifying the Potential of ChatGPT-4 Vision for Construction Progress Monitoring 20 Dec 2024 · 0 repositories · arXiv:2412.16108
-
Dexterous Manipulation Based on Prior Dexterous Grasp Pose Knowledge 20 Dec 2024 · 0 repositories · arXiv:2412.15587
-
Don't Do RAG: When Cache-Augmented Generation is All You Need for Knowledge Tasks 20 Dec 2024 · 1 repository · arXiv:2412.15605Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Explainable AI for Multivariate Time Series Pattern Exploration: Latent Space Visual Analytics with Temporal Fusion Transformer and Variational Autoencoders in Power Grid Event Diagnosis 20 Dec 2024 · 0 repositories · arXiv:2412.16098
-
Exploiting Multimodal Spatial-temporal Patterns for Video Object Tracking 20 Dec 2024 · 1 repository · arXiv:2412.15691Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 1 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 8 harvested samples) · 1 pointer-only (licence)
-
FedGAT: A Privacy-Preserving Federated Approximation Algorithm for Graph Attention Networks 20 Dec 2024 · 0 repositories · arXiv:2412.16144
-
Function Space Diversity for Uncertainty Prediction via Repulsive Last-Layer Ensembles 20 Dec 2024 · 0 repositories · arXiv:2412.15758
-
GAT-RWOS: Graph Attention-Guided Random Walk Oversampling for Imbalanced Data Classification 20 Dec 2024 · 1 repository · arXiv:2412.16394
-
Graph Structure Refinement with Energy-based Contrastive Learning 20 Dec 2024 · 0 repositories · arXiv:2412.17856