Methods › General › Attention Mechanisms › Attention › Papers, page 64
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 64 of 316: papers 6,301 to 6,400 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Glimpse: Enabling White-Box Methods to Use Proprietary Models for Zero-Shot LLM-Generated Text Detection 16 Dec 2024 · 1 repository · arXiv:2412.11506Syntology official: not harvested · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Graph-Guided Textual Explanation Generation Framework 16 Dec 2024 · 0 repositories · arXiv:2412.12318
-
GroupFace: Imbalanced Age Estimation Based on Multi-hop Attention Graph Convolutional Network and Group-aware Margin Optimization 16 Dec 2024 · 0 repositories · arXiv:2412.11450
-
High-speed and High-quality Vision Reconstruction of Spike Camera with Spike Stability Theorem 16 Dec 2024 · 0 repositories · arXiv:2412.11639
-
HResFormer: Hybrid Residual Transformer for Volumetric Medical Image Segmentation 16 Dec 2024 · 0 repositories · arXiv:2412.11458
-
IDArb: Intrinsic Decomposition for Arbitrary Number of Input Views and Illuminations 16 Dec 2024 · 0 repositories · arXiv:2412.12083
-
Inferring Functionality of Attention Heads from their Parameters 16 Dec 2024 · 1 repository · arXiv:2412.11965Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Investigating Mixture of Experts in Dense Retrieval 16 Dec 2024 · 0 repositories · arXiv:2412.11864
-
Learning Implicit Features with Flow Infused Attention for Realistic Virtual Try-On 16 Dec 2024 · 0 repositories · arXiv:2412.11435
-
LLM-RG4: Flexible and Factual Radiology Report Generation across Diverse Input Contexts 16 Dec 2024 · 1 repository · arXiv:2412.12001
-
Look Ahead Text Understanding and LLM Stitching 16 Dec 2024 · 1 repository · arXiv:2412.17836
-
Magnetic Field Data Calibration with Transformer Model Using Physical Constraints: A Scalable Method for Satellite Missions, Illustrated by Tianwen-1 16 Dec 2024 · 0 repositories · arXiv:2501.00020
-
MPQ-DM: Mixed Precision Quantization for Extremely Low Bit Diffusion Models 16 Dec 2024 · 1 repository · arXiv:2412.11549Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 2 violated, 4 with no contract checked; 5 where Syntology's instrument failed) · 2 unverified (of 13 harvested samples) · 3 pointer-only (licence)
-
No More Adam: Learning Rate Scaling at Initialization is All You Need 16 Dec 2024 · 1 repository · arXiv:2412.11768
-
No More Tuning: Prioritized Multi-Task Learning with Lagrangian Differential Multiplier Methods 16 Dec 2024 · 0 repositories · arXiv:2412.12092
-
OpenReviewer: A Specialized Large Language Model for Generating Critical Scientific Paper Reviews 16 Dec 2024 · 0 repositories · arXiv:2412.11948
-
Optimized Quran Passage Retrieval Using an Expanded QA Dataset and Fine-Tuned Language Models 16 Dec 2024 · 0 repositories · arXiv:2412.11431
-
PanSplat: 4K Panorama Synthesis with Feed-Forward Gaussian Splatting 16 Dec 2024 · 1 repository · arXiv:2412.12096
-
Priority-Aware Model-Distributed Inference at Edge Networks 16 Dec 2024 · 0 repositories · arXiv:2412.12371
-
RADARSAT Constellation Mission Compact Polarisation SAR Data for Burned Area Mapping with Deep Learning 16 Dec 2024 · 0 repositories · arXiv:2412.11561
-
RAG Playground: A Framework for Systematic Evaluation of Retrieval Strategies and Prompt Engineering in RAG Systems 16 Dec 2024 · 1 repository · arXiv:2412.12322
-
Second Language (Arabic) Acquisition of LLMs via Progressive Vocabulary Expansion 16 Dec 2024 · 0 repositories · arXiv:2412.12310
-
SegMAN: Omni-scale Context Modeling with State Space Models and Local Attention for Semantic Segmentation 16 Dec 2024 · 1 repository · arXiv:2412.11890
-
SepLLM: Accelerate Large Language Models by Compressing One Segment into One Separator 16 Dec 2024 · 1 repository · arXiv:2412.12094Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
SP²T: Sparse Proxy Attention for Dual-stream Point Transformer 16 Dec 2024 · 0 repositories · arXiv:2412.11540
-
SpeechPrune: Context-aware Token Pruning for Speech Information Retrieval 16 Dec 2024 · 1 repository · arXiv:2412.12009
-
Stepwise Reasoning Error Disruption Attack of LLMs 16 Dec 2024 · 0 repositories · arXiv:2412.11934
-
The Impact of AI Assistance on Radiology Reporting: A Pilot Study Using Simulated AI Draft Reports 16 Dec 2024 · 0 repositories · arXiv:2412.12042
-
The Open Source Advantage in Large Language Models (LLMs) 16 Dec 2024 · 0 repositories · arXiv:2412.12004
-
Token Prepending: A Training-Free Approach for Eliciting Better Sentence Embeddings from LLMs 16 Dec 2024 · 0 repositories · arXiv:2412.11556
-
Towards a Universal Synthetic Video Detector: From Face or Background Manipulations to Fully AI-Generated Content 16 Dec 2024 · 0 repositories · arXiv:2412.12278
-
Transformers Use Causal World Models in Maze-Solving Tasks 16 Dec 2024 · 0 repositories · arXiv:2412.11867
-
Ultra-High-Definition Dynamic Multi-Exposure Image Fusion via Infinite Pixel Learning 16 Dec 2024 · 0 repositories · arXiv:2412.11685
-
Unanswerability Evaluation for Retrieval Augmented Generation 16 Dec 2024 · 0 repositories · arXiv:2412.12300
-
UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer 16 Dec 2024 · 0 repositories · arXiv:2412.11836
-
A Comparative Study on Dynamic Graph Embedding based on Mamba and Transformers 15 Dec 2024 · 0 repositories · arXiv:2412.11293
-
A Contextualized BERT model for Knowledge Graph Completion 15 Dec 2024 · 0 repositories · arXiv:2412.11016
-
Analyzing the Attention Heads for Pronoun Disambiguation in Context-aware Machine Translation Models 15 Dec 2024 · 1 repository · arXiv:2412.11187
-
EEG-GMACN: Interpretable EEG Graph Mutual Attention Convolutional Network 15 Dec 2024 · 0 repositories · arXiv:2412.17834
-
From Votes to Volatility Predicting the Stock Market on Election Day 15 Dec 2024 · 1 repository · arXiv:2412.11192
-
FSTA-SNN:Frequency-based Spatial-Temporal Attention Module for Spiking Neural Networks 15 Dec 2024 · 1 repository · arXiv:2501.14744
-
MoRe: Class Patch Attention Needs Regularization for Weakly Supervised Semantic Segmentation 15 Dec 2024 · 1 repository · arXiv:2412.11076
-
Multi-Graph Co-Training for Capturing User Intent in Session-based Recommendation 15 Dec 2024 · 1 repository · arXiv:2412.11105
-
On the Generalizability of Iterative Patch Selection for Memory-Efficient High-Resolution Image Classification 15 Dec 2024 · 1 repository · arXiv:2412.11237
-
One-Shot Multilingual Font Generation Via ViT 15 Dec 2024 · 0 repositories · arXiv:2412.11342
-
RoLargeSum: A Large Dialect-Aware Romanian News Dataset for Summary, Headline, and Keyword Generation 15 Dec 2024 · 1 repository · arXiv:2412.11317
-
Smaller Language Models Are Better Instruction Evolvers 15 Dec 2024 · 1 repository · arXiv:2412.11231
-
Towards Context-aware Convolutional Network for Image Restoration 15 Dec 2024 · 0 repositories · arXiv:2412.11008
-
Transformer-Based Bearing Fault Detection using Temporal Decomposition Attention Mechanism 15 Dec 2024 · 0 repositories · arXiv:2412.11245
-
ViSymRe: Vision-guided Multimodal Symbolic Regression 15 Dec 2024 · 0 repositories · arXiv:2412.11139
-
Accelerating Retrieval-Augmented Generation 14 Dec 2024 · 0 repositories · arXiv:2412.15246
-
Attention-driven GUI Grounding: Leveraging Pretrained Multimodal Large Language Models without Fine-Tuning 14 Dec 2024 · 1 repository · arXiv:2412.10840Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Boosting ViT-based MRI Reconstruction from the Perspectives of Frequency Modulation, Spatial Purification, and Scale Diversification 14 Dec 2024 · 0 repositories · arXiv:2412.10776
-
CENTAUR: Bridging the Impossible Trinity of Privacy, Efficiency, and Performance in Privacy-Preserving Transformer Inference 14 Dec 2024 · 0 repositories · arXiv:2412.10652
-
DeMo: Decoupled Feature-Based Mixture of Experts for Multi-Modal Object Re-Identification 14 Dec 2024 · 1 repository · arXiv:2412.10650Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 3 pointer-only (licence)
-
Do large language vision models understand 3D shapes? 14 Dec 2024 · 1 repository · arXiv:2412.10908
-
Enhancing Road Crack Detection Accuracy with BsS-YOLO: Optimizing Feature Fusion and Attention Mechanisms 14 Dec 2024 · 0 repositories · arXiv:2412.10902
-
FairGP: A Scalable and Fair Graph Transformer Using Graph Partitioning 14 Dec 2024 · 1 repository · arXiv:2412.10669
-
Graph Attention Hamiltonian Neural Networks: A Lattice System Analysis Model Based on Structural Learning 14 Dec 2024 · 0 repositories · arXiv:2412.10821
-
Heterogeneous Graph Transformer for Multiple Tiny Object Tracking in RGB-T Videos 14 Dec 2024 · 1 repository · arXiv:2412.10861
-
Inference Scaling for Bridging Retrieval and Augmented Generation 14 Dec 2024 · 0 repositories · arXiv:2412.10684
-
Learning Semantic-Aware Representation in Visual-Language Models for Multi-Label Recognition with Partial Labels 14 Dec 2024 · 0 repositories · arXiv:2412.10843
-
Linked Adapters: Linking Past and Future to Present for Effective Continual Learning 14 Dec 2024 · 0 repositories · arXiv:2412.10687
-
MASV: Speaker Verification with Global and Local Context Mamba 14 Dec 2024 · 0 repositories · arXiv:2412.10989
-
MedG-KRP: Medical Graph Knowledge Representation Probing 14 Dec 2024 · 1 repository · arXiv:2412.10982
-
Memory Efficient Matting with Adaptive Token Routing 14 Dec 2024 · 1 repository · arXiv:2412.10702
-
Pop-out vs. Glue: A Study on the pre-attentive and focused attention stages in Visual Search tasks 14 Dec 2024 · 0 repositories · arXiv:2412.12198
-
RAT: Adversarial Attacks on Deep Reinforcement Agents for Targeted Behaviors 14 Dec 2024 · 0 repositories · arXiv:2412.10713
-
Sentiment and Hashtag-aware Attentive Deep Neural Network for Multimodal Post Popularity Prediction 14 Dec 2024 · 0 repositories · arXiv:2412.10737
-
StyleDiT: A Unified Framework for Diverse Child and Partner Faces Synthesis with Style Latent Diffusion Transformer 14 Dec 2024 · 0 repositories · arXiv:2412.10785
-
SusGen-GPT: A Data-Centric LLM for Financial NLP and Sustainability Report Generation 14 Dec 2024 · 1 repository · arXiv:2412.10906
-
Tokens, the oft-overlooked appetizer: Large language models, the distributional hypothesis, and meaning 14 Dec 2024 · 0 repositories · arXiv:2412.10924
-
VisDoM: Multi-Document QA with Visually Rich Elements Using Multimodal Retrieval-Augmented Generation 14 Dec 2024 · 0 repositories · arXiv:2412.10704
-
A Cascaded Dilated Convolution Approach for Mpox Lesion Classification 13 Dec 2024 · 1 repository · arXiv:2412.10106
-
Advances in Transformers for Robotic Applications: A Review 13 Dec 2024 · 0 repositories · arXiv:2412.10599
-
AMuSeD: An Attentive Deep Neural Network for Multimodal Sarcasm Detection Incorporating Bi-modal Data Augmentation 13 Dec 2024 · 0 repositories · arXiv:2412.10103
-
Arbitrary Reading Order Scene Text Spotter with Local Semantics Guidance 13 Dec 2024 · 0 repositories · arXiv:2412.10159
-
Automated Image Captioning with CNNs and Transformers 13 Dec 2024 · 1 repository · arXiv:2412.10511
-
AutoPatent: A Multi-Agent Framework for Automatic Patent Generation 13 Dec 2024 · 1 repository · arXiv:2412.09796
-
Building a Multi-modal Spatiotemporal Expert for Zero-shot Action Recognition with CLIP 13 Dec 2024 · 1 repository · arXiv:2412.09895
-
Byte Latent Transformer: Patches Scale Better Than Tokens 13 Dec 2024 · 1 repository · arXiv:2412.09871Syntology official (archive's flag): 16 ran · 16 ran (of which 0 constructed an object rather than computing a result; 16 with no instrument failure: 0 honoured, 0 violated, 16 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified (of 22 harvested samples) · 22 pointer-only (licence)
-
CognitionCapturer: Decoding Visual Stimuli From Human EEG Signal With Multimodal Information 13 Dec 2024 · 1 repository · arXiv:2412.10489
-
CrossVIT-augmented Geospatial-Intelligence Visualization System for Tracking Economic Development Dynamics 13 Dec 2024 · 1 repository · arXiv:2412.10474
-
CSL-L2M: Controllable Song-Level Lyric-to-Melody Generation Based on Conditional Transformer with Fine-Grained Lyric and Musical Controls 13 Dec 2024 · 0 repositories · arXiv:2412.09887
-
DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding 13 Dec 2024 · 1 repository · arXiv:2412.10302Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 13 harvested samples) · 3 pointer-only (licence)
-
Does Multiple Choice Have a Future in the Age of Generative AI? A Posttest-only RCT 13 Dec 2024 · 1 repository · arXiv:2412.10267
-
Dynamic Try-On: Taming Video Virtual Try-on with Dynamic Attention Mechanism 13 Dec 2024 · 0 repositories · arXiv:2412.09822
-
Edge AI-based Radio Frequency Fingerprinting for IoT Networks 13 Dec 2024 · 0 repositories · arXiv:2412.10553
-
Efficient Large-Scale Traffic Forecasting with Transformers: A Spatial Data Management Perspective 13 Dec 2024 · 3 repositories · arXiv:2412.09972Syntology official (archive's flag): 3 ran · 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 1 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples)
-
Enhancing Multimodal Large Language Models Complex Reason via Similarity Computation 13 Dec 2024 · 1 repository · arXiv:2412.09817
-
Evidence Contextualization and Counterfactual Attribution for Conversational QA over Heterogeneous Data with RAG Systems 13 Dec 2024 · 0 repositories · arXiv:2412.10571
-
FaceShield: Defending Facial Image against Deepfake Threats 13 Dec 2024 · 0 repositories · arXiv:2412.09921
-
HashEvict: A Pre-Attention KV Cache Eviction Strategy using Locality-Sensitive Hashing 13 Dec 2024 · 0 repositories · arXiv:2412.16187
-
Higher Order Transformers: Enhancing Stock Movement Prediction On Multimodal Time-Series Data 13 Dec 2024 · 1 repository · arXiv:2412.10540
-
Infinite-dimensional next-generation reservoir computing 13 Dec 2024 · 1 repository · arXiv:2412.09800Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
IQViC: In-context, Question Adaptive Vision Compressor for Long-term Video Understanding LMMs 13 Dec 2024 · 0 repositories · arXiv:2412.09907
-
Label-template based Few-Shot Text Classification with Contrastive Learning 13 Dec 2024 · 0 repositories · arXiv:2412.10110
-
LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity 13 Dec 2024 · 0 repositories · arXiv:2412.09856
-
Low-Resource Fast Text Classification Based on Intra-Class and Inter-Class Distance Calculation 13 Dec 2024 · 0 repositories · arXiv:2412.09922
-
MANGO: Multimodal Acuity traNsformer for intelliGent ICU Outcomes 13 Dec 2024 · 0 repositories · arXiv:2412.17832