Methods › General › Attention Mechanisms › Attention › Papers, page 33
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 33 of 316: papers 3,201 to 3,300 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Efficient Many-Shot In-Context Learning with Dynamic Block-Sparse Attention 11 Mar 2025 · 1 repository · arXiv:2503.08640
-
EFPC: Towards Efficient and Flexible Prompt Compression 11 Mar 2025 · 0 repositories · arXiv:2503.07956
-
ESNLIR: A Spanish Multi-Genre Dataset with Causal Relationships 11 Mar 2025 · 0 repositories · arXiv:2503.08803
-
External Knowledge Injection for CLIP-Based Class-Incremental Learning 11 Mar 2025 · 3 repositories · arXiv:2503.08510
-
From Slices to Sequences: Autoregressive Tracking Transformer for Cohesive and Consistent 3D Lymph Node Detection in CT Scans 11 Mar 2025 · 0 repositories · arXiv:2503.07933
-
GPT-PPG: A GPT-based Foundation Model for Photoplethysmography Signals 11 Mar 2025 · 0 repositories · arXiv:2503.08015
-
Gradient-guided Attention Map Editing: Towards Efficient Contextual Hallucination Mitigation 11 Mar 2025 · 0 repositories · arXiv:2503.08963
-
Guess What I am Thinking: A Benchmark for Inner Thought Reasoning of Role-Playing Language Agents 11 Mar 2025 · 0 repositories · arXiv:2503.08193
-
HOTFormerLoc: Hierarchical Octree Transformer for Versatile Lidar Place Recognition Across Ground and Aerial Views 11 Mar 2025 · 0 repositories · arXiv:2503.08140
-
In silico clinical trials in drug development: a systematic review 11 Mar 2025 · 0 repositories · arXiv:2503.08746
-
Interpretable and Robust Dialogue State Tracking via Natural Language Summarization with LLMs 11 Mar 2025 · 0 repositories · arXiv:2503.08857
-
Interpreting the Repeated Token Phenomenon in Large Language Models 11 Mar 2025 · 1 repository · arXiv:2503.08908
-
Joint Image-Instance Spatial-Temporal Attention for Few-shot Action Recognition 11 Mar 2025 · 0 repositories · arXiv:2503.14430
-
KAN-Mixers: a new deep learning architecture for image classification 11 Mar 2025 · 0 repositories · arXiv:2503.08939
-
LLM-based Corroborating and Refuting Evidence Retrieval for Scientific Claim Verification 11 Mar 2025 · 0 repositories · arXiv:2503.07937
-
LLMs Know What to Drop: Self-Attention Guided KV Cache Eviction for Efficient Long-Context Inference 11 Mar 2025 · 0 repositories · arXiv:2503.08879
-
Llms, Virtual Users, and Bias: Predicting Any Survey Question Without Human Data 11 Mar 2025 · 0 repositories · arXiv:2503.16498
-
MaskAttn-UNet: A Mask Attention-Driven Framework for Universal Low-Resolution Image Segmentation 11 Mar 2025 · 0 repositories · arXiv:2503.10686
-
MEAT: Multiview Diffusion Model for Human Generation on Megapixels with Mesh Attention 11 Mar 2025 · 1 repository · arXiv:2503.08664
-
MFRS: A Multi-Frequency Reference Series Approach to Scalable and Accurate Time-Series Forecasting 11 Mar 2025 · 0 repositories · arXiv:2503.08328
-
MVGSR: Multi-View Consistency Gaussian Splatting for Robust Surface Reconstruction 11 Mar 2025 · 0 repositories · arXiv:2503.08093
-
On Digital Optimization of Analog Self-Interference Cancellation for Full-Duplex Wireless Systems 11 Mar 2025 · 0 repositories · arXiv:2503.08357
-
OpenRAG: Optimizing RAG End-to-End via In-Context Retrieval Learning 11 Mar 2025 · 0 repositories · arXiv:2503.08398
-
Overlap-aware meta-learning attention to enhance hypergraph neural networks for node classification 11 Mar 2025 · 0 repositories · arXiv:2503.07961
-
PromptGAR: Flexible Promptive Group Activity Recognition 11 Mar 2025 · 0 repositories · arXiv:2503.08933
-
QUIET-SR: Quantum Image Enhancement Transformer for Single Image Super-Resolution 11 Mar 2025 · 0 repositories · arXiv:2503.08759
-
QuoTA: Query-oriented Token Assignment via CoT Query Decouple for Long Video Comprehension 11 Mar 2025 · 1 repository · arXiv:2503.08689Syntology official: no sample here; runs from other or unrecorded repositories · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Recent Advances in Hypergraph Neural Networks 11 Mar 2025 · 0 repositories · arXiv:2503.07959
-
Seeing What's Not There: Spurious Correlation in Multimodal LLMs 11 Mar 2025 · 0 repositories · arXiv:2503.08884
-
Shedding Light in Task Decomposition in Program Synthesis: The Driving Force of the Synthesizer Model 11 Mar 2025 · 0 repositories · arXiv:2503.08738
-
Simulating Automotive Radar with Lidar and Camera Inputs 11 Mar 2025 · 0 repositories · arXiv:2503.08068
-
STEAD: Spatio-Temporal Efficient Anomaly Detection for Time and Compute Sensitive Applications 11 Mar 2025 · 1 repository · arXiv:2503.07942
-
Stick to Facts: Towards Fidelity-oriented Product Description Generation 11 Mar 2025 · 0 repositories · arXiv:2503.08454
-
The Algorithmic State Architecture (ASA): An Integrated Framework for AI-Enabled Government 11 Mar 2025 · 0 repositories · arXiv:2503.08725
-
TransECG: Leveraging Transformers for Explainable ECG Re-identification Risk Analysis 11 Mar 2025 · 0 repositories · arXiv:2503.13495
-
Vision Transformer for Intracranial Hemorrhage Classification in CT Scans Using an Entropy-Aware Fuzzy Integral Strategy for Adaptive Scan-Level Decision Fusion 11 Mar 2025 · 0 repositories · arXiv:2503.08609
-
Visual Attention Graph 11 Mar 2025 · 0 repositories · arXiv:2503.08531
-
WISA: World Simulator Assistant for Physics-Aware Text-to-Video Generation 11 Mar 2025 · 0 repositories · arXiv:2503.08153
-
A Comprehensive Survey of Mixture-of-Experts: Algorithms, Theory, and Applications 10 Mar 2025 · 1 repository · arXiv:2503.07137Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
A Data-Centric Revisit of Pre-Trained Vision Models for Robot Learning 10 Mar 2025 · 1 repository · arXiv:2503.06960Syntology official (archive's flag): 3 ran · 3 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
A LongFormer-Based Framework for Accurate and Efficient Medical Text Summarization 10 Mar 2025 · 0 repositories · arXiv:2503.06888
-
A LSTM-Transformer Model for pulsation control of pVADs 10 Mar 2025 · 0 repositories · arXiv:2503.07110
-
A Recipe for Improving Remote Sensing VLM Zero Shot Generalization 10 Mar 2025 · 0 repositories · arXiv:2503.08722
-
AttentionSwarm: Reinforcement Learning with Attention Control Barier Function for Crazyflie Drones in Dynamic Environments 10 Mar 2025 · 0 repositories · arXiv:2503.07376
-
AttFC: Attention Fully-Connected Layer for Large-Scale Face Recognition with One GPU 10 Mar 2025 · 0 repositories · arXiv:2503.06839
-
Bot Wars Evolved: Orchestrating Competing LLMs in a Counterstrike Against Phone Scams 10 Mar 2025 · 0 repositories · arXiv:2503.07036
-
Building English ASR model with regional language support 10 Mar 2025 · 0 repositories · arXiv:2503.07522
-
CATANet: Efficient Content-Aware Token Aggregation for Lightweight Image Super-Resolution 10 Mar 2025 · 1 repository · arXiv:2503.06896
-
CtrlRAG: Black-box Adversarial Attacks Based on Masked Language Models in Retrieval-Augmented Language Generation 10 Mar 2025 · 0 repositories · arXiv:2503.06950
-
DirectTriGS: Triplane-based Gaussian Splatting Field Representation for 3D Generation 10 Mar 2025 · 0 repositories · arXiv:2503.06900
-
DreamRelation: Relation-Centric Video Customization 10 Mar 2025 · 0 repositories · arXiv:2503.07602
-
Dynamic Cross-Modal Feature Interaction Network for Hyperspectral and LiDAR Data Classification 10 Mar 2025 · 1 repository · arXiv:2503.06945
-
EasyControl: Adding Efficient and Flexible Control for Diffusion Transformer 10 Mar 2025 · 0 repositories · arXiv:2503.07027
-
EAZY: Eliminating Hallucinations in LVLMs by Zeroing out Hallucinatory Image Tokens 10 Mar 2025 · 0 repositories · arXiv:2503.07772
-
Enhancing Time Series Forecasting via Logic-Inspired Regularization 10 Mar 2025 · 0 repositories · arXiv:2503.06867
-
Erase Diffusion: Empowering Object Removal Through Calibrating Diffusion Pathways 10 Mar 2025 · 0 repositories · arXiv:2503.07026
-
Exploring Multimodal Perception in Large Language Models Through Perceptual Strength Ratings 10 Mar 2025 · 0 repositories · arXiv:2503.06980
-
Exposure Bias Reduction for Enhancing Diffusion Transformer Feature Caching 10 Mar 2025 · 1 repository · arXiv:2503.07120
-
Filter Images First, Generate Instructions Later: Pre-Instruction Data Selection for Visual Instruction Tuning 10 Mar 2025 · 0 repositories · arXiv:2503.07591
-
Find your Needle: Small Object Image Retrieval via Multi-Object Attention Optimization 10 Mar 2025 · 0 repositories · arXiv:2503.07038
-
FinTSBridge: A New Evaluation Suite for Real-world Financial Prediction with Advanced Time Series Models 10 Mar 2025 · 0 repositories · arXiv:2503.06928
-
From Idea to Implementation: Evaluating the Influence of Large Language Models in Software Development -- An Opinion Paper 10 Mar 2025 · 0 repositories · arXiv:2503.07450
-
Fully Autonomous Programming using Iterative Multi-Agent Debugging with Large Language Models 10 Mar 2025 · 0 repositories · arXiv:2503.07693
-
Implicit Reasoning in Transformers is Reasoning through Shortcuts 10 Mar 2025 · 1 repository · arXiv:2503.07604Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples) · 1 pointer-only (licence)
-
Improving cognitive diagnostics in pathology: a deep learning approach for augmenting perceptional understanding of histopathology images 10 Mar 2025 · 0 repositories · arXiv:2503.06894
-
Inorganic Catalyst Efficiency Prediction Based on EAPCR Model: A Deep Learning Solution for Multi-Source Heterogeneous Data 10 Mar 2025 · 0 repositories · arXiv:2503.07424
-
Interference-Aware Super-Constellation Design for NOMA 10 Mar 2025 · 0 repositories · arXiv:2503.07509
-
Inversion-Free Video Style Transfer with Trajectory Reset Attention Control and Content-Style Bridging 10 Mar 2025 · 0 repositories · arXiv:2503.07363
-
Large model enhanced computational ghost imaging 10 Mar 2025 · 1 repository · arXiv:2503.08710
-
Learning a Unified Degradation-aware Representation Model for Multi-modal Image Fusion 10 Mar 2025 · 0 repositories · arXiv:2503.07033
-
MADS: Multi-Attribute Document Supervision for Zero-Shot Image Classification 10 Mar 2025 · 0 repositories · arXiv:2503.06847
-
MambaFlow: A Mamba-Centric Architecture for End-to-End Optical Flow Estimation 10 Mar 2025 · 0 repositories · arXiv:2503.07046
-
MapQA: Open-domain Geospatial Question Answering on Map Data 10 Mar 2025 · 0 repositories · arXiv:2503.07871
-
Open-Set Gait Recognition from Sparse mmWave Radar Point Clouds 10 Mar 2025 · 1 repository · arXiv:2503.07435
-
Petri Net Modeling of Root Hair Response to Phosphate Starvation in Arabidopsis Thaliana 10 Mar 2025 · 0 repositories · arXiv:2503.07477
-
Post-Training Quantization for Diffusion Transformer via Hierarchical Timestep Grouping 10 Mar 2025 · 0 repositories · arXiv:2503.06930
-
Reproducibility and Artifact Consistency of the SIGIR 2022 Recommender Systems Papers Based on Message Passing 10 Mar 2025 · 0 repositories · arXiv:2503.07823
-
ResMoE: Space-efficient Compression of Mixture of Experts LLMs via Residual Restoration 10 Mar 2025 · 1 repository · arXiv:2503.06881Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 3 where Syntology's instrument failed) · 5 unverified (of 15 harvested samples) · 1 pointer-only (licence)
-
Roamify: Designing and Evaluating an LLM Based Google Chrome Extension for Personalised Itinerary Planning 10 Mar 2025 · 1 repository · arXiv:2504.10489
-
SOYO: A Tuning-Free Approach for Video Style Morphing via Style-Adaptive Interpolation in Diffusion Models 10 Mar 2025 · 0 repositories · arXiv:2503.06998
-
Split-n-Chain: Privacy-Preserving Multi-Node Split Learning with Blockchain-Based Auditability 10 Mar 2025 · 0 repositories · arXiv:2503.07570
-
Taking Notes Brings Focus? Towards Multi-Turn Multimodal Dialogue Learning 10 Mar 2025 · 0 repositories · arXiv:2503.07002
-
Talking to GDELT Through Knowledge Graphs 10 Mar 2025 · 0 repositories · arXiv:2503.07584
-
Two-stage Deep Denoising with Self-guided Noise Attention for Multimodal Medical Images 10 Mar 2025 · 0 repositories · arXiv:2503.06827
-
Using a single actor to output personalized policy for different intersections 10 Mar 2025 · 0 repositories · arXiv:2503.07678
-
VACE: All-in-One Video Creation and Editing 10 Mar 2025 · 2 repositories · arXiv:2503.07598Syntology 11 ran (of which 5 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 1 violated, 8 with no contract checked; 2 where Syntology's instrument failed) · 8 unverified (of 19 harvested samples)
-
Visual and Text Prompt Segmentation: A Novel Multi-Model Framework for Remote Sensing 10 Mar 2025 · 0 repositories · arXiv:2503.07911
-
YOLOMG: Vision-based Drone-to-Drone Detection with Appearance and Pixel-Level Motion Fusion 10 Mar 2025 · 1 repository · arXiv:2503.07115
-
A Quantitative Evaluation of the Expressivity of BMI, Pose and Gender in Body Embeddings for Recognition and Identification 9 Mar 2025 · 0 repositories · arXiv:2503.06451
-
ARMOR v0.1: Empowering Autoregressive Multimodal Understanding Model with Interleaved Multimodal Generation via Asymmetric Synergy 9 Mar 2025 · 0 repositories · arXiv:2503.06542
-
Beyond Decoder-only: Large Language Models Can be Good Encoders for Machine Translation 9 Mar 2025 · 1 repository · arXiv:2503.06594
-
Can Small Language Models Reliably Resist Jailbreak Attacks? A Comprehensive Evaluation 9 Mar 2025 · 0 repositories · arXiv:2503.06519
-
Causality Enhanced Origin-Destination Flow Prediction in Data-Scarce Cities 9 Mar 2025 · 0 repositories · arXiv:2503.06398
-
Conceptrol: Concept Control of Zero-shot Personalized Image Generation 9 Mar 2025 · 1 repository · arXiv:2503.06568
-
DiffCLIP: Differential Attention Meets CLIP 9 Mar 2025 · 1 repository · arXiv:2503.06626Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Diffusion Model Based Probabilistic Day-ahead Load Forecasting 9 Mar 2025 · 0 repositories · arXiv:2503.06697
-
DynamicID: Zero-Shot Multi-ID Image Personalization with Flexible Facial Editability 9 Mar 2025 · 0 repositories · arXiv:2503.06505
-
Effectiveness of Zero-shot-CoT in Japanese Prompts 9 Mar 2025 · 0 repositories · arXiv:2503.06765
-
Emulating Self-attention with Convolution for Efficient Image Super-Resolution 9 Mar 2025 · 1 repository · arXiv:2503.06671
-
Enhancing Layer Attention Efficiency through Pruning Redundant Retrievals 9 Mar 2025 · 0 repositories · arXiv:2503.06473