Methods › General › Attention Mechanisms › Attention › Papers, page 29
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 29 of 316: papers 2,801 to 2,900 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Meme Similarity and Emotion Detection using Multimodal Analysis 21 Mar 2025 · 0 repositories · arXiv:2503.17493
-
MSCA-Net:Multi-Scale Context Aggregation Network for Infrared Small Target Detection 21 Mar 2025 · 0 repositories · arXiv:2503.17193
-
Multi-Span Optical Power Spectrum Evolution Modeling using ML-based Multi-Decoder Attention Framework 21 Mar 2025 · 0 repositories · arXiv:2503.17072
-
On Explaining (Large) Language Models For Code Using Global Code-Based Explanations 21 Mar 2025 · 0 repositories · arXiv:2503.16771
-
PVChat: Personalized Video Chat with One-Shot Learning 21 Mar 2025 · 0 repositories · arXiv:2503.17069
-
Rankformer: A Graph Transformer for Recommendation based on Ranking Objective 21 Mar 2025 · 1 repository · arXiv:2503.16927
-
SaudiCulture: A Benchmark for Evaluating Large Language Models Cultural Competence within Saudi Arabia 21 Mar 2025 · 0 repositories · arXiv:2503.17485
-
Scoring, Remember, and Reference: Catching Camouflaged Objects in Videos 21 Mar 2025 · 0 repositories · arXiv:2503.17050
-
Symbolic Audio Classification via Modal Decision Tree Learning 21 Mar 2025 · 0 repositories · arXiv:2503.17018
-
Token Dynamics: Towards Efficient and Dynamic Video Token Representation for Video Large Language Models 21 Mar 2025 · 0 repositories · arXiv:2503.16980
-
Vision Transformer Based Semantic Communications for Next Generation Wireless Networks 21 Mar 2025 · 0 repositories · arXiv:2503.17275
-
When Words Outperform Vision: VLMs Can Self-Improve Via Text-Only Training For Human-Centered Decision Making 21 Mar 2025 · 0 repositories · arXiv:2503.16965Syntology 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Zero-Shot Styled Text Image Generation, but Make It Autoregressive 21 Mar 2025 · 0 repositories · arXiv:2503.17074
-
1000+ FPS 4D Gaussian Splatting for Dynamic Scene Rendering 20 Mar 2025 · 0 repositories · arXiv:2503.16422
-
Attention Pruning: Automated Fairness Repair of Language Models via Surrogate Simulated Annealing 20 Mar 2025 · 0 repositories · arXiv:2503.15815
-
ATTENTION2D: Communication Efficient Distributed Self-Attention Mechanism 20 Mar 2025 · 0 repositories · arXiv:2503.15758
-
Attentional Triple-Encoder Network in Spatiospectral Domains for Medical Image Segmentation 20 Mar 2025 · 0 repositories · arXiv:2503.16389
-
Binarized Mamba-Transformer for Lightweight Quad Bayer HybridEVS Demosaicing 20 Mar 2025 · 1 repository · arXiv:2503.16134
-
Bokehlicious: Photorealistic Bokeh Rendering with Controllable Apertures 20 Mar 2025 · 1 repository · arXiv:2503.16067
-
Design and Implementation of an FPGA-Based Hardware Accelerator for Transformer 20 Mar 2025 · 1 repository · arXiv:2503.16731
-
Disentangled and Interpretable Multimodal Attention Fusion for Cancer Survival Prediction 20 Mar 2025 · 0 repositories · arXiv:2503.16069
-
DocVideoQA: Towards Comprehensive Understanding of Document-Centric Videos through Question Answering 20 Mar 2025 · 0 repositories · arXiv:2503.15887
-
Don't Fight Hallucinations, Use Them: Estimating Image Realism using NLI over Atomic Facts 20 Mar 2025 · 1 repository · arXiv:2503.15948
-
DynamicVis: An Efficient and General Visual Foundation Model for Remote Sensing Image Understanding 20 Mar 2025 · 1 repository · arXiv:2503.16426
-
EDEN: Enhanced Diffusion for High-quality Large-motion Video Frame Interpolation 20 Mar 2025 · 1 repository · arXiv:2503.15831
-
EDiT: Efficient Diffusion Transformers with Linear Compressed Attention 20 Mar 2025 · 0 repositories · arXiv:2503.16726
-
Efficient ANN-Guided Distillation: Aligning Rate-based Features of Spiking Neural Networks through Hybrid Block-wise Replacement 20 Mar 2025 · 0 repositories · arXiv:2503.16572
-
Financial Analysis: Intelligent Financial Data Analysis System Based on LLM-RAG 20 Mar 2025 · 0 repositories · arXiv:2504.06279
-
FreeFlux: Understanding and Exploiting Layer-Specific Roles in RoPE-Based MMDiT for Versatile Image Editing 20 Mar 2025 · 0 repositories · arXiv:2503.16153
-
Gene42: Long-Range Genomic Foundation Model With Dense Attention 20 Mar 2025 · 0 repositories · arXiv:2503.16565
-
GraPLUS: Graph-based Placement Using Semantics for Image Composition 20 Mar 2025 · 0 repositories · arXiv:2503.15761
-
Hybrid-Level Instruction Injection for Video Token Compression in Multi-modal Large Language Models 20 Mar 2025 · 1 repository · arXiv:2503.16036Syntology official (archive's flag): 2 ran · 2 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Hyperspectral Imaging for Identifying Foreign Objects on Pork Belly 20 Mar 2025 · 0 repositories · arXiv:2503.16086
-
iFlame: Interleaving Full and Linear Attention for Efficient Mesh Generation 20 Mar 2025 · 0 repositories · arXiv:2503.16653
-
InhibiDistilbert: Knowledge Distillation for a ReLU and Addition-based Transformer 20 Mar 2025 · 0 repositories · arXiv:2503.15983
-
Iterative Optimal Attention and Local Model for Single Image Rain Streak Removal 20 Mar 2025 · 1 repository · arXiv:2503.16165
-
M2N2V2: Multi-Modal Unsupervised and Training-free Interactive Segmentation 20 Mar 2025 · 0 repositories · arXiv:2503.16254
-
MASH-VLM: Mitigating Action-Scene Hallucination in Video-LLMs through Disentangled Spatial-Temporal Representations 20 Mar 2025 · 0 repositories · arXiv:2503.15871
-
Neural Combinatorial Optimization for Real-World Routing 20 Mar 2025 · 1 repository · arXiv:2503.16159Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 6 harvested samples)
-
Neural Variable-Order Fractional Differential Equation Networks 20 Mar 2025 · 0 repositories · arXiv:2503.16207
-
Deep learning framework for action prediction reveals multi-timescale locomotor control 20 Mar 2025 · 0 repositories · arXiv:2503.16340
-
Parameters vs. Context: Fine-Grained Control of Knowledge Reliance in Language Models 20 Mar 2025 · 1 repository · arXiv:2503.15888
-
PromptHash: Affinity-Prompted Collaborative Cross-Modal Learning for Adaptive Hashing Retrieval 20 Mar 2025 · 0 repositories · arXiv:2503.16064
-
PSA-MIL: A Probabilistic Spatial Attention-Based Multiple Instance Learning for Whole Slide Image Classification 20 Mar 2025 · 1 repository · arXiv:2503.16284
-
Selective Complementary Feature Fusion and Modal Feature Compression Interaction for Brain Tumor Segmentation 20 Mar 2025 · 0 repositories · arXiv:2503.16149
-
Semantic-Guided Global-Local Collaborative Networks for Lightweight Image Super-Resolution 20 Mar 2025 · 1 repository · arXiv:2503.16056
-
SenseExpo: Efficient Autonomous Exploration with Prediction Information from Lightweight Neural Networks 20 Mar 2025 · 0 repositories · arXiv:2503.16000
-
Shining Yourself: High-Fidelity Ornaments Virtual Try-on with Diffusion Model 20 Mar 2025 · 0 repositories · arXiv:2503.16065
-
SpiLiFormer: Enhancing Spiking Transformers with Lateral Inhibition 20 Mar 2025 · 0 repositories · arXiv:2503.15986
-
STOP: Integrated Spatial-Temporal Dynamic Prompting for Video Understanding 20 Mar 2025 · 1 repository · arXiv:2503.15973
-
Temporal-Spatial Attention Network (TSAN) for DoS Attack Detection in Network Traffic 20 Mar 2025 · 0 repositories · arXiv:2503.16047
-
The Lighthouse of Language: Enhancing LLM Agents via Critique-Guided Improvement 20 Mar 2025 · 0 repositories · arXiv:2503.16024
-
Towards Lighter and Robust Evaluation for Retrieval Augmented Generation 20 Mar 2025 · 1 repository · arXiv:2503.16161Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Transfer learning from first-principles calculations to experiments with chemistry-informed domain transformation 20 Mar 2025 · 0 repositories · arXiv:2504.02848
-
Transformer-based Wireless Symbol Detection Over Fading Channels 20 Mar 2025 · 0 repositories · arXiv:2503.16594
-
Tuning LLMs by RAG Principles: Towards LLM-native Memory 20 Mar 2025 · 1 repository · arXiv:2503.16071
-
Typed-RAG: Type-aware Multi-Aspect Decomposition for Non-Factoid Question Answering 20 Mar 2025 · 1 repository · arXiv:2503.15879
-
Unify and Triumph: Polyglot, Diverse, and Self-Consistent Generation of Unit Tests with LLMs 20 Mar 2025 · 0 repositories · arXiv:2503.16144
-
UniHDSA: A Unified Relation Prediction Approach for Hierarchical Document Structure Analysis 20 Mar 2025 · 1 repository · arXiv:2503.15893
-
XAttention: Block Sparse Attention with Antidiagonal Scoring 20 Mar 2025 · 1 repository · arXiv:2503.16428
-
A Comprehensive Survey on Architectural Advances in Deep CNNs: Challenges, Applications, and Emerging Research Directions 19 Mar 2025 · 0 repositories · arXiv:2503.16546
-
A Novel Channel Boosted Residual CNN-Transformer with Regional-Boundary Learning for Breast Cancer Detection 19 Mar 2025 · 0 repositories · arXiv:2503.15008
-
Challenges and Trends in Egocentric Vision: A Survey 19 Mar 2025 · 0 repositories · arXiv:2503.15275
-
ChatGPT or A Silent Everywhere Helper: A Survey of Large Language Models 19 Mar 2025 · 0 repositories · arXiv:2503.17403
-
ChatStitch: Visualizing Through Structures via Surround-View Unsupervised Deep Image Stitching with Collaborative LLM-Agents 19 Mar 2025 · 0 repositories · arXiv:2503.14948
-
Dynamic Bi-Elman Attention Networks: A Dual-Directional Context-Aware Test-Time Learning for Text Classification 19 Mar 2025 · 1 repository · arXiv:2503.15469
-
Dynamic Power Flow Analysis and Fault Characteristics: A Graph Attention Neural Network 19 Mar 2025 · 0 repositories · arXiv:2503.15563
-
ELTEX: A Framework for Domain-Driven Synthetic Data Generation 19 Mar 2025 · 1 repository · arXiv:2503.15055
-
Enhancing Code LLM Training with Programmer Attention 19 Mar 2025 · 0 repositories · arXiv:2503.14936
-
Enhancing Pancreatic Cancer Staging with Large Language Models: The Role of Retrieval-Augmented Generation 19 Mar 2025 · 0 repositories · arXiv:2503.15664
-
Bias Evaluation and Mitigation in Retrieval-Augmented Medical Question-Answering Systems 19 Mar 2025 · 0 repositories · arXiv:2503.15454
-
Fine-Grained Open-Vocabulary Object Detection with Fined-Grained Prompts: Task, Dataset and Benchmark 19 Mar 2025 · 0 repositories · arXiv:2503.14862
-
FP4DiT: Towards Effective Floating Point Quantization for Diffusion Transformers 19 Mar 2025 · 1 repository · arXiv:2503.15465Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
GenM³: Generative Pretrained Multi-path Motion Model for Text Conditional Human Motion Generation 19 Mar 2025 · 0 repositories · arXiv:2503.14919
-
Improving Adversarial Transferability on Vision Transformers via Forward Propagation Refinement 19 Mar 2025 · 1 repository · arXiv:2503.15404
-
Lyapunov-Based Graph Neural Networks for Adaptive Control of Multi-Agent Systems 19 Mar 2025 · 0 repositories · arXiv:2503.15360
-
Optimal Transport Adapter Tuning for Bridging Modality Gaps in Few-Shot Remote Sensing Scene Classification 19 Mar 2025 · 0 repositories · arXiv:2503.14938
-
Optimizing Retrieval Strategies for Financial Question Answering Documents in Retrieval-Augmented Generation Systems 19 Mar 2025 · 1 repository · arXiv:2503.15191Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
RAG-based User Profiling for Precision Planning in Mixed-precision Over-the-Air Federated Learning 19 Mar 2025 · 0 repositories · arXiv:2503.15569
-
Shushing! Let's Imagine an Authentic Speech from the Silent Video 19 Mar 2025 · 0 repositories · arXiv:2503.14928
-
Survey on Generalization Theory for Graph Neural Networks 19 Mar 2025 · 0 repositories · arXiv:2503.15650
-
TROVE: A Challenge for Fine-Grained Text Provenance via Source Sentence Tracing and Relationship Classification 19 Mar 2025 · 0 repositories · arXiv:2503.15289
-
TruthLens:A Training-Free Paradigm for DeepFake Detection 19 Mar 2025 · 0 repositories · arXiv:2503.15342
-
Understanding the Generalization of In-Context Learning in Transformers: An Empirical Study 19 Mar 2025 · 1 repository · arXiv:2503.15579
-
USAM-Net: A U-Net-based Network for Improved Stereo Correspondence and Scene Depth Estimation using Features from a Pre-trained Image Segmentation network 19 Mar 2025 · 0 repositories · arXiv:2503.14950
-
xMOD: Cross-Modal Distillation for 2D/3D Multi-Object Discovery from 2D motion 19 Mar 2025 · 0 repositories · arXiv:2503.15022
-
A-SCoRe: Attention-based Scene Coordinate Regression for wide-ranging scenarios 18 Mar 2025 · 1 repository · arXiv:2503.13982
-
AdaST: Dynamically Adapting Encoder States in the Decoder for End-to-End Speech-to-Text Translation 18 Mar 2025 · 0 repositories · arXiv:2503.14185
-
Beyond holography: the entropic quantum gravity foundations of image processing 18 Mar 2025 · 0 repositories · arXiv:2503.14048
-
BI-RADS prediction of mammographic masses using uncertainty information extracted from a Bayesian Deep Learning model 18 Mar 2025 · 0 repositories · arXiv:2503.13999
-
Binary AddiVortes: (Bayesian) Additive Voronoi Tessellations for Binary Classification with an application to Predicting Home Mortgage Application Outcomes 18 Mar 2025 · 0 repositories · arXiv:2503.21792
-
BurTorch: Revisiting Training from First Principles by Coupling Autodiff, Math Optimization, and Systems 18 Mar 2025 · 1 repository · arXiv:2503.13795
-
COMM:Concentrated Margin Maximization for Robust Document-Level Relation Extraction 18 Mar 2025 · 0 repositories · arXiv:2503.13885
-
Comparative and Interpretative Analysis of CNN and Transformer Models in Predicting Wildfire Spread Using Remote Sensing Data 18 Mar 2025 · 1 repository · arXiv:2503.14150
-
ConSCompF: Consistency-focused Similarity Comparison Framework for Generative Large Language Models 18 Mar 2025 · 0 repositories · arXiv:2503.13923
-
CTSAC: Curriculum-Based Transformer Soft Actor-Critic for Goal-Oriented Robot Exploration 18 Mar 2025 · 0 repositories · arXiv:2503.14254
-
DARS: Dynamic Action Re-Sampling to Enhance Coding Agent Performance by Adaptive Tree Traversal 18 Mar 2025 · 2 repositories · arXiv:2503.14269Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples)
-
DefectFill: Realistic Defect Generation with Inpainting Diffusion Model for Visual Inspection 18 Mar 2025 · 0 repositories · arXiv:2503.13985
-
Dynamic Accumulated Attention Map for Interpreting Evolution of Decision-Making in Vision Transformer 18 Mar 2025 · 1 repository · arXiv:2503.14640
-
Efficient but Vulnerable: Benchmarking and Defending LLM Batch Prompting Attack 18 Mar 2025 · 0 repositories · arXiv:2503.15551