Methods › General › Attention Mechanisms › Attention › Papers, page 52
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 52 of 316: papers 5,101 to 5,200 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
In the Picture: Medical Imaging Datasets, Artifacts, and their Living Review 18 Jan 2025 · 0 repositories · arXiv:2501.10727
-
LD-DETR: Loop Decoder DEtection TRansformer for Video Moment Retrieval and Highlight Detection 18 Jan 2025 · 1 repository · arXiv:2501.10787
-
Semi-supervised Semantic Segmentation for Remote Sensing Images via Multi-scale Uncertainty Consistency and Cross-Teacher-Student Attention 18 Jan 2025 · 1 repository · arXiv:2501.10736
-
Neural Algorithmic Reasoning for Hypergraphs with Looped Transformers 18 Jan 2025 · 0 repositories · arXiv:2501.10688
-
UAV-Assisted Multi-Task Federated Learning with Task Knowledge Sharing 18 Jan 2025 · 0 repositories · arXiv:2501.10644
-
Visual RAG: Expanding MLLM visual knowledge without fine-tuning 18 Jan 2025 · 0 repositories · arXiv:2501.10834
-
4bit-Quantization in Vector-Embedding for RAG 17 Jan 2025 · 1 repository · arXiv:2501.10534
-
ACCEPT: Diagnostic Forecasting of Battery Degradation Through Contrastive Learning 17 Jan 2025 · 0 repositories · arXiv:2501.10492
-
AirRAG: Activating Intrinsic Reasoning for Retrieval Augmented Generation via Tree-based Search 17 Jan 2025 · 0 repositories · arXiv:2501.10053
-
Attention-guided Self-reflection for Zero-shot Hallucination Detection in Large Language Models 17 Jan 2025 · 0 repositories · arXiv:2501.09997
-
BBPOS: BERT-based Part-of-Speech Tagging for Uzbek 17 Jan 2025 · 0 repositories · arXiv:2501.10107
-
Bias in Decision-Making for AI's Ethical Dilemmas: A Comparative Study of ChatGPT and Claude 17 Jan 2025 · 1 repository · arXiv:2501.10484
-
Challenges and recommendations for Electronic Health Records data extraction and preparation for dynamic prediction modelling in hospitalized patients -- a practical guide 17 Jan 2025 · 0 repositories · arXiv:2501.10240
-
DiffVSR: Enhancing Real-World Video Super-Resolution with Diffusion Models for Advanced Visual Quality and Temporal Consistency 17 Jan 2025 · 0 repositories · arXiv:2501.10110
-
DPERC: Direct Parameter Estimation for Mixed Data 17 Jan 2025 · 0 repositories · arXiv:2501.10540
-
Enhancing the Reliability in Machine Learning for Gravitational Wave Parameter Estimation with Attention-Based Models 17 Jan 2025 · 0 repositories · arXiv:2501.10486
-
FiLo++: Zero-/Few-Shot Anomaly Detection by Fused Fine-Grained Descriptions and Deformable Localization 17 Jan 2025 · 1 repository · arXiv:2501.10067
-
HiMix: Reducing Computational Complexity in Large Vision-Language Models 17 Jan 2025 · 0 repositories · arXiv:2501.10318
-
LWGANet: A Lightweight Group Attention Backbone for Remote Sensing Visual Tasks 17 Jan 2025 · 1 repository · arXiv:2501.10040
-
Multi-Modal Attention Networks for Enhanced Segmentation and Depth Estimation of Subsurface Defects in Pulse Thermography 17 Jan 2025 · 1 repository · arXiv:2501.09994
-
MultiPruner: Balanced Structure Removal in Foundation Models 17 Jan 2025 · 1 repository · arXiv:2501.09949
-
PaSa: An LLM Agent for Comprehensive Academic Paper Search 17 Jan 2025 · 1 repository · arXiv:2501.10120Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Passage Segmentation of Documents for Extractive Question Answering 17 Jan 2025 · 0 repositories · arXiv:2501.09940
-
Robust Change Captioning in Remote Sensing: SECOND-CC Dataset and MModalCC Framework 17 Jan 2025 · 0 repositories · arXiv:2501.10075
-
Self-Clustering Graph Transformer Approach to Model Resting-State Functional Brain Activity 17 Jan 2025 · 0 repositories · arXiv:2501.16345
-
The R-Vessel-X Project 17 Jan 2025 · 2 repositories · arXiv:2501.10068
-
A Simple Aerial Detection Baseline of Multimodal Language Models 16 Jan 2025 · 1 repository · arXiv:2501.09720
-
A Simple Graph Contrastive Learning Framework for Short Text Classification 16 Jan 2025 · 1 repository · arXiv:2501.09219Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample) · 1 pointer-only (licence)
-
AdaFV: Rethinking of Visual-Language alignment for VLM acceleration 16 Jan 2025 · 0 repositories · arXiv:2501.09532
-
Attention based Bidirectional GRU hybrid model for inappropriate content detection in Urdu language 16 Jan 2025 · 0 repositories · arXiv:2501.09722
-
CaPa: Carve-n-Paint Synthesis for Efficient 4K Textured Mesh Generation 16 Jan 2025 · 1 repository · arXiv:2501.09433
-
Confidence Estimation for Error Detection in Text-to-SQL Systems 16 Jan 2025 · 1 repository · arXiv:2501.09527
-
Cooperative Decentralized Backdoor Attacks on Vertical Federated Learning 16 Jan 2025 · 0 repositories · arXiv:2501.09320
-
DSTIGCN: Deformable Spatial-Temporal Interaction Graph Convolution Network for Pedestrian Trajectory Prediction 16 Jan 2025 · 1 repository
-
Enhancing Lexicon-Based Text Embeddings with Large Language Models 16 Jan 2025 · 0 repositories · arXiv:2501.09749
-
EraseBench: Understanding The Ripple Effects of Concept Erasure Techniques 16 Jan 2025 · 0 repositories · arXiv:2501.09833
-
Exploring the Inquiry-Diagnosis Relationship with Advanced Patient Simulators 16 Jan 2025 · 1 repository · arXiv:2501.09484
-
Fine-Grained Image-Text Correspondence with Cost Aggregation for Open-Vocabulary Part Segmentation 16 Jan 2025 · 1 repository · arXiv:2501.09688Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 1 honoured, 0 violated, 7 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples) · 3 pointer-only (licence)
-
Free-Knots Kolmogorov-Arnold Network: On the Analysis of Spline Knots and Advancing Stability 16 Jan 2025 · 1 repository · arXiv:2501.09283
-
Generalized Single-Image-Based Morphing Attack Detection Using Deep Representations from Vision Transformer 16 Jan 2025 · 0 repositories · arXiv:2501.09817
-
HSPFormer: Hierarchical Spatial Perception Transformer for Semantic Segmentation 16 Jan 2025 · 1 repository
-
LAVCap: LLM-based Audio-Visual Captioning using Optimal Transport 16 Jan 2025 · 2 repositories · arXiv:2501.09291
-
Learnings from Scaling Visual Tokenizers for Reconstruction and Generation 16 Jan 2025 · 0 repositories · arXiv:2501.09755
-
Leveraging Scale-aware Representations for improved Concept-Representation Alignment in ViTs 16 Jan 2025 · 0 repositories · arXiv:2501.09221
-
Mitigating Hallucinations in Large Vision-Language Models via DPO: On-Policy Data Hold the Key 16 Jan 2025 · 1 repository · arXiv:2501.09695Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 2 where Syntology's instrument failed) · 5 unverified (of 16 harvested samples) · 16 pointer-only (licence)
-
Exploring the Advantages of Sparse Arrays in XL-MIMO Systems: Do Half-Wavelength Arrays Still Offer an Edge in the Near Field? 16 Jan 2025 · 0 repositories · arXiv:2501.09234
-
On Learning Informative Trajectory Embeddings for Imitation, Classification and Regression 16 Jan 2025 · 1 repository · arXiv:2501.09327
-
Perspective Transition of Large Language Models for Solving Subjective Tasks 16 Jan 2025 · 0 repositories · arXiv:2501.09265
-
Practical Continual Forgetting for Pre-trained Vision Models 16 Jan 2025 · 1 repository · arXiv:2501.09705
-
Prompt-CAM: A Simpler Interpretable Transformer for Fine-Grained Analysis 16 Jan 2025 · 1 repository · arXiv:2501.09333Syntology 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample)
-
Inference-Time Alignment in Diffusion Models with Reward-Guided Generation: Tutorial and Review 16 Jan 2025 · 0 repositories · arXiv:2501.09685
-
SE-BSFV: Online Subspace Learning based Shadow Enhancement and Background Suppression for ViSAR under Complex Background 16 Jan 2025 · 0 repositories · arXiv:2501.09341
-
Sentiment Analysis in Twitter Social Network Centered on Cryptocurrencies Using Machine Learning 16 Jan 2025 · 0 repositories · arXiv:2501.09777
-
Soft Knowledge Distillation with Multi-Dimensional Cross-Net Attention for Image Restoration Models Compression 16 Jan 2025 · 0 repositories · arXiv:2501.09321
-
Towards Robust and Realistic Human Pose Estimation via WiFi Signals 16 Jan 2025 · 1 repository · arXiv:2501.09411
-
Unified Face Matching and Physical-Digital Spoofing Attack Detection 16 Jan 2025 · 0 repositories · arXiv:2501.09635
-
Achieving Stability and Optimality: Control Strategy for a Wind Turbine Supplying an Electrolyzer in the Islanded Storage-less Microgrid 15 Jan 2025 · 0 repositories · arXiv:2501.08853
-
Agentic Retrieval-Augmented Generation: A Survey on Agentic RAG 15 Jan 2025 · 1 repository · arXiv:2501.09136
-
Attention is All You Need Until You Need Retention 15 Jan 2025 · 0 repositories · arXiv:2501.09166
-
Augmenting Human-Annotated Training Data with Large Language Model Generation and Distillation in Open-Response Assessment 15 Jan 2025 · 0 repositories · arXiv:2501.09126
-
Beyond Speaker Identity: Text Guided Target Speech Extraction 15 Jan 2025 · 1 repository · arXiv:2501.09169
-
BRIGHT-VO: Brightness-Guided Hybrid Transformer for Visual Odometry with Multi-modality Refinement Module 15 Jan 2025 · 1 repository · arXiv:2501.08659
-
Cancer-Net PCa-Seg: Benchmarking Deep Learning Models for Prostate Cancer Segmentation Using Synthetic Correlated Diffusion Imaging 15 Jan 2025 · 0 repositories · arXiv:2501.09185
-
CityDreamer4D: Compositional Generative Model of Unbounded 4D Cities 15 Jan 2025 · 1 repository · arXiv:2501.08983
-
Computationally Intensive Research: Advancing a Role for Secondary Analysis of Qualitative Data 15 Jan 2025 · 0 repositories · arXiv:2506.04230
-
CT-PatchTST: Channel-Time Patch Time-Series Transformer for Long-Term Renewable Energy Forecasting 15 Jan 2025 · 0 repositories · arXiv:2501.08620
-
Deep Self-Supervised Disturbance Mapping with the OPERA Sentinel-1 Radiometric Terrain Corrected SAR Backscatter Product 15 Jan 2025 · 1 repository · arXiv:2501.09129
-
Dynamic Portfolio Optimization via Augmented DDPG with Quantum Price Levels-Based Trading Strategy 15 Jan 2025 · 0 repositories · arXiv:2501.08528
-
DynamicFace: High-Quality and Consistent Video Face Swapping using Composable 3D Facial Priors 15 Jan 2025 · 0 repositories · arXiv:2501.08553
-
Easing Seasickness through Attention Redirection with a Mindfulness-Based Brain--Computer Interface 15 Jan 2025 · 0 repositories · arXiv:2501.08518
-
Efficient Traffic Prediction Through Spatio-Temporal Distillation 15 Jan 2025 · 1 repository · arXiv:2501.10459Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Enhanced Large Language Models for Effective Screening of Depression and Anxiety 15 Jan 2025 · 0 repositories · arXiv:2501.08769
-
Enhanced Multi-Scale Cross-Attention for Person Image Generation 15 Jan 2025 · 0 repositories · arXiv:2501.08900
-
Expanding Vietnamese SentiWordNet to Improve Performance of Vietnamese Sentiment Analysis Models 15 Jan 2025 · 0 repositories · arXiv:2501.08758
-
Feature-based One-For-All: A Universal Framework for Heterogeneous Knowledge Distillation 15 Jan 2025 · 0 repositories · arXiv:2501.08885Syntology 13 ran (of which 8 constructed an object rather than computing a result; 9 with no instrument failure: 1 honoured, 0 violated, 8 with no contract checked; 4 where Syntology's instrument failed) · 3 unverified (of 16 harvested samples)
-
Generative AI Takes a Statistics Exam: A Comparison of Performance between ChatGPT3.5, ChatGPT4, and ChatGPT4o-mini 15 Jan 2025 · 0 repositories · arXiv:2501.09171
-
GRAPPA - A Hybrid Graph Neural Network for Predicting Pure Component Vapor Pressures 15 Jan 2025 · 1 repository · arXiv:2501.08729
-
Incrementally Learning Multiple Diverse Data Domains via Multi-Source Dynamic Expansion Model 15 Jan 2025 · 0 repositories · arXiv:2501.08878
-
Information Entropy Invariance: Enhancing Length Extrapolation in Attention Mechanisms 15 Jan 2025 · 1 repository · arXiv:2501.08570Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
MAGNET: Augmenting Generative Decoders with Representation Learning and Infilling Capabilities 15 Jan 2025 · 0 repositories · arXiv:2501.08648
-
MIAFEx: An Attention-based Feature Extraction Method for Medical Image Classification 15 Jan 2025 · 0 repositories · arXiv:2501.08562
-
Multi-Class Traffic Assignment using Multi-View Heterogeneous Graph Attention Networks 15 Jan 2025 · 0 repositories · arXiv:2501.09117
-
Multi-View Transformers for Airway-To-Lung Ratio Inference on Cardiac CT Scans: The C4R Study 15 Jan 2025 · 0 repositories · arXiv:2501.08902
-
Multimodal Fake News Video Explanation: Dataset, Analysis and Evaluation 15 Jan 2025 · 0 repositories · arXiv:2501.08514
-
Ouroboros-Diffusion: Exploring Consistent Content Generation in Tuning-free Long Video Diffusion 15 Jan 2025 · 0 repositories · arXiv:2501.09019
-
Polyp detection in colonoscopy images using YOLOv11 15 Jan 2025 · 0 repositories · arXiv:2501.09051
-
RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency 15 Jan 2025 · 0 repositories · arXiv:2501.08682
-
RepVideo: Rethinking Cross-Layer Representation for Video Generation 15 Jan 2025 · 0 repositories · arXiv:2501.08994
-
Rethinking Post-Training Quantization: Introducing a Statistical Pre-Calibration Approach 15 Jan 2025 · 0 repositories · arXiv:2501.09107
-
SuperSAM: Crafting a SAM Supernetwork via Structured Pruning and Unstructured Parameter Prioritization 15 Jan 2025 · 1 repository · arXiv:2501.08504
-
SwinTExCo: Exemplar-based video colorization using Swin Transformer 15 Jan 2025 · 1 repository
-
The Impact of Big Five Personality Traits on AI Agent Decision-Making in Public Spaces: A Social Simulation Study 15 Jan 2025 · 0 repositories · arXiv:2503.15497
-
Transformed Low-rank Adaptation via Tensor Decomposition and Its Applications to Text-to-image Models 15 Jan 2025 · 0 repositories · arXiv:2501.08727
-
UNIR-Net: A Novel Approach for Restoring Underwater Images with Non-Uniform Illumination Using Synthetic Data 15 Jan 2025 · 1 repository · arXiv:2501.09053
-
A Comparative Analysis of Transformer-less Inverter Topologies for Grid-Connected PV Systems: Minimizing Leakage Current and THD 14 Jan 2025 · 0 repositories · arXiv:2501.08103
-
Lilan: A linear latent network approach for real-time solutions of stiff, nonlinear, ordinary differential equations 14 Jan 2025 · 0 repositories · arXiv:2501.08423
-
A Driver Advisory System Based on Large Language Model for High-speed Train 14 Jan 2025 · 0 repositories · arXiv:2501.07837
-
Active Sampling for Node Attribute Completion on Graphs 14 Jan 2025 · 0 repositories · arXiv:2501.08450
-
Advancing Brainwave-Based Biometrics: A Large-Scale, Multi-Session Evaluation 14 Jan 2025 · 1 repository · arXiv:2501.17866
-
ASTRID -- An Automated and Scalable TRIaD for the Evaluation of RAG-based Clinical Question Answering Systems 14 Jan 2025 · 0 repositories · arXiv:2501.08208