Methods › General › Attention Mechanisms › Attention › Papers, page 70
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 70 of 316: papers 6,901 to 7,000 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Look Every Frame All at Once: Video-Ma²mba for Efficient Long-form Video Understanding with Multi-Axis Gradient Checkpointing 29 Nov 2024 · 0 repositories · arXiv:2411.19460
-
Multi-task CNN Behavioral Embedding Model For Transaction Fraud Detection 29 Nov 2024 · 0 repositories · arXiv:2411.19457
-
On Domain-Specific Post-Training for Multimodal Large Language Models 29 Nov 2024 · 0 repositories · arXiv:2411.19930
-
RAGDiffusion: Faithful Cloth Generation via External Knowledge Assimilation 29 Nov 2024 · 0 repositories · arXiv:2411.19528
-
RL-MILP Solver: A Reinforcement Learning Approach for Solving Mixed-Integer Linear Programs with Graph Neural Networks 29 Nov 2024 · 0 repositories · arXiv:2411.19517
-
SAT-HMR: Real-Time Multi-Person 3D Mesh Estimation via Scale-Adaptive Tokens 29 Nov 2024 · 0 repositories · arXiv:2411.19824
-
SDR-GNN: Spectral Domain Reconstruction Graph Neural Network for Incomplete Multimodal Learning in Conversational Emotion Recognition 29 Nov 2024 · 1 repository · arXiv:2411.19822
-
SIMS: Simulating Stylized Human-Scene Interactions with Retrieval-Augmented Script Generation 29 Nov 2024 · 0 repositories · arXiv:2411.19921
-
T2Vid: Translating Long Text into Multi-Image is the Catalyst for Video-LLMs 29 Nov 2024 · 1 repository · arXiv:2411.19951
-
Towards Santali Linguistic Inclusion: Building the First Santali-to-English Translation Model using mT5 Transformer and Data Augmentation 29 Nov 2024 · 0 repositories · arXiv:2411.19726
-
Towards Understanding Retrieval Accuracy and Prompt Quality in RAG Systems 29 Nov 2024 · 0 repositories · arXiv:2411.19463
-
Train Once for All: A Transitional Approach for Efficient Aspect Sentiment Triplet Extraction 29 Nov 2024 · 0 repositories · arXiv:2412.00208
-
Training Agents with Weakly Supervised Feedback from Large Language Models 29 Nov 2024 · 0 repositories · arXiv:2411.19547
-
Uniform Attention Maps: Boosting Image Fidelity in Reconstruction and Editing 29 Nov 2024 · 1 repository · arXiv:2411.19652
-
V2SFlow: Video-to-Speech Generation with Speech Decomposition and Rectified Flow 29 Nov 2024 · 1 repository · arXiv:2411.19486
-
3D-WAG: Hierarchical Wavelet-Guided Autoregressive Generation for High-Fidelity 3D Shapes 28 Nov 2024 · 0 repositories · arXiv:2411.19037
-
A Lean Dataset for International Math Olympiad: Small Steps towards Writing Math Proofs for Hard Problems 28 Nov 2024 · 0 repositories · arXiv:2411.18872
-
A Survey on Automatic Online Hate Speech Detection in Low-Resource Languages 28 Nov 2024 · 0 repositories · arXiv:2411.19017
-
AMO Sampler: Enhancing Text Rendering with Overshooting 28 Nov 2024 · 1 repository · arXiv:2411.19415
-
An Extensive Evaluation of Factual Consistency in Large Language Models for Data-to-Text Generation 28 Nov 2024 · 0 repositories · arXiv:2411.19203
-
Automatic Prompt Generation and Grounding Object Detection for Zero-Shot Image Anomaly Detection 28 Nov 2024 · 0 repositories · arXiv:2411.19220
-
Beautimeter: Harnessing GPT for Assessing Architectural and Urban Beauty based on the 15 Properties of Living Structure 28 Nov 2024 · 0 repositories · arXiv:2411.19094
-
CLIP meets DINO for Tuning Zero-Shot Classifier using Unlabeled Image Collections 28 Nov 2024 · 1 repository · arXiv:2411.19346
-
Cross-Spectral Attention for Unsupervised RGB-IR Face Verification and Person Re-identification 28 Nov 2024 · 0 repositories · arXiv:2411.19215
-
DENIAHL: In-Context Features Influence LLM Needle-In-A-Haystack Abilities 28 Nov 2024 · 1 repository · arXiv:2411.19360
-
DreamBlend: Advancing Personalized Fine-tuning of Text-to-Image Diffusion Models 28 Nov 2024 · 0 repositories · arXiv:2411.19390
-
Dynamic Attention and Bi-directional Fusion for Safety Helmet Wearing Detection 28 Nov 2024 · 0 repositories · arXiv:2411.19071
-
Efficient Learning Content Retrieval with Knowledge Injection 28 Nov 2024 · 0 repositories · arXiv:2412.00125
-
Efficient Track Anything 28 Nov 2024 · 1 repository · arXiv:2411.18933Syntology 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Enhancing Parameter-Efficient Fine-Tuning of Vision Transformers through Frequency-Based Adaptation 28 Nov 2024 · 1 repository · arXiv:2411.19297
-
GRU-PFG: Extract Inter-Stock Correlation from Stock Factors with Graph Neural Network 28 Nov 2024 · 1 repository · arXiv:2411.18997
-
Habit Coach: Customising RAG-based chatbots to support behavior change 28 Nov 2024 · 0 repositories · arXiv:2411.19229
-
Improving Multi-Subject Consistency in Open-Domain Image Generation with Isolation and Reposition Attention 28 Nov 2024 · 0 repositories · arXiv:2411.19261
-
RevPRAG: Revealing Poisoning Attacks in Retrieval-Augmented Generation through LLM Activation Analysis 28 Nov 2024 · 0 repositories · arXiv:2411.18948
-
Large width penalization for neural network-based prediction interval estimation 28 Nov 2024 · 1 repository · arXiv:2411.19181
-
Locally-Focused Face Representation for Sketch-to-Image Generation Using Noise-Induced Refinement 28 Nov 2024 · 0 repositories · arXiv:2411.19005
-
MAG-V: A Multi-Agent Framework for Synthetic Data Generation and Verification 28 Nov 2024 · 0 repositories · arXiv:2412.04494
-
Marconi: Prefix Caching for the Era of Hybrid LLMs 28 Nov 2024 · 0 repositories · arXiv:2411.19379
-
MaskRIS: Semantic Distortion-aware Data Augmentation for Referring Image Segmentation 28 Nov 2024 · 1 repository · arXiv:2411.19067
-
MATATA: Weakly Supervised End-to-End MAthematical Tool-Augmented Reasoning for Tabular Applications 28 Nov 2024 · 0 repositories · arXiv:2411.18915
-
Pilot Contamination Aware Transformer for Downlink Power Control in Cell-Free Massive MIMO Networks 28 Nov 2024 · 0 repositories · arXiv:2411.19020
-
Random Sampling for Diffusion-based Adversarial Purification 28 Nov 2024 · 1 repository · arXiv:2411.18956
-
SmartLLMSentry: A Comprehensive LLM Based Smart Contract Vulnerability Detection Framework 28 Nov 2024 · 0 repositories · arXiv:2411.19234
-
SOWing Information: Cultivating Contextual Coherence with MLLMs in Image Generation 28 Nov 2024 · 0 repositories · arXiv:2411.19182
-
Sparse Attention Vectors: Generative Multimodal Model Features Are Discriminative Vision-Language Classifiers 28 Nov 2024 · 0 repositories · arXiv:2412.00142
-
T2SG: Traffic Topology Scene Graph for Topology Reasoning in Autonomous Driving 28 Nov 2024 · 0 repositories · arXiv:2411.18894
-
Talking to DINO: Bridging Self-Supervised Vision Backbones with Language for Open-Vocabulary Segmentation 28 Nov 2024 · 1 repository · arXiv:2411.19331Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 10 harvested samples) · 2 pointer-only (licence)
-
The Impact of Example Selection in Few-Shot Prompting on Automated Essay Scoring Using GPT Models 28 Nov 2024 · 0 repositories · arXiv:2411.18924
-
Tracking Progress Towards Sustainable Development Goal 6 Using Satellite Imagery 28 Nov 2024 · 0 repositories · arXiv:2411.19093
-
Trajectory Attention for Fine-grained Video Motion Control 28 Nov 2024 · 0 repositories · arXiv:2411.19324
-
Unleashing the Power of Data Synthesis in Visual Localization 28 Nov 2024 · 0 repositories · arXiv:2412.00138
-
Waterfall Transformer for Multi-person Pose Estimation 28 Nov 2024 · 0 repositories · arXiv:2411.18944
-
3D Scene Graph Guided Vision-Language Pre-training 27 Nov 2024 · 0 repositories · arXiv:2411.18666
-
Dspy-based Neural-Symbolic Pipeline to Enhance Spatial Reasoning in LLMs 27 Nov 2024 · 0 repositories · arXiv:2411.18564
-
A survey on cutting-edge relation extraction techniques based on language models 27 Nov 2024 · 0 repositories · arXiv:2411.18157
-
Addressing bias in Recommender Systems: A Case Study on Data Debiasing Techniques in Mobile Games 27 Nov 2024 · 0 repositories · arXiv:2411.18716
-
AEGIS: An Agent-based Framework for General Bug Reproduction from Issue Descriptions 27 Nov 2024 · 0 repositories · arXiv:2411.18015
-
Aligning Knowledge Concepts to Whole Slide Images for Precise Histopathology Image Analysis 27 Nov 2024 · 1 repository · arXiv:2411.18101
-
Automated Literature Review Using NLP Techniques and LLM-Based Retrieval-Augmented Generation 27 Nov 2024 · 0 repositories · arXiv:2411.18583
-
Can bidirectional encoder become the ultimate winner for downstream applications of foundation models? 27 Nov 2024 · 0 repositories · arXiv:2411.18021
-
Causal and Local Correlations Based Network for Multivariate Time Series Classification 27 Nov 2024 · 1 repository · arXiv:2411.18008
-
ChatGPT as speechwriter for the French presidents 27 Nov 2024 · 0 repositories · arXiv:2411.18382
-
Deep Fourier-embedded Network for Bi-modal Salient Object Detection 27 Nov 2024 · 1 repository · arXiv:2411.18409
-
DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models 27 Nov 2024 · 0 repositories · arXiv:2411.18659
-
DistinctAD: Distinctive Audio Description Generation in Contexts 27 Nov 2024 · 0 repositories · arXiv:2411.18180
-
DRS: Deep Question Reformulation With Structured Output 27 Nov 2024 · 1 repository · arXiv:2411.17993
-
DualCast: Disentangling Aperiodic Events from Traffic Series with a Dual-Branch Model 27 Nov 2024 · 0 repositories · arXiv:2411.18286
-
ElectroVizQA: How well do Multi-modal LLMs perform in Electronics Visual Question Answering? 27 Nov 2024 · 0 repositories · arXiv:2412.00102
-
Enhancing MMDiT-Based Text-to-Image Models for Similar Subject Generation 27 Nov 2024 · 1 repository · arXiv:2411.18301
-
Evaluating and Improving the Robustness of Security Attack Detectors Generated by LLMs 27 Nov 2024 · 1 repository · arXiv:2411.18216
-
Exploring Depth Information for Detecting Manipulated Face Videos 27 Nov 2024 · 0 repositories · arXiv:2411.18572
-
FAM Diffusion: Frequency and Attention Modulation for High-Resolution Image Generation with Stable Diffusion 27 Nov 2024 · 0 repositories · arXiv:2411.18552
-
Fine-Tuning Large Language Models for Scientific Text Classification: A Comparative Study 27 Nov 2024 · 0 repositories · arXiv:2412.00098
-
Fine-Tuning Small Embeddings for Elevated Performance 27 Nov 2024 · 0 repositories · arXiv:2411.18099
-
Foundation Models in Radiology: What, How, When, Why and Why Not 27 Nov 2024 · 0 repositories · arXiv:2411.18730
-
Grid-augmented vision: A simple yet effective approach for enhanced spatial understanding in multi-modal agents 27 Nov 2024 · 1 repository · arXiv:2411.18270
-
HAAT: Hybrid Attention Aggregation Transformer for Image Super-Resolution 27 Nov 2024 · 0 repositories · arXiv:2411.18003
-
HDI-Former: Hybrid Dynamic Interaction ANN-SNN Transformer for Object Detection Using Frames and Events 27 Nov 2024 · 0 repositories · arXiv:2411.18658
-
Heterogeneous Relationships of Subjects and Shapelets for Semi-supervised Multivariate Series Classification 27 Nov 2024 · 0 repositories · arXiv:2411.18043
-
Lightweight Gaze Estimation Model Via Fusion Global Information 27 Nov 2024 · 0 repositories · arXiv:2411.18064
-
LoCATe-GAT: Modeling Multi-Scale Local Context and Action Relationships for Zero-Shot Action Recognition 27 Nov 2024 · 1 repository
-
Mixture of Cache-Conditional Experts for Efficient Mobile Device Inference 27 Nov 2024 · 0 repositories · arXiv:2412.00099
-
Multi-task Gaze Estimation Via Unidirectional Convolution 27 Nov 2024 · 0 repositories · arXiv:2411.18061
-
Multi-Task Model Merging via Adaptive Weight Disentanglement 27 Nov 2024 · 0 repositories · arXiv:2411.18729
-
Multimodal Integration of Longitudinal Noninvasive Diagnostics for Survival Prediction in Immunotherapy Using Deep Learning 27 Nov 2024 · 1 repository · arXiv:2411.18253
-
MvKeTR: Chest CT Report Generation with Multi-View Perception and Knowledge Enhancement 27 Nov 2024 · 0 repositories · arXiv:2411.18309
-
On Importance of Code-Mixed Embeddings for Hate Speech Identification 27 Nov 2024 · 0 repositories · arXiv:2411.18577
-
PATHS: A Hierarchical Transformer for Efficient Whole Slide Image Analysis 27 Nov 2024 · 1 repository · arXiv:2411.18225
-
Perturbation Ontology based Graph Attention Networks 27 Nov 2024 · 0 repositories · arXiv:2411.18520
-
Residual Attention Single-Head Vision Transformer Network for Rolling Bearing Fault Diagnosis in Noisy Environments 27 Nov 2024 · 0 repositories · arXiv:2412.00085
-
ROICtrl: Boosting Instance Control for Visual Generation 27 Nov 2024 · 0 repositories · arXiv:2411.17949
-
RPEE-HEADS: A Novel Benchmark for Pedestrian Head Detection in Crowd Videos 27 Nov 2024 · 0 repositories · arXiv:2411.18164
-
Spectral-Spatial Transformer with Active Transfer Learning for Hyperspectral Image Classification 27 Nov 2024 · 1 repository · arXiv:2411.18115
-
Streamlining Prediction in Bayesian Deep Learning 27 Nov 2024 · 1 repository · arXiv:2411.18425Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 10 harvested samples)
-
TAPTRv3: Spatial and Temporal Context Foster Robust Tracking of Any Point in Long Video 27 Nov 2024 · 0 repositories · arXiv:2411.18671
-
The importance of visual modelling languages in generative software engineering 27 Nov 2024 · 1 repository · arXiv:2411.17976
-
Training and Evaluating Language Models with Template-based Data Generation 27 Nov 2024 · 1 repository · arXiv:2411.18104Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Training Noise Token Pruning 27 Nov 2024 · 1 repository · arXiv:2411.18092
-
TS3-Codec: Transformer-Based Simple Streaming Single Codec 27 Nov 2024 · 1 repository · arXiv:2411.18803
-
Unpacking the Individual Components of Diffusion Policy 27 Nov 2024 · 0 repositories · arXiv:2412.00084