Methods › General › Attention Mechanisms › Attention › Papers, page 49
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 49 of 316: papers 4,801 to 4,900 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
LCTG Bench: LLM Controlled Text Generation Benchmark 27 Jan 2025 · 1 repository · arXiv:2501.15875
-
LemmaHead: RAG Assisted Proof Generation Using Large Language Models 27 Jan 2025 · 0 repositories · arXiv:2501.15797
-
Leveraging Video Vision Transformer for Alzheimer's Disease Diagnosis from 3D Brain MRI 27 Jan 2025 · 0 repositories · arXiv:2501.15733
-
Long-Term Interest Clock: Fine-Grained Time Perception in Streaming Recommendation System 27 Jan 2025 · 0 repositories · arXiv:2501.15817
-
MM-Retinal V2: Transfer an Elite Knowledge Spark into Fundus Vision-Language Pretraining 27 Jan 2025 · 1 repository · arXiv:2501.15798
-
Multi-View Attention Syntactic Enhanced Graph Convolutional Network for Aspect-based Sentiment Analysis 27 Jan 2025 · 1 repository · arXiv:2501.15968
-
Object Detection for Medical Image Analysis: Insights from the RT-DETR Model 27 Jan 2025 · 0 repositories · arXiv:2501.16469
-
One-Bit Sigma-Delta DFRC Waveform Design: Using Quantization Noise for Radar Probing 27 Jan 2025 · 0 repositories · arXiv:2501.15868
-
Parametric Retrieval Augmented Generation 27 Jan 2025 · 1 repository · arXiv:2501.15915
-
PDC-ViT : Source Camera Identification using Pixel Difference Convolution and Vision Transformer 27 Jan 2025 · 0 repositories · arXiv:2501.16227
-
Phase Transitions in Large Language Models and the O(N) Model 27 Jan 2025 · 0 repositories · arXiv:2501.16241
-
Provence: efficient and robust context pruning for retrieval-augmented generation 27 Jan 2025 · 0 repositories · arXiv:2501.16214
-
RelCAT: Advancing Extraction of Clinical Inter-Entity Relationships from Unstructured Electronic Health Records 27 Jan 2025 · 1 repository · arXiv:2501.16077
-
Rethinking the Bias of Foundation Model under Long-tailed Distribution 27 Jan 2025 · 0 repositories · arXiv:2501.15955
-
Risk-Aware Distributional Intervention Policies for Language Models 27 Jan 2025 · 1 repository · arXiv:2501.15758
-
SIM: Surface-based fMRI Analysis for Inter-Subject Multimodal Decoding from Movie-Watching Experiments 27 Jan 2025 · 1 repository · arXiv:2501.16471
-
Skill-Based Labor Market Polarization in the Age of AI: A Comparative Analysis of India and the United States 27 Jan 2025 · 0 repositories · arXiv:2501.15809
-
Slot-Guided Adaptation of Pre-trained Diffusion Models for Object-Centric Learning and Compositional Generation 27 Jan 2025 · 0 repositories · arXiv:2501.15878
-
SWIFT: Mapping Sub-series with Wavelet Decomposition Improves Time Series Forecasting 27 Jan 2025 · 1 repository · arXiv:2501.16178
-
The Effect of Optimal Self-Distillation in Noisy Gaussian Mixture Model 27 Jan 2025 · 1 repository · arXiv:2501.16226Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
The Linear Attention Resurrection in Vision Transformer 27 Jan 2025 · 0 repositories · arXiv:2501.16182
-
TransPathNet: A Novel Two-Stage Framework for Indoor Radio Map Prediction 27 Jan 2025 · 0 repositories · arXiv:2501.16023
-
UniPET-SPK: A Unified Framework for Parameter-Efficient Tuning of Pre-trained Speech Models for Robust Speaker Verification 27 Jan 2025 · 0 repositories · arXiv:2501.16542
-
URAG: Implementing a Unified Hybrid RAG for Precise Answers in University Admission Chatbots -- A Case Study at HCMUT 27 Jan 2025 · 0 repositories · arXiv:2501.16276
-
Weight-based Analysis of Detokenization in Language Models: Understanding the First Stage of Inference Without Inference 27 Jan 2025 · 0 repositories · arXiv:2501.15754
-
Adapting Biomedical Abstracts into Plain language using Large Language Models 26 Jan 2025 · 0 repositories · arXiv:2501.15700
-
AI-Driven Secure Data Sharing: A Trustworthy and Privacy-Preserving Approach 26 Jan 2025 · 0 repositories · arXiv:2501.15363
-
ARWKV: Pretrain is not what we need, an RNN-Attention-Based Language Model Born from Transformer 26 Jan 2025 · 1 repository · arXiv:2501.15570
-
CE-SDWV: Effective and Efficient Concept Erasure for Text-to-Image Diffusion Models via a Semantic-Driven Word Vocabulary 26 Jan 2025 · 0 repositories · arXiv:2501.15562
-
Classifying Deepfakes Using Swin Transformers 26 Jan 2025 · 0 repositories · arXiv:2501.15656
-
Decentralized Low-Rank Fine-Tuning of Large Language Models 26 Jan 2025 · 0 repositories · arXiv:2501.15361
-
Doracamom: Joint 3D Detection and Occupancy Prediction with Multi-view 4D Radars and Cameras for Omnidirectional Perception 26 Jan 2025 · 0 repositories · arXiv:2501.15394
-
Economic Implications of Corporate Governance and Corporate Social Responsibility: Evidence from Banks in Bangladesh 26 Jan 2025 · 0 repositories · arXiv:2501.15594
-
End-to-End Target Speaker Speech Recognition Using Context-Aware Attention Mechanisms for Challenging Enrollment Scenario 26 Jan 2025 · 0 repositories · arXiv:2501.15466
-
Estimating Committor Functions via Deep Adaptive Sampling on Rare Transition Paths 26 Jan 2025 · 0 repositories · arXiv:2501.15522
-
Evaluating the Effectiveness of XAI Techniques for Encoder-Based Language Models 26 Jan 2025 · 0 repositories · arXiv:2501.15374
-
Guaranteed Multidimensional Time Series Prediction via Deterministic Tensor Completion Theory 26 Jan 2025 · 1 repository · arXiv:2501.15388
-
iFormer: Integrating ConvNet and Transformer for Mobile Application 26 Jan 2025 · 1 repository · arXiv:2501.15369Syntology official (archive's flag): 5 ran · 5 ran (of which 4 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 5 unverified (of 10 harvested samples) · 1 pointer-only (licence)
-
Improving Estonian Text Simplification through Pretrained Language Models and Custom Datasets 26 Jan 2025 · 0 repositories · arXiv:2501.15624
-
Information Consistent Pruning: How to Efficiently Search for Sparse Networks? 26 Jan 2025 · 1 repository · arXiv:2501.15592
-
OpenCharacter: Training Customizable Role-Playing LLMs with Large-Scale Synthetic Personas 26 Jan 2025 · 0 repositories · arXiv:2501.15427
-
Optimal Transport on Categorical Data for Counterfactuals using Compositional Data and Dirichlet Transport 26 Jan 2025 · 1 repository · arXiv:2501.15549
-
Quantum-Enhanced Attention Mechanism in NLP: A Hybrid Classical-Quantum Approach 26 Jan 2025 · 0 repositories · arXiv:2501.15630
-
Qwen2.5-1M Technical Report 26 Jan 2025 · 0 repositories · arXiv:2501.15383
-
RLER-TTE: An Efficient and Effective Framework for En Route Travel Time Estimation with Reinforcement Learning 26 Jan 2025 · 0 repositories · arXiv:2501.15493
-
SedarEval: Automated Evaluation using Self-Adaptive Rubrics 26 Jan 2025 · 1 repository · arXiv:2501.15595Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Self-supervised Benchmark Lottery on ImageNet: Do Marginal Improvements Translate to Improvements on Similar Datasets? 26 Jan 2025 · 0 repositories · arXiv:2501.15431
-
StagFormer: Time Staggering Transformer Decoding for RunningLayers In Parallel 26 Jan 2025 · 0 repositories · arXiv:2501.15665
-
Stroke Lesion Segmentation using Multi-Stage Cross-Scale Attention 26 Jan 2025 · 0 repositories · arXiv:2501.15423
-
TdAttenMix: Top-Down Attention Guided Mixup 26 Jan 2025 · 1 repository · arXiv:2501.15409
-
TensorLLM: Tensorising Multi-Head Attention for Enhanced Reasoning and Compression in LLMs 26 Jan 2025 · 1 repository · arXiv:2501.15674
-
Transformer^-1: Input-Adaptive Computation for Resource-Constrained Deployment 26 Jan 2025 · 0 repositories · arXiv:2501.16394
-
Visualizing Uncertainty in Translation Tasks: An Evaluation of LLM Performance and Confidence Metrics 26 Jan 2025 · 1 repository · arXiv:2501.17187
-
A Review on Self-Supervised Learning for Time Series Anomaly Detection: Recent Advances and Open Challenges 25 Jan 2025 · 1 repository · arXiv:2501.15196
-
A Training-free Synthetic Data Selection Method for Semantic Segmentation 25 Jan 2025 · 1 repository · arXiv:2501.15201
-
A Two-Stage CAE-Based Federated Learning Framework for Efficient Jamming Detection in 5G Networks 25 Jan 2025 · 0 repositories · arXiv:2501.15288
-
ABXI: Invariant Interest Adaptation for Task-Guided Cross-Domain Sequential Recommendation 25 Jan 2025 · 1 repository · arXiv:2501.15118
-
Advanced Real-Time Fraud Detection Using RAG-Based LLMs 25 Jan 2025 · 0 repositories · arXiv:2501.15290
-
AKVQ-VL: Attention-Aware KV Cache Adaptive 2-Bit Quantization for Vision-Language Models 25 Jan 2025 · 0 repositories · arXiv:2501.15021
-
An AI-Driven Live Systematic Reviews in the Brain-Heart Interconnectome: Minimizing Research Waste and Advancing Evidence Synthesis 25 Jan 2025 · 1 repository · arXiv:2501.17181
-
An Attempt to Unraveling Token Prediction Refinement and Identifying Essential Layers of Large Language Models 25 Jan 2025 · 0 repositories · arXiv:2501.15054
-
ASRank: Zero-Shot Re-Ranking with Answer Scent for Document Retrieval 25 Jan 2025 · 0 repositories · arXiv:2501.15245
-
CG-RAG: Research Question Answering by Citation Graph Retrieval-Augmented LLMs 25 Jan 2025 · 0 repositories · arXiv:2501.15067
-
Cross-modal Context Fusion and Adaptive Graph Convolutional Network for Multimodal Conversational Emotion Recognition 25 Jan 2025 · 0 repositories · arXiv:2501.15063
-
Crystal Oscillators in OSNMA-Enabled Receivers: An Implementation View for Automotive Applications 25 Jan 2025 · 0 repositories · arXiv:2501.15123
-
Deep Multimodal Learning for Real-Time DDoS Attacks Detection in Internet of Vehicles 25 Jan 2025 · 1 repository · arXiv:2501.15252
-
Exact Fit Attention in Node-Holistic Graph Convolutional Network for Improved EEG-Based Driver Fatigue Detection 25 Jan 2025 · 0 repositories · arXiv:2501.15062
-
Figurative-cum-Commonsense Knowledge Infusion for Multimodal Mental Health Meme Classification 25 Jan 2025 · 1 repository · arXiv:2501.15321
-
Generalizable Deepfake Detection via Effective Local-Global Feature Extraction 25 Jan 2025 · 1 repository · arXiv:2501.15253
-
Generating Negative Samples for Multi-Modal Recommendation 25 Jan 2025 · 0 repositories · arXiv:2501.15183
-
Group Ligands Docking to Protein Pockets 25 Jan 2025 · 0 repositories · arXiv:2501.15055
-
ILETIA: An AI-enhanced method for individualized trigger-oocyte pickup interval estimation of progestin-primed ovarian stimulation protocol 25 Jan 2025 · 0 repositories · arXiv:2501.16386
-
Improving Retrieval-Augmented Generation through Multi-Agent Reinforcement Learning 25 Jan 2025 · 1 repository · arXiv:2501.15228Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Inductive Biases for Zero-shot Systematic Generalization in Language-informed Reinforcement Learning 25 Jan 2025 · 1 repository · arXiv:2501.15270
-
Knowledge Hierarchy Guided Biological-Medical Dataset Distillation for Domain LLM Training 25 Jan 2025 · 0 repositories · arXiv:2501.15108
-
LLM Evaluation Based on Aerospace Manufacturing Expertise: Automated Generation and Multi-Model Question Answering 25 Jan 2025 · 0 repositories · arXiv:2501.17183
-
Mirage in the Eyes: Hallucination Attack on Multi-modal Large Language Models with Only Attention Sink 25 Jan 2025 · 0 repositories · arXiv:2501.15269
-
On Spectral Approach to the Synthesis of Shaping Filters 25 Jan 2025 · 0 repositories · arXiv:2501.15174
-
PolaFormer: Polarity-aware Linear Attention for Vision Transformers 25 Jan 2025 · 0 repositories · arXiv:2501.15061
-
ReInc: Scaling Training of Dynamic Graph Neural Networks 25 Jan 2025 · 0 repositories · arXiv:2501.15348
-
Reliable Pseudo-labeling via Optimal Transport with Attention for Short Text Clustering 25 Jan 2025 · 1 repository · arXiv:2501.15194
-
RotateKV: Accurate and Robust 2-Bit KV Cache Quantization for LLMs via Outlier-Aware Adaptive Rotations 25 Jan 2025 · 0 repositories · arXiv:2501.16383
-
SEAL: Scaling to Emphasize Attention for Long-Context Retrieval 25 Jan 2025 · 0 repositories · arXiv:2501.15225Syntology 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Speech Translation Refinement using Large Language Models 25 Jan 2025 · 1 repository · arXiv:2501.15090
-
SpikSSD: Better Extraction and Fusion for Object Detection with Spiking Neuron Networks 25 Jan 2025 · 1 repository · arXiv:2501.15151
-
Task-KV: Task-aware KV Cache Optimization via Semantic Differentiation of Attention Heads 25 Jan 2025 · 0 repositories · arXiv:2501.15113
-
Towards Robust Unsupervised Attention Prediction in Autonomous Driving 25 Jan 2025 · 1 repository · arXiv:2501.15045
-
TranStable: Towards Robust Pixel-level Online Video Stabilization by Jointing Transformer and CNN 25 Jan 2025 · 0 repositories · arXiv:2501.15138
-
TrustDataFilter:Leveraging Trusted Knowledge Base Data for More Effective Filtering of Unknown Information 25 Jan 2025 · 0 repositories · arXiv:2502.15714
-
Uni-Sign: Toward Unified Sign Language Understanding at Scale 25 Jan 2025 · 1 repository · arXiv:2501.15187Syntology official (archive's flag): 14 ran · 14 ran (of which 8 constructed an object rather than computing a result; 10 with no instrument failure: 1 honoured, 1 violated, 8 with no contract checked; 4 where Syntology's instrument failed) · 1 unverified (of 15 harvested samples) · 15 pointer-only (licence)
-
Using Large Language Models for education managements in Vietnamese with low resources 25 Jan 2025 · 0 repositories · arXiv:2501.15022
-
Utilizing Graph Neural Networks for Effective Link Prediction in Microservice Architectures 25 Jan 2025 · 0 repositories · arXiv:2501.15019
-
A Comprehensive Framework for Semantic Similarity Analysis of Human and AI-Generated Text Using Transformer Architectures and Ensemble Techniques 24 Jan 2025 · 0 repositories · arXiv:2501.14288
-
A Predictive Approach for Enhancing Accuracy in Remote Robotic Surgery Using Informer Model 24 Jan 2025 · 0 repositories · arXiv:2501.14678
-
Active Intracellular Mechanics: A Key to Cellular Function and Organization 24 Jan 2025 · 0 repositories · arXiv:2501.14538
-
Adaptive Progressive Attention Graph Neural Network for EEG Emotion Recognition 24 Jan 2025 · 0 repositories · arXiv:2501.14246
-
Advances in Set Function Learning: A Survey of Techniques and Applications 24 Jan 2025 · 0 repositories · arXiv:2501.14991
-
An Attentive Graph Agent for Topology-Adaptive Cyber Defence 24 Jan 2025 · 1 repository · arXiv:2501.14700
-
Automatic detection and prediction of nAMD activity change in retinal OCT using Siamese networks and Wasserstein Distance for ordinality 24 Jan 2025 · 1 repository · arXiv:2501.14323
-
BOLDreams: Dreaming with pruned in-silico fMRI Encoding Models of the Visual Cortex 24 Jan 2025 · 0 repositories · arXiv:2501.14854