Methods › General › Attention Mechanisms › Attention › Papers, page 67
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 67 of 316: papers 6,601 to 6,700 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Unseen Attack Detection in Software-Defined Networking Using a BERT-Based Large Language Model 9 Dec 2024 · 0 repositories · arXiv:2412.06239
-
VP-MEL: Visual Prompts Guided Multimodal Entity Linking 9 Dec 2024 · 0 repositories · arXiv:2412.06720
-
VQ4ALL: Efficient Neural Network Representation via a Universal Codebook 9 Dec 2024 · 0 repositories · arXiv:2412.06875
-
ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models 9 Dec 2024 · 0 repositories · arXiv:2412.06292
-
A Collaborative Multi-Agent Approach to Retrieval-Augmented Generation Across Diverse Data 8 Dec 2024 · 0 repositories · arXiv:2412.05838
-
A Two-Stage AI-Powered Motif Mining Method for Efficient Power System Topological Analysis 8 Dec 2024 · 0 repositories · arXiv:2412.05957
-
A4-Unet: Deformable Multi-Scale Attention Network for Brain Tumor Segmentation 8 Dec 2024 · 1 repository · arXiv:2412.06088
-
Are Clinical T5 Models Better for Clinical Text? 8 Dec 2024 · 1 repository · arXiv:2412.05845
-
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs 8 Dec 2024 · 1 repository · arXiv:2412.05819Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 1 violated, 3 with no contract checked; 5 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples) · 1 pointer-only (licence)
-
Curse of Attention: A Kernel-Based Perspective for Why Transformers Fail to Generalize on Time Series Forecasting and Beyond 8 Dec 2024 · 0 repositories · arXiv:2412.06061
-
Enhanced Computationally Efficient Long LoRA Inspired Perceiver Architectures for Auto-Regressive Language Modeling 8 Dec 2024 · 0 repositories · arXiv:2412.06106
-
Enhancing Content Representation for AR Image Quality Assessment Using Knowledge Distillation 8 Dec 2024 · 0 repositories · arXiv:2412.06003
-
Evaluating Robustness of LLMs on Crisis-Related Microblogs across Events, Information Types, and Linguistic Features 8 Dec 2024 · 0 repositories · arXiv:2412.10413
-
Fully Open Source Moxin-7B Technical Report 8 Dec 2024 · 1 repository · arXiv:2412.06845
-
GBR: Generative Bundle Refinement for High-fidelity Gaussian Splatting and Meshing 8 Dec 2024 · 0 repositories · arXiv:2412.05908
-
Heuristic-Induced Multimodal Risk Distribution Jailbreak Attack for Multimodal Large Language Models 8 Dec 2024 · 1 repository · arXiv:2412.05934
-
Imputation Matters: A Deeper Look into an Overlooked Step in Longitudinal Health and Behavior Sensing Research 8 Dec 2024 · 0 repositories · arXiv:2412.06018
-
KITE-DDI: A Knowledge graph Integrated Transformer Model for accurately predicting Drug-Drug Interaction Events from Drug SMILES and Biomedical Knowledge Graph 8 Dec 2024 · 0 repositories · arXiv:2412.05770
-
Language-Guided Image Tokenization for Generation 8 Dec 2024 · 0 repositories · arXiv:2412.05796
-
Learning to Correction: Explainable Feedback Generation for Visual Commonsense Reasoning Distractor 8 Dec 2024 · 1 repository · arXiv:2412.07801
-
Lightweight Spatial Embedding for Vision-based 3D Occupancy Prediction 8 Dec 2024 · 0 repositories · arXiv:2412.05976
-
LVS-Net: A Lightweight Vessels Segmentation Network for Retinal Image Analysis 8 Dec 2024 · 0 repositories · arXiv:2412.05968
-
M³-20M: A Large-Scale Multi-Modal Molecule Dataset for AI-driven Drug Design and Discovery 8 Dec 2024 · 1 repository · arXiv:2412.06847
-
Mixture-of-PageRanks: Replacing Long-Context with Real-Time, Sparse GraphRAG 8 Dec 2024 · 0 repositories · arXiv:2412.06078
-
Paddy Disease Detection and Classification Using Computer Vision Techniques: A Mobile Application to Detect Paddy Disease 8 Dec 2024 · 0 repositories · arXiv:2412.05996
-
Vision Transformer-based Semantic Communications With Importance-Aware Quantization 8 Dec 2024 · 0 repositories · arXiv:2412.06038
-
A Comparative Study on Code Generation with Transformers 7 Dec 2024 · 0 repositories · arXiv:2412.05749
-
BERTCaps: BERT Capsule for Persian Multi-Domain Sentiment Analysis 7 Dec 2024 · 0 repositories · arXiv:2412.05591
-
CharacterBox: Evaluating the Role-Playing Capabilities of LLMs in Text-Based Virtual Worlds 7 Dec 2024 · 1 repository · arXiv:2412.05631Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples)
-
Can the Rookies Cut the Tough Cookie? Exploring the Use of LLMs for SQL Equivalence Checking 7 Dec 2024 · 0 repositories · arXiv:2412.05561
-
Flex Attention: A Programming Model for Generating Optimized Attention Kernels 7 Dec 2024 · 0 repositories · arXiv:2412.05496
-
GAF-FusionNet: Multimodal ECG Analysis via Gramian Angular Fields and Split Attention 7 Dec 2024 · 1 repository · arXiv:2501.01960
-
Innovative Sentiment Analysis and Prediction of Stock Price Using FinBERT, GPT-4 and Logistic Regression: A Data-Driven Approach 7 Dec 2024 · 0 repositories · arXiv:2412.06837
-
Integrating YOLO11 and Convolution Block Attention Module for Multi-Season Segmentation of Tree Trunks and Branches in Commercial Apple Orchards 7 Dec 2024 · 0 repositories · arXiv:2412.05728
-
Jointly RS Image Deblurring and Super-Resolution with Adjustable-Kernel and Multi-Domain Attention 7 Dec 2024 · 1 repository · arXiv:2412.05696
-
KG-Retriever: Efficient Knowledge Indexing for Retrieval-Augmented Large Language Models 7 Dec 2024 · 1 repository · arXiv:2412.05547Syntology official: harvested, nothing ran · 0 ran · 3 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Learning Soft Driving Constraints from Vectorized Scene Embeddings while Imitating Expert Trajectories 7 Dec 2024 · 0 repositories · arXiv:2412.05717
-
LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods 7 Dec 2024 · 1 repository · arXiv:2412.05579
-
M³PC: Test-time Model Predictive Control for Pretrained Masked Trajectory Model 7 Dec 2024 · 1 repository · arXiv:2412.05675Syntology official (archive's flag): 2 ran · 2 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; every one of the 2 samples that ran constructed an object rather than computing a result (of 3 harvested samples)
-
Multimodal Biometric Authentication Using Camera-Based PPG and Fingerprint Fusion 7 Dec 2024 · 0 repositories · arXiv:2412.05660
-
On the Expressive Power of Modern Hopfield Networks 7 Dec 2024 · 0 repositories · arXiv:2412.05562
-
PrivAgent: Agentic-based Red-teaming for LLM Privacy Leakage 7 Dec 2024 · 1 repository · arXiv:2412.05734
-
RefSAM3D: Adapting SAM with Cross-modal Reference for 3D Medical Image Segmentation 7 Dec 2024 · 0 repositories · arXiv:2412.05605
-
Shifting NER into High Gear: The Auto-AdvER Approach 7 Dec 2024 · 0 repositories · arXiv:2412.05655
-
SLA Management in Reconfigurable Multi-Agent RAG: A Systems Approach to Question Answering 7 Dec 2024 · 0 repositories · arXiv:2412.06832
-
SMI-Editor: Edit-based SMILES Language Model with Fragment-level Supervision 7 Dec 2024 · 0 repositories · arXiv:2412.05569
-
STONet: A novel neural operator for modeling solute transport in micro-cracked reservoirs 7 Dec 2024 · 1 repository · arXiv:2412.05576
-
Towards 3D Acceleration for low-power Mixture-of-Experts and Multi-Head Attention Spiking Transformers 7 Dec 2024 · 0 repositories · arXiv:2412.05540
-
Towards Learning to Reason: Comparing LLMs with Neuro-Symbolic on Arithmetic Relations in Abstract Reasoning 7 Dec 2024 · 2 repositories · arXiv:2412.05586Syntology official (archive's flag): 1 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 2 pointer-only (licence)
-
Trimming Down Large Spiking Vision Transformers via Heterogeneous Quantization Search 7 Dec 2024 · 0 repositories · arXiv:2412.05505
-
UMSPU: Universal Multi-Size Phase Unwrapping via Mutual Self-Distillation and Adaptive Boosting Ensemble Segmenters 7 Dec 2024 · 0 repositories · arXiv:2412.05584
-
WavFusion: Towards wav2vec 2.0 Multimodal Speech Emotion Recognition 7 Dec 2024 · 0 repositories · arXiv:2412.05558
-
100% Elimination of Hallucinations on RAGTruth for GPT-4 and GPT-3.5 Turbo 6 Dec 2024 · 0 repositories · arXiv:2412.05223
-
TOBUGraph: Knowledge Graph-Based Retrieval for Enhanced LLM Performance Beyond RAG 6 Dec 2024 · 0 repositories · arXiv:2412.05447
-
Addressing Attribute Leakages in Diffusion-based Image Editing without Training 6 Dec 2024 · 0 repositories · arXiv:2412.04715
-
Are Frontier Large Language Models Suitable for Q&A in Science Centres? 6 Dec 2024 · 0 repositories · arXiv:2412.05200
-
BEExformer: A Fast Inferencing Transformer Architecture via Binarization with Multiple Early Exits 6 Dec 2024 · 0 repositories · arXiv:2412.05225
-
BIAS: A Body-based Interpretable Active Speaker Approach 6 Dec 2024 · 1 repository · arXiv:2412.05150
-
Continuous Video Process: Modeling Videos as Continuous Multi-Dimensional Processes for Video Prediction 6 Dec 2024 · 0 repositories · arXiv:2412.04929
-
DHIL-GT: Scalable Graph Transformer with Decoupled Hierarchy Labeling 6 Dec 2024 · 0 repositories · arXiv:2412.04738
-
Differentially Private Random Feature Model 6 Dec 2024 · 1 repository · arXiv:2412.04785
-
Dirac-Equation Signal Processing: Physics Boosts Topological Machine Learning 6 Dec 2024 · 0 repositories · arXiv:2412.05132
-
Enhancing Cross-Language Code Translation via Task-Specific Embedding Alignment in Retrieval-Augmented Generation 6 Dec 2024 · 0 repositories · arXiv:2412.05159
-
Enhancing LLMs for Impression Generation in Radiology Reports through a Multi-Agent System 6 Dec 2024 · 0 repositories · arXiv:2412.06828
-
Feature Group Tabular Transformer: A Novel Approach to Traffic Crash Modeling and Causality Analysis 6 Dec 2024 · 0 repositories · arXiv:2412.06825
-
How to Squeeze An Explanation Out of Your Model 6 Dec 2024 · 0 repositories · arXiv:2412.05134
-
IterL2Norm: Fast Iterative L2-Normalization 6 Dec 2024 · 0 repositories · arXiv:2412.04778
-
Ltri-LLM: Streaming Long Context Inference for LLMs with Training-Free Dynamic Triangular Attention Pattern 6 Dec 2024 · 0 repositories · arXiv:2412.04757
-
Megatron: Evasive Clean-Label Backdoor Attacks against Vision Transformer 6 Dec 2024 · 0 repositories · arXiv:2412.04776
-
NLP-ADBench: NLP Anomaly Detection Benchmark 6 Dec 2024 · 1 repository · arXiv:2412.04784
-
PCTreeS: 3D Point Cloud Tree Species Classification Using Airborne LiDAR Images 6 Dec 2024 · 0 repositories · arXiv:2412.04714
-
Privacy-Preserving Retrieval-Augmented Generation with Differential Privacy 6 Dec 2024 · 0 repositories · arXiv:2412.04697
-
QueEn: A Large Language Model for Quechua-English Translation 6 Dec 2024 · 0 repositories · arXiv:2412.05184
-
StableVC: Style Controllable Zero-Shot Voice Conversion with Conditional Flow Matching 6 Dec 2024 · 0 repositories · arXiv:2412.04724
-
Superpixel Tokenization for Vision Transformers: Preserving Semantic Integrity in Visual Tokens 6 Dec 2024 · 1 repository · arXiv:2412.04680Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Tabular data generation with tensor contraction layers and transformers 6 Dec 2024 · 1 repository · arXiv:2412.05390
-
Unsupervised Segmentation by Diffusing, Walking and Cutting 6 Dec 2024 · 0 repositories · arXiv:2412.04678
-
Verb Mirage: Unveiling and Assessing Verb Concept Hallucinations in Multimodal Large Language Models 6 Dec 2024 · 0 repositories · arXiv:2412.04939
-
Addressing Hallucinations with RAG and NMISS in Italian Healthcare LLM Chatbots 5 Dec 2024 · 0 repositories · arXiv:2412.04235
-
ARTeFACT: Benchmarking Segmentation Models on Diverse Analogue Media Damage 5 Dec 2024 · 0 repositories · arXiv:2412.04580
-
Automated LaTeX Code Generation from Handwritten Math Expressions Using Vision Transformer 5 Dec 2024 · 0 repositories · arXiv:2412.03853
-
Comprehensive Audio Query Handling System with Integrated Expert Models and Contextual Understanding 5 Dec 2024 · 0 repositories · arXiv:2412.03980
-
Cross-Self KV Cache Pruning for Efficient Vision-Language Inference 5 Dec 2024 · 1 repository · arXiv:2412.04652Syntology official (archive's flag): 1 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 1 pointer-only (licence)
-
Cubify Anything: Scaling Indoor 3D Object Detection 5 Dec 2024 · 1 repository · arXiv:2412.04458
-
DEIM: DETR with Improved Matching for Fast Convergence 5 Dec 2024 · 1 repository · arXiv:2412.04234Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 1 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 13 harvested samples) · 13 pointer-only (licence)
-
Dynamic Graph Representation with Contrastive Learning for Financial Market Prediction: Integrating Temporal Evolution and Static Relations 5 Dec 2024 · 1 repository · arXiv:2412.04034
-
Exploring AI Text Generation, Retrieval-Augmented Generation, and Detection Technologies: a Comprehensive Overview 5 Dec 2024 · 0 repositories · arXiv:2412.03933
-
Exploring Real&Synthetic Dataset and Linear Attention in Image Restoration 5 Dec 2024 · 0 repositories · arXiv:2412.03814
-
Extractive Structures Learned in Pretraining Enable Generalization on Finetuned Facts 5 Dec 2024 · 1 repository · arXiv:2412.04614Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
Final-Model-Only Data Attribution with a Unifying View of Gradient-Based Methods 5 Dec 2024 · 0 repositories · arXiv:2412.03906Syntology 0 ran · 7 unverified (of 7 harvested samples)
-
Florence-VL: Enhancing Vision-Language Models with Generative Vision Encoder and Depth-Breadth Fusion 5 Dec 2024 · 1 repository · arXiv:2412.04424
-
GRAF: Graph Retrieval Augmented by Facts for Romanian Legal Multi-Choice Question Answering 5 Dec 2024 · 0 repositories · arXiv:2412.04119
-
Guidance is All You Need: Temperature-Guided Reasoning in Large Language Models 5 Dec 2024 · 0 repositories · arXiv:2412.06822
-
HEAL: Hierarchical Embedding Alignment Loss for Improved Retrieval and Representation Learning 5 Dec 2024 · 1 repository · arXiv:2412.04661
-
How Good is ChatGPT in Giving Adaptive Guidance Using Knowledge Graphs in E-Learning Environments? 5 Dec 2024 · 0 repositories · arXiv:2412.03856
-
HyperDefect-YOLO: Enhance YOLO with HyperGraph Computation for Industrial Defect Detection 5 Dec 2024 · 0 repositories · arXiv:2412.03969
-
Integrating Various Software Artifacts for Better LLM-based Bug Localization and Program Repair 5 Dec 2024 · 1 repository · arXiv:2412.03905
-
Learning to Hash for Recommendation: A Survey 5 Dec 2024 · 1 repository · arXiv:2412.03875
-
M³D: A Multimodal, Multilingual and Multitask Dataset for Grounded Document-level Information Extraction 5 Dec 2024 · 1 repository · arXiv:2412.04026
-
MEMO: Memory-Guided Diffusion for Expressive Talking Video Generation 5 Dec 2024 · 0 repositories · arXiv:2412.04448