Methods › General › Attention Mechanisms › Attention › Papers, page 116
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 116 of 316: papers 11,501 to 11,600 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Patched RTC: evaluating LLMs for diverse software development tasks 23 Jul 2024 · 1 repository · arXiv:2407.16557
-
RedAgent: Red Teaming Large Language Models with Context-aware Autonomous Language Agent 23 Jul 2024 · 0 repositories · arXiv:2407.16667
-
Retrieval Augmented Generation or Long-Context LLMs? A Comprehensive Study and Hybrid Approach 23 Jul 2024 · 0 repositories · arXiv:2407.16833
-
Robust Privacy Amidst Innovation with Large Language Models Through a Critical Assessment of the Risks 23 Jul 2024 · 1 repository · arXiv:2407.16166
-
S-E Pipeline: A Vision Transformer (ViT) based Resilient Classification Pipeline for Medical Imaging Against Adversarial Attacks 23 Jul 2024 · 0 repositories · arXiv:2407.17587
-
SAFNet: Selective Alignment Fusion Network for Efficient HDR Imaging 23 Jul 2024 · 1 repository · arXiv:2407.16308Syntology official (archive's flag): 14 ran · 14 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 9 where Syntology's instrument failed) · 3 unverified (of 17 harvested samples) · 17 pointer-only (licence)
-
SEDS: Semantically Enhanced Dual-Stream Encoder for Sign Language Retrieval 23 Jul 2024 · 1 repository · arXiv:2407.16394Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
Self-Reasoning Assistant Learning for non-Abelian Gauge Fields Design 23 Jul 2024 · 0 repositories · arXiv:2407.16255
-
SINDER: Repairing the Singular Defects of DINOv2 23 Jul 2024 · 1 repository · arXiv:2407.16826Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample)
-
Spatiotemporal Graph Guided Multi-modal Network for Livestreaming Product Retrieval 23 Jul 2024 · 1 repository · arXiv:2407.16248
-
SPLAT: A framework for optimised GPU code-generation for SParse reguLar ATtention 23 Jul 2024 · 0 repositories · arXiv:2407.16847
-
Stock-driven Household Attention 23 Jul 2024 · 0 repositories · arXiv:2407.16141
-
Synthesizer Sound Matching Using Audio Spectrogram Transformers 23 Jul 2024 · 0 repositories · arXiv:2407.16643
-
TAPTRv2: Attention-based Position Update Improves Tracking Any Point 23 Jul 2024 · 0 repositories · arXiv:2407.16291
-
TookaBERT: A Step Forward for Persian NLU 23 Jul 2024 · 0 repositories · arXiv:2407.16382
-
TWIN V2: Scaling Ultra-Long User Behavior Sequence Modeling for Enhanced CTR Prediction at Kuaishou 23 Jul 2024 · 0 repositories · arXiv:2407.16357
-
When, Where, and What? A Novel Benchmark for Accident Anticipation and Localization with Large Language Models 23 Jul 2024 · 0 repositories · arXiv:2407.16277
-
An Empirical Comparison of Video Frame Sampling Methods for Multi-Modal RAG Retrieval 22 Jul 2024 · 0 repositories · arXiv:2408.03340
-
Attention Beats Linear for Fast Implicit Neural Representation Generation 22 Jul 2024 · 1 repository · arXiv:2407.15355
-
Attention Is All You Need But You Don't Need All Of It For Inference of Large Language Models 22 Jul 2024 · 0 repositories · arXiv:2407.15516
-
Bidirectional skip-frame prediction for video anomaly detection with intra-domain disparity-driven attention 22 Jul 2024 · 0 repositories · arXiv:2407.15424
-
Can GPT-4 learn to analyse moves in research article abstracts? 22 Jul 2024 · 0 repositories · arXiv:2407.15612
-
CarFormer: Self-Driving with Learned Object-Centric Representations 22 Jul 2024 · 0 repositories · arXiv:2407.15843
-
Counter Turing Test (CT²): Investigating AI-Generated Text Detection for Hindi -- Ranking LLMs based on Hindi AI Detectability Index (ADIₕᵢ) 22 Jul 2024 · 1 repository · arXiv:2407.15694
-
Customized Retrieval Augmented Generation and Benchmarking for EDA Tool Documentation QA 22 Jul 2024 · 0 repositories · arXiv:2407.15353
-
DiffX: Guide Your Layout to Cross-Modal Generative Modeling 22 Jul 2024 · 1 repository · arXiv:2407.15488
-
Dissecting Multiplication in Transformers: Insights into LLMs 22 Jul 2024 · 1 repository · arXiv:2407.15360
-
Efficient Multi-disparity Transformer for Light Field Image Super-resolution 22 Jul 2024 · 0 repositories · arXiv:2407.15329
-
Estimating Probability Densities with Transformer and Denoising Diffusion 22 Jul 2024 · 1 repository · arXiv:2407.15703Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 1 pointer-only (licence)
-
GFE-Mamba: Mamba-based AD Multi-modal Progression Assessment via Generative Feature Extraction from MCI 22 Jul 2024 · 1 repository · arXiv:2407.15719
-
Impacts of Anthropomorphizing Large Language Models in Learning Environments 22 Jul 2024 · 0 repositories · arXiv:2408.03945
-
Imposter.AI: Adversarial Attacks with Hidden Intentions towards Aligned Large Language Models 22 Jul 2024 · 0 repositories · arXiv:2407.15399
-
In Search of Quantum Advantage: Estimating the Number of Shots in Quantum Kernel Methods 22 Jul 2024 · 0 repositories · arXiv:2407.15776
-
Inverted Activations: Reducing Memory Footprint in Neural Network Training 22 Jul 2024 · 1 repository · arXiv:2407.15545
-
KWT-Tiny: RISC-V Accelerated, Embedded Keyword Spotting Transformer 22 Jul 2024 · 0 repositories · arXiv:2407.16026
-
Large-scale Time-Varying Portfolio Optimisation using Graph Attention Networks 22 Jul 2024 · 0 repositories · arXiv:2407.15532
-
Learning to Manipulate Anywhere: A Visual Generalizable Framework For Reinforcement Learning 22 Jul 2024 · 0 repositories · arXiv:2407.15815
-
Link Polarity Prediction from Sparse and Noisy Labels via Multiscale Social Balance 22 Jul 2024 · 1 repository · arXiv:2407.15643
-
LLMmap: Fingerprinting For Large Language Models 22 Jul 2024 · 1 repository · arXiv:2407.15847Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 7 unverified (of 16 harvested samples)
-
Local All-Pair Correspondence for Point Tracking 22 Jul 2024 · 2 repositories · arXiv:2407.15420Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 12 harvested samples) · 2 pointer-only (licence)
-
Mamba meets crack segmentation 22 Jul 2024 · 1 repository · arXiv:2407.15714
-
Mini-Sequence Transformer: Optimizing Intermediate Memory for Long Sequences Training 22 Jul 2024 · 1 repository · arXiv:2407.15892
-
MMInstruct: A High-Quality Multi-Modal Instruction Tuning Dataset with Extensive Diversity 22 Jul 2024 · 1 repository · arXiv:2407.15838Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
MoRSE: Bridging the Gap in Cybersecurity Expertise with Retrieval Augmented Generation 22 Jul 2024 · 0 repositories · arXiv:2407.15748
-
Movable Antenna-Enhanced Wireless Communications: General Architectures and Implementation Methods 22 Jul 2024 · 0 repositories · arXiv:2407.15448
-
Multi-Modality Co-Learning for Efficient Skeleton-based Action Recognition 22 Jul 2024 · 1 repository · arXiv:2407.15706Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 1 violated, 6 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
NV-Retriever: Improving text embedding models with effective hard-negative mining 22 Jul 2024 · 0 repositories · arXiv:2407.15831
-
Online Reduced-Order Data-Enabled Predictive Control 22 Jul 2024 · 0 repositories · arXiv:2407.16066
-
Poisoning with A Pill: Circumventing Detection in Federated Learning 22 Jul 2024 · 0 repositories · arXiv:2407.15389
-
Predicting the Best of N Visual Trackers 22 Jul 2024 · 1 repository · arXiv:2407.15707
-
Problems in AI, their roots in philosophy, and implications for science and society 22 Jul 2024 · 0 repositories · arXiv:2407.15671
-
Promises and Pitfalls of Generative Masked Language Modeling: Theoretical Framework and Practical Guidelines 22 Jul 2024 · 1 repository · arXiv:2407.21046
-
RadioRAG: Factual large language models for enhanced diagnostics in radiology using online retrieval augmented generation 22 Jul 2024 · 1 repository · arXiv:2407.15621
-
RazorAttention: Efficient KV Cache Compression Through Retrieval Heads 22 Jul 2024 · 0 repositories · arXiv:2407.15891
-
Region Guided Attention Network for Retinal Vessel Segmentation 22 Jul 2024 · 0 repositories · arXiv:2407.18970
-
RoadPainter: Points Are Ideal Navigators for Topology transformER 22 Jul 2024 · 0 repositories · arXiv:2407.15349
-
Robust Facial Reactions Generation: An Emotion-Aware Framework with Modality Compensation 22 Jul 2024 · 0 repositories · arXiv:2407.15798
-
SETTP: Style Extraction and Tunable Inference via Dual-level Transferable Prompt Learning 22 Jul 2024 · 0 repositories · arXiv:2407.15556
-
STAMP: Outlier-Aware Test-Time Adaptation with Stable Memory Replay 22 Jul 2024 · 1 repository · arXiv:2407.15773Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Stretching Each Dollar: Diffusion Training from Scratch on a Micro-Budget 22 Jul 2024 · 1 repository · arXiv:2407.15811Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples)
-
Test-Time Low Rank Adaptation via Confidence Maximization for Zero-Shot Generalization of Vision-Language Models 22 Jul 2024 · 1 repository · arXiv:2407.15913Syntology official (archive's flag): 2 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Towards Open-World Object-based Anomaly Detection via Self-Supervised Outlier Synthesis 22 Jul 2024 · 1 repository · arXiv:2407.15763
-
A Comparison of Language Modeling and Translation as Multilingual Pretraining Objectives 22 Jul 2024 · 1 repository · arXiv:2407.15489
-
Unlocking the Potential: Benchmarking Large Language Models in Water Engineering and Research 22 Jul 2024 · 0 repositories · arXiv:2407.21045
-
vTensor: Flexible Virtual Tensor Management for Efficient LLM Serving 22 Jul 2024 · 1 repository · arXiv:2407.15309
-
ZZU-NLP at SIGHAN-2024 dimABSA Task: Aspect-Based Sentiment Analysis with Coarse-to-Fine In-context Learning 22 Jul 2024 · 0 repositories · arXiv:2407.15341
-
A multi-level multi-label text classification dataset of 19th century Ottoman and Russian literary and critical texts 21 Jul 2024 · 0 repositories · arXiv:2407.15136
-
Answer, Assemble, Ace: Understanding How Transformers Answer Multiple Choice Questions 21 Jul 2024 · 0 repositories · arXiv:2407.15018Syntology 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Arondight: Red Teaming Large Vision Language Models with Auto-generated Multi-modal Jailbreak Prompts 21 Jul 2024 · 0 repositories · arXiv:2407.15050
-
CalibRBEV: Multi-Camera Calibration via ReversedBird's-eye-view Representations for Autonomous Driving. 21 Jul 2024 · 0 repositories
-
Decoding Multilingual Moral Preferences: Unveiling LLM's Biases Through the Moral Machine Experiment 21 Jul 2024 · 1 repository · arXiv:2407.15184
-
Efficient Visual Transformer by Learnable Token Merging 21 Jul 2024 · 1 repository · arXiv:2407.15219
-
Evidence-Based Temporal Fact Verification 21 Jul 2024 · 0 repositories · arXiv:2407.15291
-
ReAttention: Training-Free Infinite Context with Finite Attention Scope 21 Jul 2024 · 0 repositories · arXiv:2407.15176Syntology 5 ran (of which 4 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples)
-
Improving Prediction of Need for Mechanical Ventilation using Cross-Attention 21 Jul 2024 · 0 repositories · arXiv:2407.15885
-
Mask Guided Gated Convolution for Amodal Content Completion 21 Jul 2024 · 1 repository · arXiv:2407.15203
-
Point Transformer V3 Extreme: 1st Place Solution for 2024 Waymo Open Dataset Challenge in Semantic Segmentation 21 Jul 2024 · 0 repositories · arXiv:2407.15282
-
Prior Knowledge Integration via LLM Encoding and Pseudo Event Regulation for Video Moment Retrieval 21 Jul 2024 · 1 repository · arXiv:2407.15051Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 1 violated, 4 with no contract checked; 3 where Syntology's instrument failed) · 3 unverified (of 11 harvested samples) · 9 pointer-only (licence)
-
Semi-Supervised Pipe Video Temporal Defect Interval Localization 21 Jul 2024 · 0 repositories · arXiv:2407.15170
-
Token-Picker: Accelerating Attention in Text Generation with Minimized Memory Transfer via Probability Estimation 21 Jul 2024 · 0 repositories · arXiv:2407.15131
-
Toward Adaptive Reasoning in Large Language Models with Thought Rollback 21 Jul 2024 · 1 repository
-
Golden-Retriever: High-Fidelity Agentic Retrieval Augmented Generation for Industrial Knowledge Base 20 Jul 2024 · 0 repositories · arXiv:2408.00798
-
All Against Some: Efficient Integration of Large Language Models for Message Passing in Graph Neural Networks 20 Jul 2024 · 0 repositories · arXiv:2407.14996
-
Automatic Generation of Fashion Images using Prompting in Generative Machine Learning Models 20 Jul 2024 · 1 repository · arXiv:2407.14944
-
CORT: Class-Oriented Real-time Tracking for Embedded Systems 20 Jul 2024 · 0 repositories · arXiv:2407.17521
-
Differential Privacy of Cross-Attention with Provable Guarantee 20 Jul 2024 · 0 repositories · arXiv:2407.14717
-
Dual High-Order Total Variation Model for Underwater Image Restoration 20 Jul 2024 · 1 repository · arXiv:2407.14868
-
Enhancing Microgrid Performance Prediction with Attention-based Deep Learning Models 20 Jul 2024 · 0 repositories · arXiv:2407.14984
-
FairViT: Fair Vision Transformer via Adaptive Masking 20 Jul 2024 · 1 repository · arXiv:2407.14799
-
GaitMA: Pose-guided Multi-modal Feature Fusion for Gait Recognition 20 Jul 2024 · 0 repositories · arXiv:2407.14812
-
Improving Context-Aware Preference Modeling for Language Models 20 Jul 2024 · 0 repositories · arXiv:2407.14916
-
Mapping Patient Trajectories: Understanding and Visualizing Sepsis Prognostic Pathways from Patients Clinical Narratives 20 Jul 2024 · 0 repositories · arXiv:2407.21039
-
MetaAug: Meta-Data Augmentation for Post-Training Quantization 20 Jul 2024 · 2 repositories · arXiv:2407.14726
-
RGB2Point: 3D Point Cloud Generation from Single RGB Images 20 Jul 2024 · 0 repositories · arXiv:2407.14979
-
RoIPoly: Vectorized Building Outline Extraction Using Vertex and Logit Embeddings 20 Jul 2024 · 0 repositories · arXiv:2407.14920
-
Step-by-Step Reasoning to Solve Grid Puzzles: Where do LLMs Falter? 20 Jul 2024 · 1 repository · arXiv:2407.14790
-
Technical report: Improving the properties of molecules generated by LIMO 20 Jul 2024 · 0 repositories · arXiv:2407.14968
-
TraveLLM: Could you plan my new public transit route in face of a network disruption? 20 Jul 2024 · 0 repositories · arXiv:2407.14926
-
A Comparative Study of Transfer Learning for Emotion Recognition using CNN and Modified VGG16 Models 19 Jul 2024 · 0 repositories · arXiv:2407.14576
-
A Mirror Descent-Based Algorithm for Corruption-Tolerant Distributed Gradient Descent 19 Jul 2024 · 0 repositories · arXiv:2407.14111