Methods › General › Attention Mechanisms › Attention › Papers, page 14
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 14 of 316: papers 1,301 to 1,400 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Utilizing LLMs to Investigate the Disputed Role of Evidence in Electronic Cigarette Health Policy Formation in Australia and the UK 10 May 2025 · 0 repositories · arXiv:2505.06782
-
xGen-small Technical Report 10 May 2025 · 0 repositories · arXiv:2505.06496
-
A Noise-Resilient Semi-Supervised Graph Autoencoder for Overlapping Semantic Community Detection 9 May 2025 · 0 repositories · arXiv:2505.05965
-
A novel Neural-ODE model for the state of health estimation of lithium-ion battery using charging curve 9 May 2025 · 0 repositories · arXiv:2505.05803
-
A review of advancements in low-light image enhancement using deep learning 9 May 2025 · 0 repositories · arXiv:2505.05759
-
Accurate and Efficient Multivariate Time Series Forecasting via Offline Clustering 9 May 2025 · 0 repositories · arXiv:2505.05738
-
Achieving 3D Attention via Triplet Squeeze and Excitation Block 9 May 2025 · 0 repositories · arXiv:2505.05943
-
Adapting a Segmentation Foundation Model for Medical Image Classification 9 May 2025 · 0 repositories · arXiv:2505.06217
-
An empathic GPT-based chatbot to talk about mental disorders with Spanish teenagers 9 May 2025 · 0 repositories · arXiv:2505.05828
-
Attention on Multiword Expressions: A Multilingual Study of BERT-based Models with Regard to Idiomaticity and Microsyntax 9 May 2025 · 1 repository · arXiv:2505.06062
-
Camera Control at the Edge with Language Models for Scene Understanding 9 May 2025 · 0 repositories · arXiv:2505.06402
-
CellVerse: Do Large Language Models Really Understand Cell Biology? 9 May 2025 · 0 repositories · arXiv:2505.07865
-
Crowding Out The Noise: Algorithmic Collective Action Under Differential Privacy 9 May 2025 · 0 repositories · arXiv:2505.05707
-
DFEN: Dual Feature Equalization Network for Medical Image Segmentation 9 May 2025 · 1 repository · arXiv:2505.05913
-
DiGIT: Multi-Dilated Gated Encoder and Central-Adjacent Region Integrated Decoder for Temporal Action Detection Transformer 9 May 2025 · 1 repository · arXiv:2505.05711
-
Document Attribution: Examining Citation Relationships using Large Language Models 9 May 2025 · 0 repositories · arXiv:2505.06324
-
Dome-DETR: DETR with Density-Oriented Feature-Query Manipulation for Efficient Tiny Object Detection 9 May 2025 · 0 repositories · arXiv:2505.05741
-
Graph Laplacian Wavelet Transformer via Learnable Spectral Decomposition 9 May 2025 · 0 repositories · arXiv:2505.07862
-
Healthy LLMs? Benchmarking LLM Knowledge of UK Government Public Health Information 9 May 2025 · 0 repositories · arXiv:2505.06046
-
Let Humanoids Hike! Integrative Skill Development on Complex Trails 9 May 2025 · 0 repositories · arXiv:2505.06218
-
Modeling Multi-Hop Semantic Paths for Recommendation in Heterogeneous Information Networks 9 May 2025 · 0 repositories · arXiv:2505.05989
-
Multi-Agent Systems for Robotic Autonomy with LLMs 9 May 2025 · 0 repositories · arXiv:2505.05762
-
Multimodal Integrated Knowledge Transfer to Large Language Models through Preference Optimization with Biomedical Applications 9 May 2025 · 1 repository · arXiv:2505.05736
-
Register and CLS tokens yield a decoupling of local and global features in large ViTs 9 May 2025 · 0 repositories · arXiv:2505.05892
-
Short-circuiting Shortcuts: Mechanistic Investigation of Shortcuts in Text Classification 9 May 2025 · 1 repository · arXiv:2505.06032
-
Sparse Attention Remapping with Clustering for Efficient LLM Decoding on PIM 9 May 2025 · 0 repositories · arXiv:2505.05772
-
TopicVD: A Topic-Based Dataset of Video-Guided Multimodal Machine Translation for Documentaries 9 May 2025 · 0 repositories · arXiv:2505.05714
-
Topo-VM-UNetV2: Encoding Topology into Vision Mamba UNet for Polyp Segmentation 9 May 2025 · 0 repositories · arXiv:2505.06210
-
Towards Robust Few-Shot Text Classification Using Transformer Architectures and Dual Loss Strategies 9 May 2025 · 0 repositories · arXiv:2505.06145
-
Turbo-ICL: In-Context Learning-Based Turbo Equalization 9 May 2025 · 0 repositories · arXiv:2505.06175
-
UniSymNet: A Unified Symbolic Network Guided by Transformer 9 May 2025 · 0 repositories · arXiv:2505.06091
-
UniVLA: Learning to Act Anywhere with Task-centric Latent Actions 9 May 2025 · 1 repository · arXiv:2505.06111Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
What Is Next for LLMs? Next-Generation AI Computing Hardware Using Photonic Chips 9 May 2025 · 0 repositories · arXiv:2505.05794
-
AI Approaches to Qualitative and Quantitative News Analytics on NATO Unity 8 May 2025 · 0 repositories · arXiv:2505.06313
-
Benchmarking Vision, Language, & Action Models in Procedurally Generated, Open Ended Action Environments 8 May 2025 · 1 repository · arXiv:2505.05540
-
Biomed-DPT: Dual Modality Prompt Tuning for Biomedical Vision-Language Models 8 May 2025 · 1 repository · arXiv:2505.05189
-
Cardioformer: Advancing AI in ECG Analysis with Multi-Granularity Patching and ResNet 8 May 2025 · 1 repository · arXiv:2505.05538
-
CART-ELC: Oblique Decision Tree Induction via Exhaustive Search 8 May 2025 · 1 repository · arXiv:2505.05402
-
Enhanced Urdu Intent Detection with Large Language Models and Prototype-Informed Predictive Pipelines 8 May 2025 · 0 repositories · arXiv:2505.07857
-
FF-PNet: A Pyramid Network Based on Feature and Field for Brain Image Registration 8 May 2025 · 0 repositories · arXiv:2505.04938
-
GlyphMastero: A Glyph Encoder for High-Fidelity Scene Text Editing 8 May 2025 · 0 repositories · arXiv:2505.04915
-
Hide & Seek: Transformer Symmetries Obscure Sharpness & Riemannian Geometry Finds It 8 May 2025 · 0 repositories · arXiv:2505.05409
-
LiteLMGuard: Seamless and Lightweight On-Device Prompt Filtering for Safeguarding Small Language Models against Quantization-induced Risks and Vulnerabilities 8 May 2025 · 2 repositories · arXiv:2505.05619
-
Lost in OCR Translation? Vision-Based Approaches to Robust Document Retrieval 8 May 2025 · 0 repositories · arXiv:2505.05666
-
MDAA-Diff: CT-Guided Multi-Dose Adaptive Attention Diffusion Model for PET Denoising 8 May 2025 · 0 repositories · arXiv:2505.05112
-
MDE-Edit: Masked Dual-Editing for Multi-Object Image Editing via Diffusion Models 8 May 2025 · 0 repositories · arXiv:2505.05101
-
OXSeg: Multidimensional attention UNet-based lip segmentation using semi-supervised lip contours 8 May 2025 · 0 repositories · arXiv:2505.05531
-
Performance Evaluation of Large Language Models in Bangla Consumer Health Query Summarization 8 May 2025 · 0 repositories · arXiv:2505.05070
-
PillarMamba: Learning Local-Global Context for Roadside Point Cloud via Hybrid State Space Model 8 May 2025 · 0 repositories · arXiv:2505.05397
-
Pro2SAM: Mask Prompt to SAM with Grid Points for Weakly Supervised Object Localization 8 May 2025 · 0 repositories · arXiv:2505.04905
-
Progressive Inertial Poser: Progressive Real-Time Kinematic Chain Estimation for 3D Full-Body Pose from Three IMU Sensors 8 May 2025 · 0 repositories · arXiv:2505.05336
-
Prompted Meta-Learning for Few-shot Knowledge Graph Completion 8 May 2025 · 0 repositories · arXiv:2505.05684
-
QualBench: Benchmarking Chinese LLMs with Localized Professional Qualifications for Vertical Domain Evaluation 8 May 2025 · 0 repositories · arXiv:2505.05225
-
Research on Anomaly Detection Methods Based on Diffusion Models 8 May 2025 · 0 repositories · arXiv:2505.05137
-
Software Development Life Cycle Perspective: A Survey of Benchmarks for Code Large Language Models and Agents 8 May 2025 · 0 repositories · arXiv:2505.05283
-
SSH-Net: A Self-Supervised and Hybrid Network for Noisy Image Watermark Removal 8 May 2025 · 1 repository · arXiv:2505.05088
-
T-T: Table Transformer for Tagging-based Aspect Sentiment Triplet Extraction 8 May 2025 · 0 repositories · arXiv:2505.05271
-
This part looks alike this: identifying important parts of explained instances and prototypes 8 May 2025 · 1 repository · arXiv:2505.05597
-
Trading Under Uncertainty: A Distribution-Based Strategy for Futures Markets Using FutureQuant Transformer 8 May 2025 · 0 repositories · arXiv:2505.05595
-
Understanding In-context Learning of Addition via Activation Subspaces 8 May 2025 · 0 repositories · arXiv:2505.05145
-
USPR: Learning a Unified Solver for Profiled Routing 8 May 2025 · 0 repositories · arXiv:2505.05119
-
A Survey on Temporal Interaction Graph Representation Learning: Progress, Challenges, and Opportunities 7 May 2025 · 0 repositories · arXiv:2505.04461
-
ALFEE: Adaptive Large Foundation Model for EEG Representation 7 May 2025 · 0 repositories · arXiv:2505.06291
-
An Empirical Study of OpenAI API Discussions on Stack Overflow 7 May 2025 · 0 repositories · arXiv:2505.04084
-
An Enhanced YOLOv8 Model for Real-Time and Accurate Pothole Detection and Measurement 7 May 2025 · 0 repositories · arXiv:2505.04207
-
AS3D: 2D-Assisted Cross-Modal Understanding with Semantic-Spatial Scene Graphs for 3D Visual Grounding 7 May 2025 · 1 repository · arXiv:2505.04058
-
Balancing Accuracy, Calibration, and Efficiency in Active Learning with Vision Transformers Under Label Noise 7 May 2025 · 0 repositories · arXiv:2505.04375
-
Benchmarking LLM Faithfulness in RAG with Evolving Leaderboards 7 May 2025 · 1 repository · arXiv:2505.04847Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample) · 1 pointer-only (licence)
-
Bringing legal knowledge to the public by constructing a legal question bank using large-scale pre-trained language model 7 May 2025 · 0 repositories · arXiv:2505.04132
-
CountDiffusion: Text-to-Image Synthesis with Training-Free Counting-Guidance Diffusion 7 May 2025 · 0 repositories · arXiv:2505.04347
-
DeCLIP: Decoupled Learning for Open-Vocabulary Dense Perception 7 May 2025 · 1 repository · arXiv:2505.04410Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
DOTA: Deformable Optimized Transformer Architecture for End-to-End Text Recognition with Retrieval-Augmented Generation 7 May 2025 · 0 repositories · arXiv:2505.04175
-
Enhanced SCanNet with CBAM and Dice Loss for Semantic Change Detection 7 May 2025 · 1 repository · arXiv:2505.04199
-
EvEnhancer: Empowering Effectiveness, Efficiency and Generalizability for Continuous Space-Time Video Super-Resolution with Events 7 May 2025 · 1 repository · arXiv:2505.04657
-
Fine-Tuning Large Language Models and Evaluating Retrieval Methods for Improved Question Answering on Building Codes 7 May 2025 · 0 repositories · arXiv:2505.04666
-
Flower Across Time and Media: Sentiment Analysis of Tang Song Poetry and Visual Correspondence 7 May 2025 · 0 repositories · arXiv:2505.04785
-
GAPrompt: Geometry-Aware Point Cloud Prompt for 3D Vision Model 7 May 2025 · 1 repository · arXiv:2505.04119Syntology official (archive's flag): 3 ran · 3 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
GASCADE: Grouped Summarization of Adverse Drug Event for Enhanced Cancer Pharmacovigilance 7 May 2025 · 1 repository · arXiv:2505.04284
-
HDiffTG: A Lightweight Hybrid Diffusion-Transformer-GCN Architecture for 3D Human Pose Estimation 7 May 2025 · 1 repository · arXiv:2505.04276
-
HiPerRAG: High-Performance Retrieval Augmented Generation for Scientific Insights 7 May 2025 · 0 repositories · arXiv:2505.04846
-
Identities are not Interchangeable: The Problem of Overgeneralization in Fair Machine Learning 7 May 2025 · 0 repositories · arXiv:2505.04038
-
Image Restoration via Multi-domain Learning 7 May 2025 · 1 repository · arXiv:2505.05504
-
Integrated Image Reconstruction and Target Recognition based on Deep Learning Technique 7 May 2025 · 0 repositories · arXiv:2505.04836
-
Lay-Your-Scene: Natural Scene Layout Generation with Diffusion Transformers 7 May 2025 · 0 repositories · arXiv:2505.04718
-
Learning from Similarity Proportion Loss for Classifying Skeletal Muscle Recovery Stages 7 May 2025 · 0 repositories · arXiv:2505.04150
-
Lightweight RGB-D Salient Object Detection from a Speed-Accuracy Tradeoff Perspective 7 May 2025 · 1 repository · arXiv:2505.04758
-
LLM-e Guess: Can LLMs Capabilities Advance Without Hardware Progress? 7 May 2025 · 1 repository · arXiv:2505.04075
-
LONGER: Scaling Up Long Sequence Modeling in Industrial Recommenders 7 May 2025 · 0 repositories · arXiv:2505.04421
-
M2Rec: Multi-scale Mamba for Efficient Sequential Recommendation 7 May 2025 · 0 repositories · arXiv:2505.04445
-
Multi-Granular Attention based Heterogeneous Hypergraph Neural Network 7 May 2025 · 0 repositories · arXiv:2505.04340
-
Multi-turn Consistent Image Editing 7 May 2025 · 0 repositories · arXiv:2505.04320
-
Object-Shot Enhanced Grounding Network for Egocentric Video 7 May 2025 · 1 repository · arXiv:2505.04270Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 1 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 10 unverified (of 17 harvested samples) · 5 pointer-only (licence)
-
Occupancy World Model for Robots 7 May 2025 · 0 repositories · arXiv:2505.05512
-
ORBIT-2: Scaling Exascale Vision Foundation Models for Weather and Climate Downscaling 7 May 2025 · 0 repositories · arXiv:2505.04802
-
Osiris: A Lightweight Open-Source Hallucination Detection System 7 May 2025 · 0 repositories · arXiv:2505.04844
-
Personalized Risks and Regulatory Strategies of Large Language Models in Digital Advertising 7 May 2025 · 0 repositories · arXiv:2505.04665
-
Pose Estimation for Intra-cardiac Echocardiography Catheter via AI-Based Anatomical Understanding 7 May 2025 · 0 repositories · arXiv:2505.07851
-
Red Teaming the Mind of the Machine: A Systematic Evaluation of Prompt Injection and Jailbreak Vulnerabilities in LLMs 7 May 2025 · 0 repositories · arXiv:2505.04806
-
Reliable Disentanglement Multi-view Learning Against View Adversarial Attacks 7 May 2025 · 1 repository · arXiv:2505.04046
-
Retrieval Augmented Generation Evaluation for Health Documents 7 May 2025 · 0 repositories · arXiv:2505.04680