Methods › General › Attention Mechanisms › Attention › Papers, page 120
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 120 of 316: papers 11,901 to 12,000 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
BKDSNN: Enhancing the Performance of Learning-based Spiking Neural Networks Training with Blurred Knowledge Distillation 12 Jul 2024 · 1 repository · arXiv:2407.09083Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Bora: Biomedical Generalist Video Generation Model 12 Jul 2024 · 0 repositories · arXiv:2407.08944
-
DANIEL: A fast Document Attention Network for Information Extraction and Labelling of handwritten documents 12 Jul 2024 · 1 repository · arXiv:2407.09103
-
Deep Attention Driven Reinforcement Learning (DAD-RL) for Autonomous Decision-Making in Dynamic Environment 12 Jul 2024 · 1 repository · arXiv:2407.08932
-
Deep Bag-of-Words Model: An Efficient and Interpretable Relevance Architecture for Chinese E-Commerce 12 Jul 2024 · 0 repositories · arXiv:2407.09395
-
Detect, Investigate, Judge and Determine: A Novel LLM-based Framework for Few-shot Fake News Detection 12 Jul 2024 · 0 repositories · arXiv:2407.08952
-
DroneMOT: Drone-based Multi-Object Tracking Considering Detection Difficulties and Simultaneous Moving of Drones and Objects 12 Jul 2024 · 1 repository · arXiv:2407.09051
-
Enhancing Depressive Post Detection in Bangla: A Comparative Study of TF-IDF, BERT and FastText Embeddings 12 Jul 2024 · 0 repositories · arXiv:2407.09187
-
Exploring State Space and Reasoning by Elimination in Tsetlin Machines 12 Jul 2024 · 0 repositories · arXiv:2407.09162
-
Global Attention-Guided Dual-Domain Point Cloud Feature Learning for Classification and Segmentation 12 Jul 2024 · 0 repositories · arXiv:2407.08994
-
Heterogeneous Subgraph Network with Prompt Learning for Interpretable Depression Detection on Social Media 12 Jul 2024 · 0 repositories · arXiv:2407.09019
-
Human-like Episodic Memory for Infinite Context LLMs 12 Jul 2024 · 1 repository · arXiv:2407.09450Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 13 harvested samples)
-
Hybrid Reconfigurable Intelligent Surface Enabled Over the Air Index Modulation 12 Jul 2024 · 0 repositories · arXiv:2407.08985
-
Inference Optimization of Foundation Models on AI Accelerators 12 Jul 2024 · 0 repositories · arXiv:2407.09111
-
Introducing VaDA: Novel Image Segmentation Model for Maritime Object Segmentation Using New Dataset 12 Jul 2024 · 0 repositories · arXiv:2407.09005
-
Leveraging large language models for nano synthesis mechanism explanation: solid foundations or mere conjectures? 12 Jul 2024 · 1 repository · arXiv:2407.08922
-
Movie Recommendation with Poster Attention via Multi-modal Transformer Feature Fusion 12 Jul 2024 · 0 repositories · arXiv:2407.09157
-
MUSCLE: A Model Update Strategy for Compatible LLM Evolution 12 Jul 2024 · 0 repositories · arXiv:2407.09435
-
Neural-based Video Compression on Solar Dynamics Observatory Images 12 Jul 2024 · 0 repositories · arXiv:2407.15730
-
PID: Physics-Informed Diffusion Model for Infrared Image Generation 12 Jul 2024 · 1 repository · arXiv:2407.09299Syntology official (archive's flag): 10 ran · 12 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 3 violated, 6 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 13 harvested samples) · 5 pointer-only (licence)
-
Refuse Whenever You Feel Unsafe: Improving Safety in LLMs via Decoupled Refusal Training 12 Jul 2024 · 2 repositories · arXiv:2407.09121Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Region Attention Transformer for Medical Image Restoration 12 Jul 2024 · 1 repository · arXiv:2407.09268
-
Revealing the Dark Secrets of Extremely Large Kernel ConvNets on Robustness 12 Jul 2024 · 1 repository · arXiv:2407.08972Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 1 pointer-only (licence)
-
Robustness of LLMs to Perturbations in Text 12 Jul 2024 · 0 repositories · arXiv:2407.08989
-
Self-Evolving GPT: A Lifelong Autonomous Experiential Learner 12 Jul 2024 · 0 repositories · arXiv:2407.08937
-
Self-Prompt Tuning: Enable Autonomous Role-Playing in LLMs 12 Jul 2024 · 0 repositories · arXiv:2407.08995
-
Show, Don't Tell: Evaluating Large Language Models Beyond Textual Understanding with ChildPlay 12 Jul 2024 · 1 repository · arXiv:2407.11068
-
STD-PLM: Understanding Both Spatial and Temporal Properties of Spatial-Temporal Data with PLM 12 Jul 2024 · 1 repository · arXiv:2407.09096
-
Surgical Text-to-Image Generation 12 Jul 2024 · 0 repositories · arXiv:2407.09230
-
TCAN: Animating Human Images with Temporally Consistent Pose Guidance using Diffusion Models 12 Jul 2024 · 0 repositories · arXiv:2407.09012
-
TelecomGPT: A Framework to Build Telecom-Specfic Large Language Models 12 Jul 2024 · 0 repositories · arXiv:2407.09424
-
TensorTEE: Unifying Heterogeneous TEE Granularity for Efficient Secure Collaborative Tensor Computing 12 Jul 2024 · 0 repositories · arXiv:2407.08903
-
The Heterophilic Graph Learning Handbook: Benchmarks, Models, Theoretical Analysis, Applications and Challenges 12 Jul 2024 · 0 repositories · arXiv:2407.09618
-
The Two Sides of the Coin: Hallucination Generation and Detection with LLMs as Evaluators for LLMs 12 Jul 2024 · 0 repositories · arXiv:2407.09152
-
Toward Automatic Group Membership Annotation for Group Fairness Evaluation 12 Jul 2024 · 0 repositories · arXiv:2407.08926
-
Adaptive Parametric Activation 11 Jul 2024 · 1 repository · arXiv:2407.08567
-
ADMM Based Semi-Structured Pattern Pruning Framework For Transformer 11 Jul 2024 · 0 repositories · arXiv:2407.08334
-
Let Network Decide What to Learn: Symbolic Music Understanding Model Based on Large-scale Adversarial Pre-training 11 Jul 2024 · 1 repository · arXiv:2407.08306
-
Beyond Benchmarks: Evaluating Embedding Model Similarity for Retrieval Augmented Generation Systems 11 Jul 2024 · 1 repository · arXiv:2407.08275
-
Brain Tumor Segmentation in MRI Images with 3D U-Net and Contextual Transformer 11 Jul 2024 · 0 repositories · arXiv:2407.08470
-
BraTS-PEDs: Results of the Multi-Consortium International Pediatric Brain Tumor Segmentation Challenge 2023 11 Jul 2024 · 0 repositories · arXiv:2407.08855
-
Converging Paradigms: The Synergy of Symbolic and Connectionist AI in LLM-Empowered Autonomous Agents 11 Jul 2024 · 0 repositories · arXiv:2407.08516
-
CXR-Agent: Vision-language models for chest X-ray interpretation with uncertainty aware radiology reporting 11 Jul 2024 · 0 repositories · arXiv:2407.08811
-
DMM: Disparity-guided Multispectral Mamba for Oriented Object Detection in Remote Sensing 11 Jul 2024 · 1 repository · arXiv:2407.08132
-
Enrich the content of the image Using Context-Aware Copy Paste 11 Jul 2024 · 0 repositories · arXiv:2407.08151
-
ERD: Exponential Retinex decomposition based on weak space and hybrid nonconvex regularization and its denoising application 11 Jul 2024 · 0 repositories · arXiv:2407.08498
-
Explicit-NeRF-QA: A Quality Assessment Database for Explicit NeRF Model Compression 11 Jul 2024 · 0 repositories · arXiv:2407.08165
-
fairBERTs: Erasing Sensitive Information Through Semantic and Fairness-aware Perturbations 11 Jul 2024 · 0 repositories · arXiv:2407.08189
-
FairDomain: Achieving Fairness in Cross-Domain Medical Image Segmentation and Classification 11 Jul 2024 · 1 repository · arXiv:2407.08813
-
Fault Diagnosis in Power Grids with Large Language Model 11 Jul 2024 · 0 repositories · arXiv:2407.08836
-
FlashAttention-3: Fast and Accurate Attention with Asynchrony and Low-precision 11 Jul 2024 · 2 repositories · arXiv:2407.08608Syntology official (archive's flag): 16 ran · 16 ran (of which 0 constructed an object rather than computing a result; 15 with no instrument failure: 2 honoured, 0 violated, 13 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 18 harvested samples) · 10 pointer-only (licence)
-
Generalized Face Anti-spoofing via Finer Domain Partition and Disentangling Liveness-irrelevant Factors 11 Jul 2024 · 1 repository · arXiv:2407.08243
-
GPT-4 is judged more human than humans in displaced and inverted Turing tests 11 Jul 2024 · 0 repositories · arXiv:2407.08853
-
GraphMamba: An Efficient Graph Structure Learning Vision Mamba for Hyperspectral Image Classification 11 Jul 2024 · 1 repository · arXiv:2407.08255
-
GTA: A Benchmark for General Tool Agents 11 Jul 2024 · 1 repository · arXiv:2407.08713
-
HDT: Hierarchical Document Transformer 11 Jul 2024 · 0 repositories · arXiv:2407.08330
-
Hierarchical Consensus-Based Multi-Agent Reinforcement Learning for Multi-Robot Cooperation Tasks 11 Jul 2024 · 0 repositories · arXiv:2407.08164
-
How does Burrows' Delta work on medieval Chinese poetic texts? 11 Jul 2024 · 0 repositories · arXiv:2407.08099
-
HRRPGraphNet: Make HRRPs to Be Graphs for Efficient Target Recognition 11 Jul 2024 · 1 repository · arXiv:2407.08236
-
Improving Dental Diagnostics: Enhanced Convolution with Spatial Attention Mechanism 11 Jul 2024 · 0 repositories · arXiv:2407.08114
-
Investigating LLMs as Voting Assistants via Contextual Augmentation: A Case Study on the European Parliament Elections 2024 11 Jul 2024 · 0 repositories · arXiv:2407.08495
-
Knowledge distillation to effectively attain both region-of-interest and global semantics from an image where multiple objects appear 11 Jul 2024 · 1 repository · arXiv:2407.08257
-
Latent Spaces Enable Transformer-Based Dose Prediction in Complex Radiotherapy Plans 11 Jul 2024 · 1 repository · arXiv:2407.08650
-
Lifelong Histopathology Whole Slide Image Retrieval via Distance Consistency Rehearsal 11 Jul 2024 · 1 repository · arXiv:2407.08153
-
Live2Diff: Live Stream Translation via Uni-directional Attention in Video Diffusion Models 11 Jul 2024 · 0 repositories · arXiv:2407.08701
-
LLMs' morphological analyses of complex FST-generated Finnish words 11 Jul 2024 · 1 repository · arXiv:2407.08269
-
Long-range Turbulence Mitigation: A Large-scale Dataset and A Coarse-to-fine Framework 11 Jul 2024 · 0 repositories · arXiv:2407.08377
-
MAVIS: Mathematical Visual Instruction Tuning with an Automatic Data Engine 11 Jul 2024 · 3 repositories · arXiv:2407.08739
-
Multi-scale gridded Gabor attention for cirrus segmentation 11 Jul 2024 · 0 repositories · arXiv:2407.08852
-
Multimodal contrastive learning for spatial gene expression prediction using histology images 11 Jul 2024 · 1 repository · arXiv:2407.08216Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 1 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Nonverbal Interaction Detection 11 Jul 2024 · 1 repository · arXiv:2407.08133Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 1 pointer-only (licence)
-
OMR-NET: a two-stage octave multi-scale residual network for screen content image compression 11 Jul 2024 · 0 repositories · arXiv:2407.08545
-
On the (In)Security of LLM App Stores 11 Jul 2024 · 0 repositories · arXiv:2407.08422
-
Performance-Barrier Event-Triggered Control of a Class of Reaction-Diffusion PDEs 11 Jul 2024 · 0 repositories · arXiv:2407.08178
-
Predicting Heart Failure with Attention Learning Techniques Utilizing Cardiovascular Data 11 Jul 2024 · 0 repositories · arXiv:2407.08289
-
Projecting Points to Axes: Oriented Object Detection via Point-Axis Representation 11 Jul 2024 · 1 repository · arXiv:2407.08489
-
Real-Time Anomaly Detection and Reactive Planning with Large Language Models 11 Jul 2024 · 0 repositories · arXiv:2407.08735
-
SALSA: Swift Adaptive Lightweight Self-Attention for Enhanced LiDAR Place Recognition 11 Jul 2024 · 1 repository · arXiv:2407.08260
-
ScaleDepth: Decomposing Metric Depth Estimation into Scale Prediction and Relative Depth Estimation 11 Jul 2024 · 1 repository · arXiv:2407.08187
-
SEED-Story: Multimodal Long Story Generation with Large Language Model 11 Jul 2024 · 1 repository · arXiv:2407.08683Syntology official (archive's flag): 14 ran · 14 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 6 where Syntology's instrument failed) · 3 unverified (of 17 harvested samples) · 17 pointer-only (licence)
-
Segmentation-guided Attention for Visual Question Answering from Remote Sensing Images 11 Jul 2024 · 0 repositories · arXiv:2407.08669
-
Several questions of visual generation in 2024 11 Jul 2024 · 0 repositories · arXiv:2407.18290
-
Skywork-Math: Data Scaling Laws for Mathematical Reasoning in Large Language Models -- The Story Goes On 11 Jul 2024 · 0 repositories · arXiv:2407.08348
-
SLRL: Structured Latent Representation Learning for Multi-view Clustering 11 Jul 2024 · 0 repositories · arXiv:2407.08340
-
Speculative RAG: Enhancing Retrieval Augmented Generation through Drafting 11 Jul 2024 · 0 repositories · arXiv:2407.08223
-
Spiking Tucker Fusion Transformer for Audio-Visual Zero-Shot Learning 11 Jul 2024 · 0 repositories · arXiv:2407.08130
-
SRPose: Two-view Relative Pose Estimation with Sparse Keypoints 11 Jul 2024 · 1 repository · arXiv:2407.08199
-
stEnTrans: Transformer-based deep learning for spatial transcriptomics enhancement 11 Jul 2024 · 1 repository · arXiv:2407.08224
-
Synthetic Electroretinogram Signal Generation Using Conditional Generative Adversarial Network for Enhancing Classification of Autism Spectrum Disorder 11 Jul 2024 · 0 repositories · arXiv:2407.08166
-
The Synergy between Data and Multi-Modal Large Language Models: A Survey from Co-Development Perspective 11 Jul 2024 · 1 repository · arXiv:2407.08583
-
TractGraphFormer: Anatomically Informed Hybrid Graph CNN-Transformer Network for Classification from Diffusion MRI Tractography 11 Jul 2024 · 0 repositories · arXiv:2407.08883
-
Vox Populi, Vox AI? Using Language Models to Estimate German Public Opinion 11 Jul 2024 · 1 repository · arXiv:2407.08563
-
WildGaussians: 3D Gaussian Splatting in the Wild 11 Jul 2024 · 1 repository · arXiv:2407.08447Syntology official (archive's flag): 19 ran · 19 ran (of which 0 constructed an object rather than computing a result; 16 with no instrument failure: 2 honoured, 0 violated, 14 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 21 harvested samples) · 21 pointer-only (licence)
-
Adversarial Attacks and Defenses on Text-to-Image Diffusion Models: A Survey 10 Jul 2024 · 1 repository · arXiv:2407.15861
-
Arabic Automatic Story Generation with Large Language Models 10 Jul 2024 · 1 repository · arXiv:2407.07551
-
Attribute or Abstain: Large Language Models as Long Document Assistants 10 Jul 2024 · 1 repository · arXiv:2407.07799Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Deep(er) Reconstruction of Imaging Cherenkov Detectors with Swin Transformers and Normalizing Flow Models 10 Jul 2024 · 1 repository · arXiv:2407.07376
-
Deformable Feature Alignment and Refinement for Moving Infrared Dim-small Target Detection 10 Jul 2024 · 0 repositories · arXiv:2407.07289
-
Disentangling Masked Autoencoders for Unsupervised Domain Generalization 10 Jul 2024 · 1 repository · arXiv:2407.07544Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
DS@GT eRisk 2024: Sentence Transformers for Social Media Risk Assessment 10 Jul 2024 · 1 repository · arXiv:2407.08008