Methods › General › Attention Mechanisms › Attention › Papers, page 97
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 97 of 316: papers 9,601 to 9,700 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Cross-Target Stance Detection: A Survey of Techniques, Datasets, and Challenges 20 Sep 2024 · 0 repositories · arXiv:2409.13594
-
Data Augmentation for Sequential Recommendation: A Survey 20 Sep 2024 · 1 repository · arXiv:2409.13545
-
DS2TA: Denoising Spiking Transformer with Attenuated Spatiotemporal Attention 20 Sep 2024 · 0 repositories · arXiv:2409.15375
-
EMMeTT: Efficient Multimodal Machine Translation Training 20 Sep 2024 · 0 repositories · arXiv:2409.13523
-
Enhancing Large Language Models with Domain-specific Retrieval Augment Generation: A Case Study on Long-form Consumer Health Question Answering in Ophthalmology 20 Sep 2024 · 0 repositories · arXiv:2409.13902
-
FAIR GPT: A virtual consultant for research data management in ChatGPT 20 Sep 2024 · 1 repository · arXiv:2410.07108
-
GAProtoNet: A Multi-head Graph Attention-based Prototypical Network for Interpretable Text Classification 20 Sep 2024 · 1 repository · arXiv:2409.13312
-
GASA-UNet: Global Axial Self-Attention U-Net for 3D Medical Image Segmentation 20 Sep 2024 · 0 repositories · arXiv:2409.13146
-
High-dimensional learning of narrow neural networks 20 Sep 2024 · 0 repositories · arXiv:2409.13904
-
HUT: A More Computation Efficient Fine-Tuning Method With Hadamard Updated Transformation 20 Sep 2024 · 0 repositories · arXiv:2409.13501
-
Imagine yourself: Tuning-Free Personalized Image Generation 20 Sep 2024 · 0 repositories · arXiv:2409.13346
-
Improved Unet brain tumor image segmentation based on GSConv module and ECA attention mechanism 20 Sep 2024 · 0 repositories · arXiv:2409.13626
-
Large Language Model Should Understand Pinyin for Chinese ASR Error Correction 20 Sep 2024 · 0 repositories · arXiv:2409.13262
-
Learning to Compare Hardware Designs for High-Level Synthesis 20 Sep 2024 · 1 repository · arXiv:2409.13138
-
Leveraging Knowledge Graphs and LLMs to Support and Monitor Legislative Systems 20 Sep 2024 · 0 repositories · arXiv:2409.13252
-
Localized Gaussians as Self-Attention Weights for Point Clouds Correspondence 20 Sep 2024 · 0 repositories · arXiv:2409.13291
-
Multiscale Encoder and Omni-Dimensional Dynamic Convolution Enrichment in nnU-Net for Brain Tumor Segmentation 20 Sep 2024 · 1 repository · arXiv:2409.13229
-
Neural-Symbolic Collaborative Distillation: Advancing Small Language Models for Complex Reasoning Tasks 20 Sep 2024 · 1 repository · arXiv:2409.13203Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
Occupancy-Based Dual Contouring 20 Sep 2024 · 1 repository · arXiv:2409.13418
-
On-Device Collaborative Language Modeling via a Mixture of Generalists and Specialists 20 Sep 2024 · 1 repository · arXiv:2409.13931
-
OneBEV: Using One Panoramic Image for Bird's-Eye-View Semantic Mapping 20 Sep 2024 · 1 repository · arXiv:2409.13912
-
Persistent Backdoor Attacks in Continual Learning 20 Sep 2024 · 0 repositories · arXiv:2409.13864
-
PlainUSR: Chasing Faster ConvNet for Efficient Super-Resolution 20 Sep 2024 · 1 repository · arXiv:2409.13435
-
PLOT: Text-based Person Search with Part Slot Attention for Corresponding Part Discovery 20 Sep 2024 · 0 repositories · arXiv:2409.13475
-
Prompting Large Language Models for Supporting the Differential Diagnosis of Anemia 20 Sep 2024 · 0 repositories · arXiv:2409.15377
-
Robust Salient Object Detection on Compressed Images Using Convolutional Neural Networks 20 Sep 2024 · 0 repositories · arXiv:2409.13464
-
Scalable Multi-agent Reinforcement Learning for Factory-wide Dynamic Scheduling 20 Sep 2024 · 0 repositories · arXiv:2409.13571
-
ShizishanGPT: An Agricultural Large Language Model Integrating Tools and Resources 20 Sep 2024 · 1 repository · arXiv:2409.13537
-
STOP! Benchmarking Large Language Models with Sensitivity Testing on Offensive Progressions 20 Sep 2024 · 1 repository · arXiv:2409.13843
-
Tackling fluffy clouds: field boundaries detection using time series of S2 and/or S1 imagery 20 Sep 2024 · 1 repository · arXiv:2409.13568
-
TalkMosaic: Interactive PhotoMosaic with Multi-modal LLM Q&A Interactions 20 Sep 2024 · 0 repositories · arXiv:2409.13941
-
SKIntern: Internalizing Symbolic Knowledge for Distilling Better CoT Capabilities into Small Language Models 20 Sep 2024 · 1 repository · arXiv:2409.13183Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples)
-
The Impact of Large Language Models in Academia: from Writing to Speaking 20 Sep 2024 · 0 repositories · arXiv:2409.13686
-
Transfer Learning and Double U-Net Empowered Wave Propagation Model in Complex Indoor Environment 20 Sep 2024 · 0 repositories · arXiv:2409.13833
-
Transformers in Uniform TC⁰ 20 Sep 2024 · 0 repositories · arXiv:2409.13629
-
ViTGuard: Attention-aware Detection against Adversarial Examples for Vision Transformer 20 Sep 2024 · 0 repositories · arXiv:2409.13828
-
3DTopia-XL: Scaling High-quality 3D Asset Generation via Primitive Diffusion 19 Sep 2024 · 1 repository · arXiv:2409.12957Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
A dynamic vision sensor object recognition model based on trainable event-driven convolution and spiking attention mechanism 19 Sep 2024 · 0 repositories · arXiv:2409.12691
-
A Novel Perspective for Multi-modal Multi-label Skin Lesion Classification 19 Sep 2024 · 0 repositories · arXiv:2409.12390
-
CritiPrefill: A Segment-wise Criticality-based Approach for Prefilling Acceleration in LLMs 19 Sep 2024 · 1 repository · arXiv:2409.12490Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Development of a pulse oximeter robust to measurement errors, with the ability to estimate heartrate and transmit data to smartphones 19 Sep 2024 · 0 repositories · arXiv:2409.12818
-
EEG-based Decoding of Selective Visual Attention in Superimposed Videos 19 Sep 2024 · 1 repository · arXiv:2409.12562
-
EFA-YOLO: An Efficient Feature Attention Model for Fire and Flame Detection 19 Sep 2024 · 0 repositories · arXiv:2409.12635
-
End-to-end Open-vocabulary Video Visual Relationship Detection using Multi-modal Prompting 19 Sep 2024 · 0 repositories · arXiv:2409.12499
-
Enhancing E-commerce Product Title Translation with Retrieval-Augmented Generation and Large Language Models 19 Sep 2024 · 0 repositories · arXiv:2409.12880
-
Enhancing Performance and Scalability of Large-Scale Recommendation Systems with Jagged Flash Attention 19 Sep 2024 · 0 repositories · arXiv:2409.15373
-
Enhancing TinyBERT for Financial Sentiment Analysis Using GPT-Augmented FinBERT Distillation 19 Sep 2024 · 1 repository · arXiv:2409.18999
-
Evaluating Image Hallucination in Text-to-Image Generation with Question-Answering 19 Sep 2024 · 1 repository · arXiv:2409.12784
-
Exploring Large Language Models for Product Attribute Value Identification 19 Sep 2024 · 0 repositories · arXiv:2409.12695
-
Fact, Fetch, and Reason: A Unified Evaluation of Retrieval-Augmented Generation 19 Sep 2024 · 2 repositories · arXiv:2409.12941
-
FedAT: Federated Adversarial Training for Distributed Insider Threat Detection 19 Sep 2024 · 0 repositories · arXiv:2409.13083
-
Fine Tuning Large Language Models for Medicine: The Role and Importance of Direct Preference Optimization 19 Sep 2024 · 0 repositories · arXiv:2409.12741
-
FoME: A Foundation Model for EEG using Adaptive Temporal-Lateral Attention Scaling 19 Sep 2024 · 0 repositories · arXiv:2409.12454
-
Geometry-Constrained EEG Channel Selection for Brain-Assisted Speech Enhancement 19 Sep 2024 · 0 repositories · arXiv:2409.12520
-
HSIGene: A Foundation Model For Hyperspectral Image Generation 19 Sep 2024 · 1 repository · arXiv:2409.12470
-
Hybrid Ensemble Deep Graph Temporal Clustering for Spatiotemporal Data 19 Sep 2024 · 0 repositories · arXiv:2409.12590
-
Incremental and Data-Efficient Concept Formation to Support Masked Word Prediction 19 Sep 2024 · 0 repositories · arXiv:2409.12440
-
Is it Still Fair? A Comparative Evaluation of Fairness Algorithms through the Lens of Covariate Drift 19 Sep 2024 · 0 repositories · arXiv:2409.12428
-
KLDD: Kalman Filter based Linear Deformable Diffusion Model in Retinal Image Segmentation 19 Sep 2024 · 0 repositories · arXiv:2410.02808
-
KnowFormer: Revisiting Transformers for Knowledge Graph Reasoning 19 Sep 2024 · 0 repositories · arXiv:2409.12865
-
Linear Model Predictive Control for Quadrotors with An Analytically Derived Koopman Model 19 Sep 2024 · 0 repositories · arXiv:2409.12374
-
LMT-Net: Lane Model Transformer Network for Automated HD Mapping from Sparse Vehicle Observations 19 Sep 2024 · 0 repositories · arXiv:2409.12409
-
LVCD: Reference-based Lineart Video Colorization with Diffusion Models 19 Sep 2024 · 0 repositories · arXiv:2409.12960
-
MambaClinix: Hierarchical Gated Convolution and Mamba-Based U-Net for Enhanced 3D Medical Image Segmentation 19 Sep 2024 · 1 repository · arXiv:2409.12533
-
MURI: High-Quality Instruction Tuning Datasets for Low-Resource Languages via Reverse Instructions 19 Sep 2024 · 1 repository · arXiv:2409.12958
-
On the Effectiveness of LLMs for Manual Test Verifications 19 Sep 2024 · 0 repositories · arXiv:2409.12405
-
Optimizing food taste sensory evaluation through neural network-based taste electroencephalogram channel selection 19 Sep 2024 · 0 repositories · arXiv:2410.03559
-
Predicting soccer matches with complex networks and machine learning 19 Sep 2024 · 0 repositories · arXiv:2409.13098
-
Profiling Patient Transcript Using Large Language Model Reasoning Augmentation for Alzheimer's Disease Detection 19 Sep 2024 · 1 repository · arXiv:2409.12541
-
Prompting Segment Anything Model with Domain-Adaptive Prototype for Generalizable Medical Image Segmentation 19 Sep 2024 · 1 repository · arXiv:2409.12522
-
Prompts Are Programs Too! Understanding How Developers Build Software Containing Prompts 19 Sep 2024 · 0 repositories · arXiv:2409.12447
-
Real-time estimation of overt attention from dynamic features of the face using deep-learning 19 Sep 2024 · 1 repository · arXiv:2409.13084
-
Retrieval-Augmented Test Generation: How Far Are We? 19 Sep 2024 · 0 repositories · arXiv:2409.12682
-
Should RAG Chatbots Forget Unimportant Conversations? Exploring Importance and Forgetting with Psychological Insights 19 Sep 2024 · 1 repository · arXiv:2409.12524
-
Small Language Models are Equation Reasoners 19 Sep 2024 · 0 repositories · arXiv:2409.12393
-
TACO-RL: Task Aware Prompt Compression Optimization with Reinforcement Learning 19 Sep 2024 · 0 repositories · arXiv:2409.13035
-
TEAM PILOT -- Learned Feasible Extendable Set of Dynamic MRI Acquisition Trajectories 19 Sep 2024 · 0 repositories · arXiv:2409.12777
-
Text2Traj2Text: Learning-by-Synthesis Framework for Contextual Captioning of Human Movement Trajectories 19 Sep 2024 · 1 repository · arXiv:2409.12670
-
Towards Low-latency Event-based Visual Recognition with Hybrid Step-wise Distillation Spiking Neural Networks 19 Sep 2024 · 1 repository · arXiv:2409.12507
-
What Would You Ask When You First Saw a²+b²=c²? Evaluating LLM on Curiosity-Driven Questioning 19 Sep 2024 · 0 repositories · arXiv:2409.17172
-
A Controlled Study on Long Context Extension and Generalization in LLMs 18 Sep 2024 · 1 repository · arXiv:2409.12181Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 2 honoured, 0 violated, 5 with no contract checked; 4 where Syntology's instrument failed) · 3 unverified (of 14 harvested samples) · 14 pointer-only (licence)
-
Agent Aggregator with Mask Denoise Mechanism for Histopathology Whole Slide Image Analysis 18 Sep 2024 · 0 repositories · arXiv:2409.11664
-
Axial Attention Transformer Networks: A New Frontier in Breast Cancer Detection 18 Sep 2024 · 0 repositories · arXiv:2409.12347
-
BERT-VBD: Vietnamese Multi-Document Summarization Framework 18 Sep 2024 · 0 repositories · arXiv:2409.12134
-
Bridging Domain Gap for Flight-Ready Spaceborne Vision 18 Sep 2024 · 0 repositories · arXiv:2409.11661
-
Data Efficient Acoustic Scene Classification using Teacher-Informed Confusing Class Instruction 18 Sep 2024 · 0 repositories · arXiv:2409.11964
-
DPI-TTS: Directional Patch Interaction for Fast-Converging and Style Temporal Modeling in Text-to-Speech 18 Sep 2024 · 0 repositories · arXiv:2409.11835
-
DynaMo: In-Domain Dynamics Pretraining for Visuo-Motor Control 18 Sep 2024 · 0 repositories · arXiv:2409.12192
-
Extract-and-Abstract: Unifying Extractive and Abstractive Summarization within Single Encoder-Decoder Framework 18 Sep 2024 · 0 repositories · arXiv:2409.11827
-
From Lists to Emojis: How Format Bias Affects Model Alignment 18 Sep 2024 · 0 repositories · arXiv:2409.11704
-
Generalized Robot Learning Framework 18 Sep 2024 · 0 repositories · arXiv:2409.12061
-
Harnessing LLMs for API Interactions: A Framework for Classification and Synthetic Data Generation 18 Sep 2024 · 0 repositories · arXiv:2409.11703
-
ID-Free Not Risk-Free: LLM-Powered Agents Unveil Risks in ID-Free Recommender Systems 18 Sep 2024 · 0 repositories · arXiv:2409.11690
-
MAgICoRe: Multi-Agent, Iterative, Coarse-to-Fine Refinement for Reasoning 18 Sep 2024 · 1 repository · arXiv:2409.12147Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
Mastering Chess with a Transformer Model 18 Sep 2024 · 1 repository · arXiv:2409.12272
-
Measuring Sound Symbolism in Audio-visual Models 18 Sep 2024 · 0 repositories · arXiv:2409.12306
-
Monomial Matrix Group Equivariant Neural Functional Networks 18 Sep 2024 · 1 repository · arXiv:2409.11697Syntology official (archive's flag): 26 ran · 26 ran (of which 14 constructed an object rather than computing a result; 20 with no instrument failure: 0 honoured, 0 violated, 20 with no contract checked; 6 where Syntology's instrument failed) · 10 unverified (of 36 harvested samples)
-
Multi-Grid Graph Neural Networks with Self-Attention for Computational Mechanics 18 Sep 2024 · 1 repository · arXiv:2409.11899Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
NT-ViT: Neural Transcoding Vision Transformers for EEG-to-fMRI Synthesis 18 Sep 2024 · 0 repositories · arXiv:2409.11836
-
ORB-SfMLearner: ORB-Guided Self-supervised Visual Odometry with Selective Online Adaptation 18 Sep 2024 · 0 repositories · arXiv:2409.11692