Methods › General › Attention Mechanisms › Attention › Papers, page 213
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 213 of 316: papers 21,201 to 21,300 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Bokeh Rendering Based on Adaptive Depth Calibration Network 21 Feb 2023 · 0 repositories · arXiv:2302.10808
-
ChatGPT: Jack of all trades, master of none 21 Feb 2023 · 1 repository · arXiv:2302.10724Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Co-Driven Recognition of Semantic Consistency via the Fusion of Transformer and HowNet Sememes Knowledge 21 Feb 2023 · 1 repository · arXiv:2302.10570
-
Diffusion Models and Semi-Supervised Learners Benefit Mutually with Few Labels 21 Feb 2023 · 3 repositories · arXiv:2302.10586Syntology official (archive's flag): 18 ran · 18 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 1 honoured, 0 violated, 9 with no contract checked; 8 where Syntology's instrument failed) · 4 unverified (of 22 harvested samples) · 9 pointer-only (licence)
-
Edgeformers: Graph-Empowered Transformers for Representation Learning on Textual-Edge Networks 21 Feb 2023 · 1 repository · arXiv:2302.11050Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (of 7 harvested samples)
-
Hyena Hierarchy: Towards Larger Convolutional Language Models 21 Feb 2023 · 7 repositories · arXiv:2302.10866Syntology official (archive's flag): 2 ran · 5 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 5 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 4 pointer-only (licence)
-
kNN-Adapter: Efficient Domain Adaptation for Black-Box Language Models 21 Feb 2023 · 0 repositories · arXiv:2302.10879
-
Label Information Enhanced Fraud Detection against Low Homophily in Graphs 21 Feb 2023 · 1 repository · arXiv:2302.10407
-
Lightweight Real-time Semantic Segmentation Network with Efficient Transformer and CNN 21 Feb 2023 · 1 repository · arXiv:2302.10484
-
Memory-augmented Online Video Anomaly Detection 21 Feb 2023 · 1 repository · arXiv:2302.10719
-
MulGT: Multi-task Graph-Transformer with Task-aware Knowledge Injection and Domain Knowledge-driven Pooling for Whole Slide Image Analysis 21 Feb 2023 · 0 repositories · arXiv:2302.10574
-
MVMTnet: A Multi-variate Multi-modal Transformer for Multi-class Classification of Cardiac Irregularities Using ECG Waveforms and Clinical Notes 21 Feb 2023 · 1 repository · arXiv:2302.11021
-
SF2Former: Amyotrophic Lateral Sclerosis Identification From Multi-center MRI Data Using Spatial and Frequency Fusion Transformer 21 Feb 2023 · 1 repository · arXiv:2302.10859
-
Time to Embrace Natural Language Processing (NLP)-based Digital Pathology: Benchmarking NLP- and Convolutional Neural Network-based Deep Learning Pipelines 21 Feb 2023 · 0 repositories · arXiv:2302.10406
-
Because Every Sensor Is Unique, so Is Every Pair: Handling Dynamicity in Traffic Forecasting 20 Feb 2023 · 1 repository · arXiv:2302.09956
-
Boosting classification reliability of NLP transformer models in the long run 20 Feb 2023 · 0 repositories · arXiv:2302.10016
-
Exploring the Advantages of Transformers for High-Frequency Trading 20 Feb 2023 · 1 repository · arXiv:2302.13850
-
Friend Ranking in Online Games via Pre-training Edge Transformers 20 Feb 2023 · 2 repositories · arXiv:2302.10043
-
GlocalFuse-Depth: Fusing Transformers and CNNs for All-day Self-supervised Monocular Depth Estimation 20 Feb 2023 · 0 repositories · arXiv:2302.09884
-
Large-scale Multi-Modal Pre-trained Models: A Comprehensive Survey 20 Feb 2023 · 1 repository · arXiv:2302.10035
-
Optical Transformers 20 Feb 2023 · 0 repositories · arXiv:2302.10360
-
STB-VMM: Swin Transformer Based Video Motion Magnification 20 Feb 2023 · 1 repository · arXiv:2302.10001
-
ChatIE: Zero-Shot Information Extraction via Chatting with ChatGPT 20 Feb 2023 · 1 repository · arXiv:2302.10205
-
Can ChatGPT Understand Too? A Comparative Study on ChatGPT and Fine-tuned BERT 19 Feb 2023 · 1 repository · arXiv:2302.10198
-
Evaluating the Effectiveness of Pre-trained Language Models in Predicting the Helpfulness of Online Product Reviews 19 Feb 2023 · 1 repository · arXiv:2302.10199
-
MedViT: A Robust Vision Transformer for Generalized Medical Image Classification 19 Feb 2023 · 1 repository · arXiv:2302.09462Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
Mixed Hierarchy Network for Image Restoration 19 Feb 2023 · 1 repository · arXiv:2302.09554
-
Text Classification in the Wild: a Large-scale Long-tailed Name Normalization Dataset 19 Feb 2023 · 1 repository · arXiv:2302.09509
-
A Comprehensive Survey on Pretrained Foundation Models: A History from BERT to ChatGPT 18 Feb 2023 · 0 repositories · arXiv:2302.09419
-
Bag of Tricks for Effective Language Model Pretraining and Downstream Adaptation: A Case Study on GLUE 18 Feb 2023 · 0 repositories · arXiv:2302.09268
-
BBT-Fin: Comprehensive Construction of Chinese Financial Domain Pre-trained Language Model, Corpus and Benchmark 18 Feb 2023 · 2 repositories · arXiv:2302.09432
-
How Good Are GPT Models at Machine Translation? A Comprehensive Evaluation 18 Feb 2023 · 1 repository · arXiv:2302.09210Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Hyneter: Hybrid Network Transformer for Object Detection 18 Feb 2023 · 0 repositories · arXiv:2302.09365
-
Neural Attention Memory 18 Feb 2023 · 0 repositories · arXiv:2302.09422
-
VITAL: Vision Transformer Neural Networks for Accurate Smartphone Heterogeneity Resilient Indoor Localization 18 Feb 2023 · 0 repositories · arXiv:2302.09443
-
Bounding the Capabilities of Large Language Models in Open Text Generation with Prompt Constraints 17 Feb 2023 · 1 repository · arXiv:2302.09185
-
Conveying the Predicted Future to Users: A Case Study of Story Plot Prediction 17 Feb 2023 · 1 repository · arXiv:2302.09122
-
DTAAD: Dual Tcn-Attention Networks for Anomaly Detection in Multivariate Time Series Data 17 Feb 2023 · 1 repository · arXiv:2302.10753
-
GPT4MIA: Utilizing Generative Pre-trained Transformer (GPT-3) as A Plug-and-Play Transductive Model for Medical Image Analysis 17 Feb 2023 · 0 repositories · arXiv:2302.08722
-
Hate Speech and Offensive Language Detection using an Emotion-aware Shared Encoder 17 Feb 2023 · 0 repositories · arXiv:2302.08777
-
Improving Transformer-based Networks With Locality For Automatic Speaker Verification 17 Feb 2023 · 0 repositories · arXiv:2302.08639
-
Like a Good Nearest Neighbor: Practical Content Moderation and Text Classification 17 Feb 2023 · 1 repository · arXiv:2302.08957
-
EnfoMax: Domain Entropy and Mutual Information Maximization for Domain Generalized Face Anti-spoofing 17 Feb 2023 · 0 repositories · arXiv:2302.08674
-
Multiresolution Graph Transformers and Wavelet Positional Encoding for Learning Hierarchical Structures 17 Feb 2023 · 2 repositories · arXiv:2302.08647Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 2 pointer-only (licence)
-
PAC Prediction Sets for Large Language Models of Code 17 Feb 2023 · 1 repository · arXiv:2302.08703
-
Prompting Large Language Models With the Socratic Method 17 Feb 2023 · 0 repositories · arXiv:2303.08769
-
Transformer-based Generative Adversarial Networks in Computer Vision: A Comprehensive Survey 17 Feb 2023 · 0 repositories · arXiv:2302.08641
-
ViTA: A Vision Transformer Inference Accelerator for Edge Applications 17 Feb 2023 · 0 repositories · arXiv:2302.09108
-
A Transformer-based Deep Learning Algorithm to Auto-record Undocumented Clinical One-Lung Ventilation Events 16 Feb 2023 · 0 repositories · arXiv:2302.12713
-
Document Flattening: Beyond Concatenating Context for Document-Level Neural Machine Translation 16 Feb 2023 · 0 repositories · arXiv:2302.08079
-
Efficiency 360: Efficient Vision Transformers 16 Feb 2023 · 1 repository · arXiv:2302.08374
-
Foundation Models for Natural Language Processing -- Pre-trained Language Models Integrating Media 16 Feb 2023 · 0 repositories · arXiv:2302.08575
-
Hierarchical Cross-modal Transformer for RGB-D Salient Object Detection 16 Feb 2023 · 0 repositories · arXiv:2302.08052
-
For Generated Text, Is NLI-Neutral Text the Best Text? 16 Feb 2023 · 1 repository · arXiv:2302.08577
-
Learning Non-Local Spatial-Angular Correlation for Light Field Image Super-Resolution 16 Feb 2023 · 1 repository · arXiv:2302.08058
-
Marich: A Query-efficient Distributionally Equivalent Model Extraction Attack using Public Data 16 Feb 2023 · 1 repository · arXiv:2302.08466Syntology official (archive's flag): 5 ran · 5 ran (of which 1 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 0 violated, 1 with no contract checked; 3 where Syntology's instrument failed) · 3 unverified (of 8 harvested samples)
-
Prompt Tuning of Deep Neural Networks for Speaker-adaptive Visual Speech Recognition 16 Feb 2023 · 0 repositories · arXiv:2302.08102
-
Retrieval-augmented Image Captioning 16 Feb 2023 · 1 repository · arXiv:2302.08268
-
Robust Human Motion Forecasting using Transformer-based Model 16 Feb 2023 · 0 repositories · arXiv:2302.08274
-
Search-Engine-augmented Dialogue Response Generation with Cheaply Supervised Query Production 16 Feb 2023 · 1 repository · arXiv:2302.09300
-
Short-term and long-term memory self-attention network for segmentation of tumours in 3D medical images 16 Feb 2023 · 0 repositories
-
Speaker Change Detection for Transformer Transducer ASR 16 Feb 2023 · 0 repositories · arXiv:2302.08549
-
Syntactic Structure Processing in the Brain while Listening 16 Feb 2023 · 0 repositories · arXiv:2302.08589
-
TcGAN: Semantic-Aware and Structure-Preserved GANs with Individual Vision Transformer for Fast Arbitrary One-Shot Image Generation 16 Feb 2023 · 0 repositories · arXiv:2302.08047
-
URCDC-Depth: Uncertainty Rectified Cross-Distillation with CutFlip for Monocular Depth Estimation 16 Feb 2023 · 1 repository · arXiv:2302.08149Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 1 honoured, 0 violated, 4 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 11 harvested samples) · 5 pointer-only (licence)
-
Commonsense Reasoning for Conversational AI: A Survey of the State of the Art 15 Feb 2023 · 0 repositories · arXiv:2302.07926
-
Confidence Score Based Speaker Adaptation of Conformer Speech Recognition Systems 15 Feb 2023 · 1 repository · arXiv:2302.07521
-
Learning Performance-Improving Code Edits 15 Feb 2023 · 2 repositories · arXiv:2302.07867Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 2 honoured, 0 violated, 1 with no contract checked; 5 where Syntology's instrument failed) · 11 unverified (of 19 harvested samples) · 19 pointer-only (licence)
-
Pose-Oriented Transformer with Uncertainty-Guided Refinement for 2D-to-3D Human Pose Estimation 15 Feb 2023 · 0 repositories · arXiv:2302.07408
-
Self-Supervised Learning for Modeling Gamma-ray Variability in Blazars 15 Feb 2023 · 0 repositories · arXiv:2302.07700
-
Speculative Decoding with Big Little Decoder 15 Feb 2023 · 1 repository · arXiv:2302.07863Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
TFormer: A Transmission-Friendly ViT Model for IoT Devices 15 Feb 2023 · 0 repositories · arXiv:2302.07734
-
Towards Optimal Compression: Joint Pruning and Quantization 15 Feb 2023 · 0 repositories · arXiv:2302.07612
-
Tree-Based Representation and Generation of Natural and Mathematical Language 15 Feb 2023 · 1 repository · arXiv:2302.07974
-
A Modern Look at the Relationship between Sharpness and Generalization 14 Feb 2023 · 1 repository · arXiv:2302.07011
-
A Psycholinguistic Analysis of BERT's Representations of Compounds 14 Feb 2023 · 1 repository · arXiv:2302.07232
-
Deep Learning-Based Modeling of 5G Core Control Plane for 5G Network Digital Twin 14 Feb 2023 · 0 repositories · arXiv:2302.06980
-
DiffFashion: Reference-based Fashion Design with Structure-aware Transfer by Diffusion Models 14 Feb 2023 · 1 repository · arXiv:2302.06826
-
Energy Transformer 14 Feb 2023 · 4 repositories · arXiv:2302.07253Syntology official (archive's flag): 11 ran · 13 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 2 honoured, 0 violated, 5 with no contract checked; 6 where Syntology's instrument failed) · 0 unverified (of 13 harvested samples) · 3 pointer-only (licence)
-
Exploring Category Structure with Contextual Language Models and Lexical Semantic Networks 14 Feb 2023 · 0 repositories · arXiv:2302.06942
-
Few-shot learning approaches for classifying low resource domain specific software requirements 14 Feb 2023 · 0 repositories · arXiv:2302.06951
-
PolyFormer: Referring Image Segmentation as Sequential Polygon Generation 14 Feb 2023 · 1 repository · arXiv:2302.07387Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Reveal the Unknown: Out-of-Knowledge-Base Mention Discovery with Entity Linking 14 Feb 2023 · 3 repositories · arXiv:2302.07189
-
ScatterShot: Interactive In-context Example Curation for Text Transformation 14 Feb 2023 · 1 repository · arXiv:2302.07346
-
Team DETR: Guide Queries as a Professional Team in Detection Transformers 14 Feb 2023 · 1 repository · arXiv:2302.07116
-
A Comprehensive Study of Modern Architectures and Regularization Approaches on CheXpert5000 13 Feb 2023 · 0 repositories · arXiv:2302.06684
-
A Study on ReLU and Softmax in Transformer 13 Feb 2023 · 0 repositories · arXiv:2302.06461
-
A Unified View of Long-Sequence Models towards Modeling Million-Scale Dependencies 13 Feb 2023 · 0 repositories · arXiv:2302.06218
-
Anticipating Next Active Objects for Egocentric Videos 13 Feb 2023 · 0 repositories · arXiv:2302.06358
-
Diminished Diversity-of-Thought in a Standard Large Language Model 13 Feb 2023 · 0 repositories · arXiv:2302.07267
-
Can GPT-3 Perform Statutory Reasoning? 13 Feb 2023 · 1 repository · arXiv:2302.06100
-
CholecTriplet2022: Show me a tool and tell me the triplet -- an endoscopic vision challenge for surgical action triplet detection 13 Feb 2023 · 2 repositories · arXiv:2302.06294
-
VITR: Augmenting Vision Transformers with Relation-Focused Learning for Cross-Modal Information Retrieval 13 Feb 2023 · 0 repositories · arXiv:2302.06350
-
Encoding Sentence Position in Context-Aware Neural Machine Translation with Concatenation 13 Feb 2023 · 1 repository · arXiv:2302.06459
-
Linguistic ambiguity analysis in ChatGPT 13 Feb 2023 · 0 repositories · arXiv:2302.06426
-
One Transformer for All Time Series: Representing and Training with Time-Dependent Heterogeneous Tabular Data 13 Feb 2023 · 1 repository · arXiv:2302.06375Syntology 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
Simple Hardware-Efficient Long Convolutions for Sequence Modeling 13 Feb 2023 · 1 repository · arXiv:2302.06646
-
STREET: A Multi-Task Structured Reasoning and Explanation Benchmark 13 Feb 2023 · 0 repositories · arXiv:2302.06729
-
Towards Local Visual Modeling for Image Captioning 13 Feb 2023 · 1 repository · arXiv:2302.06098
-
Using SHAP Values and Machine Learning to Understand Trends in the Transient Stability Limit 13 Feb 2023 · 0 repositories · arXiv:2302.06274