Methods › General › Attention Mechanisms › Attention › Papers, page 239
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 239 of 316: papers 23,801 to 23,900 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
PASH at TREC 2021 Deep Learning Track: Generative Enhanced Model for Multi-stage Ranking 18 May 2022 · 0 repositories · arXiv:2205.11245
-
Persian Natural Language Inference: A Meta-learning approach 18 May 2022 · 1 repository · arXiv:2205.08755
-
Transformer based multiple instance learning for weakly supervised histopathology image segmentation 18 May 2022 · 1 repository · arXiv:2205.08878
-
U-Former: Improving Monaural Speech Enhancement with Multi-head Self and Cross Attention 18 May 2022 · 1 repository · arXiv:2205.08681
-
Attention-aware contrastive learning for predicting T cell receptor-antigen binding specificity 17 May 2022 · 0 repositories · arXiv:2206.11255
-
Feature Aggregation in Zero-Shot Cross-Lingual Transfer Using Multilingual BERT 17 May 2022 · 0 repositories · arXiv:2205.08497
-
HoVer-Trans: Anatomy-aware HoVer-Transformer for ROI-free Breast Cancer Diagnosis in Ultrasound Images 17 May 2022 · 0 repositories · arXiv:2205.08390
-
M6-Rec: Generative Pretrained Language Models are Open-Ended Recommender Systems 17 May 2022 · 0 repositories · arXiv:2205.08084
-
MATrIX -- Modality-Aware Transformer for Information eXtraction 17 May 2022 · 0 repositories · arXiv:2205.08094
-
MulT: An End-to-End Multitask Learning Transformer 17 May 2022 · 1 repository · arXiv:2205.08303Syntology 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
POViT: Vision Transformer for Multi-objective Design and Characterization of Nanophotonic Devices 17 May 2022 · 0 repositories · arXiv:2205.09045
-
SAMU-XLSR: Semantically-Aligned Multimodal Utterance-level Cross-Lingual Speech Representation 17 May 2022 · 0 repositories · arXiv:2205.08180
-
SEMI-FND: Stacked Ensemble Based Multimodal Inference For Faster Fake News Detection 17 May 2022 · 0 repositories · arXiv:2205.08159
-
SKILL: Structured Knowledge Infusion for Large Language Models 17 May 2022 · 0 repositories · arXiv:2205.08184
-
The Power of Fragmentation: A Hierarchical Transformer Model for Structural Segmentation in Symbolic Music Generation 17 May 2022 · 0 repositories · arXiv:2205.08579
-
ViralBERT: A User Focused BERT-Based Approach to Virality Prediction 17 May 2022 · 1 repository · arXiv:2206.10298
-
Vision Transformer Adapter for Dense Predictions 17 May 2022 · 2 repositories · arXiv:2205.08534
-
Chemical transformer compression for accelerating both training and inference of molecular modeling 16 May 2022 · 1 repository · arXiv:2205.07582
-
CONSENT: Context Sensitive Transformer for Bold Words Classification 16 May 2022 · 0 repositories · arXiv:2205.07683
-
Harnessing Multilingual Resources to Question Answering in Arabic 16 May 2022 · 0 repositories · arXiv:2205.08024
-
Heroes, Villains, and Victims, and GPT-3: Automated Extraction of Character Roles Without Training Data 16 May 2022 · 0 repositories · arXiv:2205.07557
-
The AI Teacher Test: Measuring the Pedagogical Ability of Blender and GPT-3 in Educational Dialogues 16 May 2022 · 1 repository · arXiv:2205.07540
-
Transformers in 3D Point Clouds: A Survey 16 May 2022 · 0 repositories · arXiv:2205.07417
-
What GPT Knows About Who is Who 16 May 2022 · 1 repository · arXiv:2205.07407
-
Discovering Latent Concepts Learned in BERT 15 May 2022 · 0 repositories · arXiv:2205.07237
-
Downstream Transformer Generation of Question-Answer Pairs with Preprocessing and Postprocessing Pipelines 15 May 2022 · 1 repository · arXiv:2205.07387
-
Hero-Gang Neural Model For Named Entity Recognition 15 May 2022 · 1 repository · arXiv:2205.07177
-
Learning Lip-Based Audio-Visual Speaker Embeddings with AV-HuBERT 15 May 2022 · 1 repository · arXiv:2205.07180
-
Proxyless Neural Architecture Adaptation for Supervised Learning and Self-Supervised Learning 15 May 2022 · 0 repositories · arXiv:2205.07168
-
Transkimmer: Transformer Learns to Layer-wise Skim 15 May 2022 · 1 repository · arXiv:2205.07324Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; the one sample that ran constructed an object rather than computing a result (of 2 harvested samples) · 2 pointer-only (licence)
-
Video Frame Interpolation with Transformer 15 May 2022 · 1 repository · arXiv:2205.07230
-
Dense residual Transformer for image denoising 14 May 2022 · 0 repositories · arXiv:2205.06944
-
Fake News Quick Detection on Dynamic Heterogeneous Information Networks 14 May 2022 · 0 repositories · arXiv:2205.07039
-
Naturalistic Causal Probing for Morpho-Syntax 14 May 2022 · 1 repository · arXiv:2205.07043
-
RASAT: Integrating Relational Structures into Pretrained Seq2Seq Model for Text-to-SQL 14 May 2022 · 1 repository · arXiv:2205.06983Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 9 harvested samples) · 1 pointer-only (licence)
-
Transformer Scale Gate for Semantic Segmentation 14 May 2022 · 0 repositories · arXiv:2205.07056
-
A microstructure estimation Transformer inspired by sparse representation for diffusion MRI 13 May 2022 · 1 repository · arXiv:2205.06450
-
A Study of the Attention Abnormality in Trojaned BERTs 13 May 2022 · 1 repository · arXiv:2205.08305Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Improving Contextual Representation with Gloss Regularized Pre-training 13 May 2022 · 0 repositories · arXiv:2205.06603
-
Local Attention Graph-based Transformer for Multi-target Genetic Alteration Prediction 13 May 2022 · 1 repository · arXiv:2205.06672
-
ViT5: Pretrained Text-to-Text Transformer for Vietnamese Language Generation 13 May 2022 · 1 repository · arXiv:2205.06457
-
A Generalist Agent 12 May 2022 · 3 repositories · arXiv:2205.06175Syntology 0 ran · 2 unverified (of 2 harvested samples)
-
AppTek's Submission to the IWSLT 2022 Isometric Spoken Language Translation Task 12 May 2022 · 0 repositories · arXiv:2205.05807
-
Deep Learning for Prawn Farming: Forecasting and Anomaly Detection 12 May 2022 · 0 repositories · arXiv:2205.06359
-
Entity-aware and Motion-aware Transformers for Language-driven Action Localization in Videos 12 May 2022 · 1 repository · arXiv:2205.05854
-
Group R-CNN for Weakly Semi-supervised Object Detection with Points 12 May 2022 · 1 repository · arXiv:2205.05920
-
Is the Computation of Abstract Sameness Relations Human-Like in Neural Language Models? 12 May 2022 · 0 repositories · arXiv:2205.06149
-
Multi Task Learning For Zero Shot Performance Prediction of Multilingual Models 12 May 2022 · 0 repositories · arXiv:2205.06130Syntology 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 8 harvested samples)
-
NER-MQMRC: Formulating Named Entity Recognition as Multi Question Machine Reading Comprehension 12 May 2022 · 0 repositories · arXiv:2205.05904
-
NFLAT: Non-Flat-Lattice Transformer for Chinese Named Entity Recognition 12 May 2022 · 1 repository · arXiv:2205.05832
-
Robot Cooking with Stir-fry: Bimanual Non-prehensile Manipulation of Semi-fluid Objects 12 May 2022 · 1 repository · arXiv:2205.05960
-
Efficient and Training-Free Control of Language Generation 12 May 2022 · 0 repositories · arXiv:2205.06036
-
Simple Open-Vocabulary Object Detection with Vision Transformers 12 May 2022 · 2 repositories · arXiv:2205.06230
-
Supplementary Material: Implementation and Experiments for GAU-based Model 12 May 2022 · 0 repositories · arXiv:2205.05842
-
Vision Transformer: Vit and its Derivatives 12 May 2022 · 0 repositories · arXiv:2205.11239
-
A time-varying study of Chinese investor sentiment, stock market liquidity and volatility: Based on deep learning BERT model and TVP-VAR model 11 May 2022 · 0 repositories · arXiv:2205.05719
-
AggPose: Deep Aggregation Vision Transformer for Infant Pose Estimation 11 May 2022 · 1 repository · arXiv:2205.05277Syntology official (archive's flag): 5 ran · 5 ran (of which 5 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified; every one of the 5 samples that ran constructed an object rather than computing a result (of 7 harvested samples) · 7 pointer-only (licence)
-
An Empirical Study Of Self-supervised Learning Approaches For Object Detection With Transformers 11 May 2022 · 2 repositories · arXiv:2205.05543
-
Clinical Prompt Learning with Frozen Language Models 11 May 2022 · 1 repository · arXiv:2205.05535
-
CV4Code: Sourcecode Understanding via Visual Code Representations 11 May 2022 · 0 repositories · arXiv:2205.08585
-
Query-Based Keyphrase Extraction from Long Documents 11 May 2022 · 1 repository · arXiv:2205.05391
-
Towards the Generation of Musical Explanations with GPT-3 11 May 2022 · 1 repository · arXiv:2206.08264
-
AdMix: A Mixed Sample Data Augmentation Method for Neural Machine Translation 10 May 2022 · 0 repositories · arXiv:2205.04686
-
Deep learning based Chinese text sentiment mining and stock market correlation research 10 May 2022 · 0 repositories · arXiv:2205.04743
-
DistilProtBert: A distilled protein language model used to distinguish between real proteins and their randomly shuffled counterparts 10 May 2022 · 1 repository
-
Problems with Cosine as a Measure of Embedding Similarity for High Frequency Words 10 May 2022 · 2 repositories · arXiv:2205.05092
-
Ratatouille: A tool for Novel Recipe Generation 10 May 2022 · 0 repositories · arXiv:2206.08267
-
Reduce Information Loss in Transformers for Pluralistic Image Inpainting 10 May 2022 · 1 repository · arXiv:2205.05076
-
Reducing Activation Recomputation in Large Transformer Models 10 May 2022 · 4 repositories · arXiv:2205.05198Syntology official: harvested, nothing ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 1 pointer-only (licence)
-
Spatio-Temporal Transformer for Dynamic Facial Expression Recognition in the Wild 10 May 2022 · 0 repositories · arXiv:2205.04749
-
UL2: Unifying Language Learning Paradigms 10 May 2022 · 2 repositories · arXiv:2205.05131Syntology community repositories only · 15 ran (of which 0 constructed an object rather than computing a result; 15 with no instrument failure: 0 honoured, 0 violated, 15 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 16 harvested samples)
-
Activating More Pixels in Image Super-Resolution Transformer 9 May 2022 · 2 repositories · arXiv:2205.04437
-
Automated Evaluation for Student Argumentative Writing: A Survey 9 May 2022 · 0 repositories · arXiv:2205.04083
-
Incremental-DETR: Incremental Few-Shot Object Detection via Self-Supervised Learning 9 May 2022 · 0 repositories · arXiv:2205.04042
-
Long Document Re-ranking with Modular Re-ranker 9 May 2022 · 1 repository · arXiv:2205.04275
-
Multi-segment preserving sampling for deep manifold sampler 9 May 2022 · 0 repositories · arXiv:2205.04259
-
Research on the correlation between text emotion mining and stock market based on deep learning 9 May 2022 · 0 repositories · arXiv:2205.06675
-
SwinIQA: Learned Swin Distance for Compressed Image Quality Assessment 9 May 2022 · 1 repository · arXiv:2205.04264
-
Neural Architecture Search using Property Guided Synthesis 8 May 2022 · 1 repository · arXiv:2205.03960
-
Adaptive Graph Convolutional Network Framework for Multidimensional Time Series Prediction 8 May 2022 · 0 repositories · arXiv:2205.04885
-
DxFormer: A Decoupled Automatic Diagnostic System Based on Decoder-Encoder Transformer with Dense Symptom Representations 8 May 2022 · 1 repository · arXiv:2205.03755
-
Mutual Distillation Learning Network for Trajectory-User Linking 8 May 2022 · 1 repository · arXiv:2205.03773Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
On the Use of BERT for Automated Essay Scoring: Joint Learning of Multi-Scale Essay Representation 8 May 2022 · 1 repository · arXiv:2205.03835
-
AKI-BERT: a Pre-trained Clinical Language Model for Early Prediction of Acute Kidney Injury 7 May 2022 · 1 repository · arXiv:2205.03695
-
EmotionFlow: Capture the Dialogue Level Emotion Transitions 7 May 2022 · 1 repository
-
Improving Downstream Task Performance by Treating Numbers as Entities 7 May 2022 · 0 repositories · arXiv:2205.03559
-
Vector Representations of Idioms in Conversational Systems 7 May 2022 · 0 repositories · arXiv:2205.03666
-
A Data Cartography based MixUp for Pre-trained Language Models 6 May 2022 · 1 repository · arXiv:2205.03403
-
Explaining the Effectiveness of Multi-Task Learning for Efficient Knowledge Extraction from Spine MRI Reports 6 May 2022 · 0 repositories · arXiv:2205.02979
-
Fake News Detection with Heterogeneous Transformer 6 May 2022 · 1 repository · arXiv:2205.03100
-
RCMNet: A deep learning model assists CAR-T therapy for leukemia 6 May 2022 · 0 repositories · arXiv:2205.04230
-
Stock Price Prediction Based on Natural Language Processing 6 May 2022 · 2 repositories
-
The Unreliability of Explanations in Few-shot Prompting for Textual Reasoning 6 May 2022 · 1 repository · arXiv:2205.03401
-
Transformer-Based Multi-Aspect Multi-Granularity Non-Native English Speaker Pronunciation Assessment 6 May 2022 · 1 repository · arXiv:2205.03432
-
When a sentence does not introduce a discourse entity, Transformer-based models still sometimes refer to it 6 May 2022 · 1 repository · arXiv:2205.03472
-
A Deep Learning Approach to Dst Index Prediction 5 May 2022 · 0 repositories · arXiv:2205.02447
-
BORT: Back and Denoising Reconstruction for End-to-End Task-Oriented Dialog 5 May 2022 · 1 repository · arXiv:2205.02471
-
Declaration-based Prompt Tuning for Visual Question Answering 5 May 2022 · 1 repository · arXiv:2205.02456
-
Exploiting Global and Local Hierarchies for Hierarchical Text Classification 5 May 2022 · 1 repository · arXiv:2205.02613
-
Identifying Cause-and-Effect Relationships of Manufacturing Errors using Sequence-to-Sequence Learning 5 May 2022 · 0 repositories · arXiv:2205.02827