Methods › General › Attention Mechanisms › Attention › Papers, page 249
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 249 of 316: papers 24,801 to 24,900 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
RePre: Improving Self-Supervised Vision Transformer with Reconstructive Pre-training 18 Jan 2022 · 0 repositories · arXiv:2201.06857
-
Resistance Training using Prior Bias: toward Unbiased Scene Graph Generation 18 Jan 2022 · 1 repository · arXiv:2201.06794Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Controllable Protein Design with Language Models 18 Jan 2022 · 0 repositories · arXiv:2201.07338
-
TranAD: Deep Transformer Networks for Anomaly Detection in Multivariate Time Series Data 18 Jan 2022 · 2 repositories · arXiv:2201.07284Syntology official: harvested, nothing ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
BERT vs ALBERT explained 17 Jan 2022 · 0 repositories
-
Continual Transformers: Redundancy-Free Attention for Online Inference 17 Jan 2022 · 1 repository · arXiv:2201.06268
-
Disentangled Latent Transformer for Interpretable Monocular Height Estimation 17 Jan 2022 · 0 repositories · arXiv:2201.06357
-
Korean-Specific Dataset for Table Question Answering 17 Jan 2022 · 1 repository · arXiv:2201.06223
-
Looking at the Performer from a Hopfield Point of View 17 Jan 2022 · 0 repositories
-
MuLVE, A Multi-Language Vocabulary Evaluation Data Set 17 Jan 2022 · 0 repositories · arXiv:2201.06286
-
SQUIRE: A Sequence-to-sequence Framework for Multi-hop Knowledge Graph Reasoning 17 Jan 2022 · 1 repository · arXiv:2201.06206
-
SwinUNet3D -- A Hierarchical Architecture for Deep Traffic Prediction using Shifted Window Transformers 17 Jan 2022 · 1 repository · arXiv:2201.06390Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Unintended Bias in Language Model-driven Conversational Recommendation 17 Jan 2022 · 0 repositories · arXiv:2201.06224
-
A Balanced Data Approach for Evaluating Cross-Lingual Transfer: Mapping the Linguistic Blood Bank 16 Jan 2022 · 0 repositories
-
A Multi-Granularity Opinion Summarization Method 16 Jan 2022 · 0 repositories
-
A Study of Pre-trained Language Models for Analogy Generation 16 Jan 2022 · 0 repositories
-
A Study of the Attention Abnormality in Trojaned BERTs 16 Jan 2022 · 0 repositories
-
AllWOZ: Towards Multilingual Task-Oriented Dialog Systems for All 16 Jan 2022 · 0 repositories
-
An Exploitation of Heterogeneous Graph Neural Network for Extractive Long Document Summarization 16 Jan 2022 · 0 repositories
-
Applying SoftTriple Loss for Supervised Language Model Fine Tuning 16 Jan 2022 · 0 repositories
-
AraBART: a Pretrained Arabic Sequence-to-Sequence Model for Abstractive Summarization 16 Jan 2022 · 0 repositories
-
Are Pretrained Multilingual Models Equally Fair Across Languages? 16 Jan 2022 · 0 repositories
-
Auto-regressive Text Generation with Pre-Trained Language Models: An Empirical Study on Question-type Short Text Generation 16 Jan 2022 · 0 repositories
-
AutoAttention: Automatic Attention Head Selection Through Differentiable Pruning 16 Jan 2022 · 0 repositories
-
Bridge the Gap Between CV and NLP! A Gradient-based Textual Adversarial Attack Framework 16 Jan 2022 · 0 repositories
-
Can BERT Conduct Logical Reasoning? On the Difficulty of Learning to Reason from Data 16 Jan 2022 · 0 repositories
-
Challenge for open-domain targeted sentiment analysis 16 Jan 2022 · 0 repositories
-
Context-Aware Prompt: Customize A Unique Prompt For Each Input 16 Jan 2022 · 0 repositories
-
Conventional clustering-based method for event detection on social networks 16 Jan 2022 · 0 repositories
-
DECK: Behavioral Tests to Improve Interpretability and Generalizability of BERT Models Detecting Depression from Text 16 Jan 2022 · 0 repositories
-
Divide and Conquer: Text Semantic Matching with Disentangled Keywords and Intents 16 Jan 2022 · 0 repositories
-
Do BERTs Learn to Use Browser User Interface? Exploring Multi-Step Tasks with Unified Vision-and-Language BERTs 16 Jan 2022 · 0 repositories
-
Efficient Hierarchical Domain Adaptation for Pretrained Language Models 16 Jan 2022 · 0 repositories
-
Efficient Zero-Shot Semantic Parsing with Paraphrasing from Pretrained Language Models 16 Jan 2022 · 0 repositories
-
EiCi: A New Method of Dynamic Embedding Incorporating Contextual Information in Chinese NER 16 Jan 2022 · 0 repositories
-
Elastic Weight Consolidation for Reduction of Catastrophic Forgetting in GPT-2 16 Jan 2022 · 0 repositories
-
ERNIE-Layout: Layout-Knowledge Enhanced Multi-modal Pre-training for Document Understanding 16 Jan 2022 · 1 repository
-
Event Detection via Derangement Reading Comprehension 16 Jan 2022 · 0 repositories
-
Experiments with adversarial attacks on text genres 16 Jan 2022 · 0 repositories
-
Exploring Example Selection for Few-shot Text-to-SQL Semantic Parsing 16 Jan 2022 · 0 repositories
-
Exploring the Low-Resource Transfer-Learning with mT5 model 16 Jan 2022 · 0 repositories
-
Extract, Select and Rewrite: A New Modular Summarization Method 16 Jan 2022 · 0 repositories
-
Feasibility of BERT Embeddings For Domain-Specific Knowledge Mining 16 Jan 2022 · 0 repositories
-
FedNLP: Benchmarking Federated Learning Methods for Natural Language Processing Tasks 16 Jan 2022 · 0 repositories
-
Few-Shot Semantic Parsing with Language Models Trained On Code 16 Jan 2022 · 0 repositories
-
Focus-Driven Contrastive Learning for Medical Question Summarization 16 Jan 2022 · 0 repositories
-
Global Entity Disambiguation with BERT 16 Jan 2022 · 0 repositories
-
GPL: Generative Pseudo Labeling for Unsupervised Domain Adaptation of Dense Retrieval 16 Jan 2022 · 0 repositories
-
Hierarchical Transformers Are More Efficient Language Models 16 Jan 2022 · 0 repositories
-
Identifying the Source of Vulnerability in Fragile Interpretations: A Case Study in Neural Text Classification 16 Jan 2022 · 0 repositories
-
IMPLI: Investigatng NLI Models' Performance on Figurative Language 16 Jan 2022 · 0 repositories
-
Improving Contextual Representation with Gloss Regularized Pre-training 16 Jan 2022 · 0 repositories
-
Investigating and Explaining Feature and Representation Learning in Translationese Classification 16 Jan 2022 · 0 repositories
-
Investigating the saliency of sentiment expressions in aspect-based sentiment analysis 16 Jan 2022 · 0 repositories
-
Jointly Reinforced User Simulator and Task-oriented Dialog System with Simplified Generative Architecture 16 Jan 2022 · 0 repositories
-
KAT: A Knowledge Augmented Transformer for Vision-and-Language 16 Jan 2022 · 0 repositories
-
KD-VLP: Improving End-to-End Vision-and-Language Pretraining with Object Knowledge Distillation 16 Jan 2022 · 0 repositories
-
Language Models for Code-switch Detection of te reo Māori and English in a Low-resource Setting 16 Jan 2022 · 0 repositories
-
Learning to Transpile AMR into SPARQL 16 Jan 2022 · 0 repositories
-
LoPE: Learnable Sinusoidal Positional Encoding for Improving Document Transformer Model 16 Jan 2022 · 0 repositories
-
Magic Pyramid: Accelerating Inference with Early Exiting and Token Pruning 16 Jan 2022 · 0 repositories
-
Measuring Word-Context Biases in Lexical Semantic Datasets 16 Jan 2022 · 0 repositories
-
Memory-assisted prompt editing to improve GPT-3 after deployment 16 Jan 2022 · 1 repository · arXiv:2201.06009Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Minimally-Supervised Relation Induction from Pre-trained Language Model 16 Jan 2022 · 0 repositories
-
Modeling Multi-Granularity Hierarchical Features for Relation Extraction 16 Jan 2022 · 0 repositories
-
Multi-Stage Pre-Training for Math-Understanding: μ²(AL)BERT 16 Jan 2022 · 0 repositories
-
MUST: A Framework for Training Task-oriented Dialogue Systems with Multiple User SimulaTors 16 Jan 2022 · 0 repositories
-
Natural Language Deduction through Search over Statement Compositions 16 Jan 2022 · 0 repositories · arXiv:2201.06028
-
Non-Autoregressive Neural Machine Translation with Consistency Regularization Optimized Variational Framework 16 Jan 2022 · 0 repositories
-
Patching Leaks in the Charformer for Generative Tasks 16 Jan 2022 · 0 repositories
-
PCEE-BERT: Accelerating BERT Inference via Patient and Confident Early Exiting 16 Jan 2022 · 0 repositories
-
Penguins Don’t Fly: Reasoning about Generics through Instantiations and Exceptions 16 Jan 2022 · 0 repositories
-
Polling Latent Opinions: A Method for Computational Sociolinguistics Using Transformer Language Models 16 Jan 2022 · 0 repositories
-
Probing The Linguistic Capacity of Pre-Trained Vision-Language Models 16 Jan 2022 · 0 repositories
-
Progressive Class Semantic Matching for Semi-supervised Text Classification 16 Jan 2022 · 0 repositories
-
Provably Confidential Language Modelling 16 Jan 2022 · 0 repositories
-
Re2G: Retrieve, Rerank, Generate 16 Jan 2022 · 0 repositories
-
Reframing Human-AI Collaboration for Generating Free-Text Explanations 16 Jan 2022 · 0 repositories
-
Representation Learning for Conversational Data using Discourse Mutual Information Maximization 16 Jan 2022 · 0 repositories
-
Revisiting Additive Compositionality: AND, OR, and NOT Operations with Word Embeddings 16 Jan 2022 · 0 repositories
-
Robin: A Novel Online Suicidal Text Corpus of Substantial Breadth and Scale 16 Jan 2022 · 0 repositories
-
Roof-BERT: Divide Understanding Labour and Join in Work 16 Jan 2022 · 0 repositories
-
S5 Framework: A Review of Self-Supervised Shared Semantic Space Optimization for Multimodal Zero-Shot Learning 16 Jan 2022 · 0 repositories
-
SemAttack: Natural Textual Attacks via Different Semantic Spaces 16 Jan 2022 · 0 repositories
-
Seq-GAN-BERT:Sequence Generative Adversarial Learning for Low-resource Name Entity Recognition 16 Jan 2022 · 0 repositories
-
Simple Local Attentions Remain Competitive for Long-Context Tasks 16 Jan 2022 · 0 repositories
-
Surprisingly Simple Adapter Ensembling for Zero-Shot Cross-Lingual Sequence Tagging 16 Jan 2022 · 0 repositories
-
TaCL: Improving BERT Pre-training with Token-aware Contrastive Learning 16 Jan 2022 · 0 repositories
-
Tapping BERT for Preposition Sense Disambiguation 16 Jan 2022 · 0 repositories
-
TEMPLATE: TempRel Classification Model Trained with Embedded Temporal Relation Knowledge 16 Jan 2022 · 0 repositories
-
That is a good looking car !: Visual Aspect based Sentiment Controlled Personalized Response Generation 16 Jan 2022 · 0 repositories
-
Tree Knowledge Distillation for Compressing Transformer-Based Language Models 16 Jan 2022 · 0 repositories
-
Uncovering Surprising Event Boundaries in Narratives 16 Jan 2022 · 0 repositories
-
Understand before Answer: Improve Temporal Reading Comprehension via Precise Question Understanding 16 Jan 2022 · 0 repositories
-
UnifiedSKG: Unifying and Multi-Tasking Structured Knowledge Grounding with Text-to-Text Language Models 16 Jan 2022 · 1 repository · arXiv:2201.05966Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 5 harvested samples) · 1 pointer-only (licence)
-
VEE-BERT: Accelerating BERT Inference for Named Entity Recognition via Vote Early Exiting 16 Jan 2022 · 0 repositories
-
Video Transformers: A Survey 16 Jan 2022 · 0 repositories · arXiv:2201.05991
-
WANLI: Worker and AI Collaboration for Natural Language Inference Dataset Creation 16 Jan 2022 · 1 repository · arXiv:2201.05955
-
What do tokens know about their characters and how do they know it? 16 Jan 2022 · 1 repository
-
What Role Does BERT Play in the Neural Machine Translation Encoder? 16 Jan 2022 · 0 repositories