Methods › Natural Language Processing › Transformers › XLNet › Papers, page 2
XLNet
Papers archive 2025-07-28
archive papers tagged: 167 · with a code link: 70 · where Syntology ran a sample: 10 (10 with a run with no instrument failure, 0 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (10 of 167 tagged: 10 with a run with no instrument failure, 0 where every run was a failure of Syntology's instrument)
Page 2 of 2: papers 101 to 167 of 167, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
KERMIT: Complementing Transformer Architectures with Encoders of Explicit Syntactic Interpretations 1 Nov 2020 · 1 repository
-
ERNIE-Gram: Pre-Training with Explicitly N-Gram Masked Language Modeling for Natural Language Understanding 23 Oct 2020 · 2 repositories · arXiv:2010.12148
-
AutoMeTS: The Autocomplete for Medical Text Simplification 20 Oct 2020 · 1 repository · arXiv:2010.10573
-
Performance of Transfer Learning Model vs. Traditional Neural Network in Low System Resource Environment 20 Oct 2020 · 0 repositories · arXiv:2011.07962
-
NUIG-Shubhanker@Dravidian-CodeMix-FIRE2020: Sentiment Analysis of Code-Mixed Dravidian text using XLNet 15 Oct 2020 · 0 repositories · arXiv:2010.07773
-
Aspect-based Document Similarity for Research Papers 13 Oct 2020 · 1 repository · arXiv:2010.06395Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Interpreting Attention Models with Human Visual Attention in Machine Reading Comprehension 13 Oct 2020 · 0 repositories · arXiv:2010.06396
-
Automated Concatenation of Embeddings for Structured Prediction 10 Oct 2020 · 2 repositories · arXiv:2010.05006
-
Analyzing Individual Neurons in Pre-trained Language Models 6 Oct 2020 · 1 repository · arXiv:2010.02695
-
How Effective is Task-Agnostic Data Augmentation for Pretrained Transformers? 5 Oct 2020 · 0 repositories · arXiv:2010.01764
-
PUM at SemEval-2020 Task 12: Aggregation of Transformer-based models' features for offensive language recognition 5 Oct 2020 · 0 repositories · arXiv:2010.01897
-
Examining the rhetorical capacities of neural language models 1 Oct 2020 · 0 repositories · arXiv:2010.00153
-
Accelerating Multi-Model Inference by Merging DNNs of Different Weights 28 Sep 2020 · 0 repositories · arXiv:2009.13062
-
BET: A Backtranslation Approach for Easy Data Augmentation in Transformer-based Paraphrase Identification Context 25 Sep 2020 · 1 repository · arXiv:2009.12452
-
Weird AI Yankovic: Generating Parody Lyrics 25 Sep 2020 · 0 repositories · arXiv:2009.12240
-
Fine-tuning Pre-trained Contextual Embeddings for Citation Content Analysis in Scholarly Publication 12 Sep 2020 · 0 repositories · arXiv:2009.05836
-
Compressed Deep Networks: Goodbye SVD, Hello Robust Low-Rank Approximation 11 Sep 2020 · 1 repository · arXiv:2009.05647
-
UPB at SemEval-2020 Task 6: Pretrained Language Models for Definition Extraction 11 Sep 2020 · 3 repositories · arXiv:2009.05603
-
EdinburghNLP at WNUT-2020 Task 2: Leveraging Transformers with Generalized Augmentation for Identifying Informativeness in COVID-19 Tweets 6 Sep 2020 · 0 repositories · arXiv:2009.06375
-
QiaoNing at SemEval-2020 Task 4: Commonsense Validation and Explanation system based on ensemble of language model 6 Sep 2020 · 0 repositories · arXiv:2009.02645
-
A Multitask Deep Learning Approach for User Depression Detection on Sina Weibo 26 Aug 2020 · 0 repositories · arXiv:2008.11708
-
KR-BERT: A Small-Scale Korean-Specific Language Model 10 Aug 2020 · 1 repository · arXiv:2008.03979
-
Multi-node Bert-pretraining: Cost-efficient Approach 1 Aug 2020 · 0 repositories · arXiv:2008.00177
-
Neural Machine Translation with Error Correction 21 Jul 2020 · 1 repository · arXiv:2007.10681
-
A Transformer Approach to Contextual Sarcasm Detection in Twitter 1 Jul 2020 · 0 repositories
-
Detecting Sarcasm in Conversation Context Using Transformer-Based Models 1 Jul 2020 · 0 repositories
-
Metaphor Detection Using Contextual Word Embeddings From Transformers 1 Jul 2020 · 0 repositories
-
Transferring Monolingual Model to Low-Resource Language: The Case of Tigrinya 13 Jun 2020 · 0 repositories · arXiv:2006.07698
-
A Comparative Study of Lexical Substitution Approaches based on Neural Language Models 29 May 2020 · 0 repositories · arXiv:2006.00031
-
Using Large Pretrained Language Models for Answering User Queries from Product Specifications 29 May 2020 · 0 repositories · arXiv:2005.14613
-
ImpactCite: An XLNet-based method for Citation Impact Analysis 5 May 2020 · 1 repository · arXiv:2005.06611
-
DeFormer: Decomposing Pre-trained Transformers for Faster Question Answering 2 May 2020 · 1 repository · arXiv:2005.00697Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 6 harvested samples)
-
Cross-lingual Information Retrieval with BERT 24 Apr 2020 · 1 repository · arXiv:2004.13005
-
MPNet: Masked and Permuted Pre-training for Language Understanding 20 Apr 2020 · 7 repositories · arXiv:2004.09297Syntology official (archive's flag): 2 ran · 8 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 1 honoured, 0 violated, 4 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 8 harvested samples) · 6 pointer-only (licence)
-
StereoSet: Measuring stereotypical bias in pretrained language models 20 Apr 2020 · 3 repositories · arXiv:2004.09456
-
Analyzing Redundancy in Pretrained Transformer Models 8 Apr 2020 · 1 repository · arXiv:2004.04010
-
On the Effect of Dropping Layers of Pre-trained Transformer Models 8 Apr 2020 · 4 repositories · arXiv:2004.03844Syntology official (archive's flag): 5 ran · 7 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 0 violated, 3 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples) · 1 pointer-only (licence)
-
ELECTRA: Pre-training Text Encoders as Discriminators Rather Than Generators 23 Mar 2020 · 19 repositories · arXiv:2003.10555Syntology official: no sample here; runs from other or unrecorded repositories · 31 ran (of which 7 constructed an object rather than computing a result; 18 with no instrument failure: 2 honoured, 2 violated, 14 with no contract checked; 13 where Syntology's instrument failed) · 9 unverified (of 40 harvested samples) · 10 pointer-only (licence)
-
Pairwise Multi-Class Document Classification for Semantic Relations between Wikipedia Articles 22 Mar 2020 · 4 repositories · arXiv:2003.09881
-
Stress Test Evaluation of Transformer-based Models in Natural Language Understanding Tasks 14 Feb 2020 · 0 repositories · arXiv:2002.06261
-
Resolving the Scope of Speculation and Negation using Transformer-Based Architectures 9 Jan 2020 · 1 repository · arXiv:2001.02885
-
BERT-AL: BERT for Arbitrarily Long Document Understanding 1 Jan 2020 · 0 repositories
-
Clinical XLNet: Modeling Sequential Clinical Notes and Predicting Prolonged Mechanical Ventilation 27 Dec 2019 · 3 repositories · arXiv:1912.11975
-
WaLDORf: Wasteless Language-model Distillation On Reading-comprehension 13 Dec 2019 · 0 repositories · arXiv:1912.06638
-
An Exploration of Data Augmentation and Sampling Techniques for Domain-Agnostic Question Answering 4 Dec 2019 · 0 repositories · arXiv:1912.02145
-
Evaluating Commonsense in Pre-trained Language Models 27 Nov 2019 · 1 repository · arXiv:1911.11931
-
Low Rank Factorization for Compact Multi-Head Self-Attention 26 Nov 2019 · 1 repository · arXiv:1912.00835
-
Attending to Entities for Better Text Understanding 11 Nov 2019 · 0 repositories · arXiv:1911.04361
-
FASPell: A Fast, Adaptable, Simple, Powerful Chinese Spell Checker Based On DAE-Decoder Paradigm 1 Nov 2019 · 1 repository
-
Generalizing Question Answering System with Pre-trained Language Model Fine-tuning 1 Nov 2019 · 0 repositories
-
IIT-KGP at COIN 2019: Using pre-trained Language Models for modeling Machine Comprehension 1 Nov 2019 · 0 repositories
-
Pingan Smart Health and SJTU at COIN - Shared Task: utilizing Pre-trained Language Models and Common-sense Knowledge in Machine Reading Tasks 1 Nov 2019 · 0 repositories
-
Transfer Learning from Transformers to Fake News Challenge Stance Detection (FNC-1) Task 31 Oct 2019 · 0 repositories · arXiv:1910.14353
-
Modeling Inter-Speaker Relationship in XLNet for Contextual Spoken Language Understanding 28 Oct 2019 · 0 repositories · arXiv:1910.12531
-
Speech-XLNet: Unsupervised Acoustic Model Pretraining For Self-Attention Networks 23 Oct 2019 · 0 repositories · arXiv:1910.10387
-
XL-Editor: Post-editing Sentences with XLNet 19 Oct 2019 · 0 repositories · arXiv:1910.10479
-
Multilingual Question Answering from Formatted Text applied to Conversational Agents 10 Oct 2019 · 0 repositories · arXiv:1910.04659
-
Extremely Small BERT Models from Mixed-Vocabulary Training 25 Sep 2019 · 0 repositories · arXiv:1909.11687
-
Language models and Automated Essay Scoring 18 Sep 2019 · 1 repository · arXiv:1909.09482
-
Frustratingly Easy Natural Question Answering 11 Sep 2019 · 0 repositories · arXiv:1909.05286
-
Reasoning Over Semantic-Level Graph for Fact Checking 9 Sep 2019 · 0 repositories · arXiv:1909.03745
-
Transfer Learning Robustness in Multi-Class Categorization by Fine-Tuning Pre-Trained Contextualized Language Models 8 Sep 2019 · 1 repository · arXiv:1909.03564
-
Integrating Multimodal Information in Large Pretrained Transformers 15 Aug 2019 · 1 repository · arXiv:1908.05787
-
Scalable Attentive Sentence-Pair Modeling via Distilled Sentence Embedding 14 Aug 2019 · 1 repository · arXiv:1908.05161
-
BioFLAIR: Pretrained Pooled Contextualized Embeddings for Biomedical Sequence Labeling Tasks 13 Aug 2019 · 1 repository · arXiv:1908.05760
-
ERNIE 2.0: A Continual Pre-training Framework for Language Understanding 29 Jul 2019 · 3 repositories · arXiv:1907.12412Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
XLNet: Generalized Autoregressive Pretraining for Language Understanding 19 Jun 2019 · 27 repositories · arXiv:1906.08237Syntology official (archive's flag): 1 ran · 15 ran (of which 2 constructed an object rather than computing a result; 10 with no instrument failure: 2 honoured, 0 violated, 8 with no contract checked; 5 where Syntology's instrument failed) · 9 unverified (of 24 harvested samples) · 4 pointer-only (licence)