Methods › Natural Language Processing › Transformers › RoBERTa › Papers, page 9
RoBERTa
Papers archive 2025-07-28
archive papers tagged: 913 · with a code link: 399 · where Syntology ran a sample: 87 (66 with a run with no instrument failure, 21 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (87 of 913 tagged: 66 with a run with no instrument failure, 21 where every run was a failure of Syntology's instrument)
Page 9 of 10: papers 801 to 900 of 913, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Data-Efficient Pretraining via Contrastive Self-Supervision 2 Oct 2020 · 0 repositories · arXiv:2010.01061
-
BET: A Backtranslation Approach for Easy Data Augmentation in Transformer-based Paraphrase Identification Context 25 Sep 2020 · 1 repository · arXiv:2009.12452
-
Constructing interval variables via faceted Rasch measurement and multitask deep learning: a hate speech application 22 Sep 2020 · 2 repositories · arXiv:2009.10277Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
On Data Augmentation for Extreme Multi-label Classification 22 Sep 2020 · 0 repositories · arXiv:2009.10778
-
Compositional and Lexical Semantics in RoBERTa, BERT and DistilBERT: A Case Study on CoQA 17 Sep 2020 · 0 repositories · arXiv:2009.08257
-
Efficient Transformer-based Large Scale Language Representations using Hardware-friendly Block Structured Pruning 17 Sep 2020 · 0 repositories · arXiv:2009.08065
-
Solomon at SemEval-2020 Task 11: Ensemble Architecture for Fine-Tuned Propaganda Detection in News Articles 16 Sep 2020 · 0 repositories · arXiv:2009.07473
-
BoostingBERT:Integrating Multi-Class Boosting into BERT for NLP Tasks 13 Sep 2020 · 0 repositories · arXiv:2009.05959
-
CIA_NITT at WNUT-2020 Task 2: Classification of COVID-19 Tweets Using Pre-trained Language Models 12 Sep 2020 · 0 repositories · arXiv:2009.05782
-
Compressed Deep Networks: Goodbye SVD, Hello Robust Low-Rank Approximation 11 Sep 2020 · 1 repository · arXiv:2009.05647
-
UPB at SemEval-2020 Task 6: Pretrained Language Models for Definition Extraction 11 Sep 2020 · 3 repositories · arXiv:2009.05603
-
Do Response Selection Models Really Know What's Next? Utterance Manipulation Strategies for Multi-turn Response Selection 10 Sep 2020 · 1 repository · arXiv:2009.04703
-
Comparative Study of Language Models on Cross-Domain Data with Model Agnostic Explainability 9 Sep 2020 · 0 repositories · arXiv:2009.04095
-
ERNIE at SemEval-2020 Task 10: Learning Word Emphasis Selection by Pre-trained Language Model 8 Sep 2020 · 0 repositories · arXiv:2009.03706
-
EdinburghNLP at WNUT-2020 Task 2: Leveraging Transformers with Generalized Augmentation for Identifying Informativeness in COVID-19 Tweets 6 Sep 2020 · 0 repositories · arXiv:2009.06375
-
QiaoNing at SemEval-2020 Task 4: Commonsense Validation and Explanation system based on ensemble of language model 6 Sep 2020 · 0 repositories · arXiv:2009.02645
-
Accenture at CheckThat! 2020: If you say so: Post-hoc fact-checking of claims using transformer-based models 5 Sep 2020 · 0 repositories · arXiv:2009.02431
-
Conceptualized Representation Learning for Chinese Biomedical Text Mining 25 Aug 2020 · 0 repositories · arXiv:2008.10813
-
ETC-NLG: End-to-end Topic-Conditioned Natural Language Generation 25 Aug 2020 · 1 repository · arXiv:2008.10875
-
HinglishNLP: Fine-tuned Language Models for Hinglish Sentiment Detection 22 Aug 2020 · 2 repositories · arXiv:2008.09820
-
KR-BERT: A Small-Scale Korean-Specific Language Model 10 Aug 2020 · 1 repository · arXiv:2008.03979
-
SemEval-2020 Task 10: Emphasis Selection for Written Text in Visual Media 7 Aug 2020 · 0 repositories · arXiv:2008.03274
-
aschern at SemEval-2020 Task 11: It Takes Three to Tango: RoBERTa, CRF, and Transfer Learning 6 Aug 2020 · 1 repository · arXiv:2008.02837
-
Composer Style Classification of Piano Sheet Music Images Using Language Model Pretraining 29 Jul 2020 · 1 repository · arXiv:2007.14587
-
BUT-FIT at SemEval-2020 Task 5: Automatic detection of counterfactual statements with deep pre-trained language representation models 28 Jul 2020 · 1 repository · arXiv:2007.14128
-
Variants of BERT, Random Forests and SVM approach for Multimodal Emotion-Target Sub-challenge 28 Jul 2020 · 0 repositories · arXiv:2007.13928
-
newsSweeper at SemEval-2020 Task 11: Context-Aware Rich Feature Representations For Propaganda Classification 21 Jul 2020 · 1 repository · arXiv:2007.10827
-
problemConquero at SemEval-2020 Task 12: Transformer and Soft label-based approaches 21 Jul 2020 · 1 repository · arXiv:2007.10877
-
AdapterHub: A Framework for Adapting Transformers 15 Jul 2020 · 9 repositories · arXiv:2007.07779Syntology official (archive's flag): 1 ran · 11 ran (of which 2 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 15 harvested samples) · 12 pointer-only (licence)
-
Contrastive Code Representation Learning 9 Jul 2020 · 1 repository · arXiv:2007.04973
-
Deep Contextual Embeddings for Address Classification in E-commerce 6 Jul 2020 · 0 repositories · arXiv:2007.03020
-
Robust Prediction of Punctuation and Truecasing for Medical ASR 4 Jul 2020 · 0 repositories · arXiv:2007.02025
-
A Transformer Approach to Contextual Sarcasm Detection in Twitter 1 Jul 2020 · 0 repositories
-
How does BERT's attention change when you fine-tune? An analysis methodology and a case study in negation scope 1 Jul 2020 · 0 repositories
-
IlliniMet: Illinois System for Metaphor Detection with Contextual and Linguistic Information 1 Jul 2020 · 0 repositories
-
Intermediate-Task Transfer Learning with Pretrained Language Models: When and Why Does It Work? 1 Jul 2020 · 0 repositories
-
Modelling Context and Syntactical Features for Aspect-based Sentiment Analysis 1 Jul 2020 · 1 repository
-
Neural Sarcasm Detection using Conversation Context 1 Jul 2020 · 0 repositories
-
RobertNLP at the IWPT 2020 Shared Task: Surprisingly Simple Enhanced UD Parsing for English 1 Jul 2020 · 0 repositories
-
Transformers on Sarcasm Detection with Context 1 Jul 2020 · 0 repositories
-
Want to Identify, Extract and Normalize Adverse Drug Reactions in Tweets? Use RoBERTa 29 Jun 2020 · 0 repositories · arXiv:2006.16146
-
On the Stability of Fine-tuning BERT: Misconceptions, Explanations, and Strong Baselines 8 Jun 2020 · 2 repositories · arXiv:2006.04884Syntology official (archive's flag): 8 ran · 13 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 3 honoured, 0 violated, 6 with no contract checked; 4 where Syntology's instrument failed) · 7 unverified (of 20 harvested samples) · 3 pointer-only (licence)
-
Medical Concept Normalization in User Generated Texts by Learning Target Concept Embeddings 7 Jun 2020 · 0 repositories · arXiv:2006.04014
-
DeBERTa: Decoding-enhanced BERT with Disentangled Attention 5 Jun 2020 · 14 repositories · arXiv:2006.03654Syntology official: no sample here; runs from other or unrecorded repositories · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 2 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 9 unverified (of 13 harvested samples) · 3 pointer-only (licence)
-
BERT-based Ensembles for Modeling Disclosure and Support in Conversational Social Media Text 1 Jun 2020 · 0 repositories · arXiv:2006.01222
-
Emergence of Separable Manifolds in Deep Language Representations 1 Jun 2020 · 1 repository · arXiv:2006.01095
-
Language Representation Models for Fine-Grained Sentiment Classification 27 May 2020 · 1 repository · arXiv:2005.13619
-
L2R2: Leveraging Ranking for Abductive Reasoning 22 May 2020 · 1 repository · arXiv:2005.11223
-
Robust Layout-aware IE for Visually Rich Documents with Pre-trained Language Models 22 May 2020 · 0 repositories · arXiv:2005.11017
-
BERTweet: A pre-trained language model for English Tweets 20 May 2020 · 3 repositories · arXiv:2005.10200
-
Adversarial Training for Commonsense Inference 17 May 2020 · 1 repository · arXiv:2005.08156
-
On the Robustness of Language Encoders against Grammatical Errors 12 May 2020 · 1 repository · arXiv:2005.05683Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
How Context Affects Language Models' Factual Predictions 10 May 2020 · 0 repositories · arXiv:2005.04611
-
Beyond Accuracy: Behavioral Testing of NLP models with CheckList 8 May 2020 · 4 repositories · arXiv:2005.04118Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
Unsupervised Alignment-based Iterative Evidence Retrieval for Multi-hop Question Answering 4 May 2020 · 1 repository · arXiv:2005.01218
-
Birds have four legs?! NumerSense: Probing Numerical Commonsense Knowledge of Pre-trained Language Models 2 May 2020 · 0 repositories · arXiv:2005.00683
-
IsoBN: Fine-Tuning BERT with Isotropic Batch Normalization 2 May 2020 · 1 repository · arXiv:2005.02178
-
Aggression Identification in English, Hindi and Bangla Text using BERT, RoBERTa and SVM 1 May 2020 · 1 repository
-
Discourse-Aware Unsupervised Summarization of Long Scientific Documents 1 May 2020 · 1 repository · arXiv:2005.00513
-
Intermediate-Task Transfer Learning with Pretrained Models for Natural Language Understanding: When and Why Does It Work? 1 May 2020 · 0 repositories · arXiv:2005.00628
-
Segatron: Segment-Aware Transformer for Language Modeling and Understanding 30 Apr 2020 · 1 repository · arXiv:2004.14996
-
Revisiting Pre-Trained Models for Chinese Natural Language Processing 29 Apr 2020 · 6 repositories · arXiv:2004.13922Syntology community repositories only · 21 ran (of which 1 constructed an object rather than computing a result; 16 with no instrument failure: 3 honoured, 0 violated, 13 with no contract checked; 5 where Syntology's instrument failed) · 14 unverified (of 35 harvested samples) · 4 pointer-only (licence)
-
Classification of Cuisines from Sequentially Structured Recipes 26 Apr 2020 · 0 repositories · arXiv:2004.14165
-
Masking as an Efficient Alternative to Finetuning for Pretrained Language Models 26 Apr 2020 · 0 repositories · arXiv:2004.12406
-
New Protocols and Negative Results for Textual Entailment Data Collection 24 Apr 2020 · 1 repository · arXiv:2004.11997
-
Contextualized Representations Using Textual Encyclopedic Knowledge 24 Apr 2020 · 0 repositories · arXiv:2004.12006
-
UHH-LT at SemEval-2020 Task 12: Fine-Tuning of Pre-Trained Transformer Networks for Offensive Language Detection 23 Apr 2020 · 0 repositories · arXiv:2004.11493
-
Residual Energy-Based Models for Text Generation 22 Apr 2020 · 1 repository · arXiv:2004.11714
-
Adversarial Training for Large Neural Language Models 20 Apr 2020 · 3 repositories · arXiv:2004.08994
-
StereoSet: Measuring stereotypical bias in pretrained language models 20 Apr 2020 · 3 repositories · arXiv:2004.09456
-
Learning-to-Rank with BERT in TF-Ranking 17 Apr 2020 · 0 repositories · arXiv:2004.08476
-
Training with Quantization Noise for Extreme Model Compression 15 Apr 2020 · 4 repositories · arXiv:2004.07320
-
A Simple Yet Strong Pipeline for HotpotQA 14 Apr 2020 · 0 repositories · arXiv:2004.06753
-
Robustly Pre-trained Neural Model for Direct Temporal Relation Extraction 13 Apr 2020 · 0 repositories · arXiv:2004.06216
-
DynaBERT: Dynamic BERT with Adaptive Width and Depth 8 Apr 2020 · 3 repositories · arXiv:2004.04037Syntology official (archive's flag): 5 ran · 5 ran (of which 1 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 0 violated, 1 with no contract checked; 3 where Syntology's instrument failed) · 3 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
On the Effect of Dropping Layers of Pre-trained Transformer Models 8 Apr 2020 · 4 repositories · arXiv:2004.03844Syntology official (archive's flag): 5 ran · 7 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 0 violated, 3 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples) · 1 pointer-only (licence)
-
Severing the Edge Between Before and After: Neural Architectures for Temporal Ordering of Events 8 Apr 2020 · 0 repositories · arXiv:2004.04295
-
TextGAIL: Generative Adversarial Imitation Learning for Text Generation 7 Apr 2020 · 0 repositories · arXiv:2004.13796
-
Transformers to Learn Hierarchical Contexts in Multiparty Dialogue for Span-based Question Answering 7 Apr 2020 · 1 repository · arXiv:2004.03561
-
Continual Domain-Tuning for Pretrained Language Models 5 Apr 2020 · 0 repositories · arXiv:2004.02288
-
Gestalt: a Stacking Ensemble for SQuAD2.0 2 Apr 2020 · 0 repositories · arXiv:2004.07067
-
Deep Entity Matching with Pre-Trained Language Models 1 Apr 2020 · 1 repository · arXiv:2004.00584Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
ELECTRA: Pre-training Text Encoders as Discriminators Rather Than Generators 23 Mar 2020 · 19 repositories · arXiv:2003.10555Syntology official: no sample here; runs from other or unrecorded repositories · 31 ran (of which 7 constructed an object rather than computing a result; 18 with no instrument failure: 2 honoured, 2 violated, 14 with no contract checked; 13 where Syntology's instrument failed) · 9 unverified (of 40 harvested samples) · 10 pointer-only (licence)
-
Calibration of Pre-trained Transformers 17 Mar 2020 · 1 repository · arXiv:2003.07892
-
HypoNLI: Exploring the Artificial Patterns of Hypothesis-only Bias in Natural Language Inference 5 Mar 2020 · 0 repositories · arXiv:2003.02756
-
jiant: A Software Toolkit for Research on General-Purpose Text Understanding Models 4 Mar 2020 · 6 repositories · arXiv:2003.02249Syntology official (archive's flag): 6 ran · 6 ran (of which 3 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 8 harvested samples)
-
The Microsoft Toolkit of Multi-Task Deep Neural Networks for Natural Language Understanding 19 Feb 2020 · 3 repositories · arXiv:2002.07972
-
Stress Test Evaluation of Transformer-based Models in Natural Language Understanding Tasks 14 Feb 2020 · 0 repositories · arXiv:2002.06261
-
Application of Pre-training Models in Named Entity Recognition 9 Feb 2020 · 0 repositories · arXiv:2002.08902
-
K-Adapter: Infusing Knowledge into Pre-Trained Models with Adapters 5 Feb 2020 · 2 repositories · arXiv:2002.01808Syntology 11 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 1 honoured, 0 violated, 5 with no contract checked; 5 where Syntology's instrument failed) · 3 unverified (of 14 harvested samples) · 2 pointer-only (licence)
-
Beat the AI: Investigating Adversarial Human Annotation for Reading Comprehension 2 Feb 2020 · 1 repository · arXiv:2002.00293
-
RobBERT: a Dutch RoBERTa-based Language Model 17 Jan 2020 · 1 repository · arXiv:2001.06286Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Resolving the Scope of Speculation and Negation using Transformer-Based Architectures 9 Jan 2020 · 1 repository · arXiv:2001.02885
-
BERT-AL: BERT for Arbitrarily Long Document Understanding 1 Jan 2020 · 0 repositories
-
oLMpics -- On what Language Model Pre-training Captures 31 Dec 2019 · 2 repositories · arXiv:1912.13283
-
WaLDORf: Wasteless Language-model Distillation On Reading-comprehension 13 Dec 2019 · 0 repositories · arXiv:1912.06638
-
BERT has a Moral Compass: Improvements of ethical and moral values of machines 11 Dec 2019 · 0 repositories · arXiv:1912.05238
-
Do Attention Heads in BERT Track Syntactic Dependencies? 27 Nov 2019 · 1 repository · arXiv:1911.12246
-
Evaluating Commonsense in Pre-trained Language Models 27 Nov 2019 · 1 repository · arXiv:1911.11931
-
Taking a Stance on Fake News: Towards Automatic Disinformation Assessment via Deep Bidirectional Transformer Language Models for Stance Detection 27 Nov 2019 · 0 repositories · arXiv:1911.11951