Methods › Natural Language Processing › Transformers › RoBERTa › Papers, page 6
RoBERTa
Papers archive 2025-07-28
archive papers tagged: 913 · with a code link: 399 · where Syntology ran a sample: 87 (66 with a run with no instrument failure, 21 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (87 of 913 tagged: 66 with a run with no instrument failure, 21 where every run was a failure of Syntology's instrument)
Page 6 of 10: papers 501 to 600 of 913, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Do Transformers know symbolic rules, and would we know if they did? 19 Feb 2022 · 0 repositories · arXiv:2203.00162
-
Automatic Issue Classifier: A Transfer Learning Framework for Classifying Issue Reports 12 Feb 2022 · 1 repository · arXiv:2202.06149
-
pNLP-Mixer: an Efficient all-MLP Architecture for Language 9 Feb 2022 · 1 repository · arXiv:2202.04350
-
Do Language Models Learn Position-Role Mappings? 8 Feb 2022 · 0 repositories · arXiv:2202.03611
-
Logical Reasoning for Task Oriented Dialogue Systems 8 Feb 2022 · 0 repositories · arXiv:2202.04161
-
Memory-Efficient Backpropagation through Large Linear Layers 31 Jan 2022 · 2 repositories · arXiv:2201.13195
-
Black-box Prompt Learning for Pre-trained Language Models 21 Jan 2022 · 1 repository · arXiv:2201.08531Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
Dual Contrastive Learning: Text Classification via Label-Aware Data Augmentation 21 Jan 2022 · 2 repositories · arXiv:2201.08702
-
Applying SoftTriple Loss for Supervised Language Model Fine Tuning 16 Jan 2022 · 0 repositories
-
Bridge the Gap Between CV and NLP! A Gradient-based Textual Adversarial Attack Framework 16 Jan 2022 · 0 repositories
-
DECK: Behavioral Tests to Improve Interpretability and Generalizability of BERT Models Detecting Depression from Text 16 Jan 2022 · 0 repositories
-
IMPLI: Investigatng NLI Models' Performance on Figurative Language 16 Jan 2022 · 0 repositories
-
TEMPLATE: TempRel Classification Model Trained with Embedded Temporal Relation Knowledge 16 Jan 2022 · 0 repositories
-
Uncovering Surprising Event Boundaries in Narratives 16 Jan 2022 · 0 repositories
-
Understand before Answer: Improve Temporal Reading Comprehension via Precise Question Understanding 16 Jan 2022 · 0 repositories
-
What do tokens know about their characters and how do they know it? 16 Jan 2022 · 1 repository
-
Knowledge Graph Augmented Network Towards Multiview Representation Learning for Aspect-based Sentiment Analysis 13 Jan 2022 · 1 repository · arXiv:2201.04831
-
Towards Automated Error Analysis: Learning to Characterize Errors 13 Jan 2022 · 0 repositories · arXiv:2201.05017
-
PromptBERT: Improving BERT Sentence Embeddings with Prompts 12 Jan 2022 · 1 repository · arXiv:2201.04337Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified; the one sample that ran constructed an object rather than computing a result (of 5 harvested samples) · 5 pointer-only (licence)
-
Explaining Predictive Uncertainty by Looking Back at Model Explanations 11 Jan 2022 · 0 repositories · arXiv:2201.03742
-
Black-Box Tuning for Language-Model-as-a-Service 10 Jan 2022 · 2 repositories · arXiv:2201.03514Syntology community repositories only · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
EvoMoE: An Evolutional Mixture-of-Experts Training Framework via Dense-To-Sparse Gate 29 Dec 2021 · 2 repositories · arXiv:2112.14397Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Contextual Sentence Analysis for the Sentiment Prediction on Financial Data 27 Dec 2021 · 0 repositories · arXiv:2112.13790
-
Secondary Use of Clinical Problem List Entries for Neural Network-Based Disease Code Assignment 27 Dec 2021 · 0 repositories · arXiv:2112.13756
-
Evaluating Contextual Embeddings and their Extraction Layers for Depression Assessment 27 Dec 2021 · 0 repositories · arXiv:2112.13795
-
Challenging America: Modeling language in longer time scales 17 Dec 2021 · 0 repositories
-
Knowledge-Augmented Language Models for Cause-Effect Relation Classification 16 Dec 2021 · 1 repository · arXiv:2112.08615Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Applying SoftTriple Loss for Supervised Language Model Fine Tuning 15 Dec 2021 · 0 repositories · arXiv:2112.08462
-
WECHSEL: Effective initialization of subword embeddings for cross-lingual transfer of monolingual language models 13 Dec 2021 · 1 repository · arXiv:2112.06598Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
raceBERT -- A Transformer-based Model for Predicting Race and Ethnicity from Names 7 Dec 2021 · 1 repository · arXiv:2112.03807
-
Bridging Pre-trained Models and Downstream Tasks for Source Code Understanding 4 Dec 2021 · 1 repository · arXiv:2112.02268Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 0 violated, 12 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 13 harvested samples)
-
Sentiment Analysis and Effect of COVID-19 Pandemic using College SubReddit Data 30 Nov 2021 · 1 repository · arXiv:2112.04351
-
Customer Sentiment Analysis using Weak Supervision for Customer-Agent Chat 29 Nov 2021 · 0 repositories · arXiv:2111.14282
-
ANNA: Enhanced Language Representation for Question Answering 16 Nov 2021 · 0 repositories
-
Assessing the Coherence Modeling Capabilities of Pretrained Transformer-based Language Models 16 Nov 2021 · 0 repositories
-
BERT got a Date: Introducing Transformers to Temporal Tagging 16 Nov 2021 · 0 repositories
-
Impact of Tokenization on Language Models: An Analysis for Turkish 16 Nov 2021 · 0 repositories
-
Improved grammatical error correction by ranking elementary edits 16 Nov 2021 · 1 repository
-
Interpreting the Robustness of Neural NLP Models to Textual Perturbations 16 Nov 2021 · 0 repositories
-
Learning to Ignore Adversarial Attacks 16 Nov 2021 · 0 repositories
-
NSP-BERT: A Prompt-based Zero-Shot Learner Through an Original Pre-training Task —— Next Sentence Prediction 16 Nov 2021 · 0 repositories
-
On the Robustness of Reading Comprehension Models to Entity Renaming 16 Nov 2021 · 0 repositories
-
Perturbations in the Wild: Leveraging Human-Written Text Perturbations for Realistic Adversarial Attack and Defense 16 Nov 2021 · 0 repositories
-
PromptBERT: Improving BERT Sentence Embeddings with Prompts 16 Nov 2021 · 0 repositories
-
Unsupervised multiple-choice question generation for out-of-domain Q&A fine-tuning 16 Nov 2021 · 0 repositories
-
Assessing gender bias in medical and scientific masked language models with StereoSet 15 Nov 2021 · 0 repositories · arXiv:2111.08088
-
Character-level HyperNetworks for Hate Speech Detection 11 Nov 2021 · 1 repository · arXiv:2111.06336
-
Improving Large-scale Language Models and Resources for Filipino 11 Nov 2021 · 0 repositories · arXiv:2111.06053
-
Amazon SageMaker Model Parallelism: A General and Flexible Framework for Large Model Training 10 Nov 2021 · 0 repositories · arXiv:2111.05972
-
BagBERT: BERT-based bagging-stacking for multi-topic classification 10 Nov 2021 · 1 repository · arXiv:2111.05808
-
TaCL: Improving BERT Pre-training with Token-aware Contrastive Learning 7 Nov 2021 · 2 repositories · arXiv:2111.04198
-
IBERT: Idiom Cloze-style reading comprehension with Attention 5 Nov 2021 · 0 repositories · arXiv:2112.02994
-
An Empirical Study of the Effectiveness of an Ensemble of Stand-alone Sentiment Detection Tools for Software Engineering Datasets 4 Nov 2021 · 1 repository · arXiv:2111.03196
-
An Empirical Study of Training End-to-End Vision-and-Language Transformers 3 Nov 2021 · 3 repositories · arXiv:2111.02387Syntology community repositories only · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
DSEE: Dually Sparsity-embedded Efficient Tuning of Pre-trained Language Models 30 Oct 2021 · 1 repository · arXiv:2111.00160Syntology official (archive's flag): 9 ran · 10 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 13 harvested samples) · 10 pointer-only (licence)
-
Bridge the Gap Between CV and NLP! A Gradient-based Textual Adversarial Attack Framework 28 Oct 2021 · 1 repository · arXiv:2110.15317
-
Hate and Offensive Speech Detection in Hindi and Marathi 23 Oct 2021 · 0 repositories · arXiv:2110.12200
-
IMPLI: Investigating NLI Models' Performance on Figurative Language 16 Oct 2021 · 0 repositories
-
Models In a Spelling Bee: Language Models Implicitly Learn the Character Composition of Tokens 16 Oct 2021 · 0 repositories
-
On the Robustness of Reading Comprehension Models to Entity Renaming 16 Oct 2021 · 1 repository · arXiv:2110.08555
-
WECHSEL: Effective initialization of subword embeddings for cross-lingual transfer of monolingual language models 16 Oct 2021 · 0 repositories
-
Evaluating the Faithfulness of Importance Measures in NLP by Recursively Masking Allegedly Important Tokens and Retraining 15 Oct 2021 · 1 repository · arXiv:2110.08412Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
Interpreting the Robustness of Neural NLP Models to Textual Perturbations 14 Oct 2021 · 0 repositories · arXiv:2110.07159
-
P-Adapters: Robustly Extracting Factual Information from Language Models with Diverse Prompts 14 Oct 2021 · 1 repository · arXiv:2110.07280
-
Extracting Feelings of People Regarding COVID-19 by Social Network Mining 12 Oct 2021 · 0 repositories · arXiv:2110.06151
-
Leveraging recent advances in Pre-Trained Language Models forEye-Tracking Prediction 9 Oct 2021 · 1 repository · arXiv:2110.04475
-
A Comparative Study of Transformer-Based Language Models on Extractive Question Answering 7 Oct 2021 · 0 repositories · arXiv:2110.03142
-
8-bit Optimizers via Block-wise Quantization 6 Oct 2021 · 3 repositories · arXiv:2110.02861
-
Analyzing the Impact of COVID-19 on Economy from the Perspective of Users Reviews 5 Oct 2021 · 0 repositories · arXiv:2110.02198
-
FoodChem: A food-chemical relation extraction model 5 Oct 2021 · 1 repository · arXiv:2110.02019
-
Exploiting Pre-Trained ASR Models for Alzheimer's Disease Recognition Through Spontaneous Speech 4 Oct 2021 · 0 repositories · arXiv:2110.01493
-
BERT got a Date: Introducing Transformers to Temporal Tagging 30 Sep 2021 · 1 repository · arXiv:2109.14927
-
ScaLA: Speeding-Up Fine-tuning of Pre-trained Transformer Networks via Efficient and Scalable Adversarial Perturbation 29 Sep 2021 · 0 repositories
-
Effective Use of Graph Convolution Network and Contextual Sub-Tree forCommodity News Event Extraction 27 Sep 2021 · 1 repository · arXiv:2109.12781
-
Finetuning Transformer Models to Build ASAG System 25 Sep 2021 · 0 repositories · arXiv:2109.12300
-
BERTweetFR : Domain Adaptation of Pre-Trained Language Models for French Tweets 21 Sep 2021 · 0 repositories · arXiv:2109.10234
-
MirrorWiC: On Eliciting Word-in-Context Representations from Pretrained Language Models 19 Sep 2021 · 1 repository · arXiv:2109.09237
-
Fine-Tuned Transformers Show Clusters of Similar Representations Across Layers 17 Sep 2021 · 0 repositories · arXiv:2109.08406
-
Grounding Natural Language Instructions: Can Large Language Models Capture Spatial Information? 17 Sep 2021 · 1 repository · arXiv:2109.08634
-
The futility of STILTs for the classification of lexical borrowings in Spanish 17 Sep 2021 · 0 repositories · arXiv:2109.08607
-
Efficient Domain Adaptation of Language Models via Adaptive Tokenization 15 Sep 2021 · 0 repositories · arXiv:2109.07460
-
Legal Transformer Models May Not Always Help 14 Sep 2021 · 0 repositories · arXiv:2109.06862
-
Question Answering over Electronic Devices: A New Benchmark Dataset and a Multi-Task Learning based QA Framework 13 Sep 2021 · 1 repository · arXiv:2109.05897Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples)
-
TEASEL: A Transformer-Based Speech-Prefixed Language Model 12 Sep 2021 · 1 repository · arXiv:2109.05522Syntology 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples)
-
D-REX: Dialogue Relation Extraction with Explanations 10 Sep 2021 · 1 repository · arXiv:2109.05126
-
Mining Points of Interest via Address Embeddings: An Unsupervised Approach 9 Sep 2021 · 0 repositories · arXiv:2109.04467
-
Multi-granularity Textual Adversarial Attack with Behavior Cloning 9 Sep 2021 · 1 repository · arXiv:2109.04367
-
Word-Level Coreference Resolution 9 Sep 2021 · 1 repository · arXiv:2109.04127
-
NSP-BERT: A Prompt-based Few-Shot Learner Through an Original Pre-training Task--Next Sentence Prediction 8 Sep 2021 · 1 repository · arXiv:2109.03564
-
Empathetic Dialogue Generation with Pre-trained RoBERTa-GPT2 and External Knowledge 7 Sep 2021 · 0 repositories · arXiv:2109.03004
-
How much pretraining data do language models need to learn syntax? 7 Sep 2021 · 0 repositories · arXiv:2109.03160
-
Error Detection in Large-Scale Natural Language Understanding Systems Using Transformer Models 4 Sep 2021 · 0 repositories · arXiv:2109.01754
-
ConQX: Semantic Expansion of Spoken Queries for Intent Detection based on Conditioned Text Generation 2 Sep 2021 · 0 repositories · arXiv:2109.00729
-
So Cloze yet so Far: N400 Amplitude is Better Predicted by Distributional Information than Human Predictability Judgements 2 Sep 2021 · 0 repositories · arXiv:2109.01226
-
Fight Fire with Fire: Fine-tuning Hate Detectors using Large Samples of Generated Hate Speech 1 Sep 2021 · 0 repositories · arXiv:2109.00591
-
Towards Improving Adversarial Training of NLP Models 1 Sep 2021 · 1 repository · arXiv:2109.00544Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Evaluating the Robustness of Neural Language Models to Input Perturbations 27 Aug 2021 · 1 repository · arXiv:2108.12237Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
A Statutory Article Retrieval Dataset in French 26 Aug 2021 · 1 repository · arXiv:2108.11792
-
EmoBERTa: Speaker-Aware Emotion Recognition in Conversation with RoBERTa 26 Aug 2021 · 1 repository · arXiv:2108.12009
-
Rethinking Why Intermediate-Task Fine-Tuning Works 26 Aug 2021 · 1 repository · arXiv:2108.11696