Methods › Natural Language Processing › Transformers › RoBERTa › Papers, page 5
RoBERTa
Papers archive 2025-07-28
archive papers tagged: 913 · with a code link: 399 · where Syntology ran a sample: 87 (66 with a run with no instrument failure, 21 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (87 of 913 tagged: 66 with a run with no instrument failure, 21 where every run was a failure of Syntology's instrument)
Page 5 of 10: papers 401 to 500 of 913, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Empowering Language Models with Knowledge Graph Reasoning for Question Answering 15 Nov 2022 · 0 repositories · arXiv:2211.08380
-
Xu at SemEval-2022 Task 4: Pre-BERT Neural Network Methods vs Post-BERT RoBERTa Approach for Patronizing and Condescending Language Detection 13 Nov 2022 · 1 repository · arXiv:2211.06874
-
Dark patterns in e-commerce: a dataset and its baseline evaluations 12 Nov 2022 · 1 repository · arXiv:2211.06543
-
Using Persuasive Writing Strategies to Explain and Detect Health Misinformation 11 Nov 2022 · 1 repository · arXiv:2211.05985
-
Biomedical Multi-hop Question Answering Using Knowledge Graph Embeddings and Language Models 10 Nov 2022 · 0 repositories · arXiv:2211.05351
-
Collateral facilitation in humans and language models 9 Nov 2022 · 1 repository · arXiv:2211.05198Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Mask More and Mask Later: Efficient Pre-training of Masked Language Models by Disentangling the [MASK] Token 9 Nov 2022 · 1 repository · arXiv:2211.04898
-
Discover, Explanation, Improvement: An Automatic Slice Detection Framework for Natural Language Processing 8 Nov 2022 · 0 repositories · arXiv:2211.04476
-
Exploiting prompt learning with pre-trained language models for Alzheimer's Disease detection 29 Oct 2022 · 1 repository · arXiv:2210.16539
-
Probing for targeted syntactic knowledge through grammatical error detection 28 Oct 2022 · 1 repository · arXiv:2210.16228
-
Exploring Robustness of Prefix Tuning in Noisy Data: A Case Study in Financial Sentiment Analysis 26 Oct 2022 · 0 repositories · arXiv:2211.05584
-
Exploring Euphemism Detection in Few-Shot and Zero-Shot Settings 24 Oct 2022 · 1 repository · arXiv:2210.12926
-
The Better Your Syntax, the Better Your Semantics? Probing Pretrained Language Models for the English Comparative Correlative 24 Oct 2022 · 0 repositories · arXiv:2210.13181
-
Data Augmentation for Automated Essay Scoring using Transformer Models 23 Oct 2022 · 0 repositories · arXiv:2210.12809
-
Tempo: Accelerating Transformer-Based Model Training through Memory Footprint Reduction 19 Oct 2022 · 1 repository · arXiv:2210.10246Syntology official: harvested, nothing ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
ELASTIC: Numerical Reasoning with Adaptive Symbolic Compiler 18 Oct 2022 · 1 repository · arXiv:2210.10105Syntology official (archive's flag): 2 ran · 2 ran (of which 1 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Zero-Shot Ranking Socio-Political Texts with Transformer Language Models to Reduce Close Reading Time 17 Oct 2022 · 0 repositories · arXiv:2210.09179
-
DyLoRA: Parameter Efficient Tuning of Pre-trained Models using Dynamic Search-Free Low-Rank Adaptation 14 Oct 2022 · 2 repositories · arXiv:2210.07558Syntology official: not harvested · 0 ran · 1 unverified (of 1 harvested sample)
-
Empowering the Fact-checkers! Automatic Identification of Claim Spans on Twitter 10 Oct 2022 · 1 repository · arXiv:2210.04710
-
On Task-Adaptive Pretraining for Dialogue Response Selection 8 Oct 2022 · 0 repositories · arXiv:2210.04073
-
Generating Quizzes to Support Training on Quality Management and Assurance in Space Science and Engineering 7 Oct 2022 · 0 repositories · arXiv:2210.03427
-
Matching Text and Audio Embeddings: Exploring Transfer-learning Strategies for Language-based Audio Retrieval 6 Oct 2022 · 0 repositories · arXiv:2210.02833
-
Revisiting Structured Dropout 5 Oct 2022 · 0 repositories · arXiv:2210.02570
-
The (In)Effectiveness of Intermediate Task Training For Domain Adaptation and Cross-Lingual Transfer Learning 3 Oct 2022 · 0 repositories · arXiv:2210.01091
-
Downstream Datasets Make Surprisingly Good Pretraining Corpora 28 Sep 2022 · 1 repository · arXiv:2209.14389
-
Supervised Contrastive Learning as Multi-Objective Optimization for Fine-Tuning Large Pre-trained Language Models 28 Sep 2022 · 0 repositories · arXiv:2209.14161
-
YATO: Yet Another deep learning based Text analysis Open toolkit 28 Sep 2022 · 1 repository · arXiv:2209.13877
-
Extractive Question Answering on Queries in Hindi and Tamil 27 Sep 2022 · 0 repositories · arXiv:2210.06356
-
Do ever larger octopi still amplify reporting biases? Evidence from judgments of typical colour 26 Sep 2022 · 0 repositories · arXiv:2209.12786
-
Can Transformer Models Effectively Detect Software Aspects in StackOverflow Discussion? 24 Sep 2022 · 0 repositories · arXiv:2209.12065
-
Adaptation of domain-specific transformer models with text oversampling for sentiment analysis of social media posts on Covid-19 vaccines 22 Sep 2022 · 1 repository · arXiv:2209.10966
-
AIR-JPMC@SMM4H'22: Classifying Self-Reported Intimate Partner Violence in Tweets with Multiple BERT-based Models 22 Sep 2022 · 0 repositories · arXiv:2209.10763
-
Detecting Generated Scientific Papers using an Ensemble of Transformer Models 17 Sep 2022 · 1 repository · arXiv:2209.08283
-
DECK: Behavioral Tests to Improve Interpretability and Generalizability of BERT Models Detecting Depression from Text 12 Sep 2022 · 0 repositories · arXiv:2209.05286
-
Probing for Understanding of English Verb Classes and Alternations in Large Pre-trained Language Models 11 Sep 2022 · 0 repositories · arXiv:2209.04811
-
CLaCLab at SocialDisNER: Using Medical Gazetteers for Named-Entity Recognition of Disease Mentions in Spanish Tweets 8 Sep 2022 · 1 repository · arXiv:2209.03528
-
5q032e@SMM4H'22: Transformer-based classification of premise in tweets related to COVID-19 8 Sep 2022 · 0 repositories · arXiv:2209.03851
-
Multilingual Bidirectional Unsupervised Translation Through Multilingual Finetuning and Back-Translation 6 Sep 2022 · 1 repository · arXiv:2209.02821
-
Negation detection in Dutch clinical texts: an evaluation of rule-based and machine learning methods 1 Sep 2022 · 1 repository · arXiv:2209.00470
-
Transformers with Learnable Activation Functions 30 Aug 2022 · 2 repositories · arXiv:2208.14111Syntology official (archive's flag): 1 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
Addressing Token Uniformity in Transformers via Singular Value Transformation 24 Aug 2022 · 1 repository · arXiv:2208.11790Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
SPOT: Knowledge-Enhanced Language Representations for Information Extraction 20 Aug 2022 · 0 repositories · arXiv:2208.09625
-
EmoMent: An Emotion Annotated Mental Health Corpus from two South Asian Countries 17 Aug 2022 · 0 repositories · arXiv:2208.08486
-
Transformer Encoder for Social Science 17 Aug 2022 · 1 repository · arXiv:2208.08005
-
Is Your Model Sensitive? SPeDaC: A New Benchmark for Detecting and Classifying Sensitive Personal Data 12 Aug 2022 · 0 repositories · arXiv:2208.06216
-
A Comparative Study on COVID-19 Fake News Detection Using Different Transformer Based Models 2 Aug 2022 · 0 repositories · arXiv:2208.01355
-
Enhancing Collaborative Filtering Recommender with Prompt-Based Sentiment Analysis 19 Jul 2022 · 1 repository · arXiv:2207.12883
-
Pre-trained language models with domain knowledge for biomedical extractive summarization 19 Jul 2022 · 1 repository
-
Selection Bias Induced Spurious Correlations in Large Language Models 18 Jul 2022 · 1 repository · arXiv:2207.08982
-
Overview of the Shared Task on Fake News Detection in Urdu at FIRE 2021 11 Jul 2022 · 0 repositories · arXiv:2207.05133
-
Computationally Identifying Funneling and Focusing Questions in Classroom Discourse 8 Jul 2022 · 1 repository · arXiv:2208.04715
-
Using contextual sentence analysis models to recognize ESG concepts 4 Jul 2022 · 0 repositories · arXiv:2207.01402
-
ListBERT: Learning to Rank E-commerce products with Listwise BERT 30 Jun 2022 · 0 repositories · arXiv:2206.15198
-
Exploring linguistic feature and model combination for speech recognition based automatic AD detection 28 Jun 2022 · 0 repositories · arXiv:2206.13758
-
Detecting Harmful Online Conversational Content towards LGBTQIA+ Individuals 15 Jun 2022 · 1 repository · arXiv:2207.10032
-
Always Keep your Target in Mind: Studying Semantics and Improving Performance of Neural Lexical Substitution 7 Jun 2022 · 1 repository · arXiv:2206.11815
-
An Empirical Study of IoT Security Aspects at Sentence-Level in Developer Textual Discussions 7 Jun 2022 · 0 repositories · arXiv:2206.03079
-
What do tokens know about their characters and how do they know it? 6 Jun 2022 · 1 repository · arXiv:2206.02608
-
MMTM: Multi-Tasking Multi-Decoder Transformer for Math Word Problems 2 Jun 2022 · 0 repositories · arXiv:2206.01268
-
BERT-Sort: A Zero-shot MLM Semantic Encoder on Ordinal Features for AutoML 1 Jun 2022 · 1 repository
-
The Document Vectors Using Cosine Similarity Revisited 26 May 2022 · 1 repository · arXiv:2205.13357
-
RobustLR: Evaluating Robustness to Logical Perturbation in Deductive Reasoning 25 May 2022 · 1 repository · arXiv:2205.12598Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples)
-
Do we need Label Regularization to Fine-tune Pre-trained Language Models? 25 May 2022 · 0 repositories · arXiv:2205.12428
-
VulBERTa: Simplified Source Code Pre-Training for Vulnerability Detection 25 May 2022 · 1 repository · arXiv:2205.12424
-
Partial-input baselines show that NLI models can ignore context, but they don't 24 May 2022 · 1 repository · arXiv:2205.12181
-
KOLD: Korean Offensive Language Dataset 23 May 2022 · 1 repository · arXiv:2205.11315
-
Learning to Ignore Adversarial Attacks 23 May 2022 · 0 repositories · arXiv:2205.11551
-
Outliers Dimensions that Disrupt Transformers Are Driven by Frequency 23 May 2022 · 1 repository · arXiv:2205.11380Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 8 unverified (of 16 harvested samples)
-
Parameter-Efficient Sparsity for Large Language Models Fine-Tuning 23 May 2022 · 2 repositories · arXiv:2205.11005Syntology official (archive's flag): 2 ran · 2 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; every one of the 2 samples that ran constructed an object rather than computing a result (of 2 harvested samples)
-
The Diminishing Returns of Masked Language Models to Science 23 May 2022 · 0 repositories · arXiv:2205.11342
-
Pre-training Transformer Models with Sentence-Level Objectives for Answer Sentence Selection 20 May 2022 · 0 repositories · arXiv:2205.10455
-
Naturalistic Causal Probing for Morpho-Syntax 14 May 2022 · 1 repository · arXiv:2205.07043
-
EmotionFlow: Capture the Dialogue Level Emotion Transitions 7 May 2022 · 1 repository
-
Improving Downstream Task Performance by Treating Numbers as Entities 7 May 2022 · 0 repositories · arXiv:2205.03559
-
Hyperbolic Relevance Matching for Neural Keyphrase Extraction 4 May 2022 · 1 repository · arXiv:2205.02047
-
Detecting COVID-19 Conspiracy Theories with Transformers and TF-IDF 1 May 2022 · 0 repositories · arXiv:2205.00377
-
RigoBERTa: A State-of-the-Art Language Model For Spanish 27 Apr 2022 · 0 repositories · arXiv:2205.10233
-
PLOD: An Abbreviation Detection Dataset for Scientific Documents 26 Apr 2022 · 1 repository · arXiv:2204.12061
-
Taygete at SemEval-2022 Task 4: RoBERTa based models for detecting Patronising and Condescending Language 22 Apr 2022 · 0 repositories · arXiv:2204.10519
-
Impact of Tokenization on Language Models: An Analysis for Turkish 19 Apr 2022 · 0 repositories · arXiv:2204.08832
-
UTNLP at SemEval-2022 Task 6: A Comparative Analysis of Sarcasm Detection Using Generative-based and Mutation-based Data Augmentation 18 Apr 2022 · 2 repositories · arXiv:2204.08198
-
Analysing similarities between legal court documents using natural language processing approaches based on Transformers 14 Apr 2022 · 0 repositories · arXiv:2204.07182
-
Latent Aspect Detection from Online Unsolicited Customer Reviews 14 Apr 2022 · 1 repository · arXiv:2204.06964
-
Probing for Constituency Structure in Neural Language Models 13 Apr 2022 · 1 repository · arXiv:2204.06201Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
TangoBERT: Reducing Inference Cost by using Cascaded Architecture 13 Apr 2022 · 0 repositories · arXiv:2204.06271
-
Towards Generalizable Semantic Product Search by Text Similarity Pre-training on Search Click Logs 11 Apr 2022 · 0 repositories · arXiv:2204.05231
-
Autoencoding Language Model Based Ensemble Learning for Commonsense Validation and Explanation 7 Apr 2022 · 0 repositories · arXiv:2204.03324
-
BERTuit: Understanding Spanish language in Twitter through a native transformer 7 Apr 2022 · 0 repositories · arXiv:2204.03465
-
CoCoSoDa: Effective Contrastive Learning for Code Search 7 Apr 2022 · 0 repositories · arXiv:2204.03293
-
PALBERT: Teaching ALBERT to Ponder 7 Apr 2022 · 1 repository · arXiv:2204.03276Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; the one sample that ran constructed an object rather than computing a result (of 2 harvested samples)
-
Using Synthetic Data for Conversational Response Generation in Low-resource Settings 6 Apr 2022 · 0 repositories · arXiv:2204.02653
-
Cyberbullying detection across social media platforms via platform-aware adversarial encoding 1 Apr 2022 · 0 repositories · arXiv:2204.00334
-
Effect and Analysis of Large-scale Language Model Rescoring on Competitive ASR Systems 1 Apr 2022 · 0 repositories · arXiv:2204.00212
-
CatIss: An Intelligent Tool for Categorizing Issues Reports using Transformers 31 Mar 2022 · 1 repository · arXiv:2203.17196
-
ANNA: Enhanced Language Representation for Question Answering 28 Mar 2022 · 0 repositories · arXiv:2203.14507
-
UTSA NLP at SemEval-2022 Task 4: An Exploration of Simple Ensembles of Transformers, Convolutional, and Recurrent Neural Networks 28 Mar 2022 · 0 repositories · arXiv:2203.14920
-
Mono vs Multilingual BERT: A Case Study in Hindi and Marathi Named Entity Recognition 24 Mar 2022 · 0 repositories · arXiv:2203.12907
-
A Computational Approach to Understand Mental Health from Reddit: Knowledge-aware Multitask Learning Framework 22 Mar 2022 · 0 repositories · arXiv:2203.11856
-
Perturbations in the Wild: Leveraging Human-Written Text Perturbations for Realistic Adversarial Attack and Defense 19 Mar 2022 · 1 repository · arXiv:2203.10346Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Investigating the Impact of COVID-19 on Education by Social Network Mining 13 Mar 2022 · 0 repositories · arXiv:2203.06584