Methods › Natural Language Processing › Transformers › RoBERTa › Papers, page 4
RoBERTa
Papers archive 2025-07-28
archive papers tagged: 913 · with a code link: 399 · where Syntology ran a sample: 87 (66 with a run with no instrument failure, 21 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (87 of 913 tagged: 66 with a run with no instrument failure, 21 where every run was a failure of Syntology's instrument)
Page 4 of 10: papers 301 to 400 of 913, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Named Entity Inclusion in Abstractive Text Summarization 5 Jul 2023 · 0 repositories · arXiv:2307.02570
-
Improving Language Plasticity via Pretraining with Active Forgetting 3 Jul 2023 · 1 repository · arXiv:2307.01163
-
MAT: Mixed-Strategy Game of Adversarial Training in Fine-tuning 27 Jun 2023 · 0 repositories · arXiv:2306.15826
-
Comparison of Pre-trained Language Models for Turkish Address Parsing 24 Jun 2023 · 0 repositories · arXiv:2306.13947
-
Partitioning-Guided K-Means: Extreme Empty Cluster Resolution for Extreme Model Compression 24 Jun 2023 · 0 repositories · arXiv:2306.14031
-
Resume Information Extraction via Post-OCR Text Processing 23 Jun 2023 · 0 repositories · arXiv:2306.13775
-
Named entity recognition in resumes 22 Jun 2023 · 0 repositories · arXiv:2306.13062
-
GEmo-CLAP: Gender-Attribute-Enhanced Contrastive Language-Audio Pretraining for Accurate Speech Emotion Recognition 13 Jun 2023 · 0 repositories · arXiv:2306.07848
-
Multimodal Audio-textual Architecture for Robust Spoken Language Understanding 12 Jun 2023 · 0 repositories · arXiv:2306.06819
-
EaSyGuide : ESG Issue Identification Framework leveraging Abilities of Generative Large Language Models 11 Jun 2023 · 1 repository · arXiv:2306.06662
-
Enhancing Low Resource NER Using Assisting Language And Transfer Learning 10 Jun 2023 · 0 repositories · arXiv:2306.06477
-
Implementing BERT and fine-tuned RobertA to detect AI generated news by ChatGPT 9 Jun 2023 · 0 repositories · arXiv:2306.07401
-
Prodigy: An Expeditiously Adaptive Parameter-Free Learner 9 Jun 2023 · 1 repository · arXiv:2306.06101
-
Understanding Telecom Language Through Large Language Models 9 Jun 2023 · 0 repositories · arXiv:2306.07933
-
Skill over Scale: The Case for Medium, Domain-Specific Models for SE 5 Jun 2023 · 0 repositories · arXiv:2306.03268
-
MultiLegalPile: A 689GB Multilingual Legal Corpus 3 Jun 2023 · 0 repositories · arXiv:2306.02069
-
Boosting the Performance of Transformer Architectures for Semantic Textual Similarity 1 Jun 2023 · 0 repositories · arXiv:2306.00708
-
Column Type Annotation using ChatGPT 1 Jun 2023 · 1 repository · arXiv:2306.00745
-
Make Pre-trained Model Reversible: From Parameter to Memory Efficient Fine-Tuning 1 Jun 2023 · 1 repository · arXiv:2306.00477Syntology community repositories only · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
XPhoneBERT: A Pre-trained Multilingual Model for Phoneme Representations for Text-to-Speech 31 May 2023 · 2 repositories · arXiv:2305.19709Syntology official (archive's flag): 14 ran · 14 ran (of which 0 constructed an object rather than computing a result; 14 with no instrument failure: 2 honoured, 0 violated, 12 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 18 harvested samples) · 10 pointer-only (licence)
-
PreQuant: A Task-agnostic Quantization Approach for Pre-trained Language Models 30 May 2023 · 0 repositories · arXiv:2306.00014
-
ScoNe: Benchmarking Negation Reasoning in Language Models With Fine-Tuning and In-Context Learning 30 May 2023 · 1 repository · arXiv:2305.19426
-
Abstractive Summarization as Augmentation for Document-Level Event Detection 29 May 2023 · 0 repositories · arXiv:2305.18023
-
From Adversarial Arms Race to Model-centric Evaluation: Motivating a Unified Automatic Robustness Evaluation Framework 29 May 2023 · 1 repository · arXiv:2305.18503Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Zero is Not Hero Yet: Benchmarking Zero-Shot Performance of LLMs for Financial Tasks 26 May 2023 · 1 repository · arXiv:2305.16633
-
Comparative Study of Pre-Trained BERT Models for Code-Mixed Hindi-English Data 25 May 2023 · 0 repositories · arXiv:2305.15722
-
Not wacky vs. definitely wacky: A study of scalar adverbs in pretrained language models 25 May 2023 · 0 repositories · arXiv:2305.16426
-
A Causal View of Entity Bias in (Large) Language Models 24 May 2023 · 1 repository · arXiv:2305.14695Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Complex Mathematical Symbol Definition Structures: A Dataset and Model for Coordination Resolution in Definition Extraction 24 May 2023 · 1 repository · arXiv:2305.14660
-
Context-Aware Transformer Pre-Training for Answer Sentence Selection 24 May 2023 · 0 repositories · arXiv:2305.15358
-
Extracting Psychological Indicators Using Question Answering 24 May 2023 · 0 repositories · arXiv:2305.14891
-
Ghostbuster: Detecting Text Ghostwritten by Large Language Models 24 May 2023 · 2 repositories · arXiv:2305.15047
-
Exploring Large Language Models for Classical Philology 23 May 2023 · 1 repository · arXiv:2305.13698
-
A Deeper (Autoregressive) Approach to Non-Convergent Discourse Parsing 21 May 2023 · 0 repositories · arXiv:2305.12510
-
SEntFiN 1.0: Entity-Aware Sentiment Analysis for Financial News 20 May 2023 · 0 repositories · arXiv:2305.12257
-
Ahead-of-Time P-Tuning 18 May 2023 · 0 repositories · arXiv:2305.10835
-
Knowledge Rumination for Pre-trained Language Models 15 May 2023 · 1 repository · arXiv:2305.08732
-
PESTS: Persian_English Cross Lingual Corpus for Semantic Textual Similarity 13 May 2023 · 0 repositories · arXiv:2305.07893
-
Attack Named Entity Recognition by Entity Boundary Interference 9 May 2023 · 0 repositories · arXiv:2305.05253
-
StrAE: Autoencoding for Pre-Trained Embeddings using Explicit Structure 9 May 2023 · 0 repositories · arXiv:2305.05588
-
GersteinLab at MEDIQA-Chat 2023: Clinical Note Summarization from Doctor-Patient Conversations through Fine-tuning and In-context Learning 8 May 2023 · 0 repositories · arXiv:2305.05001
-
Stanford MLab at SemEval-2023 Task 10: Exploring GloVe- and Transformer-Based Methods for the Explainable Detection of Online Sexism 7 May 2023 · 0 repositories · arXiv:2305.04356
-
On the Usage of Continual Learning for Out-of-Distribution Generalization in Pre-trained Language Models of Code 6 May 2023 · 0 repositories · arXiv:2305.04106
-
CLaC at SemEval-2023 Task 2: Comparing Span-Prediction and Sequence-Labeling approaches for NER 5 May 2023 · 0 repositories · arXiv:2305.03845
-
Using ChatGPT for Entity Matching 5 May 2023 · 1 repository · arXiv:2305.03423
-
NLP-LTU at SemEval-2023 Task 10: The Impact of Data Augmentation and Semi-Supervised Learning Techniques on Text Classification Performance on an Imbalanced Dataset 25 Apr 2023 · 0 repositories · arXiv:2304.12847
-
SocialDial: A Benchmark for Socially-Aware Dialogue Systems 24 Apr 2023 · 1 repository · arXiv:2304.12026
-
Processing Natural Language on Embedded Devices: How Well Do Transformer Models Perform? 23 Apr 2023 · 2 repositories · arXiv:2304.11520
-
Context-Dependent Embedding Utterance Representations for Emotion Recognition in Conversations 17 Apr 2023 · 1 repository · arXiv:2304.08216
-
ArguGPT: evaluating, understanding and identifying argumentative essays generated by GPT models 16 Apr 2023 · 2 repositories · arXiv:2304.07666
-
SimpLex: a lexical text simplification architecture 14 Apr 2023 · 1 repository · arXiv:2304.07002
-
Automated Mapping of CVE Vulnerability Records to MITRE CWE Weaknesses 13 Apr 2023 · 0 repositories · arXiv:2304.11130
-
tmn at SemEval-2023 Task 9: Multilingual Tweet Intimacy Detection using XLM-T, Google Translate, and Ensemble Learning 8 Apr 2023 · 1 repository · arXiv:2304.04054
-
Evaluating the Logical Reasoning Ability of ChatGPT and GPT-4 7 Apr 2023 · 1 repository · arXiv:2304.03439
-
Deep Learning for Opinion Mining and Topic Classification of Course Reviews 6 Apr 2023 · 0 repositories · arXiv:2304.03394
-
MiniRBT: A Two-stage Distilled Small Chinese Pre-trained Model 3 Apr 2023 · 1 repository · arXiv:2304.00717
-
The Other Side of Compression: Measuring Bias in Pruned Transformers 1 Apr 2023 · 1 repository
-
Extracting Thyroid Nodules Characteristics from Ultrasound Reports Using Transformer-based Natural Language Processing Methods 31 Mar 2023 · 0 repositories · arXiv:2304.00115
-
oBERTa: Improving Sparse Transfer Learning via improved initialization, distillation, and pruning regimes 30 Mar 2023 · 0 repositories · arXiv:2303.17612
-
SIGMORPHON 2023 Shared Task of Interlinear Glossing: Baseline Model 24 Mar 2023 · 1 repository · arXiv:2303.14234
-
Generate labeled training data using Prompt Programming and GPT-3. An example of Big Five Personality Classification 22 Mar 2023 · 0 repositories · arXiv:2303.12279
-
PACO: Provocation Involving Action, Culture, and Oppression 19 Mar 2023 · 0 repositories · arXiv:2303.12808
-
NoisyHate: Mining Online Human-Written Perturbations for Realistic Robustness Benchmarking of Content Moderation Models 18 Mar 2023 · 0 repositories · arXiv:2303.10430
-
Do Transformers Parse while Predicting the Masked Word? 14 Mar 2023 · 0 repositories · arXiv:2303.08117
-
Finding the Needle in a Haystack: Unsupervised Rationale Extraction from Long Text Classifiers 14 Mar 2023 · 0 repositories · arXiv:2303.07991
-
Neuro-symbolic Commonsense Social Reasoning 14 Mar 2023 · 3 repositories · arXiv:2303.08264
-
Transformer-based approaches to Sentiment Detection 13 Mar 2023 · 0 repositories · arXiv:2303.07292
-
Will Affective Computing Emerge from Foundation Models and General AI? A First Evaluation on ChatGPT 3 Mar 2023 · 0 repositories · arXiv:2303.03186
-
ToxVis: Enabling Interpretability of Implicit vs. Explicit Toxicity Detection Models with Interactive Visualization 1 Mar 2023 · 0 repositories · arXiv:2303.09402
-
Automatically Classifying Emotions based on Text: A Comparative Exploration of Different Datasets 28 Feb 2023 · 0 repositories · arXiv:2302.14727
-
HULAT at SemEval-2023 Task 10: Data augmentation for pre-trained transformers applied to the detection of sexism in social media 24 Feb 2023 · 1 repository · arXiv:2302.12840
-
Window transformer for dialogue document: a joint framework for causal emotion entailment 24 Feb 2023 · 0 repositories
-
Evaluating the Effectiveness of Pre-trained Language Models in Predicting the Helpfulness of Online Product Reviews 19 Feb 2023 · 1 repository · arXiv:2302.10199
-
DocILE Benchmark for Document Information Localization and Extraction 11 Feb 2023 · 1 repository · arXiv:2302.05658
-
CRL+: A Novel Semi-Supervised Deep Active Contrastive Representation Learning-Based Text Classification Model for Insurance Data 8 Feb 2023 · 0 repositories · arXiv:2302.04343
-
Prompting for Multimodal Hateful Meme Classification 8 Feb 2023 · 0 repositories · arXiv:2302.04156
-
What do Language Models know about word senses? Zero-Shot WSD with Language Models and Domain Inventories 7 Feb 2023 · 0 repositories · arXiv:2302.03353
-
cross-modal fusion techniques for utterance-level emotion recognition from text and speech 5 Feb 2023 · 0 repositories · arXiv:2302.02447
-
Towards Few-Shot Identification of Morality Frames using In-Context Learning 3 Feb 2023 · 0 repositories · arXiv:2302.02029
-
A benchmark for toxic comment classification on Civil Comments dataset 26 Jan 2023 · 1 repository · arXiv:2301.11125
-
A Stability Analysis of Fine-Tuning a Pre-Trained Model 24 Jan 2023 · 0 repositories · arXiv:2301.09820
-
TEDB System Description to a Shared Task on Euphemism Detection 2022 16 Jan 2023 · 1 repository · arXiv:2301.06602
-
Floods Relevancy and Identification of Location from Twitter Posts using NLP Techniques 1 Jan 2023 · 0 repositories · arXiv:2301.00321
-
Finetuning for Sarcasm Detection with a Pruned Dataset 23 Dec 2022 · 1 repository · arXiv:2212.12213
-
Analyzing Semantic Faithfulness of Language Models via Input Intervention on Question Answering 21 Dec 2022 · 1 repository · arXiv:2212.10696
-
Do CoNLL-2003 Named Entity Taggers Still Work Well in 2023? 19 Dec 2022 · 1 repository · arXiv:2212.09747
-
MANTIS at TSAR-2022 Shared Task: Improved Unsupervised Lexical Simplification with Pretrained Encoders 19 Dec 2022 · 0 repositories · arXiv:2212.09855
-
LegalRelectra: Mixed-domain Language Modeling for Long-range Legal Text Comprehension 16 Dec 2022 · 0 repositories · arXiv:2212.08204
-
Visually-augmented pretrained language models for NLP tasks without images 15 Dec 2022 · 1 repository · arXiv:2212.07937
-
Efficient Self-supervised Learning with Contextualized Target Representations for Vision, Speech and Language 14 Dec 2022 · 5 repositories · arXiv:2212.07525Syntology community repositories only · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 8 harvested samples) · 2 pointer-only (licence)
-
Classifying the Ideological Orientation of User-Submitted Texts in Social Media 12 Dec 2022 · 1 repository
-
Learning-To-Embed: Adopting Transformer based models for E-commerce Products Representation Learning 7 Dec 2022 · 0 repositories · arXiv:2212.03725
-
Modern French Poetry Generation with RoBERTa and GPT-2 6 Dec 2022 · 0 repositories · arXiv:2212.02911
-
ColD Fusion: Collaborative Descent for Distributed Multitask Finetuning 2 Dec 2022 · 0 repositories · arXiv:2212.01378
-
Revisiting Distance Metric Learning for Few-Shot Natural Language Classification 28 Nov 2022 · 0 repositories · arXiv:2211.15202
-
Finetuning BERT on Partially Annotated NER Corpora 25 Nov 2022 · 1 repository · arXiv:2211.14360
-
AF Adapter: Continual Pretraining for Building Chinese Biomedical Language Model 21 Nov 2022 · 1 repository · arXiv:2211.11363
-
Entity-Assisted Language Models for Identifying Check-worthy Sentences 19 Nov 2022 · 0 repositories · arXiv:2211.10678
-
Where did you tweet from? Inferring the origin locations of tweets based on contextual information 18 Nov 2022 · 0 repositories · arXiv:2211.16506
-
LongFNT: Long-form Speech Recognition with Factorized Neural Transducer 17 Nov 2022 · 0 repositories · arXiv:2211.09412