Methods › Natural Language Processing › Transformers › RoBERTa › Papers, page 2
RoBERTa
Papers archive 2025-07-28
archive papers tagged: 913 · with a code link: 399 · where Syntology ran a sample: 87 (66 with a run with no instrument failure, 21 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (87 of 913 tagged: 66 with a run with no instrument failure, 21 where every run was a failure of Syntology's instrument)
Page 2 of 10: papers 101 to 200 of 913, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Comparing Unidirectional, Bidirectional, and Word2vec Models for Discovering Vulnerabilities in Compiled Lifted Code 26 Sep 2024 · 0 repositories · arXiv:2409.17513
-
Improving Academic Skills Assessment with NLP and Ensemble Learning 23 Sep 2024 · 0 repositories · arXiv:2409.19013
-
Drift to Remember 21 Sep 2024 · 0 repositories · arXiv:2409.13997
-
HUT: A More Computation Efficient Fine-Tuning Method With Hadamard Updated Transformation 20 Sep 2024 · 0 repositories · arXiv:2409.13501
-
Towards understanding evolution of science through language model series 15 Sep 2024 · 1 repository · arXiv:2409.09636
-
Enhanced Online Grooming Detection Employing Context Determination and Message-Level Analysis 12 Sep 2024 · 0 repositories · arXiv:2409.07958
-
Constrained Multi-Layer Contrastive Learning for Implicit Discourse Relationship Recognition 7 Sep 2024 · 0 repositories · arXiv:2409.13716
-
CBF-LLM: Safe Control for LLM Alignment 28 Aug 2024 · 1 repository · arXiv:2408.15625
-
The Self-Contained Negation Test Set 21 Aug 2024 · 0 repositories · arXiv:2408.11469
-
A Strategy to Combine 1stGen Transformers and Open LLMs for Automatic Text Classification 19 Aug 2024 · 0 repositories · arXiv:2408.09629
-
Active Learning for Identifying Disaster-Related Tweets: A Comparison with Keyword Filtering and Generic Fine-Tuning 19 Aug 2024 · 0 repositories · arXiv:2408.09914
-
The Language of Trauma: Modeling Traumatic Event Descriptions Across Domains with Explainable AI 12 Aug 2024 · 0 repositories · arXiv:2408.05977
-
Is Child-Directed Speech Effective Training Data for Language Models? 7 Aug 2024 · 1 repository · arXiv:2408.03617Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 1 violated, 4 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples) · 2 pointer-only (licence)
-
Defining and Evaluating Decision and Composite Risk in Language Models Applied to Natural Language Inference 4 Aug 2024 · 0 repositories · arXiv:2408.01935
-
MistralBSM: Leveraging Mistral-7B for Vehicular Networks Misbehavior Detection 26 Jul 2024 · 0 repositories · arXiv:2407.18462
-
Banyan: Improved Representation Learning with Explicit Structure 25 Jul 2024 · 0 repositories · arXiv:2407.17771
-
RoBERTa, ResNeXt and BiLSTM with self-attention: The ultimate trio for customer sentiment analysis 25 Jul 2024 · 0 repositories
-
Artificial Intelligence in Extracting Diagnostic Data from Dental Records 23 Jul 2024 · 0 repositories · arXiv:2407.21050
-
Evaluating Large Language Models for Anxiety and Depression Classification using Counseling and Psychotherapy Transcripts 18 Jul 2024 · 1 repository · arXiv:2407.13228
-
Qalam : A Multimodal LLM for Arabic Optical Character and Handwriting Recognition 18 Jul 2024 · 0 repositories · arXiv:2407.13559
-
Sharif-STR at SemEval-2024 Task 1: Transformer as a Regression Model for Fine-Grained Scoring of Textual Semantic Relations 17 Jul 2024 · 1 repository · arXiv:2407.12426
-
R-SFLLM: Jamming Resilient Framework for Split Federated Learning with Large Language Models 16 Jul 2024 · 0 repositories · arXiv:2407.11654
-
Resource Management for Low-latency Cooperative Fine-tuning of Foundation Models at the Network Edge 13 Jul 2024 · 0 repositories · arXiv:2407.09873
-
Robustness of LLMs to Perturbations in Text 12 Jul 2024 · 0 repositories · arXiv:2407.08989
-
ROSA: Random Subspace Adaptation for Efficient Fine-Tuning 10 Jul 2024 · 1 repository · arXiv:2407.07802
-
An Empirical Comparison of Vocabulary Expansion and Initialization Approaches for Language Models 8 Jul 2024 · 1 repository · arXiv:2407.05841Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Using LLMs to label medical papers according to the CIViC evidence model 5 Jul 2024 · 1 repository · arXiv:2407.04466
-
HYBRINFOX at CheckThat! 2024 -- Task 1: Enhancing Language Models with Structured Information for Check-Worthiness Estimation 4 Jul 2024 · 0 repositories · arXiv:2407.03850
-
HYBRINFOX at CheckThat! 2024 -- Task 2: Enriching BERT Models with the Expert System VAGO for Subjectivity Detection 4 Jul 2024 · 0 repositories · arXiv:2407.03770
-
LLM-Generated Natural Language Meets Scaling Laws: New Explorations and Data Augmentation Methods 29 Jun 2024 · 0 repositories · arXiv:2407.00322
-
RAGBench: Explainable Benchmark for Retrieval-Augmented Generation Systems 25 Jun 2024 · 0 repositories · arXiv:2407.11005
-
This Paper Had the Smartest Reviewers -- Flattery Detection Utilising an Audio-Textual Transformer-Based Approach 25 Jun 2024 · 1 repository · arXiv:2406.17667
-
Unlocking Continual Learning Abilities in Language Models 25 Jun 2024 · 1 repository · arXiv:2406.17245Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples)
-
MixTex: Unambiguous Recognition Should Not Rely Solely on Real Data 24 Jun 2024 · 1 repository · arXiv:2406.17148
-
WellDunn: On the Robustness and Explainability of Language Models and Large Language Models in Identifying Wellness Dimensions 17 Jun 2024 · 1 repository · arXiv:2406.12058Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
ShareLoRA: Parameter Efficient and Robust Large Language Model Fine-tuning via Shared Low-Rank Adaptation 16 Jun 2024 · 1 repository · arXiv:2406.10785Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 5 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
The Devil is in the Neurons: Interpreting and Mitigating Social Biases in Pre-trained Language Models 14 Jun 2024 · 1 repository · arXiv:2406.10130
-
BPE-knockout: Pruning Pre-existing BPE Tokenisers with Backwards-compatible Morphological Semi-supervision 13 Jun 2024 · 1 repository
-
Label-aware Hard Negative Sampling Strategies with Momentum Contrastive Learning for Implicit Hate Speech Detection 12 Jun 2024 · 1 repository · arXiv:2406.07886
-
Advancing Semantic Textual Similarity Modeling: A Regression Framework with Translated ReLU and Smooth K2 Loss 8 Jun 2024 · 2 repositories · arXiv:2406.05326
-
BAMO at SemEval-2024 Task 9: BRAINTEASER: A Novel Task Defying Common Sense 7 Jun 2024 · 1 repository · arXiv:2406.04947
-
RICo: Reddit ideological communities 5 Jun 2024 · 1 repository
-
Probing the Category of Verbal Aspect in Transformer Language Models 4 Jun 2024 · 0 repositories · arXiv:2406.02335
-
Annotation Guidelines-Based Knowledge Augmentation: Towards Enhancing Large Language Models for Educational Text Classification 3 Jun 2024 · 0 repositories · arXiv:2406.00954
-
FactGenius: Combining Zero-Shot Prompting and Fuzzy Relation Mining to Improve Fact Verification with Knowledge Graphs 3 Jun 2024 · 1 repository · arXiv:2406.01311
-
RoBERTa-BiLSTM: A Context-Aware Hybrid Model for Sentiment Analysis 1 Jun 2024 · 1 repository · arXiv:2406.00367
-
Bi-Directional Transformers vs. word2vec: Discovering Vulnerabilities in Lifted Compiled Code 31 May 2024 · 0 repositories · arXiv:2405.20611
-
Ensemble Model With Bert,Roberta and Xlnet For Molecular property prediction 30 May 2024 · 0 repositories · arXiv:2406.06553
-
Performance evaluation of Reddit Comments using Machine Learning and Natural Language Processing methods in Sentiment Analysis 27 May 2024 · 0 repositories · arXiv:2405.16810
-
Incremental Comprehension of Garden-Path Sentences by Large Language Models: Semantic Interpretation, Syntactic Re-Analysis, and Attention 25 May 2024 · 0 repositories · arXiv:2405.16042
-
Comet: A Communication-efficient and Performant Approximation for Private Transformer Inference 24 May 2024 · 0 repositories · arXiv:2405.17485
-
CReMa: Crisis Response through Computational Identification and Matching of Cross-Lingual Requests and Offers Shared on Social Media 20 May 2024 · 0 repositories · arXiv:2405.11897
-
ExplainableDetector: Exploring Transformer-based Language Modeling Approach for SMS Spam Detection with Explainability Analysis 12 May 2024 · 0 repositories · arXiv:2405.08026
-
TacoERE: Cluster-aware Compression for Event Relation Extraction 11 May 2024 · 0 repositories · arXiv:2405.06890
-
Reddit-Impacts: A Named Entity Recognition Dataset for Analyzing Clinical and Social Effects of Substance Use Derived from Social Media 9 May 2024 · 0 repositories · arXiv:2405.06145
-
Detecting Anti-Semitic Hate Speech using Transformer-based Large Language Models 6 May 2024 · 0 repositories · arXiv:2405.03794
-
Structural Pruning of Pre-trained Language Models via Neural Architecture Search 3 May 2024 · 1 repository · arXiv:2405.02267
-
Early Transformers: A study on Efficient Training of Transformer Models through Early-Bird Lottery Tickets 2 May 2024 · 0 repositories · arXiv:2405.02353
-
Investigating Wit, Creativity, and Detectability of Large Language Models in Domain-Specific Writing Style Adaptation of Reddit's Showerthoughts 2 May 2024 · 1 repository · arXiv:2405.01660
-
A Named Entity Recognition and Topic Modeling-based Solution for Locating and Better Assessment of Natural Disasters in Social Media 1 May 2024 · 0 repositories · arXiv:2405.00903
-
FeDeRA:Efficient Fine-tuning of Language Models in Federated Learning Leveraging Weight Decomposition 29 Apr 2024 · 0 repositories · arXiv:2404.18848
-
Can Perplexity Predict Fine-Tuning Performance? An Investigation of Tokenization Effects on Sequential Language Models for Nepali 28 Apr 2024 · 0 repositories · arXiv:2404.18071
-
Pre-Calc: Learning to Use the Calculator Improves Numeracy in Language Models 22 Apr 2024 · 1 repository · arXiv:2404.14355Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Marking: Visual Grading with Highlighting Errors and Annotating Missing Bits 22 Apr 2024 · 0 repositories · arXiv:2404.14301
-
Do "English" Named Entity Recognizers Work Well on Global Englishes? 20 Apr 2024 · 2 repositories · arXiv:2404.13465
-
Evaluating Subword Tokenization: Alien Subword Composition and OOV Generalization Challenge 20 Apr 2024 · 1 repository · arXiv:2404.13292
-
Enabling Natural Zero-Shot Prompting on Encoder Models via Statement-Tuning 19 Apr 2024 · 0 repositories · arXiv:2404.12897
-
emrQA-msquad: A Medical Dataset Structured with the SQuAD V2.0 Framework, Enriched with emrQA Medical Information 18 Apr 2024 · 0 repositories · arXiv:2404.12050
-
Demystifying Legalese: An Automated Approach for Summarizing and Analyzing Overlaps in Privacy Policies and Terms of Service 17 Apr 2024 · 0 repositories · arXiv:2404.13087
-
Relational Graph Convolutional Networks for Sentiment Analysis 16 Apr 2024 · 0 repositories · arXiv:2404.13079
-
Comprehensive Study on German Language Models for Clinical and Biomedical Text Understanding 8 Apr 2024 · 0 repositories · arXiv:2404.05694
-
Investigating the Robustness of Modelling Decisions for Few-Shot Cross-Topic Stance Detection: A Preregistered Study 5 Apr 2024 · 1 repository · arXiv:2404.03987
-
BCAmirs at SemEval-2024 Task 4: Beyond Words: A Multimodal and Multilingual Exploration of Persuasion in Memes 3 Apr 2024 · 1 repository · arXiv:2404.03022
-
Explainable Deep Learning: A Visual Analytics Approach with Transition Matrices 29 Mar 2024 · 1 repository
-
AlloyBERT: Alloy Property Prediction with Large Language Models 28 Mar 2024 · 0 repositories · arXiv:2403.19783
-
AcTED: Automatic Acquisition of Typical Event Duration for Semi-supervised Temporal Commonsense QA 27 Mar 2024 · 0 repositories · arXiv:2403.18504
-
Evaluating Large Language Models for Health-Related Text Classification Tasks with Public Social Media Data 27 Mar 2024 · 0 repositories · arXiv:2403.19031
-
SemRoDe: Macro Adversarial Training to Learn Representations That are Robust to Word-Level Attacks 27 Mar 2024 · 1 repository · arXiv:2403.18423Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples)
-
Fingerprinting web servers through Transformer-encoded HTTP response headers 26 Mar 2024 · 1 repository · arXiv:2404.00056
-
LlamBERT: Large-scale low-cost data annotation in NLP 23 Mar 2024 · 1 repository · arXiv:2403.15938
-
Ax-to-Grind Urdu: Benchmark Dataset for Urdu Fake News Detection 20 Mar 2024 · 1 repository · arXiv:2403.14037
-
CICLe: Conformal In-Context Learning for Largescale Multi-Class Food Risk Classification 18 Mar 2024 · 1 repository · arXiv:2403.11904
-
LookupFFN: Making Transformers Compute-lite for CPU inference 12 Mar 2024 · 1 repository · arXiv:2403.07221Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
GaLore: Memory-Efficient LLM Training by Gradient Low-Rank Projection 6 Mar 2024 · 3 repositories · arXiv:2403.03507
-
EEE-QA: Exploring Effective and Efficient Question-Answer Representations 4 Mar 2024 · 1 repository · arXiv:2403.02176
-
Analysis of Privacy Leakage in Federated Large Language Models 2 Mar 2024 · 1 repository · arXiv:2403.04784Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Leveraging pre-trained language models for code generation 29 Feb 2024 · 1 repository
-
PeLLE: Encoder-based language models for Brazilian Portuguese based on open data 29 Feb 2024 · 0 repositories · arXiv:2402.19204
-
MAGPIE: Multi-Task Media-Bias Analysis Generalization for Pre-Trained Identification of Expressions 27 Feb 2024 · 1 repository · arXiv:2403.07910
-
Asymmetry in Low-Rank Adapters of Foundation Models 26 Feb 2024 · 1 repository · arXiv:2402.16842Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
SemEval-2024 Task 8: Weighted Layer Averaging RoBERTa for Black-Box Machine-Generated Text Detection 24 Feb 2024 · 1 repository · arXiv:2402.15873
-
Advancing Parameter Efficiency in Fine-tuning via Representation Editing 23 Feb 2024 · 2 repositories · arXiv:2402.15179Syntology official (archive's flag): 1 ran · 2 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; every one of the 2 samples that ran constructed an object rather than computing a result (of 2 harvested samples) · 1 pointer-only (licence)
-
Towards Efficient Active Learning in NLP via Pretrained Representations 23 Feb 2024 · 0 repositories · arXiv:2402.15613
-
Can GNN be Good Adapter for LLMs? 20 Feb 2024 · 2 repositories · arXiv:2402.12984Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
Acquiring Clean Language Models from Backdoor Poisoned Datasets by Downscaling Frequency Space 19 Feb 2024 · 1 repository · arXiv:2402.12026Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
A Curious Case of Searching for the Correlation between Training Data and Adversarial Robustness of Transformer Textual Models 18 Feb 2024 · 1 repository · arXiv:2402.11469
-
Detecting a Proxy for Potential Comorbid ADHD in People Reporting Anxiety Symptoms from Social Media Data 17 Feb 2024 · 0 repositories · arXiv:2403.05561
-
Efficient Models for the Detection of Hate, Abuse and Profanity 8 Feb 2024 · 0 repositories · arXiv:2402.05624
-
Neural Models for Source Code Synthesis and Completion 8 Feb 2024 · 0 repositories · arXiv:2402.06690
-
The Use of a Large Language Model for Cyberbullying Detection 6 Feb 2024 · 0 repositories · arXiv:2402.04088