Browse State-of-the-Art › XLM-R › Papers, page 2
XLM-R
Papers archive 2025-07-28
archive papers tagged: 221 · with a code link: 99 · where Syntology ran a sample: 20 (17 with a run with no instrument failure, 3 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (20 of 221 tagged: 17 with a run with no instrument failure, 3 where every run was a failure of Syntology's instrument)
Page 2 of 3: papers 101 to 200 of 221, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Subasa -- Adapting Language Models for Low-resourced Offensive Language Detection in Sinhala2 Apr 2025 0 repositories listed
-
AmaSQuAD: A Benchmark for Amharic Extractive Question Answering4 Feb 2025 0 repositories listed
-
Evaluating the Effectiveness of XAI Techniques for Encoder-Based Language Models26 Jan 2025 0 repositories listed
-
Comparative Approaches to Sentiment Analysis Using Datasets in Major European and Arabic Languages21 Jan 2025 0 repositories listed
-
FuocChuVIP123 at CoMeDi Shared Task: Disagreement Ranking with XLM-Roberta Sentence Embeddings and Deep Neural Regression21 Jan 2025 0 repositories listed
-
Multi-stage Training of Bilingual Islamic LLM for Neural Passage Retrieval17 Jan 2025 0 repositories listed
-
BabyLMs for isiXhosa: Data-Efficient Language Modelling in a Low-Resource Context7 Jan 2025 0 repositories listed
-
USTCCTSU at SemEval-2024 Task 1: Reducing Anisotropy for Cross-lingual Semantic Textual Relatedness Task28 Nov 2024 0 repositories listed
-
Retrofitting Large Language Models with Dynamic Tokenization27 Nov 2024 0 repositories listed
-
Transformer-Based Contextualized Language Models Joint with Neural Networks for Natural Language Inference in Vietnamese20 Nov 2024 0 repositories listed
-
29 Jul 2024 0 repositories listed Syntology 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
The Model Arena for Cross-lingual Sentiment Analysis: A Comparative Study in the Era of Large Language Models27 Jun 2024 0 repositories listed
-
Multilingual Large Language Models and Curse of Multilinguality15 Jun 2024 0 repositories listed
-
Targeted Multilingual Adaptation for Low-resource Language Families20 May 2024 0 repositories listed
-
Software Mention Recognition with a Three-Stage Framework Based on BERTology Models at SOMD 202423 Apr 2024 0 repositories listed
-
Adapting Mental Health Prediction Tasks for Cross-lingual Learning via Meta-Training and In-context Learning with Large Language Model13 Apr 2024 0 repositories listed
-
MaiNLP at SemEval-2024 Task 1: Analyzing Source Language Selection in Cross-Lingual Textual Relatedness3 Apr 2024 0 repositories listed
-
Solution for Emotion Prediction Competition of Workshop on Emotionally and Culturally Intelligent AI26 Mar 2024 0 repositories listed
-
Machines Do See Color: A Guideline to Classify Different Forms of Racist Discourse in Large Corpora17 Jan 2024 0 repositories listed
-
LinguAlchemy: Fusing Typological and Geographical Elements for Unseen Language Generalization11 Jan 2024 0 repositories listed
-
Hate Speech and Offensive Content Detection in Indo-Aryan Languages: A Battle of LSTM and Transformers9 Dec 2023 0 repositories listed
-
A Text-to-Text Model for Multilingual Offensive Language Identification6 Dec 2023 0 repositories listed
-
Zero-Shot Cross-Lingual Sentiment Classification under Distribution Shift: an Exploratory Study11 Nov 2023 0 repositories listed
-
MedAI Dialog Corpus (MEDIC): Zero-Shot Classification of Doctor and AI Responses in Health Consultations19 Oct 2023 0 repositories listed
-
9 Oct 2023 0 repositories listed
-
Mixed-Distil-BERT: Code-mixed Language Modeling for Bangla, English, and Hindi19 Sep 2023 0 repositories listed
-
Self-Distilled Quantization: Achieving High Compression Rates in Transformer-Based Language Models12 Jul 2023 0 repositories listed
-
Does mBERT understand Romansh? Evaluating word embeddings using word alignment14 Jun 2023 0 repositories listed
-
Extrapolating Multilingual Understanding Models as Multilingual Generators22 May 2023 0 repositories listed
-
DN at SemEval-2023 Task 12: Low-Resource Language Text Classification via Multilingual Pretrained Language Model Fine-tuning4 May 2023 0 repositories listed
-
USTC-NELSLIP at SemEval-2023 Task 2: Statistical Construction and Dual Adaptation of Gazetteer for Multilingual Complex NER4 May 2023 0 repositories listed
-
Transfer to a Low-Resource Language via Close Relatives: The Case Study on Faroese18 Apr 2023 0 repositories listed
-
Tollywood Emotions: Annotation of Valence-Arousal in Telugu Song Lyrics16 Mar 2023 0 repositories listed
-
Integrating Semantic Information into Sketchy Reading Module of Retro-Reader for Vietnamese Machine Reading Comprehension1 Jan 2023 0 repositories listed
-
VTCC-NLP at NL4Opt competition subtask 1: An Ensemble Pre-trained language models for Named Entity Recognition14 Dec 2022 0 repositories listed
-
Languages You Know Influence Those You Learn: Impact of Language Characteristics on Multi-Lingual Text-to-Text Transfer4 Dec 2022 0 repositories listed
-
Compressing Cross-Lingual Multi-Task Models at Qualtrics29 Nov 2022 0 repositories listed
-
L3Cube-HindBERT and DevBERT: Pre-Trained BERT Transformer models for Devanagari based Hindi and Marathi Languages21 Nov 2022 0 repositories listed
-
Legal-Tech Open Diaries: Lesson learned on how to develop and deploy light-weight models in the era of humongous Language Models24 Oct 2022 0 repositories listed
-
Extending Word-Level Quality Estimation for Post-Editing Assistance23 Sep 2022 0 repositories listed
-
SMTCE: A Social Media Text Classification Evaluation Benchmark and BERTology Models for Vietnamese21 Sep 2022 0 repositories listed
-
ALEXSIS-PT: A New Resource for Portuguese Lexical Simplification19 Sep 2022 0 repositories listed
-
Investigating Language Relationships in Multilingual Sentence Encoders Through the Lens of Linguistic Typology1 Sep 2022 0 repositories listed
-
BERTifying Sinhala -- A Comprehensive Analysis of Pre-trained Language Models for Sinhala Text Classification16 Aug 2022 0 repositories listed
-
Massively Multilingual Lexical Specialization of Multilingual Transformers1 Aug 2022 0 repositories listed
-
AsNER -- Annotated Dataset and Baseline for Assamese Named Entity recognition7 Jul 2022 0 repositories listed
-
Hitachi at SemEval-2022 Task 2: On the Effectiveness of Span-based Classification Approaches for Multilingual Idiomaticity Detection1 Jul 2022 0 repositories listed
-
Sliced at SemEval-2022 Task 11: Bigger, Better? Massively Multilingual LMs for Multilingual Complex NER on an Academic GPU Budget1 Jul 2022 0 repositories listed
-
Alexa Teacher Model: Pretraining and Distilling Multi-Billion-Parameter Encoders for Natural Language Understanding Systems15 Jun 2022 0 repositories listed
-
AsNER - Annotated Dataset and Baseline for Assamese Named Entity recognition1 Jun 2022 0 repositories listed
-
Debating Europe: A Multilingual Multi-Target Stance Classification Dataset of Online Debates1 Jun 2022 0 repositories listed
-
Analyzing the Mono- and Cross-Lingual Pretraining Dynamics of Multilingual Language Models24 May 2022 0 repositories listed
-
black[LSCDiscovery shared task] GlossReader at LSCDiscovery: Train to Select a Proper Gloss in English – Discover Lexical Semantic Change in Spanish1 May 2022 0 repositories listed
-
CUET-NLP@TamilNLP-ACL2022: Multi-Class Textual Emotion Detection from Social Media using Transformer1 May 2022 0 repositories listed
-
TeluguNER: Leveraging Multi-Domain Named Entity Recognition with Deep Transformers1 May 2022 0 repositories listed
-
Do Not Fire the Linguist: Grammatical Profiles Help Language Models Detect Semantic Change12 Apr 2022 0 repositories listed
-
Do Multilingual Language Models Capture Differing Moral Norms?18 Mar 2022 0 repositories listed
-
Cross-Lingual Ability of Multilingual Masked Language Models: A Study of Language Structure16 Mar 2022 0 repositories listed
-
Are Pretrained Multilingual Models Equally Fair Across Languages?16 Jan 2022 0 repositories listed
-
Frustratingly Simple Regularization to Improve Zero-shot Cross-lingual Robustness16 Jan 2022 0 repositories listed
-
Lexicon based Fine-tuning of Multilingual Language Models for Sentiment Analysis of Low-resource Languages16 Jan 2022 0 repositories listed
-
The Impact of Cross-Lingual Adjustment of Contextual Word Representations on Zero-Shot Transfer16 Jan 2022 0 repositories listed
-
ZeroBERTo: Leveraging Zero-Shot Text Classification by Topic Modeling4 Jan 2022 0 repositories listed
-
Multilingual Pre-training with Universal Dependency Learning1 Dec 2021 0 repositories listed
-
Combining static and contextualised multilingual embeddings16 Nov 2021 0 repositories listed
-
Gradient Sparsification For Masked Fine-Tuning of Transformers16 Nov 2021 0 repositories listed
-
Multi-stage Distillation Framework for Cross-Lingual Semantic Similarity Matching16 Nov 2021 0 repositories listed
-
Prix-LM: Pretraining for Multilingual Knowledge Base Construction16 Nov 2021 0 repositories listed
-
NICT Kyoto Submission for the WMT’21 Quality Estimation Task: Multimetric Multilingual Pretraining for Critical Error Detection1 Nov 2021 0 repositories listed
-
Saliency-based Multi-View Mixed Language Training for Zero-shot Cross-lingual Classification1 Nov 2021 0 repositories listed
-
TEET! Tunisian Dataset for Toxic Speech Detection11 Oct 2021 0 repositories listed
-
Boosting Transformers for Job Expression Extraction and Classification in a Low-Resource Setting17 Sep 2021 0 repositories listed
-
On the Universality of Deep Contextual Language Models15 Sep 2021 0 repositories listed
-
FBERT: A Neural Transformer for Identifying Offensive Content10 Sep 2021 0 repositories listed
-
Classification of Code-Mixed Text Using Capsule Networks1 Sep 2021 0 repositories listed
-
Siamese Networks for Inference in Malayalam Language Texts1 Sep 2021 0 repositories listed
-
Contributions of Transformer Attention Heads in Multi- and Cross-lingual Tasks18 Aug 2021 0 repositories listed
-
Applying Occam’s Razor to Transformer-Based Dependency Parsing: What Works, What Doesn’t, and What is Really Necessary1 Aug 2021 0 repositories listed
-
ARBERT & MARBERT: Deep Bidirectional Transformers for Arabic1 Aug 2021 0 repositories listed
-
GlossReader at SemEval-2021 Task 2: Reading Definitions Improves Contextualized Word Embeddings1 Aug 2021 0 repositories listed
-
IBM MNLP IE at CASE 2021 Task 1: Multigranular and Multilingual Event Detection on Protest News1 Aug 2021 0 repositories listed
-
利用语义关联增强的跨语言预训练模型的译文质量评估(A Cross-language Pre-trained Model with Enhanced Semantic Connection for MT Quality Estimation)1 Aug 2021 0 repositories listed
-
LIORI at SemEval-2021 Task 2: Span Prediction and Binary Classification approaches to Word-in-Context Disambiguation1 Aug 2021 0 repositories listed
-
RobertNLP at the IWPT 2021 Shared Task: Simple Enhanced UD Parsing for 17 Languages1 Aug 2021 0 repositories listed
-
SkoltechNLP at SemEval-2021 Task 2: Generating Cross-Lingual Training Data for the Word-in-Context Task1 Aug 2021 0 repositories listed
-
Team “DaDeFrNi” at CASE 2021 Task 1: Document and Sentence Classification for Protest Event Detection1 Aug 2021 0 repositories listed
-
Emotion Stimulus Detection in German News Headlines27 Jul 2021 0 repositories listed
-
TGIF: Tree-Graph Integrated-Format Parser for Enhanced UD with Two-Stage Generic- to Individual-Language Finetuning14 Jul 2021 0 repositories listed
-
A Primer on Pretrained Multilingual Language Models1 Jul 2021 0 repositories listed
-
Automatic Sexism Detection with Multilingual Transformer Models9 Jun 2021 0 repositories listed
-
How to Adapt Your Pretrained Multilingual Model to 1600 Languages3 Jun 2021 0 repositories listed
-
Diagnosing Transformers in Task-Oriented Semantic Parsing27 May 2021 0 repositories listed
-
Larger-Scale Transformers for Multilingual Masked Language Modeling2 May 2021 0 repositories listed
-
Multilingual and Zero-Shot is Closing in on Monolingual Web Register Classification1 May 2021 0 repositories listed
-
MCL@IITK at SemEval-2021 Task 2: Multilingual and Cross-lingual Word-in-Context Disambiguation using Augmented Data, Signals, and Transformers4 Apr 2021 0 repositories listed
-
Priberam Labs at the 3rd Shared Task on SlavNER1 Apr 2021 0 repositories listed
-
LightMBERT: A Simple Yet Effective Method for Multilingual BERT Distillation11 Mar 2021 0 repositories listed
-
Automatic Difficulty Classification of Arabic Sentences7 Mar 2021 0 repositories listed
-
Vyākarana: A Colorless Green Benchmark for Syntactic Evaluation in Indic Languages1 Mar 2021 0 repositories listed
Syntology lines on 2 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.