Methods › Natural Language Processing › Language Models › XLM-R › Papers, page 2
XLM-R
Papers archive 2025-07-28
archive papers tagged: 176 · with a code link: 80 · where Syntology ran a sample: 14 (12 with a run with no instrument failure, 2 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (14 of 176 tagged: 12 with a run with no instrument failure, 2 where every run was a failure of Syntology's instrument)
Page 2 of 2: papers 101 to 176 of 176, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Are Pretrained Multilingual Models Equally Fair Across Languages? 16 Jan 2022 · 0 repositories
-
Frustratingly Simple Regularization to Improve Zero-shot Cross-lingual Robustness 16 Jan 2022 · 0 repositories
-
Lexicon based Fine-tuning of Multilingual Language Models for Sentiment Analysis of Low-resource Languages 16 Jan 2022 · 0 repositories
-
The Impact of Cross-Lingual Adjustment of Contextual Word Representations on Zero-Shot Transfer 16 Jan 2022 · 0 repositories
-
ZeroBERTo: Leveraging Zero-Shot Text Classification by Topic Modeling 4 Jan 2022 · 0 repositories · arXiv:2201.01337
-
Multilingual Pre-training with Universal Dependency Learning 1 Dec 2021 · 0 repositories
-
DeBERTaV3: Improving DeBERTa using ELECTRA-Style Pre-Training with Gradient-Disentangled Embedding Sharing 18 Nov 2021 · 3 repositories · arXiv:2111.09543Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples)
-
Combining static and contextualised multilingual embeddings 16 Nov 2021 · 0 repositories
-
Multi-stage Distillation Framework for Cross-Lingual Semantic Similarity Matching 16 Nov 2021 · 0 repositories
-
Prix-LM: Pretraining for Multilingual Knowledge Base Construction 16 Nov 2021 · 0 repositories
-
NICT Kyoto Submission for the WMT’21 Quality Estimation Task: Multimetric Multilingual Pretraining for Critical Error Detection 1 Nov 2021 · 0 repositories
-
Saliency-based Multi-View Mixed Language Training for Zero-shot Cross-lingual Classification 1 Nov 2021 · 0 repositories
-
IndoNLI: A Natural Language Inference Dataset for Indonesian 27 Oct 2021 · 1 repository · arXiv:2110.14566
-
Prix-LM: Pretraining for Multilingual Knowledge Base Construction 16 Oct 2021 · 1 repository · arXiv:2110.08443
-
Towards Making the Most of Multilingual Pretraining for Zero-Shot Neural Machine Translation 16 Oct 2021 · 1 repository · arXiv:2110.08547Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
TEET! Tunisian Dataset for Toxic Speech Detection 11 Oct 2021 · 0 repositories · arXiv:2110.05287
-
Boosting Transformers for Job Expression Extraction and Classification in a Low-Resource Setting 17 Sep 2021 · 0 repositories · arXiv:2109.08597
-
On the Universality of Deep Contextual Language Models 15 Sep 2021 · 0 repositories · arXiv:2109.07140
-
FBERT: A Neural Transformer for Identifying Offensive Content 10 Sep 2021 · 0 repositories · arXiv:2109.05074
-
Classification of Code-Mixed Text Using Capsule Networks 1 Sep 2021 · 0 repositories
-
Siamese Networks for Inference in Malayalam Language Texts 1 Sep 2021 · 0 repositories
-
Cross-Lingual Text Classification of Transliterated Hindi and Malayalam 31 Aug 2021 · 1 repository · arXiv:2108.13620
-
Contributions of Transformer Attention Heads in Multi- and Cross-lingual Tasks 18 Aug 2021 · 0 repositories · arXiv:2108.08375
-
Applying Occam’s Razor to Transformer-Based Dependency Parsing: What Works, What Doesn’t, and What is Really Necessary 1 Aug 2021 · 0 repositories
-
ARBERT & MARBERT: Deep Bidirectional Transformers for Arabic 1 Aug 2021 · 0 repositories
-
COSY: COunterfactual SYntax for Cross-Lingual Understanding 1 Aug 2021 · 1 repository
-
GlossReader at SemEval-2021 Task 2: Reading Definitions Improves Contextualized Word Embeddings 1 Aug 2021 · 0 repositories
-
LIORI at SemEval-2021 Task 2: Span Prediction and Binary Classification approaches to Word-in-Context Disambiguation 1 Aug 2021 · 0 repositories
-
RobertNLP at the IWPT 2021 Shared Task: Simple Enhanced UD Parsing for 17 Languages 1 Aug 2021 · 0 repositories
-
SkoltechNLP at SemEval-2021 Task 2: Generating Cross-Lingual Training Data for the Word-in-Context Task 1 Aug 2021 · 0 repositories
-
Emotion Stimulus Detection in German News Headlines 27 Jul 2021 · 0 repositories · arXiv:2107.12920
-
TGIF: Tree-Graph Integrated-Format Parser for Enhanced UD with Two-Stage Generic- to Individual-Language Finetuning 14 Jul 2021 · 0 repositories · arXiv:2107.06907
-
A Primer on Pretrained Multilingual Language Models 1 Jul 2021 · 0 repositories · arXiv:2107.00676
-
Automatic Sexism Detection with Multilingual Transformer Models 9 Jun 2021 · 0 repositories · arXiv:2106.04908
-
Investigating Transfer Learning in Multilingual Pre-trained Language Models through Chinese Natural Language Inference 7 Jun 2021 · 1 repository · arXiv:2106.03983
-
How to Adapt Your Pretrained Multilingual Model to 1600 Languages 3 Jun 2021 · 0 repositories · arXiv:2106.02124
-
Diagnosing Transformers in Task-Oriented Semantic Parsing 27 May 2021 · 0 repositories · arXiv:2105.13496
-
XeroAlign: Zero-Shot Cross-lingual Transformer Alignment 6 May 2021 · 2 repositories · arXiv:2105.02472
-
Larger-Scale Transformers for Multilingual Masked Language Modeling 2 May 2021 · 0 repositories · arXiv:2105.00572
-
Multilingual and Zero-Shot is Closing in on Monolingual Web Register Classification 1 May 2021 · 0 repositories
-
XLM-T: Multilingual Language Models in Twitter for Sentiment Analysis and Beyond 25 Apr 2021 · 1 repository · arXiv:2104.12250Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
X-METRA-ADA: Cross-lingual Meta-Transfer Learning Adaptation to Natural Language Understanding and Question Answering 20 Apr 2021 · 1 repository · arXiv:2104.09696
-
AmericasNLI: Evaluating Zero-shot Natural Language Understanding of Pretrained Multilingual Models in Truly Low-resource Languages 18 Apr 2021 · 1 repository · arXiv:2104.08726
-
Emotion Classification in a Resource Constrained Language Using Transformer-based Approach 17 Apr 2021 · 4 repositories · arXiv:2104.08613
-
Improving Zero-Shot Cross-Lingual Transfer Learning via Robust Training 17 Apr 2021 · 1 repository · arXiv:2104.08645
-
Bilingual alignment transfers to multilingual alignment for unsupervised parallel text mining 15 Apr 2021 · 1 repository · arXiv:2104.07642
-
COVID-19 Named Entity Recognition for Vietnamese 8 Apr 2021 · 1 repository · arXiv:2104.03879
-
MCL@IITK at SemEval-2021 Task 2: Multilingual and Cross-lingual Word-in-Context Disambiguation using Augmented Data, Signals, and Transformers 4 Apr 2021 · 0 repositories · arXiv:2104.01567
-
Benchmarking Pre-trained Language Models for Multilingual NER: TraSpaS at the BSNLP2021 Shared Task 1 Apr 2021 · 1 repository
-
Challenges in Annotating and Parsing Spoken, Code-switched, Frisian-Dutch Data 1 Apr 2021 · 1 repository
-
Priberam Labs at the 3rd Shared Task on SlavNER 1 Apr 2021 · 0 repositories
-
Automatic Difficulty Classification of Arabic Sentences 7 Mar 2021 · 0 repositories · arXiv:2103.04386
-
Vyākarana: A Colorless Green Benchmark for Syntactic Evaluation in Indic Languages 1 Mar 2021 · 0 repositories · arXiv:2103.00854
-
NLP-CUET@DravidianLangTech-EACL2021: Offensive Language Detection from Multilingual Code-Mixed Text using Transformers 28 Feb 2021 · 1 repository · arXiv:2103.00455
-
Bootstrapping Multilingual AMR with Contextual Word Alignments 3 Feb 2021 · 0 repositories · arXiv:2102.02189
-
LOME: Large Ontology Multilingual Extraction 28 Jan 2021 · 0 repositories · arXiv:2101.12175
-
Distilling Large Language Models into Tiny and Effective Students using pQRNN 21 Jan 2021 · 0 repositories · arXiv:2101.08890
-
ARBERT & MARBERT: Deep Bidirectional Transformers for Arabic 27 Dec 2020 · 2 repositories · arXiv:2101.01785
-
CogALex-VI Shared Task: Transrelation - A Robust Multilingual Language Model for Multilingual Relation Identification 12 Dec 2020 · 1 repository
-
A Multilingual Reading Comprehension System for more than 100 Languages 1 Dec 2020 · 0 repositories
-
Detecting Urgency Status of Crisis Tweets: A Transfer Learning Approach for Low Resource Languages 1 Dec 2020 · 1 repository
-
SenseCluster at SemEval-2020 Task 1: Unsupervised Lexical Semantic Change Detection 1 Dec 2020 · 0 repositories
-
XGLUE: A New Benchmark Datasetfor Cross-lingual Pre-training, Understanding and Generation 1 Nov 2020 · 0 repositories
-
Applying Occam's Razor to Transformer-Based Dependency Parsing: What Works, What Doesn't, and What is Really Necessary 23 Oct 2020 · 2 repositories · arXiv:2010.12699
-
Galileo at SemEval-2020 Task 12: Multi-lingual Learning for Offensive Language Identification using Pre-trained Language Models 7 Oct 2020 · 0 repositories · arXiv:2010.03542
-
A Pilot Study of Text-to-SQL Semantic Parsing for Vietnamese 5 Oct 2020 · 1 repository · arXiv:2010.01891
-
Cross-lingual Transfer Learning for Semantic Role Labeling in Russian 1 Sep 2020 · 0 repositories
-
On Learning Universal Representations Across Languages 31 Jul 2020 · 0 repositories · arXiv:2007.15960
-
FinEst BERT and CroSloEngual BERT: less is more in multilingual models 14 Jun 2020 · 0 repositories · arXiv:2006.07890
-
English Intermediate-Task Training Improves Zero-Shot Cross-Lingual Transfer Too 3 Jun 2020 · 0 repositories
-
English Intermediate-Task Training Improves Zero-Shot Cross-Lingual Transfer Too 26 May 2020 · 0 repositories · arXiv:2005.13013
-
MAD-X: An Adapter-Based Framework for Multi-Task Cross-Lingual Transfer 30 Apr 2020 · 3 repositories · arXiv:2005.00052
-
Testing pre-trained Transformer models for Lithuanian news clustering 3 Apr 2020 · 0 repositories · arXiv:2004.03461
-
XGLUE: A New Benchmark Dataset for Cross-lingual Pre-training, Understanding and Generation 3 Apr 2020 · 2 repositories · arXiv:2004.01401
-
PhoBERT: Pre-trained language models for Vietnamese 2 Mar 2020 · 1 repository · arXiv:2003.00744
-
Unsupervised Cross-lingual Representation Learning at Scale 5 Nov 2019 · 35 repositories · arXiv:1911.02116Syntology official (archive's flag): 12 ran · 40 ran (of which 13 constructed an object rather than computing a result; 35 with no instrument failure: 4 honoured, 1 violated, 30 with no contract checked; 5 where Syntology's instrument failed) · 19 unverified (of 59 harvested samples) · 52 pointer-only (licence)