Browse State-of-the-Art › XLM-R
XLM-R
99 papers with code · 0 benchmarks · 1 dataset archive 2025-07-28
XLM-R
Description from the archive archive 2025-07-28.
Benchmarks archive 2025-07-28
No benchmark for this task in the archive.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
1 dataset whose archive record lists this task, ordered by the archive's paper count.
Subtasks archive 2025-07-28
No subtask under this task in the archive's task tree.
Parent tasks archive 2025-07-28
Most implemented papers archive 2025-07-28
30 shown of 99 papers with code (221 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
5 Nov 2019 35 repositories listed Syntology ran 25 of 59 samples · 34 unverified · 52 pointer-only (licence)We also present a detailed empirical analysis of the key factors that are required to achieve these gains, including the trade-offs between (1) positive transfer and capacity dilution and (2) the performance of high and…
-
15 Jul 2020 9 repositories listed Syntology ran 5 of 15 samples · 10 unverified · 12 pointer-only (licence)We propose AdapterHub, a framework that allows dynamic "stitching-in" of pre-trained adapters for different tasks and languages.
-
18 Apr 2022 6 repositories listed Syntology ran 1 of 10 samples · 9 unverified · 1 pointer-only (licence)We present the MASSIVE dataset--Multilingual Amazon Slu resource package (SLURP) for Slot-filling, Intent classification, and Virtual assistant Evaluation.
-
17 Apr 2021 4 repositories listedA Bengali emotion corpus consists of 6243 texts is developed for the classification task.
-
25 Jan 2023 3 repositories listedLarge multilingual language models typically rely on a single vocabulary shared across 100+ languages.
-
18 Nov 2021 3 repositories listed Syntology ran 0 of 7 samples · 7 unverifiedWe thus propose a new gradient-disentangled embedding sharing method that avoids the tug-of-war dynamics, improving both training efficiency and the quality of the pre-trained model.
-
20 May 2020 3 repositories listedWe present BERTweet, the first public large-scale pre-trained language model for English Tweets.
-
30 Apr 2020 3 repositories listedThe main goal behind state-of-the-art pre-trained multilingual models such as multilingual BERT and XLM-R is enabling and bootstrapping NLP applications in low-resource languages through zero-shot or few-shot…
-
11 Oct 2024 2 repositories listedIn this paper, we investigate the use of N-gram models and Large Pre-trained Multilingual models for Language Identification (LID) across 11 South African languages.
-
13 May 2024 2 repositories listed Syntology ran 7 of 12 samples · 5 unverified · 4 pointer-only (licence)Finally, we show that a ZeTT hypernetwork trained for a base (L)LM can also be applied to fine-tuned variants without extra training.
-
23 May 2023 2 repositories listedHowever, if we want to use a new tokenizer specialized for the target language, we cannot transfer the source model's embedding matrix.
-
22 May 2023 2 repositories listedThe benchmark includes a diverse set of datasets for low-, medium- and high-resource tasks.
-
3 Apr 2023 2 repositories listedIn addition, we examine its performance on two NLG tasks from GreekSUM, a newly introduced summarization dataset for the Greek language.
-
22 Nov 2022 2 repositories listed Syntology ran 2 of 6 samples · 4 unverified · 6 pointer-only (licence)Vision language pre-training aims to learn alignments between vision and language from a large amount of data.
-
12 Nov 2022 2 repositories listed Syntology ran 2 of 11 samples · 9 unverifiedIn this work, we present a conceptually simple and effective method to train a strong bilingual/multilingual multimodal representation model.
-
6 May 2021 2 repositories listedThe introduction of pretrained cross-lingual language models brought decisive improvements to multilingual NLP tasks.
-
31 Dec 2020 2 repositories listedWe generalize deep self-attention distillation in MiniLM (Wang et al., 2020) by only using self-attention relation distillation for task-agnostic compression of pretrained Transformers.
-
27 Dec 2020 2 repositories listedTo evaluate our models, we also introduce ARLUE, a new benchmark for multi-dialectal Arabic language understanding evaluation.
-
23 Oct 2020 2 repositories listedWe find that the choice of pre-trained embeddings has by far the greatest impact on parser performance and identify XLM-R as a robust choice across the languages in our study.
-
2 Jul 2020 2 repositories listedIn this paper, we present a Bayesian multilingual document model for learning language-independent document embeddings.
-
3 Apr 2020 2 repositories listedIn this paper, we introduce XGLUE, a new benchmark dataset that can be used to train large-scale cross-lingual pre-trained models using multilingual and bilingual corpora and evaluate their performance across a diverse…
-
15 Feb 2025 1 repository listedWhile multilingual language models like XLM-R have advanced multilingualism in NLP, they still perform poorly in extremely low-resource languages.
-
26 Sep 2024 1 repository listedHowever, this removal increases the burden on token embeddings to encode all language-specific information, which may hinder the model's ability to produce more language-neutral representations.
-
26 Sep 2024 1 repository listedContextualized embeddings based on large language models (LLMs) are available for various languages, but their coverage is often limited for lower resourced languages.
-
19 Jun 2024 1 repository listedTo our knowledge, our Vietnamese real-world dataset is the largest spoken NER dataset in the world regarding the number of entity types, featuring 18 distinct types.
-
23 May 2024 1 repository listed Syntology ran 1 of 1 samples · 0 unverified · 1 pointer-only (licence)Our analysis focuses on quantifying the \textit{alignment} and \textit{overlap} of these concepts across various languages within the latent space.
-
29 Mar 2024 1 repository listedMultilingual Language Models (MLLMs) exhibit robust cross-lingual transfer capabilities, or the ability to leverage information acquired in a source language and apply it to a target language.
-
18 Mar 2024 1 repository listedContaminated or adulterated food poses a substantial risk to human health.
-
20 Nov 2023 1 repository listedWe achieve a biomedical multilingual corpus by incorporating three granularity knowledge alignments (entity, fact, and passage levels) into monolingual corpora.
-
15 Nov 2023 1 repository listedIn this work, we present the largest benchmark to date on linguistic acceptability: Multilingual Evaluation of Linguistic Acceptability -- MELA, with 46K samples covering 10 languages from a diverse set of language…
Syntology lines on 8 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections