Methods › Natural Language Processing › Language Models › mBART
mBART
archive 2025-07-28 Description, source and code snippet are the archive's method entry.
mBART is a sequence-to-sequence denoising auto-encoder pre-trained on large-scale monolingual corpora in many languages using the BART objective. The input texts are noised by masking phrases and permuting sentences, and a single Transformer model is learned to recover the texts. Different from other pre-training approaches for machine translation, mBART pre-trains a complete autoregressive Seq2Seq model. mBART is trained once for all languages, providing a set of parameters that can be fine-tuned for any of the language pairs in both supervised and unsupervised settings, without any task-specific or language-specific modifications or initialization schemes.
Papers archive 2025-07-28
30 shown of 66, newest first. Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code; it is per sample and not a correctness claim.
-
State-of-the-Art Translation of Text-to-Gloss using mBART : A case study of Bangla 3 Apr 2025 · 0 repositories · arXiv:2504.02293
-
BeliN: A Novel Corpus for Bengali Religious News Headline Generation using Contextual Feature Fusion 2 Jan 2025 · 1 repository · arXiv:2501.01069
-
SiTSE: Sinhala Text Simplification Dataset and Evaluation 2 Dec 2024 · 1 repository · arXiv:2412.01293
-
Abstractive Summarization of Low resourced Nepali language using Multilingual Transformers 29 Sep 2024 · 0 repositories · arXiv:2409.19566
-
Exploring Fine-tuned Generative Models for Keyphrase Selection: A Case Study for Russian 16 Sep 2024 · 0 repositories · arXiv:2409.10640
-
Segment-Based Interactive Machine Translation for Pre-trained Models 9 Jul 2024 · 0 repositories · arXiv:2407.06990
-
NAIST Simultaneous Speech Translation System for IWSLT 2024 30 Jun 2024 · 0 repositories · arXiv:2407.00826
-
CantonMT: Cantonese to English NMT Platform with Fine-Tuned Models Using Synthetic Back-Translation Data 17 Mar 2024 · 1 repository · arXiv:2403.11346
-
VBART: The Turkish LLM 2 Mar 2024 · 0 repositories · arXiv:2403.01308
-
Transcription and translation of videos using fine-tuned XLSR Wav2Vec2 on custom dataset and mBART 1 Mar 2024 · 0 repositories · arXiv:2403.00212
-
Key ingredients for effective zero-shot cross-lingual knowledge transfer in generative tasks 19 Feb 2024 · 0 repositories · arXiv:2402.12279
-
End to end Hindi to English speech conversion using Bark, mBART and a finetuned XLSR Wav2Vec2 11 Jan 2024 · 0 repositories · arXiv:2401.06183
-
Exploring Automatic Text Simplification of German Narrative Documents 15 Dec 2023 · 1 repository · arXiv:2312.09907
-
Context-aware Neural Machine Translation for English-Japanese Business Scene Dialogues 20 Nov 2023 · 1 repository · arXiv:2311.11976
-
Empirical study of pretrained multilingual language models for zero-shot cross-lingual knowledge transfer in generation 15 Oct 2023 · 0 repositories · arXiv:2310.09917
-
Direct Text to Speech Translation System using Acoustic Units 14 Sep 2023 · 0 repositories · arXiv:2309.07478
-
Benchmarking Procedural Language Understanding for Low-Resource Languages: A Case Study on Turkish 13 Sep 2023 · 1 repository · arXiv:2309.06698
-
CLIPTrans: Transferring Visual Knowledge with Pre-trained Models for Multimodal Machine Translation 29 Aug 2023 · 1 repository · arXiv:2308.15226Syntology ran 5 of 7 samples · 2 unverified · 7 pointer-only (licence)
-
VBD-MT Chinese-Vietnamese Translation Systems for VLSP 2022 15 Aug 2023 · 0 repositories · arXiv:2308.07601
-
Binary and Ternary Natural Language Generation 2 Jun 2023 · 1 repository · arXiv:2306.01841Syntology ran 1 of 1 samples · 0 unverified · 1 pointer-only (licence)
-
Improving Polish to English Neural Machine Translation with Transfer Learning: Effects of Data Volume and Language Similarity 1 Jun 2023 · 0 repositories · arXiv:2306.00660
-
DEPLAIN: A German Parallel Corpus with Intralingual Translations into Plain Language for Sentence and Document Simplification 30 May 2023 · 1 repository · arXiv:2305.18939
-
Enhancing Translation for Indigenous Languages: Experiments with Multilingual Models 27 May 2023 · 0 repositories · arXiv:2305.17406
-
Extrapolating Multilingual Understanding Models as Multilingual Generators 22 May 2023 · 0 repositories · arXiv:2305.13140
-
mLongT5: A Multilingual and Efficient Text-To-Text Transformer for Longer Sequences 18 May 2023 · 1 repository · arXiv:2305.11129
-
Towards Speech Dialogue Translation Mediating Speakers of Different Languages 16 May 2023 · 1 repository · arXiv:2305.09210
-
Prompt-Learning for Cross-Lingual Relation Extraction 20 Apr 2023 · 1 repository · arXiv:2304.10354
-
Neural Machine Translation For Low Resource Languages 16 Apr 2023 · 1 repository · arXiv:2304.07869
-
PEACH: Pre-Training Sequence-to-Sequence Multilingual Models for Translation with Semi-Supervised Pseudo-Parallel Document Generation 3 Apr 2023 · 1 repository · arXiv:2304.01282
-
Summarizing Indian Languages using Multilingual Transformers based Models 29 Mar 2023 · 0 repositories · arXiv:2303.16657
Tasks archive 2025-07-28
20 shown of 76 tasks the archive attaches to papers tagged with this method, by distinct papers. A task without a page in the catalog is plain text.
Usage over time archive 2025-07-28
Components: the archive holds no method-to-method composition, so PwC's Components table cannot be rebuilt; the Papers list carries no Results column for the same reason (the archive does not join its leaderboard rows to method tags).
Categories archive 2025-07-28
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections