Papers › Syntax-based data augmentation for Hungarian-English machine translation

Syntax-based data augmentation for Hungarian-English machine translation

18 Jan 2022arXiv:2201.06876archive 2025-07-28

Attila Nagy, Patrick Nanys, Balázs Frey Konrád, Bence Bial, Judit Ács

We train Transformer-based neural machine translation models for Hungarian-English and English-Hungarian using the Hunglish2 corpus. Our best models achieve a BLEU score of 40.0 on HungarianEnglish and 33.4 on English-Hungarian. Furthermore, we present results on an ongoing work about syntax-based augmentation for neural machine translation. Both our code and models are publicly available.

PaperPDFCode

Code

attilanagy234/syntax-augmentation-nmt officialmentioned in papermentioned on GitHubpytorch report
attilanagy234/treeswap mentioned on GitHubpytorch report

Repository list and official/mentioned flags are the archive's, frozen 2025-07-28. Reachability, where shown, is from one Syntology probe window (2026-09-16 to 2026-09-18); repositories not probed show nothing. GitHub stars are not tracked.

Code Syntology ran Syntology

Not run by Syntology. Nothing on this page verifies that the listed code works.

Tasks

Data AugmentationMachine TranslationTranslation

Results from the paper archive 2025-07-28

No leaderboard rows for this paper in the archive.

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections