Browse State-of-the-Art › de-en
de-en
33 papers with code · 0 benchmarks · 1 dataset archive 2025-07-28
Benchmarks archive 2025-07-28
No benchmark for this task in the archive.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
1 dataset whose archive record lists this task, ordered by the archive's paper count.
Subtasks archive 2025-07-28
No subtask under this task in the archive's task tree.
Most implemented papers archive 2025-07-28
30 shown of 33 papers with code (82 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
9 Mar 2020 5 repositories listedSeveral papers argue that wide minima generalize better than narrow minima.
-
17 Feb 2020 5 repositories listed Syntology ran 1 of 3 samples · 2 unverified · 3 pointer-only (licence)We also apply BatchEnsemble to lifelong learning, where on Split-CIFAR-100, BatchEnsemble yields comparable performance to progressive neural networks while having a much lower computational and memory costs.
-
19 Oct 2018 3 repositories listed Syntology ran 5 of 5 samples · 0 unverified · 5 pointer-only (licence)Simultaneous translation, which translates sentences before they are finished, is useful in many scenarios but is notoriously difficult due to word-order differences.
-
9 Sep 2021 2 repositories listed Syntology ran 2 of 2 samples · 0 unverified · 1 pointer-only (licence)The success of bidirectional encoders using masked language models, such as BERT, on numerous natural language processing tasks has prompted researchers to attempt to incorporate these pre-trained models into neural…
-
10 Oct 2016 2 repositories listed Syntology ran 1 of 3 samples · 2 unverifiedWe observe that on CS-EN, FI-EN and RU-EN, the quality of the multilingual character-level translation even surpasses the models specifically trained on that language pair alone, both in terms of BLEU score and human…
-
19 Mar 2016 2 repositories listed Syntology ran 1 of 3 samples · 2 unverifiedThe existing machine translation systems, whether phrase-based or neural, have relied almost exclusively on word-level modelling with explicit segmentation.
-
5 Jun 2024 1 repository listed Syntology ran 4 of 4 samples · 0 unverifiedSimultaneous speech-to-speech translation (Simul-S2ST, a.
-
14 May 2024 1 repository listedThis article explores the adaptive relationship between Encoder Layers and Decoder Layers using the SOTA model Helsinki-NLP/opus-mt-de-en, which translates German to English.
-
17 Feb 2024 1 repository listedMinimum Bayes risk (MBR) decoding achieved state-of-the-art translation performance by using COMET, a neural metric that has a high correlation with human evaluation.
-
30 Oct 2023 1 repository listedAs part of the WMT-2023 "Test suites" shared task, in this paper we summarize the results of two test suites evaluations: MuST-SHE-WMT23 and INES.
-
20 Oct 2023 1 repository listedThere is a lack of research into capabilities of recent LLMs to generate convincing text in languages other than English and into performance of detectors of machine-generated text in multilingual settings.
-
20 Oct 2023 1 repository listedBack translation (BT) is one of the most significant technologies in NMT research fields.
-
2 Jun 2023 1 repository listedSubword tokenization is the de facto standard for tokenization in neural language models and machine translation systems.
-
15 May 2023 1 repository listedMotivated by the remarkable success of back translation in MT, we develop a back translation algorithm for ST (BT4ST) to synthesize pseudo ST data from monolingual target data.
-
25 Apr 2023 1 repository listedIt is well-known that document context is vital for resolving a range of translation ambiguities, and in fact the document setting is the most natural setting for nearly all translation.
-
15 Feb 2023 1 repository listed Syntology ran 2 of 3 samples · 1 unverifiedTo address this, we propose Big Little Decoder (BiLD), a framework that can improve inference efficiency and latency for a wide range of text generation applications.
-
20 Sep 2022 1 repository listedAs for model sizes, we scale the Transformer-Big up to the extremely large model that owns nearly 4.
-
DiMS: Distilling Multiple Steps of Iterative Non-Autoregressive Transformers for Machine Translation7 Jun 2022 1 repository listedThe student is optimized to predict the output of the teacher after multiple decoding steps while the teacher follows the student via a slow-moving average.
-
6 Jun 2022 1 repository listedWe introduce Bi-SimCut: a simple but effective training strategy to boost neural machine translation (NMT) performance.
-
17 Mar 2022 1 repository listed Syntology ran 0 of 1 samples · 1 unverifiedFor SiMT policy, GMA models the aligned source position of each target word, and accordingly waits until its aligned position to start translating.
-
17 Mar 2022 1 repository listed Syntology ran 4 of 9 samples · 5 unverifiedAccording to duality constraints, the read/write path in source-to-target and target-to-source SiMT models can be mapped to each other.
-
10 Feb 2022 1 repository listedNeural metrics have achieved impressive correlation with human judgements in the evaluation of machine translation systems, but before we can safely optimise towards such metrics, we should be aware of (and ideally…
-
16 Dec 2021 1 repository listedOn examples with a maximum source and target length of 30 from De-En, WMT'16 English-Romanian, and WMT'21 English-Chinese translation tasks, our learned order outperforms all heuristic generation orders on four out of…
-
1 Nov 2021 1 repository listedTowards keeping the consistency of data distribution with iterative decoding, an iterative training strategy is employed to further improve the capacity of rewriting.
-
26 Jul 2021 1 repository listedIn this paper, we evaluate the translation of negation both automatically and manually, in English--German (EN--DE) and English--Chinese (EN--ZH).
-
18 Apr 2021 1 repository listedIn the lowest-resource setting, we outperform GIZA++ by 8.
-
1 Mar 2021 1 repository listedIn OmniNet, instead of maintaining a strictly horizontal receptive field, each token is allowed to attend to all tokens in the entire network.
-
1 Nov 2020 1 repository listedPrior to fine-tuning, our method replaces the embedding layers of the NMT model by projecting general word embeddings induced from monolingual data in a target domain onto a source-domain embedding space.
-
1 Nov 2020 1 repository listedWe report the results of the first edition of the WMT shared task on chat translation.
-
15 Oct 2020 1 repository listedOur sentence-level model shows a 0.
Syntology lines on 9 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections