Browse State-of-the-Art › Bilingual Lexicon Induction
Bilingual Lexicon Induction
35 papers with code · 0 benchmarks · 0 datasets archive 2025-07-28
Translate words from one language to another.
Description from the archive archive 2025-07-28.
Benchmarks archive 2025-07-28
No benchmark for this task in the archive.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
No dataset record in the archive lists this task.
Subtasks archive 2025-07-28
No subtask under this task in the archive's task tree.
Parent tasks archive 2025-07-28
Most implemented papers archive 2025-07-28
30 shown of 35 papers with code (109 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
10 Feb 2024 2 repositories listedIt was recently shown that it is possible to infer such lexicon, without using any parallel data, by aligning word embeddings trained on monolingual data.
-
10 Oct 2019 2 repositories listedLearning multilingual representations of text has proven a successful method for many cross-lingual transfer learning tasks.
-
27 Aug 2018 2 repositories listedOur approach decouples learning the transformation from the source language to the target language into (a) learning rotations for language-specific embeddings to align them to a common space, and (b) learning a…
-
15 Feb 2024 1 repository listed Syntology ran 7 of 10 samples · 3 unverifiedRecent work has shown that, while large language models (LLMs) demonstrate strong word translation or bilingual lexicon induction (BLI) capabilities in few-shot setups, they still cannot match the performance of…
-
28 Oct 2023 1 repository listedWe also demonstrate the effectiveness of ProMap in re-ranking results from other BLI methods such as with aligned static word embeddings.
-
21 Oct 2023 1 repository listed Syntology ran 3 of 4 samples · 1 unverifiedBilingual Lexicon Induction (BLI) is a core task in multilingual NLP that still, to a large extent, relies on calculating cross-lingual word representations.
-
23 May 2023 1 repository listedMost existing approaches for unsupervised bilingual lexicon induction (BLI) depend on good quality static or contextual embeddings requiring large monolingual corpora for both languages.
-
19 Apr 2023 1 repository listedBilingual word lexicons are crucial tools for multilingual natural language understanding and machine translation tasks, as they facilitate the mapping of words in one language to their synonyms in another language.
-
30 Oct 2022 1 repository listed Syntology ran 0 of 4 samples · 4 unverifiedThis crucial step is done via 1) creating a word similarity dataset, comprising positive word pairs (i.
-
18 Oct 2022 1 repository listed Syntology ran 6 of 6 samples · 0 unverified · 6 pointer-only (licence)Bilingual lexicon induction induces the word translations by aligning independently trained word embeddings in two languages.
-
11 Oct 2022 1 repository listedThe ability to extract high-quality translation dictionaries from monolingual word embedding spaces depends critically on the geometric similarity of the spaces -- their degree of "isomorphism."
-
15 Mar 2022 1 repository listed Syntology ran 5 of 6 samples · 1 unverified · 4 pointer-only (licence)At Stage C1, we propose to refine standard cross-lingual linear maps between static word embeddings (WEs) via a contrastive learning objective; we also show how to integrate it into the self-learning procedure for even…
-
16 Nov 2021 1 repository listedAs Stage C1, we propose to refine standard cross-lingual linear maps between static word embeddings (WEs) via a contrastive learning objective; we also show how to integrate it into the self-learning procedure for even…
-
16 Oct 2021 1 repository listedTo achieve this, it is crucial to represent multilingual knowledge in a shared/unified space.
-
26 Sep 2021 1 repository listedAlternatively, word embeddings may be understood as nodes in a weighted graph.
-
1 Aug 2021 1 repository listedAimed at generating a seed lexicon for use in downstream natural language tasks and unsupervised methods for bilingual lexicon induction have received much attention in the academic literature recently.
-
6 Jun 2021 1 repository listed Syntology ran 7 of 18 samples · 11 unverified · 18 pointer-only (licence)Bilingual Lexicon Induction (BLI) aims to map words in one language to their translations in another, and is typically through learning linear projections to align monolingual word representation spaces.
-
2 Jun 2021 1 repository listedWe find moderate to strong positive correlations between categorical modularity and performance on the monolingual tasks of sentiment analysis and word similarity calculation and on the cross-lingual task of bilingual…
-
1 Jun 2021 1 repository listedIt is therefore recommended that this strategy be adopted as a standard for CLWE methods.
-
11 Apr 2021 1 repository listedIt is therefore recommended that this strategy be adopted as a standard for CLWE methods.
-
18 Mar 2021 1 repository listedSuccessful methods for unsupervised neural machine translation (UNMT) employ crosslingual pretraining via self-supervision, often in the form of a masked language modeling or a sequence generation task, which requires…
-
27 Oct 2020 1 repository listedWe propose a new approach for learning contextualised cross-lingual word embeddings based on a small parallel corpus (e.
-
14 Oct 2020 1 repository listedIn this paper, we propose a new semi-supervised BLI framework to encourage the interaction between the supervised signal and unsupervised alignment.
-
21 Feb 2020 1 repository listedIn this paper, we propose a self-supervised method to refine the alignment of unsupervised bilingual word embeddings.
-
4 Sep 2019 1 repository listedA series of bilingual lexicon induction (BLI) experiments with 15 diverse languages (210 language pairs) show that fully unsupervised CLWE methods still fail for a large number of language pairs (e.
-
19 Aug 2019 1 repository listedWe then propose Bilingual Lexicon Induction with Semi-Supervision (BLISS) --- a semi-supervised approach that relaxes the isometric assumption while leveraging both limited aligned bilingual lexicons and a larger set of…
-
24 Jul 2019 1 repository listedA recent research line has obtained strong results on bilingual lexicon induction by aligning independently trained word embeddings in two languages and using the resulting cross-lingual embeddings to induce word…
-
1 Jul 2019 1 repository listedRecent advances in BLI work by aligning the two word embedding spaces.
-
4 Apr 2019 1 repository listed Syntology ran 0 of 9 samples · 9 unverifiedRecent approaches to cross-lingual word embedding have generally been based on linear transformations between the sets of embedding vectors in the two languages.
-
1 Feb 2019 1 repository listedIn this work, we make the first step towards a comprehensive evaluation of cross-lingual word embeddings.
Syntology lines on 7 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections