JOURNAL ARTICLE

Bootstrapping Unsupervised Bilingual Lexicon Induction

Abstract

The task of unsupervised lexicon induction is to find translation pairs across monolingual corpora. We develop a novel method that creates seed lexicons by identifying cognates in the vocabularies of related languages on the basis of their frequency and lexical similarity. We apply bidirectional bootstrapping to a method which learns a linear mapping between context-based vector spaces. Experimental results on three language pairs show consistent improvement over prior work.

Keywords:
Bootstrapping (finance) Computer science Lexicon Natural language processing Artificial intelligence Context (archaeology) Similarity (geometry) Task (project management) Basis (linear algebra) Speech recognition Mathematics

Metrics

17
Cited By
2.98
FWCI (Field Weighted Citation Impact)
19
Refs
0.92
Citation Normalized Percentile
Is in top 1%
Is in top 10%

Citation History

Topics

Natural Language Processing Techniques
Physical Sciences →  Computer Science →  Artificial Intelligence
Topic Modeling
Physical Sciences →  Computer Science →  Artificial Intelligence
Speech and dialogue systems
Physical Sciences →  Computer Science →  Artificial Intelligence
© 2026 ScienceGate Book Chapters — All rights reserved.