PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
January 1, 2017490 citationsOpen Access

Learning bilingual word embeddings with (almost) no bilingual data

MAMikel ArtetxeGLGorka LabakaEAEneko Agirre

Key Points

Key points are not available for this paper at this time.

Abstract

Most methods to learn bilingual word embeddings rely on large parallel corpora, which is difficult to obtain for most language pairs. This has motivated an active research line to relax this requirement, with methods that use document-aligned corpora or bilingual dictionaries of a few thousand words instead. In this work, we further reduce the need of bilingual resources using a very simple self-learning approach that can be combined with any dictionary-based mapping technique. Our method exploits the structural similarity of embedding spaces, and works with as little bilingual evidence as a 25 word dictionary or even an automatically generated list of numerals, obtaining results comparable to those of systems that use richer resources.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Artetxe et al. (2017) studied this question.

synapsesocial.com/papers/6a12b91816f0ac689b9e2c89https://doi.org/10.18653/v1/p17-1042
Ask AI
Helpful
Bookmark
Share
View Full Paper