PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
August 22, 2012Computational Linguistics29 citationsOpen Access

A Joint Model to Identify and Align Bilingual Named Entities

YCYufeng ChenCZChengqing ZongKSKeh‐Yih Su

Key Points

Key points are not available for this paper at this time.

Abstract

In this article, an integrated model is derived that jointly identifies and aligns bilingual named entities (NEs) between Chinese and English. The model is motivated by the following observations: (1) whether an NE is translated semantically or phonetically depends greatly on its entity type, (2) entities within an aligned pair should share the same type, and (3) the initially detected NEs can act as anchors and provide further information while selecting NE candidates. Based on these observations, this article proposes a translation mode ratio feature (defined as the proportion of NE internal tokens that are semantically translated), enforces an entity type consistency constraint, and utilizes additional new NE likelihoods (based on the initially detected NE anchors). Experiments show that this novel method significantly outperforms the baseline. The type-insensitive F-score of identified NE pairs increases from 78.4% to 88.0% (12.2% relative improvement) in our Chinese–English NE alignment task, and the type-sensitive F-score increases from 68.4% to 83.0% (21.3% relative improvement). Furthermore, the proposed model demonstrates its robustness when it is tested across different domains. Finally, when semi-supervised learning is conducted to train the adopted English NE recognition model, the proposed model also significantly boosts the English NE recognition type-sensitive F-score.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Chen et al. (2012) studied this question.

synapsesocial.com/papers/6a1700c32fcf950e0005897ehttps://doi.org/10.1162/coli_a_00122
Ask AI
Helpful
Bookmark
Share
View Full Paper

Also Consider

Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context:

  1. 1Improved backing-off for M-gram language modeling2002 · 1,500 citations
  2. 2Proceedings of the 42nd Annual Meeting on Association for Computational Linguistics - ACL '042004 · 191 citations
  3. 3Statistical Models: Theory and Practice2006 · 1,006 citations
  4. 4Description of the LTG system used for MUC-71998 · 164 citations
  5. 5Automatic Extraction of Translational Japanese-KATAKANA and English Word Pairs from Bilingual Corpora2002 · 27 citations