PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
January 1, 200166 citationsOpen Access

Semi-supervised Maximum Entropy based approach to acronym and abbreviation normalization in medical texts

SPSerguei Pakhomov

Key Points

Key points are not available for this paper at this time.

Abstract

Text normalization is an important aspect of successful information retrieval from medical documents such as clinical notes, radiology reports and discharge summaries. In the medical domain, a significant part of the general problem of text normalization is abbreviation and acronym disambiguation. Numerous abbreviations are used routinely throughout such texts and knowing their meaning is critical to data retrieval from the document. In this paper I will demonstrate a method of automatically generating training data for Maximum Entropy (ME) modeling of abbreviations and acronyms and will show that using ME modeling is a promising technique for abbreviation and acronym normalization. I report on the results of an experiment involving training a number of ME models used to normalize abbreviations and acronyms on a sample of 10,000 rheumatology notes with ~89% accuracy.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Serguei Pakhomov (2001) studied this question.

synapsesocial.com/papers/6a155c1cd64fa333899f914dhttps://doi.org/10.3115/1073083.1073111
Ask AI
Helpful
Bookmark
Share
View Full Paper

Also Consider

Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context:

  1. 1A maximum entropy approach to natural language processing1996 · 3,116 citations
  2. 2An experiment in computational discrimination of English word senses1988 · 69 citations
  3. 3A study of abbreviations in the UMLS.2001 · 74 citations
  4. 4Speech and language processing2010 · 3,606 citations
  5. 5Evaluating the UMLS as a source of lexical knowledge for medical language processing.2001 · 83 citations