A prerequisite for all higher level information extraction tasks is the identification of unknown names in text. This paper presents a method for extracting protein names from abstracts of articles in the biomedical domain. These names present several interesting difficulties because of their variant structural characteristics and the lack of common naming standards and fixed nomenclatures in the domain.
No takes yet. Share an insight, caveat, or question.
Eriksson et al. (2002) studied this question.
Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context: