Strategies to deal with legacy ICD data must address the issue of codes created by non-taxonomist users. The NLM core subset possibly needs augmentation with concepts from certain SNOMED hierarchies, notably qualifiers, body structures, substances/products and organisms. Concept-matching software needs to utilize query expansion strategies, but these may be effective in production settings only if a large but non-redundant SNOMED subset that minimizes the proportion of extensively pre-coordinated concepts is also available.
No takes yet. Share an insight, caveat, or question.
Nadkarni et al. (2010) studied this question.
Synapse has enriched 3 closely related papers on similar clinical questions. Consider them for comparative context: