An Unsupervised Approach to Term Variant Identification

WANG Bao-xun,WANG Xiao-long,LIU Bing-quan,LI Peng
DOI: https://doi.org/10.3969/j.issn.1003-0077.2008.03.004
2008-01-01
Abstract:This paper presents an unsupervised learning strategy to identify the variants of biomedical terms.The minimum edit distance algorithm and a character-matching algorithm are first applied to identify the morphological variants and the abbreviations as the candidate variants for a given term.The system similarity model is innovatively introduced to measure the semantic context for each candidate variant.This method requires no linguistic knowledge or labor-intensive corpora,and the experiment indicates its significant improvement in recall with a reasonable precision.
What problem does this paper attempt to address?