KDnuggets : News : 2003 : n10 : item17 | PREVIOUS | NEXT |
PublicationsSubject: Paper: Terminology-driven mining of biomedical literature
Results: In this paper, we present an overview of an integrated framework for terminology-driven mining from biomedical literature. The framework integrates the following components: automatic term recognition, term variation handling, acronym acquisition, automatic discovery of term similarities and term clustering. The term variant recognition is incorporated into terminology recognition process by taking into account orthographical, morphological, syntactic, lexico-semantic and pragmatic term variations. In particular, we address acronyms as a common way of introducing term variants in biomedical papers. Term clustering is based on the automatic discovery of term similarities. We use a hybrid similarity measure, where terms are compared by using both internal and external evidence. The measure combines lexical, syntactical and contextual similarity. Experiments on terminology recognition and clustering performed on a corpus of MEDLINE abstracts recorded the precision of 98 and 71% respectively.
|
KDnuggets : News : 2003 : n10 : item17 | PREVIOUS | NEXT |
Copyright © 2003 KDnuggets. Subscribe to KDnuggets News!