Literature DB >> 24976874

A Study of the Morpho-Semantic Relationship in Medline.

W John Wilbur1, Larry Smith1.   

Abstract

Morphological analysis as applied to English has generally involved the study of rules for inflections and derivations. Recent work has attempted to derive such rules from automatic analysis of corpora. Here we study similar issues, but in the context of the biological literature. We introduce a new approach which allows us to assign probabilities of the semantic relatedness of pairs of tokens that occur in text in consequence of their relatedness as character strings. Our analysis is based on over 84 million sentences from the MEDLINE database, over 2.3 million token types that occur in MEDLINE, and enables us to identify over 36 million token type pairs which have assigned probabilities of semantic relatedness of at least 0.7 based on their similarity as strings. The quality of these predictions is tested by two different manual evaluations and found to be good.

Entities:  

Keywords:  cost function; lexical similarity; morphology; mutual information; potential; weight

Year:  2013        PMID: 24976874      PMCID: PMC4072344          DOI: 10.2174/1874133920131121001

Source DB:  PubMed          Journal:  Open Inf Syst J        ISSN: 1874-1339


  3 in total

1.  MedPost: a part-of-speech tagger for bioMedical text.

Authors:  L Smith; T Rindflesch; W J Wilbur
Journal:  Bioinformatics       Date:  2004-04-08       Impact factor: 6.937

2.  Gauging Similarity with n-Grams: Language-Independent Categorization of Text.

Authors:  M Damashek
Journal:  Science       Date:  1995-02-10       Impact factor: 47.728

3.  A family of similarity measures between two strings.

Authors:  N V Findler; J Van Leeuwen
Journal:  IEEE Trans Pattern Anal Mach Intell       Date:  1979-01       Impact factor: 6.226

  3 in total
  1 in total

1.  Better synonyms for enriching biomedical search.

Authors:  Lana Yeganova; Sun Kim; Qingyu Chen; Grigory Balasanov; W John Wilbur; Zhiyong Lu
Journal:  J Am Med Inform Assoc       Date:  2020-12-09       Impact factor: 4.497

  1 in total

北京卡尤迪生物科技股份有限公司 © 2022-2023.