Warning: Undefined array key "mm" in /www/wwwroot/www.ai-bt.com/si.php on line 10 Deprecated: trim(): Passing null to parameter #1 ($string) of type string is deprecated in /www/wwwroot/www.ai-bt.com/si.php on line 10 An analysis of statistical term strength and its use in the indexing and retrieval of molecular biology texts.

Literature DB >> 8725772

An analysis of statistical term strength and its use in the indexing and retrieval of molecular biology texts.

Abstract

The biological literature presents a difficult challenge to information processing in its complexity, diversity, and in its sheer volume. Much of the diversity resides in its technical terminology, which has also become voluminous. In an effort to deal more effectively with this large vocabulary and improve information processing, a method of focus has been developed which allows one to classify terms based on a measure of their importance in describing the content of the documents in which they occur. The measurement is called the strength of a term and is a measure of how strongly the term's occurrences correlate with the subjects of documents in the database. If term occurrences are random then there will be no correlation and the strength will be zero, but if for any subject, the term is either always present or never present its strength will be one. We give here a new, information theoretical interpretation of term strength, review some of its uses in focusing the processing of documents for information retrieval and describe new results obtained in document categorization.

Mesh：

Year: 1996 PMID： 8725772 DOI： 10.1016/0010-4825(95)00055-0

Source DB: PubMed Journal: Comput Biol Med ISSN： 0010-4825 Impact factor: 4.589

Keyword Cloud
Cited

23 in total

1. Including biological literature improves homology search.

Authors: J T Chang; S Raychaudhuri; R B Altman
Journal: Pac Symp Biocomput Date: 2001

2. UMLS concept indexing for production databases: a feasibility study.

Authors: P Nadkarni; R Chen; C Brandt
Journal: J Am Med Inform Assoc Date: 2001 Jan-Feb Impact factor: 4.497

3. The NLM Indexing Initiative.

Authors: A R Aronson; O Bodenreider; H F Chang; S M Humphrey; J G Mork; S J Nelson; T C Rindflesch; W J Wilbur
Journal: Proc AMIA Symp Date: 2000

4. What's related? Generalizing approaches to related articles in medicine.

Authors: H R Strasberg; C D Manning; T C Rindfleisch; K L Melmon
Journal: Proc AMIA Symp Date: 2000

5. Use of general-purpose negation detection to augment concept indexing of medical documents: a quantitative study using the UMLS.

Authors: P G Mutalik; A Deshpande; P M Nadkarni
Journal: J Am Med Inform Assoc Date: 2001 Nov-Dec Impact factor: 4.497

An analysis of statistical term strength and its use in the indexing and retrieval of molecular biology texts.

1. Including biological literature improves homology search.

2. UMLS concept indexing for production databases: a feasibility study.

3. The NLM Indexing Initiative.

4. What's related? Generalizing approaches to related articles in medicine.

5. Use of general-purpose negation detection to augment concept indexing of medical documents: a quantitative study using the UMLS.

6. Update on XplorMed: A web server for exploring scientific literature.

Review 7. Natural Language Processing methods and systems for biomedical ontology learning.

8. Semi-automatic indexing of full text biomedical articles.

9. A document clustering and ranking system for exploring MEDLINE citations.

10. From episodes of care to diagnosis codes: automatic text categorization for medico-economic encoding.